스캔해서 PDF로
문서를 사진으로 찍으면 스캔한 것처럼 읽히는 결과가 나와요.
이 도구는 브라우저 안에서만 실행돼요. 파일은 절대 업로드되지 않으며, 브라우저의 네트워크 탭에서 직접 확인할 수 있어요. 직접 확인해 보세요. 브라우저의 네트워크 탭을 열어 두고 지켜보면, 남은 작업이 있는지 묻는 작은 요청 하나만 보여요. 도구 이름과 해시뿐이고, 파일은 절대 포함되지 않아요.
이 도구가 하는 일
Most people do not own a scanner and everybody owns a camera, but a photograph of a page is not a scan. This closes that gap in four steps: the page is found in the frame, the perspective is flattened, the text is straightened, and the image is thresholded so it reads as a document rather than as a picture of one. On a phone the camera panel captures directly; on a desktop it falls back to the file picker, because most people there are dropping photos they already took.
Use it when something has to be submitted as a scan and there is no scanner in the building: an identity document or a certificate for an application, a signed page going back to an office, a receipt for an expenses claim, or a handwritten note that has to be filed with everything else.
작동 방식
- Take the photo here, or drop photos you have already taken. Up to 100 at a time, each up to 100 MB, and every photo becomes one page in the order they are listed.
- Leave find the page and straighten on. A detection the tool is not confident about is discarded rather than guessed at, and the whole frame is used instead.
- Choose the mode. Black and white is the default and is what makes a photograph look like a scan; grayscale suits a page with photographs on it; colour keeps a stamp or a signature in its own ink.
- Pick the page size. A4 or Letter puts every photo on a standard sheet; Match makes each page the shape of its own photo.
- Turn on the searchable-text option if you will need to find or copy words later, and pick the language.
- Press Scan. The pages are assembled into a single PDF and the download starts.
The four steps run in that order because each depends on the one before. The document is a bright quadrilateral somewhere in a frame that also contains a desk, so its outline is detected first. A homography then maps that trapezoid back to a rectangle, which is what removes the leaning look of a photo taken at an angle. Only then is the residual rotation measured, and it is measured from the text lines rather than the page edges, because lines of text give away the true angle more reliably than a torn or shadowed edge does.
Thresholding is the step that decides whether the result is usable. A photograph has uneven lighting, so a single global threshold turns one half of the page black and blows the other half out. An adaptive threshold compares each pixel to its own neighbourhood instead, which is why a phone photo comes out looking like a scan rather than like a photocopy of a shadow. It is also why black and white is the default rather than an option for the brave.
The mode chooses the encoding as well as the look, and that is where the file size goes. A thresholded page is two colours and compresses to almost nothing as PNG, while JPEG would put ringing artefacts around every letter. A grayscale or colour page is photographic and is the other way round, so those are written as JPEG at quality 0.85. A black-and-white scan is routinely several times smaller than the same page in colour.
Resolution is set in dpi and defaults to 200, which is enough for ordinary printed text. Raise it to 300 for small print, faint carbon copies or anything the recognition is struggling with. 400 stores four times the pixels of 200 and rarely reveals something the threshold at 200 had not already found, so it is worth checking one page before committing a hundred. EXIF orientation is read and applied before any of the geometry, so a photo your phone tagged sideways is not straightened into the wrong rectangle.
Recognition, when you switch it on, runs on the processed pages rather than the originals - that is the whole point of doing the image work first, because a straightened, thresholded page recognises far better than a photograph of one. The words are placed as an invisible text layer over the image, so the page looks exactly as it did and the text can be selected, searched and copied. Words the recogniser scored below 40 per cent are left out rather than added wrong, since a wrong word in a searchable document is worse than a missing one.
이 도구가 할 수 없는 일
- Page detection can fail on a low-contrast background, a page that runs off the edge of the frame, or a heavily shadowed edge. It then falls back to the whole frame and says so rather than cropping to the wrong rectangle.
- Adaptive thresholding handles uneven light, not every kind of bad light. Hard glare, a strong reflection off glossy paper or a deep shadow across the text will still be visible in the result.
- Capturing directly needs a camera the browser can reach and your permission to use it. On a device without one, the tool falls back to the file picker and works on photos taken elsewhere.
자주 묻는 질문
- 여기서 사진을 찍을 수 있나요, 아니면 미리 찍어 둬야 하나요?
- 둘 다 돼요. 휴대폰이나 태블릿에서는 카메라 패널이 페이지 안에서 열려 감지된 페이지 윤곽을 실시간으로 보여 주고, 찍으면 바로 대기열에 들어가요. 데스크톱이나 브라우저가 카메라에 접근할 수 없는 환경에서는 파일 선택기로 되돌아가니 다른 기기로 찍은 사진을 놓으면 돼요. 그 뒤의 처리는 어느 쪽이든 똑같아요.
- 진짜 스캔처럼 나오게 하려면 어떻게 찍어야 하나요?
- 종이를 평평하게 두고 화면을 거의 채우도록 찍되, 네 모서리가 모두 사진 안에 들어오게 하세요. 그래야 윤곽을 찾을 수 있어요. 밝은 조명 하나보다 고르고 은은한 빛이 나아요. 강한 조명은 임계 처리로 지울 수 없는 뚜렷한 그림자 경계를 만들거든요. 종이 위에 자기 그림자가 지지 않게 하고, 카메라는 종이와 되도록 나란히 두세요. 원근 보정이 기울기를 잡아 주지만, 기울기가 작을수록 늘려야 할 부분도 적어요.
- 엉뚱한 곳을 잘랐거나 아예 자르지 않았어요. 어떻게 하죠?
- 확신하지 못한 감지 결과는 일부러 버려요. 잘못된 자동 자르기는 아예 자르지 않는 것보다 나쁘기 때문이에요. 그럴 때는 사진 전체를 쓰고, 몇 쪽이 잘렸는지 결과에 알려 줘요. 어떤 쪽이 잘못 잘렸다면 종이와 대비되는 배경 위에서 다시 찍으세요. 사진을 이미 다른 데서 잘라 뒀다면 페이지 찾기를 아예 끄세요.
- PDF가 크게 나왔어요. 무엇을 바꿔야 하나요?
- 먼저 모드예요. 흑백은 두 가지 색만 있는 이미지를 PNG로 쓰기 때문에 같은 쪽을 컬러나 회색조로 담는 것보다 훨씬 작아요. 다음은 해상도예요. 보통의 글자에는 200dpi면 충분하고, 400dpi는 픽셀을 네 배로 저장해요. 그러고도 너무 크면 결과를 PDF 압축에 한 번 통과시키세요.
- 나중에 텍스트를 검색할 수 있나요?
- 네, 변환하기 전에 인식을 켜면 돼요. 영어와 인도네시아어, 그리고 둘을 함께 쓸 수 있어요. 인식된 단어는 페이지 이미지 위에 보이지 않는 층으로 쓰이기 때문에 문서 모습은 그대로인 채로 텍스트를 선택하고 검색하고 복사할 수 있어요. 확신도가 40퍼센트 아래로 매겨진 단어는 짐작하지 않고 빼요.
- 사진이 업로드되나요? 특히 인식을 켰을 때는요?
- 아니요. 그리고 인식이야말로 그 약속을 어기기 가장 쉬운 경우예요. 이미지 처리는 브라우저 안의 웹 워커에서 실행되고, 인식기의 WebAssembly 코어와 언어 데이터는 공개 CDN이 아니라 이 사이트 자체 출처에서 제공돼요. 영어는 약 10.9MB, 인도네시아어는 3.8MB이고 처음 실행한 뒤에는 캐시돼요. 여권을 찍은 사진은 절대 기기를 떠나지 않아요.