สแกนเป็น PDF
ถ่ายรูปเอกสารแล้วได้ไฟล์ที่อ่านเหมือนงานสแกน
เครื่องมือนี้ทำงานในเบราว์เซอร์ของคุณทั้งหมด ไฟล์ของคุณไม่ถูกอัปโหลดเลย และคุณตรวจสอบได้เองในแท็บเครือข่ายของเบราว์เซอร์ ตรวจสอบด้วยตัวเอง เปิดแท็บเครือข่ายของเบราว์เซอร์แล้วดู คุณจะเห็นคำขอเล็ก ๆ หนึ่งรายการที่ถามว่าคุณยังมีโควตางานเหลืออยู่หรือไม่ ซึ่งมีแค่ชื่อเครื่องมือกับค่าแฮช ไม่ใช่ตัวไฟล์
เครื่องมือนี้ทำอะไร
Most people do not own a scanner and everybody owns a camera, but a photograph of a page is not a scan. This closes that gap in four steps: the page is found in the frame, the perspective is flattened, the text is straightened, and the image is thresholded so it reads as a document rather than as a picture of one. On a phone the camera panel captures directly; on a desktop it falls back to the file picker, because most people there are dropping photos they already took.
Use it when something has to be submitted as a scan and there is no scanner in the building: an identity document or a certificate for an application, a signed page going back to an office, a receipt for an expenses claim, or a handwritten note that has to be filed with everything else.
วิธีทำงาน
- Take the photo here, or drop photos you have already taken. Up to 100 at a time, each up to 100 MB, and every photo becomes one page in the order they are listed.
- Leave find the page and straighten on. A detection the tool is not confident about is discarded rather than guessed at, and the whole frame is used instead.
- Choose the mode. Black and white is the default and is what makes a photograph look like a scan; grayscale suits a page with photographs on it; colour keeps a stamp or a signature in its own ink.
- Pick the page size. A4 or Letter puts every photo on a standard sheet; Match makes each page the shape of its own photo.
- Turn on the searchable-text option if you will need to find or copy words later, and pick the language.
- Press Scan. The pages are assembled into a single PDF and the download starts.
The four steps run in that order because each depends on the one before. The document is a bright quadrilateral somewhere in a frame that also contains a desk, so its outline is detected first. A homography then maps that trapezoid back to a rectangle, which is what removes the leaning look of a photo taken at an angle. Only then is the residual rotation measured, and it is measured from the text lines rather than the page edges, because lines of text give away the true angle more reliably than a torn or shadowed edge does.
Thresholding is the step that decides whether the result is usable. A photograph has uneven lighting, so a single global threshold turns one half of the page black and blows the other half out. An adaptive threshold compares each pixel to its own neighbourhood instead, which is why a phone photo comes out looking like a scan rather than like a photocopy of a shadow. It is also why black and white is the default rather than an option for the brave.
The mode chooses the encoding as well as the look, and that is where the file size goes. A thresholded page is two colours and compresses to almost nothing as PNG, while JPEG would put ringing artefacts around every letter. A grayscale or colour page is photographic and is the other way round, so those are written as JPEG at quality 0.85. A black-and-white scan is routinely several times smaller than the same page in colour.
Resolution is set in dpi and defaults to 200, which is enough for ordinary printed text. Raise it to 300 for small print, faint carbon copies or anything the recognition is struggling with. 400 stores four times the pixels of 200 and rarely reveals something the threshold at 200 had not already found, so it is worth checking one page before committing a hundred. EXIF orientation is read and applied before any of the geometry, so a photo your phone tagged sideways is not straightened into the wrong rectangle.
Recognition, when you switch it on, runs on the processed pages rather than the originals - that is the whole point of doing the image work first, because a straightened, thresholded page recognises far better than a photograph of one. The words are placed as an invisible text layer over the image, so the page looks exactly as it did and the text can be selected, searched and copied. Words the recogniser scored below 40 per cent are left out rather than added wrong, since a wrong word in a searchable document is worse than a missing one.
สิ่งที่เครื่องมือนี้ทำไม่ได้
- Page detection can fail on a low-contrast background, a page that runs off the edge of the frame, or a heavily shadowed edge. It then falls back to the whole frame and says so rather than cropping to the wrong rectangle.
- Adaptive thresholding handles uneven light, not every kind of bad light. Hard glare, a strong reflection off glossy paper or a deep shadow across the text will still be visible in the result.
- Capturing directly needs a camera the browser can reach and your permission to use it. On a device without one, the tool falls back to the file picker and works on photos taken elsewhere.
คำถามที่คนมักถาม
- ฉันถ่ายรูปที่นี่ได้เลยไหม หรือต้องมีรูปอยู่ก่อน
- ได้ทั้งสองอย่าง บนมือถือหรือแท็บเล็ต แผงกล้องจะเปิดขึ้นในหน้านี้ แสดงกรอบหน้าเอกสารที่ตรวจพบแบบสด และถ่ายเข้าคิวได้เลย ส่วนบนคอมพิวเตอร์ หรือที่ใดก็ตามที่เบราว์เซอร์เข้าถึงกล้องไม่ได้ จะกลับไปใช้ตัวเลือกไฟล์ คุณจึงวางรูปที่ถ่ายด้วยอย่างอื่นได้ การประมวลผลหลังจากนั้นเหมือนกันทั้งสองทาง
- ฉันทำอย่างไรให้ผลลัพธ์ดูเหมือนงานสแกนจริง
- วางเอกสารให้เรียบ ให้กินพื้นที่ส่วนใหญ่ของกรอบภาพ และให้มุมทั้งสี่อยู่ในรูป เพื่อให้หาขอบหน้าได้ แสงสม่ำเสมอที่ไม่ส่องตรงดีกว่าโคมไฟจ้า ซึ่งจะทิ้งขอบเงาคมที่การปรับเป็นขาวดำลบไม่ออก อย่าให้เงาของคุณเองทาบลงบนหน้าเอกสาร และถือกล้องให้ขนานกับหน้าเอกสารพอสมควร การแก้เพอร์สเปคทีฟจัดการความเอียงได้ แต่ยิ่งเอียงน้อยก็ยิ่งเหลือให้ดึงน้อย
- มันครอบตัดผิดที่ หรือไม่ครอบตัดเลย ต้องทำอย่างไรต่อ
- การตรวจจับที่เครื่องมือไม่มั่นใจจะถูกทิ้งโดยตั้งใจ เพราะการครอบตัดอัตโนมัติที่ผิดแย่กว่าการไม่ครอบตัดเลย ระบบจะใช้ทั้งกรอบภาพแทน และผลลัพธ์จะบอกคุณว่ามีกี่หน้าที่ถูกครอบตัด ถ้าหน้าไหนถูกครอบตัดผิด ให้ถ่ายใหม่บนพื้นหลังที่ตัดกับสีของกระดาษ ถ้าคุณครอบตัดรูปมาจากที่อื่นแล้ว ให้ปิดการหาขอบหน้าไปเลย
- ไฟล์ PDF ออกมาใหญ่ ฉันควรเปลี่ยนอะไร
- โหมดก่อน แบบขาวดำถูกเขียนเป็น PNG ของภาพสองสี และเล็กกว่าหน้าเดียวกันแบบสีหรือสีเทามาก ความละเอียดเป็นอย่างที่สอง 200 dpi พอสำหรับข้อความทั่วไป ส่วน 400 เก็บพิกเซลไว้มากกว่าสี่เท่า ถ้ายังใหญ่เกินไปหลังจากนั้น ให้นำผลลัพธ์ไปผ่านเครื่องมือบีบอัด PDF
- ฉันค้นหาข้อความในภายหลังได้ไหม
- ได้ ถ้าคุณเปิดการรู้จำข้อความก่อนแปลง มีให้เลือกทั้งภาษาอังกฤษ ภาษาอินโดนีเซีย และทั้งสองภาษาพร้อมกัน คำที่รู้จำได้จะถูกเขียนเป็นชั้นข้อความที่มองไม่เห็นทับบนภาพหน้าเอกสาร เอกสารจึงดูเหมือนเดิม แต่ข้อความเลือก ค้นหา และคัดลอกได้ คำที่ได้คะแนนความมั่นใจต่ำกว่า 40 เปอร์เซ็นต์จะถูกละไว้แทนที่จะถูกเดา
- รูปของฉันถูกอัปโหลดไหม โดยเฉพาะเมื่อเปิดการรู้จำข้อความ
- ไม่ และการรู้จำข้อความคือกรณีที่พลาดได้ง่ายที่สุด งานประมวลผลภาพทำงานใน Web Worker ในเบราว์เซอร์ของคุณ ส่วนแกน WebAssembly และข้อมูลภาษาของตัวรู้จำถูกให้บริการจากโดเมนของเว็บไซต์นี้เอง ไม่ใช่จาก CDN สาธารณะ ราว 10.9 MB สำหรับภาษาอังกฤษ และ 3.8 MB สำหรับภาษาอินโดนีเซีย โดยถูกแคชไว้หลังใช้ครั้งแรก รูปถ่ายหนังสือเดินทางของคุณไม่เคยออกจากเครื่อง