본문으로 건너뛰기

PDF 번역

문서의 텍스트를 다른 언어로 옮겨요.

텍스트가 AI 서비스로 전송됨

PDF는 절대 업로드되지 않아요. 텍스트는 브라우저에서 추출되어 먼저 보여 드리고, 승인한 뒤에만 전송돼요.

이 도구가 하는 일

This extracts a document's text in your browser, sends it to the AI provider you choose with your own key, and typesets the translation as a new PDF. What comes back is a clean document of translated text, not your original with the words swapped: German runs about a third longer than English and Indonesian shorter, so substituting in place would overflow some lines and leave holes in others. Paragraphs of prose survive that; images, tables, columns, headers, footers and the original pagination do not.

Use it when you need to read a document rather than reproduce it: supplier terms that arrived in a language you do not work in, a research paper you want to skim before paying to have it translated properly, an Indonesian regulation you need the substance of, or a manual whose instructions you have to follow today.

작동 방식

  1. Drop the PDF onto this page. Its text is extracted and shown to you, with an estimate of how much will be sent.
  2. Choose your provider and paste your API key. It stays in this browser and goes only to the provider it belongs to.
  3. Pick the target language. Thirteen are offered, and five of them come back as a text file rather than a PDF, for the reason below.
  4. Set the register under formality if it matters, and the paper size for the typeset result, under advanced options.
  5. Press Translate, then download the PDF, the plain text, or both.

The reason this produces a new document rather than an edited one is worth stating plainly, because plenty of tools imply otherwise. Text in a PDF is drawn at fixed coordinates; it does not sit in boxes that can grow. Replace an English sentence with its German translation and it is roughly a third too long for the space it has; replace it with Indonesian and there is a gap where the rest of the line used to be. Tools that do it anyway produce overlapping lines and clipped words. This one typesets the translation fresh, on the paper size you pick and with page numbers.

Five languages are delivered as text rather than as a PDF, and that is enforced rather than warned about. The font embedded in generated PDFs is Noto Sans, which covers Latin, Greek, Cyrillic and Vietnamese. Arabic, Chinese, Japanese, Hindi and Thai fall outside it, and a PDF built from glyphs a font does not have is a page of empty rectangles. So for those five you get a text file with a line at the top explaining why.

The text is cut into chunks of about 24,000 characters - smaller than the summariser uses, because a translation is roughly as long as its source and each request has to leave room for its own answer. Every chunk carries the same instruction: preserve meaning, structure and paragraph breaks, keep numbers, names and figures exactly, do not summarise and do not comment. Chunk boundaries fall between paragraphs, but a term rendered one way in one chunk can still come out differently in another.

Machine translation from a large model is good enough to work from and not good enough to sign. It handles ordinary prose well, technical vocabulary less reliably, and legal or contractual phrasing least reliably of all - which is to say it is weakest exactly where a translation matters most. For anything binding, this is a way to find out what a document says before engaging a translator, not a way to avoid engaging one.

이 도구가 할 수 없는 일

  • The result is re-typeset from scratch. Images, tables, columns, headers, footers and the original page breaks are not reproduced.
  • Arabic, Chinese, Japanese, Hindi and Thai are delivered as a text file, because the embedded font does not cover those scripts.
  • The document's text is sent to the AI provider you choose, using your own key. The file itself is not.
  • A scanned PDF has no text to translate; run OCR a PDF over it first.

자주 묻는 질문

번역본이 원래 레이아웃을 그대로 지키나요?
아니요, 이런 방식으로 동작하는 도구라면 어느 것도 정직하게는 못 해요. 번역문은 새 문서로 조판돼요. 문단이 순서대로, 고른 용지 크기에, 쪽 번호와 함께요. 이미지와 표, 단, 원래의 쪽 나눔은 재현되지 않아요. 번역한 글은 원문과 길이가 달라서, 원본의 고정된 자리에 그대로 넣으면 어떤 줄은 넘치고 어떤 줄은 비어요.
제 파일이 업로드되나요?
파일은 아니에요. 거기서 뽑은 텍스트가 가는데, 지정한 제공자에게만 가고 내 브라우저가 헤더에 내 키를 넣어 보내요. 그 경로에 CekPDF 서버는 없어요. 나가기 전에 무엇이 얼마나 보내지는지 패널이 그대로 보여 줘요.
중국어는 왜 PDF가 아니라 텍스트 파일로 돌아왔나요?
만들어지는 PDF에 넣는 글꼴이 Noto Sans인데, 라틴과 그리스, 키릴, 베트남어 문자는 다루지만 중국어와 일본어, 아랍어, 힌디어, 태국어는 못 다루기 때문이에요. 글리프가 없는 글꼴로 만든 PDF는 글자가 있어야 할 자리에 빈 네모가 나와요. 그런 걸 건네는 대신 번역문을 텍스트로 주고 이유를 적어 둬요. 그 글꼴이 있는 워드프로세서에 바로 붙여넣을 수 있어요.
공식 문서로 쓸 만한 품질인가요?
구속력이 있는 문서에는 안 돼요. 대형 언어 모델은 보통 산문은 잘 옮기고, 전문 용어는 덜 안정적이고, 법률 문구는 그중 가장 못 미더워요. 문서에 무슨 말이 있는지 알고, 중요한 문서인지 판단하고, 사람 번역자에게 어디에 집중할지 알려 주는 데 쓰세요. 이걸 믿고 서명하지는 마세요.
번역 한 번에 비용이 얼마나 드나요?
여기서는 안 들고 제공자가 청구해요. 번역은 결과가 입력의 일부가 아니라 입력만큼 길어서 요약보다 쪽당 비용이 커요. 텍스트는 2만 4천 자쯤씩 조각으로 나가고 조각마다 요청이 한 번이에요. 보내기 전에 패널이 크기를 어림해 주고, 글자 수 제한은 기본값 12만 자로 한 번에 쓸 수 있는 양을 묶어 줘요.

관련 도구