本文へスキップ

PDF を翻訳

文書のテキストを別の言語に翻訳します。

テキストがAIサービスへ送信されます

PDFそのものがアップロードされることはありません。テキストはブラウザーの中で抽出して画面に表示し、あなたが承認してから初めて送信されます。

このツールでできること

This extracts a document's text in your browser, sends it to the AI provider you choose with your own key, and typesets the translation as a new PDF. What comes back is a clean document of translated text, not your original with the words swapped: German runs about a third longer than English and Indonesian shorter, so substituting in place would overflow some lines and leave holes in others. Paragraphs of prose survive that; images, tables, columns, headers, footers and the original pagination do not.

Use it when you need to read a document rather than reproduce it: supplier terms that arrived in a language you do not work in, a research paper you want to skim before paying to have it translated properly, an Indonesian regulation you need the substance of, or a manual whose instructions you have to follow today.

仕組み

  1. Drop the PDF onto this page. Its text is extracted and shown to you, with an estimate of how much will be sent.
  2. Choose your provider and paste your API key. It stays in this browser and goes only to the provider it belongs to.
  3. Pick the target language. Thirteen are offered, and five of them come back as a text file rather than a PDF, for the reason below.
  4. Set the register under formality if it matters, and the paper size for the typeset result, under advanced options.
  5. Press Translate, then download the PDF, the plain text, or both.

The reason this produces a new document rather than an edited one is worth stating plainly, because plenty of tools imply otherwise. Text in a PDF is drawn at fixed coordinates; it does not sit in boxes that can grow. Replace an English sentence with its German translation and it is roughly a third too long for the space it has; replace it with Indonesian and there is a gap where the rest of the line used to be. Tools that do it anyway produce overlapping lines and clipped words. This one typesets the translation fresh, on the paper size you pick and with page numbers.

Five languages are delivered as text rather than as a PDF, and that is enforced rather than warned about. The font embedded in generated PDFs is Noto Sans, which covers Latin, Greek, Cyrillic and Vietnamese. Arabic, Chinese, Japanese, Hindi and Thai fall outside it, and a PDF built from glyphs a font does not have is a page of empty rectangles. So for those five you get a text file with a line at the top explaining why.

The text is cut into chunks of about 24,000 characters - smaller than the summariser uses, because a translation is roughly as long as its source and each request has to leave room for its own answer. Every chunk carries the same instruction: preserve meaning, structure and paragraph breaks, keep numbers, names and figures exactly, do not summarise and do not comment. Chunk boundaries fall between paragraphs, but a term rendered one way in one chunk can still come out differently in another.

Machine translation from a large model is good enough to work from and not good enough to sign. It handles ordinary prose well, technical vocabulary less reliably, and legal or contractual phrasing least reliably of all - which is to say it is weakest exactly where a translation matters most. For anything binding, this is a way to find out what a document says before engaging a translator, not a way to avoid engaging one.

このツールにできないこと

  • The result is re-typeset from scratch. Images, tables, columns, headers, footers and the original page breaks are not reproduced.
  • Arabic, Chinese, Japanese, Hindi and Thai are delivered as a text file, because the embedded font does not cover those scripts.
  • The document's text is sent to the AI provider you choose, using your own key. The file itself is not.
  • A scanned PDF has no text to translate; run OCR a PDF over it first.

よくある質問

翻訳しても元のレイアウトは保たれますか?
保たれません。この方式で動くツールで、正直にそうだと言えるものはありません。翻訳は新しい文書として組み直されます。段落は順番どおり、選んだ用紙サイズ、ページ番号付きです。画像・表・段組み・元の改ページは再現されません。翻訳された文字は元とは長さが違うので、元の固定位置に戻せば、ある行はあふれ、別の行には隙間が空きます。
ファイルはアップロードされますか?
ファイルはされません。そこから抽出したテキストはされますが、送り先はあなたが指定したプロバイダーだけで、あなたのブラウザが、ヘッダーにあなた自身のキーを入れて送ります。その経路に CekPDF のサーバーはありません。何が、どれだけ送られるのかは、送信前にパネルがそのまま表示します。
中国語がなぜ PDF ではなくテキストファイルで返ってきたのですか?
生成される PDF に埋め込まれるフォントが Noto Sans で、ラテン文字・ギリシャ文字・キリル文字・ベトナム語は扱えますが、中国語・日本語・アラビア語・ヒンディー語・タイ語は扱えないためです。グリフを持たないフォントで PDF を作ると、語があるべき場所に空の四角が並びます。それをお渡しするかわりに、説明を添えたテキストとして翻訳を出力します。必要なフォントを持つワープロに、そのまま貼り付けられます。
公式な文書に使える品質ですか?
拘束力のあるものには使えません。大規模言語モデルは普通の文章はうまく訳し、専門用語はそれより不確かで、法律の言い回しがもっとも不確かです。文書に何が書かれているかを知り、それが重要かどうかを判断し、人間の翻訳者にどこへ注力してほしいかを伝えるために使ってください。これを根拠に署名しないでください。
翻訳 1 回の費用はどのくらいですか?
ここでは無料で、請求はプロバイダーからです。翻訳は要約よりページあたりの費用が高くつきます。出力が入力の一部ではなく、ほぼ同じ長さになるからです。テキストはおよそ 24,000 文字のチャンクに分けて送られ、チャンクごとに 1 リクエストです。送信前にパネルが分量を見積もり、文字数の上限(既定で 120,000 文字)が、1 回の実行で使える額に上限をかけます。

関連ツール