Word sang PDF
Chuyển .docx sang PDF, dàn lại trang từ cấu trúc của tài liệu.
Công cụ này chạy hoàn toàn trong trình duyệt của bạn. Tệp của bạn không bao giờ được tải lên, và bạn có thể tự kiểm chứng điều đó trong tab mạng của trình duyệt. Tự kiểm chứng: mở tab mạng của trình duyệt và theo dõi. Bạn sẽ thấy một yêu cầu nhỏ hỏi xem bạn còn tác vụ nào không - một tên công cụ và một mã băm, không bao giờ là tệp.
Công cụ này làm gì
This reads a .docx directly: the file is unzipped, its XML is parsed, and the document's structure is laid out afresh by our own typesetter. Headings, paragraphs, lists, tables, images and hyperlinks all come across. What does not come across is Word's own pagination - the text is re-flowed, so page breaks will not always fall where Word put them. For a CV, a letter, a report or a contract that difference is invisible; for a document whose layout is load-bearing, Word's own Save as PDF is the honest answer.
Use it when you need a PDF and Word is not to hand: a CV that an application form insists on receiving as a PDF, a contract you want to send in a format nobody edits by accident, or a document sitting on a phone where installing an office suite is not a realistic step.
Cách hoạt động
- Drop your .docx files onto this page. Up to 20 at a time, each up to 100 MB.
- Choose the paper size, orientation and margins. Normal margins are 56 points on every side, which sits close to Word's own default.
- Leave keep hyperlinks on so web addresses become real, clickable annotations rather than blue text that does nothing.
- Leave bookmarks from headings on and the PDF outline builds itself from your heading styles.
- Under advanced options, switch on combine into one PDF if several documents should arrive as a single file.
- Press Convert. Each document downloads as its own PDF unless you asked for them merged.
The conversion is a re-layout, not a re-render. Word's pagination is the product of Word's line-breaking, hyphenation and widow rules, and reproducing it exactly means reproducing all three. Instead the document's structure is measured against the paper you chose and flowed onto pages here, with the same widow and orphan control every conversion on this site uses. A heading is never left stranded alone at the foot of a page, and a single line is never left behind by itself.
Headers, footers and page-margin content are not carried across. The document model behind every conversion here describes flowing content - paragraphs, lists, tables, images - and it has no slot for a running header or a footer that repeats. If your document depends on a letterhead in the header, convert it first and then apply the letterhead with Stamp PDFs, which puts it on every page more reliably than a header ever did.
Fonts are the other honest limit. A Unicode font is embedded by default, about 2 MB of Noto Sans and Noto Sans Mono served from this site's own origin, so curly quotes, accents and Vietnamese diacritics survive. That set does not cover Chinese, Japanese, Korean, Arabic or Indic scripts, because the faces that would are tens of megabytes each. Characters it cannot draw are reported to you rather than quietly swapped for boxes.
Bookmarks build themselves from the heading styles. Heading 1 through Heading 6 become outline entries at the matching depth, so a long report arrives with a working navigation pane that nobody had to add by hand. It is also what makes a batch conversion usable: combine twenty documents into one PDF and each file's headings sit under its own branch of the outline rather than in a flat list.
A legacy .doc is a different format entirely and is not supported. The drop zone accepts one anyway, so it can tell you what to do about it instead of refusing the file with no explanation. Open it in Word, LibreOffice or Google Docs, save it as .docx, and come back.
Công cụ này không làm được gì
- Page breaks are recalculated rather than copied. Text is re-flowed onto the paper you choose, so a break can fall a paragraph earlier or later than Word placed it.
- Headers, footers and other page-margin content are not carried across. Add what you need afterwards with Add page numbers or Stamp PDFs.
- Chinese, Japanese, Korean, Arabic and Indic scripts are outside the embedded font's coverage. Characters that cannot be drawn are reported rather than silently replaced.
- The legacy .doc format is not supported. Save the file as .docx in Word, LibreOffice or Google Docs first.
Câu hỏi thường gặp
- PDF có giống hệt tài liệu Word của bạn không?
- Gần giống, nhưng không trùng từng trang. Tiêu đề, chữ đậm và nghiêng, danh sách, bảng, hình ảnh và liên kết đều được tái tạo, sau đó phần chữ được dàn lại theo khổ giấy bạn chọn, nên chỗ ngắt trang có thể rơi sớm hoặc muộn hơn một đoạn so với Word. Với một CV, một lá thư hay một báo cáo thì khác biệt đó không đáng kể. Với biểu mẫu, chứng chỉ hay bất cứ thứ gì mà vị trí mang ý nghĩa, hãy dùng chính chức năng Save as PDF của Word.
- Bạn chuyển được tệp .doc cũ không?
- Không. Định dạng .doc là một vùng chứa nhị phân từ thập niên 1990, chẳng có gì chung với dạng XML nén của .docx, nên hỗ trợ nó đồng nghĩa với việc phải làm thêm một bộ chuyển đổi thứ hai chứ không phải một bổ sung nhỏ. Hãy mở tệp trong Word, LibreOffice hoặc Google Docs, lưu thành .docx rồi chuyển tệp đó. Công cụ này chỉ nhận tệp .doc đủ lâu để báo cho bạn biết, thay vì từ chối mà không nêu lý do.
- Tài liệu của bạn dùng một phông chữ bạn đã cài. Nó có hiển thị giống vậy không?
- Không. Phần chữ được dàn bằng Noto Sans, hoặc bằng các phông PDF tiêu chuẩn nếu bạn tắt nhúng phông. Tệp .docx chỉ nêu tên các phông nó muốn nhưng thường không chứa chúng, và việc tải một phông thương mại thay cho bạn không phải là điều công cụ này có thể hay nên làm. Vì vậy độ dài dòng khác một chút so với Word, và đó là một trong những lý do chỗ ngắt trang bị xê dịch.
- Hình ảnh trong tài liệu của bạn có được giữ lại không?
- Có. Hình ảnh nhúng trong .docx được trích ra và đặt vào mạch nội dung theo đúng kích thước ghi trong tệp, và một ảnh dùng nhiều lần chỉ được giữ một bản chứ không lưu lặp lại. Hộp văn bản, hình khối và SmartArt thì không: chúng là phần đồ hoạ đặt theo vị trí chứ không phải nội dung trôi theo mạch, mà mô hình chuyển đổi này chỉ mô tả mạch trôi.
- Tài liệu của bạn có bị tải lên để chuyển đổi không?
- Không. Tệp .docx được giải nén và phân tích trong một Web Worker bên trong trình duyệt của bạn, và tệp PDF cũng được tạo ngay tại đó. Thứ duy nhất trang này tải về là tệp phông Noto, từ chính máy chủ của trang, và một khi trình duyệt đã lưu nó vào bộ nhớ đệm thì việc chuyển đổi vẫn chạy được khi tắt kết nối. Hãy mở thẻ mạng và thử chuyển một tệp nếu bạn muốn tự kiểm tra thay vì tin lời chúng tôi.
- Bạn chuyển được nhiều tài liệu cùng lúc không?
- Có, tối đa 20 tệp trong một lần, và theo mặc định mỗi tệp thành một PDF riêng. Bật "gộp thành một PDF" trong phần tuỳ chọn nâng cao thì chúng được ghép thành một tài liệu theo thứ tự trong danh sách, với tiêu đề của mỗi tệp được giữ thành một nhánh riêng trong mục lục.