Cookies for analytics and advertising
We use cookies for analytics and advertising, both sent to Google. Refusing changes nothing you can see.Read the privacy page
DOCX to write, PDF to send — one is the document, the other is a picture of it.
Choose PDF when
Anything final — a contract, an invoice, a CV, a report. A PDF looks the same everywhere and cannot be edited by accident.
Choose DOCX when
Anything still being written or meant to be edited by somebody else. A PDF is a picture of a document; a DOCX is the document.
Converting a PDF back to DOCX recovers text, not layout. Anything laid out with care comes back approximately, and the more designed it was the worse the approximation.
DOCX has nowhere to put XMP records, where editing history and rights information sit, so that goes no further than the PDF. Worth checking before the original is deleted, and worth knowing if the point was to strip it.
Both PDF and DOCX hold several pages, so a multi-page document stays one file.
DOCX is a working format and PDF is a finished one. What comes back is editable text and objects rather than a picture of a page, which is usually the reason for the conversion and also where its limits are.
No browser reads DOCX. It is the less portable of the two, so it is worth being sure the program at the other end accepts it before sending one.
The usual programs do not overlap: PDF opens in Adobe Acrobat, Preview and LibreOffice Draw, DOCX in Microsoft Word, LibreOffice Writer and Google Docs — so whoever receives the result needs something from the second list.
The two are aimed at different work: PDF at handing a finished file over, print and archiving, DOCX at editing. That is worth weighing before converting, because the reason one exists is usually the reason the other is awkward.
PDF is Adobe's format, published in 1993. The specification is ISO 32000-2, and it is worth reading if the file has to outlive the tool that wrote it.
DOCX comes from Microsoft and dates from 2007, specified as ECMA-376. Microsoft Word, LibreOffice Writer and Google Docs all read it.
The document becomes a fixed page: PDF stores how it looks rather than how it was built.
document properties can cross over — both DOCX and PDF have somewhere to store it.
Both DOCX and PDF hold several pages, so a multi-page document stays one file.
PDF opens in every current browser. DOCX has narrower browser support than that. If the file is going onto a web page or into a form, that is usually the whole reason for the conversion.
The usual programs do not overlap: DOCX opens in Microsoft Word, LibreOffice Writer and Google Docs, PDF in Adobe Acrobat, Preview and LibreOffice Draw — so whoever receives the result needs something from the second list.
The two are aimed at different work: DOCX at editing, PDF at handing a finished file over, print and archiving. That is worth weighing before converting, because the reason one exists is usually the reason the other is awkward.
DOCX is Microsoft's format, published in 2007. The specification is ECMA-376, and it is worth reading if the file has to outlive the tool that wrote it.
PDF comes from Adobe and dates from 1993, specified as ISO 32000-2. Adobe Acrobat, Preview and LibreOffice Draw all read it.
| DOCX | ||
|---|---|---|
| Full name | Portable Document Format | Word Document |
| Extension | .docx | |
| Media type | application/pdf | application/vnd.openxmlformats-officedocument.wordprocessingml.document |
| Compression | Either, depending on the setting | Uncompressed |
| This site can write it | Yes | Yes |
Which to choose
This is not a comparison of two ways to store a document. It is the difference between a document that is still being written and one that is done, and almost every question about the pair resolves once that is clear.
A DOCX describes intent: this paragraph is a heading, this list is numbered, this text is in Calibri at eleven point. Where the words fall on the page is worked out when it is opened, by that computer, with those fonts. A PDF describes a result: this glyph, at this position, on this page, with the font embedded so the answer cannot change. One is meant to be edited; the other is meant to be final.
Because the layout is recomputed every time it opens. A missing font is substituted with something of a different width, so line breaks move, and a document that filled four pages fills five. Different versions of Word disagree about spacing. Printer drivers affect margins. Two people reading "the same" file are reading two renderings of it.
None of that can happen to a PDF, because the layout was decided once and written down, and the fonts travel inside the file. That single property is why every CV, invoice, contract, application form and official submission is a PDF, and it is the entire value of the format.
The honest answer people rarely get. A PDF contains positioned glyphs, not paragraphs. A converter has to look at where the text sits and infer the structure back: these lines are probably one paragraph, this larger text was probably a heading, these aligned runs were probably a table.
For a plain report the inference is good and the result is genuinely editable. For anything designed — a brochure, a form, a multi-column layout, a report with pull quotes — it degrades sharply, and what comes out is a document held together by text boxes and manual spacing that is harder to edit than retyping. Scanned PDFs are a separate case again: they contain pictures of text, so optical character recognition has to read them first, and it makes mistakes.
The rule that follows: never treat the PDF as your only copy of something you may need to change. Keep the DOCX.
DOCX for collaboration: tracked changes, comments, version history, and the ability for two people to work on the same file. There is no equivalent in PDF worth the name — annotation exists and is not the same thing as editing.
PDF for everything final. It prints identically, it can be signed digitally, it can be locked against changes, it holds forms that can be filled in and submitted, and it can be archived to a standard designed for exactly that. PDF/A exists because national archives needed a format that will still render in fifty years, and no word-processor format offers anything comparable.
A DOCX carries its history: author names, the organisation, the time spent editing, often tracked changes and comments that were hidden rather than accepted, sometimes earlier text that was cut. Documents have embarrassed people this way repeatedly.
A PDF is safer and not safe. Metadata records the author and the software; text under a black rectangle is still text, and copying it out is trivial — that particular mistake has appeared in court filings and government releases more than once. Proper redaction removes the content rather than covering it. If a document is sensitive, inspect it before sending, in either format.
A DOCX written with real heading styles is accessible almost by default: a screen reader can navigate the structure, because the structure is what the file records.
A PDF is only accessible if somebody made it so. Tags describing the reading order have to be present, images need alternative text, and tables need their headers marked — and a PDF exported carelessly has none of it, leaving a screen reader to read a two-column page straight across. Exporting from a well-structured Word document carries most of it over; exporting from a design tool usually does not. For public sector publication this is a legal requirement rather than a nicety.
Write in DOCX. Send PDF. Keep both — the DOCX is the version you can change next year, and the PDF is the version that will look the way you meant it to.
Convert PDF to DOCX only when the original is genuinely gone, and expect to spend time on the result. If the document was ever a Word file, asking whoever produced it for the original will take five minutes and save an hour of fixing text boxes.
Yes, with caveats. A PDF holds positioned glyphs rather than paragraphs, so a converter has to infer the structure. Plain reports come back well; designed layouts, forms and multi-column pages come back as a mess of text boxes. Scanned PDFs need OCR first, which introduces errors.
Because Word recalculates the layout on opening. A missing font is replaced by one of a different width, so line breaks move and page counts change. A PDF cannot do this — the layout was fixed when it was made and the fonts are inside the file.
PDF, unless the employer or agency specifically asks for Word. It looks the same on every machine and cannot be edited by accident. Some recruitment systems still ask for DOCX so they can parse it, in which case send what they ask for.
Safer, not safe. A DOCX carries author names, editing history and often hidden tracked changes. A PDF carries less, but its metadata still names you, and text hidden under a black rectangle can be copied straight out. Real redaction removes content rather than covering it.
DOCX by default, because a document written with real heading styles already records its structure. A PDF is only accessible if it was tagged with a reading order, alternative text and table headers — exporting from a well-structured Word file carries most of that over, and exporting from a design tool usually does not.
Small corrections are possible in a PDF editor — fixing a typo, filling a form field, adding a signature. Reflowing a paragraph or changing a layout is not, because the file describes positions rather than structure. For real editing, go back to the source document.
The claims this page makes about PDF and DOCX are checkable, and these are the documents that settle them.