PDF Word Swap

PDF to Word

Your file never leaves this device. Conversion runs in your browser. The document itself is not uploaded.

1 · Choose a PDF file

…or drop a PDF file here

.pdf · stays in this tab

2 · Convert

3 · Download

Choose a file to convert. It is not uploaded.

FAQ and format details
Why convert PDF to Word?

A contract scan that actually has a text layer, a paper you want to annotate in Word, a form letter trapped in PDF. If you need letterhead and columns, you need a different class of tool (or the original DOCX). This is the honest local path.

How this conversion runs

PDF text → HTML article with page headings → WordprocessingML paragraphs. You will see 'Page 1' style sections. Tables in the PDF rarely become Word tables.

Keep in mind for PDF → Word

This will disappoint anyone who thought 'convert' meant InDesign-quality reconstruction. Images, signatures as pictures, and vector logos will not land in the DOCX as graphics unless they were somehow in the HTML hub as data URLs (they are not, from pdf.js text). Scans need OCR first.

After you have a Word file

Open in Word, delete page labels, restyle. If the order is nonsense, the PDF was multi-column — fix in the HTML step or copy from Acrobat. For a reflowable book, PDF→EPUB after cleanup.

Why doesn't it look like the PDF?

PDF is paint. Word is a flow. We copy words, not coordinates. Tools that 'look identical' are often screenshotting pages into Word, which is also not editable text.

Will comments in the PDF appear?

No. Annotations are not extracted. Export comments from Acrobat if you need them.

This pair at a glance
Accepts
.pdf
Writes
.docx · application/vnd.openxmlformats-officedocument.wordprocessingml.document
Read fidelity
medium (PDF → HTML)
Write fidelity
high (HTML → Word)
Source size
Plan for files under 38 MB

The file is read in this tab. It is not uploaded.

Something off with this pair? Send feedback.

About PDF and Microsoft Word (Office Open XML)

Source file

PDF

A page-description format that paints glyphs and graphics onto fixed pages. Universal for print, archival, and sharing.

PDF (Portable Document Format) describes pages, not a flowing document tree. Incoming PDFs are reconstructed as HTML: designed pages keep positioned text, fonts, images, and vectors; simpler reading documents become flowing articles. Export to PDF paints HTML onto pages via pdf-lib. Pixel-perfect Word/InDesign layout is not the goal.

Extensions
.pdf
MIME types
application/pdf
Kind
binary · binary
Category
page description
Standard
Open format · since 1993
Spec
ISO 32000
Comfortable size
Up to about 38 MB in this browser converter
Read into HTML
medium fidelity · implemented — pdf.js extracts text, fonts, images, and vectors. Designed or overlapping pages stay visually positioned; simpler reading documents become flowing HTML with inferred headings, lists, and a sidebar when one exists. Scanned PDFs without a text layer will be empty.

What this format can hold

  • page layout
  • paragraphs
  • images
  • links
  • embedded fonts

Usually dropped on the way in

  • form fields
  • javascript
  • scanned pages without ocr

Apps that consume PDF

  • Adobe Acrobat / Reader (Adobe) — creates and opens on Windows, macOS, iOS, Android
  • Chrome (Google) — opens on Windows, macOS, Linux, Android
  • Edge (Microsoft) — opens on Windows, macOS
  • Preview (Apple) — opens on macOS, iOS
  • Foxit — creates and opens on Windows, macOS
  • Okular (KDE) — opens on Linux

Catalog notes

  • Encrypted PDFs cannot be opened in-browser without the password (not collected).
  • OCR for image-only scans is out of scope for v1.

Destination file

Microsoft Word (Office Open XML)

The default Word format since 2007: a ZIP package of XML parts for body text, styles, media, and document settings.

DOCX (Office Open XML WordprocessingML) replaced the binary .doc format. It is the interchange format for business documents: resumes, contracts, manuscripts. Word, Google Docs, LibreOffice, Pages, and OnlyOffice all read and write it. Macros live in a sibling .docm type, which this converter does not execute or emit.

Extensions
.docx
MIME types
application/vnd.openxmlformats-officedocument.wordprocessingml.document
Kind
zip · binary
Category
word processing
Standard
Open format · since 2006
Comfortable size
Up to about 24 MB in this browser converter
Written from HTML
high fidelity · implemented — A constrained HTML subset is mapped onto WordprocessingML runs and paragraphs.

What this format can hold

  • headings
  • paragraphs
  • lists
  • tables
  • links
  • images
  • inline formatting
  • headers footers
  • styles
  • comments
  • track changes

Usually dropped on the way out

  • css positioning
  • custom fonts as linked webfonts
  • page layout

Apps that consume Word

  • Microsoft Word (Microsoft) — creates and opens on Windows, macOS, Web, iOS, Android
  • Google Docs (Google) — creates and opens on Web, iOS, Android
  • LibreOffice Writer (The Document Foundation) — creates and opens on Windows, macOS, Linux
  • Apple Pages (Apple) — creates and opens on macOS, iOS
  • OnlyOffice — creates and opens on Windows, macOS, Linux, Web
  • WPS Office (Kingsoft) — creates and opens on Windows, macOS, Linux, Android

Catalog notes

  • .docm, .dotx, and .dotm are not accepted; strip macros before converting.
  • ZIP sniffing is shared with ODT and EPUB — the engine also checks [Content_Types].xml for WordprocessingML.
Other conversions

Other conversions from PDF: HTMLMarkdownRTFOpenDocumentEPUB

Other ways to get Word: HTMLMarkdownRTFOpenDocumentEPUB