About the PDF to Word converter
This tool extracts the selectable text from a PDF and lays it out as paragraphs in a real Microsoft Word (.docx) file, with a page break between each original PDF page — a quick way to get PDF content into an editable document without retyping it.
It's built for pulling text out of a report, contract or article you need to edit, quote from or repurpose — without copying and pasting section by section, which usually mangles line breaks and paragraph structure. If the PDF already has a real text layer (rather than being a scanned image), this gets you an editable copy in seconds.
How it works
- Upload a PDF.
- The tool reads the text layer of every page with pdf.js.
- Each line becomes a paragraph in a new document, built with the docx library, with a page break inserted at every original page boundary.
- The
.docxfile downloads automatically.
Assumptions and behaviour
- Extracts text only — the output is a plain, readable Word document, not a pixel-for-pixel copy of the PDF's layout.
- Preserves line breaks and page breaks matching the source PDF, so document length and structure stay recognisable.
- Does not reproduce fonts, colors, images, tables, columns, or precise formatting — headings and body text come out as regular paragraphs.
Limitations
- Scanned PDFs (images of text) won't extract anything — this reads the PDF's existing text layer, it doesn't perform OCR.
- Multi-column layouts, tables and complex formatting may come out of order or as unstructured lines, since PDF text has no inherent paragraph structure.
- Encrypted or password-protected PDFs can't be processed.
Privacy
The PDF is read and the .docx built entirely in your browser using pdf.js and the docx library. Nothing is uploaded or stored.

