Convert PDF to Word

Back to Home
Add a PDF file — or drag them here
PDF

About this PDF to Word converter

This tool pulls the text out of a PDF and hands it to you as an editable document. It uses PDF.js to read the text objects the file already contains — every character with its position on the page — and groups them back into lines by comparing their vertical coordinates. The result is written either as a Word-compatible .doc file or as plain .txt.

It is worth setting expectations precisely, because “PDF to Word” promises more than any converter can deliver. A PDF describes where each glyph sits on a page; it does not record that a run of text was a heading, a bullet list, a table cell or a two-column layout. Reconstructing that structure is guesswork, and this tool does not attempt it. What you get is the text, in reading order, as paragraphs — a clean starting point for rewriting rather than a faithful replica.

The text is extracted and the document assembled inside your browser. Nothing is uploaded, unlike the conversion services that require you to hand over the file.

How to convert a PDF to an editable document

  1. Load the PDF. Drop it onto the zone. The file needs a real text layer — see below if yours is a scan.
  2. Choose the output. Word document produces a .doc that Word, Google Docs and LibreOffice all open. Plain text produces a .txt with no formatting at all.
  3. Convert. Each page is read in turn and its text grouped into lines and paragraphs.
  4. Download and tidy up. Open the result and expect to redo headings, lists and tables. The words and their order will be right.

What comes across, and what does not

ElementResult
Body textExtracted accurately, in reading order
Paragraph breaksReconstructed from line positions
Page breaksPreserved
Headings and stylesLost — everything becomes body text
TablesFlattened into lines of text
Images and chartsNot included
ColumnsRead in the file’s internal order, which may interleave

If your PDF is a scan

A scanned document contains photographs of pages, not text objects. There is nothing for this tool to extract and it will report that no text was found. Use the OCR tool instead, which recognises the characters visually and produces a text file you can work from.

You can tell the difference in any PDF reader: try selecting a sentence with the mouse. If a text cursor appears and words highlight, this converter will work. If you only get a rectangular selection over the whole page, it is a scan.

Frequently asked questions

Why does my converted document look nothing like the PDF?

Because a PDF stores glyph positions, not document structure. There is no record that something was a heading, a table or a bulleted list, so no converter can restore it reliably. This tool extracts the text faithfully and leaves the formatting to you.

Nothing was extracted from my PDF. Why?

Your file is almost certainly a scan — images of pages with no text layer. Use the OCR tool, which recognises the characters visually instead.

Will tables be preserved?

No. Table cells are just positioned text in a PDF, so they come out as ordinary lines. Complex tables usually need rebuilding by hand.

What is the difference between the .doc and .txt output?

The .doc keeps paragraph and page breaks and opens in Word, Google Docs and LibreOffice. The .txt is pure text with no structure at all — useful for feeding into other software.

Are images from the PDF included?

No. Only the text layer is extracted. If you need the pages as pictures, use PDF to JPG.

Is my document uploaded?

No. PDF.js reads the text in your browser and the output file is assembled locally before downloading.

Related tools

\n\n\n