Diagram showing how to convert PDF to Word without losing formatting

You open the converted file, and the damage is immediate. Headings have become body text, a two-column layout has collapsed into one long ribbon, the table has turned into a pile of tab characters, and something has inserted a text box you cannot select. The instinct is to blame the converter — but in most cases the converter did the only thing it could.

The reason it is hard to convert PDF to Word cleanly is that the two formats disagree about what a document is. Word stores structure: this is a heading, this is a table, this paragraph flows after that one. PDF stores appearance: put this glyph at this coordinate, in this font, at this size. Converting one to the other means guessing the structure back from the positions, and how good that guess is depends almost entirely on what kind of PDF you started with.

This guide covers five ways to convert PDF to Word, which one to use for which file, the thirty-second test that tells you in advance whether formatting will survive, and how to clean up what is left.

Why formatting breaks when you convert PDF to Word

Three things account for nearly every broken conversion.

Diagram showing how to convert PDF to Word without losing formatting

Knowing which one you are hitting saves a lot of pointless retrying. Knowing which one you are hitting saves a lot of pointless retrying with different tools.

1. The PDF is a picture, not text

This is the big one, and it catches people constantly. A PDF made by scanning paper, or photographed on a phone, contains no text at all — only an image of text. There is nothing for a converter to read.

Feed that into any converter and you get one of two results: a Word file containing a single large picture, or garbled nonsense where optical character recognition has guessed badly. Neither is a formatting problem. It is a recognition problem, and it needs OCR, which is covered further down.

2. The fonts are not on your computer

PDFs usually embed their fonts, which is why they look identical everywhere. Word does not work that way — it needs the font installed to render it.

When the font is missing, Word substitutes something similar. The substitute has different letter widths, so lines rewrap, pages reflow, and a document that was four pages becomes five with headings stranded at the bottom of pages. The text is all correct; the layout has drifted.

3. The layout was never a layout

Designers building a PDF in InDesign or Illustrator place text in free-floating boxes. There are no Word-style columns, no real tables, sometimes not even paragraphs — just blocks of text positioned by eye.

A converter faced with that has to invent structure. It commonly wraps everything in text boxes or frames, which is why the result looks right at first and becomes unusable the moment you try to edit a sentence. Nothing reflows, because nothing is connected to anything else.

The thirty-second test: which kind of PDF do you have?

Do this before choosing a method. It takes less time than one failed conversion.

  1. Open the PDF in any reader.
  2. Try to select a sentence with your mouse, as if highlighting it.
  3. Press Ctrl+F and search for a word you can see on the page.
Which PDF have you got?
Text highlights, search finds it
Word, Google Docs or a PDF editor will all work
Text PDF
A box covers the page, search finds nothing
Run OCR first, then convert
Scanned image
Some pages work, some do not
OCR the scanned pages only
Mixed

If the text highlights and search finds it, you have a text-based PDF. Every method below will work, and it is only a question of how well the layout survives.

If your cursor draws a box over the whole page and search finds nothing, it is a scanned image. Skip straight to the OCR section; the other methods have nothing to work with.

If some pages behave and others do not, it is a mixed document — commonly a typed contract with scanned signature pages. You will need OCR for part of it.

Method 1: Microsoft Word itself

If you already have Word, try this first. It is free, it is already installed, and for straightforward documents it is genuinely good.

  1. Open Word.
  2. File → Open, then select the PDF.
  3. Word warns that it will convert the PDF into an editable document and that the result may not look exactly like the original. Click OK.
  4. Wait. A long or graphic-heavy PDF can take a minute or more.
  5. Save it as .docx straight away, so you are not working on an unsaved conversion.

Where it is strong: reports, letters, essays, anything that was written in Word and printed to PDF in the first place. Headings, lists and simple tables usually come through intact.

Where it struggles: multi-column magazine layouts, heavily designed brochures, and anything with text wrapped around images. It also has no OCR, so scanned PDFs produce a document containing one large picture.

One habit worth adopting: keep the PDF open side by side while you check the result. Errors in converted documents are usually obvious within the first two pages, and if the first two pages are a mess the rest will be too — stop and switch methods rather than fixing it page by page.

Method 2: Google Docs

Free, works on any machine with a browser, and useful when you do not have Word at all.

  1. Upload the PDF to Google Drive.
  2. Right-click it → Open with → Google Docs.
  3. Google converts it and opens an editable document.
  4. File → Download → Microsoft Word (.docx).

Google Docs has a real advantage here: it runs OCR automatically on scanned PDFs, so it will often pull readable text out of an image-only file that Word cannot touch at all.

The trade-off is layout. Google Docs is aggressive about simplifying — it tends to strip the design down to plain flowing text with the images dropped in. For a document you intend to rewrite anyway, that is often exactly what you want. For one where the layout matters, it is the wrong tool.

Method 3: A dedicated PDF editor

When the layout genuinely has to survive — a contract, a form, a designed report you have to edit and send back — a purpose-built PDF application will beat both free options. This is what they are engineered to do.

What you are paying for is better structure recognition: real Word tables instead of tab characters, proper columns instead of text boxes, headings mapped to heading styles, and OCR built in so scanned pages are handled in the same pass.

The two that matter

Adobe Acrobat is the reference implementation. Adobe created the PDF format, and its conversion is consistently the most faithful, particularly on documents Adobe’s own tools produced. It is also the most expensive option and is sold as a subscription.

Wondershare PDFelement is the main challenger and the one most people land on when Acrobat’s pricing does not fit. It converts, edits, OCRs and fills forms, and it is sold both as a subscription and as a perpetual licence — which matters if you convert a few documents a month rather than daily. Current Wondershare discounts are worth checking before you buy, because this is a category where the list price is rarely what people actually pay.

On the subscription-versus-perpetual question generally: if you will use the software for more than about two years, a perpetual licence almost always wins on total cost — the same arithmetic we walked through for buying a Windows licence. The catch is that perpetual usually means one major version, not free upgrades forever. Read what the licence actually covers before assuming.

How to use one properly

The step most people skip is the one that matters. Do not just hit Convert:

  1. Open the PDF in the editor.
  2. If it is scanned, run OCR first as a separate step, and pick the correct language. Running OCR and conversion blind in one click is where accuracy is lost.
  3. Check the OCR result on screen before converting. Fix obvious misreads now — it is far quicker than fixing them in Word later.
  4. Then convert to Word, choosing the “retain layout” or “flowing text” option depending on whether you need the design or the words.

That last choice is worth understanding. Retain layout reproduces the appearance using text boxes — it looks right and edits badly. Flowing text rebuilds it as a normal Word document — it edits well and looks approximate. Pick based on what you are going to do with the file, not on which preview looks nicer.

Method 4: Online converters

Free browser converters are everywhere, they need no installation, and for a one-off page of nothing sensitive they are perfectly reasonable.

Two things to be clear-eyed about.

Your file goes to someone else’s server. Never upload contracts, medical records, financial statements, anything with ID numbers, or unpublished work to a free converter. You do not know the company, you have not read their retention policy, and there is no way to verify deletion. For anything confidential, use a tool that runs on your own machine — including the free Word and Google Docs options above, since at least you know whose account it is sitting in.

Quality varies wildly. Some free sites are excellent. Others are a thin wrapper around an old open-source library with a queue and an advert. You cannot tell which from the homepage, so test with a document you do not care about first.

Watch for the usual funnel too: free for the first two pages, then a payment wall once you are invested. If you are going to pay anyway, a proper application costs about the same and keeps your documents on your own computer.

Method 5: When the PDF is scanned — OCR

Optical character recognition reads the shapes in an image and works out which letters they are. Modern OCR is very good, and it is still guessing, so the input quality decides the output quality.

What actually improves accuracy, in order of impact:

  • Scan at 300 DPI or higher. Below about 200 DPI, accuracy falls off a cliff. This single setting matters more than which OCR engine you use.
  • Straighten the page. A few degrees of skew from a phone photo costs real accuracy. Most tools have a deskew or auto-rotate option — use it.
  • Set the correct language, and set it before running, not after. An engine expecting English will mangle accented characters it was not told to expect.
  • Use black and white or greyscale, not colour, for plain documents. Less noise to interpret.
  • Expect to proofread. Even excellent OCR confuses rn with m, 0 with O, and 1 with l. Run a spell check, then read the numbers by eye — spell check will not catch a wrong digit.

If the scan is genuinely poor — faded, skewed, photographed at an angle in bad light — rescanning takes five minutes and saves an hour of correcting. No software fixes a bad original.

Which method for which document

Your document Use this Why
Text PDF, simple layout Microsoft Word Free, installed, good enough
Text PDF, you only need the words Google Docs Strips design, keeps text
Contract or form to edit and return PDF editor Keeps tables and fields intact
Multi-column or designed layout PDF editor Only tool that maps columns properly
Scanned document OCR, then convert Nothing else can read it
One page, nothing confidential Online converter Fastest, no install
Anything confidential Offline tool only The file never leaves your machine

Cleaning up the converted file in five minutes

Almost no conversion is perfect. These five passes fix the overwhelming majority of what is left, and they are quicker in this order.

Turn on formatting marks first. Ctrl+Shift+8 shows paragraph marks, spaces and tabs. Most conversion mess is invisible until you can see it — rows of tab characters pretending to be a table, or a paragraph break after every single line.

  1. Kill the manual line breaks. Converters often end every visual line with a break. Use Find and Replace with ^l (manual line break) replaced by a space, then fix the genuine paragraph ends. This alone repairs most “why will this text not wrap” problems.
  2. Clear the direct formatting. Select all, then Ctrl+Space clears character formatting and Ctrl+Q resets paragraph formatting. Brutal, but it removes the invisible inherited styling that makes converted documents behave oddly.
  3. Reapply real heading styles. Converted headings are usually just large bold text. Apply Heading 1 and Heading 2 properly — the navigation pane and table of contents depend on it.
  4. Rebuild the tables. If a table came through as tabbed text, select it and use Insert → Table → Convert Text to Table. Faster and cleaner than nudging cells.
  5. Fix images last. Set each picture’s wrap option to something sensible — converted images default to positions that jump around when you edit the text above them.

Work top to bottom and do not fix the same thing twice. The temptation is to correct a heading, then a table, then go back — each pass over the whole document is faster than jumping around.

Converting a folder full of PDFs, not one

Everything above assumes a single document. Once you are converting twenty invoices or a year of statements, clicking through a dialog twenty times is the wrong approach, and the free tools stop being free of effort.

Batch conversion is the main practical reason people buy a PDF editor rather than using Word. Point it at a folder, set the output format once, and let it work through the queue. A job that would take an afternoon by hand finishes while you do something else.

Three things to check before you start a large batch:

  • Test on three files first, chosen to be different from each other — one simple, one with a table, one scanned if you have them. If the settings are wrong, you find out after three conversions instead of two hundred.
  • Set OCR language once, at the batch level. Most tools apply it to the whole queue; getting it wrong halfway through means running the lot again.
  • Output to a new folder. Never let a batch job write next to the originals. If something goes wrong you want the source files untouched and obviously separate.

If this is a one-off — a single archive you need to convert PDF to Word and never repeat — a trial version usually covers it honestly. If it is a monthly task, buy the licence; the time saved pays for it within two or three batches.

Going back the other way, properly

Most people converting in this direction will eventually need to go back — edit in Word, then send a PDF. That trip damages documents too, just more quietly.

Use File → Save As → PDF, or Export. Do not print to a PDF printer driver unless you have a specific reason to. Save As keeps the document’s internal structure: headings stay headings, links stay clickable, and the text stays selectable and searchable. Printing to PDF flattens all of that into appearance only — which is precisely the kind of PDF that is painful to convert back later.

If the document will be read on screen, tick the accessibility or “document structure tags” option in the export dialog. It costs nothing, makes the PDF work with screen readers, and — usefully for you — makes it far easier for any future converter to reconstruct the layout.

When nothing preserves the layout

Sometimes the honest answer is that no converter will produce what you want, because the original was never structured in a way Word can express. A magazine spread with text flowing around cut-out images has no Word equivalent that is also editable.

At that point you have two sensible options, and one bad one.

Edit the PDF directly instead. If all you need is to change a date, a name or a price, do not convert at all. A PDF editor changes the text in place, in the original layout, in seconds. Converting a document you only need to tweak is a self-inflicted wound.

Rebuild it in Word on purpose. Convert with the “flowing text” option to recover the words without the broken layout, then apply your own styles. For a document you will maintain long-term, forty minutes of deliberate rebuilding beats months of fighting a converted file that was never really a Word document.

The bad option is accepting a text-box conversion for a document you have to keep editing. It looks finished and it is not — every future edit will break something, and you will pay that cost repeatedly.

Frequently asked questions

Can I convert PDF to Word for free?

Yes. Microsoft Word opens PDFs directly, and Google Docs converts them and runs OCR on scanned files. Both are free if you already have the account or the software. Paid tools buy better layout fidelity and better OCR, not the basic ability.

Why does my converted Word file look nothing like the PDF?

Usually a missing font, or a PDF built from free-floating text boxes rather than real paragraphs. Check whether the text is even selectable in the PDF first — if it is not, you are looking at a scan, and that is a different problem.

How do I convert a scanned PDF to editable Word?

Run OCR first, then convert. Google Docs does both automatically and is free. A dedicated PDF editor gives you control over language, page orientation and correction before the conversion, which matters on anything longer than a few pages.

Are free online PDF converters safe?

For a public document, generally yes. For anything confidential, no — the file is uploaded to a third-party server you know nothing about. Use an offline tool for contracts, financial documents or anything containing personal data.

Will tables survive the conversion?

Only if they were real tables in the PDF. Many PDF “tables” are just text positioned in a grid, in which case even a good converter produces tabbed text. Word’s Convert Text to Table turns that back into a proper table in about ten seconds.

Is it better to convert PDF to Word or just edit the PDF?

If you need to restructure the document, convert it. If you only need to change a few words, numbers or dates, edit the PDF directly — you keep the exact layout and skip the cleanup entirely.

The short version

Before anything else, try to select the text in your PDF. That one action tells you whether you need OCR or not, and it is the difference between a two-minute job and an hour of frustration.

If it is text and the layout is simple, open it in Word and you are finished. If you only want the words, Google Docs will hand them over. If the layout has to survive — contracts, forms, designed documents — that is what a real PDF editor is for, and it is the one case where paying is clearly worth it.

And if you only need to change a line or two, do not convert PDF to Word at all. Edit the PDF where it stands and keep the layout the original author intended.

Further reading: Microsoft explains the built-in behaviour in its guide to editing PDFs in Word.