← all tools / pdf

PDF to Text

Live

Extract the text a PDF already stores and download it as a plain text file, with a marker at the start of every page.

What it does

  • Whole document, or just the pages you name
  • Page markers so you can tell where a line came from
  • Says so plainly when a PDF is a scan, instead of returning an empty file

Your files

Gone within the hour.

Files are uploaded over HTTPS, used only to run this tool, and deleted from our server automatically about an hour later. Never sold, never used to train anything.

Frequently asked questions

Why did it say there is no text to extract?

Because the pages are pictures rather than text, which is what a scan or a photographed document is. Run OCR PDF on it first to add a text layer, then extract the text from that result.

Does the .txt keep the layout?

No. A plain text file has no columns, tables, fonts or images. If you need the layout, convert to Word instead.

Can I extract only some pages?

Yes. Type ranges like 1-3, 7, 12- in the pages field. Leaving it blank extracts the whole document, up to 500 pages a job.

Will the reading order be right?

For ordinary documents, yes: the text is sorted into reading order rather than the order the file happens to store it in. Heavily designed pages with side panels can still interleave.

More PDF tools

all tools