OCR a PDF
Make scanned PDFs searchable
Extract the text a PDF already stores and download it as a plain text file, with a marker at the start of every page.
Gone within the hour.
Files are uploaded over HTTPS, used only to run this tool, and deleted from our server automatically about an hour later. Never sold, never used to train anything.
Because the pages are pictures rather than text, which is what a scan or a photographed document is. Run OCR PDF on it first to add a text layer, then extract the text from that result.
No. A plain text file has no columns, tables, fonts or images. If you need the layout, convert to Word instead.
Yes. Type ranges like 1-3, 7, 12- in the pages field. Leaving it blank extracts the whole document, up to 500 pages a job.
For ordinary documents, yes: the text is sorted into reading order rather than the order the file happens to store it in. Heavily designed pages with side panels can still interleave.