How to Extract Text From a PDF, Even a Scanned One
Need the words out of a PDF? Learn the difference between text and scanned PDFs, four ways to copy text, and what to do when it will not select.
Read article →Copy all the text out of a PDF in one go and save it as a plain text file you can edit anywhere.
Useful for quoting, translating, searching, or pasting a document into an editor. The text file is saved in UTF-8, so accents and non-Latin scripts display correctly.
This reads the text that is already inside the PDF. A scanned document is only a picture of text, so it contains nothing to extract and would need OCR, which this tool does not do. Layout such as tables and columns is not preserved in plain text.
Opens in any editor on any device.
Accents and non-Latin scripts display properly.
Optionally mark where each page starts.
Tells you when a file has no selectable text.
The PDF is probably a scan or made of images. Those need OCR software to recognise the letters.
Text is extracted line by line. Plain text cannot keep table layout.
Yes, as long as the PDF stores real text. The output is UTF-8.
Need the words out of a PDF? Learn the difference between text and scanned PDFs, four ways to copy text, and what to do when it will not select.
Read article →
Not sure which DPI or color mode to pick? These scan settings keep text sharp and PDF files small, for contracts, receipts, books and photos.
Read article →It is free, private and takes seconds. Your files never leave your device.