Use PDF to Text - free
Frequently asked questions
What about scanned PDFs with no text layer?
Scanned PDFs contain images of text, not real characters. The PDF to Text tool will produce an empty or near-empty file. Use the OCR tool first to add a text layer.
Is the text extracted in reading order?
PDF.js extracts text in the order it appears in the content stream, which usually matches reading order for well-structured PDFs. Complex multi-column layouts may have text in a non-intuitive sequence.
Are headers, footers, and page numbers included?
Yes - all text on every page is included, including headers, footers, and page numbers. You can clean these up manually in a text editor after extraction.