Extract text from a PDF
Pull the digital text layer out of a PDF. Scanned pages are reported as empty.
This is not OCR. A scan of paper will not become words. At most 40 pages are read.
This copies the text that is already in the PDF. A scan of paper has no text layer, so those pages are reported as empty. It is not optical character recognition. At most 40 pages are read.
- 1Drop a digital PDF.
- 2Read the extract.
- 3Copy it or download a .txt file.
Digital text only. A scanned page is reported as empty. This is not OCR and not a Word file.
Rate this tool
Questions
- Why is a page blank?
- That page has no text objects. It is probably a scan or a drawing.
- Will it make a Word file?
- No. You get plain text. Layout and tables are not reconstructed.
- Nothing is uploaded. Closing the tab drops the file.
- Text is read in the tab.