All / PDFs

Extract text from a PDF

Pull the digital text layer out of a PDF. Scanned pages are reported as empty.

This is not OCR. A scan of paper will not become words. At most 40 pages are read.

This copies the text that is already in the PDF. A scan of paper has no text layer, so those pages are reported as empty. It is not optical character recognition. At most 40 pages are read.

  1. 1Drop a digital PDF.
  2. 2Read the extract.
  3. 3Copy it or download a .txt file.

Digital text only. A scanned page is reported as empty. This is not OCR and not a Word file.

Rate this tool

Questions

Why is a page blank?
That page has no text objects. It is probably a scan or a drawing.
Will it make a Word file?
No. You get plain text. Layout and tables are not reconstructed.
Nothing is uploaded. Closing the tab drops the file.
Text is read in the tab.

Related