Plain text

Extract text from a PDF

We read the text layer, not pixels. A photographed invoice with no fonts will come back empty — use OCR elsewhere for that.

Scanned pages with no text layer will come back empty. That is not OCR.

How to use it

  1. Drop the PDF

    Born-digital files (exports, invoices with fonts) work.

  2. Optional range

    Leave blank for every page.

  3. Copy or download

    The text appears on the page. Download a .txt for your editor.

FAQ

Is this OCR?

No. OCR is a different product. We only read text that is already in the PDF.

Why is the order weird?

We sort runs top-to-bottom, then left-to-right. Multi-column layouts can interleave.