Copy text from any PDF

Digital PDFs are read instantly. Scanned pages are detected and OCRed automatically. Repeated headers, footers and page numbers are stripped for you.

Drop your PDF here

Digital PDFs are read instantly from the text layer. Scanned pages are detected automatically and run through OCR — all locally.

…or press Ctrl+V / +V to paste

100% local — files never leave your browser

How it works

1

Drop your PDF

Any PDF works — reports, papers, invoices, ebooks, scans.

2

Smart extraction

Pages with a text layer are read directly. Scanned pages are rendered and OCRed, page by page, in your browser.

3

Copy clean text

Repeating headers, footers and page numbers are removed automatically. Copy or download the result.

Why this one?

Handles scanned PDFs

Each page is checked for a usable text layer; scans fall back to local OCR automatically.

Header & footer removal

Lines repeated across pages — titles, page numbers — are detected and stripped from the combined text.

Nothing is uploaded

Contracts and invoices are sensitive. Your PDF is parsed entirely on your device.

Per-page progress

Long documents show live page-by-page progress, so you always know what is happening.

Frequently asked questions

Does it work with scanned PDFs?

Yes. Each page is checked automatically: if there is no usable text layer, the page is rendered and run through OCR locally. Expect a few seconds per scanned page.

Is there a page or file-size limit?

There is no hard limit, but everything runs in your browser, so very large files depend on your device. Documents in the hundreds of pages generally work fine; scanned pages take longer than digital ones.

Will my PDF be uploaded or stored?

No. The PDF is parsed by your browser using pdf.js and local OCR. It never leaves your device.

Why does the text sometimes have odd line breaks?

PDFs store text as positioned fragments, so line breaks can be artifacts of layout. Toggle "Clean up text" to join broken lines and remove stray hyphenation.

Have a different file?