Recognise first, or extract now? One check tells you which tool

OCR creates text in a scanned PDF; Extract Text pulls existing text out as TXT or Markdown. The select-all check and when to run both.

UnboundPDF is a free suite of 54 PDF and image tools that run entirely in your browser — merge, split, compress, edit text, OCR in 126 languages, redact, sign, convert and archive to PDF/A. Your document is read and written by the page on your own device; there is no document-upload endpoint in the core tools, no account, no watermark and no daily cap. Every result can be checked — with the Network tab, or with the Document Passport the Workspace writes for a chain of steps.

Diagram comparing in-browser local processing, where a document never leaves the device, to an upload-based tool that sends the document to a server and back

The check. Open the PDF in any reader and try to select a word. Nothing selects — it is a scan: run OCR PDF first. Text selects — run Extract Text now.

OCR PDFExtract Text
InputScanned pages (pictures)PDFs that already contain text
OutputThe same PDF with a searchable text layerA .txt or .md file
AccuracyDepends on scan quality; per-page confidence shownExact — the document's own text
ThenExtract, convert to Word/Excel, searchPaste, search, feed a script
Diagram comparing in-browser local processing, where a document never leaves the device, to an upload-based tool that sends the document to a server and back
Everything described above runs the same way: the document processed in the browser tab, never sent to a server.

Running both

Scan → OCR → Extract Text is the normal path for paper documents: recognise on your device in one of 126 languages, then export the text in reading order with headings where the layout supports them. Check the confidence before trusting a number.

Extract Text with a document processed into clean text output in reading order
A page extracted to clean text — reading order kept, ready to quote or feed to a script.

What OCR does not change

The page still looks like the scan; the text layer is invisible underneath. Nothing is re-encoded.

Word or Excel instead of text

For an editable document, PDF to Word; for a table with typed cells, PDF to Excel — both after OCR on a scan.

Frequently asked questions

Can Extract Text read a scan?

No — a scan has no text until OCR writes the layer.

Does OCR replace my scan with typed text?

No. The scan stays as the picture; text is added underneath for search and copy.

Which is uploaded?

Neither — both run in your browser; OCR downloads its engine once as program code.

Try it yourself

Free, private, no account. Runs entirely in your browser.

Related articles