PDF OCR
A scanned contract, a photographed page, an old fax saved as PDF — all of them look like documents and behave like pictures: you cannot search them, cannot copy a line, and a screen reader has nothing to read aloud. This page recognises the words and puts them back where they are on the page, invisibly. The scan looks the same; everything else starts working.
Recognition is a guess, not a reading: a clean printed page comes out almost perfect, a phone photo of a crumpled receipt does not. Nothing is auto-corrected — a plausible fix to a number is a mistake you would never catch.
The scan is not uploaded: the engine and the language model are downloaded into your browser and work there. What comes out is your own scan with an invisible text layer on top — the picture looks the same, but search, copying and screen readers work.
Nearby: Image to Text · PDF to Text · PDF to Word · Compress PDF
- Free
- Files never leave your device
- No sign-up
- No file limit
How to use
- Choose the scanned PDF.
- Pick the language of the document — recognition depends on it more than on anything else.
- Press the button, wait, and download the searchable PDF or the plain text.
Good to know
What a searchable PDF actually is
It is your scan with a second layer: the recognised words, placed exactly over the printed ones and drawn with invisible ink. The page looks identical — same paper, same stamps, same handwriting in the margin — but Ctrl+F finds a name in it, a phrase can be copied, and a screen reader can read it aloud. That is why the picture is kept rather than replaced by clean text: a scan is often a document of record, and re-typesetting it would make it something else.
Recognition is a guess, not a reading
A clean printed page comes out nearly perfect. A photo of a crumpled receipt does not, and no engine changes that. So the average confidence is shown after the run, and nothing is auto-corrected: a plausible correction to a digit is the kind of error nobody ever catches. Read the text before you rely on it — especially the numbers.
Frequently asked questions
Is my scan uploaded?
No. The engine and the language model are downloaded into your browser and run there; the file stays on the device.
How much is downloaded?
Around two to three megabytes: the engine plus one language model. It is stated on the page before you press, and after the first run everything works offline.
My PDF already has text — should I run this?
No, and the page says so when it notices: recognition would replace good text with a guess. Use «PDF to text» instead.
How many pages can it handle?
Up to fifty. Beyond that a browser would spend tens of minutes, and the page says how many were done.
Related tools
- Image to Text Drop in a picture, pick the language of the writing, and press the button. A scan or a screensh…
- PDF to Text Copying from a PDF gives you words cut in half, lines broken at the column edge and a company n…
- PDF to Word Drop in a PDF and get a .docx you can edit. We are straight with you about what moves and what…
- Compress PDF Two honest ways to shrink a PDF: rebuild the file and keep the text intact, or re-render the pa…
- PDF to JPG Each page becomes its own image, ready to download one by one — usually you need a page or two,…
- Merge PDF Add the files, put them in the order you need and get one document. The pages are copied as the…
- JPG to PDF Drop your photos in, put them in the right order and get a single PDF. Ordinary JPEG files are…
- Image to PDF Drop in your pictures and get one PDF back. Photos of documents, scans, screenshots, pages shot…