Skip to content

OCR a PDF, free, in your browser

Make a scanned PDF searchable: select, copy and find its words. The text is read on your device and laid invisibly over the scan, so every page looks exactly as before.

PDF · scanned or photographed · any size

✓ no upload✓ pages unchanged✓ 11 languages✓ works offline

How OCR works here

  1. Drop the scan. A PDF from a scanner, a phone app or a fax. Pick the language it is written in; a second one if the pages mix two.
  2. Read. Each page without text is drawn at 300 dpi and read by Tesseract 5, the open source OCR engine, running in your browser. The first time, the reader and the language are fetched from this site and kept.
  3. Download. Every word goes back onto the page, invisible, exactly where the scan shows it. The result card lists the pages read and how sure the reader was.

What it does

  • Reads the text on your device: a scanned contract, a passport or a medical letter never leaves it.
  • Keeps the page as it was, pixel for pixel; the file grows by the text alone.
  • Places each word over its picture, so selecting and highlighting line up with what you see.
  • Turns along with a page scanned sideways and rotated upright.
  • Says how sure it is on average, warns for each page under 60 %, and names the words it was least sure of.

What it does not do

  • OCR guesses: small print, poor scans and unusual fonts come out with mistakes, and the result does not hide that.
  • It does not read handwriting reliably.
  • It does not straighten a crooked scan or clean up the picture; it reads the page as it is.
  • It does not make the text editable in the page itself; PDF to Word makes an editable copy.

Questions

What does OCR do to my PDF?

It adds the text it reads as an invisible layer over each scanned page. The scan stays exactly as it was. Afterwards you can search the file, select words and copy them.

How accurate is it?

On a clean scan of printed text at 300 dpi, almost every word comes out right. Blurry photos, tiny print and handwriting do worse. The result gives the average confidence, names every page under 60 % and the words it was least sure of, so you know where to check.

Which languages can it read?

English, German, French, Spanish, Italian, Portuguese, Dutch, Polish, Czech, Croatian and Russian. Pick two at once for a document that mixes them; it is slower.

Why does the first run take longer?

Your browser fetches the reader, about 1.5 MB, and the language, 0.7 to 3 MB, from this site once and keeps them. After that, a full page takes a few seconds.

Does my file leave my computer?

No. The reading runs in your browser tab. The only network requests are the site's own files, which you can check by switching the connection off after the first run.

Guides for this tool