PDF OCR

Drop a scanned PDF, a photo of a document saved as a PDF, or anything with no real text layer, and this tool reads the text visually, page by page, giving you plain text you can copy, search or download.

  • Free
  • No sign-up
  • Runs in your browser
pdf-ocr

100% private — the PDF is processed in your browser and never uploaded.

How it works

Drop your scanned PDF

A document with no selectable text.

Wait while it reads each page

A progress note shows the page and percentage.

Copy or download the text

Plain text pulled from the images.

When a PDF is really just a picture of text

A scanned contract, a photographed receipt saved as a PDF, an old document digitized years ago: these all look like text on screen, but to a computer they are just images, with no text layer to select, search or copy. Optical character recognition, OCR, reads the shapes of the letters visually and turns them back into real text, the same job a document scanner app does on a phone.

How it works

Each page of the PDF is rendered as an image at a resolution good for reading small text, and a text recognition engine, running entirely on your device, reads the letters and words from that image. The recognized text from every page is combined into one block you can copy, search or download.

The first time takes a little longer

The recognition engine and its language data, a few megabytes, are downloaded the first time you use this tool or the image-to-text tool, and kept by your browser after that. This means the very first page takes longer than the rest while that download finishes; pages after that are quicker.

Getting a good result

OCR accuracy depends heavily on the quality of the scan: a clear, high-resolution, well-lit scan with straight, unrotated text reads far better than a blurry photo taken at an angle. If the source document is tilted, the deskew tool can help before running OCR. Handwriting is read far less reliably than printed text, and some errors should be expected on any source, so always check the result rather than trusting it blindly for anything important.

What it does not do

This gives you the text, not a searchable PDF with the original images and an invisible text layer added back in, which is a more complex file format to build than a browser tool can reasonably produce. It reads what is visible and returns text only. Everything runs in your browser: the images from your PDF are never uploaded anywhere.

Frequently asked questions

How do I get text out of a scanned PDF?
Drop it here. Each page is read visually with on-device OCR, and the recognized text is combined for you to copy or download.
Why does the first page take so long?
The text recognition engine downloads the first time you use it, a few megabytes. After that, pages are read more quickly.
How accurate is the result?
It depends on the scan quality. Clear, straight, high-resolution scans read very well; blurry or tilted photos produce more mistakes. Always check the result.
Does it work on handwriting?
Not reliably. This works best on printed text; handwriting recognition is much less accurate.
Do I get a searchable PDF back, or just text?
Just the text. This does not rebuild a PDF with a hidden text layer added to the original images.
Is my PDF uploaded anywhere?
No. Every page is read entirely on your device.