PDF OCR: Scanned PDF to Text
Recognise text on scanned PDF pages locally in your browser; your PDF is not uploaded.
Your files stay in this browser. No uploads, accounts or saved document data.
Up to 50 queue items, including up to 50 PDF pages. Images above 40 megapixels are reduced before OCR. GIF: first frame only.
Printed text works best. Handwriting, blurred or very small text, tables and complex layouts may be read poorly or in a different order. Language files download on first use (one to a few MB each); your browser may cache them.
Most selected pages in a PDF already contain text. You can continue with OCR or extract the existing text directly: Open PDF to text
Review and save
Confidence is the engine's own estimate, not a measure of correctness. Compare important details with the original.
Copy all and Download TXT use this edited text. Editing an item, adding or removing items, or recognising again rebuilds this field from the item texts.
Manual text corrections are included in TXT only. The searchable PDF keeps the engine's original text layer; compare it with the page images.
How to run OCR on a PDF
- For scanned PDF to text, add your document and enter its password if it is protected.
- Enter page numbers such as 1–3, 5, or leave the range empty to select every page.
- Check the text-layer notice. If most selected pages already contain text, follow its PDF to text link for extraction without OCR.
- Choose the document language, an optional second language and a suitable layout. Adjust a page preview if its orientation or crop needs attention.
- Recognise the selected pages, correct the editable text and download TXT. Leave searchable PDF enabled if you also need that file.
When PDF OCR helps
- Make the words in a scanned contract, letter or archive page available for copying.
- Run recognition only on relevant pages instead of processing the entire document.
- Inspect each page beside its result and correct names or numbers before saving.
- Use a searchable PDF when you want to find words while retaining page images.
- Change the crop or reading layout and recognise a difficult page again.
PDF OCR questions
Does a PDF with selectable text need OCR?
Usually no. The tool checks selected pages for a text layer and points you to PDF to text when most already have one. You may still continue with OCR.
What affects recognition on scanned pages?
Sharp printed pages tend to work better than handwriting, blurred scans or complicated tables. A scan around 300 DPI is a useful target; enlarging a poor scan cannot restore missing detail. Confidence is only the engine's estimate.
How many pages can I select?
You can process up to 50 PDF pages in one run, with no more than 50 queue items overall. Large documents may also be constrained by your device's memory.
Can the PDF contain more than one language?
Choose from 20 language models. One document language is required and a second is optional. Select the languages used on the pages because detection is not automatic.
Is the PDF uploaded for recognition?
The PDF is processed locally in your browser and is not uploaded. Language files download the first time they are needed and may remain cached. This site displays ads.
What does the searchable PDF contain?
It combines the recognised text with the original page images, which can make the download large. You can also copy corrected text or save it as TXT.