How to recognize text in a scanned PDF using OCR
In short: To recognize text in a scanned PDF, run it through ocr-pdf and the document becomes selectable and searchable. Recognition quality depends on the scan, so use a clear, straight, well-lit page. Standard OCR handles printed text well but not handwriting.
Cluster
how-to guide
Step-by-step instructions for getting a PDF task from input to a reliable result.
Primary tool
OCR PDF online
Open the tool from this article and complete the operation in the current locale.
Open toolTable of contents
You receive a PDF with a scanned contract, certificate, or form. You try to select some text and nothing happens — the cursor behaves as if the page is a picture. That is exactly what it is: a scanned image with no text layer underneath. OCR (optical character recognition) adds that layer and turns the image into a working document where text can be selected, copied, and searched.
When you need OCR
OCR is useful when you have a scanned PDF and want to:
- copy a passage without retyping it by hand;
- search for a word in the document with Ctrl+F;
- extract content into another tool without manual entry;
- make an archived document searchable by keyword.
If the PDF already has a text layer — meaning you can select words — OCR is not needed.
What affects recognition quality
The quality of the output is almost entirely determined by the source scan. The key factors:
- Resolution: good OCR needs at least 150–200 dpi, ideally 300 dpi. Phone photos are often sufficient if the page is flat and the lighting is even.
- Alignment: the page should be straight. Heavily tilted text recognizes poorly.
- Lighting and contrast: shadows from the spine, yellowed paper, or faded ink all reduce accuracy.
- Language: the algorithm needs to know which language the text is in to correctly interpret letters and words.
How to run OCR, step by step
1. Open ocr-pdf and upload your scanned PDF. 2. Select the document language if the tool prompts for it. 3. Wait for processing — multi-page scans take longer. 4. Download the result. 5. Open the file and verify: can you select text? Are key details — names, dates, numbers — recognized correctly?
What can go wrong
- Many errors in names and numbers. The scan quality is too low. Try rephotographing or rescanning at a higher resolution.
- Text is not recognized at all. The PDF may contain vector graphics that look like a scan but technically are not. OCR cannot help with that.
- The file got much bigger. OCR adds a text layer and slightly increases size. If the result is too heavy to send, compress it with compress-pdf.
- Handwriting is not readable. Standard OCR targets printed text. Handwritten signatures and margin notes recognize poorly or not at all.
What to check after OCR
- Select a few lines and paste them into a text editor — confirm the text came through correctly.
- Search for key words with Ctrl+F: a name, an organization, a date.
- Compare numbers and dates in the recognized text against the original — those are the most common source of errors.
To make a scanned PDF searchable and selectable, upload it to ocr-pdf. For long documents, split the file with split-pdf first, process the parts separately, and combine the results with merge-pdf.
FAQ
More from this cluster
how-to guide
How to convert PDF to editable Word
Turn a PDF into a Word document you can actually edit: keep tables and formatting intact, deal with scans and columns, and check the result before sending.
how-to guide
How to Extract Tables From PDF to Excel
How to pull tables out of a PDF into Excel: a step-by-step walkthrough, what to do with scans, shifted columns and merged cells, and how to check the result before you trust the numbers.
how-to guide
How to Extract PDF Pages to a New File
Pick specific pages from a large PDF and save them as a separate document: online, no software to install, no loss of quality.
Related tools
OCR PDF online
OCR PDF keeps the PDF task in one browser flow: upload the source file, check options, run processing, and download the result.
PDF to text online
Extract text from a PDF into a plain text file to copy the content without formatting and layout.
Scan to PDF
Turn photos and scans of documents into one tidy PDF, ready to send or print.
Compress PDF online
Reduce the size of a PDF so it is easier to email, upload or store. Especially useful for scans and documents with images.
What to do next
All tools
PDF tools catalog: merge, compress, split, convert, rotate, protect and unlock PDF files online, all directly in your browser.
FAQ
Answers to common questions about iHatePDF: whether registration is required, how files are processed, where to check limits, and whether it's safe to upload documents.
Contact
Contact iHatePDF about processing errors, choosing a tool, security, business inquiries, and suggestions for new features.