About OCR PDF
A scan is a photograph of a document. To a computer it is pixels: you cannot search it, copy from it, or have a screen reader read it aloud. Optical character recognition looks at those pixels, identifies the characters, and attaches a text layer to the page.
The result looks identical but behaves like a real document — searchable with Ctrl+F, selectable, and usable by assistive technology.
How to use OCR PDF
- Step 1
Upload the scanned PDF
Add a document made of page images.
- Step 2
Recognise the text
Characters are detected and a text layer is written behind the image.
- Step 3
Download
Save the searchable PDF.
When it helps
- • Making an archive of scanned invoices searchable by reference number.
- • Copying a quotation out of a scanned book page.
- • Meeting accessibility requirements for published documents.
- • Preparing scans so PDF to Word or Extract Text can work on them.
Good to know
- • Accuracy depends on the scan. Straight, well-lit pages at 300 DPI or better give the best results.
- • Handwriting, decorative fonts, faint carbon copies and heavy skew all reduce accuracy.
- • Always proofread OCR output before relying on it for numbers or legal text.