Language selection
Choose only languages actually present. A large mixed-language set can slow recognition and increase ambiguous character choices.
Recognize text in scanned pages and add a searchable layer while processing the document locally.
OCR turns page images into selectable and searchable text. It is helpful for scanned letters, invoices and archives, but it is an estimate rather than a perfect transcription.
Accuracy depends on language, resolution, contrast, skew and typography. Names, account numbers and legal clauses always deserve a manual check against the visible scan.
Add the scanned PDF and choose the document language.
Select a model and resolution; start around 150–200 DPI.
Run OCR, then test search and copy several difficult passages in the result.
Choose only languages actually present. A large mixed-language set can slow recognition and increase ambiguous character choices.
Higher DPI can help small print but uses more memory and time. It cannot restore detail missing from the original scan.
OCR commonly confuses O/0, I/1, punctuation and accented characters. Never rely on it alone for financial or legal data entry.
Handwriting, decorative fonts, tables, multi-column layouts and poor scans may produce incomplete or misplaced text. OCR models are downloaded by the browser when needed.
The tool keeps the page appearance and adds an invisible text layer to the exported PDF.
Each page is rendered and analyzed on your device. Resolution, page count and available memory all affect speed.
Treat OCR as a draft. Compare important numbers, names and dates with the original image.