Recognize text in a small scanned PDF
PDF OCR renders each selected page and recognizes printed English or Arabic text with a local Tesseract.js worker. It is intentionally limited to small documents to remain reliable in mobile browsers.
How to use PDF OCR
- Choose a PDF within the displayed size and page limits.
- Select English, Arabic or both.
- Run OCR and download the page-labelled TXT file.
Get better OCR results
Use clear, upright scans with strong contrast. OCR is slower than normal text extraction; try Extract Text from PDF first when the document already contains selectable text.
Frequently asked questions
Why is PDF OCR limited to a few pages?
Rendering and recognition require significant memory, so the limit protects phones and shared browser tabs.
Does PDF OCR support handwriting?
It is designed for printed text and is not reliable for handwriting.
Are pages sent to an OCR service?
No. PDF.js and Tesseract.js run on your device using files hosted by OmniTools.