Run English-first OCR on a screenshot or scan in this tab. The Tesseract model downloads once. This is not a passport reader, not a 99% accuracy claim, and not the right tool when the PDF already has a text layer.
You need the words off a screenshot or English scan and you do not want that image on an OCR SaaS. That is this page: Tesseract in the tab, English (eng) first.
If the PDF already lets you highlight text, skip OCR and use extract PDF text. If you only needed a picture of the page, use PDF to JPG.
What OCR can and cannot read
Works often: printed English, clear screenshots, high-contrast scans.
Fails often: handwriting, low-res phone photos of glossy paper, stamps, dense tables, and languages we did not load.
This is not Google Cloud Vision and not a structured ID or invoice parser. We will not lead with passport MRZ or “证件 OCR.”
How to use it
- Drop a still (screenshot, JPEG, PNG). Typical ceiling about 32 MB.
- Wait for the English model on first use, then recognize in this tab.
- Copy the text. Clean spacing with whitespace cleaner if needed.
Honest limits
- Accuracy is “good enough to start typing,” not a court transcript.
- No searchable-PDF export.
- HEIC may need HEIC to JPG first.
Local is why you would use a browser OCR — not a claim we are the only private tool on the internet.