Scanned PDF OCR
Recognize and pull text out of scanned/image PDFs that have no text layer · 100% in your browser
1. Choose a PDF
2. Language
Recognition language
Language data downloads once on first use and is cached in your browser. Files are never sent to a server.
3. Run
Recognize and pull text out of scanned/image PDFs that have no text layer. Each page is rasterized and recognized with tesseract.js; files are never sent to a server.
What you can do
- ✓Extract text from scanned or photo PDFs
- ✓Korean+English, English, Japanese, Chinese recognition
- ✓View the result on screen, copy it, or save as TXT
- ✓Batch multiple pages (first 30)
- ✓100% in-browser — no upload
How to use
- 1Add a scanned PDF
- 2Pick the recognition language and click Recognize text
- 3Copy the result or download it as TXT
Supported formats · PDF (input) → text / TXT
Related tools