← All PDF tools
OCR PDF
OCR a PDF
Extract text from scanned pages with Tesseract running in your browser. Download a text report of what was found.
Runs in your browser. Your PDF stays on this device — nothing is uploaded for processing.
On-device OCR
Tesseract.js runs locally after model load.
Page by page
See progress as each page is recognized.
Text export
Download extracted text as a .txt file.
Private scans
Images stay in your tab during recognition.
Use the tool
How it works
Searchable text without a cloud OCR API
Pages render to canvas, then Tesseract recognizes text entirely in the browser (first run downloads the model).
Before you start
- First OCR run may download a language model.
- Clearer scans produce better results.
Three steps
- 1
Load a scan
Choose a scanned or image-heavy PDF.
- 2
Run OCR
Wait while each page is recognized locally.
- 3
Download text
Save the extracted text file.