← All PDF tools

OCR PDF

OCR a PDF

Extract text from scanned pages with Tesseract running in your browser. Download a text report of what was found.

Runs in your browser. Your PDF stays on this device — nothing is uploaded for processing.

On-device OCR

Tesseract.js runs locally after model load.

Page by page

See progress as each page is recognized.

Text export

Download extracted text as a .txt file.

Private scans

Images stay in your tab during recognition.

Use the tool

Searchable text without a cloud OCR API

Pages render to canvas, then Tesseract recognizes text entirely in the browser (first run downloads the model).

Before you start

  • First OCR run may download a language model.
  • Clearer scans produce better results.

Three steps

  1. 1

    Load a scan

    Choose a scanned or image-heavy PDF.

  2. 2

    Run OCR

    Wait while each page is recognized locally.

  3. 3

    Download text

    Save the extracted text file.