How it works
- 1Choose a scanned or image-based PDF.
- 2Run OCR; existing text is preserved and scanned pages are recognized locally.
- 3Download searchable PDF, plain text or positioned JSON.
OCR scanned PDFs locally in your browser with Tesseract.js WASM and download searchable PDF, TXT or positioned JSON results.
The OCR engine and automatic mixed-script models load from Toolardo; the PDF is processed only on this device.
No file selected.
OCR PDF reads existing text directly and uses local Tesseract.js WASM OCR for scanned pages. The original PDF stays on your device.
No. PDF rendering and OCR run locally in your browser.
The browser must download and cache the OCR engine and the required language models.