@trupdf/ocr
OCR
In-browser OCR · 100+ languages · streamed page output
Detect and extract text from scanned PDFs and images. Runs entirely in the browser via TruPDF's WASM OCR engine; no server round-trips for sensitive content.
F.01 — What you get
- Fully client-side — the OCR engine and language models load on demand.
- Per-page progress events for long documents.
- Output flows back into the annotation layer as searchable highlights.
- Works offline once the language model is cached.
Next step
See it in the demo
The live demo wires this feature into a full viewer. Open any PDF and try it.