TPTruPDF

@trupdf/ocr

OCR

In-browser OCR · 100+ languages · streamed page output

Detect and extract text from scanned PDFs and images. Runs entirely in the browser via TruPDF's WASM OCR engine; no server round-trips for sensitive content.

F.01 — What you get

  • Fully client-side — the OCR engine and language models load on demand.
  • Per-page progress events for long documents.
  • Output flows back into the annotation layer as searchable highlights.
  • Works offline once the language model is cached.

Next step

See it in the demo

The live demo wires this feature into a full viewer. Open any PDF and try it.