Native text stays native
Page evidence decides whether OCR is required. Reliable digital text is not degraded or duplicated.
Local Tesseract · Private temporary jobs
Add an invisible searchable text layer while preserving page appearance, or download plain text and structured OCR data.
Page evidence decides whether OCR is required. Reliable digital text is not degraded or duplicated.
Orientation, deskew and contrast operate on a bounded temporary raster. Original colour pages remain untouched by default.
Page strategies, languages, confidence, transformations and low-confidence warnings remain available in structured data.
About PDF OCR
The OCR workflow classifies each page, preserves reliable native text, and sends only pages that need recognition through local Tesseract. Orientation, deskew, contrast, language selection, confidence, and page ranges remain under your control.
The default searchable-PDF output preserves original page visuals and adds an invisible text layer.
Yes. Enter a page range such as 1-3,5 before starting processing.
No. OCR runs locally with installed Tesseract language data.