PDF to Text & OCR

Local PDF extraction and OCR

Extract text from digital and scanned PDFs, with OCR only when needed

Page Preview

Add a PDF or load the sample to preview pages.

Extracted Text

Choose pages and a mode, then select Extract Text.

Runs locally in your browser. The PDF, its text and any password are never uploaded. OCR is approximate and reading order is inferred from layout. The first OCR run downloads the OCR engine from this site.

PDF to Text & OCR

Extract text from digital and scanned PDFs in your browser. Uses the text layer first and OCR only when needed. Add one PDF, choose a page range, extraction mode and OCR language, then select Extract Text. Browse pages in the preview, switch between Text and JSON, and copy or download the current page or all selected pages.

Extraction Mode

Auto reads the digital text layer and runs OCR only on pages with no usable text. Text Layer never runs OCR. Force OCR renders and recognizes every selected page even when digital text exists.

OCR Language

Used by OCR only. Auto picks English or Indonesian from the first OCR page; it is not general language detection. Digital text in any language is read directly.

Privacy

Everything runs locally. The PDF, its text and any password stay in this browser tab. The OCR engine and English/Indonesian language files load from this site; there is no remote OCR fallback.

Limits

Up to 100 MB per file and 500 selected pages per run (any pages of a longer file). OCR is approximate and reading order is reconstructed from layout. Only mode, language and output preferences are saved. Limits: 100 MB per file · 500 selected pages per run · OCR renders up to 16 MP per page.

CodingTool is officially live on

Launched on StartupBaseCodingTool.dev - Free, Fast & Easy Developer Tools in One Place | Product HuntCodingTool - Featured on Startup FameFazier badgeListed on Turbo0Featured on Twelve ToolsFeatured on Findly.tools