FREE · NO ACCOUNT · ON YOUR DEVICE
Read the text in scanned PDF pages.
Recognize English or Simplified Chinese text from selected scanned PDF pages and download TXT. Free local OCR, with no account, upload or cloud storage.
How is this different from PDF to text?
PDF to text reads an existing text layer. This tool renders selected pages into temporary images and recognizes the pixels with Tesseract.js. Use it for scanned printed documents without selectable text. PDF text extraction is faster and usually more accurate when a good text layer already exists.
Does it create a searchable PDF or Word file?
No. It exports recognized UTF-8 TXT with source page separators, not a PDF with an OCR text layer or a reconstructed Word layout. Tables, columns, handwriting, unusual fonts and faint scans may be misread. Compare names, amounts and punctuation against the original before reusing the text.
What are the size and language limits?
English or Simplified Chinese plus English. One unprotected PDF up to 20 MB and 100 pages, selecting at most 10 pages per job. Rendering is bounded to 10 megapixels per page, a 6,000-pixel edge and 40 megapixels total. Choose 150 or 200 DPI. The two-minute deadline may require a smaller selection on slower devices.
Are page images saved or uploaded?
No. PDF pages and recognized text exist only in this tab’s temporary memory. The engine and language models are loaded from this site with OCR data caching disabled. Form fields and annotations are not rendered, and copying restrictions are respected. Download TXT before clearing or leaving; there is no saved OCR history.