OCR PDF / Image to Text
Extract text from scanned PDFs and images using on-device OCR.
Drop a scanned PDF or image here
or use the button below · JPG, PNG, WebP or PDF
About this tool
Run Optical Character Recognition on a scanned PDF or a photo of text and get back editable, searchable text — entirely inside your browser. Nothing is uploaded to a server: the file, the OCR model and the result all stay on your device. This is the tool to reach for when a PDF has no real text layer (a phone photo of a page, or a document that was scanned rather than exported), which is exactly the case the Extract Text tool cannot handle.
Key Features
- Works on scanned PDFs and photos of text
- 9 languages, more can be requested
- Live per-page progress, not a spinner that hides what is happening
- Copy or download the result as .txt
- 100% client-side — files never leave your device
How to Use
- Drop a scanned PDF, JPG, PNG or WebP file.
- Pick the language the text is written in.
- Click Extract Text and watch the live progress.
- Copy the result or download it as a .txt file.
Benefits
- Digitize paper documents and receipts
- Make scanned PDFs searchable and copyable
- No account, no upload, no file size games
Frequently Asked Questions
Extract Text reads the invisible text layer a PDF already has — it is instant but does nothing on a scanned or photographed page. This tool runs real OCR (Tesseract) to read the pixels themselves, so it works on scans, photos and image-only PDFs that have no text layer at all.
Yes. Sharp, well-lit, high-resolution scans give the best results. Blurry photos, low resolution, or heavily skewed pages will reduce accuracy — try to photograph documents straight-on and in good light.
No. The PDF/image is rendered and recognized entirely in your browser using WebAssembly. No file or extracted text is sent to any server.
Yes — each page is rendered and OCR'd in sequence, with a running progress bar showing which page is currently being processed.