Extract text from a scanned PDF (OCR)
Recognize and extract the text of a scanned PDF (images, no text layer) with OCR, directly in your browser. Works best with sharp pages; does not recognize handwriting or tables.
or click to select it
Recognize the text of a scanned PDF (pages that are images, with no real text layer) with OCR and turn it into text you can copy or download, directly in your browser.
Renders each page of the PDF as an image and uses an OCR engine (Tesseract) that runs entirely in your browser to recognize its text, page by page. Works best with sharp, high-contrast pages; it does not recognize handwritten text or tables, and recognition can take several seconds per page. If the PDF already has a digital text layer, the PDF to Text tool is faster and more accurate for that document.
Use cases
- Recover the text of a document scanned with a scanner or a phone.
- Convert a scanned contract or report into copyable text.
- Extract the content of an old PDF with no real text layer.
How it works
- 1
Upload your scanned PDF
Drag the file or choose it from your device.
- 2
Choose the text language
Select Spanish or English depending on the document's language.
- 3
Extract the text
Click the button and wait for OCR to recognize each page.
Frequently asked questions
Does it work with any PDF?
It's designed for scanned PDFs (pages that are images). If the PDF already has digital text, the PDF to Text tool (no OCR) is faster and more accurate.
Does it recognize handwriting or tables?
No, the OCR is designed for printed, straight-line text. Handwritten text and complex tables may be recognized inaccurately or not at all.
How many pages does it support?
OCR is an expensive process that runs in your browser, so there's a page limit per document, designed especially to protect mobile devices.
Is my PDF uploaded to any server?
No, all OCR processing happens in your own browser.