Extract text from a PDF or image in your browser
The OCR tool reads text from a PDF or image locally in your browser, then lets you copy or download the extracted text.
How OCR text extraction works
Upload a PDF or image, choose the recognition language, then run OCR. For PDFs, each page is converted in the browser before text recognition.
Useful for scanned documents
OCR helps recover text from scanned invoices, forms, administrative letters, receipts, or photographed pages.
Privacy-first OCR
The file is processed locally in your browser. For large PDFs, the tool warns you because OCR can take time and use browser resources.
Frequently asked questions
Is OCR performed on a server?
No. OCR runs in your browser with Tesseract.js.
Can I copy the extracted text?
Yes. After recognition, you can copy the text or download it as a TXT file.
Why is there a warning for large PDFs?
OCR is CPU-intensive. PDFs with more than 10 pages can take several minutes and use significant browser resources.