OCR text extraction

Upload a PDF or image, choose a recognition language, and extract text locally with Tesseract.js.

Drag & drop a PDF or image, or click to browse

PDF, PNG, JPG, WEBP, TIFF and more are supported

Extract text from a PDF or image in your browser

The OCR tool reads text from a PDF or image locally in your browser, then lets you copy or download the extracted text.

1

How OCR text extraction works

Upload a PDF or image, choose the recognition language, then run OCR. For PDFs, each page is converted in the browser before text recognition.

2

Useful for scanned documents

OCR helps recover text from scanned invoices, forms, administrative letters, receipts, or photographed pages.

3

Privacy-first OCR

The file is processed locally in your browser. For large PDFs, the tool warns you because OCR can take time and use browser resources.

Frequently asked questions

Is OCR performed on a server?

No. OCR runs in your browser with Tesseract.js.

Can I copy the extracted text?

Yes. After recognition, you can copy the text or download it as a TXT file.

Why is there a warning for large PDFs?

OCR is CPU-intensive. PDFs with more than 10 pages can take several minutes and use significant browser resources.

Massicot PDF 2026 - EN | FR