Skip to content
All Online Tools

13 tools

OCR: extract text from images and scanned PDFs

Optical character recognition runs on your device using the open-source Tesseract engine. Read text from screenshots, digitise printed documents, turn scanned PDFs into searchable ones and pull tables, receipts and invoices into Excel. Supports English, Hindi, Marathi, Gujarati, Bengali, Tamil, Telugu, Kannada, Malayalam and Punjabi.

Most used

More tools

Frequently asked questions

Is my document uploaded for OCR?

No. The OCR engine and language data are downloaded to your browser once (from a public CDN) and then your images are processed locally.

How accurate is it?

Clean printed text is usually recognised very well. Handwriting, low-light photos and decorative fonts are harder — every result shows a confidence score.

Which languages are supported?

English, Hindi, Marathi, Gujarati, Bengali, Tamil, Telugu, Kannada, Malayalam and Punjabi, including mixed English + Hindi documents.

Other categories