Skip to main content
NEW

Image to Text (OCR)

Extract text from images using optical character recognition.

How to Use Image to Text (OCR)

  1. 1Upload an image containing text — a photo, screenshot or scanned page — or paste one with Ctrl+V.
  2. 2Choose the language, or a combination such as English + Urdu for mixed documents.
  3. 3Click 'Extract Text' and watch the progress; you can cancel at any time.
  4. 4Copy the text or download it as a .txt file.

Frequently Asked Questions

What languages are supported?

English, Urdu, Arabic, Hindi, Bengali, Persian, French, Spanish, German, Italian, Portuguese, Russian, Turkish, Chinese (Simplified) and Japanese, plus combined English + Urdu, English + Arabic and English + Hindi for documents that mix scripts. Each language's data downloads once and your browser keeps it.

Does it work on scanned PDFs?

This tool reads image files. For a scanned PDF, use our PDF Make Searchable tool, or PDF to JPG first and then run each page through here.

Why is my extracted text inaccurate?

OCR accuracy depends on the image. Use a sharp, well-lit, straight-on photo with good contrast and printed (not handwritten) text, and choose the right language.

About Image to Text (OCR)

Optical Character Recognition (OCR) turns pictures of text into text you can edit, search and copy. Image to Text uses Tesseract.js, a long-established open-source OCR engine, entirely in your browser — the image is never uploaded.

For the best results use a clear, high-resolution image with good contrast and straight lines of printed text. It handles screenshots, documents and photographed signs well; handwriting and heavily stylised fonts are much harder for any OCR engine.

You May Also Like