Image to text — extract text from a photo or screenshot, offline-ready

This tool reads the words in a photo, scan or screenshot and gives you editable text, using an OCR engine that runs in your browser.

Live demo · real result from this tool

Drop images here

or paste with Ctrl+V · nothing is uploaded

First use downloads the OCR engine (3.9 MB) plus the languages ticked above from jsDelivr; your image is never uploaded.

Your result appears here — processed on your device.

Processed on your device — nothing is uploaded.

How to use Image to Text (OCR)

  1. Choose one or more images: screenshots, phone photos of a page, or scans.
  2. Tick the languages in the image: English, Hindi and/or Telugu. The download size of each is shown beside it.
  3. Keep Keep line breaks ticked to follow the layout of the original, or untick it to get one flowing paragraph.
  4. Press Extract text. The first run downloads the engine and the language data.
  5. Read the result in the Extracted text box, copy it, or save it as extracted-text.txt.

What it does and when to use it

OCR (optical character recognition) turns a picture of words into words you can edit, search and paste. Common uses:

  • Copying from a screenshot, such as an error message, a quote from a video, or text in an image you cannot select.
  • Typing up a printed page: a notice, a letter, a recipe or a page from a textbook.
  • Pulling details off a receipt or bill to paste into a spreadsheet.
  • Reading a Hindi or Telugu notice into text you can translate or search.

Because nothing leaves your device, it is also suitable for documents you would rather not upload, like letters with your address on them.

How it works

The tool uses Tesseract, a long-standing open-source OCR engine, through tesseract.js (Apache-2.0).

  1. Downloads. On first use the browser fetches the Tesseract worker and core (about 3.9 MB) and the trained data for each language you ticked, from jsDelivr. The language files are stored in your browser’s IndexedDB, so they are not fetched again next time.
  2. Recognition. Tesseract’s neural-network recogniser finds lines of text, then reads each word. With more than one language ticked, it considers all of them together.
  3. Several images. Each file is read in turn. With more than one image, each block of text starts with the file name.
  4. Result. You get the text, the number of characters, the languages used and an average confidence score.

The image itself never leaves the page. The engine is shut down when the job ends to free memory.

Worked examples

A screenshot of an error dialog. Choose the screenshot, keep English, and press Extract text. Screenshots are sharp and evenly lit, so they are the easiest case. The text appears with its line breaks, ready to paste into a search engine or a support ticket.

A photo of a two-column printed notice. Tesseract reads columns as separate blocks, but the order can come out mixed. If that happens, crop each column with the Image Cropper and run them one at a time.

A Hindi circular photographed on a phone. Untick English, tick Hindi, and run it. If the confidence is low, take the photo again straight on, in daylight, with the page filling the frame. A tilted page can be straightened first with the Image Rotator.

Limits and tips

  • Sharp and straight wins. Hold the phone parallel to the page, fill the frame, and avoid shadows from your hand.
  • Dark text on a light background reads best. White text on a dark or busy background often fails; try the Image Invert tool first.
  • Small text needs pixels. If the letters are only a few pixels tall, take a closer photo rather than zooming a screenshot.
  • Tables lose their shape. The text comes out, but columns may not line up. Check numbers carefully before you rely on them.
  • Always proofread. Even at high confidence, OCR can swap similar shapes like 0 and O or 1 and l. Do not use the output for anything important without checking it against the image.
  • Three languages only. Other scripts are not supported on this page yet.
  • Going the other way? To write corrected text onto a picture, use the Image Text Overlay tool.

Frequently asked questions

Is my image uploaded for OCR?
No. The OCR engine and language files download to your browser, and the image is read on your device. Nothing you choose is sent to a server.
Which languages can it read?
English, Hindi and Telugu. Tick one or more. Ticking only the languages in your image gives better and faster results.
Why is some of the text wrong?
OCR guesses each letter from the shapes it sees. Blur, shadows, tilted pages, small print, handwriting and fancy fonts all cause mistakes. The confidence figure gives a rough idea; below 70% the tool warns you to check carefully.
Can it read handwriting?
Not reliably. The engine is trained on printed text. Neat block capitals sometimes work, but joined-up handwriting usually comes out garbled.
How much does it download?
The first run fetches the OCR engine (about 3.9 MB) plus each language you tick, 2.95 MB for English, 1.39 MB for Hindi and 1.74 MB for Telugu. Language files are saved in your browser, so later runs start faster.
Can I get text from a PDF?
For a PDF that already has selectable text, use PDF to Text. For a scanned PDF, OCR PDF adds a searchable text layer to the pages.

Written by the ToolsRift team · Last updated

Pages are drafted with AI assistance and checked by automated tests. How we write and check pages.

Was this tool useful?

Report a problem

Start typing to search every tool.

↑ ↓ to moveEnter to openEsc to close