Image to Text Converter (OCR)

Extract text from screenshots, photos and scanned pages with OCR that runs in your browser, in English and 8 other languages.

Drop an image here, click to choose, or paste JPG, PNG or WebP. Paste a screenshot with Ctrl+V (Cmd+V on Mac) anywhere on this page.

The first time you use a language, its data file (a few MB) downloads once and is kept by your browser for later runs.

Image

Your image preview appears here.

Extracted text

OCR runs in your browser with Tesseract. Your image is never uploaded.

In short

Upload or paste a JPG, PNG or WebP image into the converter above, pick the language, and it extracts the text with Tesseract OCR running inside your browser. You can copy the result or download it as a TXT file, and the image is never uploaded to a server.

On this page
  1. How to extract text from an image
  2. Supported languages
  3. Tips for better accuracy
  4. What OCR cannot read well
  5. Privacy: what happens to your image
  6. Uses for store owners and marketers

Text trapped in an image is useless until you retype it: a supplier price list sent as a screenshot, a product label, a scanned invoice, a slide from a webinar. This converter reads the text for you with optical character recognition (OCR), so you can copy it, edit it and paste it where it belongs.

It runs Tesseract, an open source OCR engine, inside your browser. The image is processed on your own device and is never uploaded, so it is safe for invoices and internal documents.

How to extract text from an image

  1. Upload a JPG, PNG or WebP file, drag it onto the box, or paste a screenshot straight from your clipboard with Ctrl+V or Cmd+V.
  2. Choose the language of the text. English is the default.
  3. Start the conversion. The first time you use a language, its data file downloads once, so the first run takes longer.
  4. Read the result and fix any errors. OCR is rarely perfect on the first pass, especially with numbers and punctuation.
  5. Copy the text or download it as a TXT file.

Supported languages

LanguageScriptNotes
EnglishLatinThe default and usually the most accurate
Spanish, French, German, Portuguese, ItalianLatin with accentsChoose the right language so accented letters such as รฑ, รฉ, รผ and รง are read correctly
UrduArabic script, right to leftPlain printed text works best; decorative Nastaliq on posters reads less reliably
ArabicArabic script, right to leftDiacritics and very small text lower accuracy
HindiDevanagariClear print gives the best results

Pick the language that matches the text. If you run a Spanish page as English, the engine will still find words, but accented letters and some words will come out wrong. For a page that mixes languages, run it with the main language and correct the rest by hand.

Tips for better accuracy

Most OCR errors come from the image, not the engine. Tesseract's own documentation recommends at least 300 dpi for scans, straight text lines and dark text on a light background. In practice:

  • Use enough resolution. Small text in a low-resolution screenshot is the most common cause of errors. Zoom in on the page before taking the screenshot, or scan at 300 dpi.
  • Crop to the text. Remove toolbars, photos and logos around it. A small margin is fine, but large unrelated areas can turn into stray characters.
  • Keep it straight. Photograph pages flat and square on, not at an angle. Tilted lines are harder to separate.
  • Get good contrast. Dark text on a light background works best. White text on a dark banner, or text over a photo, often fails. Even lighting without shadows helps with phone photos.
  • Convert unusual formats first. iPhone photos saved as HEIC, and other formats, can be turned into JPG or PNG with the image converter.

Tip: For a scanned PDF, convert the pages to images with the PDF to JPG converter, then run each page here.

What OCR cannot read well

ContentResultWhat to do
HandwritingPoor. Tesseract is built for printed textType it by hand
Script, decorative or very bold display fontsLetters misread or skippedCrop to plain body text where possible
Text over photos or patternsMissing words and random charactersCrop tighter or find a cleaner source
Multi-column layouts and tablesText is found, but the order or columns may mixCrop one column or table section at a time
Curved or rotated textOften unreadableRotate the image or retake the photo
Math symbols and special charactersOften misreadCheck and correct by hand

Always check numbers carefully. A 0 read as O, a 1 read as l, or a missing decimal point in a price or SKU is the kind of error that costs money later.

Privacy: what happens to your image

The OCR engine runs in your browser tab. Your image is read on your device and is not sent to our server or saved. The only download is the language data file, which your browser fetches the first time you use a language and stores so later runs start faster. If you clear your browser's site data, it downloads again the next time. This makes the tool suitable for receipts, invoices and supplier documents you would not want to upload to an unknown service.

Uses for store owners and marketers

  • Supplier price lists and spec sheets sent as screenshots or photos, turned into text you can paste into a spreadsheet.
  • Product labels and packaging for ingredients, materials, care instructions and barcodes printed as text.
  • Receipts and invoices for bookkeeping, when you need the numbers without retyping.
  • Slides, infographics and social posts you want to quote or summarize, with credit to the source.

After extraction, the word counter helps you trim text for product descriptions. If you are copying product details from another online store rather than from images, there is a faster route: AM Jarvis Product Importer copies products from public Shopify and WooCommerce stores into your own store with variants, images and prices, and lets you edit and rewrite the text before import.

Frequently asked questions

Is this image to text converter free and private?

Yes. It is free with no account, and the OCR runs in your browser, so your image is never uploaded to a server. The only thing downloaded is the language data file the first time you use each language, which your browser keeps for later runs.

Can it read handwriting?

Not reliably. Tesseract, the OCR engine used here, is built for printed text, and its own documentation says settings will not make it good at handwriting. Neat block capitals sometimes come through partly, but for handwritten notes, typing is usually faster than correcting OCR output.

Why is my extracted text full of errors?

Usually the image is too small, blurry, tilted or low in contrast, or the wrong language is selected. Zoom in before taking a screenshot, crop to the text, photograph pages flat in even light and choose the language of the text. Scans at 300 dpi give the best results.

Which image formats are supported?

JPG, PNG and WebP. You can upload a file or paste a screenshot from your clipboard. For HEIC photos from an iPhone or other formats, convert them to JPG or PNG first with the image converter, and for scanned PDFs, convert each page to an image first.

Can I extract text in Urdu, Arabic or Hindi?

Yes. Along with English, you can choose Spanish, French, German, Portuguese, Italian, Urdu, Arabic and Hindi. Clear, plain printed text gives the best results. Decorative lettering, very small text and heavy diacritics in right-to-left scripts reduce accuracy, so check the output carefully.