OCR Free · no signup · no watermark

Image to Text Converter (OCR)

Convert any picture to text you can edit: photos, screenshots and scanned pages all work. The OCR engine runs inside your browser, so your images never leave your device.

Updated Krawly Editorial TeamIn-house engineers, writers & reviewers

or drop files here · paste with Ctrl+V

JPG, PNG, WEBP, BMP, GIF · up to 50 MB each · max 20 files

Processed in your browser — files are never uploaded

Quick answer

To convert an image to text, drop a JPG, PNG, WebP, BMP or GIF file into Krawly's Image to Text tool, pick the language of the text (up to 3 at once) and start OCR. The Tesseract engine reads the image locally in your browser and returns editable text you can copy or download as a .txt file. Nothing is uploaded and there is no signup.

How to use the Image to Text (OCR) tool

  1. 1

    Add your images

    Drag JPG, PNG, WebP, BMP or GIF files onto the drop area, or click to choose them. You can add up to 20 images in one run, for example a set of screenshots or photographed pages.

  2. 2

    Choose the text language

    English is selected by default. Pick the language your text is written in, or up to three languages if the image mixes them, such as English and Spanish on a bilingual menu.

  3. 3

    Run OCR

    Start recognition. On the first run your browser downloads the OCR engine and the language model, then caches them. Each image is processed on your device and gets its own text box with a mean confidence score.

  4. 4

    Review and correct

    The results are editable. Fix any misread characters directly in the text box before you use the text; a low confidence score is a hint that a line needs a closer look.

  5. 5

    Copy or download

    Click Copy to put the text on your clipboard, or Download .txt to save the results. When you process several images, the download combines them into one .txt file with a header for each image.

How does image to text (OCR) work?

OCR stands for optical character recognition: software looks at the pixels of an image, finds the areas that contain writing and works out which characters they represent. The result is real text that you can search, edit, paste into a document or feed into a spreadsheet, instead of a picture of text that you would otherwise have to retype by hand.

Krawly uses Tesseract, a long-running open-source OCR engine. It was first developed at Hewlett-Packard, released as open source in 2005 and later developed further with sponsorship from Google. Here it runs as Tesseract.js, a build compiled to WebAssembly so that it can execute inside a web page without any server. Roughly speaking, the engine cleans the image up into dark text on a light background, splits the page into blocks, lines and words, and then recognizes each line with a trained model for the language you selected.

That last step is why the language choice matters. Every language has its own model with its own alphabet, accented letters and common word shapes. If you tell the engine to read English while the image is in Polish or Vietnamese, it will still try, but letters with diacritics come out wrong. Choosing the correct language, or a combination of up to three, is one of the simplest ways to get cleaner results.

How to get accurate text from a photo or screenshot

OCR works best on clear, printed text with good contrast and straight lines. Screenshots are usually ideal because the text is sharp and perfectly horizontal. For scans, around 300 DPI is a good target: high enough that small letters have clean edges, without creating huge files. If you photograph a page with your phone, fill the frame with the document, hold the camera parallel to the paper and avoid shadows from your hand or the phone itself.

Contrast matters more than colour. Black text on white paper is easy; light grey text on a beige background, or white text over a busy photo, is much harder. If a result looks poor, try cropping the image to just the text area, rotating it so the lines are level, and increasing contrast in any photo editor before running OCR again. Crooked, curved or perspective-distorted lines, such as a page photographed near the book's spine, are a common cause of missing words.

Very small text is another frequent problem. When letters are only a few pixels tall, there is not enough detail to tell an e from a c or an l from a 1. Zooming in before taking a screenshot, or scanning at a higher resolution, usually helps more than any later editing. Keep in mind that enlarging an already small image cannot add detail that was never captured.

What OCR can't do: handwriting, tables and layout

Tesseract is built for printed text. Handwriting, cursive notes, signatures and heavily stylised or decorative fonts usually produce poor or unusable results. The same goes for low-resolution, blurry or heavily compressed photos. If your image falls into one of these groups, expect to correct a lot of the output by hand, or retype the text.

The output is plain text. Columns, tables, bold and italic styling, font sizes and images are not kept. A two-column article is flattened into lines of text, and a table becomes rows of words separated by spaces, so you may need to tidy up the structure after pasting it somewhere else. This tool does not produce Word documents or spreadsheets.

Each image also gets a mean confidence percentage. It is the engine's own estimate of how sure it was about the recognized words, not a guarantee of accuracy, but it is a useful signal: a noticeably lower score usually means the image needs to be sharper, straighter or cropped tighter.

Is it safe to OCR screenshots, IDs and receipts?

People often run OCR on sensitive material: receipts for expense reports, screenshots of chats or bank apps, invoices, contracts, ID cards and letters. Many online OCR services upload those images to a server for processing. Krawly does not. The recognition runs entirely in your browser with JavaScript and WebAssembly, and your images are never sent to Krawly's servers.

The only things your browser downloads are program files: the Tesseract engine and the language model for each language you pick, a few megabytes each. After the first run they are cached, so later runs start faster. The image and the recognized text stay in the page's memory, nothing is stored, and closing the tab discards everything.

Because the work happens on your device, speed depends on your hardware. A recent laptop handles a screenshot quickly; an older phone takes longer on large photos. The tool works in current versions of Chrome, Edge, Firefox and Safari on desktop and mobile, and it is free, with no signup, no email and no daily limit.

Tips for the best results

  • Crop the image to the text you need before running OCR. Removing logos, photos and page borders gives the engine less to misread.
  • Select every language that appears in the image, up to three. A French document with English quotations reads better with both selected.
  • For screenshots, zoom the page in before capturing so the letters are larger and sharper.
  • Check lines with unusual characters, numbers and punctuation: 0 and O, 1 and l, and rn and m are classic OCR mix-ups.
  • Have an iPhone photo in HEIC format? Convert it with the HEIC to JPG tool first, then run OCR on the JPG.
  • For a scanned PDF, use PDF to Text instead of screenshotting each page. It reads the embedded text when there is any and runs OCR only where it is needed.

Frequently asked questions

Can it read handwriting?

Not reliably. The Tesseract engine is designed for printed text, so handwritten notes, cursive and signatures usually come out as broken or wrong characters. Very neat block capitals sometimes work partially. For handwriting, retyping is often faster than correcting the OCR output.

Which languages are supported?

English, Spanish, Portuguese, French, German, Italian, Dutch, Turkish, Polish, Russian, Ukrainian, Arabic, Hindi, Japanese, Chinese (Simplified), Chinese (Traditional), Korean, Vietnamese and Indonesian. You can select up to three languages at once for images that mix them.

Are my images uploaded to a server?

No. OCR runs in your browser using Tesseract compiled to WebAssembly. Your browser downloads the engine and language files, but your images and the extracted text never leave your device. Nothing is stored, and closing the tab clears everything.

Why is the first run slower?

The first time you use the tool, your browser downloads the OCR engine and the model for each selected language, a few megabytes per language. These files are then cached, so later runs with the same languages start much faster. Picking a new language triggers a one-time download of that model.

How many images can I convert at once?

Up to 20 images per run, each up to 50 MB. Every image gets its own editable text box, and Download .txt combines all results into one text file with a header per image. There is no daily limit, so you can start another run right away.

Can I convert JPG to text or PNG to text?

Yes. JPG, PNG, WebP, BMP and GIF are accepted, up to 50 MB each. Screenshots saved as PNG are usually the cleanest input because the text is sharp. HEIC photos from iPhones are not accepted directly; convert them with the HEIC to JPG tool first and then run OCR on the JPG.

How do I copy text from image files?

Drop the image here, choose the language of the text and run OCR. The recognized text appears in an editable box; fix any misread characters, then click Copy to put it on your clipboard and paste it anywhere. For several images, Download .txt saves all results in one file.

Does it keep tables and formatting?

No. The output is plain text. Columns are flattened, tables become lines of words separated by spaces, and fonts, bold text and images are dropped. You can edit the text in the result box before copying it, or tidy it up in your own editor.

What does the confidence percentage mean?

It is the engine's average confidence across the words it recognized in that image. It is not a measured accuracy figure, but a lower score is a good sign that the image is blurry, skewed, low-contrast or in a different language than the one selected.

Private by design. Every Krawly converter runs inside your browser using JavaScript and WebAssembly. Your files are not uploaded, stored or seen by anyone — when you close the tab, they are gone.

Related converters

See all 21 free converters