Drop images or a PDF here
or click to browse — several files at once is fineor press Ctrl+V to paste
Language
Add an image or PDF to start reading text

Image to Text

Pull the text out of screenshots, photos and scanned PDFs. Recognition runs in your browser, so nothing is uploaded. Handles several files at once. Free, no signup.

What This Tool Does

Reads printed text well

Accuracy depends on image clarity, language, layout and lighting. Check the extracted text against the original, especially names, dates and numbers.

Nine languages

English, Chinese (both simplified and traditional), Japanese, Korean, German, French, Spanish and Italian. Non-English options load English alongside, because real documents are full of order numbers, emails and amounts written in Latin characters.

Scanned PDFs, page by page

A PDF is opened and each page is read in turn, so a multi-page scan comes back as text without you exporting the pages yourself first.

Several files at once

Drop in a folder's worth of screenshots and they queue up. If one of them fails, the rest carry on — a single bad file does not stop the batch.

Copy or download

Take the result straight to the clipboard, or download it as a.txt file. With several files, each one is labelled so you can tell the results apart.

How to Extract Text from an Image

1

Add your files

Drag in images or a PDF, or click to browse. JPG, PNG, WebP, GIF, BMP and PDF all work, up to 50MB each.

2

Pick the language

Choose the language of the text in the picture — not the language of this page. Getting this right matters more than anything else for accuracy.

3

Copy or download the text

Results appear as each file finishes. Copy everything to the clipboard, or download it as a.txt file.

Real examples

Image to Text · Example 1

These files were processed with this tool or its shared processing code. Compare the source, settings and downloadable results.

Getting Better Results

Set the language first

This is the single biggest factor. Reading Chinese text with the English model set will produce nonsense, no matter how clear the image is.

Crop to just the text

Photos of a page usually include the desk, a hand, part of the room. Cropping down to the paper alone gives the engine less to be confused by.

Even lighting beats bright lighting

A shadow falling across half the page hurts more than the picture being a bit dark overall. Move so the light is even rather than strong.

Keep the lines horizontal

Text that runs at an angle is read noticeably worse. Straightening the photo before uploading is worth the few seconds.

Use the original file

A screenshot of a screenshot, or a photo forwarded through several chat apps, has been compressed each time. The original is always sharper.

Printed text, not handwriting

This works on printed and typed text. Handwriting is a genuinely different problem and the results will disappoint you.

Real examples

Image to Text · More examples

These files were processed with this tool or its shared processing code. Compare the source, settings and downloadable results.

Questions About Extracting Text

Yes. OCR — optical character recognition — is the general name for reading printed text out of a picture, and that is exactly what happens here. The difference from most OCR websites is where it runs: the engine is downloaded into your browser rather than your file being uploaded to someone's server.
Accuracy depends on image clarity, language, layout and lighting. Check the extracted text against the original, especially names, dates and numbers.
No, and this is worth being clear about. The engine is built for printed and typed text. Handwriting recognition is a different problem needing different models, and trying it here will waste your time.
No. The text inside a table will be read, but the row and column structure is not preserved — you get the words, not a grid. The engine does not detect table structure at all, so there is no way to export usable Excel from it.
Selected file contents are processed in your browser without being sent to an application upload endpoint. Loading the site, fonts or processing libraries may still require network requests.
Open the page while online. Processing may continue in the loaded tab, but reloading, external assets, and uncached engines or models require a connection.
The language of the text in the image, not the language of this page. If you pick wrong the output will be nonsense even from a clear image. Non-English choices load English as well, so mixed documents with order numbers and email addresses still read correctly.
Yes. Each page is rendered and read in turn, and results appear page by page. Very long documents are capped at the first 100 pages, and you are told when that happens rather than being left to wonder.
You will see "No text found" rather than an error, because an image with no readable text is a normal result. Usually it means the picture is too blurry, the text is too small, or the wrong language is selected.
Because it should not have them. Recognition engines treat each Chinese character as a separate word and put spaces between all of them, which makes the result awkward to use. Those spaces are removed here — but spaces between Chinese and Latin text are kept, since those are real.
Yes. Drop in as many images as you like and they queue up, with results appearing as each finishes. If one file fails the rest continue, and you can stop the queue at any point without losing what has already been read.