Skip to content
Reembun
Image to text icon
Image

Image to text converter

Pull the text out of a photo, screenshot, scan or PDF, as plain text or with the headings, lists and tables kept intact.

  • English

3.0 MB downloaded once, then kept on your device.

Picking the right one matters more than any other setting.

Change this if the words come back in the wrong order.

Straightens a tilted page and evens out shadows before reading. Automatic applies it to photos and leaves screenshots alone.

Drop an image or PDF here, paste one with Ctrl+V, or

JPG, PNG, GIF, WebP, HEIC from an iPhone, and PDF. Read on your device.

Ready. Runs locally on your device.

Your files stay on your device. The tool works directly in your browser, using your device to process your files. Nothing is sent to our servers, and we never receive, store, or see your files or figures.

Share

Share this page

Share result

Your results

Use the tool above first. Whatever it works out shows up here, ready to copy or share.

Cite

Cite this page

Reembun. (2026, July 31). Image to text converter. https://reembun.com/image-to-text
Pick a style, then copy the reference. The access date is today.

How to use it

  1. Add the image or PDF

    Drop it in, paste with Ctrl+V, take a photo, or choose a file. JPG, PNG, GIF, WebP, HEIC from an iPhone and PDF all work.

  2. Set the language first

    Language of the text decides the result more than every other setting put together. Search all 102 by name, and pick two for a page that mixes them. Page layout is worth changing if the words come back out of order.

  3. Read the text

    Press Read the text. The engine and language model download once, then stay cached, so later images start instantly and work offline.

  4. Check confidence, then take it away

    Low confidence usually means the photo rather than the tool, so reshoot square on with more light. Copy it, download a Word file, or take it as plain text, Markdown or HTML.

What it reads

InputHow it is handled
JPG, JFIF, PNG, GIF, WebP, BMPDecoded by the browser, then read
HEIC and HEIF from an iPhoneDecoded here, because no browser except Safari will
PDF with real textText taken directly, with no recognition step
PDF that is a scanEvery page rendered at 200 dpi, then read
Screenshot on the clipboardPaste it with Ctrl+V anywhere on this page
A photo on your phoneThe camera button opens the rear camera

Multi-page PDFs are expanded automatically, page by page, and the results are joined in reading order.

Two output modes, and when to use each

Simple returns the engine’s own output. Line breaks land where the page broke them, which is what you want when you are quoting a paragraph or copying a serial number.

Formatted reads the geometry rather than only the characters. Every line carries a position and a height, and that is enough to recover the structure:

  • A line set noticeably larger than the body text, sitting on its own and not ending in a full stop, becomes a heading.
  • Lines that begin with a bullet or a number become a list, and consecutive items are folded into one.
  • Lines that split into the same number of columns at the same positions become a table.
  • Wrapped lines are rejoined into paragraphs, and a word broken across a line break is put back together.

You can then take that as a Word file, plain text, Markdown or HTML. The Word download is the one to use when the text has to be edited or sent on, because the headings, lists and tables arrive as real Word headings, lists and tables rather than as characters that look like them. Markdown suits a note-taking app or a CMS, since the structure survives the paste.

Why the photo matters more than the settings

Recognition failure is almost always an image problem. In rough order of how much damage each does:

Angle. A page photographed from 30 degrees off makes every character a different shape than the one the engine learned. Straighten up before you reach for any setting.

Shadow. A hand or a phone casting a line of shadow across the page splits it into two exposure zones, and the darker one usually reads as nothing at all.

Distance. Text needs to be roughly 30 pixels tall to be read reliably. Standing back and cropping later throws that away, because the pixels were never captured.

Motion. A slightly blurred photo looks fine at a glance on a phone screen and is unreadable to an engine that works letter by letter.

The confidence figure under the result is the engine’s own average across every word it read. Below 70 it is telling you plainly that it is guessing, and the fix is another photograph rather than another setting.

Page layout is a real setting, not a preset

If the words come back scrambled, the layout selector is the thing to change. Automatic detection tries to find columns and reading order, which is right for a book page and wrong for a receipt with a single narrow column of prices. Scattered text turns off layout analysis entirely, which is what a photo of a shop sign or a screenshot of a dialog box actually needs.

The privacy difference is structural

Every other free OCR site works the same way: your image is uploaded, processed on a server, and you are asked to trust a sentence about deletion. That sentence may well be true. It is unverifiable either way, and the image was still transmitted, still sat on a disk, and still passed through whatever logging that company runs.

Here the engine comes to the image. It is about 4 MB and downloads once, then your browser keeps it, so the second image you read starts instantly and works with no connection at all. That is not a marketing position. It is the only arrangement under which a payslip, a medical letter or a signed contract can go into an online OCR tool without leaving the building.

Common questions

Is my image uploaded anywhere?

No. The recognition engine itself is downloaded to your browser and runs there, which is the opposite of how almost every other OCR site works. Your photo is decoded in the tab, read in the tab, and discarded when you close it. Nothing about it reaches us, so there is no retention policy to read and no breach that could expose it.

What is the difference between Simple and Formatted?

Simple gives you the words with the line breaks exactly where the page broke them. Formatted looks at where each line sits on the page and rebuilds the document from that, so a larger line on its own becomes a heading, lines starting with bullets become a list, and columns separated by wide gaps become a table. Use Simple for a paragraph you want to quote and Formatted for anything with structure.

Which languages can it read?

All 102 languages Tesseract publishes a model for, including Arabic, Chinese, Japanese, Korean, Hindi, Thai, Russian and Greek. Search the selector by English name, Indonesian name, or the language's own name, and pick up to two at once for a page that mixes them. Getting that selector right matters more than any other setting, because the engine uses the language model to decide between characters that look almost identical. Reading Indonesian with the English model produces text that is subtly wrong throughout rather than obviously broken.

Can it read handwriting?

Rarely, and you should not rely on it. The engine was trained on printed type, so neat block capitals sometimes work and ordinary cursive almost never does. This is a limitation of the technology rather than of this tool, and any OCR site promising accurate handwriting recognition from a browser is overselling.

Why does my scanned PDF come out perfectly but my photo does not?

Because a PDF made by a computer already contains its text, and this tool takes it directly instead of recognising it. When that happens you are told so on the page, and the result is exact. A photograph has no text inside it, only pixels, so every character has to be inferred and the quality of the photo sets the ceiling.

What resolution should I photograph at?

Fill the frame with the text, hold the camera square to the page rather than at an angle, and give it even light with no shadow across the middle. Resolution matters far less than those three things. The tool rescales whatever you give it to the size the engine reads best at.

What does Photo cleanup do?

It straightens a page photographed crooked, evens out a shadow falling across one side, and separates ink from paper at every point on the page rather than by one setting for the whole of it. That last one is the important step. Recognition engines normally pick a single brightness cut for the entire image, which is correct for a scan and wrong for a photo where one corner is darker than the other. On Automatic it applies to photos and leaves screenshots untouched, because screen text is already clean and processing it would throw away detail the engine uses. The result tells you what it did, and you can switch it off.

Last reviewed