Image to Text (OCR)
Extract text from screenshots, scans and photos. Runs on your device, with an optional AI mode for handwriting.
Runs on your device by default. The optional cloud AI mode sends your input to Google Gemini.
What the Image to Text (OCR) does
OCR turns a picture of text back into text you can select, search and edit. It matters more than it sounds: a scanned contract, a screenshot of an error message, a photographed receipt — all of them hold information your computer cannot read, because to the machine they are just coloured pixels arranged in shapes. This tool offers two recognisers with genuinely different trade-offs. The default runs entirely inside your browser: the image is never uploaded, it works offline once loaded, and there is no usage limit. The optional AI mode sends the image to Google and reads things the on-device engine cannot — handwriting, table layouts, and scripts other than Latin. Which one you should use depends less on quality than on what is in the picture.
How to extract text from an image
- Upload a PNG, JPG, WebP or BMP, or paste a screenshot straight from your clipboard.
- Leave the recogniser on On-device unless you need what the AI mode adds. On-device keeps the image on your machine.
- Press Extract text. The first on-device run downloads about 9MB of recognition data; after that it is cached and near-instant.
- Read the confidence score. Below about 70% you should expect mistakes and check the result against the image.
- Correct anything wrong directly in the output box, then copy it or download it as a .txt file.
The Image to Text (OCR) runs on your device by default. If you switch to the optional cloud AI mode, your input is sent to Google's Gemini model to produce the result.
When to use it
Getting text out of a scanned PDF
A scan has no text layer, which is why converting one to Word produces an empty document and why a PDF editor cannot find any words to change. Export the page as an image, run it through here, and you have text again. This is the single most common reason people need OCR.
Copying from a screenshot
Error messages, chat threads and slides are constantly shared as images. Rather than retyping a stack trace by hand, extract it and paste it into your terminal or a search box.
Digitising receipts and invoices
Photographs of receipts are awkward: the paper curves, the lighting is uneven, and the layout is columnar. The AI mode handles all three considerably better than the on-device engine, though it means uploading the image.
Reading handwriting
Classical OCR is built around printed letterforms and does poorly on handwriting. If your image is handwritten, the on-device mode will likely return nonsense and the AI mode is the only realistic option.
Good to know
- Resolution matters more than file size. A sharp 1000px-wide crop of the text beats a 12MP photo of the whole page.
- Straighten the image first if it was photographed at an angle: on-device OCR assumes roughly horizontal lines of text.
- Crop to just the region you need. Less surrounding clutter means fewer spurious characters.
- Low contrast is the most common cause of poor results. Dark text on a light background reads far better than grey on grey.
- The on-device engine is English-only here. For other scripts, use the AI mode.
Frequently asked questions
Is my image uploaded anywhere?
Why is the first run slow?
Can it read handwriting?
Which mode should I use for something confidential?
Why is the accuracy lower than my phone's built-in scanner?
Does it keep the original layout?
Related tools
Compress Image
Shrink JPG, PNG and WebP images to a target size like 50 KB.
LLM Token Counter
Estimate BPE tokens, context limits, and API costs for GPT-4o, Claude, and Gemini.
AI Prompt Optimizer
Build structured XML system prompts with role isolation and negative rules.
AI Function Calling Schema
Build OpenAI Structured Output and Anthropic Tool Call JSON schemas.