Deni.AI Labs100% in-browser · no uploads
LIVEBackground Remover · ~26 MB Image to Text (OCR) · ~10 MB Speech to Text · ~40 MB Object Detector · ~170 MB AI Image Describer · ~250 MB Text Summarizer · ~300 MB Sentiment Analyzer · ~70 MB Text to Speech · 0 MB
AI Labs › AI Tools › Image to Text (OCR)

Image to Text (OCR)

Turn screenshots, scanned pages and photos of printed text into editable text without sending the image anywhere.

VisionModel Tesseract (English and more)Download ~10 MBSpeed 2–10 sPrivacy stays on your device
Drop, paste or click to choose an image
Ready. The model downloads the first time you run it (~10 MB).

Optical character recognition (OCR) reads the letters in a picture and returns them as plain text you can copy, search and edit. It is useful when you have a screenshot, a receipt, a scanned letter or a slide and would rather not retype it.

Everything happens on your own computer. The image is read by code running in this browser tab and is never uploaded, so it suits documents you would not hand to an online converter.

How to use it

  1. Click the file picker or drag an image (PNG, JPG or WebP) onto the drop area.
  2. Wait for the recognition engine to load. The first time, the browser downloads about 10 MB of English language data, which is cached for later visits.
  3. Watch the progress bar while the page is analysed; a clear screenshot usually takes a few seconds.
  4. Review the extracted text, correct any misread characters, then copy it to your clipboard.

How it works

The tool uses Tesseract.js, a WebAssembly port of the open-source Tesseract OCR engine. It first cleans up the image and finds the lines and words on the page, then a trained recognition model compares the shapes of the characters against its English language data and picks the most likely letters. Because the engine is compiled to WebAssembly, it runs at close to native speed inside the browser without any plug-in.

Good for

Limitations

FAQ

Is my image uploaded anywhere?

No. The image is processed locally by Tesseract.js in this tab. Only the engine and the English language data are downloaded, and they contain no information about you.

Why is the first run slower than later ones?

On the first visit the browser fetches roughly 10 MB of engine and language files. They are stored in the browser cache, so later runs start much faster unless you clear your cache.

How can I get better results?

Use a sharp, well-lit image with straight text filling most of the frame. Cropping unrelated areas and using a real screenshot both help.

More tools

1

Background Remover

Cut the subject out of a photo and save it as a transparent PNG, processed entirely on your own device.

VisionModel MODNetDownload ~26 MBIn JPG / PNG / WebPOut Transparent PNGSpeed 1–5 s
2

Object Detector

Find common objects in a photo and draw labeled boxes around them, using a detection model that runs locally.

VisionModel DETR ResNet-50Download ~170 MBIn JPG / PNG / WebPOut Labeled boxes, countsSpeed 3–10 s
3

AI Image Describer

Get a short, plain-English sentence describing what appears in a photo, generated entirely on your own device.

VisionModel ViT-GPT2 captioningDownload ~250 MBIn JPG / PNG / WebPOut One-sentence captionSpeed 2–6 s
4

Speech to Text

Transcribe English speech from an audio file or your microphone using a Whisper model that runs on your own device.

AudioModel Whisper tiny.enDownload ~40 MBIn MP3 / WAV / M4A / micOut Transcript, timestampsSpeed ~1× realtime
5

Text Summarizer

Condense a long English article or report into a few sentences, generated by a model that runs on your device.

TextModel DistilBART CNN 6-6Download ~300 MBIn English text, 60+ wordsOut Short summarySpeed 10–60 s
6

Sentiment Analyzer

Paste English text and see whether each line reads as positive or negative, with a confidence score for each.

TextModel DistilBERT SST-2Download ~70 MBIn Text, one item per lineOut Positive / negative + scoreSpeed <1 s per line