Deni.AI Labs100% in-browser · no uploads
LIVEBackground Remover · ~26 MB Image to Text (OCR) · ~10 MB Speech to Text · ~40 MB Object Detector · ~170 MB AI Image Describer · ~250 MB Text Summarizer · ~300 MB Sentiment Analyzer · ~70 MB Text to Speech · 0 MB
AI Labs › AI Tools › Text to Speech

Text to Speech

Listen to any text read aloud using the voices built into your browser and operating system, with no download required.

AudioModel Device voicesDownload 0 MBSpeed InstantPrivacy stays on your device
Ready. The model downloads the first time you run it (0 MB).

Reading on a screen for a long time is tiring, and sometimes it is easier to listen. Paste an article, a draft or a set of notes into this tool, choose a voice and a speaking speed, and your device reads it aloud. It is also a simple way to proofread your own writing, since awkward sentences are easier to hear than to see.

Unlike the other tools in the lab, this one has no AI model to download. It uses speech capabilities that already come with your browser and operating system, so it starts instantly.

How to use it

  1. Type or paste the text you want to hear into the box.
  2. Pick a voice from the list. The available voices depend on your device and browser.
  3. Adjust the rate and pitch sliders to a comfortable level.
  4. Press play, and use pause, resume or stop as needed.

How it works

The tool uses the Web Speech API, a standard browser feature, specifically its speechSynthesis interface. When you press play, the page passes your text and settings to the browser, which hands it to a speech engine provided by your operating system or browser vendor. That engine converts the words into sounds and plays them through your speakers. The voices are those installed on your device, which is why the list differs between Windows, macOS, Android, iOS and Chrome OS. Some browsers also offer online voices that are generated by the browser vendor's own service.

Good for

Limitations

FAQ

Why do I see different voices on my phone and my computer?

The voices come from your operating system and browser, not from this site. You can usually add more voices in your system's language or accessibility settings.

Can I download the speech as an MP3?

No. The browser speech feature does not provide audio files to web pages, so playback is live only.

Is my text sent anywhere?

Built-in local voices process text on your device. Some browsers, such as Chrome, also list online voices; if you choose one of those, the browser vendor may process the text on its servers.

More tools

1

Speech to Text

Transcribe English speech from an audio file or your microphone using a Whisper model that runs on your own device.

AudioModel Whisper tiny.enDownload ~40 MBIn MP3 / WAV / M4A / micOut Transcript, timestampsSpeed ~1× realtime
2

Background Remover

Cut the subject out of a photo and save it as a transparent PNG, processed entirely on your own device.

VisionModel MODNetDownload ~26 MBIn JPG / PNG / WebPOut Transparent PNGSpeed 1–5 s
3

Image to Text (OCR)

Turn screenshots, scanned pages and photos of printed text into editable text without sending the image anywhere.

VisionModel Tesseract (English and more)Download ~10 MBIn Screenshot, scan, photoOut Plain textSpeed 2–10 s
4

Object Detector

Find common objects in a photo and draw labeled boxes around them, using a detection model that runs locally.

VisionModel DETR ResNet-50Download ~170 MBIn JPG / PNG / WebPOut Labeled boxes, countsSpeed 3–10 s