Text

Text to Speech

Details

How to use Text to Speech

What the tool does, how to run it, and what to expect from the result.

How to read text aloud in your browser

Every modern browser ships the Web Speech API, which exposes the text-to-speech voices already installed on your operating system. That means no upload, no account, no per-character pricing, and no waiting for a model to generate audio: the speech starts the instant you press play.

Paste text, choose from the voices your device offers, set the rate and pitch, and listen. The trade-off is that you cannot save the result as a file, because the API does not allow it.

  • Paste the text you want read aloud.
  • Pick a voice. The list comes from your operating system, so it varies by device.
  • Set the rate (0.5x to 2x) and pitch (0 to 2).
  • Press Play, and use Pause and Stop as needed.
  • For an audio file rather than live playback, you need a server-side TTS service, since the browser API cannot export one.
Tips

Getting a better result out of Text to Speech

Specific settings and thresholds, not general advice.

  • The voices come from your operating system, not from this page. That is why the list differs between a Mac, a Windows PC, and an Android phone, and why a colleague sees voices you do not. There is no cloud model here, which is exactly why it is free, instant, and private.
  • On Chrome, several of the listed voices (the ones from Google) are actually synthesized on Google's servers, so the audio for them does leave your machine even though nothing about this page uploads it. The voices marked as local (Microsoft on Windows, Apple on macOS) run entirely offline.
  • The rate slider runs 0.5x to 2x and the pitch 0 to 2. Rates above about 1.5x get hard to follow for unfamiliar material, and pitch is the wrong lever for making a voice more pleasant: try a different voice first, since each is a different synthesis model.
  • Voices load asynchronously. On a fresh page load the dropdown can be empty for a moment because the browser fires a voiceschanged event when the list is ready. If you see no voices at all after a few seconds, the browser or the OS has no speech engines installed.
  • There is no way to save the audio. The Web Speech API speaks to your output device and exposes no recording hook, so if you need an MP3 you need a cloud TTS service or a screen recorder, not this.
Limits

What Text to Speech does not do

The honest boundary, so you do not lose time finding it yourself.

  • No audio download. The Web Speech API provides no way to capture the output as a file.
  • The voice list is whatever your OS and browser provide. There is no voice catalogue and no cloud voices to choose.
  • No SSML, no pronunciation dictionary, no per-word emphasis or pauses.
  • Pause and resume are unreliable in some browsers, particularly on mobile, where the platform may cancel the utterance instead.
At a glance

Who Text to Speech is for

A quick way to understand who this helps, what it solves, and where it connects next.

Best fit

Anyone who wants to listen to text read aloud.

Ideal for

Using the text to speech without installing anything or signing up.

FAQ

Common questions

Short answers for the questions people usually have before trying a utility like this.

Which voices are available?

Whatever is installed on your device. macOS and iOS supply Apple's system voices, Windows supplies the Microsoft ones, Android supplies Google's, and Chrome adds a set of network-backed Google voices on top. This page enumerates them with the Web Speech API and does not ship any of its own, which is why your list will differ from anyone else's.

Why does nothing play?

Three usual causes. The browser has no speech engine, in which case the tool tells you the API is unsupported. Or the voice list has not loaded yet, since it arrives asynchronously. Or your operating system has no voices installed for the selected language, which is common on stripped-down Linux installs and on some Android builds.

Can I download the speech as an MP3?

No. The Web Speech API renders audio directly to your output device and offers no hook to capture it, so no browser-based tool built on it can save a file. If you need an audio file you need a server-side TTS service, or you need to record your system audio with a separate application.

Is my text sent to a server?

This page uploads nothing. The caveat is that some of the voices your browser offers, specifically Chrome's Google voices, are synthesized remotely, so choosing one means Chrome sends the text to Google. Pick a voice your operating system provides (Apple on macOS, Microsoft on Windows) and the whole thing stays on the device.

How do I make it sound less robotic?

Change the voice, not the pitch. Each voice is a different synthesis model and the gap in quality between them is far larger than anything the rate and pitch sliders can do. Modern neural voices (the Siri voices on Apple platforms, the Natural voices on Windows 11) sound dramatically better than the legacy ones sitting next to them in the same list.

Recommendations

You Might Also Like

Nearby tools from the catalog that fit the same job or workflow.

Cleanor app

Do it all on your device

Cleanor puts these tools in one app: compress and convert images, video, and audio, work with PDFs, and scan text right on your device. Plus free up storage and clear inbox clutter with Email Cleaner. Start with a free trial.

  • iPhone
  • Android
  • Macsoon
  • Windowssoon