Text to Speech (TTS, Free)

Paste your text, pick a natural AI voice and save it as WAV or MP3 — or use your device voices instantly.

Before you start

Loading the tool...

Hearing text read aloud helps with proofreading, presentation practice and accessibility. This tool offers two modes. The "High-quality AI voice" runs Supertone's open-source Supertonic 3 model directly in this browser, producing natural, human-like narration that you can save as a WAV or MP3 file. The "Device voices" tab uses the speech synthesis built into your browser, so you can use it right away without installing anything or downloading a model.

What is the High-quality AI voice (Supertonic 3)?

Supertonic 3 is an on-device text-to-speech model that supports 31 languages, including Korean, English and Japanese. It runs inference with ONNX Runtime Web entirely inside this browser, so your text never leaves your device. Choose from 5 male and 5 female voices, a speed from 0.8x to 1.5x, and a quality setting of 4, 8 or 16 denoising steps — higher steps sound more natural but take a little longer. You can paste up to 5,000 characters at once; longer text is split into paragraph-sized chunks automatically, generated in order and stitched into a single audio file. The model files (about 400MB) download only the first time; after that they're cached in this browser. If your device supports WebGPU, it's used automatically for much faster generation, with an automatic fallback to WebAssembly (WASM) otherwise.

When to use it

How to use (High-quality AI voice)

  1. Open the "High-quality AI voice" tab and pick a language (Korean, English, Japanese or others) and a voice (Male 1-5, Female 1-5).
  2. Choose a speed (0.8 to 1.5x) and quality (4, 8 or 16 steps), then paste your text (up to 5,000 characters).
  3. Press "Pre-download" to fetch the model ahead of time, or just press "Generate AI voice" to start right away. The model (about 400MB) downloads only once.
  4. The result plays automatically; save it as a WAV or MP3 file.

How to use (Device voices)

  1. On the "Device voices" tab, paste the text into the box.
  2. Choose a voice (from the voices installed on your device), a speed (0.5 to 2.0x) and a pitch (0.5 to 2.0x).
  3. Press the play button to listen.
  4. To save the sound as a file, press "Record tab audio (experimental)" and turn on "Share tab audio" in the sharing window, or the sound won't be recorded.

Supported formats and limits

Open source used

Tips

Troubleshooting

FAQ

Can I save the speech as an audio file?

Yes. Audio made in the "High-quality AI voice" tab can be saved directly as a WAV or MP3 file. The "Device voices" tab has no built-in save feature, so use the screen-recording method instead.

Which model powers the AI voice?

It runs Supertone's open-source Supertonic 3 model directly in this browser (ONNX Runtime Web). It supports 31 languages and lets you pick from 5 male and 5 female voices.

Does the AI voice need an internet connection?

Only the first time, to download the model files (about 400MB). After that they're cached in this browser, so you can use it right away next time.

Is my text sent to a server?

No. Your text and the generated audio are processed entirely on this device. Only the model files come from Hugging Face; the text itself is never sent anywhere.

Why do device voices look different for different people?

The device voice list comes from the speech engines installed on each device and browser, so it differs from device to device. The AI voice always offers the same 10 voices for everyone.

Can I publish audio made with the AI voice?

Yes, but you must disclose that it's AI-generated, and you can't use it to impersonate someone, create deepfakes, do anything illegal, harass anyone, or spread disinformation. See the OpenRAIL-M license for details.

Can it read several paragraphs at once?

Yes. The AI voice accepts up to 5,000 characters; longer text is split into paragraph-sized chunks automatically, read in order and stitched back together.

Learn more

For more detail, see the guide: How to Use Text to Speech and Save It as a File.

Advertisement (tool-below)

More free tools