Skip to main content
AI Text to Speech turns your text into a voiceover. Pick one of about 2,000 voices, or your own, paste your script, and download the audio. Use it for videos, podcasts, ads, e-learning, and phone messages.
AI Text to Speech Generator: a text box with a welcome message for Example Yoga Studio, AI writing buttons, Use Turbo, the voice Ivan - Professional with Change voice, Clone Your Voice, Design a Voice, Advanced Settings, and Generate

Generate speech

1

Open AI Text to Speech

Click AI Text to Speech in the sidebar. Usage at the top right shows how many text-to-speech characters you’ve used this period.
2

Enter your text

Type or paste it in What text do you want to convert to speech? The counter shows how many characters you can use in one go (see Length limits).To get help with the text, use the buttons under the box: Write a script writes one from a topic, Improve, Fix grammar, and Shorten rewrite what you have. Undo brings back your previous text.
3

Pick a voice

Click Change voice and choose one. See Voices and voice cloning.
4

Adjust the settings (optional)

Turn on Use Turbo to use half the characters, or open Advanced Settings to change the file format, speed, and delivery. See below.
5

Click Generate

The audio appears at the top of Generated Audios a few seconds later. Click Play to listen.

Audio tags

Voices marked Audio tags in the voice list can act out stage directions written in square brackets, such as [laughs], [whispers], [sighs], or [exhales]. With one of these voices selected, Add audio tags adds suitable tags to your text for you.
The text box with an [exhales] tag added after the first sentence, the tooltip Insert expressive tags like [laughs] or [whispers] where they fit, and the voice Hope selected
Other voices read the brackets aloud. If your text has tags and the voice doesn’t support them, you’re offered Use a voice with audio tags. The default voice, Hope, supports audio tags.

Turbo

Use Turbo uses a faster model and counts half the characters. Try it with your voice and compare. Turbo isn’t available with audio-tag voices and Professional clones.

Advanced Settings

Advanced Settings: Format MP3 128kbps (Recommended), Speed 1.00x, Stability 70%, and Clarity 75%
Reset to defaults restores all four. Voices marked Saver always produce MP3 and ignore these settings and Turbo.

Length limits

For longer scripts, split the text and generate it in parts.

Your generated audio

Generated Audios shows your 18 newest files. All of them are in History > AI Text-to-Speech.
Generated Audios with four cards, each with the title, the voice Hope, Play, share, details, and download icons, and a player at the bottom playing the first one
Each card has Play, Share, details (ⓘ), and Download. Downloads are named after the voice and the date. Details show the text, the voice, the date, and Characters Used, and have these buttons:
Details of a generated audio: the text, the voice Hope, Created At, Characters Used 152, Status Completed, and Delete, Regenerate, Play, Share, and Download
  • Regenerate puts the text and voice back in the form, so you can change them and generate again. It doesn’t copy the Advanced Settings.
  • Share turns on a public page for the audio, with buttons to post it on LinkedIn, Facebook, X, and Reddit. On a Teams plan you can also share it with your team.
  • Delete removes the audio. Deleted audio can’t be restored.

What it costs

Text to speech uses your plan’s text-to-speech characters: Characters are only counted for audio that was created; a failed generation costs nothing. The AI writing buttons (Write a script, Improve, and so on) use words, not characters. See Audio plans and limits for how many characters each plan includes and what happens when they run out.
On paid plans, you may use the audio commercially, even after your subscription ends. Audio made on the Free plan is for non-commercial use.

For developers

The API makes speech with POST /api/generate-text-to-speech. See Text to speech in the API reference.

Voices and voice cloning

Find a voice, clone yours, or design a new one.

Talking videos

Put your voiceover on a presenter.