Skip to main content

Text to Speech Tool

Try it now! ↗️
Text to Speech turns a script into natural spoken audio using ElevenLabs voices — for voiceovers, narration, spoken-word intros, or podcast segments. Output is MP3 (44.1 kHz, 128 kbps).

How to use it

  1. Paste or write your script — up to 5,000 characters per run (a live counter keeps you honest).
  2. Browse the voice library and preview voices until one fits.
  3. Pick a model: Multilingual v2 (default, best quality), Turbo v2.5, or Flash v2.5 (fastest).
  4. Optionally fine-tune the voice settings: stability, similarity, style, speed (0.7×–1.2×), and speaker boost.
  5. Generate — the result lands in the Speech tab of your library.
Pricing is 80 credits per started 1,000 characters — a 2,500-character script costs 240 credits.

Frequently asked questions

Yes — clone it first with Voice Cloning (ElevenLabs IVC) and it appears in your voice list.
The default Multilingual v2 model speaks dozens of languages — write your script in the target language and pick a voice that suits it.
Yes — anything in your Speech library tab can be pulled into the Studio as a track, or layered into any pipeline that accepts library audio.