High-quality text-to-speech with enhanced controls and natural voices.
Approx. Price
$170.00 per 1M characters
Model Type
text-to-speech
Category
text-to-speech
Preview Examples
0
Available voices
Built-in voices accepted by NanoGPT for this model. Custom or cloned voice IDs may also be supported.
Related audio models
Compare ElevenLabs v3 with similar models from the same provider or model family.
ElevenLabs Scribe V2
Elevenlabs-Scribe-V2ElevenLabs Scribe V2 transcription with improved accuracy, word-level timestamps, and speaker identification
ElevenLabs Scribe V1
Elevenlabs-STTElevenLabs Scribe V1 transcription with word-level timestamps and speaker identification
ElevenLabs Turbo V2.5
Elevenlabs-Turbo-V2.5High quality with lowest latency, ideal for real-time applications. Supports 32 languages while maintaining natural voice quality.
ElevenLabs Sound Effects V2
fal-ai/elevenlabs/sound-effects/v2Generate sound effects and seamless loops from a text description.
Qwen Audio 3.0 TTS Flash
alibaba/qwen-audio-3-ttsFast multilingual text-to-speech with a broad set of voices and explicit language control.
Demucs Stem Separation
fal-ai/demucsSeparate up to 60 minutes / 700 MiB of stereo audio into vocals, drums, bass, guitar, piano, and other stems with Demucs.