Google Gemini 3.1 Flash text-to-speech with inline audio tag and multi-speaker prompt support.
Approx. Price
$102.42 per 1M characters
Model Type
text-to-speech
Category
text-to-speech
Preview Examples
0
Available voices
Built-in voices accepted by NanoGPT for this model. Custom or cloned voice IDs may also be supported.
Related audio models
Compare Gemini 3.1 Flash TTS Preview with similar models from the same provider or model family.
Gemini 2.5 Flash Preview TTS
gemini-2.5-flash-preview-ttsGoogle Gemini native TTS. Single and multi-speaker support via prompt.
Qwen Audio 3.0 TTS Flash
alibaba/qwen-audio-3-ttsFast multilingual text-to-speech with a broad set of voices and explicit language control.
Gemini 2.5 Pro Preview TTS
gemini-2.5-pro-preview-ttsHigher-quality Gemini TTS with controllable style and tone.
ByteDance Seed Speech TTS 2.0
bytedance/seed-speech-tts-2.0ByteDance Seed Speech TTS 2.0 for natural multilingual speech with voice instructions and delivery controls.
Alibaba Fun-ASR Flash
fun-asr-flash-2026-06-15Alibaba Cloud DashScope non-realtime speech recognition with multilingual transcription, punctuation, and sentence/word timestamps.
GPT-4o Mini TTS
gpt-4o-mini-ttsUltra-low cost text-to-speech model with voice instructions support