High-definition text-to-speech with natural pronunciation and multiple voices.
Approx. Price
$100.00 per 1M characters
Model Type
text-to-speech
Category
text-to-speech
Preview Examples
0
Available voices
Built-in voices accepted by NanoGPT for this model. Custom or cloned voice IDs may also be supported.
Related audio models
Compare MiniMax Speech 02 HD with similar models from the same provider or model family.
MiniMax Speech 2.6 HD
Minimax-Speech-2.6-HDUltra-Fast, Ultra-Human, Ultra-Smart TTS with <250ms latency, natural voice cloning, seamless multilingual support across 40+ languages, and industry-leading text normalization for flawless, expressive communication.
MiniMax Speech 2.6 Turbo
Minimax-Speech-2.6-TurboHigh-definition Text-to-Speech with natural pronunciation and crisp articulation. Supports multiple built-in voices and custom cloned voices, adjustable speed, volume, and pitch, and coverage of 40+ languages for professional audio creation.
MiniMax Speech 2.8 HD
Minimax-Speech-2.8-HDStudio-quality HD text-to-speech with expressive delivery, emotion control, pronunciation customization, and fine-grained audio controls for production-ready speech.
MiniMax Speech 2.8 Turbo
Minimax-Speech-2.8-TurboFast, cost-effective MiniMax 2.8 text-to-speech with expressive voices, emotion control, pronunciation customization, and full audio output controls.
ByteDance Seed Speech TTS 2.0
bytedance/seed-speech-tts-2.0ByteDance Seed Speech TTS 2.0 for natural multilingual speech with voice instructions and delivery controls.
MiniMax Music 02
Minimax-Music-02MiniMax Music 02 is a compact MoE music generator (230B params, 10B active) tuned for speedy, cost-effective song creation. Provide a creative prompt plus formatted lyrics to render polished, full-length tracks with configurable bitrate and sample rate.