Massively multilingual zero-shot text-to-speech with auto voice mode and optional natural-language voice descriptions.
Approx. Price
$50.00 per 1M characters ($0.005 minimum)
Model Type
text-to-speech
Category
text-to-speech
Preview Examples
0
Available voices
Built-in voices accepted by NanoGPT for this model. Custom or cloned voice IDs may also be supported.
Related audio models
Compare Omnivoice with similar models from the same provider or model family.
ACE-Step 1.5
ACE-Step-1.5ACE-Step 1.5 composes complete songs from text descriptions. Guide genre, mood, and structure with style tags and required custom lyrics. Generates up to 4 minutes of multi-track audio with vocals.
ACE-Step v1.5 Base
ACE-Step-v1.5-BaseACE-Step v1.5 Base is a Runware-hosted music model built for creator workflows, with stronger fidelity, more reliable stylistic consistency, and prompt-driven genre control for full-song generation.
ACE-Step v1.5 Turbo
ACE-Step-v1.5-TurboACE-Step v1.5 Turbo is the faster, lower-cost Runware variant for full-song generation, with broad genre coverage, improved stylistic consistency, and text-guided music creation for creator workflows.
Qwen Audio 3.0 TTS Flash
alibaba/qwen-audio-3-ttsFast multilingual text-to-speech with a broad set of voices and explicit language control.
ByteDance Seed Audio 1.0
bytedance/seed-audio-1.0ByteDance Seed Audio 1.0 generates natural audio from text, with optional preset voices, up to three reference audio clips, or a single reference image.
ByteDance Seed Speech TTS 2.0
bytedance/seed-speech-tts-2.0ByteDance Seed Speech TTS 2.0 for natural multilingual speech with voice instructions and delivery controls.