OpenAI's state-of-the-art speech recognition model
Approx. Price
<$0.001 per minute
Model Type
speech-to-text
Category
speech-to-text
Preview Examples
0
Related audio models
Compare Whisper Large V3 with similar models from the same provider or model family.
OpenAI Whisper With Video
openai-whisper-with-videoWhisper large-v3 video-to-text with translation and optional timestamps
Whisper-1
whisper-1OpenAI's original Whisper model with full format support including SRT and VTT subtitles
ACE-Step 1.5
ACE-Step-1.5ACE-Step 1.5 composes complete songs from text descriptions. Guide genre, mood, and structure with style tags and required custom lyrics. Generates up to 4 minutes of multi-track audio with vocals.
ACE-Step v1.5 Base
ACE-Step-v1.5-BaseACE-Step v1.5 Base is a Runware-hosted music model built for creator workflows, with stronger fidelity, more reliable stylistic consistency, and prompt-driven genre control for full-song generation.
ACE-Step v1.5 Turbo
ACE-Step-v1.5-TurboACE-Step v1.5 Turbo is the faster, lower-cost Runware variant for full-song generation, with broad genre coverage, improved stylistic consistency, and text-guided music creation for creator workflows.
Qwen Audio 3.0 TTS Flash
alibaba/qwen-audio-3-ttsFast multilingual text-to-speech with a broad set of voices and explicit language control.