ElevenLabs Scribe V1 transcription with word-level timestamps and speaker identification
Approx. Price
$0.051 per minute
Model Type
speech-to-text
Category
speech-to-text
Preview Examples
0
Related audio models
Compare ElevenLabs Scribe V1 with similar models from the same provider or model family.
ElevenLabs Scribe V2
Elevenlabs-Scribe-V2ElevenLabs Scribe V2 transcription with improved accuracy, word-level timestamps, and speaker identification
ElevenLabs Turbo V2.5
Elevenlabs-Turbo-V2.5High quality with lowest latency, ideal for real-time applications. Supports 32 languages while maintaining natural voice quality.
ElevenLabs v3
Elevenlabs-V3High-quality text-to-speech with enhanced controls and natural voices.
Qwen Audio 3.0 TTS Flash
alibaba/qwen-audio-3-ttsFast multilingual text-to-speech with a broad set of voices and explicit language control.
Demucs Stem Separation
fal-ai/demucsSeparate up to 60 minutes / 700 MiB of stereo audio into vocals, drums, bass, guitar, piano, and other stems with fal Demucs.
Stable Audio 3 Medium
fal-ai/stable-audio-3/medium/text-to-audioStable Audio 3 Medium generates high-quality stereo music up to 6 minutes from text prompts, trained on fully licensed data for commercial use.