Qwen3.8 Omni Flash
Qwen3.8 Omni Flash is a fast omni-modal reasoning model for understanding text, images, audio, and video. It is especially suited to meeting summaries, transcripts and subtitles, speaker-aware audiovisual analysis, and long-form content review. It returns text and supports a nearly one-million-token context window, tool calling, and structured output.
- Reasoning
- Vision
- Audio Input
- Video Input
- Tool Calling
- Structured Output
Added Sep 17, 2026
Pricing
Auto routing · per 1M tokens- Input
- $0.15
- Output
- $0.47
- Cache read
- $0.016
Specifications
- Context window
- 991.8K
- Max output
- 131.1K
- Avg output (7d)
- 233 tokens
- Longer than 18% of models
Benchmarks
No public benchmark scores for this model yet.
Providers
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…