Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is a fast omni-modal reasoning model for understanding text, images, audio, and video. It is especially suited to meeting summaries, transcripts and subtitles, speaker-aware audiovisual analysis, and long-form content review. It returns text and supports a nearly one-million-token context window, tool calling, and structured output.

  • Reasoning
  • Vision
  • Audio Input
  • Video Input
  • Tool Calling
  • Structured Output

Added Sep 17, 2026

Pricing

Auto routing · per 1M tokens
Input
$0.15
Output
$0.47
Cache read
$0.016
Compare provider prices

Specifications

Context window
991.8K
Max output
131.1K
Avg output (7d)
233 tokens
Longer than 18% of models

Benchmarks

No public benchmark scores for this model yet.

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…