Browse all Qwen text models
Provider logo

Qwen3.5 Omni Flash

qwen/qwen3.5-omni-flash
Provider logo

Qwen3.5 Omni Flash

qwen/qwen3.5-omni-flash

Qwen3.5 Omni Flash is Qwen's fast multimodal model. We verified live support for text prompts, images, audio files, and direct video URLs on Alibaba's chat-completions-compatible API. Alibaba describes Flash as a fully evolved version of Qwen3 Omni with audio input support across 60+ languages.

Added Mar 30, 2026

Context Window

49.2K

Max Output

16.4K

Input Price (Auto)

$0.43/1M

Output Price (Auto)

$1.66/1M

Cache Read (Auto)

$0.043/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

12.5

Better than 51% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

84.5%

Better than 76% of models compared

Document reasoning

AA-LCR v1.1

Long context reasoning with updated grading

52.0%

Better than 47% of models compared

Reasoning

HLE

Humanity's Last Exam

7.6%

Better than 47% of models compared

IFBench

Instruction-following benchmark

38.0%

Better than 33% of models compared

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

8.3%

Better than 42% of models compared

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

74.2%

Better than 55% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

52.0%

Better than 47% of models compared

Last updated Sep 20, 2026

Artificial Analysis

Providers

Provider information for this model’s automatic routing. These routes cannot be selected individually.

Loading provider options…

Compare Qwen3.5 Omni Flash with similar models from the same provider or model family.

Qwen3.8 Omni Flash

qwen/qwen3.8-omni-flash

Qwen3.8 Omni Flash is a fast omni-modal reasoning model for understanding text, images, audio, and video. It is especially suited to meeting summaries, transcripts and subtitles, speaker-aware audiovisual analysis, and long-form content review. It returns text and supports a nearly one-million-token context window, tool calling, and structured output.

Qwen3.8 Flash

qwen/qwen3.8-flash

Qwen3.8 Flash is Alibaba's latest fast multimodal model, with a million-token context window for coding, agentic workflows, visual understanding, long documents, codebases, and videos.

Qwen3.7 Flash

qwen/qwen3.7-flash

Qwen3.7 Flash is Qwen's fast multimodal model for coding, search and computer-use agents, visual understanding, object recognition, spatial reasoning, and stable end-to-end task execution.

Qwen3.7 Flash Thinking

qwen/qwen3.7-flash:thinking

Qwen3.7 Flash with thinking enabled for deeper multimodal reasoning, coding, search and computer-use agents, spatial reasoning, and multi-step task execution.

Qwen3.6 Flash

qwen/qwen3.6-flash

Qwen3.6 Flash is Alibaba's fast native vision-language model in the Qwen 3.6 family. It improves over 3.5 Flash with stronger coding/agent performance and better spatial intelligence, including object localization and detection.

Qwen3.5 Omni Plus

qwen/qwen3.5-omni-plus

Qwen3.5 Omni Plus is Qwen's stronger general multimodal model. We verified live support for text prompts, images, audio files, and direct video URLs on Alibaba's chat-completions-compatible API. Alibaba describes Plus as a comprehensive evolution of Qwen3 Omni with support for over 10 hours of audio input.