Browse all Qwen text models
Provider logo

Qwen3.5 Flash Thinking

qwen/qwen3.5-flash:thinking
Provider logo

Qwen3.5 Flash Thinking

qwen/qwen3.5-flash:thinking

Qwen3.5 Flash with extended reasoning enabled. The fastest and cheapest native Qwen 3.5 vision-language model.

Added Feb 24, 2026

Context Window

991.8K

Max Output

65.5K

Input Price (Auto)

$0.10/1M

Output Price (Auto)

$0.40/1M

Cache Read (Auto)

$0.050/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Vectara.

Hallucination Rate

10.5%

Better than 42% of models compared

Factual Consistency

89.5%

Better than 42% of models compared

Answer Rate

99.8%

Better than 65% of models compared

Avg Summary Length

Average generated summary length

95.0

Last updated 2026-05-11 · Matched as qwen/qwen3.5-flash-2026-02-23

Vectara Leaderboard

Providers

Provider information for this model’s automatic routing. These routes cannot be selected individually.

Loading provider options…

Compare Qwen3.5 Flash Thinking with similar models from the same provider or model family.

Qwen3.8 Omni Flash

qwen/qwen3.8-omni-flash

Qwen3.8 Omni Flash is a fast omni-modal reasoning model for understanding text, images, audio, and video. It is especially suited to meeting summaries, transcripts and subtitles, speaker-aware audiovisual analysis, and long-form content review. It returns text and supports a nearly one-million-token context window, tool calling, and structured output.

Qwen3.8 Flash

qwen/qwen3.8-flash

Qwen3.8 Flash is Alibaba's latest fast multimodal model, with a million-token context window for coding, agentic workflows, visual understanding, long documents, codebases, and videos.

Qwen3.7 Flash

qwen/qwen3.7-flash

Qwen3.7 Flash is Qwen's fast multimodal model for coding, search and computer-use agents, visual understanding, object recognition, spatial reasoning, and stable end-to-end task execution.

Qwen3.7 Flash Thinking

qwen/qwen3.7-flash:thinking

Qwen3.7 Flash with thinking enabled for deeper multimodal reasoning, coding, search and computer-use agents, spatial reasoning, and multi-step task execution.

Qwen3.6 Flash

qwen/qwen3.6-flash

Qwen3.6 Flash is Alibaba's fast native vision-language model in the Qwen 3.6 family. It improves over 3.5 Flash with stronger coding/agent performance and better spatial intelligence, including object localization and detection.

Qwen3.5 Omni Flash

qwen/qwen3.5-omni-flash

Qwen3.5 Omni Flash is Qwen's fast multimodal model. We verified live support for text prompts, images, audio files, and direct video URLs on Alibaba's chat-completions-compatible API. Alibaba describes Flash as a fully evolved version of Qwen3 Omni with audio input support across 60+ languages.