Qwen 3.5 Plus with extended reasoning. A commercial model with hybrid linear attention and sparse MoE architecture. Supports text, image, and video input with a 1M context window.
Added Feb 16, 2026
Context Window
983.6K
Max Output
65.5K
Input Price (Auto)
$0.40/1M
Output Price (Auto)
$2.40/1M
Cache Read (Auto)
$0.040/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Vectara.
Hallucination Rate
10.7%
Factual Consistency
89.3%
Answer Rate
99.8%
Avg Summary Length
Average generated summary length
92.1
Last updated 2026-05-11 · Matched as qwen/qwen3.5-plus-2026-02-15
Vectara LeaderboardProviders
Provider information for this model’s automatic routing. These routes cannot be selected individually.
Loading provider options…
Related text models
Compare Qwen3.5 Plus Thinking with similar models from the same provider or model family.
Qwen3.5 Plus
qwen/qwen3.5-plusQwen 3.5 Plus is a commercial model with hybrid linear attention and sparse MoE architecture. Supports text, image, and video input with a 1M context window.
Qwen3 Coder Plus
qwen/qwen3-coder-plusAlibaba’s proprietary upgrade to the open‑weights Qwen3 Coder 480B A35B. A coding‑first agent model with strong tool use and environment control for autonomous programming, while remaining capable at general tasks.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 262,144-token context window.
Qwen 3.8 27B Fable
qwen/qwen3.8-27b-fableQwen 3.8 27B Fable is an open-weight multimodal creative finetune for expressive dialogue, long-form storytelling, character work, and roleplay.
Qwen 3.8 27B Obliterated
qwen/qwen3.8-27b-obliteratedQwen 3.8 27B Obliterated is an open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.
Qwen 3.8 27B Obliterated Thinking
qwen/qwen3.8-27b-obliterated:thinkingQwen 3.8 27B Obliterated with thinking enabled for more deliberate creative work, coding, multimodal analysis, tool use, and long-context problem solving.