Qwen3.7 Plus with thinking mode enabled for deeper multimodal reasoning, coding, tool use, screen reading, and productivity workflows.
Added Jun 1, 2026
Context Window
983.6K
Max Output
65.5K
Avg output tokens (7d)
2.2K tokens
Input Price (Auto)
$0.40/1M
Output Price (Auto)
$1.60/1M
Cache Read (Auto)
$0.080/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
25.8
Coding Index
55.9
Agentic Index
19.7
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
17.4%
Better than 43% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
913 Elo
Better than 48% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
886 Elo
Better than 42% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
12.2%
Better than 54% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
73.0%
Better than 73% of models compared
Reasoning
HLE
Humanity's Last Exam
35.6%
Better than 87% of models compared
IFBench
Instruction-following benchmark
78.0%
Better than 97% of models compared
CritPt
Research-level physics reasoning
9.1%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
1.0%
Better than 46% of models compared
SciCode
Python programming for scientific computing
46.1%
Better than 46% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
22.5%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
27.7%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
90.0%
Better than 89% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
47.0%
Better than 94% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
93.0%
Better than 89% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
73.0%
Better than 73% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
19.3%
Last updated Sep 10, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Qwen3.7 Plus Thinking with similar models from the same provider or model family.
Qwen3.7 Plus
qwen3.7-plusQwen3.7 Plus is Alibaba's cost-effective Qwen 3.7 multimodal agent model for coding, tool use, productivity workflows, visual understanding, screen reading, and GUI interaction.
Qwen3.5 Omni Plus
qwen3.5-omni-plusQwen3.5 Omni Plus is Qwen's stronger general multimodal model. We verified live support for text prompts, images, audio files, and direct video URLs on Alibaba's chat-completions-compatible API. Alibaba describes Plus as a comprehensive evolution of Qwen3 Omni with support for over 10 hours of audio input.
Qwen3.5 Plus
qwen/qwen3.5-plusQwen 3.5 Plus is a commercial model with hybrid linear attention and sparse MoE architecture. Supports text, image, and video input with a 1M context window.
Qwen3.5 Plus Thinking
qwen/qwen3.5-plus-thinkingQwen 3.5 Plus with extended reasoning. A commercial model with hybrid linear attention and sparse MoE architecture. Supports text, image, and video input with a 1M context window.
Qwen3 Coder Plus
qwen/qwen3-coder-plusAlibaba’s proprietary upgrade to the open‑weights Qwen3 Coder 480B A35B. A coding‑first agent model with strong tool use and environment control for autonomous programming, while remaining capable at general tasks.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 262,144-token context window.