Alibaba’s proprietary upgrade to the open‑weights Qwen3 Coder 480B A35B. A coding‑first agent model with strong tool use and environment control for autonomous programming, while remaining capable at general tasks.
Added Sep 17, 2025
Context Window
128.0K
Max Output
65.5K
Input Price (Auto)
$1.00/1M
Output Price (Auto)
$5.00/1M
Cache Read (Auto)
$0.50/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
18.2
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
61.8%
Better than 39% of models compared
HLE
Humanity's Last Exam
4.5%
Better than 24% of models compared
IFBench
Instruction-following benchmark
40.5%
Better than 39% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
43.6%
Better than 48% of models compared
AA-LCR
Long context reasoning evaluation
45.3%
Better than 50% of models compared
Coding
SciCode
Python programming for scientific computing
35.9%
Better than 55% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
18.9%
Better than 59% of models compared
LiveCodeBench
Contamination-free coding benchmark
58.5%
Better than 64% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
39.3%
Better than 41% of models compared
AIME
American Invitational Mathematics Examination
47.7%
Better than 68% of models compared
Math-500
Diverse mathematical problem solving benchmark
94.2%
Better than 76% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
78.8%
Better than 62% of models compared
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Qwen3 Coder Plus with similar models from the same provider or model family.
Qwen 3 Coder 480B
qwen/qwen3-coderQwen 3 Coder 480B, a 480 billion total parameter model with 35B active, and 160 total experts with 8 active. Performs similar to Claude 4 Sonnet in coding benchmarks, but does so at a much lower price.
Qwen3.5 Plus
qwen/qwen3.5-plusQwen 3.5 Plus is a commercial model with hybrid linear attention and sparse MoE architecture. Supports text, image, and video input with a 1M context window.
Qwen3.5 Plus Thinking
qwen/qwen3.5-plus-thinkingQwen 3.5 Plus with extended reasoning. A commercial model with hybrid linear attention and sparse MoE architecture. Supports text, image, and video input with a 1M context window.
Qwen3 Coder Next
qwen/qwen3-coder-nextQwen3 Coder Next is an open-weight coding model built on Qwen3-Next-80B-A3B-Base (hybrid attention + MoE). It is agentically trained at scale on executable tasks and environment interaction, delivering strong coding and tool-use performance at lower inference cost. Native 256K context.
Qwen3 Coder Flash
qwen/qwen3-coder-flashA speed‑optimized and budget‑friendly sibling to Coder Plus. Excellent at code generation and agentic workflows (tool use, environment interaction) with solid general‑purpose ability.
Qwen 3.6 35B A3B Uncensored
qwen/qwen3.6-35b-a3b-uncensoredQwen 3.6 35B A3B Uncensored is an NVFP4 open-weight mixture-of-experts model LoRA-tuned for fewer refusals across chat, coding, tool use, and multimodal tasks.