Qwen 3 Coder 480B, a 480 billion total parameter model with 35B active, and 160 total experts with 8 active. Performs similar to Claude 4 Sonnet in coding benchmarks, but does so at a much lower price.
Added Mar 17, 2026
Model weightsContext Window
262.0K
Max Output
65.5K
Input Price (Auto)
$0.13/1M
Output Price (Auto)
$0.50/1M
Cache Read (Auto)
$0.065/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
18.2
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
61.8%
Better than 39% of models compared
HLE
Humanity's Last Exam
4.5%
Better than 24% of models compared
IFBench
Instruction-following benchmark
40.5%
Better than 39% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
43.6%
Better than 48% of models compared
AA-LCR
Long context reasoning evaluation
45.3%
Better than 50% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
SciCode
Python programming for scientific computing
35.9%
Better than 55% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
18.9%
Better than 59% of models compared
LiveCodeBench
Contamination-free coding benchmark
58.5%
Better than 64% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
39.3%
Better than 41% of models compared
AIME
American Invitational Mathematics Examination
47.7%
Better than 68% of models compared
Math-500
Diverse mathematical problem solving benchmark
94.2%
Better than 76% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
78.8%
Better than 62% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
15.7%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
44.9%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Qwen 3 Coder 480B with similar models from the same provider or model family.
Qwen3 Coder Next
qwen/qwen3-coder-nextQwen3 Coder Next is an open-weight coding model built on Qwen3-Next-80B-A3B-Base (hybrid attention + MoE). It is agentically trained at scale on executable tasks and environment interaction, delivering strong coding and tool-use performance at lower inference cost. Native 256K context.
Qwen3 Coder Flash
qwen/qwen3-coder-flashA speed‑optimized and budget‑friendly sibling to Coder Plus. Excellent at code generation and agentic workflows (tool use, environment interaction) with solid general‑purpose ability.
Qwen3 Coder Plus
qwen/qwen3-coder-plusAlibaba’s proprietary upgrade to the open‑weights Qwen3 Coder 480B A35B. A coding‑first agent model with strong tool use and environment control for autonomous programming, while remaining capable at general tasks.
Qwen 3.6 35B A3B Uncensored
qwen/qwen3.6-35b-a3b-uncensoredQwen 3.6 35B A3B Uncensored is an NVFP4 open-weight mixture-of-experts model LoRA-tuned for fewer refusals across chat, coding, tool use, and multimodal tasks.
Qwen 3.8 27B Uncensored
qwen/qwen3.8-27b-uncensoredQwen 3.8 27B Uncensored is an NVFP4 open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, tool use, and long-context work.
Qwen3.6 35B A3B Thinking
Qwen/Qwen3.6-35B-A3B:thinkingQwen3.6 35B A3B is a native vision-language MoE model with hybrid attention. Compared to Qwen3.5 35B A3B, Alibaba reports stronger agentic coding, mathematical and code reasoning, and better spatial understanding (including object localization and detection).