Qwen3 Coder Next is an open-weight coding model built on Qwen3-Next-80B-A3B-Base (hybrid attention + MoE). It is agentically trained at scale on executable tasks and environment interaction, delivering strong coding and tool-use performance at lower inference cost. Native 256K context.
Added Dec 8, 2025
Model weightsContext Window
262.1K
Max Output
65.5K
Avg output tokens (7d)
408 tokens
Input Price (Auto)
$0.12/1M
Output Price (Auto)
$0.80/1M
Cache Read (Auto)
$0.070/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
10.1
Coding Index
36.2
Agentic Index
3.6
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
1.1%
Better than 18% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
414 Elo
Better than 22% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
664 Elo
Better than 31% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
2.6%
Better than 21% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
47.0%
Better than 45% of models compared
Reasoning
HLE
Humanity's Last Exam
10.1%
Better than 54% of models compared
IFBench
Instruction-following benchmark
35.2%
Better than 27% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
0.0%
Better than 17% of models compared
SciCode
Python programming for scientific computing
36.2%
Better than 17% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
16.2%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
93.7%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
73.7%
Better than 54% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
18.2%
Better than 57% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
79.5%
Better than 69% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
47.0%
Better than 45% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
8.2%
Last updated Sep 13, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Qwen3 Coder Next with similar models from the same provider or model family.
Qwen 3 Coder 480B
qwen/qwen3-coderQwen 3 Coder 480B, a 480 billion total parameter model with 35B active, and 160 total experts with 8 active. Performs similar to Claude 4 Sonnet in coding benchmarks, but does so at a much lower price.
Qwen3 Coder Flash
qwen/qwen3-coder-flashA speed‑optimized and budget‑friendly sibling to Coder Plus. Excellent at code generation and agentic workflows (tool use, environment interaction) with solid general‑purpose ability.
Qwen3 Coder Plus
qwen/qwen3-coder-plusAlibaba’s proprietary upgrade to the open‑weights Qwen3 Coder 480B A35B. A coding‑first agent model with strong tool use and environment control for autonomous programming, while remaining capable at general tasks.
Qwen3 Next 80B A3B (Instruct)
qwen/qwen3-next-80b-a3b-instructBased on the new Qwen3‑Next architecture (hybrid attention, highly sparse MoE, training‑stability optimizations, and multi‑token prediction), the Qwen3‑Next‑80B‑A3B‑Instruct model delivers extreme efficiency with only 3B active parameters per pass. It performs comparably to Qwen3‑235B‑A22B‑Instruct‑2507 and shows clear advantages on ultra‑long context tasks (up to 256K tokens).
Qwen3 Next 80B A3B (Thinking)
qwen/qwen3-next-80b-a3b-thinkingQwen3 Next 80B A3B (Thinking)
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 262,144-token context window.