Qwen 2.5 Max is the upgraded version of Qwen Max, beating GPT-4o, Deepseek V3 and Claude 3.5 Sonnet in benchmarks.
Added Oct 1, 2024
Context Window
32.0K
Max Output
8.2K
Input Price (Auto)
$1.60/1M
Output Price (Auto)
$6.39/1M
Cache Read (Auto)
$0.80/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1373.8
Overall Rank
#177 / 399
Votes
32,417
Confidence Interval
1369.7 - 1377.9
Category Scores
Coding
#193 / 394
5,076 votes
1402.2
Math
#186 / 382
3,298 votes
1363.0
Longer Query
#170 / 377
4,522 votes
1384.9
Creative Writing
#159 / 397
4,994 votes
1353.1
Instruction Following
#183 / 399
10,937 votes
1356.9
Hard Prompts
#183 / 399
9,616 votes
1385.0
Additional Categories22
French
#139 / 278
343 votes
1411.9
Japanese
#146 / 261
727 votes
1309.3
German
#157 / 296
781 votes
1354.1
Korean
#160 / 265
465 votes
1313.7
Spanish
#160 / 279
268 votes
1373.9
Polish
#161 / 219
849 votes
1358.8
Industry Writing And Literature And Language
#166 / 398
8,174 votes
1362.2
Russian
#171 / 363
2,980 votes
1367.7
Industry Legal And Government
#172 / 372
1,856 votes
1388.3
Industry Entertainment And Sports And Media
#173 / 397
5,968 votes
1339.4
Non English
#173 / 399
14,251 votes
1360.6
Exclude Ties
#174 / 399
21,825 votes
1348.9
Chinese
#175 / 370
2,074 votes
1396.9
Industry Life And Physical And Social Science
#176 / 397
5,665 votes
1390.2
Industry Medicine And Healthcare
#176 / 368
1,588 votes
1393.6
Multi Turn
#176 / 397
4,815 votes
1373.1
Industry Business And Management And Financial Operations
#180 / 392
3,656 votes
1368.3
Industry Mathematical
#183 / 376
2,940 votes
1366.7
English
#186 / 399
18,166 votes
1382.2
Expert
#190 / 349
1,680 votes
1367.9
Industry Software And It Services
#193 / 399
8,534 votes
1397.6
Hard Prompts English
#196 / 397
5,689 votes
1386.7
Published 2026-09-02 · Matched as qwen2.5-max
LMArena DatasetProviders
Provider information for this model’s automatic routing. These routes cannot be selected individually.
Loading provider options…
Related text models
Compare Qwen 2.5 Max with similar models from the same provider or model family.
Qwen3 Max
qwen/qwen3-maxQwen3 Max. The latest Qwen 3 model (5 september 2025). Higher accuracy in coding and science, better instruction following, and optimized for tool calling.
Qwen: QvQ Max
qvq-maxQvQ Max is the top model of the Qwen series. QvQ Max is capable of thinking and reasoning, can achieve significantly enhanced performance especially on hard problems.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 262,144-token context window.
Qwen3.8 Max 0902
qwen/qwen3.8-max-0902Qwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.
Qwen3.8 Max
qwen/qwen3.8-maxQwen3.8 Max is Qwen's 2.4T-parameter flagship model for coding, knowledge work, full-stack development, data analysis, and long-running agent workflows in non-thinking mode. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.
Qwen3.8 Max Thinking
qwen/qwen3.8-max:thinkingQwen3.8 Max Thinking enables generation-time reasoning for deeper coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.