Qwen 2.5 Max is the upgraded version of Qwen Max, beating GPT-4o, Deepseek V3 and Claude 3.5 Sonnet in benchmarks.
Added Oct 1, 2024
Context Window
32.0K
Max Output
8.2K
Input Price (Auto)
$1.60/1M
Output Price (Auto)
$6.39/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1373.8
Overall Rank
#188 / 409
Votes
32,410
Confidence Interval
1369.7 - 1378.0
Category Scores
Coding
#203 / 404
5,074 votes
1401.6
Math
#196 / 393
3,298 votes
1362.8
Longer Query
#180 / 387
4,520 votes
1385.0
Creative Writing
#168 / 407
4,993 votes
1353.2
Instruction Following
#194 / 409
10,931 votes
1356.7
Hard Prompts
#191 / 409
9,614 votes
1385.2
Additional Categories22
French
#146 / 287
343 votes
1411.5
Japanese
#154 / 271
727 votes
1309.2
German
#167 / 306
780 votes
1354.8
Korean
#169 / 276
465 votes
1314.0
Polish
#169 / 229
850 votes
1358.5
Spanish
#173 / 292
268 votes
1374.2
Industry Writing And Literature And Language
#175 / 408
8,172 votes
1362.5
Russian
#180 / 373
2,979 votes
1368.3
Industry Entertainment And Sports And Media
#182 / 407
5,967 votes
1339.7
Industry Legal And Government
#182 / 382
1,854 votes
1387.9
Non English
#182 / 409
14,247 votes
1360.7
Exclude Ties
#185 / 409
21,822 votes
1349.0
Industry Medicine And Healthcare
#185 / 379
1,588 votes
1394.2
Industry Life And Physical And Social Science
#186 / 407
5,663 votes
1390.3
Multi Turn
#186 / 407
4,813 votes
1372.5
Chinese
#187 / 385
2,067 votes
1397.0
Industry Business And Management And Financial Operations
#188 / 402
3,655 votes
1368.6
English
#195 / 409
18,163 votes
1382.0
Industry Mathematical
#197 / 390
2,940 votes
1366.5
Expert
#201 / 359
1,680 votes
1368.1
Industry Software And It Services
#203 / 409
8,531 votes
1397.3
Hard Prompts English
#206 / 407
5,689 votes
1386.4
Published 2026-09-25 · Matched as qwen2.5-max
LMArena DatasetProviders
Provider information for this model’s automatic routing. These routes cannot be selected individually.
Loading provider options…
Related text models
Compare Qwen 2.5 Max with similar models from the same provider or model family.
Qwen3 Max
qwen/qwen3-maxQwen3 Max improves accuracy in coding and science, instruction following, and tool calling.
Qwen: QvQ Max
qvq-maxQvQ Max is the top model of the Qwen series. QvQ Max is capable of thinking and reasoning, can achieve significantly enhanced performance especially on hard problems.
Qwen 3.8 27B Hemingway
qwen/qwen3.8-27b-hemmingwayQwen 3.8 27B Hemingway is an open-weight NVFP4 multimodal creative finetune for long-form prose, character dialogue, storytelling, and roleplay.
Qwen 3.8 27B Cybersecurity
qwen/qwen3.8-27b-cybersecurityQwen 3.8 27B Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports optional reasoning, image understanding, tool calling, and a 262,144-token context window.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an FP8 open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 524,288-token context window.
Qwen3.8 Max 0902
qwen/qwen3.8-max-0902Qwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.