Alibaba (Qwen), proprietary
Qwen Max
Qwen Max by Alibaba (Qwen) ranks 230th of 354 ranked models on the Noometry Index as of October 2026, with a score of 34.7. Its strongest category is reasoning, where it ranks 151st. API pricing starts at $1.60 per million input tokens and $6.40 per million output tokens, with a 33K-token context window.
Last verified
Specifications
- Noometry rank
- #230 of 354
- Index score
- 34.7
- Evidence
- Confirmed 23 results
- Provider
Alibaba (Qwen)
- Released
- April 3, 2024
- Weights
- Proprietary
- Reasoning
- No
- Context window
- 33K
- Max output
- 8K
- Input price
- $1.60 / M
- Output price
- $6.40 / M
- Blended price
- $2.80 / M
- Output speed
- Not measured
- Value
- #176 of 219
- Knowledge cutoff
- April 2024
- Input
- text
Category scores
Each category score combines every public result we have in that category.
- Coding 30.7
- Reasoning 25.1
- Math 22.3
- Knowledge 30.3
- Multilingual 41.8
- Instruction Following 66.5
- Long Context 39.4
- Writing & Preference 47.8
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 30.7 | #292 | 2 |
| Reasoning | 25.1 | #151 | 1 |
| Math | 22.3 | #276 | 3 |
| Knowledge | 30.3 | #228 | 2 |
| Multilingual | 41.8 | #202 | 1 |
| Instruction Following | 66.5 | #208 | 1 |
| Long Context | 39.4 | #180 | 2 |
| Writing & Preference | 47.8 | #205 | 3 |
Strengths and weaknesses
Categories where Qwen Max places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 25.1 | +1.5 | #151 of 350, top 44% |
| Long Context | 39.4 | −1.5 | #180 of 296, top 61% |
| Writing & Preference | 47.8 | −6.0 | #205 of 312, top 66% |
Closest competitors
The models ranked just above and below Qwen Max. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| DeepSeek-R1-Distill-Qwen-32B | #226 | 35.5 | — | — | Compare |
| Magistral Medium | #227 | 35.2 | $2.75 | 0 | Compare |
| Gemini 2.0 Flash (Feb 2025) | #228 | 35.1 | — | 92 | Compare |
| C4ai Aya Expanse 8b | #229 | 34.9 | — | — | Compare |
| Claude 3.5 Sonnet | #231 | 34.6 | — | — | Compare |
| Qwen3 Coder Next | #232 | 34.3 | $0.29 | — | Compare |
| Devstral Small 2505 | #233 | 34.3 | $0.15 | 88 | Compare |
| Qwen1.5-110B | #234 | 34.2 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Aider Polyglot | 21.8% | #34 of 44, top 78% | Epoch AI | ||
| LMArena Coding | 1288 | #206 of 294, top 71% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1269 | #206 of 297, top 70% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 16.1% | #130 of 173, top 76% | Epoch AI | 2025-04-01 | |
| LMArena Math | 1275 | #194 of 285, top 69% | LMArena | 2026-10-08 | |
| MATH Level 5 | 67.2% | #33 of 79, top 42% | Epoch AI | 2025-04-01 | |
| FrontierMath (Feb 2025 set) | 1% | #60 of 68, top 89% | Epoch AI | 2025-04-02 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 56.1% | #118 of 186, top 64% | Epoch AI | 2025-04-01 | |
| LMArena Expert | 1248 | #197 of 273, top 73% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1263 | #202 of 297, top 69% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1254 | #205 of 285, top 72% | LMArena | 2026-10-08 | |
| LMArena French | 1330 | #152 of 223, top 69% | LMArena | 2026-10-08 | |
| LMArena German | 1254 | #175 of 231, top 76% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1205 | #159 of 211, top 76% | LMArena | 2026-10-08 | |
| LMArena Korean | 1142 | #181 of 213, top 85% | LMArena | 2026-10-08 | |
| LMArena Russian | 1274 | #194 of 283, top 69% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1290 | #164 of 226, top 73% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1262 | #201 of 298, top 68% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Fiction.LiveBench | 66.7% | #22 of 47, top 47% | Epoch AI | ||
| LMArena Longer Query | 1288 | #199 of 291, top 69% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1282 | #208 of 297, top 71% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1248 | #208 of 295, top 71% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1277 | #203 of 295, top 69% | LMArena | 2026-10-08 |
API pricing by provider
Compare Qwen Max
- Qwen Max vs C4ai Aya Expanse 8b
- Qwen Max vs Claude 3.5 Sonnet
- Qwen Max vs Gemini 2.0 Flash (Feb 2025)
- Qwen Max vs Qwen3 Coder Next
- Qwen Max vs Magistral Medium
- Qwen Max vs Devstral Small 2505
- Qwen Max vs GPT-6 Astra
- Qwen Max vs Claude Fable 5.1
- Qwen Max vs Gemini 3.8 Flash
- Qwen Max vs Kimi K3
- Qwen Max vs Grok 4.6
- Qwen Max vs GLM-5.3
- Qwen Max vs Muse Spark 1.3
- Qwen Max vs DeepSeek V4 Pro
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.7 Max51.5
- Qwen3.6 Max Preview51.5
- Qwen3.6 Plus47.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
Frequently asked questions
How good is Qwen Max?
Qwen Max by Alibaba (Qwen) ranks 230th of 354 ranked models on the Noometry Index as of October 2026, with a score of 34.7. Its strongest category is reasoning, where it ranks 151st. API pricing starts at $1.60 per million input tokens and $6.40 per million output tokens, with a 33K-token context window.
How much does Qwen Max cost?
Qwen Max costs $1.60 per million input tokens and $6.40 per million output tokens on Alibaba (Qwen)'s own API.
What is Qwen Max's context window?
Qwen Max accepts up to 33K tokens of input and can write up to 8K tokens in one response.
Is Qwen Max open source?
No. Qwen Max is proprietary and available only through Alibaba (Qwen)'s API and partner platforms.
What are Qwen Max's strengths and weaknesses?
Relative to other ranked models, Qwen Max places best in reasoning, long context, writing & preference and lowest in coding, math, knowledge.
What is Qwen Max best at?
Its best category is reasoning, where it ranks 151st on Noometry.