Alibaba (Qwen), proprietary
Qwen3.6 Plus
Qwen3.6 Plus by Alibaba (Qwen) ranks 62nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 47.5. Its strongest category is knowledge, where it ranks 45th. API pricing starts at $0.50 per million input tokens and $3 per million output tokens, with a 1M-token context window.
Last verified
Specifications
- Noometry rank
- #62 of 354
- Index score
- 47.5
- Evidence
- Confirmed 37 results
- Provider
Alibaba (Qwen)
- Released
- March 31, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 1M
- Max output
- 66K
- Input price
- $0.50 / M
- Output price
- $3 / M
- Blended price
- $1.13 / M
- Output speed
- Not measured
- Value
- #108 of 219
- Knowledge cutoff
- April 2025
- Input
- text, image, video
Category scores
Each category score combines every public result we have in that category.
- Coding 40.8
- Reasoning 29.3
- Math 51.8
- Knowledge 56.1
- Multilingual 53.3
- Instruction Following 75.0
- Long Context 45.2
- Writing & Preference 62.2
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 40.8 | #130 | 4 |
| Reasoning | 29.3 | #93 | 8 |
| Math | 51.8 | #54 | 3 |
| Knowledge | 56.1 | #45 | 3 |
| Multilingual | 53.3 | #70 | 1 |
| Instruction Following | 75.0 | #74 | 1 |
| Long Context | 45.2 | #49 | 2 |
| Writing & Preference | 62.2 | #82 | 3 |
Strengths and weaknesses
Categories where Qwen3.6 Plus places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 56.1 | +18.8 | #45 of 314, top 15% |
| Math | 51.8 | +15.2 | #54 of 327, top 17% |
| Long Context | 45.2 | +4.2 | #49 of 296, top 17% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 40.8 | +2.1 | #130 of 340, top 39% |
| Reasoning | 29.3 | +5.7 | #93 of 350, top 27% |
| Writing & Preference | 62.2 | +8.4 | #82 of 312, top 27% |
Closest competitors
The models ranked just above and below Qwen3.6 Plus. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Step 5 Preview | #58 | 47.9 | $1.43 | — | Compare |
| GLM-5.1 | #59 | 47.8 | $2.15 | — | Compare |
| Kimi K2.6 | #60 | 47.7 | $1.71 | — | Compare |
| o3 | #61 | 47.5 | $3.50 | 3 | Compare |
| Inkling-Small | #63 | 46.5 | $0.64 | — | Compare |
| GPT-5 Pro | #64 | 46.4 | $41.25 | 5 | Compare |
| Grok 4.20 Multi-Agent | #65 | 46.2 | $1.56 | — | Compare |
| GLM-5 | #66 | 46.1 | $1.55 | 23 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SWE-bench Verified | 57.9% | #29 of 32, top 91% | Epoch AI | 2026-05-14 | |
| LMArena WebDev | 1461 | #55 of 113, top 49% | LMArena | 2026-10-08 | |
| SciCode | 40.7% | #74 of 121, top 62% | Epoch AI | ||
| LMArena Coding | 1467 | #61 of 294, top 21% | LMArena | 2026-10-08 | |
| ALE-Bench | 670.15 | #72 of 105, top 69% | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Vending-Bench 2 | 5,115 | #28 of 60, top 47% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| NYT Connections (extended) | 60.3% | #57 of 91, top 63% | Lech Mazur benchmarks | ||
| CritPt | 2.9% | #69 of 134, top 52% | Epoch AI | ||
| Chess Puzzles | 17% | #67 of 129, top 52% | Epoch AI | 2026-08-07 | |
| Thematic Generalization | 59.5% | #13 of 23, top 57% | Lech Mazur benchmarks | ||
| LMArena Hard Prompts | 1449 | #63 of 297, top 22% | LMArena | 2026-10-08 | |
| Mystery Game Puzzles | 12% | #58 of 74, top 79% | none | Epoch AI | 2026-08-27 |
| DTBench | 81.9% | #67 of 151, top 45% | Epoch AI | ||
| LMCA | 33.1% | #68 of 125, top 55% | Epoch AI | ||
| Epoch Capabilities Index | 147.65 | #63 of 213, top 30% | Epoch AI | 2026-03-31 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 38.2% | #53 of 81, top 66% | Epoch AI | 2026-08-30 | |
| FrontierMath (Tiers 1-3) | 32.3% | none | Epoch AI | 2026-08-29 | |
| OTIS Mock AIME 2024-2025 | 93.3% | #43 of 173, top 25% | Epoch AI | 2026-08-07 | |
| LMArena Math | 1450 | #61 of 285, top 22% | LMArena | 2026-10-08 | |
| FrontierMath (Feb 2025 set) | 26.2% | #23 of 68, top 34% | Epoch AI | 2026-05-12 | |
| FrontierMath Tier 4 (v1) | 8.3% | #21 of 55, top 39% | Epoch AI | 2026-05-12 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 88.4% | #47 of 186, top 26% | Epoch AI | 2026-08-07 | |
| SimpleQA Verified | 44.1% | #37 of 77, top 49% | Epoch AI | 2026-08-27 | |
| LMArena Expert | 1454 | #66 of 273, top 25% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1424 | #70 of 297, top 24% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1477 | #69 of 285, top 25% | LMArena | 2026-10-08 | |
| LMArena French | 1455 | #64 of 223, top 29% | LMArena | 2026-10-08 | |
| LMArena German | 1452 | #42 of 231, top 19% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1389 | #70 of 211, top 34% | LMArena | 2026-10-08 | |
| LMArena Korean | 1379 | #80 of 213, top 38% | LMArena | 2026-10-08 | |
| LMArena Russian | 1434 | #62 of 283, top 22% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1432 | #81 of 226, top 36% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1425 | #67 of 298, top 23% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| CL-bench | 20.3% | #7 of 19, top 37% | Epoch AI | ||
| LMArena Longer Query | 1439 | #65 of 291, top 23% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1437 | #69 of 297, top 24% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1404 | #70 of 295, top 24% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1438 | #74 of 295, top 26% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| alibaba | $0.50 | $3 | $0.05 | 2026-10-10 |
| openrouter | $0.33 | $1.95 | — | 2026-10-10 |
| together | $0.50 | $3 | — | 2026-10-10 |
Compare Qwen3.6 Plus
- Qwen3.6 Plus vs Qwen3.5 Plus
- Qwen3.6 Plus vs o3
- Qwen3.6 Plus vs Inkling-Small
- Qwen3.6 Plus vs Kimi K2.6
- Qwen3.6 Plus vs GPT-5 Pro
- Qwen3.6 Plus vs GLM-5.1
- Qwen3.6 Plus vs Grok 4.20 Multi-Agent
- Qwen3.6 Plus vs GPT-6 Astra
- Qwen3.6 Plus vs Claude Fable 5.1
- Qwen3.6 Plus vs Gemini 3.8 Flash
- Qwen3.6 Plus vs Kimi K3
- Qwen3.6 Plus vs Grok 4.6
- Qwen3.6 Plus vs GLM-5.3
- Qwen3.6 Plus vs Muse Spark 1.3
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.7 Max51.5
- Qwen3.6 Max Preview51.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
- Qwen3 Max43.7
Frequently asked questions
How good is Qwen3.6 Plus?
Qwen3.6 Plus by Alibaba (Qwen) ranks 62nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 47.5. Its strongest category is knowledge, where it ranks 45th. API pricing starts at $0.50 per million input tokens and $3 per million output tokens, with a 1M-token context window.
How much does Qwen3.6 Plus cost?
Qwen3.6 Plus costs $0.50 per million input tokens and $3 per million output tokens on Alibaba (Qwen)'s own API, with cached input at $0.05.
What is Qwen3.6 Plus's context window?
Qwen3.6 Plus accepts up to 1M tokens of input and can write up to 66K tokens in one response.
Is Qwen3.6 Plus open source?
No. Qwen3.6 Plus is proprietary and available only through Alibaba (Qwen)'s API and partner platforms.
What are Qwen3.6 Plus's strengths and weaknesses?
Relative to other ranked models, Qwen3.6 Plus places best in knowledge, math, long context and lowest in coding, reasoning, writing & preference.
What is Qwen3.6 Plus best at?
Its best category is knowledge, where it ranks 45th on Noometry.