Alibaba (Qwen), proprietary

# Qwen3.8 Max

> Qwen3.8 Max by Alibaba (Qwen), released August 2026. Ranked #22 of 354 with a Noometry Index of 56.8. API: $2 in / $6 out per M tokens. 1M context. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/qwen3-8-max
- Last updated: 2026-10-10
- Title: Qwen3.8 Max Benchmarks, Price & Rank (October 2026)

Qwen3.8 Max by Alibaba (Qwen) ranks 22nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 56.8. Its strongest category is agentic & tool use, where it ranks 14th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 1M-token context window.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #22 of 354
- **Index score:** 56.8
- **Evidence:** Confirmed 39 results
- **Provider:** [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba)
- **Released:** August 2, 2026
- **Weights:** Proprietary
- **Reasoning:** Yes
- **Context window:** 1M
- **Max output:** 131K
- **Input price:** $2 / M
- **Output price:** $6 / M
- **Blended price:** $3 / M
- **Output speed:** Not measured
- **Value:** #154 of 219
- **Knowledge cutoff:** Unknown
- **Input:** text, image, video, pdf

## Category scores

Each category score combines every public result we have in that category.

Qwen3.8 Max category scores

1.  Coding 53.5
2.  Agentic & Tool Use 45.4
3.  Reasoning 54.4
4.  Math 73.2
5.  Knowledge 61.7
6.  Multimodal 37.2
7.  Multilingual 56.7
8.  Instruction Following 77.6
9.  Long Context 45.6
10.  Writing & Preference 67.1
11.  020406080

Qwen3.8 Max category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 53.5 | #29 | 5 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 45.4 | #14 | 3 |
| [Reasoning](https://noometry.com/best/reasoning) | 54.4 | #26 | 7 |
| [Math](https://noometry.com/best/math) | 73.2 | #20 | 5 |
| [Knowledge](https://noometry.com/best/knowledge) | 61.7 | #27 | 3 |
| [Multimodal](https://noometry.com/best/multimodal) | 37.2 | #75 | 2 |
| [Multilingual](https://noometry.com/best/multilingual) | 56.7 | #18 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 77.6 | #17 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 45.6 | #31 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 67.1 | #30 | 3 |

## Strengths and weaknesses

Categories where Qwen3.8 Max places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Qwen3.8 Max: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Instruction Following](https://noometry.com/best/instruction-following) | 77.6 | +6.3 | #17 of 305, top 6% |
| [Multilingual](https://noometry.com/best/multilingual) | 56.7 | +9.3 | #18 of 297, top 7% |
| [Math](https://noometry.com/best/math) | 73.2 | +36.6 | #20 of 327, top 7% |

### Weakest categories

Qwen3.8 Max: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Multimodal](https://noometry.com/best/multimodal) | 37.2 | −1.4 | #75 of 128, top 59% |
| [Long Context](https://noometry.com/best/long-context) | 45.6 | +4.7 | #31 of 296, top 11% |
| [Writing & Preference](https://noometry.com/best/writing) | 67.1 | +13.3 | #30 of 312, top 10% |

## Closest competitors

The models ranked just above and below Qwen3.8 Max. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen3.8 Max
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [GPT-5.4 Pro](https://noometry.com/models/gpt-5-4-pro) | #18 | 58.9 | $67.50 | — | [Compare](https://noometry.com/compare/gpt-5-4-pro-vs-qwen3-8-max) |
| [Claude Opus 4.7](https://noometry.com/models/claude-opus-4-7) | #19 | 58.3 | $10 | 33 | [Compare](https://noometry.com/compare/claude-opus-4-7-vs-qwen3-8-max) |
| [Claude Opus 4.6](https://noometry.com/models/claude-opus-4-6) | #20 | 58.2 | $10 | 19 | [Compare](https://noometry.com/compare/claude-opus-4-6-vs-qwen3-8-max) |
| [Grok 4.6](https://noometry.com/models/grok-4-6) | #21 | 56.9 | $3 | — | [Compare](https://noometry.com/compare/grok-4-6-vs-qwen3-8-max) |
| [Gemini 3.1 Pro Preview](https://noometry.com/models/gemini-3-1-pro-preview) | #23 | 56.7 | $4.50 | — | [Compare](https://noometry.com/compare/gemini-3-1-pro-preview-vs-qwen3-8-max) |
| [Gemini 4 Argon](https://noometry.com/models/gemini-4-argon) | #24 | 56.5 | — | — | [Compare](https://noometry.com/compare/gemini-4-argon-vs-qwen3-8-max) |
| [Grok 4.5](https://noometry.com/models/grok-4-5) | #25 | 55.0 | $3 | 4 | [Compare](https://noometry.com/compare/grok-4-5-vs-qwen3-8-max) |
| [GLM-5.3](https://noometry.com/models/glm-5-3) | #26 | 54.8 | $2.15 | — | [Compare](https://noometry.com/compare/glm-5-3-vs-qwen3-8-max) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Qwen3.8 Max Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [DeepSWE](https://noometry.com/benchmarks/deepswe) | 57.5% | #18 of 29, top 63% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev) | 1674 | #9 of 113, top 8% |  | [LMArena](https://lmarena.ai/leaderboard/webdev) | 2026-10-08 |
| [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev) | 1672 |  |  | [LMArena](https://lmarena.ai/leaderboard/webdev) | 2026-10-08 |
| [FrontierSWE](https://noometry.com/benchmarks/frontierswe) | 17.8% | #16 of 18, top 89% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [FrontierSWE](https://noometry.com/benchmarks/frontierswe) | 15.8% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [SciCode](https://noometry.com/benchmarks/scicode) | 52.1% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [SciCode](https://noometry.com/benchmarks/scicode) | 53.2% | #33 of 121, top 28% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1502 | #19 of 294, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Agentic & Tool Use

Qwen3.8 Max Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [APEX-Agents](https://noometry.com/benchmarks/apex-agents) | 63.3% | #11 of 49, top 23% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [τ²-bench Banking](https://noometry.com/benchmarks/tau2-banking) | 55.1% | Best of 26 | xhigh | [τ²-bench](https://taubench.com/) | 2026-08-04 |
| [GDP.pdf](https://noometry.com/benchmarks/gdp-pdf) | 23.2% | #15 of 36, top 42% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Reasoning

Qwen3.8 Max Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [NYT Connections (extended)](https://noometry.com/benchmarks/nyt-connections) | 88.3% | #24 of 91, top 27% |  | [Lech Mazur benchmarks](https://github.com/lechmazur/nyt-connections) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 17.7% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 20% | #24 of 134, top 18% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 29% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-04 |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 40% | #22 of 129, top 18% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-02 |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1496 | #11 of 297, top 4% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [Mystery Game Puzzles](https://noometry.com/benchmarks/mystery-game-puzzles) | 38% | #14 of 74, top 19% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-05 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 92% | #29 of 151, top 20% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMCA](https://noometry.com/benchmarks/lmca) | 46.2% | #30 of 125, top 24% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 156.41 | #22 of 213, top 11% |  | [Epoch AI](https://epoch.ai/eci) | 2026-08-02 |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 155.05 |  |  | [Epoch AI](https://epoch.ai/eci) | 2026-09-01 |

### Math

Qwen3.8 Max Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [FrontierMath (Tiers 1-3)](https://noometry.com/benchmarks/frontiermath) | 65.6% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-02 |
| [FrontierMath (Tiers 1-3)](https://noometry.com/benchmarks/frontiermath) | 74.7% | #19 of 81, top 24% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-04 |
| [FrontierMath Tier 4](https://noometry.com/benchmarks/frontiermath-tier-4) | 34.1% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-02 |
| [FrontierMath Tier 4](https://noometry.com/benchmarks/frontiermath-tier-4) | 46.3% | #21 of 63, top 34% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-04 |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 99.4% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-04 |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 100% | #11 of 173, top 7% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-02 |
| [ProofBench](https://noometry.com/benchmarks/proofbench) | 58% | #22 of 77, top 29% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1499 | #13 of 285, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Qwen3.8 Max Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 92.3% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-02 |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 92.7% | #21 of 186, top 12% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-04 |
| [SimpleQA Verified](https://noometry.com/benchmarks/simpleqa-verified) | 45.8% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [SimpleQA Verified](https://noometry.com/benchmarks/simpleqa-verified) | 47.3% | #31 of 77, top 41% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-02 |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1507 | #18 of 273, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multimodal

Qwen3.8 Max Multimodal benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Vision](https://noometry.com/benchmarks/arena-vision) | 1314 | #9 of 122, top 8% |  | [LMArena](https://lmarena.ai/leaderboard/vision) | 2026-10-09 |
| [Furniture Assembly](https://noometry.com/benchmarks/furniture-assembly) | 20% | #31 of 31, top 100% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-10 |

### Multilingual

Qwen3.8 Max Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1472 | #18 of 297, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1538 | #9 of 285, top 4% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1503 | #11 of 223, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1483 | #19 of 231, top 9% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1467 | #19 of 211, top 10% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1461 | #9 of 213, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1481 | #19 of 283, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1492 | #10 of 226, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Qwen3.8 Max Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1479 | #14 of 298, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Qwen3.8 Max Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1489 | #13 of 291, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Qwen3.8 Max Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1483 | #12 of 297, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1479 | #12 of 295, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1489 | #11 of 295, top 4% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## API pricing by provider

Qwen3.8 Max API prices
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
| --- | --- | --- | --- | --- |
| [alibaba](https://www.alibabacloud.com/help/en/model-studio/models) | $2 | $6 | $0.25 | 2026-10-10 |
| [deepinfra](https://deepinfra.com/models) | $1.65 | $4.95 | $0.21 | 2026-10-10 |
| [fireworks](https://fireworks.ai/docs/) | $2 | $6 | $0.25 | 2026-10-10 |
| [openrouter](https://openrouter.ai/qwen/qwen3.8-max-0902) | $2 | $6 | $0.25 | 2026-10-10 |

[All Alibaba (Qwen) API prices →](https://noometry.com/llm-pricing/alibaba) [Estimate your cost →](https://noometry.com/tools/cost-calculator)

## Compare Qwen3.8 Max

-   [Qwen3.8 Max vs Qwen3.7 Max](https://noometry.com/compare/qwen3-7-max-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Gemini 3.1 Pro Preview](https://noometry.com/compare/gemini-3-1-pro-preview-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Claude Opus 4.6](https://noometry.com/compare/claude-opus-4-6-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Gemini 4 Argon](https://noometry.com/compare/gemini-4-argon-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Claude Opus 4.7](https://noometry.com/compare/claude-opus-4-7-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Grok 4.5](https://noometry.com/compare/grok-4-5-vs-qwen3-8-max)
-   [Qwen3.8 Max vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-qwen3-8-max)
-   [Qwen3.8 Max vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-qwen3-8-max)
-   [Qwen3.8 Max vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-qwen3-8-max)
-   [Qwen3.8 Max vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-qwen3-8-max)

## Other Alibaba (Qwen) models

-   [Qwen3.7 Max](https://noometry.com/models/qwen3-7-max)51.5
-   [Qwen3.6 Max Preview](https://noometry.com/models/qwen3-6-max-preview)51.5
-   [Qwen3.6 Plus](https://noometry.com/models/qwen3-6-plus)47.5
-   [Qwen3.5 397B-A17B](https://noometry.com/models/qwen3-5-397b-a17b)46.0
-   [Qwen3.8 27B](https://noometry.com/models/qwen3-8-27b)46.0
-   [Qwen3.5 Max Preview](https://noometry.com/models/qwen3-5-max-preview)45.3
-   [Qwen3.7 Plus](https://noometry.com/models/qwen3-7-plus)45.3
-   [Qwen3 Max](https://noometry.com/models/qwen3-max)43.7

## Frequently asked questions

### How good is Qwen3.8 Max?

Qwen3.8 Max by Alibaba (Qwen) ranks 22nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 56.8. Its strongest category is agentic & tool use, where it ranks 14th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 1M-token context window.

### How much does Qwen3.8 Max cost?

Qwen3.8 Max costs $2 per million input tokens and $6 per million output tokens on Alibaba (Qwen)'s own API, with cached input at $0.25.

### What is Qwen3.8 Max's context window?

Qwen3.8 Max accepts up to 1M tokens of input and can write up to 131K tokens in one response.

### Is Qwen3.8 Max open source?

No. Qwen3.8 Max is proprietary and available only through Alibaba (Qwen)'s API and partner platforms.

### What are Qwen3.8 Max's strengths and weaknesses?

Relative to other ranked models, Qwen3.8 Max places best in instruction following, multilingual, math and lowest in multimodal, long context, writing & preference.

### What is Qwen3.8 Max best at?

Its best category is agentic & tool use, where it ranks 14th on Noometry.

### Cite this page

Noometry. (2026). Qwen3.8 Max benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/qwen3-8-max

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/qwen3-8-max.md).
