Model comparison

Qwen3.5 Max Preview vs Qwen3.6 Plus

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 45.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Qwen3.6 Plus Alibaba (Qwen)

47.5

Rank #62 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Qwen3.5 Max Preview scores higher in 6 categories and Qwen3.6 Plus in 2 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3.6 Plus leads 56.1 to 41.8.

Side by side

Qwen3.5 Max Preview and Qwen3.6 Plus specifications
Qwen3.5 Max PreviewQwen3.6 Plus
ProviderAlibaba (Qwen)Alibaba (Qwen)
Noometry Index45.347.5
Released—2026-03-31
WeightsProprietaryProprietary
Context window—1M
Max output—66K
Input $ / M tokens—$0.50
Output $ / M tokens—$3
Results tracked1737

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 44.0 (#77), Qwen3.6 Plus: 40.8 (#130)

Coding benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Coding14871467
SWE-bench Verified—57.9%
LMArena WebDev—1461
SciCode—40.7%
ALE-Bench—670.15

Agentic & Tool Use Not comparable

Qwen3.5 Max Preview: —, Qwen3.6 Plus: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
Vending-Bench 2—5,115

Reasoning Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 30.8 (#84), Qwen3.6 Plus: 29.3 (#93)

Reasoning benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Hard Prompts14831449
NYT Connections (extended)—60.3%
CritPt—2.9%
Chess Puzzles—17%
Thematic Generalization—59.5%
Mystery Game Puzzles—12%
DTBench—81.9%
LMCA—33.1%
Epoch Capabilities Index—147.65

Math Qwen3.6 Plus leads

Qwen3.5 Max Preview: 40.1 (#94), Qwen3.6 Plus: 51.8 (#54)

Math benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Math14741450
FrontierMath (Tiers 1-3)—38.2%
OTIS Mock AIME 2024-2025—93.3%
FrontierMath (Feb 2025 set)—26.2%
FrontierMath Tier 4 (v1)—8.3%

Knowledge Qwen3.6 Plus leads

Qwen3.5 Max Preview: 41.8 (#107), Qwen3.6 Plus: 56.1 (#45)

Knowledge benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Expert14891454
GPQA Diamond—88.4%
SimpleQA Verified—44.1%

Multilingual Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 56.2 (#22), Qwen3.6 Plus: 53.3 (#70)

Multilingual benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Non-English14651424
LMArena Chinese15341477
LMArena French14841455
LMArena German14871452
LMArena Japanese14951389
LMArena Korean14381379
LMArena Russian14711434
LMArena Spanish14701432

Instruction Following Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 77.0 (#31), Qwen3.6 Plus: 75.0 (#74)

Instruction Following benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Instruction Following14671425

Long Context Too close to call

Qwen3.5 Max Preview: 45.2 (#45), Qwen3.6 Plus: 45.2 (#49)

Long Context benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Longer Query14761439
CL-bench—20.3%

Writing & Preference Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 66.0 (#41), Qwen3.6 Plus: 62.2 (#82)

Writing & Preference benchmarks
BenchmarkQwen3.5 Max PreviewQwen3.6 Plus
LMArena Text14701437
LMArena Creative Writing14641404
LMArena Multi-Turn14781438

Frequently asked questions

Is Qwen3.5 Max Preview better than Qwen3.6 Plus?

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 45.3 on the Noometry Index.

Is Qwen3.5 Max Preview or Qwen3.6 Plus better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 40.8 in the Noometry coding category.

How many benchmarks do Qwen3.5 Max Preview and Qwen3.6 Plus share?

17 benchmarks have published results for both models. Qwen3.5 Max Preview has 17 scored results on Noometry and Qwen3.6 Plus has 37.

Related comparisons

Go deeper