Model comparison

Qwen3-30B-A3B vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 38.9 on the Noometry Index.

Last verified . 17 shared benchmarks.

Qwen3-30B-A3B Alibaba (Qwen)

38.9

Rank #179 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Qwen3-30B-A3B scores higher in 1 category and Qwen3.5 Max Preview in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Qwen3.5 Max Preview leads 45.2 to 31.0.
  • Qwen3-30B-A3B has downloadable open weights; the other is API-only.

Side by side

Qwen3-30B-A3B and Qwen3.5 Max Preview specifications
Qwen3-30B-A3BQwen3.5 Max Preview
ProviderAlibaba (Qwen)Alibaba (Qwen)
Noometry Index38.945.3
Released2025-04-28—
WeightsOpenProprietary
Context window41K—
Max output16K—
Input $ / M tokens$0.12—
Output $ / M tokens$0.50—
Results tracked3217

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 37.5 (#194), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Coding14161487
SciCode33.3%—
WeirdML29.8%—

Agentic & Tool Use Not comparable

Qwen3-30B-A3B: 29.8 (#82), Qwen3.5 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
Berkeley Function Calling Leaderboard41.4%—

Reasoning Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 22.2 (#204), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Hard Prompts13981483
Kagi LLM Benchmark54.9%—
CritPt0.3%—
Chess Puzzles8%—
DTBench69.3%—
LMCA22.4%—
Epoch Capabilities Index139.63—

Math Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 37.4 (#157), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Math13941474
MathArena Final-Answer Competitions47.8%—
OTIS Mock AIME 2024-202570.3%—

Knowledge Too close to call

Qwen3-30B-A3B: 41.8 (#105), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Expert13961489
GPQA Diamond70.1%—
Confabulations12.3%—

Multilingual Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 49.5 (#132), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Non-English13721465
LMArena Chinese14331534
LMArena French14181484
LMArena German13801487
LMArena Japanese13371495
LMArena Korean13311438
LMArena Russian13701471
LMArena Spanish14041470

Instruction Following Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 72.0 (#142), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Instruction Following13631467

Long Context Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 31.0 (#283), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Longer Query13791476
Fiction.LiveBench40.6%—

Writing & Preference Qwen3.5 Max Preview leads

Qwen3-30B-A3B: 55.6 (#143), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkQwen3-30B-A3BQwen3.5 Max Preview
LMArena Text13841470
LMArena Creative Writing13171464
LMArena Multi-Turn13781478
Short-Story Creative Writing75.3%—

Frequently asked questions

Is Qwen3-30B-A3B better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 38.9 on the Noometry Index.

Is Qwen3-30B-A3B or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 37.5 in the Noometry coding category.

How many benchmarks do Qwen3-30B-A3B and Qwen3.5 Max Preview share?

17 benchmarks have published results for both models. Qwen3-30B-A3B has 32 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper