Model comparison

Mistral Small 3.2 vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 31.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • The widest gap is in writing & preference, where Qwen3.5 Max Preview leads 66.0 to 45.0.
  • Mistral Small 3.2 has downloadable open weights; the other is API-only.

Side by side

Mistral Small 3.2 and Qwen3.5 Max Preview specifications
Mistral Small 3.2Qwen3.5 Max Preview
ProviderMistral AIAlibaba (Qwen)
Noometry Index31.245.3
Released2025-06-20—
WeightsOpenProprietary
Context window256K—
Max output16K—
Input $ / M tokens$0.0938—
Output $ / M tokens$0.25—
Results tracked617

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Small 3.2: —, Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
LMArena Coding—1487

Reasoning Qwen3.5 Max Preview leads

Mistral Small 3.2: 18.1 (#287), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
Kagi LLM Benchmark40.4%—
Chess Puzzles1%—
LMArena Hard Prompts—1483
Epoch Capabilities Index131.74—

Math Qwen3.5 Max Preview leads

Mistral Small 3.2: 26.3 (#260), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
OTIS Mock AIME 2024-202530.3%—
LMArena Math—1474

Knowledge Qwen3.5 Max Preview leads

Mistral Small 3.2: 26.7 (#256), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
GPQA Diamond49.1%—
LMArena Expert—1489

Multilingual Not comparable

Mistral Small 3.2: —, Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
LMArena Non-English—1465
LMArena Chinese—1534
LMArena French—1484
LMArena German—1487
LMArena Japanese—1495
LMArena Korean—1438
LMArena Russian—1471
LMArena Spanish—1470

Instruction Following Not comparable

Mistral Small 3.2: —, Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
LMArena Instruction Following—1467

Long Context Not comparable

Mistral Small 3.2: —, Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
LMArena Longer Query—1476

Writing & Preference Qwen3.5 Max Preview leads

Mistral Small 3.2: 45.0 (#224), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMistral Small 3.2Qwen3.5 Max Preview
LMArena Text—1470
LMArena Creative Writing—1464
EQ-Bench Creative Writing1255—
LMArena Multi-Turn—1478

Frequently asked questions

Is Mistral Small 3.2 better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 31.2 on the Noometry Index.

How many benchmarks do Mistral Small 3.2 and Qwen3.5 Max Preview share?

0 benchmarks have published results for both models. Mistral Small 3.2 has 6 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper