Model comparison

Mistral Nemo vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 26.4 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • The widest gap is in writing & preference, where Qwen3.5 Max Preview leads 66.0 to 28.5.
  • Mistral Nemo has downloadable open weights; the other is API-only.

Side by side

Mistral Nemo and Qwen3.5 Max Preview specifications
Mistral NemoQwen3.5 Max Preview
ProviderMistral AIAlibaba (Qwen)
Noometry Index26.445.3
Released2024-07-01—
WeightsOpenProprietary
Context window128K—
Max output128K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.15—
Results tracked1017

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Nemo: —, Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Coding—1487

Agentic & Tool Use Not comparable

Mistral Nemo: 23.5 (#125), Qwen3.5 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
Berkeley Function Calling Leaderboard27.6%—
BALROG17.6%—

Reasoning Qwen3.5 Max Preview leads

Mistral Nemo: 20.7 (#232), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Hard Prompts—1483
DTBench48.6%—
Epoch Capabilities Index118.68—
PIQA83.5%—

Math Qwen3.5 Max Preview leads

Mistral Nemo: 25.5 (#268), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Math—1474
MATH Level 510.8%—
GSM8K84.2%—

Knowledge Qwen3.5 Max Preview leads

Mistral Nemo: 12.3 (#298), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
GPQA Diamond29.9%—
LMArena Expert—1489
BoolQ82.5%—

Multilingual Not comparable

Mistral Nemo: —, Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Non-English—1465
LMArena Chinese—1534
LMArena French—1484
LMArena German—1487
LMArena Japanese—1495
LMArena Korean—1438
LMArena Russian—1471
LMArena Spanish—1470

Instruction Following Not comparable

Mistral Nemo: —, Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Instruction Following—1467

Long Context Not comparable

Mistral Nemo: —, Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Longer Query—1476

Writing & Preference Qwen3.5 Max Preview leads

Mistral Nemo: 28.5 (#296), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMistral NemoQwen3.5 Max Preview
LMArena Text—1470
LMArena Creative Writing—1464
EQ-Bench Creative Writing881—
LMArena Multi-Turn—1478

Frequently asked questions

Is Mistral Nemo better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 26.4 on the Noometry Index.

How many benchmarks do Mistral Nemo and Qwen3.5 Max Preview share?

0 benchmarks have published results for both models. Mistral Nemo has 10 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper