Model comparison

Mistral Large 4 vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 43.1 on the Noometry Index.

Last verified . 12 shared benchmarks.

Mistral Large 4 Mistral AI

43.1

Rank #99 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Mistral Large 4 scores higher in 2 categories and Qwen3.5 Max Preview in 6 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Qwen3.5 Max Preview leads 30.8 to 22.5.

Side by side

Mistral Large 4 and Qwen3.5 Max Preview specifications
Mistral Large 4Qwen3.5 Max Preview
ProviderMistral AIAlibaba (Qwen)
Noometry Index43.145.3
Released2026-10-06—
WeightsProprietaryProprietary
Context window1.05M—
Max output262K—
Input $ / M tokens$0.68—
Output $ / M tokens$2.09—
Results tracked1517

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Large 4 leads

Mistral Large 4: 48.6 (#57), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Coding14751487
LMArena WebDev1541—

Reasoning Qwen3.5 Max Preview leads

Mistral Large 4: 22.5 (#192), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Hard Prompts14441483
NYT Connections (extended)27.4%—

Math Too close to call

Mistral Large 4: 40.4 (#91), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Math14881474

Knowledge Qwen3.5 Max Preview leads

Mistral Large 4: 36.6 (#166), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Expert14471489
SimpleQA Verified20%—

Multilingual Qwen3.5 Max Preview leads

Mistral Large 4: 52.6 (#82), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Non-English14151465
LMArena Chinese14911534
LMArena Russian14141471
LMArena French—1484
LMArena German—1487
LMArena Japanese—1495
LMArena Korean—1438
LMArena Spanish—1470

Instruction Following Qwen3.5 Max Preview leads

Mistral Large 4: 75.0 (#76), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Instruction Following14241467

Long Context Qwen3.5 Max Preview leads

Mistral Large 4: 43.6 (#89), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Longer Query14291476

Writing & Preference Qwen3.5 Max Preview leads

Mistral Large 4: 60.4 (#97), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMistral Large 4Qwen3.5 Max Preview
LMArena Text14271470
LMArena Creative Writing13611464
LMArena Multi-Turn14241478

Frequently asked questions

Is Mistral Large 4 better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 43.1 on the Noometry Index.

Is Mistral Large 4 or Qwen3.5 Max Preview better for coding?

Mistral Large 4 scores higher on coding benchmarks: 48.6 versus 44.0 in the Noometry coding category.

How many benchmarks do Mistral Large 4 and Qwen3.5 Max Preview share?

12 benchmarks have published results for both models. Mistral Large 4 has 15 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper