Model comparison

Mistral 7B vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 23.0 on the Noometry Index.

Last verified . 16 shared benchmarks.

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Mistral 7B scores higher in 0 categories and Qwen3.5 Max Preview in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen3.5 Max Preview leads 66.0 to 30.7.
  • Mistral 7B has downloadable open weights; the other is API-only.

Side by side

Mistral 7B and Qwen3.5 Max Preview specifications
Mistral 7BQwen3.5 Max Preview
ProviderMistral AIAlibaba (Qwen)
Noometry Index23.045.3
Released2023-09-27—
WeightsOpenProprietary
Context window8K—
Max output8K—
Input $ / M tokens$0.25—
Output $ / M tokens$0.25—
Results tracked3717

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Mistral 7B: 26.4 (#326), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Coding10821487
BigCodeBench Instruct19.5%—
BigCodeBench Complete27.3%—
HumanEval+36%—
MBPP+42.1%—

Reasoning Qwen3.5 Max Preview leads

Mistral 7B: 13.1 (#336), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Hard Prompts10671483
Chess Puzzles0%—
DTBench42.5%—
Adversarial NLI47.1%—
BIG-Bench Hard56.1%—
Epoch Capabilities Index112.21—
HellaSwag81%—
PIQA83%—
WinoGrande75.3%—

Math Qwen3.5 Max Preview leads

Mistral 7B: 8.1 (#325), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Math10851474
OTIS Mock AIME 2024-20250.3%—
MATH Level 53.7%—
GSM8K54.4%—

Knowledge Qwen3.5 Max Preview leads

Mistral 7B: 7.4 (#311), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Expert10361489
GPQA Diamond15.2%—
ARC (AI2) Challenge78.6%—
BoolQ87.4%—
MMLU62.5%—
OpenBookQA79.8%—
TriviaQA75.2%—

Multilingual Qwen3.5 Max Preview leads

Mistral 7B: 25.8 (#283), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Non-English10121465
LMArena Chinese10091534
LMArena French10371484
LMArena German9871487
LMArena Japanese8781495
LMArena Russian10181471
LMArena Spanish10261470
LMArena Korean—1438

Instruction Following Qwen3.5 Max Preview leads

Mistral 7B: 54.2 (#280), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Instruction Following10601467

Long Context Qwen3.5 Max Preview leads

Mistral 7B: 32.2 (#271), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Longer Query10601476

Writing & Preference Qwen3.5 Max Preview leads

Mistral 7B: 30.7 (#286), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMistral 7BQwen3.5 Max Preview
LMArena Text10901470
LMArena Creative Writing10681464
LMArena Multi-Turn10621478

Frequently asked questions

Is Mistral 7B better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 23.0 on the Noometry Index.

Is Mistral 7B or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 26.4 in the Noometry coding category.

How many benchmarks do Mistral 7B and Qwen3.5 Max Preview share?

16 benchmarks have published results for both models. Mistral 7B has 37 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper