Model comparison

Gemma 1.1 7b IT vs Mistral Medium 3.5

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 31.3 on the Noometry Index.

Last verified . 16 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 1 category and Mistral Medium 3.5 in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium 3.5 leads 58.5 to 30.4.

Side by side

Gemma 1.1 7b IT and Mistral Medium 3.5 specifications
Gemma 1.1 7b ITMistral Medium 3.5
ProviderGoogleMistral AI
Noometry Index31.340.2
Released——
WeightsOpenOpen
Context window—262K
Max output—210K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked1922

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 31.5 (#284), Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Coding10841461
LMArena WebDev—1264
HumanEval+35.4%—
MBPP+45%—

Reasoning Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 20.5 (#238), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Hard Prompts10711436
Kagi LLM Benchmark—41.4%
NYT Connections (extended)—12.9%
Epoch Capabilities Index—141.35

Math Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 32.0 (#220), Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Math11071431

Knowledge Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 28.3 (#247), Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Expert10391432

Multimodal Not comparable

Gemma 1.1 7b IT: —, Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Vision—1223

Multilingual Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 28.1 (#273), Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Non-English10521404
LMArena Chinese10611442
LMArena French10651448
LMArena German10541451
LMArena Korean9881385
LMArena Russian10461395
LMArena Spanish10491409
LMArena Japanese971—

Instruction Following Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 54.0 (#283), Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Instruction Following10571415

Long Context Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 32.1 (#272), Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Longer Query10561415

Writing & Preference Mistral Medium 3.5 leads

Gemma 1.1 7b IT: 30.4 (#288), Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITMistral Medium 3.5
LMArena Text10941421
LMArena Creative Writing10601374
LMArena Multi-Turn10401423
EQ-Bench 4—993

Frequently asked questions

Is Gemma 1.1 7b IT better than Mistral Medium 3.5?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 31.3 on the Noometry Index.

Is Gemma 1.1 7b IT or Mistral Medium 3.5 better for coding?

Mistral Medium 3.5 scores higher on coding benchmarks: 36.0 versus 31.5 in the Noometry coding category.

How many benchmarks do Gemma 1.1 7b IT and Mistral Medium 3.5 share?

16 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper