Model comparison

Gemma 7B vs Mistral Medium 3.1

Mistral Medium 3.1 is the stronger model overall, scoring 31.9 to 30.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in writing & preference, where Mistral Medium 3.1 leads 55.5 to 27.1.
  • Gemma 7B has downloadable open weights; the other is API-only.

Side by side

Gemma 7B and Mistral Medium 3.1 specifications
Gemma 7BMistral Medium 3.1
ProviderGoogleMistral AI
Noometry Index30.031.9
Released2024-02-21—
WeightsOpenProprietary
Context window—131K
Max output—105K
Input $ / M tokens—$0.40
Output $ / M tokens—$2
Results tracked273

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 7B: 30.5 (#294), Mistral Medium 3.1: —

Coding benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Coding1048—
HumanEval+28.7%—
MBPP+43.4%—

Reasoning Gemma 7B leads

Gemma 7B: 19.9 (#249), Mistral Medium 3.1: 10.6 (#341)

Reasoning benchmarks
BenchmarkGemma 7BMistral Medium 3.1
NYT Connections (extended)—6.5%
Thematic Generalization—20.3%
LMArena Hard Prompts1042—
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Not comparable

Gemma 7B: 31.2 (#228), Mistral Medium 3.1: —

Math benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Math1066—
GSM8K46.4%—

Knowledge Not comparable

Gemma 7B: 27.3 (#252), Mistral Medium 3.1: —

Knowledge benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Expert1001—
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multilingual Not comparable

Gemma 7B: 25.1 (#287), Mistral Medium 3.1: —

Multilingual benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Non-English999—
LMArena Chinese1035—
LMArena French1025—
LMArena Russian993—

Instruction Following Not comparable

Gemma 7B: 51.5 (#295), Mistral Medium 3.1: —

Instruction Following benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Instruction Following1017—

Long Context Not comparable

Gemma 7B: 31.1 (#282), Mistral Medium 3.1: —

Long Context benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Longer Query1022—

Writing & Preference Mistral Medium 3.1 leads

Gemma 7B: 27.1 (#302), Mistral Medium 3.1: 55.5 (#145)

Writing & Preference benchmarks
BenchmarkGemma 7BMistral Medium 3.1
LMArena Text1056—
LMArena Creative Writing1024—
EQ-Bench Creative Writing—1476
LMArena Multi-Turn963—

Frequently asked questions

Is Gemma 7B better than Mistral Medium 3.1?

Mistral Medium 3.1 is the stronger model overall, scoring 31.9 to 30.0 on the Noometry Index.

How many benchmarks do Gemma 7B and Mistral Medium 3.1 share?

0 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Mistral Medium 3.1 has 3.

Related comparisons

Go deeper