Model comparison

Gemma 7B vs Mistral Nemo

Gemma 7B is the stronger model overall, scoring 30.0 to 26.4 on the Noometry Index.

Last verified . 4 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Gemma 7B scores higher in 2 categories and Mistral Nemo in 2 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemma 7B leads 27.3 to 12.3.

Side by side

Gemma 7B and Mistral Nemo specifications
Gemma 7BMistral Nemo
ProviderGoogleMistral AI
Noometry Index30.026.4
Released2024-02-212024-07-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.15
Results tracked2710

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 7B: 30.5 (#294), Mistral Nemo: —

Coding benchmarks
BenchmarkGemma 7BMistral Nemo
LMArena Coding1048—
HumanEval+28.7%—
MBPP+43.4%—

Agentic & Tool Use Not comparable

Gemma 7B: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkGemma 7BMistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Too close to call

Gemma 7B: 19.9 (#249), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkGemma 7BMistral Nemo
Epoch Capabilities Index111.99118.68
PIQA81.2%83.5%
LMArena Hard Prompts1042—
DTBench—48.6%
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
HellaSwag82.2%—
WinoGrande79%—

Math Gemma 7B leads

Gemma 7B: 31.2 (#228), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkGemma 7BMistral Nemo
GSM8K46.4%84.2%
LMArena Math1066—
MATH Level 5—10.8%

Knowledge Gemma 7B leads

Gemma 7B: 27.3 (#252), Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkGemma 7BMistral Nemo
BoolQ83.2%82.5%
GPQA Diamond—29.9%
LMArena Expert1001—
ARC (AI2) Challenge78.3%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multilingual Not comparable

Gemma 7B: 25.1 (#287), Mistral Nemo: —

Multilingual benchmarks
BenchmarkGemma 7BMistral Nemo
LMArena Non-English999—
LMArena Chinese1035—
LMArena French1025—
LMArena Russian993—

Instruction Following Not comparable

Gemma 7B: 51.5 (#295), Mistral Nemo: —

Instruction Following benchmarks
BenchmarkGemma 7BMistral Nemo
LMArena Instruction Following1017—

Long Context Not comparable

Gemma 7B: 31.1 (#282), Mistral Nemo: —

Long Context benchmarks
BenchmarkGemma 7BMistral Nemo
LMArena Longer Query1022—

Writing & Preference Mistral Nemo leads

Gemma 7B: 27.1 (#302), Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkGemma 7BMistral Nemo
LMArena Text1056—
LMArena Creative Writing1024—
EQ-Bench Creative Writing—881
LMArena Multi-Turn963—

Frequently asked questions

Is Gemma 7B better than Mistral Nemo?

Gemma 7B is the stronger model overall, scoring 30.0 to 26.4 on the Noometry Index.

How many benchmarks do Gemma 7B and Mistral Nemo share?

4 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper