Model comparison

Gemma 1.1 2b IT vs Mistral Nemo

Gemma 1.1 2b IT is the stronger model overall, scoring 29.3 to 26.4 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemma 1.1 2b IT Google

29.3

Rank #313 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • The widest gap is in knowledge, where Gemma 1.1 2b IT leads 26.5 to 12.3.

Side by side

Gemma 1.1 2b IT and Mistral Nemo specifications
Gemma 1.1 2b ITMistral Nemo
ProviderGoogleMistral AI
Noometry Index29.326.4
Released—2024-07-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.15
Results tracked1610

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 1.1 2b IT: 30.1 (#299), Mistral Nemo: —

Coding benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Coding1034—
HumanEval+17.7%—
MBPP+23.3%—

Agentic & Tool Use Not comparable

Gemma 1.1 2b IT: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Mistral Nemo leads

Gemma 1.1 2b IT: 19.1 (#270), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Hard Prompts1005—
DTBench—48.6%
Epoch Capabilities Index—118.68
PIQA—83.5%

Math Gemma 1.1 2b IT leads

Gemma 1.1 2b IT: 30.8 (#232), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Math1047—
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge Gemma 1.1 2b IT leads

Gemma 1.1 2b IT: 26.5 (#258), Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
GPQA Diamond—29.9%
LMArena Expert970—
BoolQ—82.5%

Multilingual Not comparable

Gemma 1.1 2b IT: 24.6 (#289), Mistral Nemo: —

Multilingual benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Non-English988—
LMArena Chinese1012—
LMArena German944—
LMArena Korean899—
LMArena Russian990—

Instruction Following Not comparable

Gemma 1.1 2b IT: 49.9 (#299), Mistral Nemo: —

Instruction Following benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Instruction Following992—

Long Context Not comparable

Gemma 1.1 2b IT: 30.6 (#286), Mistral Nemo: —

Long Context benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Longer Query1003—

Writing & Preference Mistral Nemo leads

Gemma 1.1 2b IT: 25.1 (#306), Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkGemma 1.1 2b ITMistral Nemo
LMArena Text1022—
LMArena Creative Writing998—
EQ-Bench Creative Writing—881
LMArena Multi-Turn959—

Frequently asked questions

Is Gemma 1.1 2b IT better than Mistral Nemo?

Gemma 1.1 2b IT is the stronger model overall, scoring 29.3 to 26.4 on the Noometry Index.

How many benchmarks do Gemma 1.1 2b IT and Mistral Nemo share?

0 benchmarks have published results for both models. Gemma 1.1 2b IT has 16 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper