Model comparison

Gemma 3n E4b IT vs Mistral 7B

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 23.0 on the Noometry Index.

Last verified . 16 shared benchmarks.

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemma 3n E4b IT scores higher in 8 categories and Mistral 7B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Gemma 3n E4b IT leads 35.1 to 8.1.

Side by side

Gemma 3n E4b IT and Mistral 7B specifications
Gemma 3n E4b ITMistral 7B
ProviderGoogleMistral AI
Noometry Index37.323.0
Released—2023-09-27
WeightsOpenOpen
Context window—8K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.25
Results tracked1837

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 3n E4b IT leads

Gemma 3n E4b IT: 37.0 (#198), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Coding12681082
BigCodeBench Instruct—19.5%
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Gemma 3n E4b IT leads

Gemma 3n E4b IT: 19.9 (#247), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Hard Prompts12841067
Kagi LLM Benchmark31.5%—
Chess Puzzles—0%
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Gemma 3n E4b IT leads

Gemma 3n E4b IT: 35.1 (#188), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Math12511085
OTIS Mock AIME 2024-2025—0.3%
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Gemma 3n E4b IT leads

Gemma 3n E4b IT: 34.2 (#198), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Expert12461036
GPQA Diamond—15.2%
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Gemma 3n E4b IT leads

Gemma 3n E4b IT: 43.4 (#183), Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Non-English12851012
LMArena Chinese13091009
LMArena French13301037
LMArena German1311987
LMArena Japanese1272878
LMArena Russian12881018
LMArena Spanish13051026
LMArena Korean1259—

Instruction Following Gemma 3n E4b IT leads

Gemma 3n E4b IT: 66.1 (#210), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Instruction Following12551060

Long Context Gemma 3n E4b IT leads

Gemma 3n E4b IT: 38.7 (#191), Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Longer Query12761060

Writing & Preference Gemma 3n E4b IT leads

Gemma 3n E4b IT: 50.1 (#186), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkGemma 3n E4b ITMistral 7B
LMArena Text13061090
LMArena Creative Writing12871068
LMArena Multi-Turn12761062

Frequently asked questions

Is Gemma 3n E4b IT better than Mistral 7B?

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 23.0 on the Noometry Index.

Is Gemma 3n E4b IT or Mistral 7B better for coding?

Gemma 3n E4b IT scores higher on coding benchmarks: 37.0 versus 26.4 in the Noometry coding category.

How many benchmarks do Gemma 3n E4b IT and Mistral 7B share?

16 benchmarks have published results for both models. Gemma 3n E4b IT has 18 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper