Model comparison

Gemma 3n E4b IT vs Mistral Medium 3.1

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 31.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in reasoning, where Gemma 3n E4b IT leads 19.9 to 10.6.
  • Gemma 3n E4b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 3n E4b IT and Mistral Medium 3.1 specifications
Gemma 3n E4b ITMistral Medium 3.1
ProviderGoogleMistral AI
Noometry Index37.331.9
Released——
WeightsOpenProprietary
Context window—131K
Max output—105K
Input $ / M tokens—$0.40
Output $ / M tokens—$2
Results tracked183

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 3n E4b IT: 37.0 (#198), Mistral Medium 3.1: —

Coding benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Coding1268—

Reasoning Gemma 3n E4b IT leads

Gemma 3n E4b IT: 19.9 (#247), Mistral Medium 3.1: 10.6 (#341)

Reasoning benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
Kagi LLM Benchmark31.5%—
NYT Connections (extended)—6.5%
Thematic Generalization—20.3%
LMArena Hard Prompts1284—

Math Not comparable

Gemma 3n E4b IT: 35.1 (#188), Mistral Medium 3.1: —

Math benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Math1251—

Knowledge Not comparable

Gemma 3n E4b IT: 34.2 (#198), Mistral Medium 3.1: —

Knowledge benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Expert1246—

Multilingual Not comparable

Gemma 3n E4b IT: 43.4 (#183), Mistral Medium 3.1: —

Multilingual benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Non-English1285—
LMArena Chinese1309—
LMArena French1330—
LMArena German1311—
LMArena Japanese1272—
LMArena Korean1259—
LMArena Russian1288—
LMArena Spanish1305—

Instruction Following Not comparable

Gemma 3n E4b IT: 66.1 (#210), Mistral Medium 3.1: —

Instruction Following benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Instruction Following1255—

Long Context Not comparable

Gemma 3n E4b IT: 38.7 (#191), Mistral Medium 3.1: —

Long Context benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Longer Query1276—

Writing & Preference Mistral Medium 3.1 leads

Gemma 3n E4b IT: 50.1 (#186), Mistral Medium 3.1: 55.5 (#145)

Writing & Preference benchmarks
BenchmarkGemma 3n E4b ITMistral Medium 3.1
LMArena Text1306—
LMArena Creative Writing1287—
EQ-Bench Creative Writing—1476
LMArena Multi-Turn1276—

Frequently asked questions

Is Gemma 3n E4b IT better than Mistral Medium 3.1?

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 31.9 on the Noometry Index.

How many benchmarks do Gemma 3n E4b IT and Mistral Medium 3.1 share?

0 benchmarks have published results for both models. Gemma 3n E4b IT has 18 scored results on Noometry and Mistral Medium 3.1 has 3.

Related comparisons

Go deeper