Model comparison

Gemma 3n E4b IT vs Olmo 2 0325 32b Instruct

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 32.7 on the Noometry Index.

Last verified . 11 shared benchmarks.

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Gemma 3n E4b IT scores higher in 7 categories and Olmo 2 0325 32b Instruct in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemma 3n E4b IT leads 34.2 to 19.5.

Side by side

Gemma 3n E4b IT and Olmo 2 0325 32b Instruct specifications
Gemma 3n E4b ITOlmo 2 0325 32b Instruct
ProviderGoogleAllen Institute for AI (Ai2)
Noometry Index37.332.7
Released——
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1816

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 3n E4b IT leads

Gemma 3n E4b IT: 37.0 (#198), Olmo 2 0325 32b Instruct: 35.2 (#227)

Coding benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Coding12681210

Reasoning Olmo 2 0325 32b Instruct leads

Gemma 3n E4b IT: 19.9 (#247), Olmo 2 0325 32b Instruct: 23.6 (#175)

Reasoning benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Hard Prompts12841208
Kagi LLM Benchmark31.5%—

Math Gemma 3n E4b IT leads

Gemma 3n E4b IT: 35.1 (#188), Olmo 2 0325 32b Instruct: 26.8 (#255)

Math benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Math12511208
Omni-MATH—16.1%

Knowledge Gemma 3n E4b IT leads

Gemma 3n E4b IT: 34.2 (#198), Olmo 2 0325 32b Instruct: 19.5 (#279)

Knowledge benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
MMLU-Pro—41.4%
GPQA (HELM)—28.7%
LMArena Expert1246—

Multilingual Gemma 3n E4b IT leads

Gemma 3n E4b IT: 43.4 (#183), Olmo 2 0325 32b Instruct: 34.8 (#248)

Multilingual benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Non-English12851160
LMArena Chinese13091192
LMArena Russian12881187
LMArena French1330—
LMArena German1311—
LMArena Japanese1272—
LMArena Korean1259—
LMArena Spanish1305—

Instruction Following Gemma 3n E4b IT leads

Gemma 3n E4b IT: 66.1 (#210), Olmo 2 0325 32b Instruct: 61.5 (#244)

Instruction Following benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Instruction Following12551186
IFEval—78%

Long Context Gemma 3n E4b IT leads

Gemma 3n E4b IT: 38.7 (#191), Olmo 2 0325 32b Instruct: 36.2 (#234)

Long Context benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Longer Query12761194

Writing & Preference Gemma 3n E4b IT leads

Gemma 3n E4b IT: 50.1 (#186), Olmo 2 0325 32b Instruct: 42.1 (#236)

Writing & Preference benchmarks
BenchmarkGemma 3n E4b ITOlmo 2 0325 32b Instruct
LMArena Text13061218
LMArena Creative Writing12871199
LMArena Multi-Turn12761221
WildBench—73.4%

Frequently asked questions

Is Gemma 3n E4b IT better than Olmo 2 0325 32b Instruct?

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 32.7 on the Noometry Index.

Is Gemma 3n E4b IT or Olmo 2 0325 32b Instruct better for coding?

Gemma 3n E4b IT scores higher on coding benchmarks: 37.0 versus 35.2 in the Noometry coding category.

How many benchmarks do Gemma 3n E4b IT and Olmo 2 0325 32b Instruct share?

11 benchmarks have published results for both models. Gemma 3n E4b IT has 18 scored results on Noometry and Olmo 2 0325 32b Instruct has 16.

Related comparisons

Go deeper