Model comparison

Gemini 2.0 Pro vs Olmo 3 32b Think

Gemini 2.0 Pro and Olmo 3 32b Think score almost the same on the Noometry Index (39.1 vs 38.7), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Gemini 2.0 Pro Google

39.1

Rank #173 Confirmed

Summary

  • The widest gap is in long context, where Olmo 3 32b Think leads 39.4 to 29.2.
  • Olmo 3 32b Think has downloadable open weights; the other is API-only.

Side by side

Gemini 2.0 Pro and Olmo 3 32b Think specifications
Gemini 2.0 ProOlmo 3 32b Think
ProviderGoogleAllen Institute for AI (Ai2)
Noometry Index39.138.7
Released2025-02-05—
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1414

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemini 2.0 Pro: 37.8 (#187), Olmo 3 32b Think: 38.6 (#172)

Coding benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
Aider Polyglot35.6%—
LiveBench Coding63.5%—
LMArena Coding—1319

Reasoning Olmo 3 32b Think leads

Gemini 2.0 Pro: 22.3 (#198), Olmo 3 32b Think: 25.9 (#140)

Reasoning benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
EnigmaEval0.7%—
LiveBench Reasoning60.1%—
LMArena Hard Prompts—1302
LiveBench Data Analysis68%—
Epoch Capabilities Index135.06—
LiveBench65.1%—

Math Gemini 2.0 Pro leads

Gemini 2.0 Pro: 39.7 (#100), Olmo 3 32b Think: 36.5 (#165)

Math benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
LiveBench Math71%—
LMArena Math—1316
MATH Level 583.5%—

Knowledge Gemini 2.0 Pro leads

Gemini 2.0 Pro: 36.5 (#167), Olmo 3 32b Think: 35.0 (#190)

Knowledge benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
GPQA Diamond65.7%—
Confabulations18.4%—
LMArena Expert—1273

Multilingual Not comparable

Gemini 2.0 Pro: —, Olmo 3 32b Think: 41.2 (#210)

Multilingual benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
LMArena Non-English—1255
LMArena Chinese—1300
LMArena French—1291
LMArena German—1290
LMArena Russian—1254

Instruction Following Gemini 2.0 Pro leads

Gemini 2.0 Pro: 75.5 (#59), Olmo 3 32b Think: 67.2 (#198)

Instruction Following benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
LiveBench Instruction Following83.4%—
LMArena Instruction Following—1275

Long Context Olmo 3 32b Think leads

Gemini 2.0 Pro: 29.2 (#292), Olmo 3 32b Think: 39.4 (#182)

Long Context benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
Fiction.LiveBench41.7%—
LMArena Longer Query—1296

Writing & Preference Gemini 2.0 Pro leads

Gemini 2.0 Pro: 52.7 (#165), Olmo 3 32b Think: 49.1 (#193)

Writing & Preference benchmarks
BenchmarkGemini 2.0 ProOlmo 3 32b Think
LMArena Text—1300
LMArena Creative Writing—1256
LMArena Multi-Turn—1290
LiveBench Language44.9%—

Frequently asked questions

Is Gemini 2.0 Pro better than Olmo 3 32b Think?

Gemini 2.0 Pro and Olmo 3 32b Think score almost the same on the Noometry Index (39.1 vs 38.7), so choose on price, context window or the category you care about most.

Is Gemini 2.0 Pro or Olmo 3 32b Think better for coding?

They score almost the same on coding (37.8 vs 38.6); test both on your own repository before choosing.

How many benchmarks do Gemini 2.0 Pro and Olmo 3 32b Think share?

0 benchmarks have published results for both models. Gemini 2.0 Pro has 14 scored results on Noometry and Olmo 3 32b Think has 14.

Related comparisons

Go deeper