Model comparison

Ministral 8B vs Olmo 3.1 32b Think

Olmo 3.1 32b Think is the stronger model overall, scoring 37.9 to 28.2 on the Noometry Index.

Last verified . 12 shared benchmarks.

Ministral 8B Mistral AI

28.2

Rank #325 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Ministral 8B scores higher in 0 categories and Olmo 3.1 32b Think in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Olmo 3.1 32b Think leads 35.7 to 12.6.

Side by side

Ministral 8B and Olmo 3.1 32b Think specifications
Ministral 8BOlmo 3.1 32b Think
ProviderMistral AIAllen Institute for AI (Ai2)
Noometry Index28.237.9
Released2024-10-01—
WeightsOpenOpen
Context window262K—
Max output262K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.15—
Results tracked1715

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Olmo 3.1 32b Think leads

Ministral 8B: 35.0 (#230), Olmo 3.1 32b Think: 37.7 (#189)

Coding benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Coding12021291

Agentic & Tool Use Not comparable

Ministral 8B: 16.4 (#148), Olmo 3.1 32b Think: —

Agentic & Tool Use benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
Berkeley Function Calling Leaderboard11.1%—

Reasoning Olmo 3.1 32b Think leads

Ministral 8B: 18.4 (#281), Olmo 3.1 32b Think: 25.2 (#150)

Reasoning benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Hard Prompts11911272
DTBench45.7%—

Math Olmo 3.1 32b Think leads

Ministral 8B: 25.7 (#267), Olmo 3.1 32b Think: 36.3 (#168)

Math benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Math11881305
MATH Level 514.9%—

Knowledge Olmo 3.1 32b Think leads

Ministral 8B: 12.6 (#297), Olmo 3.1 32b Think: 35.7 (#181)

Knowledge benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Expert11701295
GPQA Diamond27.1%—
Vectara Hallucination Rate7.4%—

Multilingual Olmo 3.1 32b Think leads

Ministral 8B: 35.1 (#247), Olmo 3.1 32b Think: 38.1 (#231)

Multilingual benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Non-English11651209
LMArena Chinese11931242
LMArena Russian11951193
LMArena French—1260
LMArena German—1262
LMArena Spanish—1289

Instruction Following Olmo 3.1 32b Think leads

Ministral 8B: 60.5 (#250), Olmo 3.1 32b Think: 65.6 (#218)

Instruction Following benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Instruction Following11611247

Long Context Olmo 3.1 32b Think leads

Ministral 8B: 36.7 (#227), Olmo 3.1 32b Think: 38.6 (#195)

Long Context benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Longer Query12121272

Writing & Preference Olmo 3.1 32b Think leads

Ministral 8B: 39.6 (#246), Olmo 3.1 32b Think: 46.2 (#220)

Writing & Preference benchmarks
BenchmarkMinistral 8BOlmo 3.1 32b Think
LMArena Text11911272
LMArena Creative Writing11751226
LMArena Multi-Turn11661252

Frequently asked questions

Is Ministral 8B better than Olmo 3.1 32b Think?

Olmo 3.1 32b Think is the stronger model overall, scoring 37.9 to 28.2 on the Noometry Index.

Is Ministral 8B or Olmo 3.1 32b Think better for coding?

Olmo 3.1 32b Think scores higher on coding benchmarks: 37.7 versus 35.0 in the Noometry coding category.

How many benchmarks do Ministral 8B and Olmo 3.1 32b Think share?

12 benchmarks have published results for both models. Ministral 8B has 17 scored results on Noometry and Olmo 3.1 32b Think has 15.

Related comparisons

Go deeper