Model comparison

Amazon Nova Micro vs Olmo 3.1 32b Think

Olmo 3.1 32b Think is the stronger model overall, scoring 37.9 to 30.4 on the Noometry Index.

Last verified . 15 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Summary

  • They share 15 benchmarks with published results for both. Amazon Nova Micro scores higher in 0 categories and Olmo 3.1 32b Think in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Olmo 3.1 32b Think leads 36.3 to 26.9.
  • Olmo 3.1 32b Think has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Micro and Olmo 3.1 32b Think specifications
Amazon Nova MicroOlmo 3.1 32b Think
ProviderAmazonAllen Institute for AI (Ai2)
Noometry Index30.437.9
Released2024-12-03—
WeightsProprietaryOpen
Context window128K—
Max output10K—
Input $ / M tokens$0.035—
Output $ / M tokens$0.14—
Results tracked3215

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Olmo 3.1 32b Think leads

Amazon Nova Micro: 30.5 (#295), Olmo 3.1 32b Think: 37.7 (#189)

Coding benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Coding12181291
LiveBench Coding20.2%—

Agentic & Tool Use Not comparable

Amazon Nova Micro: 22.1 (#132), Olmo 3.1 32b Think: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
Berkeley Function Calling Leaderboard22.3%—

Reasoning Olmo 3.1 32b Think leads

Amazon Nova Micro: 17.4 (#294), Olmo 3.1 32b Think: 25.2 (#150)

Reasoning benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Hard Prompts11911272
LiveBench Reasoning25.1%—
LiveBench Data Analysis34%—
LiveBench29.6%—

Math Olmo 3.1 32b Think leads

Amazon Nova Micro: 26.9 (#254), Olmo 3.1 32b Think: 36.3 (#168)

Math benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Math12061305
Omni-MATH21.4%—
LiveBench Math34.5%—

Knowledge Olmo 3.1 32b Think leads

Amazon Nova Micro: 29.6 (#237), Olmo 3.1 32b Think: 35.7 (#181)

Knowledge benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Expert11841295
MMLU-Pro51.1%—
Vectara Hallucination Rate5.5%—
GPQA (HELM)38.3%—
MMLU70.8%—

Multilingual Olmo 3.1 32b Think leads

Amazon Nova Micro: 36.5 (#239), Olmo 3.1 32b Think: 38.1 (#231)

Multilingual benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Non-English11861209
LMArena Chinese12091242
LMArena French12381260
LMArena German11921262
LMArena Russian11851193
LMArena Spanish12251289
LMArena Japanese1154—
LMArena Korean1150—

Instruction Following Olmo 3.1 32b Think leads

Amazon Nova Micro: 56.3 (#272), Olmo 3.1 32b Think: 65.6 (#218)

Instruction Following benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Instruction Following11741247
LiveBench Instruction Following48%—
IFEval76%—

Long Context Olmo 3.1 32b Think leads

Amazon Nova Micro: 36.5 (#229), Olmo 3.1 32b Think: 38.6 (#195)

Long Context benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Longer Query12051272

Writing & Preference Olmo 3.1 32b Think leads

Amazon Nova Micro: 39.5 (#247), Olmo 3.1 32b Think: 46.2 (#220)

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroOlmo 3.1 32b Think
LMArena Text12081272
LMArena Creative Writing11721226
LMArena Multi-Turn11781252
WildBench74.3%—
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than Olmo 3.1 32b Think?

Olmo 3.1 32b Think is the stronger model overall, scoring 37.9 to 30.4 on the Noometry Index.

Is Amazon Nova Micro or Olmo 3.1 32b Think better for coding?

Olmo 3.1 32b Think scores higher on coding benchmarks: 37.7 versus 30.5 in the Noometry coding category.

How many benchmarks do Amazon Nova Micro and Olmo 3.1 32b Think share?

15 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Olmo 3.1 32b Think has 15.

Related comparisons

Go deeper