Model comparison

Amazon Nova Micro vs Gemma 4 31B IT

Gemma 4 31B IT is the stronger model overall, scoring 43.5 to 30.4 on the Noometry Index. Amazon Nova Micro costs 2.5× less per token, which makes it the better buy when Gemma 4 31B IT's lead doesn't matter for your workload.

Last verified . 15 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Gemma 4 31B IT Google

43.5

Rank #90 Confirmed

Summary

  • They share 15 benchmarks with published results for both. Amazon Nova Micro scores higher in 0 categories and Gemma 4 31B IT in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemma 4 31B IT leads 60.5 to 39.5.
  • Amazon Nova Micro is cheaper at $0.035 / $0.14 per million input/output tokens, against $0.09 / $0.34 for Gemma 4 31B IT.
  • Gemma 4 31B IT accepts more context: 262K tokens versus 128K.
  • Gemma 4 31B IT has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Micro and Gemma 4 31B IT specifications
Amazon Nova MicroGemma 4 31B IT
ProviderAmazonGoogle
Noometry Index30.443.5
Released2024-12-032026-04-02
WeightsProprietaryOpen
Context window128K262K
Max output10K33K
Input $ / M tokens$0.035$0.09
Output $ / M tokens$0.14$0.34
Results tracked3235

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 4 31B IT leads

Amazon Nova Micro: 30.5 (#295), Gemma 4 31B IT: 42.3 (#108)

Coding benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Coding12181459
LMArena WebDev—1366
SciCode—43.4%
WeirdML—52.3%
LiveBench Coding20.2%—
ALE-Bench—925.5

Agentic & Tool Use Not comparable

Amazon Nova Micro: 22.1 (#132), Gemma 4 31B IT: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
Berkeley Function Calling Leaderboard22.3%—

Reasoning Gemma 4 31B IT leads

Amazon Nova Micro: 17.4 (#294), Gemma 4 31B IT: 27.2 (#122)

Reasoning benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Hard Prompts11911448
Kagi LLM Benchmark—63.5%
NYT Connections (extended)—70.6%
CritPt—1.4%
Chess Puzzles—5%
Thematic Generalization—53%
LiveBench Reasoning25.1%—
DTBench—82.7%
LiveBench Data Analysis34%—
LMCA—39.3%
Surface Evolver Bench—30.6%
Epoch Capabilities Index—142.74
LiveBench29.6%—

Math Gemma 4 31B IT leads

Amazon Nova Micro: 26.9 (#254), Gemma 4 31B IT: 43.2 (#81)

Math benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Math12061465
OTIS Mock AIME 2024-2025—73.3%
Omni-MATH21.4%—
LiveBench Math34.5%—

Knowledge Gemma 4 31B IT leads

Amazon Nova Micro: 29.6 (#237), Gemma 4 31B IT: 37.9 (#151)

Knowledge benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
Vectara Hallucination Rate5.5%7.4%
LMArena Expert11841465
GPQA Diamond—75.8%
SimpleQA Verified—10.4%
MMLU-Pro51.1%—
GPQA (HELM)38.3%—
MMLU70.8%—

Multimodal Not comparable

Amazon Nova Micro: —, Gemma 4 31B IT: 41.6 (#34)

Multimodal benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Vision—1277
LMArena Document—1425

Multilingual Gemma 4 31B IT leads

Amazon Nova Micro: 36.5 (#239), Gemma 4 31B IT: 53.8 (#57)

Multilingual benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Non-English11861431
LMArena Chinese12091476
LMArena French12381435
LMArena Russian11851460
LMArena Spanish12251444
LMArena German1192—
LMArena Japanese1154—
LMArena Korean1150—

Instruction Following Gemma 4 31B IT leads

Amazon Nova Micro: 56.3 (#272), Gemma 4 31B IT: 75.5 (#61)

Instruction Following benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Instruction Following11741433
LiveBench Instruction Following48%—
IFEval76%—

Long Context Gemma 4 31B IT leads

Amazon Nova Micro: 36.5 (#229), Gemma 4 31B IT: 44.2 (#71)

Long Context benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Longer Query12051446

Writing & Preference Gemma 4 31B IT leads

Amazon Nova Micro: 39.5 (#247), Gemma 4 31B IT: 60.5 (#96)

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroGemma 4 31B IT
LMArena Text12081443
LMArena Creative Writing11721415
LMArena Multi-Turn11781452
EQ-Bench Creative Writing—1368
WildBench74.3%—
EQ-Bench 4—1120
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than Gemma 4 31B IT?

Gemma 4 31B IT is the stronger model overall, scoring 43.5 to 30.4 on the Noometry Index. Amazon Nova Micro costs 2.5× less per token, which makes it the better buy when Gemma 4 31B IT's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Micro or Gemma 4 31B IT?

Amazon Nova Micro is cheaper. It lists at $0.035 per million input tokens and $0.14 per million output tokens; Gemma 4 31B IT lists at $0.09 and $0.34.

Is Amazon Nova Micro or Gemma 4 31B IT better for coding?

Gemma 4 31B IT scores higher on coding benchmarks: 42.3 versus 30.5 in the Noometry coding category.

Which has the bigger context window?

Gemma 4 31B IT does, with 262K tokens against 128K.

How many benchmarks do Amazon Nova Micro and Gemma 4 31B IT share?

15 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Gemma 4 31B IT has 35.

Related comparisons

Go deeper