Model comparison

Amazon Nova Micro vs Mercury 2.5

Mercury 2.5 is the stronger model overall, scoring 33.5 to 30.4 on the Noometry Index.

Last verified . 0 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Mercury 2.5 Inception

33.5

Rank #242 Reported

Summary

  • The widest gap is in coding, where Mercury 2.5 leads 39.5 to 30.5.
  • Amazon Nova Micro is cheaper at $0.035 / $0.14 per million input/output tokens, against $0.04 / $0.15 for Mercury 2.5.
  • Mercury 2.5 accepts more context: 260K tokens versus 128K.

Side by side

Amazon Nova Micro and Mercury 2.5 specifications
Amazon Nova MicroMercury 2.5
ProviderAmazonInception
Noometry Index30.433.5
Released2024-12-032026-09-08
WeightsProprietaryProprietary
Context window128K260K
Max output10K66K
Input $ / M tokens$0.035$0.04
Output $ / M tokens$0.14$0.15
Results tracked324

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury 2.5 leads

Amazon Nova Micro: 30.5 (#295), Mercury 2.5: 39.5 (#156)

Coding benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
SciCode—38.5%
LiveBench Coding20.2%—
LMArena Coding1218—
ALE-Bench—301.65

Agentic & Tool Use Not comparable

Amazon Nova Micro: 22.1 (#132), Mercury 2.5: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
Berkeley Function Calling Leaderboard22.3%—

Reasoning Mercury 2.5 leads

Amazon Nova Micro: 17.4 (#294), Mercury 2.5: 22.4 (#193)

Reasoning benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
CritPt—0%
LiveBench Reasoning25.1%—
LMArena Hard Prompts1191—
LiveBench Data Analysis34%—
LiveBench29.6%—

Math Amazon Nova Micro leads

Amazon Nova Micro: 26.9 (#254), Mercury 2.5: 23.3 (#272)

Math benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
ProofBench—3%
Omni-MATH21.4%—
LiveBench Math34.5%—
LMArena Math1206—

Knowledge Not comparable

Amazon Nova Micro: 29.6 (#237), Mercury 2.5: —

Knowledge benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
MMLU-Pro51.1%—
Vectara Hallucination Rate5.5%—
GPQA (HELM)38.3%—
LMArena Expert1184—
MMLU70.8%—

Multilingual Not comparable

Amazon Nova Micro: 36.5 (#239), Mercury 2.5: —

Multilingual benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
LMArena Non-English1186—
LMArena Chinese1209—
LMArena French1238—
LMArena German1192—
LMArena Japanese1154—
LMArena Korean1150—
LMArena Russian1185—
LMArena Spanish1225—

Instruction Following Not comparable

Amazon Nova Micro: 56.3 (#272), Mercury 2.5: —

Instruction Following benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
LiveBench Instruction Following48%—
IFEval76%—
LMArena Instruction Following1174—

Long Context Not comparable

Amazon Nova Micro: 36.5 (#229), Mercury 2.5: —

Long Context benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
LMArena Longer Query1205—

Writing & Preference Not comparable

Amazon Nova Micro: 39.5 (#247), Mercury 2.5: —

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroMercury 2.5
LMArena Text1208—
LMArena Creative Writing1172—
WildBench74.3%—
LMArena Multi-Turn1178—
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than Mercury 2.5?

Mercury 2.5 is the stronger model overall, scoring 33.5 to 30.4 on the Noometry Index.

Which is cheaper, Amazon Nova Micro or Mercury 2.5?

Amazon Nova Micro is cheaper. It lists at $0.035 per million input tokens and $0.14 per million output tokens; Mercury 2.5 lists at $0.04 and $0.15.

Is Amazon Nova Micro or Mercury 2.5 better for coding?

Mercury 2.5 scores higher on coding benchmarks: 39.5 versus 30.5 in the Noometry coding category.

Which has the bigger context window?

Mercury 2.5 does, with 260K tokens against 128K.

How many benchmarks do Amazon Nova Micro and Mercury 2.5 share?

0 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Mercury 2.5 has 4.

Related comparisons

Go deeper