Model comparison

Amazon Nova Lite vs Mercury 2.5

Mercury 2.5 is the stronger model overall, scoring 33.5 to 31.9 on the Noometry Index.

Last verified . 1 shared benchmarks.

Amazon Nova Lite Amazon

31.9

Rank #265 Confirmed

Mercury 2.5 Inception

33.5

Rank #242 Reported

Summary

  • They share 1 benchmark with published results for both. Amazon Nova Lite scores higher in 1 category and Mercury 2.5 in 2 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Mercury 2.5 leads 39.5 to 32.5.
  • Mercury 2.5 is cheaper at $0.04 / $0.15 per million input/output tokens, against $0.06 / $0.24 for Amazon Nova Lite.
  • Amazon Nova Lite accepts more context: 300K tokens versus 260K.

Side by side

Amazon Nova Lite and Mercury 2.5 specifications
Amazon Nova LiteMercury 2.5
ProviderAmazonInception
Noometry Index31.933.5
Released2024-12-032026-09-08
WeightsProprietaryProprietary
Context window300K260K
Max output10K66K
Input $ / M tokens$0.06$0.04
Output $ / M tokens$0.24$0.15
Results tracked344

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury 2.5 leads

Amazon Nova Lite: 32.5 (#270), Mercury 2.5: 39.5 (#156)

Coding benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
ALE-Bench236.25301.65
SciCode—38.5%
LiveBench Coding27.5%—
LMArena Coding1239—

Reasoning Mercury 2.5 leads

Amazon Nova Lite: 19.3 (#260), Mercury 2.5: 22.4 (#193)

Reasoning benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
CritPt—0%
LiveBench Reasoning36.7%—
LMArena Hard Prompts1220—
LiveBench Data Analysis37.2%—
LiveBench36.4%—

Math Amazon Nova Lite leads

Amazon Nova Lite: 27.8 (#247), Mercury 2.5: 23.3 (#272)

Math benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
ProofBench—3%
Omni-MATH23.3%—
LiveBench Math36.7%—
LMArena Math1227—

Knowledge Not comparable

Amazon Nova Lite: 26.3 (#259), Mercury 2.5: —

Knowledge benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
Humanity's Last Exam3.6%—
MMLU-Pro60%—
Vectara Hallucination Rate6.1%—
GPQA (HELM)39.7%—
LMArena Expert1201—
MMLU77%—

Multimodal Not comparable

Amazon Nova Lite: 25.5 (#123), Mercury 2.5: —

Multimodal benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
LMArena Vision990—

Multilingual Not comparable

Amazon Nova Lite: 38.0 (#232), Mercury 2.5: —

Multilingual benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
LMArena Non-English1208—
LMArena Chinese1225—
LMArena French1237—
LMArena German1229—
LMArena Japanese1153—
LMArena Korean1153—
LMArena Russian1216—
LMArena Spanish1226—

Instruction Following Not comparable

Amazon Nova Lite: 59.3 (#255), Mercury 2.5: —

Instruction Following benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
LiveBench Instruction Following54.1%—
IFEval77.6%—
LMArena Instruction Following1205—

Long Context Not comparable

Amazon Nova Lite: 37.4 (#217), Mercury 2.5: —

Long Context benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
LMArena Longer Query1234—

Writing & Preference Not comparable

Amazon Nova Lite: 42.1 (#237), Mercury 2.5: —

Writing & Preference benchmarks
BenchmarkAmazon Nova LiteMercury 2.5
LMArena Text1229—
LMArena Creative Writing1197—
WildBench75%—
LMArena Multi-Turn1199—
LiveBench Language25.9%—

Frequently asked questions

Is Amazon Nova Lite better than Mercury 2.5?

Mercury 2.5 is the stronger model overall, scoring 33.5 to 31.9 on the Noometry Index.

Which is cheaper, Amazon Nova Lite or Mercury 2.5?

Mercury 2.5 is cheaper. It lists at $0.04 per million input tokens and $0.15 per million output tokens; Amazon Nova Lite lists at $0.06 and $0.24.

Is Amazon Nova Lite or Mercury 2.5 better for coding?

Mercury 2.5 scores higher on coding benchmarks: 39.5 versus 32.5 in the Noometry coding category.

Which has the bigger context window?

Amazon Nova Lite does, with 300K tokens against 260K.

How many benchmarks do Amazon Nova Lite and Mercury 2.5 share?

1 benchmark has published results for both models. Amazon Nova Lite has 34 scored results on Noometry and Mercury 2.5 has 4.

Related comparisons

Go deeper