Model comparison

Amazon Nova Lite vs Qwen2.5 Plus 1127

Qwen2.5 Plus 1127 is the stronger model overall, scoring 38.8 to 31.9 on the Noometry Index.

Last verified . 14 shared benchmarks.

Amazon Nova Lite Amazon

31.9

Rank #265 Confirmed

Qwen2.5 Plus 1127 Alibaba (Qwen)

38.8

Rank #181 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Amazon Nova Lite scores higher in 0 categories and Qwen2.5 Plus 1127 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen2.5 Plus 1127 leads 35.5 to 26.3.

Side by side

Amazon Nova Lite and Qwen2.5 Plus 1127 specifications
Amazon Nova LiteQwen2.5 Plus 1127
ProviderAmazonAlibaba (Qwen)
Noometry Index31.938.8
Released2024-12-03—
WeightsProprietaryProprietary
Context window300K—
Max output10K—
Input $ / M tokens$0.06—
Output $ / M tokens$0.24—
Results tracked3414

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 32.5 (#270), Qwen2.5 Plus 1127: 38.5 (#175)

Coding benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Coding12391314
LiveBench Coding27.5%—
ALE-Bench236.25—

Reasoning Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 19.3 (#260), Qwen2.5 Plus 1127: 25.9 (#141)

Reasoning benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Hard Prompts12201299
LiveBench Reasoning36.7%—
LiveBench Data Analysis37.2%—
LiveBench36.4%—

Math Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 27.8 (#247), Qwen2.5 Plus 1127: 36.1 (#174)

Math benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Math12271298
Omni-MATH23.3%—
LiveBench Math36.7%—

Knowledge Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 26.3 (#259), Qwen2.5 Plus 1127: 35.5 (#183)

Knowledge benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Expert12011289
Humanity's Last Exam3.6%—
MMLU-Pro60%—
Vectara Hallucination Rate6.1%—
GPQA (HELM)39.7%—
MMLU77%—

Multimodal Not comparable

Amazon Nova Lite: 25.5 (#123), Qwen2.5 Plus 1127: —

Multimodal benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Vision990—

Multilingual Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 38.0 (#232), Qwen2.5 Plus 1127: 41.9 (#201)

Multilingual benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Non-English12081265
LMArena Chinese12251314
LMArena German12291231
LMArena Japanese11531207
LMArena Russian12161271
LMArena French1237—
LMArena Korean1153—
LMArena Spanish1226—

Instruction Following Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 59.3 (#255), Qwen2.5 Plus 1127: 67.2 (#199)

Instruction Following benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Instruction Following12051275
LiveBench Instruction Following54.1%—
IFEval77.6%—

Long Context Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 37.4 (#217), Qwen2.5 Plus 1127: 39.2 (#184)

Long Context benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Longer Query12341292

Writing & Preference Qwen2.5 Plus 1127 leads

Amazon Nova Lite: 42.1 (#237), Qwen2.5 Plus 1127: 49.4 (#192)

Writing & Preference benchmarks
BenchmarkAmazon Nova LiteQwen2.5 Plus 1127
LMArena Text12291299
LMArena Creative Writing11971262
LMArena Multi-Turn11991299
WildBench75%—
LiveBench Language25.9%—

Frequently asked questions

Is Amazon Nova Lite better than Qwen2.5 Plus 1127?

Qwen2.5 Plus 1127 is the stronger model overall, scoring 38.8 to 31.9 on the Noometry Index.

Is Amazon Nova Lite or Qwen2.5 Plus 1127 better for coding?

Qwen2.5 Plus 1127 scores higher on coding benchmarks: 38.5 versus 32.5 in the Noometry coding category.

How many benchmarks do Amazon Nova Lite and Qwen2.5 Plus 1127 share?

14 benchmarks have published results for both models. Amazon Nova Lite has 34 scored results on Noometry and Qwen2.5 Plus 1127 has 14.

Related comparisons

Go deeper