Model comparison

Amazon Nova Micro vs Qwen2.5-Max

Qwen2.5-Max is the stronger model overall, scoring 40.7 to 30.4 on the Noometry Index.

Last verified . 24 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Qwen2.5-Max Alibaba (Qwen)

40.7

Rank #146 Confirmed

Summary

  • They share 24 benchmarks with published results for both. Amazon Nova Micro scores higher in 0 categories and Qwen2.5-Max in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen2.5-Max leads 55.4 to 39.5.
  • The biggest single-benchmark swing is LiveBench Coding: 20.2% for Amazon Nova Micro and 64.4% for Qwen2.5-Max.

Side by side

Amazon Nova Micro and Qwen2.5-Max specifications
Amazon Nova MicroQwen2.5-Max
ProviderAmazonAlibaba (Qwen)
Noometry Index30.440.7
Released2024-12-032025-01-25
WeightsProprietaryProprietary
Context window128K—
Max output10K—
Input $ / M tokens$0.035—
Output $ / M tokens$0.14—
Results tracked3227

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen2.5-Max leads

Amazon Nova Micro: 30.5 (#295), Qwen2.5-Max: 41.8 (#117)

Coding benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LiveBench Coding20.2%64.4%
LMArena Coding12181359

Agentic & Tool Use Not comparable

Amazon Nova Micro: 22.1 (#132), Qwen2.5-Max: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
Berkeley Function Calling Leaderboard22.3%—

Reasoning Qwen2.5-Max leads

Amazon Nova Micro: 17.4 (#294), Qwen2.5-Max: 25.6 (#147)

Reasoning benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LiveBench Reasoning25.1%51.4%
LMArena Hard Prompts11911360
LiveBench Data Analysis34%67.9%
LiveBench29.6%62.3%
Epoch Capabilities Index—132.53

Math Qwen2.5-Max leads

Amazon Nova Micro: 26.9 (#254), Qwen2.5-Max: 36.9 (#162)

Math benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LiveBench Math34.5%58.4%
LMArena Math12061369
Omni-MATH21.4%—

Knowledge Qwen2.5-Max leads

Amazon Nova Micro: 29.6 (#237), Qwen2.5-Max: 35.3 (#186)

Knowledge benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LMArena Expert11841337
MMLU-Pro51.1%—
Confabulations—21.8%
Vectara Hallucination Rate5.5%—
GPQA (HELM)38.3%—
MMLU70.8%—

Multilingual Qwen2.5-Max leads

Amazon Nova Micro: 36.5 (#239), Qwen2.5-Max: 48.1 (#146)

Multilingual benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LMArena Non-English11861352
LMArena Chinese12091382
LMArena French12381396
LMArena German11921350
LMArena Japanese11541300
LMArena Korean11501304
LMArena Russian11851353
LMArena Spanish12251377

Instruction Following Qwen2.5-Max leads

Amazon Nova Micro: 56.3 (#272), Qwen2.5-Max: 71.3 (#152)

Instruction Following benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LiveBench Instruction Following48%75.3%
LMArena Instruction Following11741335
IFEval76%—

Long Context Qwen2.5-Max leads

Amazon Nova Micro: 36.5 (#229), Qwen2.5-Max: 41.4 (#142)

Long Context benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LMArena Longer Query12051358

Writing & Preference Qwen2.5-Max leads

Amazon Nova Micro: 39.5 (#247), Qwen2.5-Max: 55.4 (#146)

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroQwen2.5-Max
LMArena Text12081367
LMArena Creative Writing11721339
LMArena Multi-Turn11781364
LiveBench Language15.8%56.3%
Short-Story Creative Writing—72.9%
WildBench74.3%—

Frequently asked questions

Is Amazon Nova Micro better than Qwen2.5-Max?

Qwen2.5-Max is the stronger model overall, scoring 40.7 to 30.4 on the Noometry Index.

Is Amazon Nova Micro or Qwen2.5-Max better for coding?

Qwen2.5-Max scores higher on coding benchmarks: 41.8 versus 30.5 in the Noometry coding category.

How many benchmarks do Amazon Nova Micro and Qwen2.5-Max share?

24 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Qwen2.5-Max has 27.

Related comparisons

Go deeper