Model comparison

Amazon Nova Micro vs Claude Sonnet 4.5

Claude Sonnet 4.5 is the stronger model overall, scoring 44.1 to 30.4 on the Noometry Index. Amazon Nova Micro costs 98× less per token, which makes it the better buy when Claude Sonnet 4.5's lead doesn't matter for your workload.

Last verified . 24 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Claude Sonnet 4.5 Anthropic

44.1

Rank #81 Confirmed

Summary

  • They share 24 benchmarks with published results for both. Amazon Nova Micro scores higher in 0 categories and Claude Sonnet 4.5 in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Claude Sonnet 4.5 leads 66.5 to 39.5.
  • The biggest single-benchmark swing is Berkeley Function Calling Leaderboard: 22.3% for Amazon Nova Micro and 73.2% for Claude Sonnet 4.5.
  • Amazon Nova Micro is cheaper at $0.035 / $0.14 per million input/output tokens, against $3 / $15 for Claude Sonnet 4.5.
  • Claude Sonnet 4.5 accepts more context: 200K tokens versus 128K.

Side by side

Amazon Nova Micro and Claude Sonnet 4.5 specifications
Amazon Nova MicroClaude Sonnet 4.5
ProviderAmazonAnthropic
Noometry Index30.444.1
Released2024-12-032025-09-29
WeightsProprietaryProprietary
Context window128K200K
Max output10K64K
Input $ / M tokens$0.035$3
Output $ / M tokens$0.14$15
Results tracked3273

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Sonnet 4.5 leads

Amazon Nova Micro: 30.5 (#295), Claude Sonnet 4.5: 47.3 (#61)

Coding benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
LMArena Coding12181489
SWE-bench Verified—71.3%
SWE-bench Verified (bash only)—71.4%
LMArena WebDev—1393
SWE-bench Multilingual—67%
SciCode—44.7%
GSO—14.7%
WeirdML—47.7%
LiveBench Coding20.2%—
ALE-Bench—796.15
AlgoTune—1.52

Agentic & Tool Use Claude Sonnet 4.5 leads

Amazon Nova Micro: 22.1 (#132), Claude Sonnet 4.5: 38.3 (#32)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
Berkeley Function Calling Leaderboard22.3%73.2%
Terminal-Bench—46.5%
GDPval—42.5%
Remote Labor Index—2.1%
τ²-bench Airline—72%
τ²-bench Banking—25.3%
τ²-bench Retail—72.4%
τ²-bench Telecom—84.9%
Cybench—60%
DeepResearch Bench—52.6%
OSWorld—62.9%
LMArena Search—1159
METR Time Horizons—67.4%
Vending-Bench 2—3,839

Reasoning Claude Sonnet 4.5 leads

Amazon Nova Micro: 17.4 (#294), Claude Sonnet 4.5: 26.9 (#125)

Reasoning benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
LMArena Hard Prompts11911462
ARC-AGI-2—13.6%
SimpleBench—54.3%
Kagi LLM Benchmark—57.9%
NYT Connections (extended)—37.3%
ARC-AGI-1—63.7%
CritPt—1.1%
Chess Puzzles—12%
EnigmaEval—6%
EBR-Bench—2.4%
LiveBench Reasoning25.1%—
Mystery Game Puzzles—17%
DTBench—83.2%
LiveBench Data Analysis34%—
LMCA—38.8%
Epoch Capabilities Index—146.84
ForecastBench—61.9
LiveBench29.6%—

Math Claude Sonnet 4.5 leads

Amazon Nova Micro: 26.9 (#254), Claude Sonnet 4.5: 32.3 (#216)

Math benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
Omni-MATH21.4%55.3%
LMArena Math12061449
FrontierMath (Tiers 1-3)—23.9%
FrontierMath Tier 4—2.4%
OTIS Mock AIME 2024-2025—77.8%
ProofBench—19%
LiveBench Math34.5%—
MATH Level 5—97.7%
FrontierMath (Feb 2025 set)—15.2%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Claude Sonnet 4.5 leads

Amazon Nova Micro: 29.6 (#237), Claude Sonnet 4.5: 48.4 (#76)

Knowledge benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
MMLU-Pro51.1%86.9%
Vectara Hallucination Rate5.5%12%
GPQA (HELM)38.3%68.6%
LMArena Expert11841482
GPQA Diamond—82.3%
Humanity's Last Exam—13.7%
SimpleQA Verified—30.7%
MMLU70.8%—

Multimodal Not comparable

Amazon Nova Micro: —, Claude Sonnet 4.5: 34.8 (#89)

Multimodal benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
VPCT—39.8%
LMArena Document—1450

Multilingual Claude Sonnet 4.5 leads

Amazon Nova Micro: 36.5 (#239), Claude Sonnet 4.5: 53.4 (#69)

Multilingual benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
LMArena Non-English11861425
LMArena Chinese12091459
LMArena French12381458
LMArena German11921427
LMArena Japanese11541390
LMArena Korean11501403
LMArena Russian11851437
LMArena Spanish12251457

Instruction Following Claude Sonnet 4.5 leads

Amazon Nova Micro: 56.3 (#272), Claude Sonnet 4.5: 75.0 (#78)

Instruction Following benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
IFEval76%85%
LMArena Instruction Following11741459
LiveBench Instruction Following48%—

Long Context Claude Sonnet 4.5 leads

Amazon Nova Micro: 36.5 (#229), Claude Sonnet 4.5: 45.2 (#46)

Long Context benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
LMArena Longer Query12051476

Writing & Preference Claude Sonnet 4.5 leads

Amazon Nova Micro: 39.5 (#247), Claude Sonnet 4.5: 66.5 (#34)

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroClaude Sonnet 4.5
LMArena Text12081439
LMArena Creative Writing11721442
WildBench74.3%85.4%
LMArena Multi-Turn11781465
EQ-Bench Creative Writing—1678
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than Claude Sonnet 4.5?

Claude Sonnet 4.5 is the stronger model overall, scoring 44.1 to 30.4 on the Noometry Index. Amazon Nova Micro costs 98× less per token, which makes it the better buy when Claude Sonnet 4.5's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Micro or Claude Sonnet 4.5?

Amazon Nova Micro is cheaper. It lists at $0.035 per million input tokens and $0.14 per million output tokens; Claude Sonnet 4.5 lists at $3 and $15.

Is Amazon Nova Micro or Claude Sonnet 4.5 better for coding?

Claude Sonnet 4.5 scores higher on coding benchmarks: 47.3 versus 30.5 in the Noometry coding category.

Which has the bigger context window?

Claude Sonnet 4.5 does, with 200K tokens against 128K.

How many benchmarks do Amazon Nova Micro and Claude Sonnet 4.5 share?

24 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Claude Sonnet 4.5 has 73.

Related comparisons

Go deeper