Model comparison

Amazon Nova Pro vs Claude Opus 4

Claude Opus 4 is the stronger model overall, scoring 43.1 to 31.0 on the Noometry Index. Amazon Nova Pro costs 21× less per token, which makes it the better buy when Claude Opus 4's lead doesn't matter for your workload.

Last verified . 28 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

Claude Opus 4 Anthropic

43.1

Rank #100 Confirmed

Summary

  • They share 28 benchmarks with published results for both. Amazon Nova Pro scores higher in 0 categories and Claude Opus 4 in 10 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in agentic & tool use, where Claude Opus 4 leads 34.8 to 16.7.
  • The biggest single-benchmark swing is Omni-MATH: 24.2% for Amazon Nova Pro and 61.6% for Claude Opus 4.
  • Amazon Nova Pro is cheaper at $0.80 / $3.20 per million input/output tokens, against $15 / $75 for Claude Opus 4.
  • Amazon Nova Pro accepts more context: 300K tokens versus 200K.

Side by side

Amazon Nova Pro and Claude Opus 4 specifications
Amazon Nova ProClaude Opus 4
ProviderAmazonAnthropic
Noometry Index31.043.1
Released2024-12-032025-05-22
WeightsProprietaryProprietary
Context window300K200K
Max output10K32K
Input $ / M tokens$0.80$15
Output $ / M tokens$3.20$75
Results tracked3856

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Opus 4 leads

Amazon Nova Pro: 35.1 (#229), Claude Opus 4: 47.2 (#62)

Coding benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
LMArena Coding12701442
SWE-bench Verified—70.7%
SWE-bench Verified (bash only)—67.6%
Aider Polyglot—72%
GSO—6.9%
WeirdML—43.7%
LiveBench Coding38.1%—
AlgoTune—1.33

Agentic & Tool Use Claude Opus 4 leads

Amazon Nova Pro: 16.7 (#147), Claude Opus 4: 34.8 (#42)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—
Cybench—38%
DeepResearch Bench—46.8%
LMArena Search—1127
METR Time Horizons—63.9%

Reasoning Claude Opus 4 leads

Amazon Nova Pro: 20.0 (#243), Claude Opus 4: 27.3 (#121)

Reasoning benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
LMArena Hard Prompts12461399
Epoch Capabilities Index123.8142.67
ARC-AGI-2—8.6%
SimpleBench—58.8%
Kagi LLM Benchmark—74.3%
ARC-AGI-1—35.7%
CritPt—0.3%
EnigmaEval—5.6%
LiveBench Reasoning32.6%—
DTBench—81.6%
LiveBench Data Analysis48.3%—
LMCA—37.4%
ForecastBench—61.1
LiveBench43.5%—

Math Claude Opus 4 leads

Amazon Nova Pro: 28.5 (#243), Claude Opus 4: 42.0 (#86)

Math benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
Omni-MATH24.2%61.6%
LMArena Math12521390
OTIS Mock AIME 2024-2025—64.4%
LiveBench Math38%—
MATH Level 5—85%
FrontierMath (Feb 2025 set)—4.5%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Claude Opus 4 leads

Amazon Nova Pro: 27.4 (#250), Claude Opus 4: 44.0 (#88)

Knowledge benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
Humanity's Last Exam4.4%10.7%
MMLU-Pro67.3%87.5%
Confabulations30.1%15.9%
Vectara Hallucination Rate5.1%12%
GPQA (HELM)44.6%70.8%
LMArena Expert12111386
GPQA Diamond—76.3%
MMLU82%—

Multimodal Claude Opus 4 leads

Amazon Nova Pro: 25.0 (#126), Claude Opus 4: 31.5 (#106)

Multimodal benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
LMArena Vision9801192
GeoBench—49%
VPCT—38%

Multilingual Claude Opus 4 leads

Amazon Nova Pro: 39.7 (#223), Claude Opus 4: 48.8 (#138)

Multilingual benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
LMArena Non-English12341362
LMArena Chinese12441386
LMArena French12711372
LMArena German12431391
LMArena Japanese12001331
LMArena Korean12031321
LMArena Russian12401392
LMArena Spanish11821389

Instruction Following Claude Opus 4 leads

Amazon Nova Pro: 64.9 (#226), Claude Opus 4: 77.1 (#28)

Instruction Following benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
IFEval81.5%91.8%
LMArena Instruction Following12351406
LiveBench Instruction Following67.1%—

Long Context Claude Opus 4 leads

Amazon Nova Pro: 38.1 (#205), Claude Opus 4: 39.6 (#172)

Long Context benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
LMArena Longer Query12551422
Fiction.LiveBench—61.1%

Writing & Preference Claude Opus 4 leads

Amazon Nova Pro: 43.9 (#226), Claude Opus 4: 61.2 (#89)

Writing & Preference benchmarks
BenchmarkAmazon Nova ProClaude Opus 4
LMArena Text12591377
LMArena Creative Writing12121387
Short-Story Creative Writing60.5%83.6%
WildBench77.7%85.2%
LMArena Multi-Turn12461396
EQ-Bench Creative Writing—1580
LiveBench Language37%—

Frequently asked questions

Is Amazon Nova Pro better than Claude Opus 4?

Claude Opus 4 is the stronger model overall, scoring 43.1 to 31.0 on the Noometry Index. Amazon Nova Pro costs 21× less per token, which makes it the better buy when Claude Opus 4's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Pro or Claude Opus 4?

Amazon Nova Pro is cheaper. It lists at $0.80 per million input tokens and $3.20 per million output tokens; Claude Opus 4 lists at $15 and $75.

Is Amazon Nova Pro or Claude Opus 4 better for coding?

Claude Opus 4 scores higher on coding benchmarks: 47.2 versus 35.1 in the Noometry coding category.

Which has the bigger context window?

Amazon Nova Pro does, with 300K tokens against 200K.

How many benchmarks do Amazon Nova Pro and Claude Opus 4 share?

28 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and Claude Opus 4 has 56.

Related comparisons

Go deeper