Model comparison

Claude 3.7 Sonnet vs Nova 2.0 Pro Preview

Claude 3.7 Sonnet is the stronger model overall, scoring 39.5 to 33.4 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 3.7 Sonnet Anthropic

39.5

Rank #164 Confirmed

Nova 2.0 Pro Preview Amazon

33.4

Rank #244 Reported

Summary

  • The widest gap is in agentic & tool use, where Claude 3.7 Sonnet leads 34.1 to 21.9.

Side by side

Claude 3.7 Sonnet and Nova 2.0 Pro Preview specifications
Claude 3.7 SonnetNova 2.0 Pro Preview
ProviderAnthropicAmazon
Noometry Index39.533.4
Released2025-02-242025-12-02
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked583

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Claude 3.7 Sonnet: 40.6 (#136), Nova 2.0 Pro Preview: 40.8 (#133)

Coding benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
SWE-bench Verified61%—
SWE-bench Verified (bash only)52.8%—
Aider Polyglot64.9%—
SciCode—42.7%
GSO3.8%—
LiveBench Coding74.5%—
LMArena Coding1361—
CadEval54%—

Agentic & Tool Use Claude 3.7 Sonnet leads

Claude 3.7 Sonnet: 34.1 (#50), Nova 2.0 Pro Preview: 21.9 (#136)

Agentic & Tool Use benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
TheAgentCompany30.9%—
Cybench20%—
DeepResearch Bench43.6%—
OSWorld35.8%—
GDP.pdf—2%
METR Time Horizons60%—

Reasoning Nova 2.0 Pro Preview leads

Claude 3.7 Sonnet: 18.6 (#277), Nova 2.0 Pro Preview: 22.4 (#194)

Reasoning benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
ARC-AGI-20.9%—
SimpleBench46.4%—
ARC-AGI-128.6%—
CritPt—0%
EnigmaEval4.2%—
LiveBench Reasoning87.8%—
LMArena Hard Prompts1333—
LiveBench Data Analysis74%—
Epoch Capabilities Index141.16—
ForecastBench61.8—
LiveBench76.1%—

Math Not comparable

Claude 3.7 Sonnet: 37.5 (#153), Nova 2.0 Pro Preview: —

Math benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
OTIS Mock AIME 2024-202557.8%—
Omni-MATH33%—
LiveBench Math79%—
LMArena Math1337—
MATH Level 591.2%—
FrontierMath (Feb 2025 set)4.1%—

Knowledge Not comparable

Claude 3.7 Sonnet: 39.8 (#130), Nova 2.0 Pro Preview: —

Knowledge benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
GPQA Diamond79.7%—
Humanity's Last Exam8%—
MMLU-Pro78.4%—
Confabulations14.7%—
GPQA (HELM)60.8%—
LMArena Expert1321—

Multimodal Not comparable

Claude 3.7 Sonnet: 33.7 (#95), Nova 2.0 Pro Preview: —

Multimodal benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
LMArena Vision1169—
GeoBench68%—
VPCT39%—
SpatialViz-Bench33.9%—

Multilingual Not comparable

Claude 3.7 Sonnet: 44.1 (#179), Nova 2.0 Pro Preview: —

Multilingual benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
LMArena Non-English1296—
LMArena Chinese1299—
LMArena French1303—
LMArena German1301—
LMArena Japanese1267—
LMArena Korean1249—
LMArena Russian1311—
LMArena Spanish1298—

Instruction Following Not comparable

Claude 3.7 Sonnet: 72.9 (#125), Nova 2.0 Pro Preview: —

Instruction Following benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
LiveBench Instruction Following81.3%—
IFEval83.4%—
LMArena Instruction Following1352—

Long Context Not comparable

Claude 3.7 Sonnet: 50.3 (#10), Nova 2.0 Pro Preview: —

Long Context benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
Fiction.LiveBench83.3%—
LMArena Longer Query1373—

Writing & Preference Not comparable

Claude 3.7 Sonnet: 54.4 (#150), Nova 2.0 Pro Preview: —

Writing & Preference benchmarks
BenchmarkClaude 3.7 SonnetNova 2.0 Pro Preview
LMArena Text1314—
LMArena Creative Writing1332—
Short-Story Creative Writing81.1%—
EQ-Bench Creative Writing1412—
WildBench81.4%—
LMArena Multi-Turn1339—
LiveBench Language59.9%—

Frequently asked questions

Is Claude 3.7 Sonnet better than Nova 2.0 Pro Preview?

Claude 3.7 Sonnet is the stronger model overall, scoring 39.5 to 33.4 on the Noometry Index.

Is Claude 3.7 Sonnet or Nova 2.0 Pro Preview better for coding?

They score almost the same on coding (40.6 vs 40.8); test both on your own repository before choosing.

How many benchmarks do Claude 3.7 Sonnet and Nova 2.0 Pro Preview share?

0 benchmarks have published results for both models. Claude 3.7 Sonnet has 58 scored results on Noometry and Nova 2.0 Pro Preview has 3.

Related comparisons

Go deeper