Model comparison

Amazon Nova Pro vs DeepSeek-R1-Distill-Qwen-32B

DeepSeek-R1-Distill-Qwen-32B is the stronger model overall, scoring 35.5 to 31.0 on the Noometry Index.

Last verified . 8 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

DeepSeek-R1-Distill-Qwen-32B DeepSeek

35.5

Rank #226 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Amazon Nova Pro scores higher in 2 categories and DeepSeek-R1-Distill-Qwen-32B in 5 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in agentic & tool use, where DeepSeek-R1-Distill-Qwen-32B leads 28.1 to 16.7.
  • The biggest single-benchmark swing is LiveBench Math: 38% for Amazon Nova Pro and 59.4% for DeepSeek-R1-Distill-Qwen-32B.
  • DeepSeek-R1-Distill-Qwen-32B has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Pro and DeepSeek-R1-Distill-Qwen-32B specifications
Amazon Nova ProDeepSeek-R1-Distill-Qwen-32B
ProviderAmazonDeepSeek
Noometry Index31.035.5
Released2024-12-032025-01-20
WeightsProprietaryOpen
Context window300K—
Max output10K—
Input $ / M tokens$0.80—
Output $ / M tokens$3.20—
Results tracked3814

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek-R1-Distill-Qwen-32B leads

Amazon Nova Pro: 35.1 (#229), DeepSeek-R1-Distill-Qwen-32B: 36.1 (#212)

Coding benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LiveBench Coding38.1%33.7%
BigCodeBench Instruct—43.9%
LMArena Coding1270—
BigCodeBench Complete—54.9%

Agentic & Tool Use DeepSeek-R1-Distill-Qwen-32B leads

Amazon Nova Pro: 16.7 (#147), DeepSeek-R1-Distill-Qwen-32B: 28.1 (#94)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—
BALROG—19.5%

Reasoning Amazon Nova Pro leads

Amazon Nova Pro: 20.0 (#243), DeepSeek-R1-Distill-Qwen-32B: 18.2 (#284)

Reasoning benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LiveBench Reasoning32.6%52.3%
LiveBench Data Analysis48.3%45.4%
Epoch Capabilities Index123.8137.44
LiveBench43.5%45.5%
Chess Puzzles—1%
LMArena Hard Prompts1246—

Math DeepSeek-R1-Distill-Qwen-32B leads

Amazon Nova Pro: 28.5 (#243), DeepSeek-R1-Distill-Qwen-32B: 34.5 (#194)

Math benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LiveBench Math38%59.4%
OTIS Mock AIME 2024-2025—55.6%
Omni-MATH24.2%—
LMArena Math1252—

Knowledge DeepSeek-R1-Distill-Qwen-32B leads

Amazon Nova Pro: 27.4 (#250), DeepSeek-R1-Distill-Qwen-32B: 35.7 (#182)

Knowledge benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
GPQA Diamond—64.1%
Humanity's Last Exam4.4%—
MMLU-Pro67.3%—
Confabulations30.1%—
Vectara Hallucination Rate5.1%—
GPQA (HELM)44.6%—
LMArena Expert1211—
MMLU82%—

Multimodal Not comparable

Amazon Nova Pro: 25.0 (#126), DeepSeek-R1-Distill-Qwen-32B: —

Multimodal benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LMArena Vision980—

Multilingual Not comparable

Amazon Nova Pro: 39.7 (#223), DeepSeek-R1-Distill-Qwen-32B: —

Multilingual benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LMArena Non-English1234—
LMArena Chinese1244—
LMArena French1271—
LMArena German1243—
LMArena Japanese1200—
LMArena Korean1203—
LMArena Russian1240—
LMArena Spanish1182—

Instruction Following Amazon Nova Pro leads

Amazon Nova Pro: 64.9 (#226), DeepSeek-R1-Distill-Qwen-32B: 61.6 (#243)

Instruction Following benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LiveBench Instruction Following67.1%55.7%
IFEval81.5%—
LMArena Instruction Following1235—

Long Context Not comparable

Amazon Nova Pro: 38.1 (#205), DeepSeek-R1-Distill-Qwen-32B: —

Long Context benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LMArena Longer Query1255—

Writing & Preference DeepSeek-R1-Distill-Qwen-32B leads

Amazon Nova Pro: 43.9 (#226), DeepSeek-R1-Distill-Qwen-32B: 49.6 (#188)

Writing & Preference benchmarks
BenchmarkAmazon Nova ProDeepSeek-R1-Distill-Qwen-32B
LiveBench Language37%26.8%
LMArena Text1259—
LMArena Creative Writing1212—
Short-Story Creative Writing60.5%—
WildBench77.7%—
LMArena Multi-Turn1246—

Frequently asked questions

Is Amazon Nova Pro better than DeepSeek-R1-Distill-Qwen-32B?

DeepSeek-R1-Distill-Qwen-32B is the stronger model overall, scoring 35.5 to 31.0 on the Noometry Index.

Is Amazon Nova Pro or DeepSeek-R1-Distill-Qwen-32B better for coding?

DeepSeek-R1-Distill-Qwen-32B scores higher on coding benchmarks: 36.1 versus 35.1 in the Noometry coding category.

How many benchmarks do Amazon Nova Pro and DeepSeek-R1-Distill-Qwen-32B share?

8 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and DeepSeek-R1-Distill-Qwen-32B has 14.

Related comparisons

Go deeper