Model comparison

Amazon Nova Pro vs Qwen3.8 Max

Qwen3.8 Max is the stronger model overall, scoring 56.8 to 31.0 on the Noometry Index. Amazon Nova Pro costs 2.1× less per token, which makes it the better buy when Qwen3.8 Max's lead doesn't matter for your workload.

Last verified . 19 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

Qwen3.8 Max Alibaba (Qwen)

56.8

Rank #22 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Amazon Nova Pro scores higher in 0 categories and Qwen3.8 Max in 10 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in math, where Qwen3.8 Max leads 73.2 to 28.5.
  • Amazon Nova Pro is cheaper at $0.80 / $3.20 per million input/output tokens, against $2 / $6 for Qwen3.8 Max.
  • Qwen3.8 Max accepts more context: 1M tokens versus 300K.

Side by side

Amazon Nova Pro and Qwen3.8 Max specifications
Amazon Nova ProQwen3.8 Max
ProviderAmazonAlibaba (Qwen)
Noometry Index31.056.8
Released2024-12-032026-08-02
WeightsProprietaryProprietary
Context window300K1M
Max output10K131K
Input $ / M tokens$0.80$2
Output $ / M tokens$3.20$6
Results tracked3839

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.8 Max leads

Amazon Nova Pro: 35.1 (#229), Qwen3.8 Max: 53.5 (#29)

Coding benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Coding12701502
DeepSWE—57.5%
LMArena WebDev—1674
FrontierSWE—17.8%
SciCode—53.2%
LiveBench Coding38.1%—

Agentic & Tool Use Qwen3.8 Max leads

Amazon Nova Pro: 16.7 (#147), Qwen3.8 Max: 45.4 (#14)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
APEX-Agents—63.3%
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—
τ²-bench Banking—55.1%
GDP.pdf—23.2%

Reasoning Qwen3.8 Max leads

Amazon Nova Pro: 20.0 (#243), Qwen3.8 Max: 54.4 (#26)

Reasoning benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Hard Prompts12461496
Epoch Capabilities Index123.8156.41
NYT Connections (extended)—88.3%
CritPt—20%
Chess Puzzles—40%
LiveBench Reasoning32.6%—
Mystery Game Puzzles—38%
DTBench—92%
LiveBench Data Analysis48.3%—
LMCA—46.2%
LiveBench43.5%—

Math Qwen3.8 Max leads

Amazon Nova Pro: 28.5 (#243), Qwen3.8 Max: 73.2 (#20)

Math benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Math12521499
FrontierMath (Tiers 1-3)—74.7%
FrontierMath Tier 4—46.3%
OTIS Mock AIME 2024-2025—100%
ProofBench—58%
Omni-MATH24.2%—
LiveBench Math38%—

Knowledge Qwen3.8 Max leads

Amazon Nova Pro: 27.4 (#250), Qwen3.8 Max: 61.7 (#27)

Knowledge benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Expert12111507
GPQA Diamond—92.7%
Humanity's Last Exam4.4%—
SimpleQA Verified—47.3%
MMLU-Pro67.3%—
Confabulations30.1%—
Vectara Hallucination Rate5.1%—
GPQA (HELM)44.6%—
MMLU82%—

Multimodal Qwen3.8 Max leads

Amazon Nova Pro: 25.0 (#126), Qwen3.8 Max: 37.2 (#75)

Multimodal benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Vision9801314
Furniture Assembly—20%

Multilingual Qwen3.8 Max leads

Amazon Nova Pro: 39.7 (#223), Qwen3.8 Max: 56.7 (#18)

Multilingual benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Non-English12341472
LMArena Chinese12441538
LMArena French12711503
LMArena German12431483
LMArena Japanese12001467
LMArena Korean12031461
LMArena Russian12401481
LMArena Spanish11821492

Instruction Following Qwen3.8 Max leads

Amazon Nova Pro: 64.9 (#226), Qwen3.8 Max: 77.6 (#17)

Instruction Following benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Instruction Following12351479
LiveBench Instruction Following67.1%—
IFEval81.5%—

Long Context Qwen3.8 Max leads

Amazon Nova Pro: 38.1 (#205), Qwen3.8 Max: 45.6 (#31)

Long Context benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Longer Query12551489

Writing & Preference Qwen3.8 Max leads

Amazon Nova Pro: 43.9 (#226), Qwen3.8 Max: 67.1 (#30)

Writing & Preference benchmarks
BenchmarkAmazon Nova ProQwen3.8 Max
LMArena Text12591483
LMArena Creative Writing12121479
LMArena Multi-Turn12461489
Short-Story Creative Writing60.5%—
WildBench77.7%—
LiveBench Language37%—

Frequently asked questions

Is Amazon Nova Pro better than Qwen3.8 Max?

Qwen3.8 Max is the stronger model overall, scoring 56.8 to 31.0 on the Noometry Index. Amazon Nova Pro costs 2.1× less per token, which makes it the better buy when Qwen3.8 Max's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Pro or Qwen3.8 Max?

Amazon Nova Pro is cheaper. It lists at $0.80 per million input tokens and $3.20 per million output tokens; Qwen3.8 Max lists at $2 and $6.

Is Amazon Nova Pro or Qwen3.8 Max better for coding?

Qwen3.8 Max scores higher on coding benchmarks: 53.5 versus 35.1 in the Noometry coding category.

Which has the bigger context window?

Qwen3.8 Max does, with 1M tokens against 300K.

How many benchmarks do Amazon Nova Pro and Qwen3.8 Max share?

19 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and Qwen3.8 Max has 39.

Related comparisons

Go deeper