Model comparison

Amazon Nova Pro vs Qwen3.5-Flash

Qwen3.5-Flash is the stronger model overall, scoring 42.5 to 31.0 on the Noometry Index.

Last verified . 19 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

Qwen3.5-Flash Alibaba (Qwen)

42.5

Rank #112 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Amazon Nova Pro scores higher in 1 category and Qwen3.5-Flash in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3.5-Flash leads 43.2 to 27.4.
  • The biggest single-benchmark swing is Vectara Hallucination Rate: 5.1% for Amazon Nova Pro and 10.5% for Qwen3.5-Flash.
  • Qwen3.5-Flash is cheaper at $0.10 / $0.40 per million input/output tokens, against $0.80 / $3.20 for Amazon Nova Pro.
  • Qwen3.5-Flash accepts more context: 1M tokens versus 300K.

Side by side

Amazon Nova Pro and Qwen3.5-Flash specifications
Amazon Nova ProQwen3.5-Flash
ProviderAmazonAlibaba (Qwen)
Noometry Index31.042.5
Released2024-12-032026-02-23
WeightsProprietaryProprietary
Context window300K1M
Max output10K66K
Input $ / M tokens$0.80$0.10
Output $ / M tokens$3.20$0.40
Results tracked3832

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Amazon Nova Pro: 35.1 (#229), Qwen3.5-Flash: 34.2 (#242)

Coding benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Coding12701412
LMArena WebDev—1244
LiveBench Coding38.1%—
ALE-Bench—221.8

Agentic & Tool Use Not comparable

Amazon Nova Pro: 16.7 (#147), Qwen3.5-Flash: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—
Vending-Bench 2—462.69

Reasoning Qwen3.5-Flash leads

Amazon Nova Pro: 20.0 (#243), Qwen3.5-Flash: 33.7 (#72)

Reasoning benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Hard Prompts12461403
Epoch Capabilities Index123.8143.98
Chess Puzzles—21%
LiveBench Reasoning32.6%—
Mystery Game Puzzles—20%
DTBench—82.9%
LiveBench Data Analysis48.3%—
LMCA—29.1%
LiveBench43.5%—

Math Qwen3.5-Flash leads

Amazon Nova Pro: 28.5 (#243), Qwen3.5-Flash: 37.4 (#158)

Math benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Math12521407
FrontierMath (Tiers 1-3)—18.2%
OTIS Mock AIME 2024-2025—84.4%
Omni-MATH24.2%—
LiveBench Math38%—
FrontierMath (Feb 2025 set)—6.2%
FrontierMath Tier 4 (v1)—0%

Knowledge Qwen3.5-Flash leads

Amazon Nova Pro: 27.4 (#250), Qwen3.5-Flash: 43.2 (#93)

Knowledge benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
Vectara Hallucination Rate5.1%10.5%
LMArena Expert12111407
GPQA Diamond—82.3%
Humanity's Last Exam4.4%—
SimpleQA Verified—20.3%
MMLU-Pro67.3%—
Confabulations30.1%—
GPQA (HELM)44.6%—
MMLU82%—

Multimodal Not comparable

Amazon Nova Pro: 25.0 (#126), Qwen3.5-Flash: —

Multimodal benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Vision980—

Multilingual Qwen3.5-Flash leads

Amazon Nova Pro: 39.7 (#223), Qwen3.5-Flash: 50.5 (#121)

Multilingual benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Non-English12341385
LMArena Chinese12441446
LMArena French12711412
LMArena German12431390
LMArena Japanese12001368
LMArena Korean12031344
LMArena Russian12401379
LMArena Spanish11821400

Instruction Following Qwen3.5-Flash leads

Amazon Nova Pro: 64.9 (#226), Qwen3.5-Flash: 72.6 (#139)

Instruction Following benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Instruction Following12351374
LiveBench Instruction Following67.1%—
IFEval81.5%—

Long Context Qwen3.5-Flash leads

Amazon Nova Pro: 38.1 (#205), Qwen3.5-Flash: 42.4 (#124)

Long Context benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Longer Query12551392

Writing & Preference Qwen3.5-Flash leads

Amazon Nova Pro: 43.9 (#226), Qwen3.5-Flash: 57.9 (#122)

Writing & Preference benchmarks
BenchmarkAmazon Nova ProQwen3.5-Flash
LMArena Text12591397
LMArena Creative Writing12121343
LMArena Multi-Turn12461393
Short-Story Creative Writing60.5%—
WildBench77.7%—
LiveBench Language37%—

Frequently asked questions

Is Amazon Nova Pro better than Qwen3.5-Flash?

Qwen3.5-Flash is the stronger model overall, scoring 42.5 to 31.0 on the Noometry Index.

Which is cheaper, Amazon Nova Pro or Qwen3.5-Flash?

Qwen3.5-Flash is cheaper. It lists at $0.10 per million input tokens and $0.40 per million output tokens; Amazon Nova Pro lists at $0.80 and $3.20.

Is Amazon Nova Pro or Qwen3.5-Flash better for coding?

They score almost the same on coding (35.1 vs 34.2); test both on your own repository before choosing.

Which has the bigger context window?

Qwen3.5-Flash does, with 1M tokens against 300K.

How many benchmarks do Amazon Nova Pro and Qwen3.5-Flash share?

19 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and Qwen3.5-Flash has 32.

Related comparisons

Go deeper