Model comparison

Nova 2.0 Pro Preview vs Qwen3.6 Plus

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 33.4 on the Noometry Index.

Last verified . 2 shared benchmarks.

Nova 2.0 Pro Preview Amazon

33.4

Rank #244 Reported

Qwen3.6 Plus Alibaba (Qwen)

47.5

Rank #62 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Nova 2.0 Pro Preview scores higher in 0 categories and Qwen3.6 Plus in 2 categories; one gap is clear of the uncertainty.
  • The widest gap is in reasoning, where Qwen3.6 Plus leads 29.3 to 22.4.

Side by side

Nova 2.0 Pro Preview and Qwen3.6 Plus specifications
Nova 2.0 Pro PreviewQwen3.6 Plus
ProviderAmazonAlibaba (Qwen)
Noometry Index33.447.5
Released2025-12-022026-03-31
WeightsProprietaryProprietary
Context window—1M
Max output—66K
Input $ / M tokens—$0.50
Output $ / M tokens—$3
Results tracked337

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Nova 2.0 Pro Preview: 40.8 (#133), Qwen3.6 Plus: 40.8 (#130)

Coding benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
SciCode42.7%40.7%
SWE-bench Verified—57.9%
LMArena WebDev—1461
LMArena Coding—1467
ALE-Bench—670.15

Agentic & Tool Use Not comparable

Nova 2.0 Pro Preview: 21.9 (#136), Qwen3.6 Plus: —

Agentic & Tool Use benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
GDP.pdf2%—
Vending-Bench 2—5,115

Reasoning Qwen3.6 Plus leads

Nova 2.0 Pro Preview: 22.4 (#194), Qwen3.6 Plus: 29.3 (#93)

Reasoning benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
CritPt0%2.9%
NYT Connections (extended)—60.3%
Chess Puzzles—17%
Thematic Generalization—59.5%
LMArena Hard Prompts—1449
Mystery Game Puzzles—12%
DTBench—81.9%
LMCA—33.1%
Epoch Capabilities Index—147.65

Math Not comparable

Nova 2.0 Pro Preview: —, Qwen3.6 Plus: 51.8 (#54)

Math benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
FrontierMath (Tiers 1-3)—38.2%
OTIS Mock AIME 2024-2025—93.3%
LMArena Math—1450
FrontierMath (Feb 2025 set)—26.2%
FrontierMath Tier 4 (v1)—8.3%

Knowledge Not comparable

Nova 2.0 Pro Preview: —, Qwen3.6 Plus: 56.1 (#45)

Knowledge benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
GPQA Diamond—88.4%
SimpleQA Verified—44.1%
LMArena Expert—1454

Multilingual Not comparable

Nova 2.0 Pro Preview: —, Qwen3.6 Plus: 53.3 (#70)

Multilingual benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
LMArena Non-English—1424
LMArena Chinese—1477
LMArena French—1455
LMArena German—1452
LMArena Japanese—1389
LMArena Korean—1379
LMArena Russian—1434
LMArena Spanish—1432

Instruction Following Not comparable

Nova 2.0 Pro Preview: —, Qwen3.6 Plus: 75.0 (#74)

Instruction Following benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
LMArena Instruction Following—1425

Long Context Not comparable

Nova 2.0 Pro Preview: —, Qwen3.6 Plus: 45.2 (#49)

Long Context benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
CL-bench—20.3%
LMArena Longer Query—1439

Writing & Preference Not comparable

Nova 2.0 Pro Preview: —, Qwen3.6 Plus: 62.2 (#82)

Writing & Preference benchmarks
BenchmarkNova 2.0 Pro PreviewQwen3.6 Plus
LMArena Text—1437
LMArena Creative Writing—1404
LMArena Multi-Turn—1438

Frequently asked questions

Is Nova 2.0 Pro Preview better than Qwen3.6 Plus?

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 33.4 on the Noometry Index.

Is Nova 2.0 Pro Preview or Qwen3.6 Plus better for coding?

They score almost the same on coding (40.8 vs 40.8); test both on your own repository before choosing.

How many benchmarks do Nova 2.0 Pro Preview and Qwen3.6 Plus share?

2 benchmarks have published results for both models. Nova 2.0 Pro Preview has 3 scored results on Noometry and Qwen3.6 Plus has 37.

Related comparisons

Go deeper