Model comparison

Amazon Nova Experimental Chat 10 20 vs Qwen3-Next 80B-A3B Instruct

Amazon Nova Experimental Chat 10 20 and Qwen3-Next 80B-A3B Instruct score almost the same on the Noometry Index (42.1 vs 43.0), so choose on price, context window or the category you care about most.

Last verified . 17 shared benchmarks.

Summary

  • They share 17 benchmarks with published results for both. Amazon Nova Experimental Chat 10 20 scores higher in 3 categories and Qwen3-Next 80B-A3B Instruct in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Amazon Nova Experimental Chat 10 20 leads 41.7 to 37.0.
  • Qwen3-Next 80B-A3B Instruct has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Experimental Chat 10 20 and Qwen3-Next 80B-A3B Instruct specifications
Amazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
ProviderAmazonAlibaba (Qwen)
Noometry Index42.143.0
Released—2025-09
WeightsProprietaryOpen
Context window—131K
Max output—33K
Input $ / M tokens—$0.50
Output $ / M tokens—$2
Results tracked1725

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Amazon Nova Experimental Chat 10 20: 41.6 (#123), Qwen3-Next 80B-A3B Instruct: 42.5 (#98)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Coding14121440

Reasoning Qwen3-Next 80B-A3B Instruct leads

Amazon Nova Experimental Chat 10 20: 28.4 (#106), Qwen3-Next 80B-A3B Instruct: 31.1 (#81)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Hard Prompts13961428
Kagi LLM Benchmark—66.7%

Math Too close to call

Amazon Nova Experimental Chat 10 20: 39.0 (#119), Qwen3-Next 80B-A3B Instruct: 38.8 (#126)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Math14251440
Omni-MATH—46.7%

Knowledge Qwen3-Next 80B-A3B Instruct leads

Amazon Nova Experimental Chat 10 20: 38.5 (#144), Qwen3-Next 80B-A3B Instruct: 41.8 (#106)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Expert13861417
MMLU-Pro—78.6%
Vectara Hallucination Rate—9.3%
GPQA (HELM)—63%

Multilingual Qwen3-Next 80B-A3B Instruct leads

Amazon Nova Experimental Chat 10 20: 49.5 (#131), Qwen3-Next 80B-A3B Instruct: 52.1 (#93)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Non-English13721407
LMArena Chinese14041460
LMArena French14251413
LMArena German13911417
LMArena Japanese13601395
LMArena Korean13241364
LMArena Russian13711404
LMArena Spanish13821435

Instruction Following Amazon Nova Experimental Chat 10 20 leads

Amazon Nova Experimental Chat 10 20: 72.1 (#141), Qwen3-Next 80B-A3B Instruct: 70.8 (#159)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Instruction Following13651389
IFEval—81%

Long Context Amazon Nova Experimental Chat 10 20 leads

Amazon Nova Experimental Chat 10 20: 41.7 (#133), Qwen3-Next 80B-A3B Instruct: 37.0 (#223)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Longer Query13701403
Fiction.LiveBench—55.6%

Writing & Preference Qwen3-Next 80B-A3B Instruct leads

Amazon Nova Experimental Chat 10 20: 56.6 (#136), Qwen3-Next 80B-A3B Instruct: 58.0 (#121)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20Qwen3-Next 80B-A3B Instruct
LMArena Text13941417
LMArena Creative Writing13181334
LMArena Multi-Turn13641416
WildBench—80.7%

Frequently asked questions

Is Amazon Nova Experimental Chat 10 20 better than Qwen3-Next 80B-A3B Instruct?

Amazon Nova Experimental Chat 10 20 and Qwen3-Next 80B-A3B Instruct score almost the same on the Noometry Index (42.1 vs 43.0), so choose on price, context window or the category you care about most.

Is Amazon Nova Experimental Chat 10 20 or Qwen3-Next 80B-A3B Instruct better for coding?

They score almost the same on coding (41.6 vs 42.5); test both on your own repository before choosing.

How many benchmarks do Amazon Nova Experimental Chat 10 20 and Qwen3-Next 80B-A3B Instruct share?

17 benchmarks have published results for both models. Amazon Nova Experimental Chat 10 20 has 17 scored results on Noometry and Qwen3-Next 80B-A3B Instruct has 25.

Related comparisons

Go deeper