Model comparison

Amazon Nova Experimental Chat 10 09 vs Qwen3.5 27B

Amazon Nova Experimental Chat 10 09 and Qwen3.5 27B score almost the same on the Noometry Index (41.9 vs 41.9), so choose on price, context window or the category you care about most.

Last verified . 9 shared benchmarks.

Qwen3.5 27B Alibaba (Qwen)

41.9

Rank #127 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Amazon Nova Experimental Chat 10 09 scores higher in 1 category and Qwen3.5 27B in 5 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen3.5 27B leads 59.3 to 54.7.
  • Qwen3.5 27B has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Experimental Chat 10 09 and Qwen3.5 27B specifications
Amazon Nova Experimental Chat 10 09Qwen3.5 27B
ProviderAmazonAlibaba (Qwen)
Noometry Index41.941.9
Released—2026-02-23
WeightsProprietaryOpen
Context window—262K
Max output—66K
Input $ / M tokens—$0.30
Output $ / M tokens—$2.40
Results tracked928

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Amazon Nova Experimental Chat 10 09 leads

Amazon Nova Experimental Chat 10 09: 40.2 (#145), Qwen3.5 27B: 38.9 (#168)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Coding13701427
LMArena WebDev—1358
WeirdML—39.5%
ALE-Bench—349.45

Agentic & Tool Use Not comparable

Amazon Nova Experimental Chat 10 09: —, Qwen3.5 27B: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
Vending-Bench 2—201.98

Reasoning Too close to call

Amazon Nova Experimental Chat 10 09: 27.3 (#120), Qwen3.5 27B: 27.5 (#117)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Hard Prompts13561414
NYT Connections (extended)—47.9%
Thematic Generalization—45.5%
DTBench—82.4%
LMCA—34%

Math Not comparable

Amazon Nova Experimental Chat 10 09: —, Qwen3.5 27B: 38.8 (#127)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
MathArena Final-Answer Competitions—56.7%
LMArena Math—1429

Knowledge Not comparable

Amazon Nova Experimental Chat 10 09: —, Qwen3.5 27B: 38.0 (#150)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
Vectara Hallucination Rate—12.1%
LMArena Expert—1428

Multimodal Not comparable

Amazon Nova Experimental Chat 10 09: —, Qwen3.5 27B: 39.4 (#59)

Multimodal benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Vision—1241

Multilingual Qwen3.5 27B leads

Amazon Nova Experimental Chat 10 09: 47.3 (#150), Qwen3.5 27B: 50.8 (#115)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Non-English13401390
LMArena Chinese13411478
LMArena French—1410
LMArena German—1393
LMArena Japanese—1345
LMArena Korean—1358
LMArena Russian—1390
LMArena Spanish—1407

Instruction Following Qwen3.5 27B leads

Amazon Nova Experimental Chat 10 09: 69.4 (#172), Qwen3.5 27B: 73.5 (#119)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Instruction Following13151393

Long Context Qwen3.5 27B leads

Amazon Nova Experimental Chat 10 09: 40.5 (#152), Qwen3.5 27B: 43.1 (#106)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Longer Query13331413

Writing & Preference Qwen3.5 27B leads

Amazon Nova Experimental Chat 10 09: 54.7 (#149), Qwen3.5 27B: 59.3 (#111)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Qwen3.5 27B
LMArena Text13641409
LMArena Creative Writing13071362
LMArena Multi-Turn13551410

Frequently asked questions

Is Amazon Nova Experimental Chat 10 09 better than Qwen3.5 27B?

Amazon Nova Experimental Chat 10 09 and Qwen3.5 27B score almost the same on the Noometry Index (41.9 vs 41.9), so choose on price, context window or the category you care about most.

Is Amazon Nova Experimental Chat 10 09 or Qwen3.5 27B better for coding?

Amazon Nova Experimental Chat 10 09 scores higher on coding benchmarks: 40.2 versus 38.9 in the Noometry coding category.

How many benchmarks do Amazon Nova Experimental Chat 10 09 and Qwen3.5 27B share?

9 benchmarks have published results for both models. Amazon Nova Experimental Chat 10 09 has 9 scored results on Noometry and Qwen3.5 27B has 28.

Related comparisons

Go deeper