Model comparison

Amazon Nova Experimental Chat 11 10 vs GLM-4.6

Amazon Nova Experimental Chat 11 10 is the stronger model overall, scoring 43.0 to 41.4 on the Noometry Index.

Last verified . 17 shared benchmarks.

GLM-4.6 Z.ai (Zhipu)

41.4

Rank #135 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Amazon Nova Experimental Chat 11 10 scores higher in 2 categories and GLM-4.6 in 6 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Amazon Nova Experimental Chat 11 10 leads 29.1 to 23.7.
  • GLM-4.6 has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Experimental Chat 11 10 and GLM-4.6 specifications
Amazon Nova Experimental Chat 11 10GLM-4.6
ProviderAmazonZ.ai (Zhipu)
Noometry Index43.041.4
Released—2025-09-30
WeightsProprietaryOpen
Context window—205K
Max output—131K
Input $ / M tokens—$0.60
Output $ / M tokens—$2.20
Results tracked1729

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Amazon Nova Experimental Chat 11 10 leads

Amazon Nova Experimental Chat 11 10: 42.3 (#107), GLM-4.6: 40.1 (#148)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Coding14331449
SWE-bench Verified (bash only)—55.4%
LMArena WebDev—1340
SciCode—38.4%
ALE-Bench—340.82

Agentic & Tool Use Not comparable

Amazon Nova Experimental Chat 11 10: —, GLM-4.6: 32.3 (#66)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
Terminal-Bench—24.5%
Berkeley Function Calling Leaderboard—72.4%

Reasoning Amazon Nova Experimental Chat 11 10 leads

Amazon Nova Experimental Chat 11 10: 29.1 (#95), GLM-4.6: 23.7 (#172)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Hard Prompts14221440
Kagi LLM Benchmark—47.4%
CritPt—1.1%

Math Too close to call

Amazon Nova Experimental Chat 11 10: 39.1 (#114), GLM-4.6: 39.1 (#111)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Math14311432
FrontierMath (Feb 2025 set)—3.8%
FrontierMath Tier 4 (v1)—2.1%

Knowledge Too close to call

Amazon Nova Experimental Chat 11 10: 39.9 (#127), GLM-4.6: 40.2 (#124)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Expert14311431
Vectara Hallucination Rate—9.5%

Multilingual GLM-4.6 leads

Amazon Nova Experimental Chat 11 10: 51.9 (#98), GLM-4.6: 53.5 (#66)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Non-English14051426
LMArena Chinese14631499
LMArena French14491459
LMArena German13951447
LMArena Japanese13831393
LMArena Korean13841400
LMArena Russian13941419
LMArena Spanish14441436

Instruction Following GLM-4.6 leads

Amazon Nova Experimental Chat 11 10: 73.2 (#122), GLM-4.6: 74.3 (#98)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Instruction Following13871410

Long Context Too close to call

Amazon Nova Experimental Chat 11 10: 42.5 (#121), GLM-4.6: 43.4 (#94)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Longer Query13951422

Writing & Preference GLM-4.6 leads

Amazon Nova Experimental Chat 11 10: 58.8 (#114), GLM-4.6: 61.1 (#90)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10GLM-4.6
LMArena Text14151440
LMArena Creative Writing13491411
LMArena Multi-Turn13801427
EQ-Bench Creative Writing—1411

Frequently asked questions

Is Amazon Nova Experimental Chat 11 10 better than GLM-4.6?

Amazon Nova Experimental Chat 11 10 is the stronger model overall, scoring 43.0 to 41.4 on the Noometry Index.

Is Amazon Nova Experimental Chat 11 10 or GLM-4.6 better for coding?

Amazon Nova Experimental Chat 11 10 scores higher on coding benchmarks: 42.3 versus 40.1 in the Noometry coding category.

How many benchmarks do Amazon Nova Experimental Chat 11 10 and GLM-4.6 share?

17 benchmarks have published results for both models. Amazon Nova Experimental Chat 11 10 has 17 scored results on Noometry and GLM-4.6 has 29.

Related comparisons

Go deeper