Model comparison

Amazon Nova Experimental Chat 10 09 vs Grok 4.1 Fast

Amazon Nova Experimental Chat 10 09 and Grok 4.1 Fast score almost the same on the Noometry Index (41.9 vs 41.4), so choose on price, context window or the category you care about most.

Last verified . 9 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Amazon Nova Experimental Chat 10 09 scores higher in 1 category and Grok 4.1 Fast in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 27.3.

Side by side

Amazon Nova Experimental Chat 10 09 and Grok 4.1 Fast specifications
Amazon Nova Experimental Chat 10 09Grok 4.1 Fast
ProviderAmazonxAI
Noometry Index41.941.4
Released—2025-06-27
WeightsProprietaryProprietary
Context window—128K
Max output—30K
Input $ / M tokens—$0.20
Output $ / M tokens—$0.50
Results tracked932

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Amazon Nova Experimental Chat 10 09 leads

Amazon Nova Experimental Chat 10 09: 40.2 (#145), Grok 4.1 Fast: 34.1 (#245)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Coding13701411
LMArena WebDev—1242
ALE-Bench—394.93

Agentic & Tool Use Not comparable

Amazon Nova Experimental Chat 10 09: —, Grok 4.1 Fast: 36.3 (#39)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
Berkeley Function Calling Leaderboard—69.6%
τ²-bench Banking—13.1%
LMArena Search—1171
Vending-Bench 2—1,107

Reasoning Grok 4.1 Fast leads

Amazon Nova Experimental Chat 10 09: 27.3 (#120), Grok 4.1 Fast: 43.4 (#49)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Hard Prompts13561407
SimpleBench—56%
NYT Connections (extended)—87.4%
DTBench—87.7%
ForecastBench—61

Math Not comparable

Amazon Nova Experimental Chat 10 09: —, Grok 4.1 Fast: 31.9 (#221)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
MathArena Final-Answer Competitions—60.9%
ProofBench—4%
LMArena Math—1408

Knowledge Not comparable

Amazon Nova Experimental Chat 10 09: —, Grok 4.1 Fast: 33.1 (#207)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
Vectara Hallucination Rate—17.8%
LMArena Expert—1399

Multimodal Not comparable

Amazon Nova Experimental Chat 10 09: —, Grok 4.1 Fast: 37.0 (#76)

Multimodal benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Vision—1201

Multilingual Grok 4.1 Fast leads

Amazon Nova Experimental Chat 10 09: 47.3 (#150), Grok 4.1 Fast: 51.0 (#114)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Non-English13401391
LMArena Chinese13411441
LMArena French—1415
LMArena German—1404
LMArena Japanese—1349
LMArena Korean—1361
LMArena Russian—1387
LMArena Spanish—1413

Instruction Following Grok 4.1 Fast leads

Amazon Nova Experimental Chat 10 09: 69.4 (#172), Grok 4.1 Fast: 72.7 (#133)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Instruction Following13151376

Long Context Grok 4.1 Fast leads

Amazon Nova Experimental Chat 10 09: 40.5 (#152), Grok 4.1 Fast: 42.4 (#126)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Longer Query13331390

Writing & Preference Grok 4.1 Fast leads

Amazon Nova Experimental Chat 10 09: 54.7 (#149), Grok 4.1 Fast: 57.2 (#131)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09Grok 4.1 Fast
LMArena Text13641408
LMArena Creative Writing13071394
LMArena Multi-Turn13551389
EQ-Bench Creative Writing—1327

Frequently asked questions

Is Amazon Nova Experimental Chat 10 09 better than Grok 4.1 Fast?

Amazon Nova Experimental Chat 10 09 and Grok 4.1 Fast score almost the same on the Noometry Index (41.9 vs 41.4), so choose on price, context window or the category you care about most.

Is Amazon Nova Experimental Chat 10 09 or Grok 4.1 Fast better for coding?

Amazon Nova Experimental Chat 10 09 scores higher on coding benchmarks: 40.2 versus 34.1 in the Noometry coding category.

How many benchmarks do Amazon Nova Experimental Chat 10 09 and Grok 4.1 Fast share?

9 benchmarks have published results for both models. Amazon Nova Experimental Chat 10 09 has 9 scored results on Noometry and Grok 4.1 Fast has 32.

Related comparisons

Go deeper