Model comparison

Amazon Nova Experimental Chat 10 09 vs GPT-5.4 nano

Amazon Nova Experimental Chat 10 09 and GPT-5.4 nano score almost the same on the Noometry Index (41.9 vs 41.9), so choose on price, context window or the category you care about most.

Last verified . 9 shared benchmarks.

GPT-5.4 nano OpenAI

41.9

Rank #125 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Amazon Nova Experimental Chat 10 09 scores higher in 1 category and GPT-5.4 nano in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Amazon Nova Experimental Chat 10 09 leads 27.3 to 23.7.

Side by side

Amazon Nova Experimental Chat 10 09 and GPT-5.4 nano specifications
Amazon Nova Experimental Chat 10 09GPT-5.4 nano
ProviderAmazonOpenAI
Noometry Index41.941.9
Released—2026-03-17
WeightsProprietaryProprietary
Context window—400K
Max output—128K
Input $ / M tokens—$0.20
Output $ / M tokens—$1.25
Results tracked940

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-5.4 nano leads

Amazon Nova Experimental Chat 10 09: 40.2 (#145), GPT-5.4 nano: 43.6 (#84)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Coding13701405
SciCode—46.9%
WeirdML—49.2%
ALE-Bench—1,005

Reasoning Amazon Nova Experimental Chat 10 09 leads

Amazon Nova Experimental Chat 10 09: 27.3 (#120), GPT-5.4 nano: 23.7 (#173)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Hard Prompts13561381
ARC-AGI-2—5.7%
Kagi LLM Benchmark—39.7%
ARC-AGI-1—51.5%
CritPt—9.3%
Chess Puzzles—30%
Mystery Game Puzzles—9%
DTBench—80.3%
LMCA—36.9%
Epoch Capabilities Index—145.81
ForecastBench—57.3

Math Not comparable

Amazon Nova Experimental Chat 10 09: —, GPT-5.4 nano: 40.9 (#88)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
FrontierMath (Tiers 1-3)—44.9%
FrontierMath Tier 4—12.2%
OTIS Mock AIME 2024-2025—87.8%
ProofBench—5%
LMArena Math—1406
FrontierMath (Feb 2025 set)—25.9%
FrontierMath Tier 4 (v1)—6.3%

Knowledge Not comparable

Amazon Nova Experimental Chat 10 09: —, GPT-5.4 nano: 41.9 (#103)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
GPQA Diamond—78.5%
SimpleQA Verified—11.7%
Vectara Hallucination Rate—3.1%
LMArena Expert—1396

Multimodal Not comparable

Amazon Nova Experimental Chat 10 09: —, GPT-5.4 nano: 36.7 (#78)

Multimodal benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Vision—1196

Multilingual GPT-5.4 nano leads

Amazon Nova Experimental Chat 10 09: 47.3 (#150), GPT-5.4 nano: 48.6 (#140)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Non-English13401359
LMArena Chinese13411392
LMArena French—1396
LMArena German—1367
LMArena Japanese—1343
LMArena Korean—1320
LMArena Russian—1363
LMArena Spanish—1371

Instruction Following GPT-5.4 nano leads

Amazon Nova Experimental Chat 10 09: 69.4 (#172), GPT-5.4 nano: 71.9 (#144)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Instruction Following13151362

Long Context GPT-5.4 nano leads

Amazon Nova Experimental Chat 10 09: 40.5 (#152), GPT-5.4 nano: 41.6 (#137)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Longer Query13331366

Writing & Preference GPT-5.4 nano leads

Amazon Nova Experimental Chat 10 09: 54.7 (#149), GPT-5.4 nano: 55.7 (#142)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GPT-5.4 nano
LMArena Text13641372
LMArena Creative Writing13071314
LMArena Multi-Turn13551382

Frequently asked questions

Is Amazon Nova Experimental Chat 10 09 better than GPT-5.4 nano?

Amazon Nova Experimental Chat 10 09 and GPT-5.4 nano score almost the same on the Noometry Index (41.9 vs 41.9), so choose on price, context window or the category you care about most.

Is Amazon Nova Experimental Chat 10 09 or GPT-5.4 nano better for coding?

GPT-5.4 nano scores higher on coding benchmarks: 43.6 versus 40.2 in the Noometry coding category.

How many benchmarks do Amazon Nova Experimental Chat 10 09 and GPT-5.4 nano share?

9 benchmarks have published results for both models. Amazon Nova Experimental Chat 10 09 has 9 scored results on Noometry and GPT-5.4 nano has 40.

Related comparisons

Go deeper