Model comparison

GPT-5 vs GPT-5 Nano

GPT-5 is the stronger model overall, scoring 50.9 to 33.5 on the Noometry Index. GPT-5 Nano costs 25× less per token, which makes it the better buy when GPT-5's lead doesn't matter for your workload.

Last verified . 48 shared benchmarks.

GPT-5 OpenAI

50.9

Rank #45 Confirmed

GPT-5 Nano OpenAI

33.5

Rank #241 Confirmed

Summary

  • They share 48 benchmarks with published results for both. GPT-5 scores higher in 9 categories and GPT-5 Nano in 1 category; 10 gaps are clear of the uncertainty.
  • The widest gap is in long context, where GPT-5 leads 69.5 to 31.3.
  • The biggest single-benchmark swing is Fiction.LiveBench: 97.2% for GPT-5 and 44.4% for GPT-5 Nano.
  • GPT-5 Nano is cheaper at $0.05 / $0.40 per million input/output tokens, against $1.25 / $10 for GPT-5.

Side by side

GPT-5 and GPT-5 Nano specifications
GPT-5GPT-5 Nano
ProviderOpenAIOpenAI
Noometry Index50.933.5
Released2025-08-072025-08-07
WeightsProprietaryProprietary
Context window400K400K
Max output128K128K
Input $ / M tokens$1.25$0.05
Output $ / M tokens$10$0.40
Results tracked6949

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-5 leads

GPT-5: 50.3 (#47), GPT-5 Nano: 33.6 (#254)

Coding benchmarks
BenchmarkGPT-5GPT-5 Nano
SWE-bench Verified (bash only)65%34.8%
WeirdML60.7%38.1%
LMArena Coding14361351
ALE-Bench1,162718.67
SWE-bench Verified73.6%—
Aider Polyglot88%—
LMArena WebDev1418—
SciCode42.9%—
GSO6.9%—
AlgoTune1.67—

Agentic & Tool Use GPT-5 leads

GPT-5: 33.1 (#56), GPT-5 Nano: 25.8 (#106)

Agentic & Tool Use benchmarks
BenchmarkGPT-5GPT-5 Nano
Terminal-Bench49.6%21.8%
Berkeley Function Calling Leaderboard—51.5%
GDPval34.8%—
Remote Labor Index1.7%—
DeepResearch Bench49.6%—
BALROG32.8%—
LMArena Search1133—
METR Time Horizons69.6%—

Reasoning GPT-5 leads

GPT-5: 38.3 (#64), GPT-5 Nano: 16.3 (#306)

Reasoning benchmarks
BenchmarkGPT-5GPT-5 Nano
ARC-AGI-29.9%2.6%
Kagi LLM Benchmark72.7%62.2%
ARC-AGI-165.7%20.7%
Chess Puzzles37%27%
LMArena Hard Prompts14161328
Mystery Game Puzzles23%9%
DTBench90.7%62.7%
LMCA40%7.9%
Epoch Capabilities Index150139.38
ForecastBench61.459.1
SimpleBench56.7%—
CritPt12.6%—
EnigmaEval10.5%—
EBR-Bench12.7%—

Math GPT-5 leads

GPT-5: 55.0 (#44), GPT-5 Nano: 29.4 (#241)

Math benchmarks
BenchmarkGPT-5GPT-5 Nano
FrontierMath (Tiers 1-3)55.4%20%
FrontierMath Tier 422%2.4%
OTIS Mock AIME 2024-202591.4%81.1%
ProofBench18%12%
Omni-MATH64.7%54.6%
LMArena Math14071317
MATH Level 598.1%95.2%
FrontierMath (Feb 2025 set)32.4%8.3%
FrontierMath Tier 4 (v1)12.5%2.1%

Knowledge GPT-5 leads

GPT-5: 56.6 (#43), GPT-5 Nano: 35.9 (#178)

Knowledge benchmarks
BenchmarkGPT-5GPT-5 Nano
GPQA Diamond86.2%69.4%
SimpleQA Verified50.1%11.7%
MMLU-Pro86.3%77.8%
Vectara Hallucination Rate14.7%10.5%
GPQA (HELM)79.2%67.9%
LMArena Expert14191321
Humanity's Last Exam25.3%—
Confabulations10.3%—

Multimodal GPT-5 leads

GPT-5: 46.8 (#13), GPT-5 Nano: 31.3 (#108)

Multimodal benchmarks
BenchmarkGPT-5GPT-5 Nano
LMArena Vision12321159
VPCT66%37.2%
GeoBench81%—

Multilingual GPT-5 leads

GPT-5: 51.4 (#110), GPT-5 Nano: 45.3 (#172)

Multilingual benchmarks
BenchmarkGPT-5GPT-5 Nano
LMArena Non-English13971313
LMArena Chinese14221356
LMArena German14161327
LMArena Japanese14091226
LMArena Korean13601269
LMArena Russian14061296
LMArena Spanish13991360
LMArena French1410—

Instruction Following GPT-5 Nano leads

GPT-5: 73.8 (#113), GPT-5 Nano: 75.0 (#79)

Instruction Following benchmarks
BenchmarkGPT-5GPT-5 Nano
IFEval87.5%93.2%
LMArena Instruction Following13881306

Long Context GPT-5 leads

GPT-5: 69.5 (#2), GPT-5 Nano: 31.3 (#281)

Long Context benchmarks
BenchmarkGPT-5GPT-5 Nano
Fiction.LiveBench97.2%44.4%
LMArena Longer Query13991312

Writing & Preference GPT-5 leads

GPT-5: 63.4 (#65), GPT-5 Nano: 39.1 (#249)

Writing & Preference benchmarks
BenchmarkGPT-5GPT-5 Nano
LMArena Text14061320
LMArena Creative Writing13651249
EQ-Bench Creative Writing1627705
WildBench85.7%80.6%
LMArena Multi-Turn14261311
Short-Story Creative Writing86%—

Frequently asked questions

Is GPT-5 better than GPT-5 Nano?

GPT-5 is the stronger model overall, scoring 50.9 to 33.5 on the Noometry Index. GPT-5 Nano costs 25× less per token, which makes it the better buy when GPT-5's lead doesn't matter for your workload.

Which is cheaper, GPT-5 or GPT-5 Nano?

GPT-5 Nano is cheaper. It lists at $0.05 per million input tokens and $0.40 per million output tokens; GPT-5 lists at $1.25 and $10.

Is GPT-5 or GPT-5 Nano better for coding?

GPT-5 scores higher on coding benchmarks: 50.3 versus 33.6 in the Noometry coding category.

Which has the bigger context window?

Both accept 400K tokens.

How many benchmarks do GPT-5 and GPT-5 Nano share?

48 benchmarks have published results for both models. GPT-5 has 69 scored results on Noometry and GPT-5 Nano has 49.

Related comparisons

Go deeper