Model comparison

GPT-5 Nano vs o1

o1 is the stronger model overall, scoring 40.9 to 33.5 on the Noometry Index. GPT-5 Nano costs 191× less per token, which makes it the better buy when o1's lead doesn't matter for your workload.

Last verified . 31 shared benchmarks.

GPT-5 Nano OpenAI

33.5

Rank #241 Confirmed

o1 OpenAI

40.9

Rank #143 Confirmed

Summary

  • They share 31 benchmarks with published results for both. GPT-5 Nano scores higher in 2 categories and o1 in 8 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in long context, where o1 leads 50.3 to 31.3.
  • The biggest single-benchmark swing is Fiction.LiveBench: 44.4% for GPT-5 Nano and 83.3% for o1.
  • GPT-5 Nano is cheaper at $0.05 / $0.40 per million input/output tokens, against $15 / $60 for o1.
  • GPT-5 Nano accepts more context: 400K tokens versus 200K.

Side by side

GPT-5 Nano and o1 specifications
GPT-5 Nanoo1
ProviderOpenAIOpenAI
Noometry Index33.540.9
Released2025-08-072024-09-12
WeightsProprietaryProprietary
Context window400K200K
Max output128K100K
Input $ / M tokens$0.05$15
Output $ / M tokens$0.40$60
Results tracked4952

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding o1 leads

GPT-5 Nano: 33.6 (#254), o1: 46.1 (#70)

Coding benchmarks
BenchmarkGPT-5 Nanoo1
WeirdML38.1%47.6%
LMArena Coding13511367
SWE-bench Verified (bash only)34.8%—
Aider Polyglot—61.7%
LiveBench Coding—69.7%
CadEval—56%
ALE-Bench718.67—
HumanEval+—89%
MBPP+—80.2%

Agentic & Tool Use GPT-5 Nano leads

GPT-5 Nano: 25.8 (#106), o1: 24.6 (#117)

Agentic & Tool Use benchmarks
BenchmarkGPT-5 Nanoo1
Terminal-Bench21.8%—
Berkeley Function Calling Leaderboard51.5%—
Cybench—10%
METR Time Horizons—51.1%

Reasoning o1 leads

GPT-5 Nano: 16.3 (#306), o1: 27.9 (#111)

Reasoning benchmarks
BenchmarkGPT-5 Nanoo1
ARC-AGI-120.7%30.7%
Chess Puzzles27%15%
LMArena Hard Prompts13281371
DTBench62.7%74.7%
LMCA7.9%22.3%
Epoch Capabilities Index139.38141.91
ARC-AGI-22.6%—
SimpleBench—41.7%
Kagi LLM Benchmark62.2%—
EnigmaEval—5.7%
LiveBench Reasoning—91.6%
Mystery Game Puzzles9%—
LiveBench Data Analysis—65.5%
ForecastBench59.1—
LiveBench—75.7%

Math o1 leads

GPT-5 Nano: 29.4 (#241), o1: 36.1 (#175)

Knowledge o1 leads

GPT-5 Nano: 35.9 (#178), o1: 41.5 (#110)

Knowledge benchmarks
BenchmarkGPT-5 Nanoo1
GPQA Diamond69.4%76.8%
SimpleQA Verified11.7%41.1%
LMArena Expert13211361
Humanity's Last Exam—8%
MMLU-Pro77.8%—
Confabulations—11.7%
Vectara Hallucination Rate10.5%—
GPQA (HELM)67.9%—

Multimodal o1 leads

GPT-5 Nano: 31.3 (#108), o1: 34.2 (#93)

Multimodal benchmarks
BenchmarkGPT-5 Nanoo1
LMArena Vision11591168
VPCT37.2%37%
GeoBench—80%
SpatialViz-Bench—41.4%

Multilingual o1 leads

GPT-5 Nano: 45.3 (#172), o1: 48.6 (#142)

Multilingual benchmarks
BenchmarkGPT-5 Nanoo1
LMArena Non-English13131358
LMArena Chinese13561394
LMArena German13271337
LMArena Japanese12261346
LMArena Korean12691396
LMArena Russian12961356
LMArena Spanish13601345
LMArena French—1344

Instruction Following Too close to call

GPT-5 Nano: 75.0 (#79), o1: 74.8 (#86)

Instruction Following benchmarks
BenchmarkGPT-5 Nanoo1
LMArena Instruction Following13061367
LiveBench Instruction Following—81.5%
IFEval93.2%—

Long Context o1 leads

GPT-5 Nano: 31.3 (#281), o1: 50.3 (#9)

Long Context benchmarks
BenchmarkGPT-5 Nanoo1
Fiction.LiveBench44.4%83.3%
LMArena Longer Query13121378

Writing & Preference o1 leads

GPT-5 Nano: 39.1 (#249), o1: 55.6 (#144)

Writing & Preference benchmarks
BenchmarkGPT-5 Nanoo1
LMArena Text13201366
LMArena Creative Writing12491348
LMArena Multi-Turn13111369
Short-Story Creative Writing—70.2%
EQ-Bench Creative Writing705—
WildBench80.6%—
LiveBench Language—65.4%

Frequently asked questions

Is GPT-5 Nano better than o1?

o1 is the stronger model overall, scoring 40.9 to 33.5 on the Noometry Index. GPT-5 Nano costs 191× less per token, which makes it the better buy when o1's lead doesn't matter for your workload.

Which is cheaper, GPT-5 Nano or o1?

GPT-5 Nano is cheaper. It lists at $0.05 per million input tokens and $0.40 per million output tokens; o1 lists at $15 and $60.

Is GPT-5 Nano or o1 better for coding?

o1 scores higher on coding benchmarks: 46.1 versus 33.6 in the Noometry coding category.

Which has the bigger context window?

GPT-5 Nano does, with 400K tokens against 200K.

How many benchmarks do GPT-5 Nano and o1 share?

31 benchmarks have published results for both models. GPT-5 Nano has 49 scored results on Noometry and o1 has 52.

Related comparisons

Go deeper