Model comparison

GPT-5 Nano vs gpt-oss-20b

GPT-5 Nano is the stronger model overall, scoring 33.5 to 32.5 on the Noometry Index. gpt-oss-20b costs 3.8× less per token, which makes it the better buy when GPT-5 Nano's lead doesn't matter for your workload.

Last verified . 32 shared benchmarks.

GPT-5 Nano OpenAI

33.5

Rank #241 Confirmed

gpt-oss-20b OpenAI

32.5

Rank #255 Confirmed

Summary

  • They share 32 benchmarks with published results for both. GPT-5 Nano scores higher in 5 categories and gpt-oss-20b in 4 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in agentic & tool use, where GPT-5 Nano leads 25.8 to 9.3.
  • The biggest single-benchmark swing is Chess Puzzles: 27% for GPT-5 Nano and 4% for gpt-oss-20b.
  • gpt-oss-20b is cheaper at $0.018 / $0.09 per million input/output tokens, against $0.05 / $0.40 for GPT-5 Nano.
  • GPT-5 Nano accepts more context: 400K tokens versus 131K.
  • gpt-oss-20b has downloadable open weights; the other is API-only.

Side by side

GPT-5 Nano and gpt-oss-20b specifications
GPT-5 Nanogpt-oss-20b
ProviderOpenAIOpenAI
Noometry Index33.532.5
Released2025-08-072025-08-05
WeightsProprietaryOpen
Context window400K131K
Max output128K16K
Input $ / M tokens$0.05$0.018
Output $ / M tokens$0.40$0.09
Results tracked4934

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding gpt-oss-20b leads

GPT-5 Nano: 33.6 (#254), gpt-oss-20b: 37.6 (#192)

Coding benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
WeirdML38.1%40.9%
LMArena Coding13511306
ALE-Bench718.67566.05
SWE-bench Verified (bash only)34.8%—
SciCode—34.4%

Agentic & Tool Use GPT-5 Nano leads

GPT-5 Nano: 25.8 (#106), gpt-oss-20b: 9.3 (#154)

Agentic & Tool Use benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
Terminal-Bench21.8%3.4%
Berkeley Function Calling Leaderboard51.5%—

Reasoning gpt-oss-20b leads

GPT-5 Nano: 16.3 (#306), gpt-oss-20b: 19.3 (#261)

Reasoning benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
Kagi LLM Benchmark62.2%53.2%
Chess Puzzles27%4%
LMArena Hard Prompts13281274
DTBench62.7%68%
LMCA7.9%14.5%
Epoch Capabilities Index139.38137.82
ARC-AGI-22.6%—
ARC-AGI-120.7%—
CritPt—1.4%
Mystery Game Puzzles9%—
ForecastBench59.1—

Math gpt-oss-20b leads

GPT-5 Nano: 29.4 (#241), gpt-oss-20b: 39.4 (#103)

Math benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
OTIS Mock AIME 2024-202581.1%65.3%
Omni-MATH54.6%56.5%
LMArena Math13171317
FrontierMath (Tiers 1-3)20%—
FrontierMath Tier 42.4%—
ProofBench12%—
MATH Level 595.2%—
FrontierMath (Feb 2025 set)8.3%—
FrontierMath Tier 4 (v1)2.1%—

Knowledge GPT-5 Nano leads

GPT-5 Nano: 35.9 (#178), gpt-oss-20b: 34.6 (#195)

Knowledge benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
GPQA Diamond69.4%60.8%
MMLU-Pro77.8%74%
GPQA (HELM)67.9%59.4%
LMArena Expert13211258
SimpleQA Verified11.7%—
Vectara Hallucination Rate10.5%—

Multimodal Not comparable

GPT-5 Nano: 31.3 (#108), gpt-oss-20b: —

Multimodal benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
LMArena Vision1159—
VPCT37.2%—

Multilingual GPT-5 Nano leads

GPT-5 Nano: 45.3 (#172), gpt-oss-20b: 42.2 (#197)

Multilingual benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
LMArena Non-English13131268
LMArena Chinese13561314
LMArena German13271255
LMArena Japanese12261244
LMArena Korean12691236
LMArena Russian12961278
LMArena Spanish13601267

Instruction Following GPT-5 Nano leads

GPT-5 Nano: 75.0 (#79), gpt-oss-20b: 61.8 (#240)

Instruction Following benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
IFEval93.2%73.2%
LMArena Instruction Following13061236

Long Context gpt-oss-20b leads

GPT-5 Nano: 31.3 (#281), gpt-oss-20b: 37.9 (#209)

Long Context benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
LMArena Longer Query13121250
Fiction.LiveBench44.4%—

Writing & Preference GPT-5 Nano leads

GPT-5 Nano: 39.1 (#249), gpt-oss-20b: 35.5 (#265)

Writing & Preference benchmarks
BenchmarkGPT-5 Nanogpt-oss-20b
LMArena Text13201287
LMArena Creative Writing12491201
EQ-Bench Creative Writing705666
WildBench80.6%73.7%
LMArena Multi-Turn13111268

Frequently asked questions

Is GPT-5 Nano better than gpt-oss-20b?

GPT-5 Nano is the stronger model overall, scoring 33.5 to 32.5 on the Noometry Index. gpt-oss-20b costs 3.8× less per token, which makes it the better buy when GPT-5 Nano's lead doesn't matter for your workload.

Which is cheaper, GPT-5 Nano or gpt-oss-20b?

gpt-oss-20b is cheaper. It lists at $0.018 per million input tokens and $0.09 per million output tokens; GPT-5 Nano lists at $0.05 and $0.40.

Is GPT-5 Nano or gpt-oss-20b better for coding?

gpt-oss-20b scores higher on coding benchmarks: 37.6 versus 33.6 in the Noometry coding category.

Which has the bigger context window?

GPT-5 Nano does, with 400K tokens against 131K.

How many benchmarks do GPT-5 Nano and gpt-oss-20b share?

32 benchmarks have published results for both models. GPT-5 Nano has 49 scored results on Noometry and gpt-oss-20b has 34.

Related comparisons

Go deeper