Model comparison

GPT-5.6 Terra vs GPT-5 Mini

GPT-5.6 Terra is the stronger model overall, scoring 59.2 to 41.8 on the Noometry Index. GPT-5 Mini costs 6.5× less per token, which makes it the better buy when GPT-5.6 Terra's lead doesn't matter for your workload.

Last verified . 38 shared benchmarks.

GPT-5.6 Terra OpenAI

59.2

Rank #17 Confirmed

GPT-5 Mini OpenAI

41.8

Rank #128 Confirmed

Summary

  • They share 38 benchmarks with published results for both. GPT-5.6 Terra scores higher in 10 categories and GPT-5 Mini in 0 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where GPT-5.6 Terra leads 60.7 to 23.9.
  • The biggest single-benchmark swing is ARC-AGI-2: 83.9% for GPT-5.6 Terra and 4.4% for GPT-5 Mini.
  • GPT-5 Mini is cheaper at $0.25 / $2 per million input/output tokens, against $2 / $12 for GPT-5.6 Terra.
  • GPT-5.6 Terra accepts more context: 1.05M tokens versus 400K.

Side by side

GPT-5.6 Terra and GPT-5 Mini specifications
GPT-5.6 TerraGPT-5 Mini
ProviderOpenAIOpenAI
Noometry Index59.241.8
Released2026-07-092025-08-07
WeightsProprietaryProprietary
Context window1.05M400K
Max output128K128K
Input $ / M tokens$2$0.25
Output $ / M tokens$12$2
Results tracked5260

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-5.6 Terra leads

GPT-5.6 Terra: 57.7 (#19), GPT-5 Mini: 40.1 (#146)

Coding benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
SciCode55%39.2%
WeirdML78.3%52.7%
LMArena Coding14841406
ALE-Bench1,951799.77
SWE-bench Verified—64.7%
DeepSWE69.6%—
FrontierCode41.3%—
SWE-bench Verified (bash only)—59.8%
CursorBench41.3%—
LMArena WebDev1522—
SWE-bench Multilingual—39.7%
AlgoTune—1.38

Agentic & Tool Use GPT-5.6 Terra leads

GPT-5.6 Terra: 40.1 (#25), GPT-5 Mini: 31.1 (#70)

Agentic & Tool Use benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
Vending-Bench 27,343-31.18
Terminal-Bench—34.8%
APEX-Agents58.2%—
Berkeley Function Calling Leaderboard—55.5%
BALROG53.2%—
GDP.pdf24.7%—

Reasoning GPT-5.6 Terra leads

GPT-5.6 Terra: 60.7 (#21), GPT-5 Mini: 23.9 (#168)

Reasoning benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
ARC-AGI-283.9%4.4%
Kagi LLM Benchmark51.3%70.3%
ARC-AGI-196.5%54.3%
CritPt30%0%
Chess Puzzles54%30%
LMArena Hard Prompts14681380
Mystery Game Puzzles35%10%
DTBench93.3%80.5%
LMCA55%34.2%
Epoch Capabilities Index159.62145.52
SimpleBench48.9%—
NYT Connections (extended)78.4%—
EnigmaEval—8.2%
Surface Evolver Bench83.8%—
ForecastBench—61

Math GPT-5.6 Terra leads

GPT-5.6 Terra: 81.6 (#12), GPT-5 Mini: 46.7 (#69)

Math benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
FrontierMath (Tiers 1-3)86%46.7%
FrontierMath Tier 470.7%12.2%
OTIS Mock AIME 2024-202599.7%86.7%
ProofBench74%9%
LMArena Math14661378
Omni-MATH—72.2%
MATH Level 5—97.8%
FrontierMath (Feb 2025 set)—27.2%
FrontierMath Tier 4 (v1)—6.3%

Knowledge GPT-5.6 Terra leads

GPT-5.6 Terra: 61.2 (#30), GPT-5 Mini: 45.6 (#86)

Knowledge benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
GPQA Diamond93.3%75%
SimpleQA Verified43.2%21.6%
LMArena Expert14921379
Humanity's Last Exam—19.4%
MMLU-Pro—83.5%
Confabulations—13.3%
Vectara Hallucination Rate—12.9%
GPQA (HELM)—75.6%

Multimodal GPT-5.6 Terra leads

GPT-5.6 Terra: 47.3 (#11), GPT-5 Mini: 35.6 (#85)

Multimodal benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
LMArena Vision12711202
VPCT—40.2%
Blueprint-Bench 230.8%—
Furniture Assembly54.2%—
LMArena Document1472—

Multilingual GPT-5.6 Terra leads

GPT-5.6 Terra: 54.4 (#44), GPT-5 Mini: 48.9 (#137)

Multilingual benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
LMArena Non-English14391363
LMArena Chinese15131385
LMArena French14711386
LMArena German14601366
LMArena Japanese14571341
LMArena Korean14251308
LMArena Russian14501362
LMArena Spanish14481355

Instruction Following Too close to call

GPT-5.6 Terra: 76.4 (#40), GPT-5 Mini: 76.2 (#46)

Instruction Following benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
LMArena Instruction Following14541357
IFEval—92.7%

Long Context GPT-5.6 Terra leads

GPT-5.6 Terra: 44.4 (#68), GPT-5 Mini: 41.9 (#132)

Long Context benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
LMArena Longer Query14511355
Fiction.LiveBench—69.4%

Writing & Preference GPT-5.6 Terra leads

GPT-5.6 Terra: 70.2 (#23), GPT-5 Mini: 55.2 (#148)

Writing & Preference benchmarks
BenchmarkGPT-5.6 TerraGPT-5 Mini
LMArena Text14471373
LMArena Creative Writing14101325
EQ-Bench Creative Writing18551313
LMArena Multi-Turn14491363
Short-Story Creative Writing—83.1%
WildBench—85.5%
EQ-Bench 41234—

Frequently asked questions

Is GPT-5.6 Terra better than GPT-5 Mini?

GPT-5.6 Terra is the stronger model overall, scoring 59.2 to 41.8 on the Noometry Index. GPT-5 Mini costs 6.5× less per token, which makes it the better buy when GPT-5.6 Terra's lead doesn't matter for your workload.

Which is cheaper, GPT-5.6 Terra or GPT-5 Mini?

GPT-5 Mini is cheaper. It lists at $0.25 per million input tokens and $2 per million output tokens; GPT-5.6 Terra lists at $2 and $12.

Is GPT-5.6 Terra or GPT-5 Mini better for coding?

GPT-5.6 Terra scores higher on coding benchmarks: 57.7 versus 40.1 in the Noometry coding category.

Which has the bigger context window?

GPT-5.6 Terra does, with 1.05M tokens against 400K.

How many benchmarks do GPT-5.6 Terra and GPT-5 Mini share?

38 benchmarks have published results for both models. GPT-5.6 Terra has 52 scored results on Noometry and GPT-5 Mini has 60.

Related comparisons

Go deeper