Model comparison

Claude Fable 5 vs GPT-5.5 Instant

Claude Fable 5 is the stronger model overall, scoring 66.8 to 42.7 on the Noometry Index.

Last verified . 27 shared benchmarks.

Claude Fable 5 Anthropic

66.8

Rank #5 Confirmed

GPT-5.5 Instant OpenAI

42.7

Rank #110 Confirmed

Summary

  • They share 27 benchmarks with published results for both. Claude Fable 5 scores higher in 9 categories and GPT-5.5 Instant in 0 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in math, where Claude Fable 5 leads 88.5 to 26.5.
  • The biggest single-benchmark swing is FrontierMath Tier 4: 90.2% for Claude Fable 5 and 2.4% for GPT-5.5 Instant.

Side by side

Claude Fable 5 and GPT-5.5 Instant specifications
Claude Fable 5GPT-5.5 Instant
ProviderAnthropicOpenAI
Noometry Index66.842.7
Released2026-06-072026-05-05
WeightsProprietaryProprietary
Context window1M—
Max output128K—
Input $ / M tokens$10—
Output $ / M tokens$50—
Results tracked6227

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Fable 5 leads

Claude Fable 5: 70.6 (#4), GPT-5.5 Instant: 44.3 (#74)

Coding benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
SciCode61%48.6%
LMArena Coding15191433
DeepSWE69.9%—
FrontierCode53.5%—
LMArena WebDev1625—
FrontierSWE47%—
GSO78.4%—
WeirdML91.9%—
MirrorCode63.9%—
ALE-Bench2,041—

Agentic & Tool Use Not comparable

Claude Fable 5: 54.0 (#2), GPT-5.5 Instant: —

Agentic & Tool Use benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
APEX-Agents63.6%—
Remote Labor Index16.1%—
τ²-bench Banking39.7%—
PostTrainBench41.8%—
GBAEval74.5%—
GDP.pdf30%—
LMArena Search1230—
Vending-Bench 25,680—

Reasoning Claude Fable 5 leads

Claude Fable 5: 76.8 (#6), GPT-5.5 Instant: 24.9 (#155)

Reasoning benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
CritPt28.6%0%
Chess Puzzles41%12%
LMArena Hard Prompts15081426
Epoch Capabilities Index162.06142.52
ARC-AGI-289.2%—
SimpleBench81.9%—
Kagi LLM Benchmark91.4%—
NYT Connections (extended)92.7%—
ARC-AGI-198.5%—
EnigmaEval39.3%—
EBR-Bench39.5%—
Mystery Game Puzzles52%—
DTBench98.4%—
LMCA61.1%—
Surface Evolver Bench95%—
Bench to the Future 30.13—

Math Claude Fable 5 leads

Claude Fable 5: 88.5 (#5), GPT-5.5 Instant: 26.5 (#259)

Math benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
FrontierMath (Tiers 1-3)87%26.3%
FrontierMath Tier 490.2%2.4%
OTIS Mock AIME 2024-2025100%68.1%
LMArena Math15191420
ProofBench95%—
FrontierMath Erdős0%—

Knowledge Claude Fable 5 leads

Claude Fable 5: 62.2 (#25), GPT-5.5 Instant: 48.9 (#74)

Knowledge benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
GPQA Diamond85.9%82.5%
LMArena Expert15341409
SimpleQA Verified70.7%—

Multimodal Claude Fable 5 leads

Claude Fable 5: 45.3 (#17), GPT-5.5 Instant: 40.0 (#52)

Multimodal benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
LMArena Vision13241250
LMArena Document14961403
Blueprint-Bench 238.6%—
Furniture Assembly35.8%—

Multilingual Claude Fable 5 leads

Claude Fable 5: 57.3 (#9), GPT-5.5 Instant: 52.8 (#80)

Multilingual benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
LMArena Non-English14811417
LMArena Chinese15431456
LMArena French15051428
LMArena German14861411
LMArena Japanese15061408
LMArena Korean14881392
LMArena Russian15041431
LMArena Spanish14981429

Instruction Following Claude Fable 5 leads

Claude Fable 5: 78.6 (#8), GPT-5.5 Instant: 74.2 (#100)

Instruction Following benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
LMArena Instruction Following15021406

Long Context Claude Fable 5 leads

Claude Fable 5: 46.3 (#23), GPT-5.5 Instant: 43.4 (#96)

Long Context benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
LMArena Longer Query15091422

Writing & Preference Claude Fable 5 leads

Claude Fable 5: 75.9 (#5), GPT-5.5 Instant: 61.8 (#85)

Writing & Preference benchmarks
BenchmarkClaude Fable 5GPT-5.5 Instant
LMArena Text14911419
LMArena Creative Writing14941419
LMArena Multi-Turn15041433
EQ-Bench Creative Writing1943—
EQ-Bench 41340—

Frequently asked questions

Is Claude Fable 5 better than GPT-5.5 Instant?

Claude Fable 5 is the stronger model overall, scoring 66.8 to 42.7 on the Noometry Index.

Is Claude Fable 5 or GPT-5.5 Instant better for coding?

Claude Fable 5 scores higher on coding benchmarks: 70.6 versus 44.3 in the Noometry coding category.

How many benchmarks do Claude Fable 5 and GPT-5.5 Instant share?

27 benchmarks have published results for both models. Claude Fable 5 has 62 scored results on Noometry and GPT-5.5 Instant has 27.

Related comparisons

Go deeper