Model comparison

Claude 3.5 Sonnet vs Claude Fable 5

Claude Fable 5 is the stronger model overall, scoring 66.8 to 34.6 on the Noometry Index.

Last verified . 27 shared benchmarks.

Claude 3.5 Sonnet Anthropic

34.6

Rank #231 Confirmed

Claude Fable 5 Anthropic

66.8

Rank #5 Confirmed

Summary

  • They share 27 benchmarks with published results for both. Claude 3.5 Sonnet scores higher in 0 categories and Claude Fable 5 in 10 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in math, where Claude Fable 5 leads 88.5 to 19.2.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 8.5% for Claude 3.5 Sonnet and 100% for Claude Fable 5.

Side by side

Claude 3.5 Sonnet and Claude Fable 5 specifications
Claude 3.5 SonnetClaude Fable 5
ProviderAnthropicAnthropic
Noometry Index34.666.8
Released2024-06-202026-06-07
WeightsProprietaryProprietary
Context window—1M
Max output—128K
Input $ / M tokens—$10
Output $ / M tokens—$50
Results tracked6062

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Fable 5 leads

Claude 3.5 Sonnet: 39.0 (#165), Claude Fable 5: 70.6 (#4)

Coding benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
GSO4.6%78.4%
WeirdML40%91.9%
LMArena Coding13421519
DeepSWE—69.9%
FrontierCode—53.5%
Aider Polyglot51.6%—
LMArena WebDev—1625
FrontierSWE—47%
SciCode—61%
BigCodeBench Instruct46.8%—
LiveBench Coding67.1%—
MirrorCode—63.9%
BigCodeBench Complete58.6%—
CadEval48%—
ALE-Bench—2,041
HumanEval+81.7%—
MBPP+74.3%—

Agentic & Tool Use Claude Fable 5 leads

Claude 3.5 Sonnet: 32.3 (#67), Claude Fable 5: 54.0 (#2)

Agentic & Tool Use benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
APEX-Agents—63.6%
Remote Labor Index—16.1%
TheAgentCompany24%—
τ²-bench Banking—39.7%
Cybench17.5%—
PostTrainBench—41.8%
BALROG32.6%—
GBAEval—74.5%
GDP.pdf—30%
LMArena Search—1230
METR Time Horizons45.2%—
Vending-Bench 2—5,680

Reasoning Claude Fable 5 leads

Claude 3.5 Sonnet: 23.1 (#183), Claude Fable 5: 76.8 (#6)

Reasoning benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
SimpleBench41.4%81.9%
EnigmaEval0.9%39.3%
LMArena Hard Prompts13051508
DTBench67.8%98.4%
Epoch Capabilities Index133.55162.06
ARC-AGI-2—89.2%
Kagi LLM Benchmark—91.4%
NYT Connections (extended)—92.7%
ARC-AGI-1—98.5%
CritPt—28.6%
Chess Puzzles—41%
EBR-Bench—39.5%
LiveBench Reasoning56.7%—
Mystery Game Puzzles—52%
LiveBench Data Analysis55%—
LMCA—61.1%
Surface Evolver Bench—95%
Bench to the Future 3—0.13
ForecastBench60.7—
LiveBench59%—

Math Claude Fable 5 leads

Claude 3.5 Sonnet: 19.2 (#288), Claude Fable 5: 88.5 (#5)

Math benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
OTIS Mock AIME 2024-20258.5%100%
LMArena Math13071519
FrontierMath (Tiers 1-3)—87%
FrontierMath Tier 4—90.2%
ProofBench—95%
Omni-MATH27.6%—
LiveBench Math52.3%—
MATH Level 556.9%—
FrontierMath (Feb 2025 set)2.1%—
FrontierMath Erdős—0%
FrontierMath Tier 4 (v1)0%—

Knowledge Claude Fable 5 leads

Claude 3.5 Sonnet: 28.6 (#245), Claude Fable 5: 62.2 (#25)

Knowledge benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
GPQA Diamond55.3%85.9%
LMArena Expert12651534
Humanity's Last Exam4.1%—
SimpleQA Verified—70.7%
MMLU-Pro77.7%—
Confabulations19.9%—
GPQA (HELM)56.5%—
MMLU87.3%—

Multimodal Claude Fable 5 leads

Claude 3.5 Sonnet: 26.5 (#120), Claude Fable 5: 45.3 (#17)

Multimodal benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
LMArena Vision11251324
Video-MME60%—
GeoBench62%—
VPCT33%—
Blueprint-Bench 2—38.6%
Furniture Assembly—35.8%
LMArena Document—1496

Multilingual Claude Fable 5 leads

Claude 3.5 Sonnet: 43.2 (#185), Claude Fable 5: 57.3 (#9)

Multilingual benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
LMArena Non-English12831481
LMArena Chinese12721543
LMArena French13051505
LMArena German12971486
LMArena Japanese12341506
LMArena Korean12001488
LMArena Russian13061504
LMArena Spanish12901498

Instruction Following Claude Fable 5 leads

Claude 3.5 Sonnet: 68.8 (#182), Claude Fable 5: 78.6 (#8)

Instruction Following benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
LMArena Instruction Following12971502
LiveBench Instruction Following69.3%—
IFEval85.5%—

Long Context Claude Fable 5 leads

Claude 3.5 Sonnet: 39.9 (#167), Claude Fable 5: 46.3 (#23)

Long Context benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
LMArena Longer Query13111509

Writing & Preference Claude Fable 5 leads

Claude 3.5 Sonnet: 52.9 (#164), Claude Fable 5: 75.9 (#5)

Writing & Preference benchmarks
BenchmarkClaude 3.5 SonnetClaude Fable 5
LMArena Text12981491
LMArena Creative Writing12921494
EQ-Bench Creative Writing14511943
LMArena Multi-Turn13261504
Short-Story Creative Writing80.3%—
WildBench79.2%—
EQ-Bench 4—1340
LiveBench Language53.8%—

Frequently asked questions

Is Claude 3.5 Sonnet better than Claude Fable 5?

Claude Fable 5 is the stronger model overall, scoring 66.8 to 34.6 on the Noometry Index.

Is Claude 3.5 Sonnet or Claude Fable 5 better for coding?

Claude Fable 5 scores higher on coding benchmarks: 70.6 versus 39.0 in the Noometry coding category.

How many benchmarks do Claude 3.5 Sonnet and Claude Fable 5 share?

27 benchmarks have published results for both models. Claude 3.5 Sonnet has 60 scored results on Noometry and Claude Fable 5 has 62.

Related comparisons

Go deeper