Model comparison

Claude Fable 5 vs Deepseek Coder v2

Claude Fable 5 is the stronger model overall, scoring 66.8 to 35.9 on the Noometry Index.

Last verified . 17 shared benchmarks.

Claude Fable 5 Anthropic

66.8

Rank #5 Confirmed

Deepseek Coder v2 DeepSeek

35.9

Rank #220 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Claude Fable 5 scores higher in 8 categories and Deepseek Coder v2 in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Claude Fable 5 leads 88.5 to 34.9.
  • Deepseek Coder v2 has downloadable open weights; the other is API-only.

Side by side

Claude Fable 5 and Deepseek Coder v2 specifications
Claude Fable 5Deepseek Coder v2
ProviderAnthropicDeepSeek
Noometry Index66.835.9
Released2026-06-072024-06-17
WeightsProprietaryOpen
Context window1M—
Max output128K—
Input $ / M tokens$10—
Output $ / M tokens$50—
Results tracked6224

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Fable 5 leads

Claude Fable 5: 70.6 (#4), Deepseek Coder v2: 38.1 (#183)

Coding benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Coding15191251
DeepSWE69.9%—
FrontierCode53.5%—
LMArena WebDev1625—
FrontierSWE47%—
SciCode61%—
GSO78.4%—
WeirdML91.9%—
BigCodeBench Instruct—48.2%
MirrorCode63.9%—
BigCodeBench Complete—59.7%
ALE-Bench2,041—
HumanEval+—82.3%
MBPP+—75.1%

Agentic & Tool Use Not comparable

Claude Fable 5: 54.0 (#2), Deepseek Coder v2: —

Agentic & Tool Use benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
APEX-Agents63.6%—
Remote Labor Index16.1%—
τ²-bench Banking39.7%—
PostTrainBench41.8%—
GBAEval74.5%—
GDP.pdf30%—
LMArena Search1230—
Vending-Bench 25,680—

Reasoning Claude Fable 5 leads

Claude Fable 5: 76.8 (#6), Deepseek Coder v2: 23.6 (#176)

Reasoning benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Hard Prompts15081207
ARC-AGI-289.2%—
SimpleBench81.9%—
Kagi LLM Benchmark91.4%—
NYT Connections (extended)92.7%—
ARC-AGI-198.5%—
CritPt28.6%—
Chess Puzzles41%—
EnigmaEval39.3%—
EBR-Bench39.5%—
Mystery Game Puzzles52%—
DTBench98.4%—
LMCA61.1%—
Surface Evolver Bench95%—
Bench to the Future 30.13—
Epoch Capabilities Index162.06—
WinoGrande—83.7%

Math Claude Fable 5 leads

Claude Fable 5: 88.5 (#5), Deepseek Coder v2: 34.9 (#190)

Math benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Math15191241
FrontierMath (Tiers 1-3)87%—
FrontierMath Tier 490.2%—
OTIS Mock AIME 2024-2025100%—
ProofBench95%—
FrontierMath Erdős0%—
GSM8K—94.5%

Knowledge Claude Fable 5 leads

Claude Fable 5: 62.2 (#25), Deepseek Coder v2: 32.3 (#212)

Knowledge benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Expert15341181
GPQA Diamond85.9%—
SimpleQA Verified70.7%—
ARC (AI2) Challenge—64.3%

Multimodal Not comparable

Claude Fable 5: 45.3 (#17), Deepseek Coder v2: —

Multimodal benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Vision1324—
Blueprint-Bench 238.6%—
Furniture Assembly35.8%—
LMArena Document1496—

Multilingual Claude Fable 5 leads

Claude Fable 5: 57.3 (#9), Deepseek Coder v2: 36.3 (#240)

Multilingual benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Non-English14811182
LMArena Chinese15431201
LMArena French15051185
LMArena German14861164
LMArena Japanese15061126
LMArena Korean14881104
LMArena Russian15041188
LMArena Spanish14981153

Instruction Following Claude Fable 5 leads

Claude Fable 5: 78.6 (#8), Deepseek Coder v2: 61.7 (#242)

Instruction Following benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Instruction Following15021180

Long Context Claude Fable 5 leads

Claude Fable 5: 46.3 (#23), Deepseek Coder v2: 37.0 (#224)

Long Context benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Longer Query15091219

Writing & Preference Claude Fable 5 leads

Claude Fable 5: 75.9 (#5), Deepseek Coder v2: 38.2 (#253)

Writing & Preference benchmarks
BenchmarkClaude Fable 5Deepseek Coder v2
LMArena Text14911191
LMArena Creative Writing14941120
LMArena Multi-Turn15041177
EQ-Bench Creative Writing1943—
EQ-Bench 41340—

Frequently asked questions

Is Claude Fable 5 better than Deepseek Coder v2?

Claude Fable 5 is the stronger model overall, scoring 66.8 to 35.9 on the Noometry Index.

Is Claude Fable 5 or Deepseek Coder v2 better for coding?

Claude Fable 5 scores higher on coding benchmarks: 70.6 versus 38.1 in the Noometry coding category.

How many benchmarks do Claude Fable 5 and Deepseek Coder v2 share?

17 benchmarks have published results for both models. Claude Fable 5 has 62 scored results on Noometry and Deepseek Coder v2 has 24.

Related comparisons

Go deeper