Model comparison

Claude Instant vs Claude Sonnet 5.5

Claude Sonnet 5.5 has enough public results to be ranked (#10); Claude Instant does not yet, so treat this comparison as directional.

Last verified . 1 shared benchmarks.

Claude Instant Anthropic

29.5

Unranked Sparse

Claude Sonnet 5.5 Anthropic

61.9

Rank #10 Confirmed

Summary

  • They share 1 benchmark with published results for both. Claude Instant scores higher in 0 categories and Claude Sonnet 5.5 in 1 category; one gap is clear of the uncertainty.
  • The widest gap is in reasoning, where Claude Sonnet 5.5 leads 54.0 to 19.6.

Side by side

Claude Instant and Claude Sonnet 5.5 specifications
Claude InstantClaude Sonnet 5.5
ProviderAnthropicAnthropic
Noometry Index29.561.9
Released2023-08-092026-09-28
WeightsProprietaryProprietary
Context window—1M
Max output—128K
Input $ / M tokens—$2
Output $ / M tokens—$10
Results tracked732

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Claude Instant: —, Claude Sonnet 5.5: 67.3 (#6)

Coding benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
FrontierCode—52.1%
CursorBench—55.5%
LMArena WebDev—1774
FrontierSWE—61.9%
SciCode—61%
LMArena Coding—1513
ALE-Bench—1,819
HumanEval+50.6%—

Agentic & Tool Use Not comparable

Claude Instant: —, Claude Sonnet 5.5: 45.0 (#16)

Agentic & Tool Use benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
APEX-Agents—75.5%

Reasoning Claude Sonnet 5.5 leads

Claude Instant: 19.6, Claude Sonnet 5.5: 54.0 (#28)

Reasoning benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
Epoch Capabilities Index120.28165.03
NYT Connections (extended)—80.5%
CritPt—31.4%
LMArena Hard Prompts—1495
Mystery Game Puzzles—65%
DTBench45.8%—

Math Not comparable

Claude Instant: —, Claude Sonnet 5.5: 87.9 (#6)

Math benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
FrontierMath (Tiers 1-3)—88.8%
FrontierMath Tier 4—80.5%
OTIS Mock AIME 2024-2025—100%
ProofBench—100%
LMArena Math—1510
FrontierMath Erdős—2.9%
GSM8K86.7%—

Knowledge Not comparable

Claude Instant: —, Claude Sonnet 5.5: 66.0 (#12)

Knowledge benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
GPQA Diamond—95.6%
SimpleQA Verified—46.5%
LMArena Expert—1540
ARC (AI2) Challenge86.3%—
MMLU73.4%—
TriviaQA78.9%—

Multimodal Not comparable

Claude Instant: —, Claude Sonnet 5.5: 51.5 (#6)

Multimodal benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
LMArena Vision—1289
Furniture Assembly—75%

Multilingual Not comparable

Claude Instant: —, Claude Sonnet 5.5: 55.3 (#30)

Multilingual benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
LMArena Non-English—1452
LMArena Chinese—1522
LMArena Russian—1451

Instruction Following Not comparable

Claude Instant: —, Claude Sonnet 5.5: 78.3 (#11)

Instruction Following benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
LMArena Instruction Following—1495

Long Context Not comparable

Claude Instant: —, Claude Sonnet 5.5: 45.9 (#28)

Long Context benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
LMArena Longer Query—1498

Writing & Preference Not comparable

Claude Instant: —, Claude Sonnet 5.5: 66.0 (#40)

Writing & Preference benchmarks
BenchmarkClaude InstantClaude Sonnet 5.5
LMArena Text—1471
LMArena Creative Writing—1465
LMArena Multi-Turn—1474

Frequently asked questions

Is Claude Instant better than Claude Sonnet 5.5?

Claude Sonnet 5.5 has enough public results to be ranked (#10); Claude Instant does not yet, so treat this comparison as directional.

How many benchmarks do Claude Instant and Claude Sonnet 5.5 share?

1 benchmark has published results for both models. Claude Instant has 7 scored results on Noometry and Claude Sonnet 5.5 has 32.

Related comparisons

Go deeper