Anthropic, proprietary

Claude 2

Claude 2 by Anthropic ranks 346th of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.0. Its strongest category is reasoning, where it ranks 216th.

Last verified

Specifications

Noometry rank
#346 of 354
Index score
25.0
Evidence
Reported 8 results
Provider
Anthropic
Released
July 11, 2023
Weights
Proprietary
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Claude 2 category scores
  1. Reasoning 21.7
  2. Math 9.3
  3. Knowledge 16.9
Claude 2 category ranks
CategoryScoreRankResults
Reasoning21.7#2161
Math9.3#3202
Knowledge16.9#2871

Strengths and weaknesses

Categories where Claude 2 places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Claude 2: strongest categories
CategoryScorevs medianRank
Reasoning21.7−1.9#216 of 350, top 62%

Weakest categories

Claude 2: weakest categories
CategoryScorevs medianRank
Math9.3−27.3#320 of 327, top 98%

Closest competitors

The models ranked just above and below Claude 2. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Claude 2
ModelRankScoreBlended $/MSpeed
Dolly 2.0-12b#34225.5——Compare
GPT-4o mini#34325.5$0.26120Compare
Llama 3-8B#34425.5——Compare
Claude 2.1#34525.2——Compare
DeepSeek LLM 67B#34724.9——Compare
Llama 13b#34824.4——Compare
Llama 2-70B#34924.4——Compare
GPT-3.5-turbo#35023.2$0.75—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Claude 2 Coding benchmark results
BenchmarkScorePositionSettingSourceDate
HumanEval+61.6%#27 of 45, top 60%mar 2024EvalPlus

Reasoning

Claude 2 Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
DTBench51.9%#131 of 151, top 87%Epoch AI
Epoch Capabilities Index120.13#168 of 213, top 79%Epoch AI2023-07-11

Math

Claude 2 Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20252.5%#157 of 173, top 91%Epoch AI2025-03-07
MATH Level 511.7%#70 of 79, top 89%Epoch AI2025-01-27

Knowledge

Claude 2 Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond34.7%#159 of 186, top 86%Epoch AI2025-01-27
MMLU78.5%#23 of 81, top 29%Epoch AI
TriviaQA87.5%#2 of 25, top 8%Epoch AI

Compare Claude 2

Other Anthropic models

Frequently asked questions

How good is Claude 2?

Claude 2 by Anthropic ranks 346th of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.0. Its strongest category is reasoning, where it ranks 216th.

Is Claude 2 open source?

No. Claude 2 is proprietary and available only through Anthropic's API and partner platforms.

What are Claude 2's strengths and weaknesses?

Relative to other ranked models, Claude 2 places best in reasoning and lowest in math.

What is Claude 2 best at?

Its best category is reasoning, where it ranks 216th on Noometry.