Anthropic, proprietary

Claude 2.1

Claude 2.1 by Anthropic ranks 345th of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.2. Its strongest category is reasoning, where it ranks 221st.

Last verified

Specifications

Noometry rank
#345 of 354
Index score
25.2
Evidence
Reported 7 results
Provider
Anthropic
Released
November 21, 2023
Weights
Proprietary
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Claude 2.1 category scores
  1. Coding 26.2
  2. Reasoning 21.4
  3. Math 10.2
  4. Knowledge 15.4
Claude 2.1 category ranks
CategoryScoreRankResults
Coding26.2#3271
Reasoning21.4#2211
Math10.2#3151
Knowledge15.4#2921

Strengths and weaknesses

Categories where Claude 2.1 places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Claude 2.1: strongest categories
CategoryScorevs medianRank
Reasoning21.4−2.2#221 of 350, top 64%
Knowledge15.4−21.9#292 of 314, top 93%

Weakest categories

Claude 2.1: weakest categories
CategoryScorevs medianRank
Math10.2−26.4#315 of 327, top 97%
Coding26.2−12.5#327 of 340, top 97%

Closest competitors

The models ranked just above and below Claude 2.1. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Claude 2.1
ModelRankScoreBlended $/MSpeed
Gemma 2 9B#34125.9——Compare
Dolly 2.0-12b#34225.5——Compare
GPT-4o mini#34325.5$0.26120Compare
Llama 3-8B#34425.5——Compare
Claude 2#34625.0——Compare
DeepSeek LLM 67B#34724.9——Compare
Llama 13b#34824.4——Compare
Llama 2-70B#34924.4——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Claude 2.1 Coding benchmark results
BenchmarkScorePositionSettingSourceDate
WeirdML7.1%#116 of 119, top 98%Epoch AI

Reasoning

Claude 2.1 Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
DTBench51%#133 of 151, top 89%Epoch AI
Epoch Capabilities Index119.27#172 of 213, top 81%Epoch AI2023-11-21
ForecastBench54.2#68 of 72, top 95%Epoch AI

Math

Claude 2.1 Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20251.9%#161 of 173, top 94%Epoch AI2025-03-07

Knowledge

Claude 2.1 Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond33%#164 of 186, top 89%Epoch AI2025-01-27
MMLU73.5%#37 of 81, top 46%Epoch AI

Compare Claude 2.1

Other Anthropic models

Frequently asked questions

How good is Claude 2.1?

Claude 2.1 by Anthropic ranks 345th of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.2. Its strongest category is reasoning, where it ranks 221st.

Is Claude 2.1 open source?

No. Claude 2.1 is proprietary and available only through Anthropic's API and partner platforms.

What are Claude 2.1's strengths and weaknesses?

Relative to other ranked models, Claude 2.1 places best in reasoning, knowledge and lowest in math, coding.

What is Claude 2.1 best at?

Its best category is reasoning, where it ranks 221st on Noometry.