Anthropic, proprietary
Claude 2
Claude 2 by Anthropic ranks 346th of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.0. Its strongest category is reasoning, where it ranks 216th.
Last verified
Specifications
- Noometry rank
- #346 of 354
- Index score
- 25.0
- Evidence
- Reported 8 results
- Provider
- Anthropic
- Released
- July 11, 2023
- Weights
- Proprietary
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Reasoning 21.7
- Math 9.3
- Knowledge 16.9
Strengths and weaknesses
Categories where Claude 2 places highest and lowest among the models ranked in each, with its score against that category's median.
Closest competitors
The models ranked just above and below Claude 2. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Dolly 2.0-12b | #342 | 25.5 | — | — | Compare |
| GPT-4o mini | #343 | 25.5 | $0.26 | 120 | Compare |
| Llama 3-8B | #344 | 25.5 | — | — | Compare |
| Claude 2.1 | #345 | 25.2 | — | — | Compare |
| DeepSeek LLM 67B | #347 | 24.9 | — | — | Compare |
| Llama 13b | #348 | 24.4 | — | — | Compare |
| Llama 2-70B | #349 | 24.4 | — | — | Compare |
| GPT-3.5-turbo | #350 | 23.2 | $0.75 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| HumanEval+ | 61.6% | #27 of 45, top 60% | mar 2024 | EvalPlus |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| DTBench | 51.9% | #131 of 151, top 87% | Epoch AI | ||
| Epoch Capabilities Index | 120.13 | #168 of 213, top 79% | Epoch AI | 2023-07-11 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 2.5% | #157 of 173, top 91% | Epoch AI | 2025-03-07 | |
| MATH Level 5 | 11.7% | #70 of 79, top 89% | Epoch AI | 2025-01-27 |
Knowledge
Compare Claude 2
- Claude 2 vs Claude 2.1
- Claude 2 vs DeepSeek LLM 67B
- Claude 2 vs Llama 3-8B
- Claude 2 vs Llama 13b
- Claude 2 vs GPT-4o mini
- Claude 2 vs Llama 2-70B
- Claude 2 vs GPT-6 Astra
- Claude 2 vs Gemini 3.8 Flash
- Claude 2 vs Kimi K3
- Claude 2 vs Grok 4.6
- Claude 2 vs Qwen3.8 Max
- Claude 2 vs GLM-5.3
- Claude 2 vs Muse Spark 1.3
- Claude 2 vs DeepSeek V4 Pro
Other Anthropic models
- Claude Fable 5.169.0
- Claude Opus 5.568.6
- Claude Opus 567.8
- Claude Fable 566.8
- Claude Sonnet 5.561.9
- Claude Opus 4.860.7
- Claude Opus 4.758.3
- Claude Opus 4.658.2
Frequently asked questions
How good is Claude 2?
Claude 2 by Anthropic ranks 346th of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.0. Its strongest category is reasoning, where it ranks 216th.
Is Claude 2 open source?
No. Claude 2 is proprietary and available only through Anthropic's API and partner platforms.
What are Claude 2's strengths and weaknesses?
Relative to other ranked models, Claude 2 places best in reasoning and lowest in math.
What is Claude 2 best at?
Its best category is reasoning, where it ranks 216th on Noometry.