Model comparison

Claude 2 vs o1-pro

o1-pro is the stronger model overall, scoring 31.5 to 25.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 2 Anthropic

25.0

Rank #346 Reported

o1-pro OpenAI

31.5

Rank #271 Reported

Summary

  • The widest gap is in knowledge, where o1-pro leads 29.7 to 16.9.

Side by side

Claude 2 and o1-pro specifications
Claude 2o1-pro
ProviderAnthropicOpenAI
Noometry Index25.031.5
Released2023-07-112025-03-19
WeightsProprietaryProprietary
Context window—200K
Max output—100K
Input $ / M tokens—$150
Output $ / M tokens—$600
Results tracked83

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Claude 2: —, o1-pro: —

Coding benchmarks
BenchmarkClaude 2o1-pro
HumanEval+61.6%—

Reasoning Claude 2 leads

Claude 2: 21.7 (#216), o1-pro: 20.4 (#239)

Reasoning benchmarks
BenchmarkClaude 2o1-pro
ARC-AGI-1—23.3%
EnigmaEval—6.1%
DTBench51.9%—
Epoch Capabilities Index120.13—

Math Not comparable

Claude 2: 9.3 (#320), o1-pro: —

Math benchmarks
BenchmarkClaude 2o1-pro
OTIS Mock AIME 2024-20252.5%—
MATH Level 511.7%—

Knowledge o1-pro leads

Claude 2: 16.9 (#287), o1-pro: 29.7 (#234)

Knowledge benchmarks
BenchmarkClaude 2o1-pro
GPQA Diamond34.7%—
Humanity's Last Exam—8.1%
MMLU78.5%—
TriviaQA87.5%—

Frequently asked questions

Is Claude 2 better than o1-pro?

o1-pro is the stronger model overall, scoring 31.5 to 25.0 on the Noometry Index.

How many benchmarks do Claude 2 and o1-pro share?

0 benchmarks have published results for both models. Claude 2 has 8 scored results on Noometry and o1-pro has 3.

Related comparisons

Go deeper