Model comparison

Granite 3.1 2b Instruct vs o1-pro

Granite 3.1 2b Instruct is the stronger model overall, scoring 33.2 to 31.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 3.1 2b Instruct IBM

33.2

Rank #247 Confirmed

o1-pro OpenAI

31.5

Rank #271 Reported

Summary

  • Granite 3.1 2b Instruct has downloadable open weights; the other is API-only.

Side by side

Granite 3.1 2b Instruct and o1-pro specifications
Granite 3.1 2b Instructo1-pro
ProviderIBMOpenAI
Noometry Index33.231.5
Released—2025-03-19
WeightsOpenProprietary
Context window—200K
Max output—100K
Input $ / M tokens—$150
Output $ / M tokens—$600
Results tracked123

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 3.1 2b Instruct: 33.4 (#257), o1-pro: —

Coding benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
LMArena Coding1149—

Reasoning Granite 3.1 2b Instruct leads

Granite 3.1 2b Instruct: 22.0 (#209), o1-pro: 20.4 (#239)

Reasoning benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
ARC-AGI-1—23.3%
EnigmaEval—6.1%
LMArena Hard Prompts1138—

Math Not comparable

Granite 3.1 2b Instruct: 33.1 (#206), o1-pro: —

Math benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
LMArena Math1159—

Knowledge Granite 3.1 2b Instruct leads

Granite 3.1 2b Instruct: 30.8 (#224), o1-pro: 29.7 (#234)

Knowledge benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
Humanity's Last Exam—8.1%
LMArena Expert1131—

Multilingual Not comparable

Granite 3.1 2b Instruct: 29.1 (#269), o1-pro: —

Multilingual benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
LMArena Non-English1068—
LMArena Chinese1139—
LMArena Russian1063—

Instruction Following Not comparable

Granite 3.1 2b Instruct: 57.7 (#264), o1-pro: —

Instruction Following benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
LMArena Instruction Following1116—

Long Context Not comparable

Granite 3.1 2b Instruct: 35.0 (#244), o1-pro: —

Long Context benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
LMArena Longer Query1155—

Writing & Preference Not comparable

Granite 3.1 2b Instruct: 34.1 (#274), o1-pro: —

Writing & Preference benchmarks
BenchmarkGranite 3.1 2b Instructo1-pro
LMArena Text1127—
LMArena Creative Writing1116—
LMArena Multi-Turn1099—

Frequently asked questions

Is Granite 3.1 2b Instruct better than o1-pro?

Granite 3.1 2b Instruct is the stronger model overall, scoring 33.2 to 31.5 on the Noometry Index.

How many benchmarks do Granite 3.1 2b Instruct and o1-pro share?

0 benchmarks have published results for both models. Granite 3.1 2b Instruct has 12 scored results on Noometry and o1-pro has 3.

Related comparisons

Go deeper