IBM, open weights

Granite 3.1 8b Instruct

Granite 3.1 8b Instruct by IBM ranks 258th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.4. Its strongest category is agentic & tool use, where it ranks 120th.

Last verified

Specifications

Noometry rank
#258 of 354
Index score
32.4
Evidence
Confirmed 13 results
Provider
IBM
Released
Unknown
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Granite 3.1 8b Instruct category scores
  1. Coding 34.5
  2. Agentic & Tool Use 24.1
  3. Reasoning 22.1
  4. Math 33.0
  5. Knowledge 31.1
  6. Multilingual 30.9
  7. Instruction Following 58.6
  8. Long Context 35.2
  9. Writing & Preference 35.5
Granite 3.1 8b Instruct category ranks
CategoryScoreRankResults
Coding34.5#2331
Agentic & Tool Use24.1#1201
Reasoning22.1#2071
Math33.0#2091
Knowledge31.1#2201
Multilingual30.9#2601
Instruction Following58.6#2591
Long Context35.2#2411
Writing & Preference35.5#2663

Strengths and weaknesses

Categories where Granite 3.1 8b Instruct places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Granite 3.1 8b Instruct: strongest categories
CategoryScorevs medianRank
Reasoning22.1−1.5#207 of 350, top 60%
Math33.0−3.6#209 of 327, top 64%
Coding34.5−4.2#233 of 340, top 69%

Weakest categories

Granite 3.1 8b Instruct: weakest categories
CategoryScorevs medianRank
Multilingual30.9−16.5#260 of 297, top 88%
Writing & Preference35.5−18.2#266 of 312, top 86%
Instruction Following58.6−12.6#259 of 305, top 85%

Closest competitors

The models ranked just above and below Granite 3.1 8b Instruct. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Granite 3.1 8b Instruct
ModelRankScoreBlended $/MSpeed
Olmo 2 0325 32b Instruct#25432.7——Compare
gpt-oss-20b#25532.5$0.03696Compare
Laguna M.1#25632.5——Compare
Command R+#25732.4$4.38—Compare
Pixtral Large#25932.2$3—Compare
Falcon-180B#26032.2——Compare
Gemini 1.5 Pro (May 2024)#26132.1——Compare
Gemma 3 12B#26232.1$0.075—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Granite 3.1 8b Instruct Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1186#246 of 294, top 84%LMArena2026-10-08

Agentic & Tool Use

Granite 3.1 8b Instruct Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Berkeley Function Calling Leaderboard27.1%#39 of 49, top 80%fcBerkeley Function Calling Leaderboard

Reasoning

Granite 3.1 8b Instruct Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1145#256 of 297, top 87%LMArena2026-10-08

Math

Granite 3.1 8b Instruct Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1152#250 of 285, top 88%LMArena2026-10-08

Knowledge

Granite 3.1 8b Instruct Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1142#241 of 273, top 89%LMArena2026-10-08

Multilingual

Granite 3.1 8b Instruct Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1099#260 of 297, top 88%LMArena2026-10-08
LMArena Chinese1145#248 of 285, top 88%LMArena2026-10-08
LMArena Russian1092#258 of 283, top 92%LMArena2026-10-08

Instruction Following

Granite 3.1 8b Instruct Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1131#257 of 298, top 87%LMArena2026-10-08

Long Context

Granite 3.1 8b Instruct Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1162#250 of 291, top 86%LMArena2026-10-08

Writing & Preference

Granite 3.1 8b Instruct Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1150#258 of 297, top 87%LMArena2026-10-08
LMArena Creative Writing1129#258 of 295, top 88%LMArena2026-10-08
LMArena Multi-Turn1108#265 of 295, top 90%LMArena2026-10-08

Compare Granite 3.1 8b Instruct

Other IBM models

Frequently asked questions

How good is Granite 3.1 8b Instruct?

Granite 3.1 8b Instruct by IBM ranks 258th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.4. Its strongest category is agentic & tool use, where it ranks 120th.

Is Granite 3.1 8b Instruct open source?

Yes. Granite 3.1 8b Instruct's weights are downloadable; check the license for commercial terms.

What are Granite 3.1 8b Instruct's strengths and weaknesses?

Relative to other ranked models, Granite 3.1 8b Instruct places best in reasoning, math, coding and lowest in multilingual, writing & preference, instruction following.

What is Granite 3.1 8b Instruct best at?

Its best category is agentic & tool use, where it ranks 120th on Noometry.