Allen Institute for AI (Ai2), open weights

Olmo 3.1 32b Instruct

Olmo 3.1 32b Instruct by Allen Institute for AI (Ai2) ranks 168th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.4. Its strongest category is reasoning, where it ranks 132nd.

Last verified

Specifications

Noometry rank
#168 of 354
Index score
39.4
Evidence
Confirmed 16 results
Released
Unknown
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Olmo 3.1 32b Instruct category scores
  1. Coding 39.5
  2. Reasoning 26.4
  3. Math 36.3
  4. Knowledge 36.1
  5. Multilingual 42.6
  6. Instruction Following 68.6
  7. Long Context 39.9
  8. Writing & Preference 50.2
Olmo 3.1 32b Instruct category ranks
CategoryScoreRankResults
Coding39.5#1571
Reasoning26.4#1321
Math36.3#1671
Knowledge36.1#1751
Multilingual42.6#1911
Instruction Following68.6#1871
Long Context39.9#1661
Writing & Preference50.2#1853

Strengths and weaknesses

Categories where Olmo 3.1 32b Instruct places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Olmo 3.1 32b Instruct: strongest categories
CategoryScorevs medianRank
Reasoning26.4+2.8#132 of 350, top 38%
Coding39.5+0.8#157 of 340, top 47%
Math36.3−0.3#167 of 327, top 52%

Weakest categories

Olmo 3.1 32b Instruct: weakest categories
CategoryScorevs medianRank
Multilingual42.6−4.8#191 of 297, top 65%
Instruction Following68.6−2.7#187 of 305, top 62%
Writing & Preference50.2−3.6#185 of 312, top 60%

Closest competitors

The models ranked just above and below Olmo 3.1 32b Instruct. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Olmo 3.1 32b Instruct
ModelRankScoreBlended $/MSpeed
Claude 3.7 Sonnet#16439.5——Compare
Claude Haiku 4.5#16539.5$2—Compare
DeepSeek-V3#16639.5$0.4173Compare
Grok 4 Fast#16739.4—577Compare
Granite 4.2 3b#16939.4——Compare
Gemini 2.5 Flash#17039.3$0.85152Compare
Step 2 16k Exp 202412#17139.2——Compare
Qwen3 32B#17239.2$1.2286Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Olmo 3.1 32b Instruct Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1347#175 of 294, top 60%LMArena2026-10-08

Reasoning

Olmo 3.1 32b Instruct Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1322#179 of 297, top 61%LMArena2026-10-08

Math

Olmo 3.1 32b Instruct Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1305#180 of 285, top 64%LMArena2026-10-08

Knowledge

Olmo 3.1 32b Instruct Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1308#174 of 273, top 64%LMArena2026-10-08

Multilingual

Olmo 3.1 32b Instruct Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1275#191 of 297, top 65%LMArena2026-10-08
LMArena Chinese1304#187 of 285, top 66%LMArena2026-10-08
LMArena French1328#154 of 223, top 70%LMArena2026-10-08
LMArena German1282#159 of 231, top 69%LMArena2026-10-08
LMArena Korean1206#158 of 213, top 75%LMArena2026-10-08
LMArena Russian1268#197 of 283, top 70%LMArena2026-10-08
LMArena Spanish1336#151 of 226, top 67%LMArena2026-10-08

Instruction Following

Olmo 3.1 32b Instruct Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1299#180 of 298, top 61%LMArena2026-10-08

Long Context

Olmo 3.1 32b Instruct Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1312#178 of 291, top 62%LMArena2026-10-08

Writing & Preference

Olmo 3.1 32b Instruct Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1311#185 of 297, top 63%LMArena2026-10-08
LMArena Creative Writing1264#196 of 295, top 67%LMArena2026-10-08
LMArena Multi-Turn1309#183 of 295, top 63%LMArena2026-10-08

Compare Olmo 3.1 32b Instruct

Other Allen Institute for AI (Ai2) models

Frequently asked questions

How good is Olmo 3.1 32b Instruct?

Olmo 3.1 32b Instruct by Allen Institute for AI (Ai2) ranks 168th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.4. Its strongest category is reasoning, where it ranks 132nd.

Is Olmo 3.1 32b Instruct open source?

Yes. Olmo 3.1 32b Instruct's weights are downloadable; check the license for commercial terms.

What are Olmo 3.1 32b Instruct's strengths and weaknesses?

Relative to other ranked models, Olmo 3.1 32b Instruct places best in reasoning, coding, math and lowest in multilingual, instruction following, writing & preference.

What is Olmo 3.1 32b Instruct best at?

Its best category is reasoning, where it ranks 132nd on Noometry.