Allen Institute for AI (Ai2), open weights

# Olmo 3.1 32b Think

> Olmo 3.1 32b Think by Allen Institute for AI (Ai2). Ranked #191 of 354 with a Noometry Index of 37.9. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/olmo-3-1-32b-think
- Last updated: 2026-10-10
- Title: Olmo 3.1 32b Think Benchmarks, Price & Rank (October 2026)

Olmo 3.1 32b Think by Allen Institute for AI (Ai2) ranks 191st of 354 ranked models on the Noometry Index as of October 2026, with a score of 37.9. Its strongest category is reasoning, where it ranks 150th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #191 of 354
- **Index score:** 37.9
- **Evidence:** Confirmed 15 results
- **Provider:** [![](/logos/ai2.svg) Allen Institute for AI (Ai2)](https://noometry.com/providers/ai2)
- **Released:** Unknown
- **Weights:** Open weights
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Olmo 3.1 32b Think category scores

1.  Coding 37.7
2.  Reasoning 25.2
3.  Math 36.3
4.  Knowledge 35.7
5.  Multilingual 38.1
6.  Instruction Following 65.6
7.  Long Context 38.6
8.  Writing & Preference 46.2
9.  020406080

Olmo 3.1 32b Think category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 37.7 | #189 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 25.2 | #150 | 1 |
| [Math](https://noometry.com/best/math) | 36.3 | #168 | 1 |
| [Knowledge](https://noometry.com/best/knowledge) | 35.7 | #181 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 38.1 | #231 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 65.6 | #218 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 38.6 | #195 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 46.2 | #220 | 3 |

## Strengths and weaknesses

Categories where Olmo 3.1 32b Think places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Olmo 3.1 32b Think: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 25.2 | +1.6 | #150 of 350, top 43% |
| [Math](https://noometry.com/best/math) | 36.3 | −0.3 | #168 of 327, top 52% |
| [Coding](https://noometry.com/best/coding) | 37.7 | −1.0 | #189 of 340, top 56% |

### Weakest categories

Olmo 3.1 32b Think: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Multilingual](https://noometry.com/best/multilingual) | 38.1 | −9.3 | #231 of 297, top 78% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 65.6 | −5.7 | #218 of 305, top 72% |
| [Writing & Preference](https://noometry.com/best/writing) | 46.2 | −7.5 | #220 of 312, top 71% |

## Closest competitors

The models ranked just above and below Olmo 3.1 32b Think. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Olmo 3.1 32b Think
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Sonar](https://noometry.com/models/sonar) | #187 | 38.5 | $1 | — | [Compare](https://noometry.com/compare/olmo-3-1-32b-think-vs-sonar) |
| [MiniMax-M2.5](https://noometry.com/models/minimax-m2-5) | #188 | 38.3 | $0.52 | 49 | [Compare](https://noometry.com/compare/minimax-m2-5-vs-olmo-3-1-32b-think) |
| [Nova Premier 1.0](https://noometry.com/models/nova-premier-1-0) | #189 | 38.3 | $5 | 9 | [Compare](https://noometry.com/compare/nova-premier-1-0-vs-olmo-3-1-32b-think) |
| [Qwen3-Coder 480B-A35B Instruct](https://noometry.com/models/qwen3-coder-480b-a35b-instruct) | #190 | 38.1 | $3 | 67 | [Compare](https://noometry.com/compare/olmo-3-1-32b-think-vs-qwen3-coder-480b-a35b-instruct) |
| [GPT-5-Codex](https://noometry.com/models/gpt-5-codex) | #192 | 37.9 | $3.44 | 95 | [Compare](https://noometry.com/compare/gpt-5-codex-vs-olmo-3-1-32b-think) |
| [Hunyuan Standard 2025 02 10](https://noometry.com/models/hunyuan-standard) | #193 | 37.9 | — | — | [Compare](https://noometry.com/compare/hunyuan-standard-vs-olmo-3-1-32b-think) |
| [Gemini 2.0 Flash-Lite](https://noometry.com/models/gemini-2-0-flash-lite) | #194 | 37.8 | — | — | [Compare](https://noometry.com/compare/gemini-2-0-flash-lite-vs-olmo-3-1-32b-think) |
| [DeepSeek-R1-Distill-Llama-70B](https://noometry.com/models/deepseek-r1-distill-llama-70b) | #195 | 37.8 | — | 18 | [Compare](https://noometry.com/compare/deepseek-r1-distill-llama-70b-vs-olmo-3-1-32b-think) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Olmo 3.1 32b Think Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1291 | #203 of 294, top 70% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Reasoning

Olmo 3.1 32b Think Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1272 | #203 of 297, top 69% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Math

Olmo 3.1 32b Think Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1305 | #181 of 285, top 64% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Olmo 3.1 32b Think Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1295 | #179 of 273, top 66% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multilingual

Olmo 3.1 32b Think Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1209 | #231 of 297, top 78% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1242 | #214 of 285, top 76% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1260 | #180 of 223, top 81% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1262 | #168 of 231, top 73% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1193 | #237 of 283, top 84% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1289 | #166 of 226, top 74% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Olmo 3.1 32b Think Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1247 | #216 of 298, top 73% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Olmo 3.1 32b Think Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1272 | #209 of 291, top 72% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Olmo 3.1 32b Think Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1272 | #216 of 297, top 73% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1226 | #222 of 295, top 76% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1252 | #223 of 295, top 76% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## Compare Olmo 3.1 32b Think

-   [Olmo 3.1 32b Think vs Qwen3-Coder 480B-A35B Instruct](https://noometry.com/compare/olmo-3-1-32b-think-vs-qwen3-coder-480b-a35b-instruct)
-   [Olmo 3.1 32b Think vs GPT-5-Codex](https://noometry.com/compare/gpt-5-codex-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Nova Premier 1.0](https://noometry.com/compare/nova-premier-1-0-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Hunyuan Standard 2025 02 10](https://noometry.com/compare/hunyuan-standard-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs MiniMax-M2.5](https://noometry.com/compare/minimax-m2-5-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Gemini 2.0 Flash-Lite](https://noometry.com/compare/gemini-2-0-flash-lite-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Qwen3.8 Max](https://noometry.com/compare/olmo-3-1-32b-think-vs-qwen3-8-max)
-   [Olmo 3.1 32b Think vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-olmo-3-1-32b-think)
-   [Olmo 3.1 32b Think vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-olmo-3-1-32b-think)

## Other Allen Institute for AI (Ai2) models

-   [Olmo 3.1 32b Instruct](https://noometry.com/models/olmo-3-1-32b-instruct)39.4
-   [Molmo 2 8b](https://noometry.com/models/molmo-2-8b)39.1
-   [Olmo 3 32b Think](https://noometry.com/models/olmo-3-32b-think)38.7
-   [Llama 3.1 Tulu 3 8b](https://noometry.com/models/llama-3-1-tulu-3-8b)35.7
-   [Tulu 3 (Tülu 3) 70B](https://noometry.com/models/tulu-3-70b)33.0
-   [Olmo 2 0325 32b Instruct](https://noometry.com/models/olmo-2-0325-32b-instruct)32.7
-   [Olmo 7b Instruct](https://noometry.com/models/olmo-7b-instruct)30.3
-   [OLMo 2 Furious 13B](https://noometry.com/models/olmo-2-furious-13b)29.7

## Frequently asked questions

### How good is Olmo 3.1 32b Think?

Olmo 3.1 32b Think by Allen Institute for AI (Ai2) ranks 191st of 354 ranked models on the Noometry Index as of October 2026, with a score of 37.9. Its strongest category is reasoning, where it ranks 150th.

### Is Olmo 3.1 32b Think open source?

Yes. Olmo 3.1 32b Think's weights are downloadable; check the license for commercial terms.

### What are Olmo 3.1 32b Think's strengths and weaknesses?

Relative to other ranked models, Olmo 3.1 32b Think places best in reasoning, math, coding and lowest in multilingual, instruction following, writing & preference.

### What is Olmo 3.1 32b Think best at?

Its best category is reasoning, where it ranks 150th on Noometry.

### Cite this page

Noometry. (2026). Olmo 3.1 32b Think benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/olmo-3-1-32b-think

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/olmo-3-1-32b-think.md).
