xAI, proprietary

# Grok 4.1

> Grok 4.1 by xAI, released November 2025. Ranked #134 of 354 with a Noometry Index of 41.5. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/grok-4-1
- Last updated: 2026-10-10
- Title: Grok 4.1 Benchmarks, Price & Rank (October 2026) | Noometry

Grok 4.1 by xAI ranks 134th of 354 ranked models on the Noometry Index as of October 2026, with a score of 41.5. Its strongest category is agentic & tool use, where it ranks 49th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #134 of 354
- **Index score:** 41.5
- **Evidence:** Confirmed 19 results
- **Provider:** [xAI](https://noometry.com/providers/xai)
- **Released:** November 17, 2025
- **Weights:** Proprietary
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Grok 4.1 category scores

1.  Coding 33.7
2.  Agentic & Tool Use 34.1
3.  Reasoning 29.5
4.  Math 38.9
5.  Knowledge 39.5
6.  Multilingual 53.4
7.  Instruction Following 73.8
8.  Long Context 43.2
9.  Writing & Preference 62.4
10.  020406080

Grok 4.1 category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 33.7 | #253 | 2 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 34.1 | #49 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 29.5 | #91 | 1 |
| [Math](https://noometry.com/best/math) | 38.9 | #120 | 1 |
| [Knowledge](https://noometry.com/best/knowledge) | 39.5 | #133 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 53.4 | #68 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 73.8 | #111 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 43.2 | #100 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 62.4 | #75 | 3 |

## Strengths and weaknesses

Categories where Grok 4.1 places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Grok 4.1: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Multilingual](https://noometry.com/best/multilingual) | 53.4 | +6.0 | #68 of 297, top 23% |
| [Writing & Preference](https://noometry.com/best/writing) | 62.4 | +8.6 | #75 of 312, top 25% |
| [Reasoning](https://noometry.com/best/reasoning) | 29.5 | +5.9 | #91 of 350, top 26% |

### Weakest categories

Grok 4.1: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 33.7 | −5.0 | #253 of 340, top 75% |
| [Knowledge](https://noometry.com/best/knowledge) | 39.5 | +2.2 | #133 of 314, top 43% |
| [Math](https://noometry.com/best/math) | 38.9 | +2.4 | #120 of 327, top 37% |

## Closest competitors

The models ranked just above and below Grok 4.1. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Grok 4.1
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Granite 4.2 30b](https://noometry.com/models/granite-4-2-30b) | #130 | 41.8 | — | — | [Compare](https://noometry.com/compare/granite-4-2-30b-vs-grok-4-1) |
| [Muse Glimmer](https://noometry.com/models/muse-glimmer) | #131 | 41.7 | — | — | [Compare](https://noometry.com/compare/grok-4-1-vs-muse-glimmer) |
| [o4-mini](https://noometry.com/models/o4-mini) | #132 | 41.6 | $1.93 | 6 | [Compare](https://noometry.com/compare/grok-4-1-vs-o4-mini) |
| [Gemini 3.5 Flash Lite](https://noometry.com/models/gemini-3-5-flash-lite) | #133 | 41.5 | $0.85 | — | [Compare](https://noometry.com/compare/gemini-3-5-flash-lite-vs-grok-4-1) |
| [GLM-4.6](https://noometry.com/models/glm-4-6) | #135 | 41.4 | $1 | 12 | [Compare](https://noometry.com/compare/glm-4-6-vs-grok-4-1) |
| [Grok 4.1 Fast](https://noometry.com/models/grok-4-1-fast) | #136 | 41.4 | $0.28 | — | [Compare](https://noometry.com/compare/grok-4-1-vs-grok-4-1-fast) |
| [GLM-4.6V](https://noometry.com/models/glm-4-6v) | #137 | 41.3 | $0.45 | — | [Compare](https://noometry.com/compare/glm-4-6v-vs-grok-4-1) |
| [MiMo-V2-Flash](https://noometry.com/models/mimo-v2-flash) | #138 | 41.3 | $0.18 | — | [Compare](https://noometry.com/compare/grok-4-1-vs-mimo-v2-flash) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Grok 4.1 Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev) | 1214 | #108 of 113, top 96% | thinking | [LMArena](https://lmarena.ai/leaderboard/webdev) | 2026-10-08 |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1445 | #93 of 294, top 32% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Agentic & Tool Use

Grok 4.1 Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Cybench](https://noometry.com/benchmarks/cybench) | 39% | #6 of 21, top 29% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Reasoning

Grok 4.1 Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1435 | #85 of 297, top 29% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Math

Grok 4.1 Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1422 | #99 of 285, top 35% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Grok 4.1 Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1417 | #109 of 273, top 40% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multilingual

Grok 4.1 Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1425 | #68 of 297, top 23% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1465 | #81 of 285, top 29% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1448 | #74 of 223, top 34% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1446 | #51 of 231, top 23% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1397 | #62 of 211, top 30% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1407 | #46 of 213, top 22% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1434 | #61 of 283, top 22% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1438 | #69 of 226, top 31% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Grok 4.1 Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1400 | #106 of 298, top 36% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Grok 4.1 Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1416 | #95 of 291, top 33% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Grok 4.1 Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1437 | #70 of 297, top 24% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1411 | #63 of 295, top 22% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1437 | #75 of 295, top 26% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## Compare Grok 4.1

-   [Grok 4.1 vs Grok 4](https://noometry.com/compare/grok-4-vs-grok-4-1)
-   [Grok 4.1 vs Gemini 3.5 Flash Lite](https://noometry.com/compare/gemini-3-5-flash-lite-vs-grok-4-1)
-   [Grok 4.1 vs GLM-4.6](https://noometry.com/compare/glm-4-6-vs-grok-4-1)
-   [Grok 4.1 vs o4-mini](https://noometry.com/compare/grok-4-1-vs-o4-mini)
-   [Grok 4.1 vs Grok 4.1 Fast](https://noometry.com/compare/grok-4-1-vs-grok-4-1-fast)
-   [Grok 4.1 vs Muse Glimmer](https://noometry.com/compare/grok-4-1-vs-muse-glimmer)
-   [Grok 4.1 vs GLM-4.6V](https://noometry.com/compare/glm-4-6v-vs-grok-4-1)
-   [Grok 4.1 vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-grok-4-1)
-   [Grok 4.1 vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-grok-4-1)
-   [Grok 4.1 vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-grok-4-1)
-   [Grok 4.1 vs Kimi K3](https://noometry.com/compare/grok-4-1-vs-kimi-k3)
-   [Grok 4.1 vs Qwen3.8 Max](https://noometry.com/compare/grok-4-1-vs-qwen3-8-max)
-   [Grok 4.1 vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-grok-4-1)
-   [Grok 4.1 vs Muse Spark 1.3](https://noometry.com/compare/grok-4-1-vs-muse-spark-1-3)

## Other xAI models

-   [Grok 4.6](https://noometry.com/models/grok-4-6)56.9
-   [Grok 4.5](https://noometry.com/models/grok-4-5)55.0
-   [Grok 4.7](https://noometry.com/models/grok-4-7)53.1
-   [Grok 4.20 (Non-Reasoning)](https://noometry.com/models/grok-4-20)48.6
-   [Grok 4](https://noometry.com/models/grok-4)48.1
-   [Grok 4.20 Multi-Agent](https://noometry.com/models/grok-4-20-multi-agent)46.2
-   [Grok 4.3](https://noometry.com/models/grok-4-3)43.8
-   [Grok 4.1 Fast](https://noometry.com/models/grok-4-1-fast)41.4

## Frequently asked questions

### How good is Grok 4.1?

Grok 4.1 by xAI ranks 134th of 354 ranked models on the Noometry Index as of October 2026, with a score of 41.5. Its strongest category is agentic & tool use, where it ranks 49th.

### Is Grok 4.1 open source?

No. Grok 4.1 is proprietary and available only through xAI's API and partner platforms.

### What are Grok 4.1's strengths and weaknesses?

Relative to other ranked models, Grok 4.1 places best in multilingual, writing & preference, reasoning and lowest in coding, knowledge, math.

### What is Grok 4.1 best at?

Its best category is agentic & tool use, where it ranks 49th on Noometry.

### Cite this page

Noometry. (2026). Grok 4.1 benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/grok-4-1

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/grok-4-1.md).
