Model comparison

Grok 4.1 vs MiniMax M1

Grok 4.1 is the stronger model overall, scoring 41.5 to 40.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Grok 4.1 xAI

41.5

Rank #134 Confirmed

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Grok 4.1 scores higher in 7 categories and MiniMax M1 in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Grok 4.1 leads 62.4 to 53.1.
  • MiniMax M1 has downloadable open weights; the other is API-only.

Side by side

Grok 4.1 and MiniMax M1 specifications
Grok 4.1MiniMax M1
ProviderxAIMiniMax
Noometry Index41.540.3
Released2025-11-172025-06-13
WeightsProprietaryOpen
Context window—1M
Max output—40K
Input $ / M tokens—$0.55
Output $ / M tokens—$2.20
Results tracked1918

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiniMax M1 leads

Grok 4.1: 33.7 (#253), MiniMax M1: 39.9 (#153)

Coding benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Coding14451359
LMArena WebDev1214—

Agentic & Tool Use Not comparable

Grok 4.1: 34.1 (#49), MiniMax M1: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1MiniMax M1
Cybench39%—

Reasoning Grok 4.1 leads

Grok 4.1: 29.5 (#91), MiniMax M1: 26.9 (#126)

Reasoning benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Hard Prompts14351339

Math Grok 4.1 leads

Grok 4.1: 38.9 (#120), MiniMax M1: 37.5 (#151)

Math benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Math14221361

Knowledge Grok 4.1 leads

Grok 4.1: 39.5 (#133), MiniMax M1: 36.4 (#170)

Knowledge benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Expert14171317

Multilingual Grok 4.1 leads

Grok 4.1: 53.4 (#68), MiniMax M1: 45.8 (#163)

Multilingual benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Non-English14251319
LMArena Chinese14651360
LMArena French14481370
LMArena German14461350
LMArena Japanese13971217
LMArena Korean14071266
LMArena Russian14341329
LMArena Spanish14381353

Instruction Following Grok 4.1 leads

Grok 4.1: 73.8 (#111), MiniMax M1: 69.3 (#174)

Instruction Following benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Instruction Following14001312

Long Context Grok 4.1 leads

Grok 4.1: 43.2 (#100), MiniMax M1: 41.4 (#141)

Long Context benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Longer Query14161326
Fiction.LiveBench—69.4%

Writing & Preference Grok 4.1 leads

Grok 4.1: 62.4 (#75), MiniMax M1: 53.1 (#161)

Writing & Preference benchmarks
BenchmarkGrok 4.1MiniMax M1
LMArena Text14371343
LMArena Creative Writing14111298
LMArena Multi-Turn14371335

Frequently asked questions

Is Grok 4.1 better than MiniMax M1?

Grok 4.1 is the stronger model overall, scoring 41.5 to 40.3 on the Noometry Index.

Is Grok 4.1 or MiniMax M1 better for coding?

MiniMax M1 scores higher on coding benchmarks: 39.9 versus 33.7 in the Noometry coding category.

How many benchmarks do Grok 4.1 and MiniMax M1 share?

17 benchmarks have published results for both models. Grok 4.1 has 19 scored results on Noometry and MiniMax M1 has 18.

Related comparisons

Go deeper