Model comparison

Grok 2 Mini 2024 08 13 vs Grok 4.1 Fast

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 37.7 on the Noometry Index.

Last verified . 17 shared benchmarks.

Grok 2 Mini 2024 08 13 xAI

37.7

Rank #198 Confirmed

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Grok 2 Mini 2024 08 13 scores higher in 3 categories and Grok 4.1 Fast in 5 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 24.8.

Side by side

Grok 2 Mini 2024 08 13 and Grok 4.1 Fast specifications
Grok 2 Mini 2024 08 13Grok 4.1 Fast
ProviderxAIxAI
Noometry Index37.741.4
Released2024-08-132025-06-27
WeightsProprietaryProprietary
Context window—128K
Max output—30K
Input $ / M tokens—$0.20
Output $ / M tokens—$0.50
Results tracked1732

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok 2 Mini 2024 08 13 leads

Grok 2 Mini 2024 08 13: 37.0 (#199), Grok 4.1 Fast: 34.1 (#245)

Coding benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Coding12691411
LMArena WebDev—1242
ALE-Bench—394.93

Agentic & Tool Use Not comparable

Grok 2 Mini 2024 08 13: —, Grok 4.1 Fast: 36.3 (#39)

Agentic & Tool Use benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
Berkeley Function Calling Leaderboard—69.6%
τ²-bench Banking—13.1%
LMArena Search—1171
Vending-Bench 2—1,107

Reasoning Grok 4.1 Fast leads

Grok 2 Mini 2024 08 13: 24.8 (#159), Grok 4.1 Fast: 43.4 (#49)

Reasoning benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Hard Prompts12551407
SimpleBench—56%
NYT Connections (extended)—87.4%
DTBench—87.7%
ForecastBench—61

Math Grok 2 Mini 2024 08 13 leads

Grok 2 Mini 2024 08 13: 35.4 (#185), Grok 4.1 Fast: 31.9 (#221)

Math benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Math12651408
MathArena Final-Answer Competitions—60.9%
ProofBench—4%

Knowledge Too close to call

Grok 2 Mini 2024 08 13: 33.9 (#200), Grok 4.1 Fast: 33.1 (#207)

Knowledge benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Expert12381399
Vectara Hallucination Rate—17.8%

Multimodal Not comparable

Grok 2 Mini 2024 08 13: —, Grok 4.1 Fast: 37.0 (#76)

Multimodal benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Vision—1201

Multilingual Grok 4.1 Fast leads

Grok 2 Mini 2024 08 13: 41.4 (#208), Grok 4.1 Fast: 51.0 (#114)

Multilingual benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Non-English12571391
LMArena Chinese12621441
LMArena French12861415
LMArena German12731404
LMArena Japanese12131349
LMArena Korean11951361
LMArena Russian12611387
LMArena Spanish12791413

Instruction Following Grok 4.1 Fast leads

Grok 2 Mini 2024 08 13: 65.5 (#220), Grok 4.1 Fast: 72.7 (#133)

Instruction Following benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Instruction Following12451376

Long Context Grok 4.1 Fast leads

Grok 2 Mini 2024 08 13: 38.4 (#196), Grok 4.1 Fast: 42.4 (#126)

Long Context benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Longer Query12661390

Writing & Preference Grok 4.1 Fast leads

Grok 2 Mini 2024 08 13: 47.3 (#211), Grok 4.1 Fast: 57.2 (#131)

Writing & Preference benchmarks
BenchmarkGrok 2 Mini 2024 08 13Grok 4.1 Fast
LMArena Text12811408
LMArena Creative Writing12421394
LMArena Multi-Turn12651389
EQ-Bench Creative Writing—1327

Frequently asked questions

Is Grok 2 Mini 2024 08 13 better than Grok 4.1 Fast?

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 37.7 on the Noometry Index.

Is Grok 2 Mini 2024 08 13 or Grok 4.1 Fast better for coding?

Grok 2 Mini 2024 08 13 scores higher on coding benchmarks: 37.0 versus 34.1 in the Noometry coding category.

How many benchmarks do Grok 2 Mini 2024 08 13 and Grok 4.1 Fast share?

17 benchmarks have published results for both models. Grok 2 Mini 2024 08 13 has 17 scored results on Noometry and Grok 4.1 Fast has 32.

Related comparisons

Go deeper