Model comparison

Grok 4.1 Fast vs Grok 4.5

Grok 4.5 is the stronger model overall, scoring 55.0 to 41.4 on the Noometry Index. Grok 4.1 Fast costs 11× less per token, which makes it the better buy when Grok 4.5's lead doesn't matter for your workload.

Last verified . 28 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Grok 4.5 xAI

55.0

Rank #25 Confirmed

Summary

  • They share 28 benchmarks with published results for both. Grok 4.1 Fast scores higher in 0 categories and Grok 4.5 in 10 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Grok 4.5 leads 62.3 to 33.1.
  • The biggest single-benchmark swing is τ²-bench Banking: 13.1% for Grok 4.1 Fast and 47.9% for Grok 4.5.
  • Grok 4.1 Fast is cheaper at $0.20 / $0.50 per million input/output tokens, against $2 / $6 for Grok 4.5.
  • Grok 4.5 accepts more context: 500K tokens versus 128K.

Side by side

Grok 4.1 Fast and Grok 4.5 specifications
Grok 4.1 FastGrok 4.5
ProviderxAIxAI
Noometry Index41.455.0
Released2025-06-272026-07-08
WeightsProprietaryProprietary
Context window128K500K
Max output30K500K
Input $ / M tokens$0.20$2
Output $ / M tokens$0.50$6
Results tracked3252

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok 4.5 leads

Grok 4.1 Fast: 34.1 (#245), Grok 4.5: 52.2 (#35)

Coding benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena WebDev12421553
LMArena Coding14111474
ALE-Bench394.931,309
DeepSWE—53.8%
FrontierCode—42.4%
SciCode—54.1%
WeirdML—46.4%

Agentic & Tool Use Grok 4.5 leads

Grok 4.1 Fast: 36.3 (#39), Grok 4.5: 44.4 (#17)

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
τ²-bench Banking13.1%47.9%
LMArena Search11711213
Vending-Bench 21,1073,887
APEX-Agents—56.2%
Berkeley Function Calling Leaderboard69.6%—
PostTrainBench—23.4%
GBAEval—65.4%
GDP.pdf—14%

Reasoning Grok 4.5 leads

Grok 4.1 Fast: 43.4 (#49), Grok 4.5: 56.1 (#25)

Reasoning benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
SimpleBench56%70%
NYT Connections (extended)87.4%79.9%
LMArena Hard Prompts14071462
DTBench87.7%96.5%
ARC-AGI-2—52.6%
Kagi LLM Benchmark—83.5%
ARC-AGI-1—87.2%
CritPt—15.4%
Chess Puzzles—36%
LMCA—45.2%
Surface Evolver Bench—74.4%
Epoch Capabilities Index—153.92
ForecastBench61—

Math Grok 4.5 leads

Grok 4.1 Fast: 31.9 (#221), Grok 4.5: 60.9 (#35)

Math benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
ProofBench4%31%
LMArena Math14081459
FrontierMath (Tiers 1-3)—57.2%
FrontierMath Tier 4—24.4%
MathArena Final-Answer Competitions60.9%—
OTIS Mock AIME 2024-2025—97.8%

Knowledge Grok 4.5 leads

Grok 4.1 Fast: 33.1 (#207), Grok 4.5: 62.3 (#24)

Knowledge benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena Expert13991466
GPQA Diamond—93.4%
SimpleQA Verified—48.3%
Vectara Hallucination Rate17.8%—

Multimodal Too close to call

Grok 4.1 Fast: 37.0 (#76), Grok 4.5: 37.6 (#72)

Multimodal benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena Vision12011288
Blueprint-Bench 2—27.3%
Furniture Assembly—22.5%
LMArena Document—1452

Multilingual Grok 4.5 leads

Grok 4.1 Fast: 51.0 (#114), Grok 4.5: 54.4 (#42)

Multilingual benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena Non-English13911440
LMArena Chinese14411496
LMArena French14151456
LMArena German14041446
LMArena Japanese13491428
LMArena Korean13611404
LMArena Russian13871448
LMArena Spanish14131450

Instruction Following Grok 4.5 leads

Grok 4.1 Fast: 72.7 (#133), Grok 4.5: 76.0 (#48)

Instruction Following benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena Instruction Following13761446

Long Context Grok 4.5 leads

Grok 4.1 Fast: 42.4 (#126), Grok 4.5: 44.8 (#56)

Long Context benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena Longer Query13901463

Writing & Preference Grok 4.5 leads

Grok 4.1 Fast: 57.2 (#131), Grok 4.5: 65.8 (#42)

Writing & Preference benchmarks
BenchmarkGrok 4.1 FastGrok 4.5
LMArena Text14081448
LMArena Creative Writing13941442
EQ-Bench Creative Writing13271579
LMArena Multi-Turn13891456

Frequently asked questions

Is Grok 4.1 Fast better than Grok 4.5?

Grok 4.5 is the stronger model overall, scoring 55.0 to 41.4 on the Noometry Index. Grok 4.1 Fast costs 11× less per token, which makes it the better buy when Grok 4.5's lead doesn't matter for your workload.

Which is cheaper, Grok 4.1 Fast or Grok 4.5?

Grok 4.1 Fast is cheaper. It lists at $0.20 per million input tokens and $0.50 per million output tokens; Grok 4.5 lists at $2 and $6.

Is Grok 4.1 Fast or Grok 4.5 better for coding?

Grok 4.5 scores higher on coding benchmarks: 52.2 versus 34.1 in the Noometry coding category.

Which has the bigger context window?

Grok 4.5 does, with 500K tokens against 128K.

How many benchmarks do Grok 4.1 Fast and Grok 4.5 share?

28 benchmarks have published results for both models. Grok 4.1 Fast has 32 scored results on Noometry and Grok 4.5 has 52.

Related comparisons

Go deeper