Model comparison

Grok 4.1 Fast vs Magistral Medium

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 35.2 on the Noometry Index.

Last verified . 17 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Grok 4.1 Fast scores higher in 5 categories and Magistral Medium in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 8.6.
  • Grok 4.1 Fast is cheaper at $0.20 / $0.50 per million input/output tokens, against $2 / $5 for Magistral Medium.
  • Magistral Medium accepts more context: 262K tokens versus 128K.
  • Magistral Medium has downloadable open weights; the other is API-only.

Side by side

Grok 4.1 Fast and Magistral Medium specifications
Grok 4.1 FastMagistral Medium
ProviderxAIMistral AI
Noometry Index41.435.2
Released2025-06-272025-03-17
WeightsProprietaryOpen
Context window128K262K
Max output30K16K
Input $ / M tokens$0.20$2
Output $ / M tokens$0.50$5
Results tracked3222

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Grok 4.1 Fast: 34.1 (#245), Magistral Medium: 39.1 (#161)

Coding benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Coding14111319
LMArena WebDev1242—
SciCode—39.2%
ALE-Bench394.93—

Agentic & Tool Use Not comparable

Grok 4.1 Fast: 36.3 (#39), Magistral Medium: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
Berkeley Function Calling Leaderboard69.6%—
τ²-bench Banking13.1%—
LMArena Search1171—
Vending-Bench 21,107—

Reasoning Grok 4.1 Fast leads

Grok 4.1 Fast: 43.4 (#49), Magistral Medium: 8.6 (#348)

Reasoning benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Hard Prompts14071267
ARC-AGI-2—0%
SimpleBench56%—
Kagi LLM Benchmark—16.2%
NYT Connections (extended)87.4%—
ARC-AGI-1—6.1%
CritPt—0.3%
DTBench87.7%—
ForecastBench61—

Math Magistral Medium leads

Grok 4.1 Fast: 31.9 (#221), Magistral Medium: 35.1 (#189)

Math benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Math14081250
MathArena Final-Answer Competitions60.9%—
ProofBench4%—

Knowledge Too close to call

Grok 4.1 Fast: 33.1 (#207), Magistral Medium: 33.5 (#202)

Knowledge benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Expert13991223
Vectara Hallucination Rate17.8%—

Multimodal Not comparable

Grok 4.1 Fast: 37.0 (#76), Magistral Medium: —

Multimodal benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Vision1201—

Multilingual Grok 4.1 Fast leads

Grok 4.1 Fast: 51.0 (#114), Magistral Medium: 39.6 (#224)

Multilingual benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Non-English13911232
LMArena Chinese14411227
LMArena French14151267
LMArena German14041248
LMArena Japanese13491175
LMArena Korean13611125
LMArena Russian13871224
LMArena Spanish14131271

Instruction Following Grok 4.1 Fast leads

Grok 4.1 Fast: 72.7 (#133), Magistral Medium: 66.0 (#211)

Instruction Following benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Instruction Following13761254

Long Context Grok 4.1 Fast leads

Grok 4.1 Fast: 42.4 (#126), Magistral Medium: 39.3 (#183)

Long Context benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Longer Query13901295

Writing & Preference Grok 4.1 Fast leads

Grok 4.1 Fast: 57.2 (#131), Magistral Medium: 46.3 (#219)

Writing & Preference benchmarks
BenchmarkGrok 4.1 FastMagistral Medium
LMArena Text14081255
LMArena Creative Writing13941245
LMArena Multi-Turn13891275
EQ-Bench Creative Writing1327—

Frequently asked questions

Is Grok 4.1 Fast better than Magistral Medium?

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 35.2 on the Noometry Index.

Which is cheaper, Grok 4.1 Fast or Magistral Medium?

Grok 4.1 Fast is cheaper. It lists at $0.20 per million input tokens and $0.50 per million output tokens; Magistral Medium lists at $2 and $5.

Is Grok 4.1 Fast or Magistral Medium better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 34.1 in the Noometry coding category.

Which has the bigger context window?

Magistral Medium does, with 262K tokens against 128K.

How many benchmarks do Grok 4.1 Fast and Magistral Medium share?

17 benchmarks have published results for both models. Grok 4.1 Fast has 32 scored results on Noometry and Magistral Medium has 22.

Related comparisons

Go deeper