Model comparison

Grok 4.1 Fast vs Mistral

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 29.9 on the Noometry Index.

Last verified . 17 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Mistral Mistral AI

29.9

Rank #303 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Grok 4.1 Fast scores higher in 8 categories and Mistral in 0 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 22.2.

Side by side

Grok 4.1 Fast and Mistral specifications
Grok 4.1 FastMistral
ProviderxAIMistral AI
Noometry Index41.429.9
Released2025-06-27—
WeightsProprietaryProprietary
Context window128K—
Max output30K—
Input $ / M tokens$0.20—
Output $ / M tokens$0.50—
Results tracked3222

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Grok 4.1 Fast: 34.1 (#245), Mistral: 33.8 (#250)

Coding benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Coding14111162
LMArena WebDev1242—
ALE-Bench394.93—

Agentic & Tool Use Not comparable

Grok 4.1 Fast: 36.3 (#39), Mistral: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1 FastMistral
Berkeley Function Calling Leaderboard69.6%—
τ²-bench Banking13.1%—
LMArena Search1171—
Vending-Bench 21,107—

Reasoning Grok 4.1 Fast leads

Grok 4.1 Fast: 43.4 (#49), Mistral: 22.2 (#200)

Reasoning benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Hard Prompts14071149
SimpleBench56%—
NYT Connections (extended)87.4%—
DTBench87.7%—
ForecastBench61—

Math Grok 4.1 Fast leads

Grok 4.1 Fast: 31.9 (#221), Mistral: 22.3 (#278)

Math benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Math14081180
MathArena Final-Answer Competitions60.9%—
ProofBench4%—
Omni-MATH—7.2%

Knowledge Grok 4.1 Fast leads

Grok 4.1 Fast: 33.1 (#207), Mistral: 16.6 (#288)

Knowledge benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Expert13991125
MMLU-Pro—27.7%
Vectara Hallucination Rate17.8%—
GPQA (HELM)—30.3%

Multimodal Not comparable

Grok 4.1 Fast: 37.0 (#76), Mistral: —

Multimodal benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Vision1201—

Multilingual Grok 4.1 Fast leads

Grok 4.1 Fast: 51.0 (#114), Mistral: 32.8 (#254)

Multilingual benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Non-English13911129
LMArena Chinese14411109
LMArena French14151180
LMArena German14041155
LMArena Japanese13491013
LMArena Korean13611032
LMArena Russian13871168
LMArena Spanish14131143

Instruction Following Grok 4.1 Fast leads

Grok 4.1 Fast: 72.7 (#133), Mistral: 52.6 (#288)

Instruction Following benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Instruction Following13761152
IFEval—56.8%

Long Context Grok 4.1 Fast leads

Grok 4.1 Fast: 42.4 (#126), Mistral: 35.0 (#245)

Long Context benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Longer Query13901153

Writing & Preference Grok 4.1 Fast leads

Grok 4.1 Fast: 57.2 (#131), Mistral: 37.0 (#260)

Writing & Preference benchmarks
BenchmarkGrok 4.1 FastMistral
LMArena Text14081165
LMArena Creative Writing13941158
LMArena Multi-Turn13891147
EQ-Bench Creative Writing1327—
WildBench—66%

Frequently asked questions

Is Grok 4.1 Fast better than Mistral?

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 29.9 on the Noometry Index.

Is Grok 4.1 Fast or Mistral better for coding?

They score almost the same on coding (34.1 vs 33.8); test both on your own repository before choosing.

How many benchmarks do Grok 4.1 Fast and Mistral share?

17 benchmarks have published results for both models. Grok 4.1 Fast has 32 scored results on Noometry and Mistral has 22.

Related comparisons

Go deeper