Model comparison

Grok 4.1 Fast vs Mistral Medium 3.5

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 40.2 on the Noometry Index.

Last verified . 19 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Grok 4.1 Fast scores higher in 1 category and Mistral Medium 3.5 in 8 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 17.3.
  • The biggest single-benchmark swing is NYT Connections (extended): 87.4% for Grok 4.1 Fast and 12.9% for Mistral Medium 3.5.
  • Grok 4.1 Fast is cheaper at $0.20 / $0.50 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium 3.5.
  • Mistral Medium 3.5 accepts more context: 262K tokens versus 128K.
  • Mistral Medium 3.5 has downloadable open weights; the other is API-only.

Side by side

Grok 4.1 Fast and Mistral Medium 3.5 specifications
Grok 4.1 FastMistral Medium 3.5
ProviderxAIMistral AI
Noometry Index41.440.2
Released2025-06-27—
WeightsProprietaryOpen
Context window128K262K
Max output30K210K
Input $ / M tokens$0.20$1.50
Output $ / M tokens$0.50$7.50
Results tracked3222

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Medium 3.5 leads

Grok 4.1 Fast: 34.1 (#245), Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena WebDev12421264
LMArena Coding14111461
ALE-Bench394.93—

Agentic & Tool Use Not comparable

Grok 4.1 Fast: 36.3 (#39), Mistral Medium 3.5: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
Berkeley Function Calling Leaderboard69.6%—
τ²-bench Banking13.1%—
LMArena Search1171—
Vending-Bench 21,107—

Reasoning Grok 4.1 Fast leads

Grok 4.1 Fast: 43.4 (#49), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
NYT Connections (extended)87.4%12.9%
LMArena Hard Prompts14071436
SimpleBench56%—
Kagi LLM Benchmark—41.4%
DTBench87.7%—
Epoch Capabilities Index—141.35
ForecastBench61—

Math Mistral Medium 3.5 leads

Grok 4.1 Fast: 31.9 (#221), Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Math14081431
MathArena Final-Answer Competitions60.9%—
ProofBench4%—

Knowledge Mistral Medium 3.5 leads

Grok 4.1 Fast: 33.1 (#207), Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Expert13991432
Vectara Hallucination Rate17.8%—

Multimodal Mistral Medium 3.5 leads

Grok 4.1 Fast: 37.0 (#76), Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Vision12011223

Multilingual Too close to call

Grok 4.1 Fast: 51.0 (#114), Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Non-English13911404
LMArena Chinese14411442
LMArena French14151448
LMArena German14041451
LMArena Korean13611385
LMArena Russian13871395
LMArena Spanish14131409
LMArena Japanese1349—

Instruction Following Mistral Medium 3.5 leads

Grok 4.1 Fast: 72.7 (#133), Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Instruction Following13761415

Long Context Too close to call

Grok 4.1 Fast: 42.4 (#126), Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Longer Query13901415

Writing & Preference Mistral Medium 3.5 leads

Grok 4.1 Fast: 57.2 (#131), Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkGrok 4.1 FastMistral Medium 3.5
LMArena Text14081421
LMArena Creative Writing13941374
LMArena Multi-Turn13891423
EQ-Bench Creative Writing1327—
EQ-Bench 4—993

Frequently asked questions

Is Grok 4.1 Fast better than Mistral Medium 3.5?

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 40.2 on the Noometry Index.

Which is cheaper, Grok 4.1 Fast or Mistral Medium 3.5?

Grok 4.1 Fast is cheaper. It lists at $0.20 per million input tokens and $0.50 per million output tokens; Mistral Medium 3.5 lists at $1.50 and $7.50.

Is Grok 4.1 Fast or Mistral Medium 3.5 better for coding?

Mistral Medium 3.5 scores higher on coding benchmarks: 36.0 versus 34.1 in the Noometry coding category.

Which has the bigger context window?

Mistral Medium 3.5 does, with 262K tokens against 128K.

How many benchmarks do Grok 4.1 Fast and Mistral Medium 3.5 share?

19 benchmarks have published results for both models. Grok 4.1 Fast has 32 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper