Model comparison

Grok 4.1 Fast vs Mistral Small 3.2

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 31.2 on the Noometry Index. Mistral Small 3.2 costs 2.1× less per token, which makes it the better buy when Grok 4.1 Fast's lead doesn't matter for your workload.

Last verified . 1 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • They share 1 benchmark with published results for both. Grok 4.1 Fast scores higher in 4 categories and Mistral Small 3.2 in 0 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 18.1.
  • Mistral Small 3.2 is cheaper at $0.0938 / $0.25 per million input/output tokens, against $0.20 / $0.50 for Grok 4.1 Fast.
  • Mistral Small 3.2 accepts more context: 256K tokens versus 128K.
  • Mistral Small 3.2 has downloadable open weights; the other is API-only.

Side by side

Grok 4.1 Fast and Mistral Small 3.2 specifications
Grok 4.1 FastMistral Small 3.2
ProviderxAIMistral AI
Noometry Index41.431.2
Released2025-06-272025-06-20
WeightsProprietaryOpen
Context window128K256K
Max output30K16K
Input $ / M tokens$0.20$0.0938
Output $ / M tokens$0.50$0.25
Results tracked326

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Grok 4.1 Fast: 34.1 (#245), Mistral Small 3.2: —

Coding benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
LMArena WebDev1242—
LMArena Coding1411—
ALE-Bench394.93—

Agentic & Tool Use Not comparable

Grok 4.1 Fast: 36.3 (#39), Mistral Small 3.2: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
Berkeley Function Calling Leaderboard69.6%—
τ²-bench Banking13.1%—
LMArena Search1171—
Vending-Bench 21,107—

Reasoning Grok 4.1 Fast leads

Grok 4.1 Fast: 43.4 (#49), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
SimpleBench56%—
Kagi LLM Benchmark—40.4%
NYT Connections (extended)87.4%—
Chess Puzzles—1%
LMArena Hard Prompts1407—
DTBench87.7%—
Epoch Capabilities Index—131.74
ForecastBench61—

Math Grok 4.1 Fast leads

Grok 4.1 Fast: 31.9 (#221), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
MathArena Final-Answer Competitions60.9%—
OTIS Mock AIME 2024-2025—30.3%
ProofBench4%—
LMArena Math1408—

Knowledge Grok 4.1 Fast leads

Grok 4.1 Fast: 33.1 (#207), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
GPQA Diamond—49.1%
Vectara Hallucination Rate17.8%—
LMArena Expert1399—

Multimodal Not comparable

Grok 4.1 Fast: 37.0 (#76), Mistral Small 3.2: —

Multimodal benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
LMArena Vision1201—

Multilingual Not comparable

Grok 4.1 Fast: 51.0 (#114), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
LMArena Non-English1391—
LMArena Chinese1441—
LMArena French1415—
LMArena German1404—
LMArena Japanese1349—
LMArena Korean1361—
LMArena Russian1387—
LMArena Spanish1413—

Instruction Following Not comparable

Grok 4.1 Fast: 72.7 (#133), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
LMArena Instruction Following1376—

Long Context Not comparable

Grok 4.1 Fast: 42.4 (#126), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
LMArena Longer Query1390—

Writing & Preference Grok 4.1 Fast leads

Grok 4.1 Fast: 57.2 (#131), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkGrok 4.1 FastMistral Small 3.2
EQ-Bench Creative Writing13271255
LMArena Text1408—
LMArena Creative Writing1394—
LMArena Multi-Turn1389—

Frequently asked questions

Is Grok 4.1 Fast better than Mistral Small 3.2?

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 31.2 on the Noometry Index. Mistral Small 3.2 costs 2.1× less per token, which makes it the better buy when Grok 4.1 Fast's lead doesn't matter for your workload.

Which is cheaper, Grok 4.1 Fast or Mistral Small 3.2?

Mistral Small 3.2 is cheaper. It lists at $0.0938 per million input tokens and $0.25 per million output tokens; Grok 4.1 Fast lists at $0.20 and $0.50.

Which has the bigger context window?

Mistral Small 3.2 does, with 256K tokens against 128K.

How many benchmarks do Grok 4.1 Fast and Mistral Small 3.2 share?

1 benchmark has published results for both models. Grok 4.1 Fast has 32 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper