Model comparison

Grok Build 0.1 vs Mistral Medium 3

Grok Build 0.1 has enough public results to be ranked (#216); Mistral Medium 3 does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Grok Build 0.1 xAI

36.4

Rank #216 Reported

Mistral Medium 3 Mistral AI

37.0

Unranked Sparse

Summary

  • Mistral Medium 3 is cheaper at $0.40 / $2 per million input/output tokens, against $1 / $2 for Grok Build 0.1.
  • Grok Build 0.1 accepts more context: 256K tokens versus 131K.

Side by side

Grok Build 0.1 and Mistral Medium 3 specifications
Grok Build 0.1Mistral Medium 3
ProviderxAIMistral AI
Noometry Index36.437.0
Released2026-04-162025-05-07
WeightsProprietaryProprietary
Context window256K131K
Max output256K131K
Input $ / M tokens$1$0.40
Output $ / M tokens$2$2
Results tracked32

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Grok Build 0.1: 43.1 (#91), Mistral Medium 3: —

Coding benchmarks
BenchmarkGrok Build 0.1Mistral Medium 3
SciCode50.2%—

Agentic & Tool Use Not comparable

Grok Build 0.1: 22.7 (#129), Mistral Medium 3: —

Agentic & Tool Use benchmarks
BenchmarkGrok Build 0.1Mistral Medium 3
GBAEval2.4%—

Reasoning Not comparable

Grok Build 0.1: 32.2 (#77), Mistral Medium 3: —

Reasoning benchmarks
BenchmarkGrok Build 0.1Mistral Medium 3
CritPt9.1%—
Epoch Capabilities Index—134.07

Knowledge Not comparable

Grok Build 0.1: —, Mistral Medium 3: 34.6

Knowledge benchmarks
BenchmarkGrok Build 0.1Mistral Medium 3
Confabulations—21.9%

Frequently asked questions

Is Grok Build 0.1 better than Mistral Medium 3?

Grok Build 0.1 has enough public results to be ranked (#216); Mistral Medium 3 does not yet, so treat this comparison as directional.

Which is cheaper, Grok Build 0.1 or Mistral Medium 3?

Mistral Medium 3 is cheaper. It lists at $0.40 per million input tokens and $2 per million output tokens; Grok Build 0.1 lists at $1 and $2.

Which has the bigger context window?

Grok Build 0.1 does, with 256K tokens against 131K.

How many benchmarks do Grok Build 0.1 and Mistral Medium 3 share?

0 benchmarks have published results for both models. Grok Build 0.1 has 3 scored results on Noometry and Mistral Medium 3 has 2.

Related comparisons

Go deeper