Model comparison

Gemini 1.5 Pro (May 2024) vs Mistral Medium

Mistral Medium is the stronger model overall, scoring 36.3 to 32.1 on the Noometry Index.

Last verified . 24 shared benchmarks.

Gemini 1.5 Pro (May 2024) Google

32.1

Rank #261 Confirmed

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Summary

  • They share 24 benchmarks with published results for both. Gemini 1.5 Pro (May 2024) scores higher in 3 categories and Mistral Medium in 7 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Mistral Medium leads 24.0 to 12.3.
  • The biggest single-benchmark swing is WeirdML: 22.2% for Gemini 1.5 Pro (May 2024) and 43.7% for Mistral Medium.
  • Mistral Medium has downloadable open weights; the other is API-only.

Side by side

Gemini 1.5 Pro (May 2024) and Mistral Medium specifications
Gemini 1.5 Pro (May 2024)Mistral Medium
ProviderGoogleMistral AI
Noometry Index32.136.3
Released2024-02-152023-12-11
WeightsProprietaryOpen
Context window—262K
Max output—262K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked4536

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemini 1.5 Pro (May 2024): 34.2 (#241), Mistral Medium: 34.2 (#243)

Coding benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
WeirdML22.2%43.7%
LMArena Coding12941434
FrontierCode—8%
SciCode—40.2%
BigCodeBench Instruct43.8%—
BigCodeBench Complete57.5%—
CadEval34%—
ALE-Bench—763.98
HumanEval+79.3%—
MBPP+74.6%—

Agentic & Tool Use Mistral Medium leads

Gemini 1.5 Pro (May 2024): 17.9 (#145), Mistral Medium: 28.3 (#90)

Agentic & Tool Use benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
Berkeley Function Calling Leaderboard—37.7%
TheAgentCompany3.4%—
Cybench7.5%—
BALROG21%—

Reasoning Mistral Medium leads

Gemini 1.5 Pro (May 2024): 12.3 (#338), Mistral Medium: 24.0 (#167)

Reasoning benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
LMArena Hard Prompts12961426
DTBench59%75.5%
ARC-AGI-20.8%—
SimpleBench27.1%—
Kagi LLM Benchmark—50%
CritPt—0%
LMCA—26.1%
Surface Evolver Bench—26.9%
BIG-Bench Hard89.2%—
Epoch Capabilities Index131.73—
ForecastBench58.4—

Math Mistral Medium leads

Gemini 1.5 Pro (May 2024): 25.8 (#266), Mistral Medium: 28.1 (#245)

Math benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
OTIS Mock AIME 2024-202523.1%32.2%
LMArena Math13151408
MATH Level 570.4%81.6%
ProofBench—9%
Omni-MATH36.4%—
FrontierMath (Feb 2025 set)—0.3%

Knowledge Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 29.4 (#239), Mistral Medium: 25.0 (#265)

Knowledge benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
GPQA Diamond57.2%59.5%
Humanity's Last Exam4.6%4.5%
LMArena Expert12791408
MMLU-Pro73.7%—
Confabulations13.5%—
Vectara Hallucination Rate—22.7%
GPQA (HELM)53.4%—
MMLU86.9%—

Multimodal Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 36.8 (#77), Mistral Medium: 35.3 (#88)

Multimodal benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
LMArena Vision11611172
Video-MME75%—

Multilingual Mistral Medium leads

Gemini 1.5 Pro (May 2024): 45.3 (#174), Mistral Medium: 52.1 (#91)

Multilingual benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
LMArena Non-English13121408
LMArena Chinese13311447
LMArena French13021459
LMArena German12861432
LMArena Japanese12921378
LMArena Korean12981380
LMArena Russian13201411
LMArena Spanish13111433

Instruction Following Mistral Medium leads

Gemini 1.5 Pro (May 2024): 68.6 (#185), Mistral Medium: 73.7 (#116)

Instruction Following benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
LMArena Instruction Following12971398
IFEval83.7%—

Long Context Mistral Medium leads

Gemini 1.5 Pro (May 2024): 39.8 (#169), Mistral Medium: 42.9 (#114)

Long Context benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
LMArena Longer Query13081406

Writing & Preference Mistral Medium leads

Gemini 1.5 Pro (May 2024): 52.4 (#172), Mistral Medium: 60.0 (#103)

Writing & Preference benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mistral Medium
LMArena Text13191424
LMArena Creative Writing13331391
LMArena Multi-Turn12961418
Short-Story Creative Writing—77.3%
WildBench81.3%—

Frequently asked questions

Is Gemini 1.5 Pro (May 2024) better than Mistral Medium?

Mistral Medium is the stronger model overall, scoring 36.3 to 32.1 on the Noometry Index.

Is Gemini 1.5 Pro (May 2024) or Mistral Medium better for coding?

They score almost the same on coding (34.2 vs 34.2); test both on your own repository before choosing.

How many benchmarks do Gemini 1.5 Pro (May 2024) and Mistral Medium share?

24 benchmarks have published results for both models. Gemini 1.5 Pro (May 2024) has 45 scored results on Noometry and Mistral Medium has 36.

Related comparisons

Go deeper