Model comparison

Gemini 1.5 Pro (May 2024) vs Mixtral 8x7B

Gemini 1.5 Pro (May 2024) is the stronger model overall, scoring 32.1 to 27.1 on the Noometry Index.

Last verified . 30 shared benchmarks.

Gemini 1.5 Pro (May 2024) Google

32.1

Rank #261 Confirmed

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Summary

  • They share 30 benchmarks with published results for both. Gemini 1.5 Pro (May 2024) scores higher in 7 categories and Mixtral 8x7B in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemini 1.5 Pro (May 2024) leads 29.4 to 11.0.
  • The biggest single-benchmark swing is MATH Level 5: 70.4% for Gemini 1.5 Pro (May 2024) and 10% for Mixtral 8x7B.
  • Mixtral 8x7B has downloadable open weights; the other is API-only.

Side by side

Gemini 1.5 Pro (May 2024) and Mixtral 8x7B specifications
Gemini 1.5 Pro (May 2024)Mixtral 8x7B
ProviderGoogleMistral AI
Noometry Index32.127.1
Released2024-02-152023-12-11
WeightsProprietaryOpen
Context window—32K
Max output—32K
Input $ / M tokens—$0.70
Output $ / M tokens—$0.70
Results tracked4538

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 34.2 (#241), Mixtral 8x7B: 32.8 (#269)

Coding benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
LMArena Coding12941126
HumanEval+79.3%39.6%
MBPP+74.6%49.7%
WeirdML22.2%—
BigCodeBench Instruct43.8%—
BigCodeBench Complete57.5%—
CadEval34%—

Agentic & Tool Use Not comparable

Gemini 1.5 Pro (May 2024): 17.9 (#145), Mixtral 8x7B: —

Agentic & Tool Use benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
TheAgentCompany3.4%—
Cybench7.5%—
BALROG21%—

Reasoning Mixtral 8x7B leads

Gemini 1.5 Pro (May 2024): 12.3 (#338), Mixtral 8x7B: 18.2 (#285)

Reasoning benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
LMArena Hard Prompts12961115
DTBench59%49.6%
Epoch Capabilities Index131.73118.47
ForecastBench58.456.3
ARC-AGI-20.8%—
SimpleBench27.1%—
Adversarial NLI—55.2%
BIG-Bench Hard89.2%—
HellaSwag—86.7%
PIQA—83.6%
WinoGrande—77.2%

Math Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 25.8 (#266), Mixtral 8x7B: 18.8 (#289)

Math benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
Omni-MATH36.4%10.5%
LMArena Math13151147
MATH Level 570.4%10%
OTIS Mock AIME 2024-202523.1%—
GSM8K—74.4%

Knowledge Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 29.4 (#239), Mixtral 8x7B: 11.0 (#301)

Knowledge benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
GPQA Diamond57.2%30.6%
MMLU-Pro73.7%33.5%
GPQA (HELM)53.4%29.6%
LMArena Expert12791088
MMLU86.9%70.6%
Humanity's Last Exam4.6%—
Confabulations13.5%—
ARC (AI2) Challenge—87.3%
OpenBookQA—85.8%
TriviaQA—82.2%

Multimodal Not comparable

Gemini 1.5 Pro (May 2024): 36.8 (#77), Mixtral 8x7B: —

Multimodal benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
LMArena Vision1161—
Video-MME75%—

Multilingual Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 45.3 (#174), Mixtral 8x7B: 29.6 (#266)

Multilingual benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
LMArena Non-English13121077
LMArena Chinese13311055
LMArena French13021166
LMArena German12861114
LMArena Japanese1292931
LMArena Korean1298968
LMArena Russian13201090
LMArena Spanish13111111

Instruction Following Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 68.6 (#185), Mixtral 8x7B: 51.0 (#297)

Instruction Following benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
IFEval83.7%57.5%
LMArena Instruction Following12971109

Long Context Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 39.8 (#169), Mixtral 8x7B: 33.4 (#260)

Long Context benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
LMArena Longer Query13081103

Writing & Preference Gemini 1.5 Pro (May 2024) leads

Gemini 1.5 Pro (May 2024): 52.4 (#172), Mixtral 8x7B: 34.2 (#270)

Writing & Preference benchmarks
BenchmarkGemini 1.5 Pro (May 2024)Mixtral 8x7B
LMArena Text13191132
LMArena Creative Writing13331109
WildBench81.3%67.3%
LMArena Multi-Turn12961115

Frequently asked questions

Is Gemini 1.5 Pro (May 2024) better than Mixtral 8x7B?

Gemini 1.5 Pro (May 2024) is the stronger model overall, scoring 32.1 to 27.1 on the Noometry Index.

Is Gemini 1.5 Pro (May 2024) or Mixtral 8x7B better for coding?

Gemini 1.5 Pro (May 2024) scores higher on coding benchmarks: 34.2 versus 32.8 in the Noometry coding category.

How many benchmarks do Gemini 1.5 Pro (May 2024) and Mixtral 8x7B share?

30 benchmarks have published results for both models. Gemini 1.5 Pro (May 2024) has 45 scored results on Noometry and Mixtral 8x7B has 38.

Related comparisons

Go deeper