Model comparison

Gemma 7B vs Pixtral Large

Pixtral Large is the stronger model overall, scoring 32.2 to 30.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in writing & preference, where Pixtral Large leads 32.9 to 27.1.

Side by side

Gemma 7B and Pixtral Large specifications
Gemma 7BPixtral Large
ProviderGoogleMistral AI
Noometry Index30.032.2
Released2024-02-212024-11-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$2
Output $ / M tokens—$6
Results tracked273

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 7B: 30.5 (#294), Pixtral Large: —

Coding benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Coding1048—
HumanEval+28.7%—
MBPP+43.4%—

Reasoning Pixtral Large leads

Gemma 7B: 19.9 (#249), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkGemma 7BPixtral Large
EnigmaEval—0.8%
LMArena Hard Prompts1042—
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Not comparable

Gemma 7B: 31.2 (#228), Pixtral Large: —

Math benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Math1066—
GSM8K46.4%—

Knowledge Not comparable

Gemma 7B: 27.3 (#252), Pixtral Large: —

Knowledge benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Expert1001—
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multimodal Not comparable

Gemma 7B: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Vision—1089

Multilingual Not comparable

Gemma 7B: 25.1 (#287), Pixtral Large: —

Multilingual benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Non-English999—
LMArena Chinese1035—
LMArena French1025—
LMArena Russian993—

Instruction Following Not comparable

Gemma 7B: 51.5 (#295), Pixtral Large: —

Instruction Following benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Instruction Following1017—

Long Context Not comparable

Gemma 7B: 31.1 (#282), Pixtral Large: —

Long Context benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Longer Query1022—

Writing & Preference Pixtral Large leads

Gemma 7B: 27.1 (#302), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkGemma 7BPixtral Large
LMArena Text1056—
LMArena Creative Writing1024—
EQ-Bench Creative Writing—988
LMArena Multi-Turn963—

Frequently asked questions

Is Gemma 7B better than Pixtral Large?

Pixtral Large is the stronger model overall, scoring 32.2 to 30.0 on the Noometry Index.

How many benchmarks do Gemma 7B and Pixtral Large share?

0 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper