Model comparison

Mistral vs Pixtral Large

Pixtral Large is the stronger model overall, scoring 32.2 to 29.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mistral Mistral AI

29.9

Rank #303 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in writing & preference, where Mistral leads 37.0 to 32.9.
  • Pixtral Large has downloadable open weights; the other is API-only.

Side by side

Mistral and Pixtral Large specifications
MistralPixtral Large
ProviderMistral AIMistral AI
Noometry Index29.932.2
Released—2024-11-01
WeightsProprietaryOpen
Context window—128K
Max output—128K
Input $ / M tokens—$2
Output $ / M tokens—$6
Results tracked223

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral: 33.8 (#250), Pixtral Large: —

Coding benchmarks
BenchmarkMistralPixtral Large
LMArena Coding1162—

Reasoning Too close to call

Mistral: 22.2 (#200), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkMistralPixtral Large
EnigmaEval—0.8%
LMArena Hard Prompts1149—

Math Not comparable

Mistral: 22.3 (#278), Pixtral Large: —

Math benchmarks
BenchmarkMistralPixtral Large
Omni-MATH7.2%—
LMArena Math1180—

Knowledge Not comparable

Mistral: 16.6 (#288), Pixtral Large: —

Knowledge benchmarks
BenchmarkMistralPixtral Large
MMLU-Pro27.7%—
GPQA (HELM)30.3%—
LMArena Expert1125—

Multimodal Not comparable

Mistral: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkMistralPixtral Large
LMArena Vision—1089

Multilingual Not comparable

Mistral: 32.8 (#254), Pixtral Large: —

Multilingual benchmarks
BenchmarkMistralPixtral Large
LMArena Non-English1129—
LMArena Chinese1109—
LMArena French1180—
LMArena German1155—
LMArena Japanese1013—
LMArena Korean1032—
LMArena Russian1168—
LMArena Spanish1143—

Instruction Following Not comparable

Mistral: 52.6 (#288), Pixtral Large: —

Instruction Following benchmarks
BenchmarkMistralPixtral Large
IFEval56.8%—
LMArena Instruction Following1152—

Long Context Not comparable

Mistral: 35.0 (#245), Pixtral Large: —

Long Context benchmarks
BenchmarkMistralPixtral Large
LMArena Longer Query1153—

Writing & Preference Mistral leads

Mistral: 37.0 (#260), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkMistralPixtral Large
LMArena Text1165—
LMArena Creative Writing1158—
EQ-Bench Creative Writing—988
WildBench66%—
LMArena Multi-Turn1147—

Frequently asked questions

Is Mistral better than Pixtral Large?

Pixtral Large is the stronger model overall, scoring 32.2 to 29.9 on the Noometry Index.

How many benchmarks do Mistral and Pixtral Large share?

0 benchmarks have published results for both models. Mistral has 22 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper