Model comparison

Mistral Medium 3.5 vs Pixtral Large

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 32.2 on the Noometry Index.

Last verified . 1 shared benchmarks.

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • They share 1 benchmark with published results for both. Mistral Medium 3.5 scores higher in 2 categories and Pixtral Large in 1 category; 3 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium 3.5 leads 58.5 to 32.9.
  • Both cost about the same: $1.50 input and $7.50 output per million tokens.
  • Mistral Medium 3.5 accepts more context: 262K tokens versus 128K.

Side by side

Mistral Medium 3.5 and Pixtral Large specifications
Mistral Medium 3.5Pixtral Large
ProviderMistral AIMistral AI
Noometry Index40.232.2
Released—2024-11-01
WeightsOpenOpen
Context window262K128K
Max output210K128K
Input $ / M tokens$1.50$2
Output $ / M tokens$7.50$6
Results tracked223

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Medium 3.5: 36.0 (#213), Pixtral Large: —

Coding benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena WebDev1264—
LMArena Coding1461—

Reasoning Pixtral Large leads

Mistral Medium 3.5: 17.3 (#295), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
Kagi LLM Benchmark41.4%—
NYT Connections (extended)12.9%—
EnigmaEval—0.8%
LMArena Hard Prompts1436—
Epoch Capabilities Index141.35—

Math Not comparable

Mistral Medium 3.5: 39.1 (#113), Pixtral Large: —

Math benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Math1431—

Knowledge Not comparable

Mistral Medium 3.5: 40.0 (#126), Pixtral Large: —

Knowledge benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Expert1432—

Multimodal Mistral Medium 3.5 leads

Mistral Medium 3.5: 38.3 (#65), Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Vision12231089

Multilingual Not comparable

Mistral Medium 3.5: 51.9 (#100), Pixtral Large: —

Multilingual benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Non-English1404—
LMArena Chinese1442—
LMArena French1448—
LMArena German1451—
LMArena Korean1385—
LMArena Russian1395—
LMArena Spanish1409—

Instruction Following Not comparable

Mistral Medium 3.5: 74.6 (#90), Pixtral Large: —

Instruction Following benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Instruction Following1415—

Long Context Not comparable

Mistral Medium 3.5: 43.2 (#103), Pixtral Large: —

Long Context benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Longer Query1415—

Writing & Preference Mistral Medium 3.5 leads

Mistral Medium 3.5: 58.5 (#117), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkMistral Medium 3.5Pixtral Large
LMArena Text1421—
LMArena Creative Writing1374—
EQ-Bench Creative Writing—988
EQ-Bench 4993—
LMArena Multi-Turn1423—

Frequently asked questions

Is Mistral Medium 3.5 better than Pixtral Large?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 32.2 on the Noometry Index.

Which is cheaper, Mistral Medium 3.5 or Pixtral Large?

Pixtral Large is cheaper. It lists at $2 per million input tokens and $6 per million output tokens; Mistral Medium 3.5 lists at $1.50 and $7.50.

Which has the bigger context window?

Mistral Medium 3.5 does, with 262K tokens against 128K.

How many benchmarks do Mistral Medium 3.5 and Pixtral Large share?

1 benchmark has published results for both models. Mistral Medium 3.5 has 22 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper