Model comparison

Pixtral Large vs Qwen1.5-7B

Pixtral Large and Qwen1.5-7B score almost the same on the Noometry Index (32.2 vs 31.4), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Qwen1.5-7B Alibaba (Qwen)

31.4

Rank #273 Confirmed

Summary

  • The widest gap is in writing & preference, where Pixtral Large leads 32.9 to 29.6.

Side by side

Pixtral Large and Qwen1.5-7B specifications
Pixtral LargeQwen1.5-7B
ProviderMistral AIAlibaba (Qwen)
Noometry Index32.231.4
Released2024-11-012024-02-04
WeightsOpenOpen
Context window128K—
Max output128K—
Input $ / M tokens$2—
Output $ / M tokens$6—
Results tracked313

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Pixtral Large: —, Qwen1.5-7B: 32.2 (#276)

Coding benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Coding—1107

Reasoning Pixtral Large leads

Pixtral Large: 21.7 (#218), Qwen1.5-7B: 20.4 (#240)

Reasoning benchmarks
BenchmarkPixtral LargeQwen1.5-7B
EnigmaEval0.8%—
LMArena Hard Prompts—1065

Math Not comparable

Pixtral Large: —, Qwen1.5-7B: 31.4 (#224)

Math benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Math—1080

Knowledge Not comparable

Pixtral Large: —, Qwen1.5-7B: 28.7 (#243)

Knowledge benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Expert—1055
MMLU—62.6%

Multimodal Not comparable

Pixtral Large: 30.6 (#111), Qwen1.5-7B: —

Multimodal benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Vision1089—

Multilingual Not comparable

Pixtral Large: —, Qwen1.5-7B: 28.5 (#271)

Multilingual benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Non-English—1058
LMArena Chinese—1141
LMArena Russian—1006

Instruction Following Not comparable

Pixtral Large: —, Qwen1.5-7B: 54.1 (#281)

Instruction Following benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Instruction Following—1058

Long Context Not comparable

Pixtral Large: —, Qwen1.5-7B: 33.1 (#266)

Long Context benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Longer Query—1090

Writing & Preference Pixtral Large leads

Pixtral Large: 32.9 (#278), Qwen1.5-7B: 29.6 (#293)

Writing & Preference benchmarks
BenchmarkPixtral LargeQwen1.5-7B
LMArena Text—1083
LMArena Creative Writing—1035
EQ-Bench Creative Writing988—
LMArena Multi-Turn—1062

Frequently asked questions

Is Pixtral Large better than Qwen1.5-7B?

Pixtral Large and Qwen1.5-7B score almost the same on the Noometry Index (32.2 vs 31.4), so choose on price, context window or the category you care about most.

How many benchmarks do Pixtral Large and Qwen1.5-7B share?

0 benchmarks have published results for both models. Pixtral Large has 3 scored results on Noometry and Qwen1.5-7B has 13.

Related comparisons

Go deeper