Model comparison

Pixtral Large vs Qwen1.5-14B

Pixtral Large and Qwen1.5-14B score almost the same on the Noometry Index (32.2 vs 32.7), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Qwen1.5-14B Alibaba (Qwen)

32.7

Rank #253 Confirmed

Side by side

Pixtral Large and Qwen1.5-14B specifications
Pixtral LargeQwen1.5-14B
ProviderMistral AIAlibaba (Qwen)
Noometry Index32.232.7
Released2024-11-012024-02-04
WeightsOpenOpen
Context window128K—
Max output128K—
Input $ / M tokens$2—
Output $ / M tokens$6—
Results tracked317

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Pixtral Large: —, Qwen1.5-14B: 33.1 (#263)

Coding benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Coding—1138

Reasoning Too close to call

Pixtral Large: 21.7 (#218), Qwen1.5-14B: 21.4 (#223)

Reasoning benchmarks
BenchmarkPixtral LargeQwen1.5-14B
EnigmaEval0.8%—
LMArena Hard Prompts—1113

Math Not comparable

Pixtral Large: —, Qwen1.5-14B: 32.4 (#215)

Math benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Math—1125

Knowledge Not comparable

Pixtral Large: —, Qwen1.5-14B: 29.8 (#232)

Knowledge benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Expert—1094
MMLU—68.6%

Multimodal Not comparable

Pixtral Large: 30.6 (#111), Qwen1.5-14B: —

Multimodal benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Vision1089—

Multilingual Not comparable

Pixtral Large: —, Qwen1.5-14B: 30.7 (#262)

Multilingual benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Non-English—1095
LMArena Chinese—1147
LMArena French—1116
LMArena German—1043
LMArena Japanese—1019
LMArena Russian—1046
LMArena Spanish—1085

Instruction Following Not comparable

Pixtral Large: —, Qwen1.5-14B: 56.8 (#271)

Instruction Following benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Instruction Following—1102

Long Context Not comparable

Pixtral Large: —, Qwen1.5-14B: 33.7 (#257)

Long Context benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Longer Query—1113

Writing & Preference Too close to call

Pixtral Large: 32.9 (#278), Qwen1.5-14B: 33.6 (#276)

Writing & Preference benchmarks
BenchmarkPixtral LargeQwen1.5-14B
LMArena Text—1128
LMArena Creative Writing—1091
EQ-Bench Creative Writing988—
LMArena Multi-Turn—1110

Frequently asked questions

Is Pixtral Large better than Qwen1.5-14B?

Pixtral Large and Qwen1.5-14B score almost the same on the Noometry Index (32.2 vs 32.7), so choose on price, context window or the category you care about most.

How many benchmarks do Pixtral Large and Qwen1.5-14B share?

0 benchmarks have published results for both models. Pixtral Large has 3 scored results on Noometry and Qwen1.5-14B has 17.

Related comparisons

Go deeper