Model comparison

Pixtral Large vs Qwen Plus

Qwen Plus is the stronger model overall, scoring 37.1 to 32.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • The widest gap is in writing & preference, where Qwen Plus leads 52.2 to 32.9.
  • Qwen Plus is cheaper at $0.40 / $1.20 per million input/output tokens, against $2 / $6 for Pixtral Large.
  • Qwen Plus accepts more context: 1M tokens versus 128K.
  • Pixtral Large has downloadable open weights; the other is API-only.

Side by side

Pixtral Large and Qwen Plus specifications
Pixtral LargeQwen Plus
ProviderMistral AIAlibaba (Qwen)
Noometry Index32.237.1
Released2024-11-012024-01-25
WeightsOpenProprietary
Context window128K1M
Max output128K33K
Input $ / M tokens$2$0.40
Output $ / M tokens$6$1.20
Results tracked320

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Pixtral Large: —, Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkPixtral LargeQwen Plus
LMArena Coding—1328

Reasoning Qwen Plus leads

Pixtral Large: 21.7 (#218), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkPixtral LargeQwen Plus
Kagi LLM Benchmark—63.3%
EnigmaEval0.8%—
LMArena Hard Prompts—1317
DTBench—81.1%
LMCA—24%

Math Not comparable

Pixtral Large: —, Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkPixtral LargeQwen Plus
OTIS Mock AIME 2024-2025—17.8%
LMArena Math—1326
MATH Level 5—65.3%
FrontierMath (Feb 2025 set)—1.7%

Knowledge Not comparable

Pixtral Large: —, Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkPixtral LargeQwen Plus
GPQA Diamond—48.1%
LMArena Expert—1328

Multimodal Not comparable

Pixtral Large: 30.6 (#111), Qwen Plus: —

Multimodal benchmarks
BenchmarkPixtral LargeQwen Plus
LMArena Vision1089—

Multilingual Not comparable

Pixtral Large: —, Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkPixtral LargeQwen Plus
LMArena Non-English—1310
LMArena Chinese—1347
LMArena Japanese—1251
LMArena Russian—1323

Instruction Following Not comparable

Pixtral Large: —, Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkPixtral LargeQwen Plus
LMArena Instruction Following—1303

Long Context Not comparable

Pixtral Large: —, Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkPixtral LargeQwen Plus
LMArena Longer Query—1324

Writing & Preference Qwen Plus leads

Pixtral Large: 32.9 (#278), Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkPixtral LargeQwen Plus
LMArena Text—1326
LMArena Creative Writing—1293
EQ-Bench Creative Writing988—
LMArena Multi-Turn—1336

Frequently asked questions

Is Pixtral Large better than Qwen Plus?

Qwen Plus is the stronger model overall, scoring 37.1 to 32.2 on the Noometry Index.

Which is cheaper, Pixtral Large or Qwen Plus?

Qwen Plus is cheaper. It lists at $0.40 per million input tokens and $1.20 per million output tokens; Pixtral Large lists at $2 and $6.

Which has the bigger context window?

Qwen Plus does, with 1M tokens against 128K.

How many benchmarks do Pixtral Large and Qwen Plus share?

0 benchmarks have published results for both models. Pixtral Large has 3 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper