Model comparison

Pixtral Large vs Qwen3.6 27B

Qwen3.6 27B is the stronger model overall, scoring 42.2 to 32.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Qwen3.6 27B Alibaba (Qwen)

42.2

Rank #117 Confirmed

Summary

  • The widest gap is in writing & preference, where Qwen3.6 27B leads 50.3 to 32.9.
  • Qwen3.6 27B is cheaper at $0.60 / $3.60 per million input/output tokens, against $2 / $6 for Pixtral Large.
  • Qwen3.6 27B accepts more context: 262K tokens versus 128K.

Side by side

Pixtral Large and Qwen3.6 27B specifications
Pixtral LargeQwen3.6 27B
ProviderMistral AIAlibaba (Qwen)
Noometry Index32.242.2
Released2024-11-012026-04-22
WeightsOpenOpen
Context window128K262K
Max output128K66K
Input $ / M tokens$2$0.60
Output $ / M tokens$6$3.60
Results tracked311

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Pixtral Large: —, Qwen3.6 27B: 39.1 (#163)

Coding benchmarks
BenchmarkPixtral LargeQwen3.6 27B
SciCode—37.3%

Reasoning Qwen3.6 27B leads

Pixtral Large: 21.7 (#218), Qwen3.6 27B: 25.0 (#153)

Reasoning benchmarks
BenchmarkPixtral LargeQwen3.6 27B
CritPt—0.9%
Chess Puzzles—22%
EnigmaEval0.8%—
Mystery Game Puzzles—7%
DTBench—78.1%
LMCA—34.5%
Epoch Capabilities Index—146.5

Math Not comparable

Pixtral Large: —, Qwen3.6 27B: 48.5 (#62)

Math benchmarks
BenchmarkPixtral LargeQwen3.6 27B
FrontierMath (Tiers 1-3)—35.1%
OTIS Mock AIME 2024-2025—91.1%

Knowledge Not comparable

Pixtral Large: —, Qwen3.6 27B: 52.4 (#63)

Knowledge benchmarks
BenchmarkPixtral LargeQwen3.6 27B
GPQA Diamond—85.9%

Multimodal Not comparable

Pixtral Large: 30.6 (#111), Qwen3.6 27B: —

Multimodal benchmarks
BenchmarkPixtral LargeQwen3.6 27B
LMArena Vision1089—

Writing & Preference Qwen3.6 27B leads

Pixtral Large: 32.9 (#278), Qwen3.6 27B: 50.3 (#181)

Writing & Preference benchmarks
BenchmarkPixtral LargeQwen3.6 27B
EQ-Bench Creative Writing988—
EQ-Bench 4—1026

Frequently asked questions

Is Pixtral Large better than Qwen3.6 27B?

Qwen3.6 27B is the stronger model overall, scoring 42.2 to 32.2 on the Noometry Index.

Which is cheaper, Pixtral Large or Qwen3.6 27B?

Qwen3.6 27B is cheaper. It lists at $0.60 per million input tokens and $3.60 per million output tokens; Pixtral Large lists at $2 and $6.

Which has the bigger context window?

Qwen3.6 27B does, with 262K tokens against 128K.

How many benchmarks do Pixtral Large and Qwen3.6 27B share?

0 benchmarks have published results for both models. Pixtral Large has 3 scored results on Noometry and Qwen3.6 27B has 11.

Related comparisons

Go deeper