Model comparison

Mistral vs Qwen3.6 27B

Qwen3.6 27B is the stronger model overall, scoring 42.2 to 29.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mistral Mistral AI

29.9

Rank #303 Confirmed

Qwen3.6 27B Alibaba (Qwen)

42.2

Rank #117 Confirmed

Summary

  • The widest gap is in knowledge, where Qwen3.6 27B leads 52.4 to 16.6.
  • Qwen3.6 27B has downloadable open weights; the other is API-only.

Side by side

Mistral and Qwen3.6 27B specifications
MistralQwen3.6 27B
ProviderMistral AIAlibaba (Qwen)
Noometry Index29.942.2
Released—2026-04-22
WeightsProprietaryOpen
Context window—262K
Max output—66K
Input $ / M tokens—$0.60
Output $ / M tokens—$3.60
Results tracked2211

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 27B leads

Mistral: 33.8 (#250), Qwen3.6 27B: 39.1 (#163)

Coding benchmarks
BenchmarkMistralQwen3.6 27B
SciCode—37.3%
LMArena Coding1162—

Reasoning Qwen3.6 27B leads

Mistral: 22.2 (#200), Qwen3.6 27B: 25.0 (#153)

Reasoning benchmarks
BenchmarkMistralQwen3.6 27B
CritPt—0.9%
Chess Puzzles—22%
LMArena Hard Prompts1149—
Mystery Game Puzzles—7%
DTBench—78.1%
LMCA—34.5%
Epoch Capabilities Index—146.5

Math Qwen3.6 27B leads

Mistral: 22.3 (#278), Qwen3.6 27B: 48.5 (#62)

Math benchmarks
BenchmarkMistralQwen3.6 27B
FrontierMath (Tiers 1-3)—35.1%
OTIS Mock AIME 2024-2025—91.1%
Omni-MATH7.2%—
LMArena Math1180—

Knowledge Qwen3.6 27B leads

Mistral: 16.6 (#288), Qwen3.6 27B: 52.4 (#63)

Knowledge benchmarks
BenchmarkMistralQwen3.6 27B
GPQA Diamond—85.9%
MMLU-Pro27.7%—
GPQA (HELM)30.3%—
LMArena Expert1125—

Multilingual Not comparable

Mistral: 32.8 (#254), Qwen3.6 27B: —

Multilingual benchmarks
BenchmarkMistralQwen3.6 27B
LMArena Non-English1129—
LMArena Chinese1109—
LMArena French1180—
LMArena German1155—
LMArena Japanese1013—
LMArena Korean1032—
LMArena Russian1168—
LMArena Spanish1143—

Instruction Following Not comparable

Mistral: 52.6 (#288), Qwen3.6 27B: —

Instruction Following benchmarks
BenchmarkMistralQwen3.6 27B
IFEval56.8%—
LMArena Instruction Following1152—

Long Context Not comparable

Mistral: 35.0 (#245), Qwen3.6 27B: —

Long Context benchmarks
BenchmarkMistralQwen3.6 27B
LMArena Longer Query1153—

Writing & Preference Qwen3.6 27B leads

Mistral: 37.0 (#260), Qwen3.6 27B: 50.3 (#181)

Writing & Preference benchmarks
BenchmarkMistralQwen3.6 27B
LMArena Text1165—
LMArena Creative Writing1158—
WildBench66%—
EQ-Bench 4—1026
LMArena Multi-Turn1147—

Frequently asked questions

Is Mistral better than Qwen3.6 27B?

Qwen3.6 27B is the stronger model overall, scoring 42.2 to 29.9 on the Noometry Index.

Is Mistral or Qwen3.6 27B better for coding?

Qwen3.6 27B scores higher on coding benchmarks: 39.1 versus 33.8 in the Noometry coding category.

How many benchmarks do Mistral and Qwen3.6 27B share?

0 benchmarks have published results for both models. Mistral has 22 scored results on Noometry and Qwen3.6 27B has 11.

Related comparisons

Go deeper