# Mixtral 8x22B vs Qwen Turbo

> Mixtral 8x22B and Qwen Turbo score almost the same on the Noometry Index (27.1 vs 27.1), so choose on price, context window or the category you care about most.

- Canonical page: https://noometry.com/compare/mixtral-8x22b-vs-qwen-turbo
- Last updated: 2026-10-10
- Shared benchmarks: 2

## Summary

- They share 2 benchmarks with published results for both. Mixtral 8x22B scores higher in 1 category and Qwen Turbo in 1 category; 2 gaps are clear of the uncertainty.
- The widest gap is in math, where Mixtral 8x22B leads 22.9 to 15.3.
- The biggest single-benchmark swing is MATH Level 5: 24.2% for Mixtral 8x22B and 56.2% for Qwen Turbo.
- Qwen Turbo is cheaper at $0.05 / $0.20 per million input/output tokens, against $2 / $6 for Mixtral 8x22B.
- Qwen Turbo accepts more context: 1M tokens versus 64K.
- Mixtral 8x22B has downloadable open weights; the other is API-only.

## Snapshot

| | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| Provider | Mistral AI | Alibaba (Qwen) |
| Noometry Index | 27.1 | 27.1 |
| Rank | 333 | 335 |
| Context | 64K | 1M |
| Input $/M | $2 | $0.05 |
| Output $/M | $6 | $0.20 |
| Weights | Open | Proprietary |

## Coding

- Mixtral 8x22B: 24.2 (#329)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| WeirdML | 3.2% | — |
| BigCodeBench Instruct | 40.6% | — |
| LMArena Coding | 1166 | — |
| BigCodeBench Complete | 50.2% | — |
| HumanEval+ | 72% | — |
| MBPP+ | 64.3% | — |

## Agentic & Tool Use

- Mixtral 8x22B: 23.1 (#127)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| Cybench | 7.5% | — |

## Reasoning

- Mixtral 8x22B: 19.9 (#248)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| LMArena Hard Prompts | 1150 | — |
| DTBench | 55.1% | — |
| Epoch Capabilities Index | 122.03 | — |
| ForecastBench | 56.3 | — |

## Math

- Mixtral 8x22B: 22.9 (#275)
- Qwen Turbo: 15.3 (#297)

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| MATH Level 5 | 24.2% | 56.2% |
| OTIS Mock AIME 2024-2025 | — | 6.1% |
| Omni-MATH | 16.3% | — |
| LMArena Math | 1184 | — |

## Knowledge

- Mixtral 8x22B: 15.1 (#293)
- Qwen Turbo: 22.2 (#272)

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| GPQA Diamond | 34.1% | 41.8% |
| MMLU-Pro | 46% | — |
| GPQA (HELM) | 33.4% | — |
| LMArena Expert | 1113 | — |
| MMLU | 77.8% | — |

## Multilingual

- Mixtral 8x22B: 32.8 (#255)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| LMArena Non-English | 1128 | — |
| LMArena Chinese | 1116 | — |
| LMArena French | 1166 | — |
| LMArena German | 1141 | — |
| LMArena Japanese | 1037 | — |
| LMArena Korean | 1057 | — |
| LMArena Russian | 1158 | — |
| LMArena Spanish | 1151 | — |

## Instruction Following

- Mixtral 8x22B: 57.7 (#266)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| IFEval | 72.4% | — |
| LMArena Instruction Following | 1147 | — |

## Long Context

- Mixtral 8x22B: 34.7 (#247)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| LMArena Longer Query | 1144 | — |

## Writing & Preference

- Mixtral 8x22B: 36.9 (#262)
- Qwen Turbo: —

| Benchmark | Mixtral 8x22B | Qwen Turbo |
|---|---|---|
| LMArena Text | 1162 | — |
| LMArena Creative Writing | 1141 | — |
| WildBench | 71.1% | — |
| LMArena Multi-Turn | 1130 | — |

## FAQ

### Is Mixtral 8x22B better than Qwen Turbo?

Mixtral 8x22B and Qwen Turbo score almost the same on the Noometry Index (27.1 vs 27.1), so choose on price, context window or the category you care about most.

### Which is cheaper, Mixtral 8x22B or Qwen Turbo?

Qwen Turbo is cheaper. It lists at $0.05 per million input tokens and $0.20 per million output tokens; Mixtral 8x22B lists at $2 and $6.

### Which has the bigger context window?

Qwen Turbo does, with 1M tokens against 64K.

### How many benchmarks do Mixtral 8x22B and Qwen Turbo share?

2 benchmarks have published results for both models. Mixtral 8x22B has 34 scored results on Noometry and Qwen Turbo has 3.
