Model comparison
Mixtral 8x22B vs Pixtral Large
Pixtral Large is the stronger model overall, scoring 32.2 to 27.1 on the Noometry Index.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in writing & preference, where Mixtral 8x22B leads 36.9 to 32.9.
- Both cost about the same: $2 input and $6 output per million tokens.
- Pixtral Large accepts more context: 128K tokens versus 64K.
Side by side
| Mixtral 8x22B | Pixtral Large | |
|---|---|---|
| Provider | Mistral AI | Mistral AI |
| Noometry Index | 27.1 | 32.2 |
| Released | 2024-04-17 | 2024-11-01 |
| Weights | Open | Open |
| Context window | 64K | 128K |
| Max output | 64K | 128K |
| Input $ / M tokens | $2 | $2 |
| Output $ / M tokens | $6 | $6 |
| Results tracked | 34 | 3 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Mixtral 8x22B: 24.2 (#329), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| WeirdML | 3.2% | — |
| BigCodeBench Instruct | 40.6% | — |
| LMArena Coding | 1166 | — |
| BigCodeBench Complete | 50.2% | — |
| HumanEval+ | 72% | — |
| MBPP+ | 64.3% | — |
Agentic & Tool Use Not comparable
Mixtral 8x22B: 23.1 (#127), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| Cybench | 7.5% | — |
Reasoning Pixtral Large leads
Mixtral 8x22B: 19.9 (#248), Pixtral Large: 21.7 (#218)
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| EnigmaEval | — | 0.8% |
| LMArena Hard Prompts | 1150 | — |
| DTBench | 55.1% | — |
| Epoch Capabilities Index | 122.03 | — |
| ForecastBench | 56.3 | — |
Math Not comparable
Mixtral 8x22B: 22.9 (#275), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| Omni-MATH | 16.3% | — |
| LMArena Math | 1184 | — |
| MATH Level 5 | 24.2% | — |
Knowledge Not comparable
Mixtral 8x22B: 15.1 (#293), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| GPQA Diamond | 34.1% | — |
| MMLU-Pro | 46% | — |
| GPQA (HELM) | 33.4% | — |
| LMArena Expert | 1113 | — |
| MMLU | 77.8% | — |
Multimodal Not comparable
Mixtral 8x22B: —, Pixtral Large: 30.6 (#111)
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| LMArena Vision | — | 1089 |
Multilingual Not comparable
Mixtral 8x22B: 32.8 (#255), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| LMArena Non-English | 1128 | — |
| LMArena Chinese | 1116 | — |
| LMArena French | 1166 | — |
| LMArena German | 1141 | — |
| LMArena Japanese | 1037 | — |
| LMArena Korean | 1057 | — |
| LMArena Russian | 1158 | — |
| LMArena Spanish | 1151 | — |
Instruction Following Not comparable
Mixtral 8x22B: 57.7 (#266), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| IFEval | 72.4% | — |
| LMArena Instruction Following | 1147 | — |
Long Context Not comparable
Mixtral 8x22B: 34.7 (#247), Pixtral Large: —
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| LMArena Longer Query | 1144 | — |
Writing & Preference Mixtral 8x22B leads
Mixtral 8x22B: 36.9 (#262), Pixtral Large: 32.9 (#278)
| Benchmark | Mixtral 8x22B | Pixtral Large |
|---|---|---|
| LMArena Text | 1162 | — |
| LMArena Creative Writing | 1141 | — |
| EQ-Bench Creative Writing | — | 988 |
| WildBench | 71.1% | — |
| LMArena Multi-Turn | 1130 | — |
Frequently asked questions
Is Mixtral 8x22B better than Pixtral Large?
Pixtral Large is the stronger model overall, scoring 32.2 to 27.1 on the Noometry Index.
Which is cheaper, Mixtral 8x22B or Pixtral Large?
Pixtral Large is cheaper. It lists at $2 per million input tokens and $6 per million output tokens; Mixtral 8x22B lists at $2 and $6.
Which has the bigger context window?
Pixtral Large does, with 128K tokens against 64K.
How many benchmarks do Mixtral 8x22B and Pixtral Large share?
0 benchmarks have published results for both models. Mixtral 8x22B has 34 scored results on Noometry and Pixtral Large has 3.