Mistral AI, open weights
Mixtral 8x7B
Mixtral 8x7B by Mistral AI ranks 334th of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.1. Its strongest category is long context, where it ranks 260th. API pricing starts at $0.70 per million input tokens and $0.70 per million output tokens, with a 32K-token context window.
Last verified
Specifications
- Noometry rank
- #334 of 354
- Index score
- 27.1
- Evidence
- Confirmed 38 results
- Provider
Mistral AI
- Released
- December 11, 2023
- Weights
- Open weights
- Reasoning
- No
- Context window
- 32K
- Max output
- 32K
- Input price
- $0.70 / M
- Output price
- $0.70 / M
- Blended price
- $0.70 / M
- Output speed
- Not measured
- Value
- #121 of 219
- Knowledge cutoff
- January 2024
- Input
- text
Category scores
Each category score combines every public result we have in that category.
- Coding 32.8
- Reasoning 18.2
- Math 18.8
- Knowledge 11.0
- Multilingual 29.6
- Instruction Following 51.0
- Long Context 33.4
- Writing & Preference 34.2
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 32.8 | #269 | 1 |
| Reasoning | 18.2 | #285 | 2 |
| Math | 18.8 | #289 | 3 |
| Knowledge | 11.0 | #301 | 4 |
| Multilingual | 29.6 | #266 | 1 |
| Instruction Following | 51.0 | #297 | 2 |
| Long Context | 33.4 | #260 | 1 |
| Writing & Preference | 34.2 | #270 | 4 |
Strengths and weaknesses
Categories where Mixtral 8x7B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 32.8 | −6.0 | #269 of 340, top 80% |
| Reasoning | 18.2 | −5.4 | #285 of 350, top 82% |
| Writing & Preference | 34.2 | −19.6 | #270 of 312, top 87% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Instruction Following | 51.0 | −20.3 | #297 of 305, top 98% |
| Knowledge | 11.0 | −26.3 | #301 of 314, top 96% |
| Multilingual | 29.6 | −17.8 | #266 of 297, top 90% |
Closest competitors
The models ranked just above and below Mixtral 8x7B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Llama 4 Scout | #330 | 27.7 | $0.15 | 272 | Compare |
| Llama 3.2 90B | #331 | 27.5 | — | — | Compare |
| Gemini 1.0 Pro | #332 | 27.3 | — | — | Compare |
| Mixtral 8x22B | #333 | 27.1 | $3 | — | Compare |
| Qwen Turbo | #335 | 27.1 | $0.0875 | — | Compare |
| Qwen3-1.7B | #336 | 26.6 | — | — | Compare |
| Mistral Nemo | #337 | 26.4 | $0.15 | — | Compare |
| Ministral 3B | #338 | 26.2 | $0.10 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1126 | #263 of 294, top 90% | LMArena | 2026-10-08 | |
| HumanEval+ | 39.6% | #37 of 45, top 83% | EvalPlus | ||
| MBPP+ | 49.7% | #33 of 38, top 87% | EvalPlus |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1115 | #261 of 297, top 88% | LMArena | 2026-10-08 | |
| DTBench | 49.6% | #138 of 151, top 92% | Epoch AI | ||
| Adversarial NLI | 55.2% | #5 of 9, top 56% | Epoch AI | ||
| Epoch Capabilities Index | 118.47 | #176 of 213, top 83% | Epoch AI | 2023-12-11 | |
| ForecastBench | 56.3 | #65 of 72, top 91% | Epoch AI | ||
| HellaSwag | 86.7% | #7 of 29, top 25% | Epoch AI | ||
| PIQA | 83.6% | #9 of 27, top 34% | Epoch AI | ||
| WinoGrande | 77.2% | #19 of 43, top 45% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Omni-MATH | 10.5% | #56 of 57, top 99% | HELM Capabilities | ||
| LMArena Math | 1147 | #253 of 285, top 89% | LMArena | 2026-10-08 | |
| MATH Level 5 | 10% | #74 of 79, top 94% | Epoch AI | 2025-01-27 | |
| MATH Level 5 | 9.3% | Epoch AI | 2025-01-27 | ||
| GSM8K | 74.4% | #14 of 38, top 37% | Epoch AI |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 29.8% | Epoch AI | 2025-01-27 | ||
| GPQA Diamond | 30.6% | #169 of 186, top 91% | Epoch AI | 2025-01-27 | |
| MMLU-Pro | 33.5% | #56 of 58, top 97% | HELM Capabilities | ||
| GPQA (HELM) | 29.6% | #55 of 57, top 97% | HELM Capabilities | ||
| LMArena Expert | 1088 | #253 of 273, top 93% | LMArena | 2026-10-08 | |
| ARC (AI2) Challenge | 87.3% | #8 of 39, top 21% | Epoch AI | ||
| MMLU | 70.6% | #44 of 81, top 55% | Epoch AI | ||
| OpenBookQA | 85.8% | #5 of 19, top 27% | Epoch AI | ||
| TriviaQA | 82.2% | #9 of 25, top 36% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1077 | #266 of 297, top 90% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1055 | #266 of 285, top 94% | LMArena | 2026-10-08 | |
| LMArena French | 1166 | #200 of 223, top 90% | LMArena | 2026-10-08 | |
| LMArena German | 1114 | #208 of 231, top 91% | LMArena | 2026-10-08 | |
| LMArena Japanese | 931 | #207 of 211, top 99% | LMArena | 2026-10-08 | |
| LMArena Korean | 968 | #204 of 213, top 96% | LMArena | 2026-10-08 | |
| LMArena Russian | 1090 | #260 of 283, top 92% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1111 | #214 of 226, top 95% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| IFEval | 57.5% | #56 of 57, top 99% | HELM Capabilities | ||
| LMArena Instruction Following | 1109 | #266 of 298, top 90% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1103 | #266 of 291, top 92% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1132 | #262 of 297, top 89% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1109 | #261 of 295, top 89% | LMArena | 2026-10-08 | |
| WildBench | 67.3% | #54 of 57, top 95% | HELM Capabilities | ||
| LMArena Multi-Turn | 1115 | #260 of 295, top 89% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| mistral | $0.70 | $0.70 | — | 2026-10-10 |
Compare Mixtral 8x7B
- Mixtral 8x7B vs Mixtral 8x22B
- Mixtral 8x7B vs Qwen Turbo
- Mixtral 8x7B vs Gemini 1.0 Pro
- Mixtral 8x7B vs Qwen3-1.7B
- Mixtral 8x7B vs Llama 3.2 90B
- Mixtral 8x7B vs Mistral Nemo
- Mixtral 8x7B vs GPT-6 Astra
- Mixtral 8x7B vs Claude Fable 5.1
- Mixtral 8x7B vs Gemini 3.8 Flash
- Mixtral 8x7B vs Kimi K3
- Mixtral 8x7B vs Grok 4.6
- Mixtral 8x7B vs Qwen3.8 Max
- Mixtral 8x7B vs GLM-5.3
- Mixtral 8x7B vs Muse Spark 1.3
Other Mistral AI models
- Mistral Large 443.1
- Mistral Medium 3.540.2
- Mistral Large 339.1
- Mistral Medium36.3
- Magistral Medium35.2
- Devstral Small 250534.3
- Mistral Small33.4
- Pixtral Large32.2
Frequently asked questions
How good is Mixtral 8x7B?
Mixtral 8x7B by Mistral AI ranks 334th of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.1. Its strongest category is long context, where it ranks 260th. API pricing starts at $0.70 per million input tokens and $0.70 per million output tokens, with a 32K-token context window.
How much does Mixtral 8x7B cost?
Mixtral 8x7B costs $0.70 per million input tokens and $0.70 per million output tokens on Mistral AI's own API.
What is Mixtral 8x7B's context window?
Mixtral 8x7B accepts up to 32K tokens of input and can write up to 32K tokens in one response.
Is Mixtral 8x7B open source?
Yes. Mixtral 8x7B's weights are downloadable; check the license for commercial terms.
What are Mixtral 8x7B's strengths and weaknesses?
Relative to other ranked models, Mixtral 8x7B places best in coding, reasoning, writing & preference and lowest in instruction following, knowledge, multilingual.
What is Mixtral 8x7B best at?
Its best category is long context, where it ranks 260th on Noometry.