Mistral AI, open weights
Mistral Large 3
Mistral Large 3 by Mistral AI ranks 176th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.1. Its strongest category is multimodal, where it ranks 66th. API pricing starts at $0.25 per million input tokens and $0.75 per million output tokens, with a 262K-token context window.
Last verified
Specifications
- Noometry rank
- #176 of 354
- Index score
- 39.1
- Evidence
- Confirmed 24 results
- Provider
Mistral AI
- Released
- December 2, 2025
- Weights
- Open weights
- Reasoning
- No
- Context window
- 262K
- Max output
- 8K
- Input price
- $0.25 / M
- Output price
- $0.75 / M
- Blended price
- $0.38 / M
- Output speed
- 7 tokens/s Kagi
- Value
- #57 of 219
- Knowledge cutoff
- November 2024
- Input
- text, image
Category scores
Each category score combines every public result we have in that category.
- Coding 34.4
- Reasoning 15.2
- Math 38.7
- Knowledge 36.0
- Multimodal 38.2
- Multilingual 52.5
- Instruction Following 74.0
- Long Context 43.1
- Writing & Preference 60.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 34.4 | #237 | 2 |
| Reasoning | 15.2 | #319 | 4 |
| Math | 38.7 | #129 | 1 |
| Knowledge | 36.0 | #177 | 2 |
| Multimodal | 38.2 | #66 | 1 |
| Multilingual | 52.5 | #84 | 1 |
| Instruction Following | 74.0 | #108 | 1 |
| Long Context | 43.1 | #105 | 1 |
| Writing & Preference | 60.0 | #101 | 4 |
Strengths and weaknesses
Categories where Mistral Large 3 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 52.5 | +5.1 | #84 of 297, top 29% |
| Writing & Preference | 60.0 | +6.2 | #101 of 312, top 33% |
| Instruction Following | 74.0 | +2.8 | #108 of 305, top 36% |
Closest competitors
The models ranked just above and below Mistral Large 3. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Qwen3 32B | #172 | 39.2 | $1.22 | 86 | Compare |
| Gemini 2.0 Pro | #173 | 39.1 | — | — | Compare |
| Molmo 2 8b | #174 | 39.1 | — | — | Compare |
| Mercury 2 | #175 | 39.1 | $0.38 | — | Compare |
| GLM-4.5-Air | #177 | 38.9 | $0.43 | 160 | Compare |
| MiniMax-M2.1 | #178 | 38.9 | $0.52 | — | Compare |
| Qwen3-30B-A3B | #179 | 38.9 | $0.21 | 42 | Compare |
| GLM-4.7-Flash | #180 | 38.8 | $0.15 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena WebDev | 1230 | #106 of 113, top 94% | LMArena | 2026-10-08 | |
| LMArena Coding | 1448 | #88 of 294, top 30% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Kagi LLM Benchmark | 50.9% | #63 of 99, top 64% | Kagi LLM Benchmark | ||
| NYT Connections (extended) | 7.5% | #90 of 91, top 99% | Lech Mazur benchmarks | ||
| Thematic Generalization | 23% | #22 of 23, top 96% | Lech Mazur benchmarks | ||
| LMArena Hard Prompts | 1429 | #89 of 297, top 30% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1414 | #110 of 285, top 39% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Vectara Hallucination Rate (lower is better) | 14.5% | #85 of 96, top 89% | Vectara Hallucination Leaderboard | ||
| LMArena Expert | 1421 | #107 of 273, top 40% | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1221 | #71 of 122, top 59% | LMArena | 2026-10-09 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1413 | #84 of 297, top 29% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1447 | #105 of 285, top 37% | LMArena | 2026-10-08 | |
| LMArena French | 1455 | #63 of 223, top 29% | LMArena | 2026-10-08 | |
| LMArena German | 1437 | #61 of 231, top 27% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1394 | #64 of 211, top 31% | LMArena | 2026-10-08 | |
| LMArena Korean | 1384 | #72 of 213, top 34% | LMArena | 2026-10-08 | |
| LMArena Russian | 1411 | #92 of 283, top 33% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1440 | #67 of 226, top 30% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1403 | #102 of 298, top 35% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1413 | #100 of 291, top 35% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1428 | #83 of 297, top 28% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1386 | #100 of 295, top 34% | LMArena | 2026-10-08 | |
| EQ-Bench Creative Writing | 1412 | #69 of 115, top 60% | EQ-Bench | ||
| LMArena Multi-Turn | 1429 | #83 of 295, top 29% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| bedrock | $0.50 | $1.50 | — | 2026-10-10 |
| openrouter | $0.25 | $0.75 | $0.025 | 2026-10-10 |
Compare Mistral Large 3
- Mistral Large 3 vs Mistral Large
- Mistral Large 3 vs Mercury 2
- Mistral Large 3 vs GLM-4.5-Air
- Mistral Large 3 vs Molmo 2 8b
- Mistral Large 3 vs MiniMax-M2.1
- Mistral Large 3 vs Gemini 2.0 Pro
- Mistral Large 3 vs Qwen3-30B-A3B
- Mistral Large 3 vs GPT-6 Astra
- Mistral Large 3 vs Claude Fable 5.1
- Mistral Large 3 vs Gemini 3.8 Flash
- Mistral Large 3 vs Kimi K3
- Mistral Large 3 vs Grok 4.6
- Mistral Large 3 vs Qwen3.8 Max
- Mistral Large 3 vs GLM-5.3
Other Mistral AI models
- Mistral Large 443.1
- Mistral Medium 3.540.2
- Mistral Medium36.3
- Magistral Medium35.2
- Devstral Small 250534.3
- Mistral Small33.4
- Pixtral Large32.2
- Mistral Large31.9
Frequently asked questions
How good is Mistral Large 3?
Mistral Large 3 by Mistral AI ranks 176th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.1. Its strongest category is multimodal, where it ranks 66th. API pricing starts at $0.25 per million input tokens and $0.75 per million output tokens, with a 262K-token context window.
How much does Mistral Large 3 cost?
Mistral Large 3 costs $0.25 per million input tokens and $0.75 per million output tokens on openrouter, with cached input at $0.025.
What is Mistral Large 3's context window?
Mistral Large 3 accepts up to 262K tokens of input and can write up to 8K tokens in one response.
Is Mistral Large 3 open source?
Yes. Mistral Large 3's weights are downloadable; check the license for commercial terms.
How fast is Mistral Large 3?
Mistral Large 3 generated about 7 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.
What are Mistral Large 3's strengths and weaknesses?
Relative to other ranked models, Mistral Large 3 places best in multilingual, writing & preference, instruction following and lowest in reasoning, coding, knowledge.
What is Mistral Large 3 best at?
Its best category is multimodal, where it ranks 66th on Noometry.