Model comparison
Mistral Large vs Nova Premier 1.0
Nova Premier 1.0 is the stronger model overall, scoring 38.3 to 31.9 on the Noometry Index. Mistral Large costs 1.7× less per token, which makes it the better buy when Nova Premier 1.0's lead doesn't matter for your workload.
Last verified . 5 shared benchmarks.
Summary
- They share 5 benchmarks with published results for both. Mistral Large scores higher in 1 category and Nova Premier 1.0 in 4 categories; 5 gaps are clear of the uncertainty.
- The widest gap is in math, where Nova Premier 1.0 leads 33.8 to 18.2.
- The biggest single-benchmark swing is MMLU-Pro: 59.9% for Mistral Large and 72.6% for Nova Premier 1.0.
- Mistral Large is cheaper at $2 / $6 per million input/output tokens, against $2.50 / $12.50 for Nova Premier 1.0.
- Nova Premier 1.0 accepts more context: 1M tokens versus 131K.
- Mistral Large has downloadable open weights; the other is API-only.
Side by side
| Mistral Large | Nova Premier 1.0 | |
|---|---|---|
| Provider | Mistral AI | Amazon |
| Noometry Index | 31.9 | 38.3 |
| Released | 2024-02-26 | 2025-04-30 |
| Weights | Open | Proprietary |
| Context window | 131K | 1M |
| Max output | 16K | 10K |
| Input $ / M tokens | $2 | $2.50 |
| Output $ / M tokens | $6 | $12.50 |
| Results tracked | 51 | 6 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Mistral Large: 34.3 (#240), Nova Premier 1.0: —
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| SciCode | 36.2% | — |
| BigCodeBench Instruct | 30% | — |
| LiveBench Coding | 47.1% | — |
| LMArena Coding | 1277 | — |
| BigCodeBench Complete | 38.3% | — |
| ALE-Bench | 264.7 | — |
| HumanEval+ | 62.2% | — |
| MBPP+ | 59.5% | — |
Agentic & Tool Use Not comparable
Mistral Large: 28.6 (#89), Nova Premier 1.0: —
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| Berkeley Function Calling Leaderboard | 38.4% | — |
Reasoning Nova Premier 1.0 leads
Mistral Large: 15.8 (#310), Nova Premier 1.0: 23.0 (#185)
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| SimpleBench | 22.5% | — |
| Kagi LLM Benchmark | — | 44.8% |
| CritPt | 0% | — |
| LiveBench Reasoning | 43.5% | — |
| LMArena Hard Prompts | 1257 | — |
| DTBench | 65.1% | — |
| LiveBench Data Analysis | 50.1% | — |
| LMCA | 16.7% | — |
| Epoch Capabilities Index | 128.52 | — |
| ForecastBench | 57.1 | — |
| LiveBench | 48.4% | — |
Math Nova Premier 1.0 leads
Mistral Large: 18.2 (#291), Nova Premier 1.0: 33.8 (#199)
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| Omni-MATH | 28.1% | 35% |
| OTIS Mock AIME 2024-2025 | 8.5% | — |
| LiveBench Math | 42.5% | — |
| LMArena Math | 1262 | — |
| MATH Level 5 | 50.3% | — |
| FrontierMath (Feb 2025 set) | 0.3% | — |
Knowledge Nova Premier 1.0 leads
Mistral Large: 30.1 (#230), Nova Premier 1.0: 35.8 (#180)
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| MMLU-Pro | 59.9% | 72.6% |
| GPQA (HELM) | 43.5% | 51.8% |
| GPQA Diamond | 51.3% | — |
| Confabulations | 21.4% | — |
| Vectara Hallucination Rate | 4.5% | — |
| LMArena Expert | 1232 | — |
| MMLU | 80% | — |
Multilingual Not comparable
Mistral Large: 40.0 (#219), Nova Premier 1.0: —
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| LMArena Non-English | 1237 | — |
| LMArena Chinese | 1240 | — |
| LMArena French | 1325 | — |
| LMArena German | 1254 | — |
| LMArena Japanese | 1188 | — |
| LMArena Korean | 1202 | — |
| LMArena Russian | 1257 | — |
| LMArena Spanish | 1268 | — |
Instruction Following Mistral Large leads
Mistral Large: 67.9 (#191), Nova Premier 1.0: 66.8 (#204)
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| IFEval | 87.7% | 80.3% |
| LiveBench Instruction Following | 67.9% | — |
| LMArena Instruction Following | 1249 | — |
Long Context Not comparable
Mistral Large: 38.3 (#199), Nova Premier 1.0: —
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| LMArena Longer Query | 1261 | — |
Writing & Preference Nova Premier 1.0 leads
Mistral Large: 40.7 (#242), Nova Premier 1.0: 51.0 (#178)
| Benchmark | Mistral Large | Nova Premier 1.0 |
|---|---|---|
| WildBench | 80.1% | 78.8% |
| LMArena Text | 1266 | — |
| LMArena Creative Writing | 1243 | — |
| Short-Story Creative Writing | 69% | — |
| EQ-Bench Creative Writing | 985 | — |
| LMArena Multi-Turn | 1260 | — |
| LiveBench Language | 39.4% | — |
Frequently asked questions
Is Mistral Large better than Nova Premier 1.0?
Nova Premier 1.0 is the stronger model overall, scoring 38.3 to 31.9 on the Noometry Index. Mistral Large costs 1.7× less per token, which makes it the better buy when Nova Premier 1.0's lead doesn't matter for your workload.
Which is cheaper, Mistral Large or Nova Premier 1.0?
Mistral Large is cheaper. It lists at $2 per million input tokens and $6 per million output tokens; Nova Premier 1.0 lists at $2.50 and $12.50.
Which has the bigger context window?
Nova Premier 1.0 does, with 1M tokens against 131K.
How many benchmarks do Mistral Large and Nova Premier 1.0 share?
5 benchmarks have published results for both models. Mistral Large has 51 scored results on Noometry and Nova Premier 1.0 has 6.