Model comparison

Ministral 8B vs Mistral Large 3

Mistral Large 3 is the stronger model overall, scoring 39.1 to 28.2 on the Noometry Index. Ministral 8B costs 2.5× less per token, which makes it the better buy when Mistral Large 3's lead doesn't matter for your workload.

Last verified . 13 shared benchmarks.

Ministral 8B Mistral AI

28.2

Rank #325 Confirmed

Mistral Large 3 Mistral AI

39.1

Rank #176 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Ministral 8B scores higher in 2 categories and Mistral Large 3 in 6 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Mistral Large 3 leads 36.0 to 12.6.
  • The biggest single-benchmark swing is Vectara Hallucination Rate: 7.4% for Ministral 8B and 14.5% for Mistral Large 3.
  • Ministral 8B is cheaper at $0.15 / $0.15 per million input/output tokens, against $0.25 / $0.75 for Mistral Large 3.

Side by side

Ministral 8B and Mistral Large 3 specifications
Ministral 8BMistral Large 3
ProviderMistral AIMistral AI
Noometry Index28.239.1
Released2024-10-012025-12-02
WeightsOpenOpen
Context window262K262K
Max output262K8K
Input $ / M tokens$0.15$0.25
Output $ / M tokens$0.15$0.75
Results tracked1724

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Ministral 8B: 35.0 (#230), Mistral Large 3: 34.4 (#237)

Coding benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Coding12021448
LMArena WebDev—1230

Agentic & Tool Use Not comparable

Ministral 8B: 16.4 (#148), Mistral Large 3: —

Agentic & Tool Use benchmarks
BenchmarkMinistral 8BMistral Large 3
Berkeley Function Calling Leaderboard11.1%—

Reasoning Ministral 8B leads

Ministral 8B: 18.4 (#281), Mistral Large 3: 15.2 (#319)

Reasoning benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Hard Prompts11911429
Kagi LLM Benchmark—50.9%
NYT Connections (extended)—7.5%
Thematic Generalization—23%
DTBench45.7%—

Math Mistral Large 3 leads

Ministral 8B: 25.7 (#267), Mistral Large 3: 38.7 (#129)

Math benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Math11881414
MATH Level 514.9%—

Knowledge Mistral Large 3 leads

Ministral 8B: 12.6 (#297), Mistral Large 3: 36.0 (#177)

Knowledge benchmarks
BenchmarkMinistral 8BMistral Large 3
Vectara Hallucination Rate7.4%14.5%
LMArena Expert11701421
GPQA Diamond27.1%—

Multimodal Not comparable

Ministral 8B: —, Mistral Large 3: 38.2 (#66)

Multimodal benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Vision—1221

Multilingual Mistral Large 3 leads

Ministral 8B: 35.1 (#247), Mistral Large 3: 52.5 (#84)

Multilingual benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Non-English11651413
LMArena Chinese11931447
LMArena Russian11951411
LMArena French—1455
LMArena German—1437
LMArena Japanese—1394
LMArena Korean—1384
LMArena Spanish—1440

Instruction Following Mistral Large 3 leads

Ministral 8B: 60.5 (#250), Mistral Large 3: 74.0 (#108)

Instruction Following benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Instruction Following11611403

Long Context Mistral Large 3 leads

Ministral 8B: 36.7 (#227), Mistral Large 3: 43.1 (#105)

Long Context benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Longer Query12121413

Writing & Preference Mistral Large 3 leads

Ministral 8B: 39.6 (#246), Mistral Large 3: 60.0 (#101)

Writing & Preference benchmarks
BenchmarkMinistral 8BMistral Large 3
LMArena Text11911428
LMArena Creative Writing11751386
LMArena Multi-Turn11661429
EQ-Bench Creative Writing—1412

Frequently asked questions

Is Ministral 8B better than Mistral Large 3?

Mistral Large 3 is the stronger model overall, scoring 39.1 to 28.2 on the Noometry Index. Ministral 8B costs 2.5× less per token, which makes it the better buy when Mistral Large 3's lead doesn't matter for your workload.

Which is cheaper, Ministral 8B or Mistral Large 3?

Ministral 8B is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Mistral Large 3 lists at $0.25 and $0.75.

Is Ministral 8B or Mistral Large 3 better for coding?

They score almost the same on coding (35.0 vs 34.4); test both on your own repository before choosing.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Ministral 8B and Mistral Large 3 share?

13 benchmarks have published results for both models. Ministral 8B has 17 scored results on Noometry and Mistral Large 3 has 24.

Related comparisons

Go deeper