Model comparison

MiniMax M1 vs Mistral Large

MiniMax M1 is the stronger model overall, scoring 40.3 to 31.9 on the Noometry Index.

Last verified . 17 shared benchmarks.

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Mistral Large Mistral AI

31.9

Rank #263 Confirmed

Summary

  • They share 17 benchmarks with published results for both. MiniMax M1 scores higher in 8 categories and Mistral Large in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where MiniMax M1 leads 37.5 to 18.2.
  • MiniMax M1 is cheaper at $0.55 / $2.20 per million input/output tokens, against $2 / $6 for Mistral Large.
  • MiniMax M1 accepts more context: 1M tokens versus 131K.

Side by side

MiniMax M1 and Mistral Large specifications
MiniMax M1Mistral Large
ProviderMiniMaxMistral AI
Noometry Index40.331.9
Released2025-06-132024-02-26
WeightsOpenOpen
Context window1M131K
Max output40K16K
Input $ / M tokens$0.55$2
Output $ / M tokens$2.20$6
Results tracked1851

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiniMax M1 leads

MiniMax M1: 39.9 (#153), Mistral Large: 34.3 (#240)

Coding benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Coding13591277
SciCode—36.2%
BigCodeBench Instruct—30%
LiveBench Coding—47.1%
BigCodeBench Complete—38.3%
ALE-Bench—264.7
HumanEval+—62.2%
MBPP+—59.5%

Agentic & Tool Use Not comparable

MiniMax M1: —, Mistral Large: 28.6 (#89)

Agentic & Tool Use benchmarks
BenchmarkMiniMax M1Mistral Large
Berkeley Function Calling Leaderboard—38.4%

Reasoning MiniMax M1 leads

MiniMax M1: 26.9 (#126), Mistral Large: 15.8 (#310)

Reasoning benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Hard Prompts13391257
SimpleBench—22.5%
CritPt—0%
LiveBench Reasoning—43.5%
DTBench—65.1%
LiveBench Data Analysis—50.1%
LMCA—16.7%
Epoch Capabilities Index—128.52
ForecastBench—57.1
LiveBench—48.4%

Math MiniMax M1 leads

MiniMax M1: 37.5 (#151), Mistral Large: 18.2 (#291)

Math benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Math13611262
OTIS Mock AIME 2024-2025—8.5%
Omni-MATH—28.1%
LiveBench Math—42.5%
MATH Level 5—50.3%
FrontierMath (Feb 2025 set)—0.3%

Knowledge MiniMax M1 leads

MiniMax M1: 36.4 (#170), Mistral Large: 30.1 (#230)

Knowledge benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Expert13171232
GPQA Diamond—51.3%
MMLU-Pro—59.9%
Confabulations—21.4%
Vectara Hallucination Rate—4.5%
GPQA (HELM)—43.5%
MMLU—80%

Multilingual MiniMax M1 leads

MiniMax M1: 45.8 (#163), Mistral Large: 40.0 (#219)

Multilingual benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Non-English13191237
LMArena Chinese13601240
LMArena French13701325
LMArena German13501254
LMArena Japanese12171188
LMArena Korean12661202
LMArena Russian13291257
LMArena Spanish13531268

Instruction Following MiniMax M1 leads

MiniMax M1: 69.3 (#174), Mistral Large: 67.9 (#191)

Instruction Following benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Instruction Following13121249
LiveBench Instruction Following—67.9%
IFEval—87.7%

Long Context MiniMax M1 leads

MiniMax M1: 41.4 (#141), Mistral Large: 38.3 (#199)

Long Context benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Longer Query13261261
Fiction.LiveBench69.4%—

Writing & Preference MiniMax M1 leads

MiniMax M1: 53.1 (#161), Mistral Large: 40.7 (#242)

Writing & Preference benchmarks
BenchmarkMiniMax M1Mistral Large
LMArena Text13431266
LMArena Creative Writing12981243
LMArena Multi-Turn13351260
Short-Story Creative Writing—69%
EQ-Bench Creative Writing—985
WildBench—80.1%
LiveBench Language—39.4%

Frequently asked questions

Is MiniMax M1 better than Mistral Large?

MiniMax M1 is the stronger model overall, scoring 40.3 to 31.9 on the Noometry Index.

Which is cheaper, MiniMax M1 or Mistral Large?

MiniMax M1 is cheaper. It lists at $0.55 per million input tokens and $2.20 per million output tokens; Mistral Large lists at $2 and $6.

Is MiniMax M1 or Mistral Large better for coding?

MiniMax M1 scores higher on coding benchmarks: 39.9 versus 34.3 in the Noometry coding category.

Which has the bigger context window?

MiniMax M1 does, with 1M tokens against 131K.

How many benchmarks do MiniMax M1 and Mistral Large share?

17 benchmarks have published results for both models. MiniMax M1 has 18 scored results on Noometry and Mistral Large has 51.

Related comparisons

Go deeper