Model comparison

Devstral Small 2505 vs MiniMax M1

MiniMax M1 is the stronger model overall, scoring 40.3 to 34.3 on the Noometry Index. Devstral Small 2505 costs 6.4× less per token, which makes it the better buy when MiniMax M1's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Devstral Small 2505 Mistral AI

34.3

Rank #233 Reported

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Summary

  • The widest gap is in reasoning, where MiniMax M1 leads 26.9 to 19.7.
  • Devstral Small 2505 is cheaper at $0.10 / $0.30 per million input/output tokens, against $0.55 / $2.20 for MiniMax M1.
  • MiniMax M1 accepts more context: 1M tokens versus 128K.

Side by side

Devstral Small 2505 and MiniMax M1 specifications
Devstral Small 2505MiniMax M1
ProviderMistral AIMiniMax
Noometry Index34.340.3
Released2025-05-072025-06-13
WeightsOpenOpen
Context window128K1M
Max output128K40K
Input $ / M tokens$0.10$0.55
Output $ / M tokens$0.30$2.20
Results tracked418

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Devstral Small 2505: 38.9 (#166), MiniMax M1: 39.9 (#153)

Coding benchmarks
BenchmarkDevstral Small 2505MiniMax M1
SWE-bench Verified (bash only)56.4%—
SciCode28.8%—
LMArena Coding—1359

Reasoning MiniMax M1 leads

Devstral Small 2505: 19.7 (#252), MiniMax M1: 26.9 (#126)

Reasoning benchmarks
BenchmarkDevstral Small 2505MiniMax M1
Kagi LLM Benchmark37.7%—
CritPt0%—
LMArena Hard Prompts—1339

Math Not comparable

Devstral Small 2505: —, MiniMax M1: 37.5 (#151)

Math benchmarks
BenchmarkDevstral Small 2505MiniMax M1
LMArena Math—1361

Knowledge Not comparable

Devstral Small 2505: —, MiniMax M1: 36.4 (#170)

Knowledge benchmarks
BenchmarkDevstral Small 2505MiniMax M1
LMArena Expert—1317

Multilingual Not comparable

Devstral Small 2505: —, MiniMax M1: 45.8 (#163)

Multilingual benchmarks
BenchmarkDevstral Small 2505MiniMax M1
LMArena Non-English—1319
LMArena Chinese—1360
LMArena French—1370
LMArena German—1350
LMArena Japanese—1217
LMArena Korean—1266
LMArena Russian—1329
LMArena Spanish—1353

Instruction Following Not comparable

Devstral Small 2505: —, MiniMax M1: 69.3 (#174)

Instruction Following benchmarks
BenchmarkDevstral Small 2505MiniMax M1
LMArena Instruction Following—1312

Long Context Not comparable

Devstral Small 2505: —, MiniMax M1: 41.4 (#141)

Long Context benchmarks
BenchmarkDevstral Small 2505MiniMax M1
Fiction.LiveBench—69.4%
LMArena Longer Query—1326

Writing & Preference Not comparable

Devstral Small 2505: —, MiniMax M1: 53.1 (#161)

Writing & Preference benchmarks
BenchmarkDevstral Small 2505MiniMax M1
LMArena Text—1343
LMArena Creative Writing—1298
LMArena Multi-Turn—1335

Frequently asked questions

Is Devstral Small 2505 better than MiniMax M1?

MiniMax M1 is the stronger model overall, scoring 40.3 to 34.3 on the Noometry Index. Devstral Small 2505 costs 6.4× less per token, which makes it the better buy when MiniMax M1's lead doesn't matter for your workload.

Which is cheaper, Devstral Small 2505 or MiniMax M1?

Devstral Small 2505 is cheaper. It lists at $0.10 per million input tokens and $0.30 per million output tokens; MiniMax M1 lists at $0.55 and $2.20.

Is Devstral Small 2505 or MiniMax M1 better for coding?

They score almost the same on coding (38.9 vs 39.9); test both on your own repository before choosing.

Which has the bigger context window?

MiniMax M1 does, with 1M tokens against 128K.

How many benchmarks do Devstral Small 2505 and MiniMax M1 share?

0 benchmarks have published results for both models. Devstral Small 2505 has 4 scored results on Noometry and MiniMax M1 has 18.

Related comparisons

Go deeper