Model comparison

DeepSeek-V3 vs MiniMax M1

DeepSeek-V3 and MiniMax M1 score almost the same on the Noometry Index (39.5 vs 40.3), so choose on price, context window or the category you care about most.

Last verified . 18 shared benchmarks.

DeepSeek-V3 DeepSeek

39.5

Rank #166 Confirmed

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Summary

  • They share 18 benchmarks with published results for both. DeepSeek-V3 scores higher in 5 categories and MiniMax M1 in 3 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in long context, where MiniMax M1 leads 41.4 to 34.0.
  • The biggest single-benchmark swing is Fiction.LiveBench: 50% for DeepSeek-V3 and 69.4% for MiniMax M1.
  • DeepSeek-V3 is cheaper at $0.24 / $0.90 per million input/output tokens, against $0.55 / $2.20 for MiniMax M1.
  • MiniMax M1 accepts more context: 1M tokens versus 164K.

Side by side

DeepSeek-V3 and MiniMax M1 specifications
DeepSeek-V3MiniMax M1
ProviderDeepSeekMiniMax
Noometry Index39.540.3
Released2024-12-262025-06-13
WeightsOpenOpen
Context window164K1M
Max output164K40K
Input $ / M tokens$0.24$0.55
Output $ / M tokens$0.90$2.20
Results tracked6018

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek-V3 leads

DeepSeek-V3: 42.3 (#106), MiniMax M1: 39.9 (#153)

Coding benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Coding13681359
Aider Polyglot55.1%—
SciCode35.8%—
WeirdML36.1%—
BigCodeBench Instruct50%—
LiveBench Coding70.9%—
BigCodeBench Complete62.2%—
HumanEval+86.6%—
MBPP+73%—

Agentic & Tool Use Not comparable

DeepSeek-V3: —, MiniMax M1: —

Agentic & Tool Use benchmarks
BenchmarkDeepSeek-V3MiniMax M1
METR Time Horizons49.6%—

Reasoning MiniMax M1 leads

DeepSeek-V3: 20.5 (#236), MiniMax M1: 26.9 (#126)

Reasoning benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Hard Prompts13651339
SimpleBench27.2%—
Kagi LLM Benchmark52.3%—
CritPt0%—
LiveBench Reasoning65.8%—
DTBench64.8%—
LiveBench Data Analysis60.9%—
LMCA15.5%—
BIG-Bench Hard87.5%—
Epoch Capabilities Index135.94—
ForecastBench59.1—
HellaSwag88.9%—
LiveBench66.9%—
PIQA84.7%—
WinoGrande85.2%—

Math MiniMax M1 leads

DeepSeek-V3: 32.1 (#219), MiniMax M1: 37.5 (#151)

Math benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Math13731361
OTIS Mock AIME 2024-202537.8%—
Omni-MATH40.3%—
LiveBench Math73.5%—
MATH Level 575.5%—
FrontierMath (Feb 2025 set)1.7%—

Knowledge DeepSeek-V3 leads

DeepSeek-V3: 37.5 (#155), MiniMax M1: 36.4 (#170)

Knowledge benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Expert13511317
GPQA Diamond67.6%—
MMLU-Pro72.3%—
Confabulations26.1%—
Vectara Hallucination Rate6.1%—
GPQA (HELM)53.8%—
ARC (AI2) Challenge95.3%—
MMLU87.2%—
TriviaQA82.9%—

Multilingual DeepSeek-V3 leads

DeepSeek-V3: 48.5 (#143), MiniMax M1: 45.8 (#163)

Multilingual benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Non-English13581319
LMArena Chinese13911360
LMArena French13851370
LMArena German13741350
LMArena Japanese13331217
LMArena Korean13191266
LMArena Russian13731329
LMArena Spanish13581353

Instruction Following DeepSeek-V3 leads

DeepSeek-V3: 72.8 (#130), MiniMax M1: 69.3 (#174)

Instruction Following benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Instruction Following13451312
LiveBench Instruction Following81.5%—
IFEval83.2%—

Long Context MiniMax M1 leads

DeepSeek-V3: 34.0 (#253), MiniMax M1: 41.4 (#141)

Long Context benchmarks
BenchmarkDeepSeek-V3MiniMax M1
Fiction.LiveBench50%69.4%
LMArena Longer Query13521326

Writing & Preference DeepSeek-V3 leads

DeepSeek-V3: 57.4 (#130), MiniMax M1: 53.1 (#161)

Writing & Preference benchmarks
BenchmarkDeepSeek-V3MiniMax M1
LMArena Text13751343
LMArena Creative Writing13641298
LMArena Multi-Turn13891335
Short-Story Creative Writing77%—
EQ-Bench Creative Writing1472—
WildBench83%—
LiveBench Language49.1%—

Frequently asked questions

Is DeepSeek-V3 better than MiniMax M1?

DeepSeek-V3 and MiniMax M1 score almost the same on the Noometry Index (39.5 vs 40.3), so choose on price, context window or the category you care about most.

Which is cheaper, DeepSeek-V3 or MiniMax M1?

DeepSeek-V3 is cheaper. It lists at $0.24 per million input tokens and $0.90 per million output tokens; MiniMax M1 lists at $0.55 and $2.20.

Is DeepSeek-V3 or MiniMax M1 better for coding?

DeepSeek-V3 scores higher on coding benchmarks: 42.3 versus 39.9 in the Noometry coding category.

Which has the bigger context window?

MiniMax M1 does, with 1M tokens against 164K.

How many benchmarks do DeepSeek-V3 and MiniMax M1 share?

18 benchmarks have published results for both models. DeepSeek-V3 has 60 scored results on Noometry and MiniMax M1 has 18.

Related comparisons

Go deeper