Model comparison

MiniMax M1 vs Qwen3 14B

MiniMax M1 is the stronger model overall, scoring 40.3 to 35.5 on the Noometry Index. Qwen3 14B costs 1.6× less per token, which makes it the better buy when MiniMax M1's lead doesn't matter for your workload.

Last verified . 1 shared benchmarks.

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Qwen3 14B Alibaba (Qwen)

35.5

Rank #225 Confirmed

Summary

  • They share 1 benchmark with published results for both. MiniMax M1 scores higher in 3 categories and Qwen3 14B in 2 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where MiniMax M1 leads 26.9 to 18.5.
  • The biggest single-benchmark swing is Fiction.LiveBench: 69.4% for MiniMax M1 and 62.5% for Qwen3 14B.
  • Qwen3 14B is cheaper at $0.35 / $1.40 per million input/output tokens, against $0.55 / $2.20 for MiniMax M1.
  • MiniMax M1 accepts more context: 1M tokens versus 131K.

Side by side

MiniMax M1 and Qwen3 14B specifications
MiniMax M1Qwen3 14B
ProviderMiniMaxAlibaba (Qwen)
Noometry Index40.335.5
Released2025-06-132025-04
WeightsOpenOpen
Context window1M131K
Max output40K8K
Input $ / M tokens$0.55$0.35
Output $ / M tokens$2.20$1.40
Results tracked1812

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiniMax M1 leads

MiniMax M1: 39.9 (#153), Qwen3 14B: 37.3 (#195)

Coding benchmarks
BenchmarkMiniMax M1Qwen3 14B
SciCode—31.6%
LMArena Coding1359—

Agentic & Tool Use Not comparable

MiniMax M1: —, Qwen3 14B: 29.6 (#83)

Agentic & Tool Use benchmarks
BenchmarkMiniMax M1Qwen3 14B
Berkeley Function Calling Leaderboard—41%

Reasoning MiniMax M1 leads

MiniMax M1: 26.9 (#126), Qwen3 14B: 18.5 (#280)

Reasoning benchmarks
BenchmarkMiniMax M1Qwen3 14B
Kagi LLM Benchmark—49.1%
CritPt—0%
Chess Puzzles—4%
LMArena Hard Prompts1339—
DTBench—64%
LMCA—18.2%
Epoch Capabilities Index—138.23

Math Qwen3 14B leads

MiniMax M1: 37.5 (#151), Qwen3 14B: 38.6 (#133)

Math benchmarks
BenchmarkMiniMax M1Qwen3 14B
OTIS Mock AIME 2024-2025—66.4%
LMArena Math1361—

Knowledge Qwen3 14B leads

MiniMax M1: 36.4 (#170), Qwen3 14B: 39.3 (#134)

Knowledge benchmarks
BenchmarkMiniMax M1Qwen3 14B
GPQA Diamond—63.8%
Vectara Hallucination Rate—5.4%
LMArena Expert1317—

Multilingual Not comparable

MiniMax M1: 45.8 (#163), Qwen3 14B: —

Multilingual benchmarks
BenchmarkMiniMax M1Qwen3 14B
LMArena Non-English1319—
LMArena Chinese1360—
LMArena French1370—
LMArena German1350—
LMArena Japanese1217—
LMArena Korean1266—
LMArena Russian1329—
LMArena Spanish1353—

Instruction Following Not comparable

MiniMax M1: 69.3 (#174), Qwen3 14B: —

Instruction Following benchmarks
BenchmarkMiniMax M1Qwen3 14B
LMArena Instruction Following1312—

Long Context MiniMax M1 leads

MiniMax M1: 41.4 (#141), Qwen3 14B: 38.1 (#204)

Long Context benchmarks
BenchmarkMiniMax M1Qwen3 14B
Fiction.LiveBench69.4%62.5%
LMArena Longer Query1326—

Writing & Preference Not comparable

MiniMax M1: 53.1 (#161), Qwen3 14B: —

Writing & Preference benchmarks
BenchmarkMiniMax M1Qwen3 14B
LMArena Text1343—
LMArena Creative Writing1298—
LMArena Multi-Turn1335—

Frequently asked questions

Is MiniMax M1 better than Qwen3 14B?

MiniMax M1 is the stronger model overall, scoring 40.3 to 35.5 on the Noometry Index. Qwen3 14B costs 1.6× less per token, which makes it the better buy when MiniMax M1's lead doesn't matter for your workload.

Which is cheaper, MiniMax M1 or Qwen3 14B?

Qwen3 14B is cheaper. It lists at $0.35 per million input tokens and $1.40 per million output tokens; MiniMax M1 lists at $0.55 and $2.20.

Is MiniMax M1 or Qwen3 14B better for coding?

MiniMax M1 scores higher on coding benchmarks: 39.9 versus 37.3 in the Noometry coding category.

Which has the bigger context window?

MiniMax M1 does, with 1M tokens against 131K.

How many benchmarks do MiniMax M1 and Qwen3 14B share?

1 benchmark has published results for both models. MiniMax M1 has 18 scored results on Noometry and Qwen3 14B has 12.

Related comparisons

Go deeper