Model comparison

Llama 13b vs MiniMax M1

MiniMax M1 is the stronger model overall, scoring 40.3 to 24.4 on the Noometry Index.

Last verified . 8 shared benchmarks.

Llama 13b Meta

24.4

Rank #348 Confirmed

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Llama 13b scores higher in 0 categories and MiniMax M1 in 6 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where MiniMax M1 leads 53.1 to 13.8.

Side by side

Llama 13b and MiniMax M1 specifications
Llama 13bMiniMax M1
ProviderMetaMiniMax
Noometry Index24.440.3
Released2023-02-242025-06-13
WeightsOpenOpen
Context window—1M
Max output—40K
Input $ / M tokens—$0.55
Output $ / M tokens—$2.20
Results tracked2118

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiniMax M1 leads

Llama 13b: 21.4 (#337), MiniMax M1: 39.9 (#153)

Coding benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Coding6831359

Reasoning MiniMax M1 leads

Llama 13b: 14.0 (#329), MiniMax M1: 26.9 (#126)

Reasoning benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Hard Prompts7281339
BIG-Bench Hard37.9%—
Epoch Capabilities Index100.58—
HellaSwag79.2%—
LAMBADA75.2%—
PIQA80.1%—
WinoGrande73%—

Math MiniMax M1 leads

Llama 13b: 26.7 (#256), MiniMax M1: 37.5 (#151)

Math benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Math8381361
GSM8K20.6%—

Knowledge Not comparable

Llama 13b: —, MiniMax M1: 36.4 (#170)

Knowledge benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Expert—1317
ARC (AI2) Challenge52.7%—
BoolQ78.7%—
MMLU47.7%—
OpenBookQA56.4%—
TriviaQA77.9%—

Multimodal Not comparable

Llama 13b: —, MiniMax M1: —

Multimodal benchmarks
BenchmarkLlama 13bMiniMax M1
ScienceQA43.3%—

Multilingual MiniMax M1 leads

Llama 13b: 16.6 (#297), MiniMax M1: 45.8 (#163)

Multilingual benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Non-English8191319
LMArena Chinese—1360
LMArena French—1370
LMArena German—1350
LMArena Japanese—1217
LMArena Korean—1266
LMArena Russian—1329
LMArena Spanish—1353

Instruction Following MiniMax M1 leads

Llama 13b: 36.7 (#305), MiniMax M1: 69.3 (#174)

Instruction Following benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Instruction Following7811312

Long Context Not comparable

Llama 13b: —, MiniMax M1: 41.4 (#141)

Long Context benchmarks
BenchmarkLlama 13bMiniMax M1
Fiction.LiveBench—69.4%
LMArena Longer Query—1326

Writing & Preference MiniMax M1 leads

Llama 13b: 13.8 (#312), MiniMax M1: 53.1 (#161)

Writing & Preference benchmarks
BenchmarkLlama 13bMiniMax M1
LMArena Text8341343
LMArena Creative Writing7941298
LMArena Multi-Turn7531335

Frequently asked questions

Is Llama 13b better than MiniMax M1?

MiniMax M1 is the stronger model overall, scoring 40.3 to 24.4 on the Noometry Index.

Is Llama 13b or MiniMax M1 better for coding?

MiniMax M1 scores higher on coding benchmarks: 39.9 versus 21.4 in the Noometry coding category.

How many benchmarks do Llama 13b and MiniMax M1 share?

8 benchmarks have published results for both models. Llama 13b has 21 scored results on Noometry and MiniMax M1 has 18.

Related comparisons

Go deeper