Model comparison

DeepSeek Coder 1.3B vs Mistral Small 3.1

Mistral Small 3.1 has enough public results to be ranked (#269); DeepSeek Coder 1.3B does not yet, so treat this comparison as directional.

Last verified . 1 shared benchmarks.

DeepSeek Coder 1.3B DeepSeek

35.0

Unranked Sparse

Mistral Small 3.1 Mistral AI

31.7

Rank #269 Confirmed

Summary

  • They share 1 benchmark with published results for both. DeepSeek Coder 1.3B scores higher in 0 categories and Mistral Small 3.1 in 1 category; one gap is clear of the uncertainty.
  • The widest gap is in coding, where Mistral Small 3.1 leads 38.3 to 31.2.

Side by side

DeepSeek Coder 1.3B and Mistral Small 3.1 specifications
DeepSeek Coder 1.3BMistral Small 3.1
ProviderDeepSeekMistral AI
Noometry Index35.031.7
Released2023-11-022025-03-17
WeightsOpenOpen
Context window—128K
Max output—102K
Input $ / M tokens—$0.35
Output $ / M tokens—$0.56
Results tracked928

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Small 3.1 leads

DeepSeek Coder 1.3B: 31.2 (#287), Mistral Small 3.1: 38.3 (#179)

Coding benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
BigCodeBench Instruct22.8%—
LMArena Coding—1309
BigCodeBench Complete29.6%—
HumanEval+60.4%—
MBPP+54.8%—

Reasoning Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 19.7 (#254)

Reasoning benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
Epoch Capabilities Index63.6127.48
Chess Puzzles—1%
LMArena Hard Prompts—1278
WinoGrande53.3%—

Math Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 14.7 (#301)

Math benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
OTIS Mock AIME 2024-2025—3.9%
Omni-MATH—24.8%
LMArena Math—1262
GSM8K4.4%—

Knowledge Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 22.6 (#271)

Knowledge benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
GPQA Diamond—41.9%
MMLU-Pro—61%
GPQA (HELM)—39.2%
LMArena Expert—1257
ARC (AI2) Challenge25.4%—
MMLU25.8%—

Multimodal Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 33.2 (#99)

Multimodal benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
LMArena Vision—1136

Multilingual Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 41.2 (#209)

Multilingual benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
LMArena Non-English—1255
LMArena Chinese—1253
LMArena French—1273
LMArena German—1266
LMArena Japanese—1208
LMArena Korean—1206
LMArena Russian—1263
LMArena Spanish—1283

Instruction Following Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 63.6 (#230)

Instruction Following benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
IFEval—75%
LMArena Instruction Following—1264

Long Context Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 39.5 (#178)

Long Context benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
LMArena Longer Query—1299

Writing & Preference Not comparable

DeepSeek Coder 1.3B: —, Mistral Small 3.1: 37.0 (#259)

Writing & Preference benchmarks
BenchmarkDeepSeek Coder 1.3BMistral Small 3.1
LMArena Text—1277
LMArena Creative Writing—1253
EQ-Bench Creative Writing—761
WildBench—78.8%
LMArena Multi-Turn—1270

Frequently asked questions

Is DeepSeek Coder 1.3B better than Mistral Small 3.1?

Mistral Small 3.1 has enough public results to be ranked (#269); DeepSeek Coder 1.3B does not yet, so treat this comparison as directional.

Is DeepSeek Coder 1.3B or Mistral Small 3.1 better for coding?

Mistral Small 3.1 scores higher on coding benchmarks: 38.3 versus 31.2 in the Noometry coding category.

How many benchmarks do DeepSeek Coder 1.3B and Mistral Small 3.1 share?

1 benchmark has published results for both models. DeepSeek Coder 1.3B has 9 scored results on Noometry and Mistral Small 3.1 has 28.

Related comparisons

Go deeper