Model comparison

DeepSeek Coder 6.7B vs DeepSeek-R1-Distill-Qwen-14B

DeepSeek-R1-Distill-Qwen-14B has enough public results to be ranked (#252); DeepSeek Coder 6.7B does not yet, so treat this comparison as directional.

Last verified . 3 shared benchmarks.

DeepSeek Coder 6.7B DeepSeek

37.6

Unranked Sparse

DeepSeek-R1-Distill-Qwen-14B DeepSeek

32.7

Rank #252 Confirmed

Summary

  • They share 3 benchmarks with published results for both. DeepSeek Coder 6.7B scores higher in 0 categories and DeepSeek-R1-Distill-Qwen-14B in 1 category; one gap is clear of the uncertainty.

Side by side

DeepSeek Coder 6.7B and DeepSeek-R1-Distill-Qwen-14B specifications
DeepSeek Coder 6.7BDeepSeek-R1-Distill-Qwen-14B
ProviderDeepSeekDeepSeek
Noometry Index37.632.7
Released2023-11-022025-01-20
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked97

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek-R1-Distill-Qwen-14B leads

DeepSeek Coder 6.7B: 35.8 (#219), DeepSeek-R1-Distill-Qwen-14B: 36.9 (#200)

Coding benchmarks
BenchmarkDeepSeek Coder 6.7BDeepSeek-R1-Distill-Qwen-14B
BigCodeBench Instruct35.5%38.1%
BigCodeBench Complete43.8%48.4%
HumanEval+71.3%—
MBPP+65.6%—

Reasoning Not comparable

DeepSeek Coder 6.7B: —, DeepSeek-R1-Distill-Qwen-14B: 19.2 (#263)

Reasoning benchmarks
BenchmarkDeepSeek Coder 6.7BDeepSeek-R1-Distill-Qwen-14B
Epoch Capabilities Index89.62135.43
Chess Puzzles—1%
WinoGrande57.6%—

Math Not comparable

DeepSeek Coder 6.7B: —, DeepSeek-R1-Distill-Qwen-14B: 35.5 (#184)

Math benchmarks
BenchmarkDeepSeek Coder 6.7BDeepSeek-R1-Distill-Qwen-14B
OTIS Mock AIME 2024-2025—50.6%
MATH Level 5—87.1%
GSM8K21.3%—

Knowledge Not comparable

DeepSeek Coder 6.7B: —, DeepSeek-R1-Distill-Qwen-14B: 24.1 (#270)

Knowledge benchmarks
BenchmarkDeepSeek Coder 6.7BDeepSeek-R1-Distill-Qwen-14B
GPQA Diamond—44.7%
ARC (AI2) Challenge36.4%—
MMLU36.4%—

Frequently asked questions

Is DeepSeek Coder 6.7B better than DeepSeek-R1-Distill-Qwen-14B?

DeepSeek-R1-Distill-Qwen-14B has enough public results to be ranked (#252); DeepSeek Coder 6.7B does not yet, so treat this comparison as directional.

Is DeepSeek Coder 6.7B or DeepSeek-R1-Distill-Qwen-14B better for coding?

DeepSeek-R1-Distill-Qwen-14B scores higher on coding benchmarks: 36.9 versus 35.8 in the Noometry coding category.

How many benchmarks do DeepSeek Coder 6.7B and DeepSeek-R1-Distill-Qwen-14B share?

3 benchmarks have published results for both models. DeepSeek Coder 6.7B has 9 scored results on Noometry and DeepSeek-R1-Distill-Qwen-14B has 7.

Related comparisons

Go deeper