DeepSeek, open weights

DeepSeek LLM 67B

DeepSeek LLM 67B by DeepSeek ranks 347th of 354 ranked models on the Noometry Index as of October 2026, with a score of 24.9. Its strongest category is long context, where it ranks 265th.

Last verified

Specifications

Noometry rank
#347 of 354
Index score
24.9
Evidence
Confirmed 15 results
Provider
DeepSeek
Released
November 29, 2023
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

DeepSeek LLM 67B category scores
  1. Coding 31.9
  2. Reasoning 16.5
  3. Math 8.7
  4. Knowledge 7.0
  5. Multilingual 29.4
  6. Instruction Following 55.4
  7. Long Context 33.1
  8. Writing & Preference 31.6
DeepSeek LLM 67B category ranks
CategoryScoreRankResults
Coding31.9#2781
Reasoning16.5#3042
Math8.7#3243
Knowledge7.0#3131
Multilingual29.4#2671
Instruction Following55.4#2771
Long Context33.1#2651
Writing & Preference31.6#2823

Strengths and weaknesses

Categories where DeepSeek LLM 67B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

DeepSeek LLM 67B: strongest categories
CategoryScorevs medianRank
Coding31.9−6.9#278 of 340, top 82%
Reasoning16.5−7.1#304 of 350, top 87%
Long Context33.1−7.8#265 of 296, top 90%

Weakest categories

DeepSeek LLM 67B: weakest categories
CategoryScorevs medianRank
Knowledge7.0−30.3#313 of 314, top 100%
Math8.7−27.9#324 of 327, top 100%
Instruction Following55.4−15.9#277 of 305, top 91%

Closest competitors

The models ranked just above and below DeepSeek LLM 67B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to DeepSeek LLM 67B
ModelRankScoreBlended $/MSpeed
GPT-4o mini#34325.5$0.26120Compare
Llama 3-8B#34425.5——Compare
Claude 2.1#34525.2——Compare
Claude 2#34625.0——Compare
Llama 13b#34824.4——Compare
Llama 2-70B#34924.4——Compare
GPT-3.5-turbo#35023.2$0.75—Compare
Mistral 7B#35123.0$0.25—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

DeepSeek LLM 67B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1096#271 of 294, top 93%LMArena2026-10-08

Reasoning

DeepSeek LLM 67B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Chess Puzzles0%#108 of 129, top 84%Epoch AI2026-08-28
LMArena Hard Prompts1070#277 of 297, top 94%LMArena2026-10-08
Epoch Capabilities Index110.5#191 of 213, top 90%Epoch AI2023-11-29

Math

DeepSeek LLM 67B Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20250.8%#169 of 173, top 98%Epoch AI2026-08-28
LMArena Math1108#265 of 285, top 93%LMArena2026-10-08
MATH Level 56.4%#75 of 79, top 95%Epoch AI2025-01-27

Knowledge

DeepSeek LLM 67B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond24.6%#181 of 186, top 98%Epoch AI2025-01-27

Multilingual

DeepSeek LLM 67B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1073#267 of 297, top 90%LMArena2026-10-08
LMArena Chinese1132#252 of 285, top 89%LMArena2026-10-08

Instruction Following

DeepSeek LLM 67B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1079#273 of 298, top 92%LMArena2026-10-08

Long Context

DeepSeek LLM 67B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1092#270 of 291, top 93%LMArena2026-10-08

Writing & Preference

DeepSeek LLM 67B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1105#272 of 297, top 92%LMArena2026-10-08
LMArena Creative Writing1067#276 of 295, top 94%LMArena2026-10-08
LMArena Multi-Turn1082#270 of 295, top 92%LMArena2026-10-08

Compare DeepSeek LLM 67B

Other DeepSeek models

Frequently asked questions

How good is DeepSeek LLM 67B?

DeepSeek LLM 67B by DeepSeek ranks 347th of 354 ranked models on the Noometry Index as of October 2026, with a score of 24.9. Its strongest category is long context, where it ranks 265th.

Is DeepSeek LLM 67B open source?

Yes. DeepSeek LLM 67B's weights are downloadable; check the license for commercial terms.

What are DeepSeek LLM 67B's strengths and weaknesses?

Relative to other ranked models, DeepSeek LLM 67B places best in coding, reasoning, long context and lowest in knowledge, math, instruction following.

What is DeepSeek LLM 67B best at?

Its best category is long context, where it ranks 265th on Noometry.