DeepSeek, open weights
DeepSeek LLM 67B
DeepSeek LLM 67B by DeepSeek ranks 347th of 354 ranked models on the Noometry Index as of October 2026, with a score of 24.9. Its strongest category is long context, where it ranks 265th.
Last verified
Specifications
- Noometry rank
- #347 of 354
- Index score
- 24.9
- Evidence
- Confirmed 15 results
- Provider
DeepSeek
- Released
- November 29, 2023
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 31.9
- Reasoning 16.5
- Math 8.7
- Knowledge 7.0
- Multilingual 29.4
- Instruction Following 55.4
- Long Context 33.1
- Writing & Preference 31.6
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 31.9 | #278 | 1 |
| Reasoning | 16.5 | #304 | 2 |
| Math | 8.7 | #324 | 3 |
| Knowledge | 7.0 | #313 | 1 |
| Multilingual | 29.4 | #267 | 1 |
| Instruction Following | 55.4 | #277 | 1 |
| Long Context | 33.1 | #265 | 1 |
| Writing & Preference | 31.6 | #282 | 3 |
Strengths and weaknesses
Categories where DeepSeek LLM 67B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 31.9 | −6.9 | #278 of 340, top 82% |
| Reasoning | 16.5 | −7.1 | #304 of 350, top 87% |
| Long Context | 33.1 | −7.8 | #265 of 296, top 90% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 7.0 | −30.3 | #313 of 314, top 100% |
| Math | 8.7 | −27.9 | #324 of 327, top 100% |
| Instruction Following | 55.4 | −15.9 | #277 of 305, top 91% |
Closest competitors
The models ranked just above and below DeepSeek LLM 67B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| GPT-4o mini | #343 | 25.5 | $0.26 | 120 | Compare |
| Llama 3-8B | #344 | 25.5 | — | — | Compare |
| Claude 2.1 | #345 | 25.2 | — | — | Compare |
| Claude 2 | #346 | 25.0 | — | — | Compare |
| Llama 13b | #348 | 24.4 | — | — | Compare |
| Llama 2-70B | #349 | 24.4 | — | — | Compare |
| GPT-3.5-turbo | #350 | 23.2 | $0.75 | — | Compare |
| Mistral 7B | #351 | 23.0 | $0.25 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1096 | #271 of 294, top 93% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Chess Puzzles | 0% | #108 of 129, top 84% | Epoch AI | 2026-08-28 | |
| LMArena Hard Prompts | 1070 | #277 of 297, top 94% | LMArena | 2026-10-08 | |
| Epoch Capabilities Index | 110.5 | #191 of 213, top 90% | Epoch AI | 2023-11-29 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 0.8% | #169 of 173, top 98% | Epoch AI | 2026-08-28 | |
| LMArena Math | 1108 | #265 of 285, top 93% | LMArena | 2026-10-08 | |
| MATH Level 5 | 6.4% | #75 of 79, top 95% | Epoch AI | 2025-01-27 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 24.6% | #181 of 186, top 98% | Epoch AI | 2025-01-27 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1073 | #267 of 297, top 90% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1132 | #252 of 285, top 89% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1079 | #273 of 298, top 92% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1092 | #270 of 291, top 93% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1105 | #272 of 297, top 92% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1067 | #276 of 295, top 94% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1082 | #270 of 295, top 92% | LMArena | 2026-10-08 |
Compare DeepSeek LLM 67B
- DeepSeek LLM 67B vs Claude 2
- DeepSeek LLM 67B vs Llama 13b
- DeepSeek LLM 67B vs Claude 2.1
- DeepSeek LLM 67B vs Llama 2-70B
- DeepSeek LLM 67B vs Llama 3-8B
- DeepSeek LLM 67B vs GPT-3.5-turbo
- DeepSeek LLM 67B vs GPT-6 Astra
- DeepSeek LLM 67B vs Claude Fable 5.1
- DeepSeek LLM 67B vs Gemini 3.8 Flash
- DeepSeek LLM 67B vs Kimi K3
- DeepSeek LLM 67B vs Grok 4.6
- DeepSeek LLM 67B vs Qwen3.8 Max
- DeepSeek LLM 67B vs GLM-5.3
- DeepSeek LLM 67B vs Muse Spark 1.3
Other DeepSeek models
Frequently asked questions
How good is DeepSeek LLM 67B?
DeepSeek LLM 67B by DeepSeek ranks 347th of 354 ranked models on the Noometry Index as of October 2026, with a score of 24.9. Its strongest category is long context, where it ranks 265th.
Is DeepSeek LLM 67B open source?
Yes. DeepSeek LLM 67B's weights are downloadable; check the license for commercial terms.
What are DeepSeek LLM 67B's strengths and weaknesses?
Relative to other ranked models, DeepSeek LLM 67B places best in coding, reasoning, long context and lowest in knowledge, math, instruction following.
What is DeepSeek LLM 67B best at?
Its best category is long context, where it ranks 265th on Noometry.