Model comparison

DBRX vs DeepSeek LLM 67B

DBRX is the stronger model overall, scoring 29.4 to 24.9 on the Noometry Index.

Last verified . 12 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

DeepSeek LLM 67B DeepSeek

24.9

Rank #347 Confirmed

Summary

  • They share 12 benchmarks with published results for both. DBRX scores higher in 7 categories and DeepSeek LLM 67B in 1 category; 6 gaps are clear of the uncertainty.
  • The widest gap is in math, where DBRX leads 24.3 to 8.7.
  • The biggest single-benchmark swing is GPQA Diamond: 32.9% for DBRX and 24.6% for DeepSeek LLM 67B.

Side by side

DBRX and DeepSeek LLM 67B specifications
DBRXDeepSeek LLM 67B
ProviderDatabricksDeepSeek
Noometry Index29.424.9
Released2024-03-272023-11-29
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2115

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DBRX leads

DBRX: 32.9 (#266), DeepSeek LLM 67B: 31.9 (#278)

Coding benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Coding11321096
HumanEval+70.1%—
MBPP+55.8%—

Reasoning DBRX leads

DBRX: 21.4 (#222), DeepSeek LLM 67B: 16.5 (#304)

Reasoning benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Hard Prompts11131070
Chess Puzzles—0%
Epoch Capabilities Index—110.5

Math DBRX leads

DBRX: 24.3 (#269), DeepSeek LLM 67B: 8.7 (#324)

Math benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Math11451108
MATH Level 511.7%6.4%
OTIS Mock AIME 2024-2025—0.8%

Knowledge DBRX leads

DBRX: 14.9 (#294), DeepSeek LLM 67B: 7.0 (#313)

Knowledge benchmarks
BenchmarkDBRXDeepSeek LLM 67B
GPQA Diamond32.9%24.6%
LMArena Expert1076—

Multilingual Too close to call

DBRX: 29.3 (#268), DeepSeek LLM 67B: 29.4 (#267)

Multilingual benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Non-English10711073
LMArena Chinese10681132
LMArena French1096—
LMArena German1057—
LMArena Japanese990—
LMArena Korean993—
LMArena Russian1078—
LMArena Spanish1064—

Instruction Following DBRX leads

DBRX: 57.5 (#270), DeepSeek LLM 67B: 55.4 (#277)

Instruction Following benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Instruction Following11121079

Long Context Too close to call

DBRX: 33.7 (#258), DeepSeek LLM 67B: 33.1 (#265)

Long Context benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Longer Query11121092

Writing & Preference DBRX leads

DBRX: 33.6 (#275), DeepSeek LLM 67B: 31.6 (#282)

Writing & Preference benchmarks
BenchmarkDBRXDeepSeek LLM 67B
LMArena Text11191105
LMArena Creative Writing11041067
LMArena Multi-Turn11111082

Frequently asked questions

Is DBRX better than DeepSeek LLM 67B?

DBRX is the stronger model overall, scoring 29.4 to 24.9 on the Noometry Index.

Is DBRX or DeepSeek LLM 67B better for coding?

DBRX scores higher on coding benchmarks: 32.9 versus 31.9 in the Noometry coding category.

How many benchmarks do DBRX and DeepSeek LLM 67B share?

12 benchmarks have published results for both models. DBRX has 21 scored results on Noometry and DeepSeek LLM 67B has 15.

Related comparisons

Go deeper