Model comparison

DBRX vs DeepSeek-V2 (MoE-236B, May 2024)

DBRX has enough public results to be ranked (#311); DeepSeek-V2 (MoE-236B, May 2024) does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

Summary

  • The widest gap is in coding, where DeepSeek-V2 (MoE-236B, May 2024) leads 40.4 to 32.9.

Side by side

DBRX and DeepSeek-V2 (MoE-236B, May 2024) specifications
DBRXDeepSeek-V2 (MoE-236B, May 2024)
ProviderDatabricksDeepSeek
Noometry Index29.440.3
Released2024-03-272024-05-07
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2110

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek-V2 (MoE-236B, May 2024) leads

DBRX: 32.9 (#266), DeepSeek-V2 (MoE-236B, May 2024): 40.4 (#139)

Coding benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
BigCodeBench Instruct—48.9%
LMArena Coding1132—
BigCodeBench Complete—59.4%
HumanEval+70.1%—
MBPP+55.8%—

Reasoning Not comparable

DBRX: 21.4 (#222), DeepSeek-V2 (MoE-236B, May 2024): —

Reasoning benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
LMArena Hard Prompts1113—
BIG-Bench Hard—78.8%
Epoch Capabilities Index—124.77
HellaSwag—87.1%
PIQA—83.9%
WinoGrande—86.3%

Math Not comparable

DBRX: 24.3 (#269), DeepSeek-V2 (MoE-236B, May 2024): —

Math benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
LMArena Math1145—
MATH Level 511.7%—

Knowledge Not comparable

DBRX: 14.9 (#294), DeepSeek-V2 (MoE-236B, May 2024): —

Knowledge benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
GPQA Diamond32.9%—
LMArena Expert1076—
ARC (AI2) Challenge—92.2%
MMLU—78.4%
TriviaQA—80%

Multilingual Not comparable

DBRX: 29.3 (#268), DeepSeek-V2 (MoE-236B, May 2024): —

Multilingual benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
LMArena Non-English1071—
LMArena Chinese1068—
LMArena French1096—
LMArena German1057—
LMArena Japanese990—
LMArena Korean993—
LMArena Russian1078—
LMArena Spanish1064—

Instruction Following Not comparable

DBRX: 57.5 (#270), DeepSeek-V2 (MoE-236B, May 2024): —

Instruction Following benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
LMArena Instruction Following1112—

Long Context Not comparable

DBRX: 33.7 (#258), DeepSeek-V2 (MoE-236B, May 2024): —

Long Context benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
LMArena Longer Query1112—

Writing & Preference Not comparable

DBRX: 33.6 (#275), DeepSeek-V2 (MoE-236B, May 2024): —

Writing & Preference benchmarks
BenchmarkDBRXDeepSeek-V2 (MoE-236B, May 2024)
LMArena Text1119—
LMArena Creative Writing1104—
LMArena Multi-Turn1111—

Frequently asked questions

Is DBRX better than DeepSeek-V2 (MoE-236B, May 2024)?

DBRX has enough public results to be ranked (#311); DeepSeek-V2 (MoE-236B, May 2024) does not yet, so treat this comparison as directional.

Is DBRX or DeepSeek-V2 (MoE-236B, May 2024) better for coding?

DeepSeek-V2 (MoE-236B, May 2024) scores higher on coding benchmarks: 40.4 versus 32.9 in the Noometry coding category.

How many benchmarks do DBRX and DeepSeek-V2 (MoE-236B, May 2024) share?

0 benchmarks have published results for both models. DBRX has 21 scored results on Noometry and DeepSeek-V2 (MoE-236B, May 2024) has 10.

Related comparisons

Go deeper