Model comparison

DBRX vs Llama 3.2 3B

DBRX and Llama 3.2 3B score almost the same on the Noometry Index (29.4 vs 28.9), so choose on price, context window or the category you care about most.

Last verified . 13 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

Llama 3.2 3B Meta

28.9

Rank #321 Confirmed

Summary

  • They share 13 benchmarks with published results for both. DBRX scores higher in 6 categories and Llama 3.2 3B in 2 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Llama 3.2 3B leads 29.7 to 14.9.

Side by side

DBRX and Llama 3.2 3B specifications
DBRXLlama 3.2 3B
ProviderDatabricksMeta
Noometry Index29.428.9
Released2024-03-272024-09-24
WeightsOpenOpen
Context window—131K
Max output—118K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.33
Results tracked2118

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DBRX leads

DBRX: 32.9 (#266), Llama 3.2 3B: 27.6 (#319)

Coding benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Coding11321098
BigCodeBench Instruct—23.4%
BigCodeBench Complete—28.3%
HumanEval+70.1%—
MBPP+55.8%—

Agentic & Tool Use Not comparable

DBRX: —, Llama 3.2 3B: 20.1 (#143)

Agentic & Tool Use benchmarks
BenchmarkDBRXLlama 3.2 3B
Berkeley Function Calling Leaderboard—21.9%
BALROG—10.1%

Reasoning Too close to call

DBRX: 21.4 (#222), Llama 3.2 3B: 21.0 (#228)

Reasoning benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Hard Prompts11131095

Math Llama 3.2 3B leads

DBRX: 24.3 (#269), Llama 3.2 3B: 32.4 (#214)

Math benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Math11451126
MATH Level 511.7%—

Knowledge Llama 3.2 3B leads

DBRX: 14.9 (#294), Llama 3.2 3B: 29.7 (#235)

Knowledge benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Expert10761090
GPQA Diamond32.9%—

Multilingual DBRX leads

DBRX: 29.3 (#268), Llama 3.2 3B: 26.2 (#281)

Multilingual benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Non-English10711019
LMArena Chinese10681017
LMArena German10571056
LMArena Russian1078949
LMArena French1096—
LMArena Japanese990—
LMArena Korean993—
LMArena Spanish1064—

Instruction Following DBRX leads

DBRX: 57.5 (#270), Llama 3.2 3B: 56.0 (#275)

Instruction Following benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Instruction Following11121089

Long Context Too close to call

DBRX: 33.7 (#258), Llama 3.2 3B: 33.4 (#261)

Long Context benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Longer Query11121100

Writing & Preference DBRX leads

DBRX: 33.6 (#275), Llama 3.2 3B: 24.7 (#307)

Writing & Preference benchmarks
BenchmarkDBRXLlama 3.2 3B
LMArena Text11191110
LMArena Creative Writing11041094
LMArena Multi-Turn11111105
EQ-Bench Creative Writing—595

Frequently asked questions

Is DBRX better than Llama 3.2 3B?

DBRX and Llama 3.2 3B score almost the same on the Noometry Index (29.4 vs 28.9), so choose on price, context window or the category you care about most.

Is DBRX or Llama 3.2 3B better for coding?

DBRX scores higher on coding benchmarks: 32.9 versus 27.6 in the Noometry coding category.

How many benchmarks do DBRX and Llama 3.2 3B share?

13 benchmarks have published results for both models. DBRX has 21 scored results on Noometry and Llama 3.2 3B has 18.

Related comparisons

Go deeper