Model comparison

DBRX vs Mistral Small 3.2

Mistral Small 3.2 is the stronger model overall, scoring 31.2 to 29.4 on the Noometry Index.

Last verified . 1 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • They share 1 benchmark with published results for both. DBRX scores higher in 1 category and Mistral Small 3.2 in 3 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Mistral Small 3.2 leads 26.7 to 14.9.
  • The biggest single-benchmark swing is GPQA Diamond: 32.9% for DBRX and 49.1% for Mistral Small 3.2.

Side by side

DBRX and Mistral Small 3.2 specifications
DBRXMistral Small 3.2
ProviderDatabricksMistral AI
Noometry Index29.431.2
Released2024-03-272025-06-20
WeightsOpenOpen
Context window—256K
Max output—16K
Input $ / M tokens—$0.0938
Output $ / M tokens—$0.25
Results tracked216

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

DBRX: 32.9 (#266), Mistral Small 3.2: —

Coding benchmarks
BenchmarkDBRXMistral Small 3.2
LMArena Coding1132—
HumanEval+70.1%—
MBPP+55.8%—

Reasoning DBRX leads

DBRX: 21.4 (#222), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkDBRXMistral Small 3.2
Kagi LLM Benchmark—40.4%
Chess Puzzles—1%
LMArena Hard Prompts1113—
Epoch Capabilities Index—131.74

Math Mistral Small 3.2 leads

DBRX: 24.3 (#269), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkDBRXMistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
LMArena Math1145—
MATH Level 511.7%—

Knowledge Mistral Small 3.2 leads

DBRX: 14.9 (#294), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkDBRXMistral Small 3.2
GPQA Diamond32.9%49.1%
LMArena Expert1076—

Multilingual Not comparable

DBRX: 29.3 (#268), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkDBRXMistral Small 3.2
LMArena Non-English1071—
LMArena Chinese1068—
LMArena French1096—
LMArena German1057—
LMArena Japanese990—
LMArena Korean993—
LMArena Russian1078—
LMArena Spanish1064—

Instruction Following Not comparable

DBRX: 57.5 (#270), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkDBRXMistral Small 3.2
LMArena Instruction Following1112—

Long Context Not comparable

DBRX: 33.7 (#258), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkDBRXMistral Small 3.2
LMArena Longer Query1112—

Writing & Preference Mistral Small 3.2 leads

DBRX: 33.6 (#275), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkDBRXMistral Small 3.2
LMArena Text1119—
LMArena Creative Writing1104—
EQ-Bench Creative Writing—1255
LMArena Multi-Turn1111—

Frequently asked questions

Is DBRX better than Mistral Small 3.2?

Mistral Small 3.2 is the stronger model overall, scoring 31.2 to 29.4 on the Noometry Index.

How many benchmarks do DBRX and Mistral Small 3.2 share?

1 benchmark has published results for both models. DBRX has 21 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper