Model comparison

DBRX vs Gemini 2.5 Pro

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 29.4 on the Noometry Index.

Last verified . 19 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

Gemini 2.5 Pro Google

45.0

Rank #75 Confirmed

Summary

  • They share 19 benchmarks with published results for both. DBRX scores higher in 0 categories and Gemini 2.5 Pro in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemini 2.5 Pro leads 56.0 to 14.9.
  • The biggest single-benchmark swing is MATH Level 5: 11.7% for DBRX and 95.9% for Gemini 2.5 Pro.
  • DBRX has downloadable open weights; the other is API-only.

Side by side

DBRX and Gemini 2.5 Pro specifications
DBRXGemini 2.5 Pro
ProviderDatabricksGoogle
Noometry Index29.445.0
Released2024-03-272025-03-25
WeightsOpenProprietary
Context window—1.05M
Max output—66K
Input $ / M tokens—$1.25
Output $ / M tokens—$10
Results tracked2178

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 2.5 Pro leads

DBRX: 32.9 (#266), Gemini 2.5 Pro: 42.4 (#101)

Coding benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Coding11321452
SWE-bench Verified—57.6%
SWE-bench Verified (bash only)—53.6%
Aider Polyglot—83.1%
LMArena WebDev—1227
SciCode—42.8%
GSO—3.9%
WeirdML—54%
LiveBench Coding—85.9%
CadEval—64%
ALE-Bench—785.52
AlgoTune—1.51
HumanEval+70.1%—
MBPP+55.8%—

Agentic & Tool Use Not comparable

DBRX: —, Gemini 2.5 Pro: 29.2 (#88)

Agentic & Tool Use benchmarks
BenchmarkDBRXGemini 2.5 Pro
Terminal-Bench—32.6%
GDPval—23.3%
Remote Labor Index—0.8%
TheAgentCompany—30.3%
τ²-bench Banking—13.7%
DeepResearch Bench—42.8%
BALROG—43.3%
LMArena Search—1142
METR Time Horizons—55.4%
Vending-Bench 2—573.64

Reasoning Gemini 2.5 Pro leads

DBRX: 21.4 (#222), Gemini 2.5 Pro: 28.8 (#99)

Reasoning benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Hard Prompts11131455
ARC-AGI-2—4.9%
SimpleBench—62.4%
Kagi LLM Benchmark—70.3%
ARC-AGI-1—41%
CritPt—2%
Chess Puzzles—20%
EnigmaEval—5.6%
LiveBench Reasoning—89.8%
DTBench—82.4%
LiveBench Data Analysis—79.9%
LMCA—34.8%
Epoch Capabilities Index—145.32
ForecastBench—61.3
LiveBench—82.3%

Math Gemini 2.5 Pro leads

DBRX: 24.3 (#269), Gemini 2.5 Pro: 32.5 (#213)

Math benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Math11451450
MATH Level 511.7%95.9%
FrontierMath (Tiers 1-3)—24.6%
FrontierMath Tier 4—0%
OTIS Mock AIME 2024-2025—84.7%
Omni-MATH—41.6%
LiveBench Math—90.2%
FrontierMath (Feb 2025 set)—14.1%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Gemini 2.5 Pro leads

DBRX: 14.9 (#294), Gemini 2.5 Pro: 56.0 (#46)

Knowledge benchmarks
BenchmarkDBRXGemini 2.5 Pro
GPQA Diamond32.9%85.3%
LMArena Expert10761452
Humanity's Last Exam—21.6%
MMLU-Pro—86.3%
Confabulations—10.6%
Vectara Hallucination Rate—7%
GPQA (HELM)—74.9%

Multimodal Not comparable

DBRX: —, Gemini 2.5 Pro: 45.2 (#18)

Multimodal benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Vision—1263
GeoBench—86%
VPCT—48%
LMArena Document—1421
SpatialViz-Bench—44.7%

Multilingual Gemini 2.5 Pro leads

DBRX: 29.3 (#268), Gemini 2.5 Pro: 55.3 (#31)

Multilingual benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Non-English10711451
LMArena Chinese10681507
LMArena French10961472
LMArena German10571487
LMArena Japanese9901461
LMArena Korean9931434
LMArena Russian10781461
LMArena Spanish10641473

Instruction Following Gemini 2.5 Pro leads

DBRX: 57.5 (#270), Gemini 2.5 Pro: 75.0 (#75)

Instruction Following benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Instruction Following11121437
LiveBench Instruction Following—80.6%
IFEval—84%

Long Context Gemini 2.5 Pro leads

DBRX: 33.7 (#258), Gemini 2.5 Pro: 59.8 (#5)

Long Context benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Longer Query11121449
Fiction.LiveBench—91.7%

Writing & Preference Gemini 2.5 Pro leads

DBRX: 33.6 (#275), Gemini 2.5 Pro: 63.7 (#62)

Writing & Preference benchmarks
BenchmarkDBRXGemini 2.5 Pro
LMArena Text11191458
LMArena Creative Writing11041454
LMArena Multi-Turn11111453
Short-Story Creative Writing—83.8%
EQ-Bench Creative Writing—1421
WildBench—85.7%
LiveBench Language—67.8%

Frequently asked questions

Is DBRX better than Gemini 2.5 Pro?

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 29.4 on the Noometry Index.

Is DBRX or Gemini 2.5 Pro better for coding?

Gemini 2.5 Pro scores higher on coding benchmarks: 42.4 versus 32.9 in the Noometry coding category.

How many benchmarks do DBRX and Gemini 2.5 Pro share?

19 benchmarks have published results for both models. DBRX has 21 scored results on Noometry and Gemini 2.5 Pro has 78.

Related comparisons

Go deeper