Model comparison

DBRX vs Qwen Max

Qwen Max is the stronger model overall, scoring 34.7 to 29.4 on the Noometry Index.

Last verified . 19 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

Qwen Max Alibaba (Qwen)

34.7

Rank #230 Confirmed

Summary

  • They share 19 benchmarks with published results for both. DBRX scores higher in 2 categories and Qwen Max in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen Max leads 30.3 to 14.9.
  • The biggest single-benchmark swing is MATH Level 5: 11.7% for DBRX and 67.2% for Qwen Max.
  • DBRX has downloadable open weights; the other is API-only.

Side by side

DBRX and Qwen Max specifications
DBRXQwen Max
ProviderDatabricksAlibaba (Qwen)
Noometry Index29.434.7
Released2024-03-272024-04-03
WeightsOpenProprietary
Context window—33K
Max output—8K
Input $ / M tokens—$1.60
Output $ / M tokens—$6.40
Results tracked2123

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DBRX leads

DBRX: 32.9 (#266), Qwen Max: 30.7 (#292)

Coding benchmarks
BenchmarkDBRXQwen Max
LMArena Coding11321288
Aider Polyglot—21.8%
HumanEval+70.1%—
MBPP+55.8%—

Reasoning Qwen Max leads

DBRX: 21.4 (#222), Qwen Max: 25.1 (#151)

Reasoning benchmarks
BenchmarkDBRXQwen Max
LMArena Hard Prompts11131269

Math DBRX leads

DBRX: 24.3 (#269), Qwen Max: 22.3 (#276)

Math benchmarks
BenchmarkDBRXQwen Max
LMArena Math11451275
MATH Level 511.7%67.2%
OTIS Mock AIME 2024-2025—16.1%
FrontierMath (Feb 2025 set)—1%

Knowledge Qwen Max leads

DBRX: 14.9 (#294), Qwen Max: 30.3 (#228)

Knowledge benchmarks
BenchmarkDBRXQwen Max
GPQA Diamond32.9%56.1%
LMArena Expert10761248

Multilingual Qwen Max leads

DBRX: 29.3 (#268), Qwen Max: 41.8 (#202)

Multilingual benchmarks
BenchmarkDBRXQwen Max
LMArena Non-English10711263
LMArena Chinese10681254
LMArena French10961330
LMArena German10571254
LMArena Japanese9901205
LMArena Korean9931142
LMArena Russian10781274
LMArena Spanish10641290

Instruction Following Qwen Max leads

DBRX: 57.5 (#270), Qwen Max: 66.5 (#208)

Instruction Following benchmarks
BenchmarkDBRXQwen Max
LMArena Instruction Following11121262

Long Context Qwen Max leads

DBRX: 33.7 (#258), Qwen Max: 39.4 (#180)

Long Context benchmarks
BenchmarkDBRXQwen Max
LMArena Longer Query11121288
Fiction.LiveBench—66.7%

Writing & Preference Qwen Max leads

DBRX: 33.6 (#275), Qwen Max: 47.8 (#205)

Writing & Preference benchmarks
BenchmarkDBRXQwen Max
LMArena Text11191282
LMArena Creative Writing11041248
LMArena Multi-Turn11111277

Frequently asked questions

Is DBRX better than Qwen Max?

Qwen Max is the stronger model overall, scoring 34.7 to 29.4 on the Noometry Index.

Is DBRX or Qwen Max better for coding?

DBRX scores higher on coding benchmarks: 32.9 versus 30.7 in the Noometry coding category.

How many benchmarks do DBRX and Qwen Max share?

19 benchmarks have published results for both models. DBRX has 21 scored results on Noometry and Qwen Max has 23.

Related comparisons

Go deeper