Model comparison

DBRX vs Qwen3.5 Plus

Qwen3.5 Plus is the stronger model overall, scoring 42.9 to 29.4 on the Noometry Index.

Last verified . 1 shared benchmarks.

DBRX Databricks

29.4

Rank #311 Confirmed

Qwen3.5 Plus Alibaba (Qwen)

42.9

Rank #106 Confirmed

Summary

  • They share 1 benchmark with published results for both. DBRX scores higher in 0 categories and Qwen3.5 Plus in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3.5 Plus leads 46.0 to 14.9.
  • The biggest single-benchmark swing is GPQA Diamond: 32.9% for DBRX and 84.8% for Qwen3.5 Plus.
  • DBRX has downloadable open weights; the other is API-only.

Side by side

DBRX and Qwen3.5 Plus specifications
DBRXQwen3.5 Plus
ProviderDatabricksAlibaba (Qwen)
Noometry Index29.442.9
Released2024-03-272026-02-16
WeightsOpenProprietary
Context window—1M
Max output—66K
Input $ / M tokens—$0.40
Output $ / M tokens—$2.40
Results tracked2115

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

DBRX: 32.9 (#266), Qwen3.5 Plus: —

Coding benchmarks
BenchmarkDBRXQwen3.5 Plus
LMArena Coding1132—
ALE-Bench—621.92
HumanEval+70.1%—
MBPP+55.8%—

Agentic & Tool Use Not comparable

DBRX: —, Qwen3.5 Plus: —

Agentic & Tool Use benchmarks
BenchmarkDBRXQwen3.5 Plus
Vending-Bench 2—0.54

Reasoning Qwen3.5 Plus leads

DBRX: 21.4 (#222), Qwen3.5 Plus: 32.8 (#74)

Reasoning benchmarks
BenchmarkDBRXQwen3.5 Plus
Chess Puzzles—22%
LMArena Hard Prompts1113—
Mystery Game Puzzles—17%
DTBench—80.5%
LMCA—36.4%
Epoch Capabilities Index—146.78

Math Qwen3.5 Plus leads

DBRX: 24.3 (#269), Qwen3.5 Plus: 49.6 (#61)

Math benchmarks
BenchmarkDBRXQwen3.5 Plus
OTIS Mock AIME 2024-2025—86.7%
LMArena Math1145—
MATH Level 511.7%—
FrontierMath (Feb 2025 set)—21%
FrontierMath Tier 4 (v1)—2.1%

Knowledge Qwen3.5 Plus leads

DBRX: 14.9 (#294), Qwen3.5 Plus: 46.0 (#83)

Knowledge benchmarks
BenchmarkDBRXQwen3.5 Plus
GPQA Diamond32.9%84.8%
SimpleQA Verified—25.4%
Vectara Hallucination Rate—10.7%
LMArena Expert1076—

Multilingual Not comparable

DBRX: 29.3 (#268), Qwen3.5 Plus: —

Multilingual benchmarks
BenchmarkDBRXQwen3.5 Plus
LMArena Non-English1071—
LMArena Chinese1068—
LMArena French1096—
LMArena German1057—
LMArena Japanese990—
LMArena Korean993—
LMArena Russian1078—
LMArena Spanish1064—

Instruction Following Not comparable

DBRX: 57.5 (#270), Qwen3.5 Plus: —

Instruction Following benchmarks
BenchmarkDBRXQwen3.5 Plus
LMArena Instruction Following1112—

Long Context Qwen3.5 Plus leads

DBRX: 33.7 (#258), Qwen3.5 Plus: 43.0 (#113)

Long Context benchmarks
BenchmarkDBRXQwen3.5 Plus
CL-bench—19.8%
CL-bench Life—12.4%
LMArena Longer Query1112—

Writing & Preference Not comparable

DBRX: 33.6 (#275), Qwen3.5 Plus: —

Writing & Preference benchmarks
BenchmarkDBRXQwen3.5 Plus
LMArena Text1119—
LMArena Creative Writing1104—
LMArena Multi-Turn1111—

Frequently asked questions

Is DBRX better than Qwen3.5 Plus?

Qwen3.5 Plus is the stronger model overall, scoring 42.9 to 29.4 on the Noometry Index.

How many benchmarks do DBRX and Qwen3.5 Plus share?

1 benchmark has published results for both models. DBRX has 21 scored results on Noometry and Qwen3.5 Plus has 15.

Related comparisons

Go deeper