Model comparison

ERNIE 5.0 0110 vs Qwen Max

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 34.7 on the Noometry Index.

Last verified . 17 shared benchmarks.

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

Qwen Max Alibaba (Qwen)

34.7

Rank #230 Confirmed

Summary

  • They share 17 benchmarks with published results for both. ERNIE 5.0 0110 scores higher in 7 categories and Qwen Max in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where ERNIE 5.0 0110 leads 39.3 to 22.3.

Side by side

ERNIE 5.0 0110 and Qwen Max specifications
ERNIE 5.0 0110Qwen Max
ProviderBaiduAlibaba (Qwen)
Noometry Index41.834.7
Released—2024-04-03
WeightsProprietaryProprietary
Context window—33K
Max output—8K
Input $ / M tokens—$1.60
Output $ / M tokens—$6.40
Results tracked2023

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 43.0 (#94), Qwen Max: 30.7 (#292)

Coding benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Coding14551288
Aider Polyglot—21.8%

Reasoning Qwen Max leads

ERNIE 5.0 0110: 17.0 (#297), Qwen Max: 25.1 (#151)

Reasoning benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Hard Prompts14451269
NYT Connections (extended)10.3%—
Thematic Generalization41.7%—

Math ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 39.3 (#110), Qwen Max: 22.3 (#276)

Math benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Math14371275
OTIS Mock AIME 2024-2025—16.1%
MATH Level 5—67.2%
FrontierMath (Feb 2025 set)—1%

Knowledge ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 39.8 (#128), Qwen Max: 30.3 (#228)

Knowledge benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Expert14281248
GPQA Diamond—56.1%

Multimodal Not comparable

ERNIE 5.0 0110: 39.9 (#53), Qwen Max: —

Multimodal benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Vision1249—

Multilingual ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 54.1 (#49), Qwen Max: 41.8 (#202)

Multilingual benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Non-English14361263
LMArena Chinese15121254
LMArena French14671330
LMArena German14601254
LMArena Japanese13821205
LMArena Korean14061142
LMArena Russian14461274
LMArena Spanish14731290

Instruction Following ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 74.5 (#92), Qwen Max: 66.5 (#208)

Instruction Following benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Instruction Following14131262

Long Context ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 43.4 (#95), Qwen Max: 39.4 (#180)

Long Context benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Longer Query14221288
Fiction.LiveBench—66.7%

Writing & Preference ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 63.1 (#66), Qwen Max: 47.8 (#205)

Writing & Preference benchmarks
BenchmarkERNIE 5.0 0110Qwen Max
LMArena Text14451282
LMArena Creative Writing14261248
LMArena Multi-Turn14341277

Frequently asked questions

Is ERNIE 5.0 0110 better than Qwen Max?

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 34.7 on the Noometry Index.

Is ERNIE 5.0 0110 or Qwen Max better for coding?

ERNIE 5.0 0110 scores higher on coding benchmarks: 43.0 versus 30.7 in the Noometry coding category.

How many benchmarks do ERNIE 5.0 0110 and Qwen Max share?

17 benchmarks have published results for both models. ERNIE 5.0 0110 has 20 scored results on Noometry and Qwen Max has 23.

Related comparisons

Go deeper