Model comparison

ERNIE 5.0 0110 vs Qwen2.5 Plus 1127

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 38.8 on the Noometry Index.

Last verified . 14 shared benchmarks.

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

Qwen2.5 Plus 1127 Alibaba (Qwen)

38.8

Rank #181 Confirmed

Summary

  • They share 14 benchmarks with published results for both. ERNIE 5.0 0110 scores higher in 7 categories and Qwen2.5 Plus 1127 in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where ERNIE 5.0 0110 leads 63.1 to 49.4.

Side by side

ERNIE 5.0 0110 and Qwen2.5 Plus 1127 specifications
ERNIE 5.0 0110Qwen2.5 Plus 1127
ProviderBaiduAlibaba (Qwen)
Noometry Index41.838.8
Released——
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2014

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 43.0 (#94), Qwen2.5 Plus 1127: 38.5 (#175)

Coding benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Coding14551314

Reasoning Qwen2.5 Plus 1127 leads

ERNIE 5.0 0110: 17.0 (#297), Qwen2.5 Plus 1127: 25.9 (#141)

Reasoning benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Hard Prompts14451299
NYT Connections (extended)10.3%—
Thematic Generalization41.7%—

Math ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 39.3 (#110), Qwen2.5 Plus 1127: 36.1 (#174)

Math benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Math14371298

Knowledge ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 39.8 (#128), Qwen2.5 Plus 1127: 35.5 (#183)

Knowledge benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Expert14281289

Multimodal Not comparable

ERNIE 5.0 0110: 39.9 (#53), Qwen2.5 Plus 1127: —

Multimodal benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Vision1249—

Multilingual ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 54.1 (#49), Qwen2.5 Plus 1127: 41.9 (#201)

Multilingual benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Non-English14361265
LMArena Chinese15121314
LMArena German14601231
LMArena Japanese13821207
LMArena Russian14461271
LMArena French1467—
LMArena Korean1406—
LMArena Spanish1473—

Instruction Following ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 74.5 (#92), Qwen2.5 Plus 1127: 67.2 (#199)

Instruction Following benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Instruction Following14131275

Long Context ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 43.4 (#95), Qwen2.5 Plus 1127: 39.2 (#184)

Long Context benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Longer Query14221292

Writing & Preference ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 63.1 (#66), Qwen2.5 Plus 1127: 49.4 (#192)

Writing & Preference benchmarks
BenchmarkERNIE 5.0 0110Qwen2.5 Plus 1127
LMArena Text14451299
LMArena Creative Writing14261262
LMArena Multi-Turn14341299

Frequently asked questions

Is ERNIE 5.0 0110 better than Qwen2.5 Plus 1127?

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 38.8 on the Noometry Index.

Is ERNIE 5.0 0110 or Qwen2.5 Plus 1127 better for coding?

ERNIE 5.0 0110 scores higher on coding benchmarks: 43.0 versus 38.5 in the Noometry coding category.

How many benchmarks do ERNIE 5.0 0110 and Qwen2.5 Plus 1127 share?

14 benchmarks have published results for both models. ERNIE 5.0 0110 has 20 scored results on Noometry and Qwen2.5 Plus 1127 has 14.

Related comparisons

Go deeper