Model comparison

ERNIE 5.0 0110 vs Qwen1.5-72B

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 30.8 on the Noometry Index.

Last verified . 17 shared benchmarks.

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

Qwen1.5-72B Alibaba (Qwen)

30.8

Rank #285 Confirmed

Summary

  • They share 17 benchmarks with published results for both. ERNIE 5.0 0110 scores higher in 7 categories and Qwen1.5-72B in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where ERNIE 5.0 0110 leads 39.8 to 11.5.
  • Qwen1.5-72B has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.0 0110 and Qwen1.5-72B specifications
ERNIE 5.0 0110Qwen1.5-72B
ProviderBaiduAlibaba (Qwen)
Noometry Index41.830.8
Released—2024-02-04
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2022

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 43.0 (#94), Qwen1.5-72B: 31.9 (#277)

Coding benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Coding14551165
BigCodeBench Instruct—33.2%
BigCodeBench Complete—40.3%
HumanEval+—59.1%
MBPP+—61.6%

Reasoning Qwen1.5-72B leads

ERNIE 5.0 0110: 17.0 (#297), Qwen1.5-72B: 22.2 (#203)

Reasoning benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Hard Prompts14451148
NYT Connections (extended)10.3%—
Thematic Generalization41.7%—

Math ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 39.3 (#110), Qwen1.5-72B: 33.2 (#205)

Math benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Math14371164

Knowledge ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 39.8 (#128), Qwen1.5-72B: 11.5 (#300)

Knowledge benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Expert14281136
GPQA Diamond—28.8%

Multimodal Not comparable

ERNIE 5.0 0110: 39.9 (#53), Qwen1.5-72B: —

Multimodal benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Vision1249—

Multilingual ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 54.1 (#49), Qwen1.5-72B: 33.2 (#253)

Multilingual benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Non-English14361135
LMArena Chinese15121186
LMArena French14671159
LMArena German14601084
LMArena Japanese13821061
LMArena Korean14061050
LMArena Russian14461104
LMArena Spanish14731110

Instruction Following ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 74.5 (#92), Qwen1.5-72B: 59.3 (#256)

Instruction Following benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Instruction Following14131141

Long Context ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 43.4 (#95), Qwen1.5-72B: 35.1 (#243)

Long Context benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Longer Query14221157

Writing & Preference ERNIE 5.0 0110 leads

ERNIE 5.0 0110: 63.1 (#66), Qwen1.5-72B: 37.3 (#258)

Writing & Preference benchmarks
BenchmarkERNIE 5.0 0110Qwen1.5-72B
LMArena Text14451166
LMArena Creative Writing14261137
LMArena Multi-Turn14341160

Frequently asked questions

Is ERNIE 5.0 0110 better than Qwen1.5-72B?

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 30.8 on the Noometry Index.

Is ERNIE 5.0 0110 or Qwen1.5-72B better for coding?

ERNIE 5.0 0110 scores higher on coding benchmarks: 43.0 versus 31.9 in the Noometry coding category.

How many benchmarks do ERNIE 5.0 0110 and Qwen1.5-72B share?

17 benchmarks have published results for both models. ERNIE 5.0 0110 has 20 scored results on Noometry and Qwen1.5-72B has 22.

Related comparisons

Go deeper