Model comparison

ERNIE 5.0 0110 vs GLM-5.1

GLM-5.1 is the stronger model overall, scoring 47.8 to 41.8 on the Noometry Index.

Last verified . 19 shared benchmarks.

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

GLM-5.1 Z.ai (Zhipu)

47.8

Rank #59 Confirmed

Summary

  • They share 19 benchmarks with published results for both. ERNIE 5.0 0110 scores higher in 0 categories and GLM-5.1 in 8 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where GLM-5.1 leads 39.1 to 17.0.
  • The biggest single-benchmark swing is NYT Connections (extended): 10.3% for ERNIE 5.0 0110 and 77.7% for GLM-5.1.
  • GLM-5.1 has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.0 0110 and GLM-5.1 specifications
ERNIE 5.0 0110GLM-5.1
ProviderBaiduZ.ai (Zhipu)
Noometry Index41.847.8
Released—2026-04-07
WeightsProprietaryOpen
Context window—200K
Max output—131K
Input $ / M tokens—$1.40
Output $ / M tokens—$4.40
Results tracked2041

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-5.1 leads

ERNIE 5.0 0110: 43.0 (#94), GLM-5.1: 48.7 (#55)

Coding benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Coding14551485
SWE-bench Verified—74.2%
LMArena WebDev—1508
SciCode—43.8%
WeirdML—57.1%
ALE-Bench—887.1

Agentic & Tool Use Not comparable

ERNIE 5.0 0110: —, GLM-5.1: 24.9 (#113)

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
APEX-Agents—40.9%
ExploitBench—18.1%
GBAEval—0%
Vending-Bench 2—5,634

Reasoning GLM-5.1 leads

ERNIE 5.0 0110: 17.0 (#297), GLM-5.1: 39.1 (#60)

Reasoning benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
NYT Connections (extended)10.3%77.7%
Thematic Generalization41.7%69.8%
LMArena Hard Prompts14451472
SimpleBench—55.1%
CritPt—4.6%
Chess Puzzles—19%
Epoch Capabilities Index—149.84

Math GLM-5.1 leads

ERNIE 5.0 0110: 39.3 (#110), GLM-5.1: 49.7 (#60)

Knowledge GLM-5.1 leads

ERNIE 5.0 0110: 39.8 (#128), GLM-5.1: 54.9 (#50)

Knowledge benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Expert14281476
GPQA Diamond—89.9%
SimpleQA Verified—34%

Multimodal Not comparable

ERNIE 5.0 0110: 39.9 (#53), GLM-5.1: —

Multimodal benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Vision1249—

Multilingual Too close to call

ERNIE 5.0 0110: 54.1 (#49), GLM-5.1: 55.0 (#36)

Multilingual benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Non-English14361447
LMArena Chinese15121515
LMArena French14671474
LMArena German14601465
LMArena Japanese13821434
LMArena Korean14061418
LMArena Russian14461454
LMArena Spanish14731469

Instruction Following GLM-5.1 leads

ERNIE 5.0 0110: 74.5 (#92), GLM-5.1: 76.3 (#42)

Instruction Following benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Instruction Following14131451

Long Context GLM-5.1 leads

ERNIE 5.0 0110: 43.4 (#95), GLM-5.1: 44.9 (#53)

Long Context benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Longer Query14221466

Writing & Preference GLM-5.1 leads

ERNIE 5.0 0110: 63.1 (#66), GLM-5.1: 66.9 (#31)

Writing & Preference benchmarks
BenchmarkERNIE 5.0 0110GLM-5.1
LMArena Text14451461
LMArena Creative Writing14261453
LMArena Multi-Turn14341472
EQ-Bench Creative Writing—1592

Frequently asked questions

Is ERNIE 5.0 0110 better than GLM-5.1?

GLM-5.1 is the stronger model overall, scoring 47.8 to 41.8 on the Noometry Index.

Is ERNIE 5.0 0110 or GLM-5.1 better for coding?

GLM-5.1 scores higher on coding benchmarks: 48.7 versus 43.0 in the Noometry coding category.

How many benchmarks do ERNIE 5.0 0110 and GLM-5.1 share?

19 benchmarks have published results for both models. ERNIE 5.0 0110 has 20 scored results on Noometry and GLM-5.1 has 41.

Related comparisons

Go deeper