Model comparison

Chatgpt 4o Latest 20250326 vs ERNIE 5.1

Chatgpt 4o Latest 20250326 and ERNIE 5.1 score almost the same on the Noometry Index (43.8 vs 43.8), so choose on price, context window or the category you care about most.

Last verified . 17 shared benchmarks.

Chatgpt 4o Latest 20250326 OpenAI

43.8

Rank #82 Confirmed

ERNIE 5.1 Baidu

43.8

Rank #83 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Chatgpt 4o Latest 20250326 scores higher in 1 category and ERNIE 5.1 in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Chatgpt 4o Latest 20250326 leads 33.8 to 21.9.

Side by side

Chatgpt 4o Latest 20250326 and ERNIE 5.1 specifications
Chatgpt 4o Latest 20250326ERNIE 5.1
ProviderOpenAIBaidu
Noometry Index43.843.8
Released——
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2119

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 41.6 (#122), ERNIE 5.1: 44.0 (#76)

Coding benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Coding14131488

Agentic & Tool Use Not comparable

Chatgpt 4o Latest 20250326: —, ERNIE 5.1: —

Agentic & Tool Use benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Search—1227

Reasoning Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 33.8 (#71), ERNIE 5.1: 21.9 (#211)

Reasoning benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Hard Prompts14241481
Kagi LLM Benchmark75%—
NYT Connections (extended)—23.4%

Math ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 38.6 (#134), ERNIE 5.1: 40.3 (#92)

Math benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Math14071481

Knowledge ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 39.2 (#137), ERNIE 5.1: 41.9 (#102)

Knowledge benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Expert14011492
Confabulations16.6%—

Multimodal Not comparable

Chatgpt 4o Latest 20250326: 39.6 (#58), ERNIE 5.1: —

Multimodal benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Vision1243—

Multilingual ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 52.9 (#76), ERNIE 5.1: 55.5 (#29)

Multilingual benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Non-English14191454
LMArena Chinese14571508
LMArena French14461488
LMArena German14231470
LMArena Japanese14051422
LMArena Korean13961427
LMArena Russian14291459
LMArena Spanish14341473

Instruction Following ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 74.0 (#107), ERNIE 5.1: 76.7 (#37)

Instruction Following benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Instruction Following14031460

Long Context ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 43.1 (#107), ERNIE 5.1: 44.7 (#59)

Long Context benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Longer Query14131462

Writing & Preference ERNIE 5.1 leads

Chatgpt 4o Latest 20250326: 62.6 (#72), ERNIE 5.1: 65.1 (#52)

Writing & Preference benchmarks
BenchmarkChatgpt 4o Latest 20250326ERNIE 5.1
LMArena Text14291468
LMArena Creative Writing14051441
LMArena Multi-Turn14541471
EQ-Bench Creative Writing1501—

Frequently asked questions

Is Chatgpt 4o Latest 20250326 better than ERNIE 5.1?

Chatgpt 4o Latest 20250326 and ERNIE 5.1 score almost the same on the Noometry Index (43.8 vs 43.8), so choose on price, context window or the category you care about most.

Is Chatgpt 4o Latest 20250326 or ERNIE 5.1 better for coding?

ERNIE 5.1 scores higher on coding benchmarks: 44.0 versus 41.6 in the Noometry coding category.

How many benchmarks do Chatgpt 4o Latest 20250326 and ERNIE 5.1 share?

17 benchmarks have published results for both models. Chatgpt 4o Latest 20250326 has 21 scored results on Noometry and ERNIE 5.1 has 19.

Related comparisons

Go deeper