Model comparison

Claude 2.1 vs ERNIE 5.0 0110

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 25.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 2.1 Anthropic

25.2

Rank #345 Reported

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

Summary

  • The widest gap is in math, where ERNIE 5.0 0110 leads 39.3 to 10.2.

Side by side

Claude 2.1 and ERNIE 5.0 0110 specifications
Claude 2.1ERNIE 5.0 0110
ProviderAnthropicBaidu
Noometry Index25.241.8
Released2023-11-21—
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked720

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.0 0110 leads

Claude 2.1: 26.2 (#327), ERNIE 5.0 0110: 43.0 (#94)

Coding benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
WeirdML7.1%—
LMArena Coding—1455

Reasoning Claude 2.1 leads

Claude 2.1: 21.4 (#221), ERNIE 5.0 0110: 17.0 (#297)

Reasoning benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
NYT Connections (extended)—10.3%
Thematic Generalization—41.7%
LMArena Hard Prompts—1445
DTBench51%—
Epoch Capabilities Index119.27—
ForecastBench54.2—

Math ERNIE 5.0 0110 leads

Claude 2.1: 10.2 (#315), ERNIE 5.0 0110: 39.3 (#110)

Math benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
OTIS Mock AIME 2024-20251.9%—
LMArena Math—1437

Knowledge ERNIE 5.0 0110 leads

Claude 2.1: 15.4 (#292), ERNIE 5.0 0110: 39.8 (#128)

Knowledge benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
GPQA Diamond33%—
LMArena Expert—1428
MMLU73.5%—

Multimodal Not comparable

Claude 2.1: —, ERNIE 5.0 0110: 39.9 (#53)

Multimodal benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
LMArena Vision—1249

Multilingual Not comparable

Claude 2.1: —, ERNIE 5.0 0110: 54.1 (#49)

Multilingual benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
LMArena Non-English—1436
LMArena Chinese—1512
LMArena French—1467
LMArena German—1460
LMArena Japanese—1382
LMArena Korean—1406
LMArena Russian—1446
LMArena Spanish—1473

Instruction Following Not comparable

Claude 2.1: —, ERNIE 5.0 0110: 74.5 (#92)

Instruction Following benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
LMArena Instruction Following—1413

Long Context Not comparable

Claude 2.1: —, ERNIE 5.0 0110: 43.4 (#95)

Long Context benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
LMArena Longer Query—1422

Writing & Preference Not comparable

Claude 2.1: —, ERNIE 5.0 0110: 63.1 (#66)

Writing & Preference benchmarks
BenchmarkClaude 2.1ERNIE 5.0 0110
LMArena Text—1445
LMArena Creative Writing—1426
LMArena Multi-Turn—1434

Frequently asked questions

Is Claude 2.1 better than ERNIE 5.0 0110?

ERNIE 5.0 0110 is the stronger model overall, scoring 41.8 to 25.2 on the Noometry Index.

Is Claude 2.1 or ERNIE 5.0 0110 better for coding?

ERNIE 5.0 0110 scores higher on coding benchmarks: 43.0 versus 26.2 in the Noometry coding category.

How many benchmarks do Claude 2.1 and ERNIE 5.0 0110 share?

0 benchmarks have published results for both models. Claude 2.1 has 7 scored results on Noometry and ERNIE 5.0 0110 has 20.

Related comparisons

Go deeper