Model comparison

ERNIE 5.1 vs Gemini 1.5 Pro (May 2024)

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 32.1 on the Noometry Index.

Last verified . 17 shared benchmarks.

ERNIE 5.1 Baidu

43.8

Rank #83 Confirmed

Gemini 1.5 Pro (May 2024) Google

32.1

Rank #261 Confirmed

Summary

  • They share 17 benchmarks with published results for both. ERNIE 5.1 scores higher in 8 categories and Gemini 1.5 Pro (May 2024) in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where ERNIE 5.1 leads 40.3 to 25.8.

Side by side

ERNIE 5.1 and Gemini 1.5 Pro (May 2024) specifications
ERNIE 5.1Gemini 1.5 Pro (May 2024)
ProviderBaiduGoogle
Noometry Index43.832.1
Released—2024-02-15
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1945

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.1 leads

ERNIE 5.1: 44.0 (#76), Gemini 1.5 Pro (May 2024): 34.2 (#241)

Coding benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Coding14881294
WeirdML—22.2%
BigCodeBench Instruct—43.8%
BigCodeBench Complete—57.5%
CadEval—34%
HumanEval+—79.3%
MBPP+—74.6%

Agentic & Tool Use Not comparable

ERNIE 5.1: —, Gemini 1.5 Pro (May 2024): 17.9 (#145)

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
TheAgentCompany—3.4%
Cybench—7.5%
BALROG—21%
LMArena Search1227—

Reasoning ERNIE 5.1 leads

ERNIE 5.1: 21.9 (#211), Gemini 1.5 Pro (May 2024): 12.3 (#338)

Reasoning benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Hard Prompts14811296
ARC-AGI-2—0.8%
SimpleBench—27.1%
NYT Connections (extended)23.4%—
DTBench—59%
BIG-Bench Hard—89.2%
Epoch Capabilities Index—131.73
ForecastBench—58.4

Math ERNIE 5.1 leads

ERNIE 5.1: 40.3 (#92), Gemini 1.5 Pro (May 2024): 25.8 (#266)

Math benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Math14811315
OTIS Mock AIME 2024-2025—23.1%
Omni-MATH—36.4%
MATH Level 5—70.4%

Knowledge ERNIE 5.1 leads

ERNIE 5.1: 41.9 (#102), Gemini 1.5 Pro (May 2024): 29.4 (#239)

Knowledge benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Expert14921279
GPQA Diamond—57.2%
Humanity's Last Exam—4.6%
MMLU-Pro—73.7%
Confabulations—13.5%
GPQA (HELM)—53.4%
MMLU—86.9%

Multimodal Not comparable

ERNIE 5.1: —, Gemini 1.5 Pro (May 2024): 36.8 (#77)

Multimodal benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Vision—1161
Video-MME—75%

Multilingual ERNIE 5.1 leads

ERNIE 5.1: 55.5 (#29), Gemini 1.5 Pro (May 2024): 45.3 (#174)

Multilingual benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Non-English14541312
LMArena Chinese15081331
LMArena French14881302
LMArena German14701286
LMArena Japanese14221292
LMArena Korean14271298
LMArena Russian14591320
LMArena Spanish14731311

Instruction Following ERNIE 5.1 leads

ERNIE 5.1: 76.7 (#37), Gemini 1.5 Pro (May 2024): 68.6 (#185)

Instruction Following benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Instruction Following14601297
IFEval—83.7%

Long Context ERNIE 5.1 leads

ERNIE 5.1: 44.7 (#59), Gemini 1.5 Pro (May 2024): 39.8 (#169)

Long Context benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Longer Query14621308

Writing & Preference ERNIE 5.1 leads

ERNIE 5.1: 65.1 (#52), Gemini 1.5 Pro (May 2024): 52.4 (#172)

Writing & Preference benchmarks
BenchmarkERNIE 5.1Gemini 1.5 Pro (May 2024)
LMArena Text14681319
LMArena Creative Writing14411333
LMArena Multi-Turn14711296
WildBench—81.3%

Frequently asked questions

Is ERNIE 5.1 better than Gemini 1.5 Pro (May 2024)?

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 32.1 on the Noometry Index.

Is ERNIE 5.1 or Gemini 1.5 Pro (May 2024) better for coding?

ERNIE 5.1 scores higher on coding benchmarks: 44.0 versus 34.2 in the Noometry coding category.

How many benchmarks do ERNIE 5.1 and Gemini 1.5 Pro (May 2024) share?

17 benchmarks have published results for both models. ERNIE 5.1 has 19 scored results on Noometry and Gemini 1.5 Pro (May 2024) has 45.

Related comparisons

Go deeper