Model comparison

ERNIE 5.1 vs Granite 4.2 30b

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 41.8 on the Noometry Index.

Last verified . 11 shared benchmarks.

ERNIE 5.1 Baidu

43.8

Rank #83 Confirmed

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Summary

  • They share 11 benchmarks with published results for both. ERNIE 5.1 scores higher in 6 categories and Granite 4.2 30b in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where ERNIE 5.1 leads 65.1 to 53.8.
  • Granite 4.2 30b has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.1 and Granite 4.2 30b specifications
ERNIE 5.1Granite 4.2 30b
ProviderBaiduIBM
Noometry Index43.841.8
Released——
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1911

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.1 leads

ERNIE 5.1: 44.0 (#76), Granite 4.2 30b: 41.0 (#126)

Coding benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Coding14881396

Agentic & Tool Use Not comparable

ERNIE 5.1: —, Granite 4.2 30b: —

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Search1227—

Reasoning Granite 4.2 30b leads

ERNIE 5.1: 21.9 (#211), Granite 4.2 30b: 27.8 (#112)

Reasoning benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Hard Prompts14811374
NYT Connections (extended)23.4%—

Math Not comparable

ERNIE 5.1: 40.3 (#92), Granite 4.2 30b: —

Math benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Math1481—

Knowledge ERNIE 5.1 leads

ERNIE 5.1: 41.9 (#102), Granite 4.2 30b: 39.1 (#138)

Knowledge benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Expert14921406

Multilingual ERNIE 5.1 leads

ERNIE 5.1: 55.5 (#29), Granite 4.2 30b: 47.3 (#151)

Multilingual benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Non-English14541340
LMArena Chinese15081414
LMArena Russian14591343
LMArena French1488—
LMArena German1470—
LMArena Japanese1422—
LMArena Korean1427—
LMArena Spanish1473—

Instruction Following ERNIE 5.1 leads

ERNIE 5.1: 76.7 (#37), Granite 4.2 30b: 71.2 (#155)

Instruction Following benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Instruction Following14601347

Long Context ERNIE 5.1 leads

ERNIE 5.1: 44.7 (#59), Granite 4.2 30b: 41.4 (#140)

Long Context benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Longer Query14621359

Writing & Preference ERNIE 5.1 leads

ERNIE 5.1: 65.1 (#52), Granite 4.2 30b: 53.8 (#156)

Writing & Preference benchmarks
BenchmarkERNIE 5.1Granite 4.2 30b
LMArena Text14681361
LMArena Creative Writing14411288
LMArena Multi-Turn14711339

Frequently asked questions

Is ERNIE 5.1 better than Granite 4.2 30b?

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 41.8 on the Noometry Index.

Is ERNIE 5.1 or Granite 4.2 30b better for coding?

ERNIE 5.1 scores higher on coding benchmarks: 44.0 versus 41.0 in the Noometry coding category.

How many benchmarks do ERNIE 5.1 and Granite 4.2 30b share?

11 benchmarks have published results for both models. ERNIE 5.1 has 19 scored results on Noometry and Granite 4.2 30b has 11.

Related comparisons

Go deeper