Model comparison

ERNIE 5.1 vs Trinity Large Thinking

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 38.6 on the Noometry Index.

Last verified . 18 shared benchmarks.

ERNIE 5.1 Baidu

43.8

Rank #83 Confirmed

Trinity Large Thinking Arcee AI

38.6

Rank #185 Confirmed

Summary

  • They share 18 benchmarks with published results for both. ERNIE 5.1 scores higher in 8 categories and Trinity Large Thinking in 0 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where ERNIE 5.1 leads 65.1 to 53.8.
  • The biggest single-benchmark swing is NYT Connections (extended): 23.4% for ERNIE 5.1 and 16.5% for Trinity Large Thinking.
  • Trinity Large Thinking has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.1 and Trinity Large Thinking specifications
ERNIE 5.1Trinity Large Thinking
ProviderBaiduArcee AI
Noometry Index43.838.6
Released—2026-04-01
WeightsProprietaryOpen
Context window—262K
Max output—80K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.80
Results tracked1924

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.1 leads

ERNIE 5.1: 44.0 (#76), Trinity Large Thinking: 34.1 (#244)

Coding benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Coding14881381
LMArena WebDev—1238
SciCode—36.1%

Agentic & Tool Use Not comparable

ERNIE 5.1: —, Trinity Large Thinking: —

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Search1227—

Reasoning ERNIE 5.1 leads

ERNIE 5.1: 21.9 (#211), Trinity Large Thinking: 16.9 (#298)

Reasoning benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
NYT Connections (extended)23.4%16.5%
LMArena Hard Prompts14811350
CritPt—0.9%
Thematic Generalization—41.6%
Surface Evolver Bench—15.6%

Math ERNIE 5.1 leads

ERNIE 5.1: 40.3 (#92), Trinity Large Thinking: 37.6 (#149)

Math benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Math14811366

Knowledge Too close to call

ERNIE 5.1: 41.9 (#102), Trinity Large Thinking: 40.9 (#113)

Knowledge benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Expert14921360
Vectara Hallucination Rate—6.9%

Multilingual ERNIE 5.1 leads

ERNIE 5.1: 55.5 (#29), Trinity Large Thinking: 46.2 (#160)

Multilingual benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Non-English14541325
LMArena Chinese15081373
LMArena French14881374
LMArena German14701356
LMArena Japanese14221311
LMArena Korean14271306
LMArena Russian14591337
LMArena Spanish14731357

Instruction Following ERNIE 5.1 leads

ERNIE 5.1: 76.7 (#37), Trinity Large Thinking: 70.5 (#162)

Instruction Following benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Instruction Following14601334

Long Context ERNIE 5.1 leads

ERNIE 5.1: 44.7 (#59), Trinity Large Thinking: 41.3 (#144)

Long Context benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Longer Query14621355

Writing & Preference ERNIE 5.1 leads

ERNIE 5.1: 65.1 (#52), Trinity Large Thinking: 53.8 (#158)

Writing & Preference benchmarks
BenchmarkERNIE 5.1Trinity Large Thinking
LMArena Text14681340
LMArena Creative Writing14411320
LMArena Multi-Turn14711342

Frequently asked questions

Is ERNIE 5.1 better than Trinity Large Thinking?

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 38.6 on the Noometry Index.

Is ERNIE 5.1 or Trinity Large Thinking better for coding?

ERNIE 5.1 scores higher on coding benchmarks: 44.0 versus 34.1 in the Noometry coding category.

How many benchmarks do ERNIE 5.1 and Trinity Large Thinking share?

18 benchmarks have published results for both models. ERNIE 5.1 has 19 scored results on Noometry and Trinity Large Thinking has 24.

Related comparisons

Go deeper