Model comparison

ERNIE 5.0 0110 vs Kimi K2.5

Kimi K2.5 is the stronger model overall, scoring 48.1 to 41.8 on the Noometry Index.

Last verified . 20 shared benchmarks.

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

Kimi K2.5 Moonshot AI

48.1

Rank #57 Confirmed

Summary

  • They share 20 benchmarks with published results for both. ERNIE 5.0 0110 scores higher in 1 category and Kimi K2.5 in 8 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Kimi K2.5 leads 31.2 to 17.0.
  • The biggest single-benchmark swing is NYT Connections (extended): 10.3% for ERNIE 5.0 0110 and 69.9% for Kimi K2.5.
  • Kimi K2.5 has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.0 0110 and Kimi K2.5 specifications
ERNIE 5.0 0110Kimi K2.5
ProviderBaiduMoonshot AI
Noometry Index41.848.1
Released—2026-01-27
WeightsProprietaryOpen
Context window—262K
Max output—262K
Input $ / M tokens—$0.45
Output $ / M tokens—$2.25
Results tracked2051

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.5 leads

ERNIE 5.0 0110: 43.0 (#94), Kimi K2.5: 48.8 (#53)

Coding benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Coding14551474
SWE-bench Verified—73.8%
SWE-bench Verified (bash only)—70.8%
LMArena WebDev—1437
SWE-bench Multilingual—67.3%
SciCode—49%
WeirdML—45.6%
ALE-Bench—821.65

Agentic & Tool Use Not comparable

ERNIE 5.0 0110: —, Kimi K2.5: 34.2 (#48)

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
Terminal-Bench—43.2%
OSWorld—63.3%
Vending-Bench 2—1,198

Reasoning Kimi K2.5 leads

ERNIE 5.0 0110: 17.0 (#297), Kimi K2.5: 31.2 (#80)

Reasoning benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
NYT Connections (extended)10.3%69.9%
Thematic Generalization41.7%69.4%
LMArena Hard Prompts14451453
ARC-AGI-2—11.8%
SimpleBench—46.8%
Kagi LLM Benchmark—78.5%
ARC-AGI-1—65.3%
CritPt—3.1%
Chess Puzzles—12%
EnigmaEval—3.4%
Epoch Capabilities Index—148.03

Math Kimi K2.5 leads

ERNIE 5.0 0110: 39.3 (#110), Kimi K2.5: 51.8 (#53)

Math benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Math14371470
MathArena Final-Answer Competitions—62.3%
OTIS Mock AIME 2024-2025—92.2%
FrontierMath (Feb 2025 set)—27.9%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Kimi K2.5 leads

ERNIE 5.0 0110: 39.8 (#128), Kimi K2.5: 53.6 (#56)

Knowledge benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Expert14281466
GPQA Diamond—87.6%
Humanity's Last Exam—24.4%
SimpleQA Verified—34.3%
Vectara Hallucination Rate—14.2%

Multimodal Kimi K2.5 leads

ERNIE 5.0 0110: 39.9 (#53), Kimi K2.5: 41.1 (#39)

Multimodal benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Vision12491269
LMArena Document—1430

Multilingual Too close to call

ERNIE 5.0 0110: 54.1 (#49), Kimi K2.5: 53.9 (#53)

Multilingual benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Non-English14361433
LMArena Chinese15121495
LMArena French14671454
LMArena German14601441
LMArena Japanese13821421
LMArena Korean14061410
LMArena Russian14461435
LMArena Spanish14731450

Instruction Following Too close to call

ERNIE 5.0 0110: 74.5 (#92), Kimi K2.5: 75.3 (#64)

Instruction Following benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Instruction Following14131431

Long Context Kimi K2.5 leads

ERNIE 5.0 0110: 43.4 (#95), Kimi K2.5: 52.1 (#7)

Long Context benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Longer Query14221445
Fiction.LiveBench—86.1%
CL-bench—19.3%
CL-bench Life—13.2%

Writing & Preference Kimi K2.5 leads

ERNIE 5.0 0110: 63.1 (#66), Kimi K2.5: 65.1 (#53)

Writing & Preference benchmarks
BenchmarkERNIE 5.0 0110Kimi K2.5
LMArena Text14451445
LMArena Creative Writing14261423
LMArena Multi-Turn14341444
EQ-Bench Creative Writing—1579

Frequently asked questions

Is ERNIE 5.0 0110 better than Kimi K2.5?

Kimi K2.5 is the stronger model overall, scoring 48.1 to 41.8 on the Noometry Index.

Is ERNIE 5.0 0110 or Kimi K2.5 better for coding?

Kimi K2.5 scores higher on coding benchmarks: 48.8 versus 43.0 in the Noometry coding category.

How many benchmarks do ERNIE 5.0 0110 and Kimi K2.5 share?

20 benchmarks have published results for both models. ERNIE 5.0 0110 has 20 scored results on Noometry and Kimi K2.5 has 51.

Related comparisons

Go deeper