Model comparison

ERNIE 5.0 0110 vs Muse Spark 1.2

Muse Spark 1.2 is the stronger model overall, scoring 50.3 to 41.8 on the Noometry Index.

Last verified . 16 shared benchmarks.

ERNIE 5.0 0110 Baidu

41.8

Rank #129 Confirmed

Muse Spark 1.2 Meta

50.3

Rank #48 Confirmed

Summary

  • They share 16 benchmarks with published results for both. ERNIE 5.0 0110 scores higher in 0 categories and Muse Spark 1.2 in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Muse Spark 1.2 leads 51.3 to 17.0.
  • The biggest single-benchmark swing is NYT Connections (extended): 10.3% for ERNIE 5.0 0110 and 79.2% for Muse Spark 1.2.

Side by side

ERNIE 5.0 0110 and Muse Spark 1.2 specifications
ERNIE 5.0 0110Muse Spark 1.2
ProviderBaiduMeta
Noometry Index41.850.3
Released—2026-08-05
WeightsProprietaryProprietary
Context window—1.05M
Max output—131K
Input $ / M tokens—$1.25
Output $ / M tokens—$4.25
Results tracked2031

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Muse Spark 1.2 leads

ERNIE 5.0 0110: 43.0 (#94), Muse Spark 1.2: 49.2 (#51)

Coding benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Coding14551495
DeepSWE—54.9%
LMArena WebDev—1533
FrontierSWE—12%
SciCode—56.4%
WeirdML—60.3%

Agentic & Tool Use Not comparable

ERNIE 5.0 0110: —, Muse Spark 1.2: 29.4 (#87)

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
APEX-Agents—36.4%
GDP.pdf—16%

Reasoning Muse Spark 1.2 leads

ERNIE 5.0 0110: 17.0 (#297), Muse Spark 1.2: 51.3 (#34)

Reasoning benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
NYT Connections (extended)10.3%79.2%
LMArena Hard Prompts14451486
SimpleBench—74.5%
CritPt—17.7%
Thematic Generalization41.7%—
DTBench—94.7%
LMCA—48.4%
Epoch Capabilities Index—154.87

Math Muse Spark 1.2 leads

ERNIE 5.0 0110: 39.3 (#110), Muse Spark 1.2: 46.4 (#70)

Math benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Math14371471
ProofBench—43%

Knowledge Muse Spark 1.2 leads

ERNIE 5.0 0110: 39.8 (#128), Muse Spark 1.2: 54.1 (#53)

Knowledge benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Expert14281480
SimpleQA Verified—60.3%

Multimodal Muse Spark 1.2 leads

ERNIE 5.0 0110: 39.9 (#53), Muse Spark 1.2: 43.4 (#25)

Multimodal benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Vision12491305

Multilingual Muse Spark 1.2 leads

ERNIE 5.0 0110: 54.1 (#49), Muse Spark 1.2: 57.1 (#11)

Multilingual benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Non-English14361478
LMArena Chinese15121511
LMArena French14671513
LMArena Russian14461487
LMArena Spanish14731498
LMArena German1460—
LMArena Japanese1382—
LMArena Korean1406—

Instruction Following Muse Spark 1.2 leads

ERNIE 5.0 0110: 74.5 (#92), Muse Spark 1.2: 76.7 (#36)

Instruction Following benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Instruction Following14131461

Long Context Muse Spark 1.2 leads

ERNIE 5.0 0110: 43.4 (#95), Muse Spark 1.2: 45.2 (#48)

Long Context benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Longer Query14221475

Writing & Preference Muse Spark 1.2 leads

ERNIE 5.0 0110: 63.1 (#66), Muse Spark 1.2: 72.3 (#14)

Writing & Preference benchmarks
BenchmarkERNIE 5.0 0110Muse Spark 1.2
LMArena Text14451482
LMArena Creative Writing14261449
LMArena Multi-Turn14341494
EQ-Bench Creative Writing—1840

Frequently asked questions

Is ERNIE 5.0 0110 better than Muse Spark 1.2?

Muse Spark 1.2 is the stronger model overall, scoring 50.3 to 41.8 on the Noometry Index.

Is ERNIE 5.0 0110 or Muse Spark 1.2 better for coding?

Muse Spark 1.2 scores higher on coding benchmarks: 49.2 versus 43.0 in the Noometry coding category.

How many benchmarks do ERNIE 5.0 0110 and Muse Spark 1.2 share?

16 benchmarks have published results for both models. ERNIE 5.0 0110 has 20 scored results on Noometry and Muse Spark 1.2 has 31.

Related comparisons

Go deeper