Model comparison

ERNIE 5.1 vs Qwen3.5 35B-A3B

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 42.0 on the Noometry Index.

Last verified . 17 shared benchmarks.

ERNIE 5.1 Baidu

43.8

Rank #83 Confirmed

Qwen3.5 35B-A3B Alibaba (Qwen)

42.0

Rank #123 Confirmed

Summary

  • They share 17 benchmarks with published results for both. ERNIE 5.1 scores higher in 6 categories and Qwen3.5 35B-A3B in 2 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in coding, where ERNIE 5.1 leads 44.0 to 33.8.
  • Qwen3.5 35B-A3B has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.1 and Qwen3.5 35B-A3B specifications
ERNIE 5.1Qwen3.5 35B-A3B
ProviderBaiduAlibaba (Qwen)
Noometry Index43.842.0
Released—2026-02-01
WeightsProprietaryOpen
Context window—262K
Max output—66K
Input $ / M tokens—$0.25
Output $ / M tokens—$2
Results tracked1928

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.1 leads

ERNIE 5.1: 44.0 (#76), Qwen3.5 35B-A3B: 33.8 (#251)

Coding benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Coding14881410
LMArena WebDev—1254
SciCode—29.3%

Agentic & Tool Use Not comparable

ERNIE 5.1: —, Qwen3.5 35B-A3B: —

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Search1227—

Reasoning Qwen3.5 35B-A3B leads

ERNIE 5.1: 21.9 (#211), Qwen3.5 35B-A3B: 24.6 (#161)

Reasoning benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Hard Prompts14811400
NYT Connections (extended)23.4%—
CritPt—0.6%
Chess Puzzles—10%
DTBench—80%
LMCA—29.5%
Epoch Capabilities Index—142.52

Math Too close to call

ERNIE 5.1: 40.3 (#92), Qwen3.5 35B-A3B: 39.9 (#97)

Math benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Math14811404
MathArena Final-Answer Competitions—56%
OTIS Mock AIME 2024-2025—70%

Knowledge Qwen3.5 35B-A3B leads

ERNIE 5.1: 41.9 (#102), Qwen3.5 35B-A3B: 47.8 (#79)

Knowledge benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Expert14921408
GPQA Diamond—83.5%
Vectara Hallucination Rate—10.5%

Multilingual ERNIE 5.1 leads

ERNIE 5.1: 55.5 (#29), Qwen3.5 35B-A3B: 50.0 (#127)

Multilingual benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Non-English14541378
LMArena Chinese15081457
LMArena French14881412
LMArena German14701367
LMArena Japanese14221325
LMArena Korean14271356
LMArena Russian14591376
LMArena Spanish14731392

Instruction Following ERNIE 5.1 leads

ERNIE 5.1: 76.7 (#37), Qwen3.5 35B-A3B: 72.8 (#128)

Instruction Following benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Instruction Following14601379

Long Context ERNIE 5.1 leads

ERNIE 5.1: 44.7 (#59), Qwen3.5 35B-A3B: 42.4 (#127)

Long Context benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Longer Query14621389

Writing & Preference ERNIE 5.1 leads

ERNIE 5.1: 65.1 (#52), Qwen3.5 35B-A3B: 57.9 (#124)

Writing & Preference benchmarks
BenchmarkERNIE 5.1Qwen3.5 35B-A3B
LMArena Text14681395
LMArena Creative Writing14411346
LMArena Multi-Turn14711390

Frequently asked questions

Is ERNIE 5.1 better than Qwen3.5 35B-A3B?

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 42.0 on the Noometry Index.

Is ERNIE 5.1 or Qwen3.5 35B-A3B better for coding?

ERNIE 5.1 scores higher on coding benchmarks: 44.0 versus 33.8 in the Noometry coding category.

How many benchmarks do ERNIE 5.1 and Qwen3.5 35B-A3B share?

17 benchmarks have published results for both models. ERNIE 5.1 has 19 scored results on Noometry and Qwen3.5 35B-A3B has 28.

Related comparisons

Go deeper