Model comparison

ERNIE 5.1 vs Mistral 7B

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 23.0 on the Noometry Index.

Last verified . 16 shared benchmarks.

ERNIE 5.1 Baidu

43.8

Rank #83 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 16 benchmarks with published results for both. ERNIE 5.1 scores higher in 8 categories and Mistral 7B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where ERNIE 5.1 leads 41.9 to 7.4.
  • Mistral 7B has downloadable open weights; the other is API-only.

Side by side

ERNIE 5.1 and Mistral 7B specifications
ERNIE 5.1Mistral 7B
ProviderBaiduMistral AI
Noometry Index43.823.0
Released—2023-09-27
WeightsProprietaryOpen
Context window—8K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.25
Results tracked1937

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding ERNIE 5.1 leads

ERNIE 5.1: 44.0 (#76), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Coding14881082
BigCodeBench Instruct—19.5%
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Agentic & Tool Use Not comparable

ERNIE 5.1: —, Mistral 7B: —

Agentic & Tool Use benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Search1227—

Reasoning ERNIE 5.1 leads

ERNIE 5.1: 21.9 (#211), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Hard Prompts14811067
NYT Connections (extended)23.4%—
Chess Puzzles—0%
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math ERNIE 5.1 leads

ERNIE 5.1: 40.3 (#92), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Math14811085
OTIS Mock AIME 2024-2025—0.3%
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge ERNIE 5.1 leads

ERNIE 5.1: 41.9 (#102), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Expert14921036
GPQA Diamond—15.2%
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual ERNIE 5.1 leads

ERNIE 5.1: 55.5 (#29), Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Non-English14541012
LMArena Chinese15081009
LMArena French14881037
LMArena German1470987
LMArena Japanese1422878
LMArena Russian14591018
LMArena Spanish14731026
LMArena Korean1427—

Instruction Following ERNIE 5.1 leads

ERNIE 5.1: 76.7 (#37), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Instruction Following14601060

Long Context ERNIE 5.1 leads

ERNIE 5.1: 44.7 (#59), Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Longer Query14621060

Writing & Preference ERNIE 5.1 leads

ERNIE 5.1: 65.1 (#52), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkERNIE 5.1Mistral 7B
LMArena Text14681090
LMArena Creative Writing14411068
LMArena Multi-Turn14711062

Frequently asked questions

Is ERNIE 5.1 better than Mistral 7B?

ERNIE 5.1 is the stronger model overall, scoring 43.8 to 23.0 on the Noometry Index.

Is ERNIE 5.1 or Mistral 7B better for coding?

ERNIE 5.1 scores higher on coding benchmarks: 44.0 versus 26.4 in the Noometry coding category.

How many benchmarks do ERNIE 5.1 and Mistral 7B share?

16 benchmarks have published results for both models. ERNIE 5.1 has 19 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper