Model comparison

Hunyuan Large 2025 02 10 vs Mistral 7B

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 23.0 on the Noometry Index.

Last verified . 12 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan Large 2025 02 10 scores higher in 8 categories and Mistral 7B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Hunyuan Large 2025 02 10 leads 35.1 to 7.4.
  • Mistral 7B has downloadable open weights; the other is API-only.

Side by side

Hunyuan Large 2025 02 10 and Mistral 7B specifications
Hunyuan Large 2025 02 10Mistral 7B
ProviderTencentMistral AI
Noometry Index38.623.0
Released—2023-09-27
WeightsProprietaryOpen
Context window—8K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.25
Results tracked1237

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 38.2 (#181), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Coding13071082
BigCodeBench Instruct—19.5%
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 25.5 (#148), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Hard Prompts12861067
Chess Puzzles—0%
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.8 (#178), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Math12811085
OTIS Mock AIME 2024-2025—0.3%
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.1 (#188), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Expert12761036
GPQA Diamond—15.2%
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 42.0 (#200), Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Non-English12651012
LMArena Chinese13461009
LMArena Russian12661018
LMArena French—1037
LMArena German—987
LMArena Japanese—878
LMArena Spanish—1026

Instruction Following Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 67.3 (#197), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Instruction Following12771060

Long Context Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 40.8 (#149), Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Longer Query13411060

Writing & Preference Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 48.7 (#197), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkHunyuan Large 2025 02 10Mistral 7B
LMArena Text12881090
LMArena Creative Writing12641068
LMArena Multi-Turn12841062

Frequently asked questions

Is Hunyuan Large 2025 02 10 better than Mistral 7B?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 23.0 on the Noometry Index.

Is Hunyuan Large 2025 02 10 or Mistral 7B better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 26.4 in the Noometry coding category.

How many benchmarks do Hunyuan Large 2025 02 10 and Mistral 7B share?

12 benchmarks have published results for both models. Hunyuan Large 2025 02 10 has 12 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper