Model comparison

Hunyuan Turbos 20250226 vs Mistral 7B

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 23.0 on the Noometry Index.

Last verified . 15 shared benchmarks.

Hunyuan Turbos 20250226 Tencent

41.3

Rank #139 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 15 benchmarks with published results for both. Hunyuan Turbos 20250226 scores higher in 8 categories and Mistral 7B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Hunyuan Turbos 20250226 leads 37.0 to 7.4.
  • Mistral 7B has downloadable open weights; the other is API-only.

Side by side

Hunyuan Turbos 20250226 and Mistral 7B specifications
Hunyuan Turbos 20250226Mistral 7B
ProviderTencentMistral AI
Noometry Index41.323.0
Released—2023-09-27
WeightsProprietaryOpen
Context window—8K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.25
Results tracked1637

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 40.0 (#152), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Coding13611082
BigCodeBench Instruct—19.5%
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 27.8 (#113), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Hard Prompts13741067
Chess Puzzles—0%
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 37.5 (#154), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Math13591085
OTIS Mock AIME 2024-2025—0.3%
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 37.0 (#161), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Expert13391036
GPQA Diamond—15.2%
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 48.9 (#136), Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Non-English13631012
LMArena Chinese14171009
LMArena French13911037
LMArena German1355987
LMArena Japanese1342878
LMArena Russian13681018
LMArena Korean1351—
LMArena Spanish—1026

Instruction Following Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 71.0 (#158), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Instruction Following13441060

Long Context Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 41.6 (#136), Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Longer Query13661060

Writing & Preference Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 57.4 (#128), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkHunyuan Turbos 20250226Mistral 7B
LMArena Text13771090
LMArena Creative Writing13591068
LMArena Multi-Turn13871062

Frequently asked questions

Is Hunyuan Turbos 20250226 better than Mistral 7B?

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 23.0 on the Noometry Index.

Is Hunyuan Turbos 20250226 or Mistral 7B better for coding?

Hunyuan Turbos 20250226 scores higher on coding benchmarks: 40.0 versus 26.4 in the Noometry coding category.

How many benchmarks do Hunyuan Turbos 20250226 and Mistral 7B share?

15 benchmarks have published results for both models. Hunyuan Turbos 20250226 has 16 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper