Model comparison

Hunyuan T1 20250711 vs Mistral 7B

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 23.0 on the Noometry Index.

Last verified . 12 shared benchmarks.

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan T1 20250711 scores higher in 8 categories and Mistral 7B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Hunyuan T1 20250711 leads 38.8 to 7.4.
  • Mistral 7B has downloadable open weights; the other is API-only.

Side by side

Hunyuan T1 20250711 and Mistral 7B specifications
Hunyuan T1 20250711Mistral 7B
ProviderTencentMistral AI
Noometry Index42.523.0
Released—2023-09-27
WeightsProprietaryOpen
Context window—8K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.25
Results tracked1337

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 40.9 (#129), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Coding13901082
BigCodeBench Instruct—19.5%
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 28.5 (#103), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Hard Prompts13991067
Chess Puzzles—0%
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 38.7 (#130), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Math14141085
OTIS Mock AIME 2024-2025—0.3%
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 38.8 (#141), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Expert13951036
GPQA Diamond—15.2%
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 51.2 (#112), Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Non-English13951012
LMArena Chinese14251009
LMArena Russian13851018
LMArena French—1037
LMArena German—987
LMArena Japanese—878
LMArena Korean1406—
LMArena Spanish—1026

Instruction Following Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 72.6 (#138), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Instruction Following13741060

Long Context Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 42.2 (#128), Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Longer Query13841060

Writing & Preference Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 59.5 (#109), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkHunyuan T1 20250711Mistral 7B
LMArena Text14011090
LMArena Creative Writing13921068
LMArena Multi-Turn13931062

Frequently asked questions

Is Hunyuan T1 20250711 better than Mistral 7B?

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 23.0 on the Noometry Index.

Is Hunyuan T1 20250711 or Mistral 7B better for coding?

Hunyuan T1 20250711 scores higher on coding benchmarks: 40.9 versus 26.4 in the Noometry coding category.

How many benchmarks do Hunyuan T1 20250711 and Mistral 7B share?

12 benchmarks have published results for both models. Hunyuan T1 20250711 has 13 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper