Model comparison

Hunyuan Large 2025 02 10 vs Mixtral 8x22B

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 27.1 on the Noometry Index.

Last verified . 12 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Mixtral 8x22B Mistral AI

27.1

Rank #333 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan Large 2025 02 10 scores higher in 8 categories and Mixtral 8x22B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Hunyuan Large 2025 02 10 leads 35.1 to 15.1.
  • Mixtral 8x22B has downloadable open weights; the other is API-only.

Side by side

Hunyuan Large 2025 02 10 and Mixtral 8x22B specifications
Hunyuan Large 2025 02 10Mixtral 8x22B
ProviderTencentMistral AI
Noometry Index38.627.1
Released—2024-04-17
WeightsProprietaryOpen
Context window—64K
Max output—64K
Input $ / M tokens—$2
Output $ / M tokens—$6
Results tracked1234

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 38.2 (#181), Mixtral 8x22B: 24.2 (#329)

Coding benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Coding13071166
WeirdML—3.2%
BigCodeBench Instruct—40.6%
BigCodeBench Complete—50.2%
HumanEval+—72%
MBPP+—64.3%

Agentic & Tool Use Not comparable

Hunyuan Large 2025 02 10: —, Mixtral 8x22B: 23.1 (#127)

Agentic & Tool Use benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
Cybench—7.5%

Reasoning Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 25.5 (#148), Mixtral 8x22B: 19.9 (#248)

Reasoning benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Hard Prompts12861150
DTBench—55.1%
Epoch Capabilities Index—122.03
ForecastBench—56.3

Math Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.8 (#178), Mixtral 8x22B: 22.9 (#275)

Math benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Math12811184
Omni-MATH—16.3%
MATH Level 5—24.2%

Knowledge Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.1 (#188), Mixtral 8x22B: 15.1 (#293)

Knowledge benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Expert12761113
GPQA Diamond—34.1%
MMLU-Pro—46%
GPQA (HELM)—33.4%
MMLU—77.8%

Multilingual Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 42.0 (#200), Mixtral 8x22B: 32.8 (#255)

Multilingual benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Non-English12651128
LMArena Chinese13461116
LMArena Russian12661158
LMArena French—1166
LMArena German—1141
LMArena Japanese—1037
LMArena Korean—1057
LMArena Spanish—1151

Instruction Following Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 67.3 (#197), Mixtral 8x22B: 57.7 (#266)

Instruction Following benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Instruction Following12771147
IFEval—72.4%

Long Context Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 40.8 (#149), Mixtral 8x22B: 34.7 (#247)

Long Context benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Longer Query13411144

Writing & Preference Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 48.7 (#197), Mixtral 8x22B: 36.9 (#262)

Writing & Preference benchmarks
BenchmarkHunyuan Large 2025 02 10Mixtral 8x22B
LMArena Text12881162
LMArena Creative Writing12641141
LMArena Multi-Turn12841130
WildBench—71.1%

Frequently asked questions

Is Hunyuan Large 2025 02 10 better than Mixtral 8x22B?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 27.1 on the Noometry Index.

Is Hunyuan Large 2025 02 10 or Mixtral 8x22B better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 24.2 in the Noometry coding category.

How many benchmarks do Hunyuan Large 2025 02 10 and Mixtral 8x22B share?

12 benchmarks have published results for both models. Hunyuan Large 2025 02 10 has 12 scored results on Noometry and Mixtral 8x22B has 34.

Related comparisons

Go deeper