Model comparison

Falcon-180B vs Hunyuan Large 2025 02 10

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 32.2 on the Noometry Index.

Last verified . 6 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Summary

  • They share 6 benchmarks with published results for both. Falcon-180B scores higher in 0 categories and Hunyuan Large 2025 02 10 in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan Large 2025 02 10 leads 48.7 to 29.1.
  • Falcon-180B has downloadable open weights; the other is API-only.

Side by side

Falcon-180B and Hunyuan Large 2025 02 10 specifications
Falcon-180BHunyuan Large 2025 02 10
ProviderTechnology Innovation InstituteTencent
Noometry Index32.238.6
Released2023-09-06—
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1612

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Falcon-180B: —, Hunyuan Large 2025 02 10: 38.2 (#181)

Coding benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Coding—1307

Reasoning Hunyuan Large 2025 02 10 leads

Falcon-180B: 19.1 (#269), Hunyuan Large 2025 02 10: 25.5 (#148)

Reasoning benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Hard Prompts10071286
Epoch Capabilities Index112.13—
HellaSwag89%—
LAMBADA79.8%—
PIQA84.9%—
WinoGrande87.1%—

Math Not comparable

Falcon-180B: —, Hunyuan Large 2025 02 10: 35.8 (#178)

Math benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Math—1281
GSM8K54.4%—

Knowledge Not comparable

Falcon-180B: —, Hunyuan Large 2025 02 10: 35.1 (#188)

Knowledge benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Expert—1276
ARC (AI2) Challenge67.8%—
BoolQ89%—
MMLU70.6%—
OpenBookQA64.2%—

Multilingual Hunyuan Large 2025 02 10 leads

Falcon-180B: 25.2 (#286), Hunyuan Large 2025 02 10: 42.0 (#200)

Multilingual benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Non-English10001265
LMArena Chinese—1346
LMArena Russian—1266

Instruction Following Hunyuan Large 2025 02 10 leads

Falcon-180B: 53.4 (#286), Hunyuan Large 2025 02 10: 67.3 (#197)

Instruction Following benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Instruction Following10471277

Long Context Not comparable

Falcon-180B: —, Hunyuan Large 2025 02 10: 40.8 (#149)

Long Context benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Longer Query—1341

Writing & Preference Hunyuan Large 2025 02 10 leads

Falcon-180B: 29.1 (#295), Hunyuan Large 2025 02 10: 48.7 (#197)

Writing & Preference benchmarks
BenchmarkFalcon-180BHunyuan Large 2025 02 10
LMArena Text10541288
LMArena Creative Writing10891264
LMArena Multi-Turn10131284

Frequently asked questions

Is Falcon-180B better than Hunyuan Large 2025 02 10?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 32.2 on the Noometry Index.

How many benchmarks do Falcon-180B and Hunyuan Large 2025 02 10 share?

6 benchmarks have published results for both models. Falcon-180B has 16 scored results on Noometry and Hunyuan Large 2025 02 10 has 12.

Related comparisons

Go deeper