Model comparison
Falcon-180B vs Hunyuan Large 2025 02 10
Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 32.2 on the Noometry Index.
Last verified . 6 shared benchmarks.
Summary
- They share 6 benchmarks with published results for both. Falcon-180B scores higher in 0 categories and Hunyuan Large 2025 02 10 in 4 categories; 4 gaps are clear of the uncertainty.
- The widest gap is in writing & preference, where Hunyuan Large 2025 02 10 leads 48.7 to 29.1.
- Falcon-180B has downloadable open weights; the other is API-only.
Side by side
| Falcon-180B | Hunyuan Large 2025 02 10 | |
|---|---|---|
| Provider | Technology Innovation Institute | Tencent |
| Noometry Index | 32.2 | 38.6 |
| Released | 2023-09-06 | — |
| Weights | Open | Proprietary |
| Context window | — | — |
| Max output | — | — |
| Input $ / M tokens | — | — |
| Output $ / M tokens | — | — |
| Results tracked | 16 | 12 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Falcon-180B: —, Hunyuan Large 2025 02 10: 38.2 (#181)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Coding | — | 1307 |
Reasoning Hunyuan Large 2025 02 10 leads
Falcon-180B: 19.1 (#269), Hunyuan Large 2025 02 10: 25.5 (#148)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Hard Prompts | 1007 | 1286 |
| Epoch Capabilities Index | 112.13 | — |
| HellaSwag | 89% | — |
| LAMBADA | 79.8% | — |
| PIQA | 84.9% | — |
| WinoGrande | 87.1% | — |
Math Not comparable
Falcon-180B: —, Hunyuan Large 2025 02 10: 35.8 (#178)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Math | — | 1281 |
| GSM8K | 54.4% | — |
Knowledge Not comparable
Falcon-180B: —, Hunyuan Large 2025 02 10: 35.1 (#188)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Expert | — | 1276 |
| ARC (AI2) Challenge | 67.8% | — |
| BoolQ | 89% | — |
| MMLU | 70.6% | — |
| OpenBookQA | 64.2% | — |
Multilingual Hunyuan Large 2025 02 10 leads
Falcon-180B: 25.2 (#286), Hunyuan Large 2025 02 10: 42.0 (#200)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Non-English | 1000 | 1265 |
| LMArena Chinese | — | 1346 |
| LMArena Russian | — | 1266 |
Instruction Following Hunyuan Large 2025 02 10 leads
Falcon-180B: 53.4 (#286), Hunyuan Large 2025 02 10: 67.3 (#197)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Instruction Following | 1047 | 1277 |
Long Context Not comparable
Falcon-180B: —, Hunyuan Large 2025 02 10: 40.8 (#149)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Longer Query | — | 1341 |
Writing & Preference Hunyuan Large 2025 02 10 leads
Falcon-180B: 29.1 (#295), Hunyuan Large 2025 02 10: 48.7 (#197)
| Benchmark | Falcon-180B | Hunyuan Large 2025 02 10 |
|---|---|---|
| LMArena Text | 1054 | 1288 |
| LMArena Creative Writing | 1089 | 1264 |
| LMArena Multi-Turn | 1013 | 1284 |
Frequently asked questions
Is Falcon-180B better than Hunyuan Large 2025 02 10?
Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 32.2 on the Noometry Index.
How many benchmarks do Falcon-180B and Hunyuan Large 2025 02 10 share?
6 benchmarks have published results for both models. Falcon-180B has 16 scored results on Noometry and Hunyuan Large 2025 02 10 has 12.