Model comparison
DeepSeek-V3.1-Terminus vs Hunyuan T1 20250711
DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 score almost the same on the Noometry Index (43.1 vs 42.5), so choose on price, context window or the category you care about most.
Last verified . 10 shared benchmarks.
Summary
- They share 10 benchmarks with published results for both. DeepSeek-V3.1-Terminus scores higher in 5 categories and Hunyuan T1 20250711 in 2 categories; 5 gaps are clear of the uncertainty.
- DeepSeek-V3.1-Terminus has downloadable open weights; the other is API-only.
Side by side
| DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 | |
|---|---|---|
| Provider | DeepSeek | Tencent |
| Noometry Index | 43.1 | 42.5 |
| Released | 2025-09-22 | — |
| Weights | Open | Proprietary |
| Context window | 164K | — |
| Max output | 147K | — |
| Input $ / M tokens | $0.27 | — |
| Output $ / M tokens | $1 | — |
| Results tracked | 16 | 13 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding DeepSeek-V3.1-Terminus leads
DeepSeek-V3.1-Terminus: 42.0 (#113), Hunyuan T1 20250711: 40.9 (#129)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Coding | 1426 | 1390 |
| SciCode | 40.6% | — |
| ALE-Bench | 745.17 | — |
Reasoning Hunyuan T1 20250711 leads
DeepSeek-V3.1-Terminus: 26.4 (#133), Hunyuan T1 20250711: 28.5 (#103)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Hard Prompts | 1426 | 1399 |
| Kagi LLM Benchmark | 57.4% | — |
| CritPt | 1.7% | — |
| DTBench | 81.3% | — |
| LMCA | 28.6% | — |
Math Too close to call
DeepSeek-V3.1-Terminus: 38.5 (#137), Hunyuan T1 20250711: 38.7 (#130)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Math | 1402 | 1414 |
Knowledge Not comparable
DeepSeek-V3.1-Terminus: —, Hunyuan T1 20250711: 38.8 (#141)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Expert | — | 1395 |
Multilingual Too close to call
DeepSeek-V3.1-Terminus: 52.1 (#92), Hunyuan T1 20250711: 51.2 (#112)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Non-English | 1407 | 1395 |
| LMArena Russian | 1436 | 1385 |
| LMArena Chinese | — | 1425 |
| LMArena Korean | — | 1406 |
Instruction Following DeepSeek-V3.1-Terminus leads
DeepSeek-V3.1-Terminus: 74.0 (#106), Hunyuan T1 20250711: 72.6 (#138)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Instruction Following | 1404 | 1374 |
Long Context DeepSeek-V3.1-Terminus leads
DeepSeek-V3.1-Terminus: 43.4 (#97), Hunyuan T1 20250711: 42.2 (#128)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Longer Query | 1421 | 1384 |
Writing & Preference DeepSeek-V3.1-Terminus leads
DeepSeek-V3.1-Terminus: 61.0 (#92), Hunyuan T1 20250711: 59.5 (#109)
| Benchmark | DeepSeek-V3.1-Terminus | Hunyuan T1 20250711 |
|---|---|---|
| LMArena Text | 1419 | 1401 |
| LMArena Creative Writing | 1403 | 1392 |
| LMArena Multi-Turn | 1411 | 1393 |
Frequently asked questions
Is DeepSeek-V3.1-Terminus better than Hunyuan T1 20250711?
DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 score almost the same on the Noometry Index (43.1 vs 42.5), so choose on price, context window or the category you care about most.
Is DeepSeek-V3.1-Terminus or Hunyuan T1 20250711 better for coding?
DeepSeek-V3.1-Terminus scores higher on coding benchmarks: 42.0 versus 40.9 in the Noometry coding category.
How many benchmarks do DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 share?
10 benchmarks have published results for both models. DeepSeek-V3.1-Terminus has 16 scored results on Noometry and Hunyuan T1 20250711 has 13.