Model comparison

DeepSeek-V3.1-Terminus vs Hunyuan T1 20250711

DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 score almost the same on the Noometry Index (43.1 vs 42.5), so choose on price, context window or the category you care about most.

Last verified . 10 shared benchmarks.

DeepSeek-V3.1-Terminus DeepSeek

43.1

Rank #97 Confirmed

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Summary

  • They share 10 benchmarks with published results for both. DeepSeek-V3.1-Terminus scores higher in 5 categories and Hunyuan T1 20250711 in 2 categories; 5 gaps are clear of the uncertainty.
  • DeepSeek-V3.1-Terminus has downloadable open weights; the other is API-only.

Side by side

DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 specifications
DeepSeek-V3.1-TerminusHunyuan T1 20250711
ProviderDeepSeekTencent
Noometry Index43.142.5
Released2025-09-22—
WeightsOpenProprietary
Context window164K—
Max output147K—
Input $ / M tokens$0.27—
Output $ / M tokens$1—
Results tracked1613

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek-V3.1-Terminus leads

DeepSeek-V3.1-Terminus: 42.0 (#113), Hunyuan T1 20250711: 40.9 (#129)

Coding benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Coding14261390
SciCode40.6%—
ALE-Bench745.17—

Reasoning Hunyuan T1 20250711 leads

DeepSeek-V3.1-Terminus: 26.4 (#133), Hunyuan T1 20250711: 28.5 (#103)

Reasoning benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Hard Prompts14261399
Kagi LLM Benchmark57.4%—
CritPt1.7%—
DTBench81.3%—
LMCA28.6%—

Math Too close to call

DeepSeek-V3.1-Terminus: 38.5 (#137), Hunyuan T1 20250711: 38.7 (#130)

Math benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Math14021414

Knowledge Not comparable

DeepSeek-V3.1-Terminus: —, Hunyuan T1 20250711: 38.8 (#141)

Knowledge benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Expert—1395

Multilingual Too close to call

DeepSeek-V3.1-Terminus: 52.1 (#92), Hunyuan T1 20250711: 51.2 (#112)

Multilingual benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Non-English14071395
LMArena Russian14361385
LMArena Chinese—1425
LMArena Korean—1406

Instruction Following DeepSeek-V3.1-Terminus leads

DeepSeek-V3.1-Terminus: 74.0 (#106), Hunyuan T1 20250711: 72.6 (#138)

Instruction Following benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Instruction Following14041374

Long Context DeepSeek-V3.1-Terminus leads

DeepSeek-V3.1-Terminus: 43.4 (#97), Hunyuan T1 20250711: 42.2 (#128)

Long Context benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Longer Query14211384

Writing & Preference DeepSeek-V3.1-Terminus leads

DeepSeek-V3.1-Terminus: 61.0 (#92), Hunyuan T1 20250711: 59.5 (#109)

Writing & Preference benchmarks
BenchmarkDeepSeek-V3.1-TerminusHunyuan T1 20250711
LMArena Text14191401
LMArena Creative Writing14031392
LMArena Multi-Turn14111393

Frequently asked questions

Is DeepSeek-V3.1-Terminus better than Hunyuan T1 20250711?

DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 score almost the same on the Noometry Index (43.1 vs 42.5), so choose on price, context window or the category you care about most.

Is DeepSeek-V3.1-Terminus or Hunyuan T1 20250711 better for coding?

DeepSeek-V3.1-Terminus scores higher on coding benchmarks: 42.0 versus 40.9 in the Noometry coding category.

How many benchmarks do DeepSeek-V3.1-Terminus and Hunyuan T1 20250711 share?

10 benchmarks have published results for both models. DeepSeek-V3.1-Terminus has 16 scored results on Noometry and Hunyuan T1 20250711 has 13.

Related comparisons

Go deeper