Model comparison

DeepSeek-V3.1 vs Hunyuan T1 20250711

DeepSeek-V3.1 and Hunyuan T1 20250711 score almost the same on the Noometry Index (42.8 vs 42.5), so choose on price, context window or the category you care about most.

Last verified . 13 shared benchmarks.

DeepSeek-V3.1 DeepSeek

42.8

Rank #108 Confirmed

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Summary

  • They share 13 benchmarks with published results for both. DeepSeek-V3.1 scores higher in 5 categories and Hunyuan T1 20250711 in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Hunyuan T1 20250711 leads 42.2 to 36.3.
  • DeepSeek-V3.1 has downloadable open weights; the other is API-only.

Side by side

DeepSeek-V3.1 and Hunyuan T1 20250711 specifications
DeepSeek-V3.1Hunyuan T1 20250711
ProviderDeepSeekTencent
Noometry Index42.842.5
Released2025-08-21—
WeightsOpenProprietary
Context window164K—
Max output8K—
Input $ / M tokens$0.25—
Output $ / M tokens$0.95—
Results tracked2713

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

DeepSeek-V3.1: 40.3 (#144), Hunyuan T1 20250711: 40.9 (#129)

Coding benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Coding14171390
WeirdML38.4%—

Reasoning Too close to call

DeepSeek-V3.1: 27.9 (#110), Hunyuan T1 20250711: 28.5 (#103)

Reasoning benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Hard Prompts14171399
SimpleBench40%—
Kagi LLM Benchmark53.2%—
DTBench82.7%—
LMCA24.3%—
Epoch Capabilities Index139.92—
ForecastBench58—

Math Too close to call

DeepSeek-V3.1: 38.9 (#122), Hunyuan T1 20250711: 38.7 (#130)

Math benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Math14201414

Knowledge DeepSeek-V3.1 leads

DeepSeek-V3.1: 43.7 (#90), Hunyuan T1 20250711: 38.8 (#141)

Knowledge benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Expert14051395
Vectara Hallucination Rate5.5%—

Multilingual Too close to call

DeepSeek-V3.1: 51.6 (#106), Hunyuan T1 20250711: 51.2 (#112)

Multilingual benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Non-English14001395
LMArena Chinese14691425
LMArena Korean13371406
LMArena Russian14051385
LMArena French1447—
LMArena German1411—
LMArena Japanese1378—
LMArena Spanish1431—

Instruction Following DeepSeek-V3.1 leads

DeepSeek-V3.1: 73.9 (#110), Hunyuan T1 20250711: 72.6 (#138)

Instruction Following benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Instruction Following14001374

Long Context Hunyuan T1 20250711 leads

DeepSeek-V3.1: 36.3 (#232), Hunyuan T1 20250711: 42.2 (#128)

Long Context benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Longer Query14221384
Fiction.LiveBench52.8%—

Writing & Preference Too close to call

DeepSeek-V3.1: 60.3 (#98), Hunyuan T1 20250711: 59.5 (#109)

Writing & Preference benchmarks
BenchmarkDeepSeek-V3.1Hunyuan T1 20250711
LMArena Text14201401
LMArena Creative Writing14011392
LMArena Multi-Turn14081393
EQ-Bench Creative Writing1436—

Frequently asked questions

Is DeepSeek-V3.1 better than Hunyuan T1 20250711?

DeepSeek-V3.1 and Hunyuan T1 20250711 score almost the same on the Noometry Index (42.8 vs 42.5), so choose on price, context window or the category you care about most.

Is DeepSeek-V3.1 or Hunyuan T1 20250711 better for coding?

They score almost the same on coding (40.3 vs 40.9); test both on your own repository before choosing.

How many benchmarks do DeepSeek-V3.1 and Hunyuan T1 20250711 share?

13 benchmarks have published results for both models. DeepSeek-V3.1 has 27 scored results on Noometry and Hunyuan T1 20250711 has 13.

Related comparisons

Go deeper