Model comparison

DeepSeek LLM 67B vs Hunyuan Large 2025 02 10

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 24.9 on the Noometry Index.

Last verified . 10 shared benchmarks.

DeepSeek LLM 67B DeepSeek

24.9

Rank #347 Confirmed

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Summary

  • They share 10 benchmarks with published results for both. DeepSeek LLM 67B scores higher in 0 categories and Hunyuan Large 2025 02 10 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Hunyuan Large 2025 02 10 leads 35.1 to 7.0.
  • DeepSeek LLM 67B has downloadable open weights; the other is API-only.

Side by side

DeepSeek LLM 67B and Hunyuan Large 2025 02 10 specifications
DeepSeek LLM 67BHunyuan Large 2025 02 10
ProviderDeepSeekTencent
Noometry Index24.938.6
Released2023-11-29—
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1512

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 31.9 (#278), Hunyuan Large 2025 02 10: 38.2 (#181)

Coding benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Coding10961307

Reasoning Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 16.5 (#304), Hunyuan Large 2025 02 10: 25.5 (#148)

Reasoning benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Hard Prompts10701286
Chess Puzzles0%—
Epoch Capabilities Index110.5—

Math Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 8.7 (#324), Hunyuan Large 2025 02 10: 35.8 (#178)

Math benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Math11081281
OTIS Mock AIME 2024-20250.8%—
MATH Level 56.4%—

Knowledge Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 7.0 (#313), Hunyuan Large 2025 02 10: 35.1 (#188)

Knowledge benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
GPQA Diamond24.6%—
LMArena Expert—1276

Multilingual Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 29.4 (#267), Hunyuan Large 2025 02 10: 42.0 (#200)

Multilingual benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Non-English10731265
LMArena Chinese11321346
LMArena Russian—1266

Instruction Following Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 55.4 (#277), Hunyuan Large 2025 02 10: 67.3 (#197)

Instruction Following benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Instruction Following10791277

Long Context Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 33.1 (#265), Hunyuan Large 2025 02 10: 40.8 (#149)

Long Context benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Longer Query10921341

Writing & Preference Hunyuan Large 2025 02 10 leads

DeepSeek LLM 67B: 31.6 (#282), Hunyuan Large 2025 02 10: 48.7 (#197)

Writing & Preference benchmarks
BenchmarkDeepSeek LLM 67BHunyuan Large 2025 02 10
LMArena Text11051288
LMArena Creative Writing10671264
LMArena Multi-Turn10821284

Frequently asked questions

Is DeepSeek LLM 67B better than Hunyuan Large 2025 02 10?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 24.9 on the Noometry Index.

Is DeepSeek LLM 67B or Hunyuan Large 2025 02 10 better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 31.9 in the Noometry coding category.

How many benchmarks do DeepSeek LLM 67B and Hunyuan Large 2025 02 10 share?

10 benchmarks have published results for both models. DeepSeek LLM 67B has 15 scored results on Noometry and Hunyuan Large 2025 02 10 has 12.

Related comparisons

Go deeper