Model comparison

Hunyuan Large 2025 02 10 vs Inkling

Inkling is the stronger model overall, scoring 44.1 to 38.6 on the Noometry Index.

Last verified . 12 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Inkling Thinking Machines Lab

44.1

Rank #80 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan Large 2025 02 10 scores higher in 2 categories and Inkling in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Inkling leads 55.1 to 35.1.
  • Inkling has downloadable open weights; the other is API-only.

Side by side

Hunyuan Large 2025 02 10 and Inkling specifications
Hunyuan Large 2025 02 10Inkling
ProviderTencentThinking Machines Lab
Noometry Index38.644.1
Released—2026-07-15
WeightsProprietaryOpen
Context window—66K
Max output—66K
Input $ / M tokens—$1.87
Output $ / M tokens—$4.68
Results tracked1241

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 38.2 (#181), Inkling: 34.5 (#234)

Coding benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Coding13071464
FrontierCode—14%
LMArena WebDev—1413
FrontierSWE—4.1%
SciCode—47%
WeirdML—32.3%
ALE-Bench—946

Agentic & Tool Use Not comparable

Hunyuan Large 2025 02 10: —, Inkling: 29.6 (#85)

Agentic & Tool Use benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
APEX-Agents—33.8%
τ²-bench Banking—25%

Reasoning Inkling leads

Hunyuan Large 2025 02 10: 25.5 (#148), Inkling: 40.4 (#56)

Reasoning benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Hard Prompts12861451
ARC-AGI-2—36.5%
SimpleBench—50%
ARC-AGI-1—79.5%
CritPt—5.4%
Chess Puzzles—21%
DTBench—87.5%
LMCA—37.6%
Epoch Capabilities Index—148.54

Math Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.8 (#178), Inkling: 31.3 (#225)

Math benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Math12811479
FrontierMath (Tiers 1-3)—33.3%
FrontierMath Tier 4—4.9%
OTIS Mock AIME 2024-2025—88.9%
ProofBench—0%

Knowledge Inkling leads

Hunyuan Large 2025 02 10: 35.1 (#188), Inkling: 55.1 (#49)

Knowledge benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Expert12761465
GPQA Diamond—88.3%
SimpleQA Verified—40.3%

Multilingual Inkling leads

Hunyuan Large 2025 02 10: 42.0 (#200), Inkling: 54.0 (#52)

Multilingual benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Non-English12651434
LMArena Chinese13461490
LMArena Russian12661429
LMArena French—1458
LMArena German—1446
LMArena Japanese—1429
LMArena Korean—1404
LMArena Spanish—1448

Instruction Following Inkling leads

Hunyuan Large 2025 02 10: 67.3 (#197), Inkling: 75.1 (#71)

Instruction Following benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Instruction Following12771426

Long Context Inkling leads

Hunyuan Large 2025 02 10: 40.8 (#149), Inkling: 43.8 (#86)

Long Context benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Longer Query13411434

Writing & Preference Inkling leads

Hunyuan Large 2025 02 10: 48.7 (#197), Inkling: 65.2 (#51)

Writing & Preference benchmarks
BenchmarkHunyuan Large 2025 02 10Inkling
LMArena Text12881441
LMArena Creative Writing12641387
LMArena Multi-Turn12841436
EQ-Bench Creative Writing—1611
EQ-Bench 4—1226

Frequently asked questions

Is Hunyuan Large 2025 02 10 better than Inkling?

Inkling is the stronger model overall, scoring 44.1 to 38.6 on the Noometry Index.

Is Hunyuan Large 2025 02 10 or Inkling better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 34.5 in the Noometry coding category.

How many benchmarks do Hunyuan Large 2025 02 10 and Inkling share?

12 benchmarks have published results for both models. Hunyuan Large 2025 02 10 has 12 scored results on Noometry and Inkling has 41.

Related comparisons

Go deeper