Model comparison

Granite 4.1 8b vs Hunyuan Large 2025 02 10

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 37.4 on the Noometry Index.

Last verified . 12 shared benchmarks.

Granite 4.1 8b IBM

37.4

Rank #205 Confirmed

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Granite 4.1 8b scores higher in 3 categories and Hunyuan Large 2025 02 10 in 5 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Hunyuan Large 2025 02 10 leads 38.2 to 30.1.
  • Granite 4.1 8b has downloadable open weights; the other is API-only.

Side by side

Granite 4.1 8b and Hunyuan Large 2025 02 10 specifications
Granite 4.1 8bHunyuan Large 2025 02 10
ProviderIBMTencent
Noometry Index37.438.6
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1312

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

Granite 4.1 8b: 30.1 (#297), Hunyuan Large 2025 02 10: 38.2 (#181)

Coding benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Coding13121307
LMArena WebDev1192—

Reasoning Too close to call

Granite 4.1 8b: 25.7 (#143), Hunyuan Large 2025 02 10: 25.5 (#148)

Reasoning benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Hard Prompts12931286

Math Too close to call

Granite 4.1 8b: 36.4 (#166), Hunyuan Large 2025 02 10: 35.8 (#178)

Math benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Math13121281

Knowledge Granite 4.1 8b leads

Granite 4.1 8b: 36.1 (#174), Hunyuan Large 2025 02 10: 35.1 (#188)

Knowledge benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Expert13091276

Multilingual Too close to call

Granite 4.1 8b: 41.7 (#204), Hunyuan Large 2025 02 10: 42.0 (#200)

Multilingual benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Non-English12611265
LMArena Chinese13371346
LMArena Russian12401266

Instruction Following Too close to call

Granite 4.1 8b: 66.9 (#203), Hunyuan Large 2025 02 10: 67.3 (#197)

Instruction Following benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Instruction Following12691277

Long Context Hunyuan Large 2025 02 10 leads

Granite 4.1 8b: 38.7 (#193), Hunyuan Large 2025 02 10: 40.8 (#149)

Long Context benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Longer Query12751341

Writing & Preference Too close to call

Granite 4.1 8b: 48.1 (#204), Hunyuan Large 2025 02 10: 48.7 (#197)

Writing & Preference benchmarks
BenchmarkGranite 4.1 8bHunyuan Large 2025 02 10
LMArena Text12901288
LMArena Creative Writing12511264
LMArena Multi-Turn12681284

Frequently asked questions

Is Granite 4.1 8b better than Hunyuan Large 2025 02 10?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 37.4 on the Noometry Index.

Is Granite 4.1 8b or Hunyuan Large 2025 02 10 better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 30.1 in the Noometry coding category.

How many benchmarks do Granite 4.1 8b and Hunyuan Large 2025 02 10 share?

12 benchmarks have published results for both models. Granite 4.1 8b has 13 scored results on Noometry and Hunyuan Large 2025 02 10 has 12.

Related comparisons

Go deeper