Model comparison

Granite 3.1 8b Instruct vs Hunyuan T1 20250711

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 32.4 on the Noometry Index.

Last verified . 12 shared benchmarks.

Granite 3.1 8b Instruct IBM

32.4

Rank #258 Confirmed

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Granite 3.1 8b Instruct scores higher in 0 categories and Hunyuan T1 20250711 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan T1 20250711 leads 59.5 to 35.5.
  • Granite 3.1 8b Instruct has downloadable open weights; the other is API-only.

Side by side

Granite 3.1 8b Instruct and Hunyuan T1 20250711 specifications
Granite 3.1 8b InstructHunyuan T1 20250711
ProviderIBMTencent
Noometry Index32.442.5
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1313

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 34.5 (#233), Hunyuan T1 20250711: 40.9 (#129)

Coding benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Coding11861390

Agentic & Tool Use Not comparable

Granite 3.1 8b Instruct: 24.1 (#120), Hunyuan T1 20250711: —

Agentic & Tool Use benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
Berkeley Function Calling Leaderboard27.1%—

Reasoning Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 22.1 (#207), Hunyuan T1 20250711: 28.5 (#103)

Reasoning benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Hard Prompts11451399

Math Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 33.0 (#209), Hunyuan T1 20250711: 38.7 (#130)

Math benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Math11521414

Knowledge Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 31.1 (#220), Hunyuan T1 20250711: 38.8 (#141)

Knowledge benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Expert11421395

Multilingual Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 30.9 (#260), Hunyuan T1 20250711: 51.2 (#112)

Multilingual benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Non-English10991395
LMArena Chinese11451425
LMArena Russian10921385
LMArena Korean—1406

Instruction Following Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 58.6 (#259), Hunyuan T1 20250711: 72.6 (#138)

Instruction Following benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Instruction Following11311374

Long Context Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 35.2 (#241), Hunyuan T1 20250711: 42.2 (#128)

Long Context benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Longer Query11621384

Writing & Preference Hunyuan T1 20250711 leads

Granite 3.1 8b Instruct: 35.5 (#266), Hunyuan T1 20250711: 59.5 (#109)

Writing & Preference benchmarks
BenchmarkGranite 3.1 8b InstructHunyuan T1 20250711
LMArena Text11501401
LMArena Creative Writing11291392
LMArena Multi-Turn11081393

Frequently asked questions

Is Granite 3.1 8b Instruct better than Hunyuan T1 20250711?

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 32.4 on the Noometry Index.

Is Granite 3.1 8b Instruct or Hunyuan T1 20250711 better for coding?

Hunyuan T1 20250711 scores higher on coding benchmarks: 40.9 versus 34.5 in the Noometry coding category.

How many benchmarks do Granite 3.1 8b Instruct and Hunyuan T1 20250711 share?

12 benchmarks have published results for both models. Granite 3.1 8b Instruct has 13 scored results on Noometry and Hunyuan T1 20250711 has 13.

Related comparisons

Go deeper