Model comparison

Granite 3.1 2b Instruct vs Hunyuan Turbos 20250226

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 33.2 on the Noometry Index.

Last verified . 12 shared benchmarks.

Granite 3.1 2b Instruct IBM

33.2

Rank #247 Confirmed

Hunyuan Turbos 20250226 Tencent

41.3

Rank #139 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Granite 3.1 2b Instruct scores higher in 0 categories and Hunyuan Turbos 20250226 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan Turbos 20250226 leads 57.4 to 34.1.
  • Granite 3.1 2b Instruct has downloadable open weights; the other is API-only.

Side by side

Granite 3.1 2b Instruct and Hunyuan Turbos 20250226 specifications
Granite 3.1 2b InstructHunyuan Turbos 20250226
ProviderIBMTencent
Noometry Index33.241.3
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1216

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 33.4 (#257), Hunyuan Turbos 20250226: 40.0 (#152)

Coding benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Coding11491361

Reasoning Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 22.0 (#209), Hunyuan Turbos 20250226: 27.8 (#113)

Reasoning benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Hard Prompts11381374

Math Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 33.1 (#206), Hunyuan Turbos 20250226: 37.5 (#154)

Math benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Math11591359

Knowledge Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 30.8 (#224), Hunyuan Turbos 20250226: 37.0 (#161)

Knowledge benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Expert11311339

Multilingual Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 29.1 (#269), Hunyuan Turbos 20250226: 48.9 (#136)

Multilingual benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Non-English10681363
LMArena Chinese11391417
LMArena Russian10631368
LMArena French—1391
LMArena German—1355
LMArena Japanese—1342
LMArena Korean—1351

Instruction Following Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 57.7 (#264), Hunyuan Turbos 20250226: 71.0 (#158)

Instruction Following benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Instruction Following11161344

Long Context Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 35.0 (#244), Hunyuan Turbos 20250226: 41.6 (#136)

Long Context benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Longer Query11551366

Writing & Preference Hunyuan Turbos 20250226 leads

Granite 3.1 2b Instruct: 34.1 (#274), Hunyuan Turbos 20250226: 57.4 (#128)

Writing & Preference benchmarks
BenchmarkGranite 3.1 2b InstructHunyuan Turbos 20250226
LMArena Text11271377
LMArena Creative Writing11161359
LMArena Multi-Turn10991387

Frequently asked questions

Is Granite 3.1 2b Instruct better than Hunyuan Turbos 20250226?

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 33.2 on the Noometry Index.

Is Granite 3.1 2b Instruct or Hunyuan Turbos 20250226 better for coding?

Hunyuan Turbos 20250226 scores higher on coding benchmarks: 40.0 versus 33.4 in the Noometry coding category.

How many benchmarks do Granite 3.1 2b Instruct and Hunyuan Turbos 20250226 share?

12 benchmarks have published results for both models. Granite 3.1 2b Instruct has 12 scored results on Noometry and Hunyuan Turbos 20250226 has 16.

Related comparisons

Go deeper