Model comparison

Granite 4.2 30b vs Step 1o Turbo 202506

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 39.7 on the Noometry Index.

Last verified . 11 shared benchmarks.

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Step 1o Turbo 202506 StepFun

39.7

Rank #160 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Granite 4.2 30b scores higher in 7 categories and Step 1o Turbo 202506 in 0 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Granite 4.2 30b leads 39.1 to 36.1.
  • Granite 4.2 30b has downloadable open weights; the other is API-only.

Side by side

Granite 4.2 30b and Step 1o Turbo 202506 specifications
Granite 4.2 30bStep 1o Turbo 202506
ProviderIBMStepFun
Noometry Index41.839.7
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1114

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Granite 4.2 30b: 41.0 (#126), Step 1o Turbo 202506: 39.2 (#160)

Coding benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Coding13961339

Reasoning Granite 4.2 30b leads

Granite 4.2 30b: 27.8 (#112), Step 1o Turbo 202506: 26.8 (#129)

Reasoning benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Hard Prompts13741335

Math Not comparable

Granite 4.2 30b: —, Step 1o Turbo 202506: 36.6 (#164)

Math benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Math—1318

Knowledge Granite 4.2 30b leads

Granite 4.2 30b: 39.1 (#138), Step 1o Turbo 202506: 36.1 (#176)

Knowledge benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Expert14061308

Multimodal Not comparable

Granite 4.2 30b: —, Step 1o Turbo 202506: 36.1 (#80)

Multimodal benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Vision—1186

Multilingual Granite 4.2 30b leads

Granite 4.2 30b: 47.3 (#151), Step 1o Turbo 202506: 45.3 (#173)

Multilingual benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Non-English13401313
LMArena Chinese14141380
LMArena Russian13431327
LMArena German—1314

Instruction Following Granite 4.2 30b leads

Granite 4.2 30b: 71.2 (#155), Step 1o Turbo 202506: 69.2 (#176)

Instruction Following benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Instruction Following13471310

Long Context Too close to call

Granite 4.2 30b: 41.4 (#140), Step 1o Turbo 202506: 41.0 (#146)

Long Context benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Longer Query13591348

Writing & Preference Too close to call

Granite 4.2 30b: 53.8 (#156), Step 1o Turbo 202506: 53.1 (#160)

Writing & Preference benchmarks
BenchmarkGranite 4.2 30bStep 1o Turbo 202506
LMArena Text13611336
LMArena Creative Writing12881307
LMArena Multi-Turn13391340

Frequently asked questions

Is Granite 4.2 30b better than Step 1o Turbo 202506?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 39.7 on the Noometry Index.

Is Granite 4.2 30b or Step 1o Turbo 202506 better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 39.2 in the Noometry coding category.

How many benchmarks do Granite 4.2 30b and Step 1o Turbo 202506 share?

11 benchmarks have published results for both models. Granite 4.2 30b has 11 scored results on Noometry and Step 1o Turbo 202506 has 14.

Related comparisons

Go deeper