Model comparison

Granite 4.2 30b vs Qwen3.7 Flash

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 39.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Qwen3.7 Flash Alibaba (Qwen)

39.9

Rank #156 Confirmed

Summary

  • The widest gap is in knowledge, where Qwen3.7 Flash leads 48.9 to 39.1.
  • Granite 4.2 30b has downloadable open weights; the other is API-only.

Side by side

Granite 4.2 30b and Qwen3.7 Flash specifications
Granite 4.2 30bQwen3.7 Flash
ProviderIBMAlibaba (Qwen)
Noometry Index41.839.9
Released—2026-07-15
WeightsOpenProprietary
Context window—1M
Max output—131K
Input $ / M tokens—$0.03
Output $ / M tokens—$0.13
Results tracked117

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.2 30b: 41.0 (#126), Qwen3.7 Flash: —

Coding benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
LMArena Coding1396—

Reasoning Too close to call

Granite 4.2 30b: 27.8 (#112), Qwen3.7 Flash: 28.2 (#108)

Reasoning benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
NYT Connections (extended)—43.8%
Chess Puzzles—23%
LMArena Hard Prompts1374—
Mystery Game Puzzles—15%
Epoch Capabilities Index—144.64

Math Not comparable

Granite 4.2 30b: —, Qwen3.7 Flash: 38.3 (#140)

Math benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
FrontierMath (Tiers 1-3)—19.3%
OTIS Mock AIME 2024-2025—86.7%

Knowledge Qwen3.7 Flash leads

Granite 4.2 30b: 39.1 (#138), Qwen3.7 Flash: 48.9 (#75)

Knowledge benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
GPQA Diamond—82.3%
LMArena Expert1406—

Multilingual Not comparable

Granite 4.2 30b: 47.3 (#151), Qwen3.7 Flash: —

Multilingual benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
LMArena Non-English1340—
LMArena Chinese1414—
LMArena Russian1343—

Instruction Following Not comparable

Granite 4.2 30b: 71.2 (#155), Qwen3.7 Flash: —

Instruction Following benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
LMArena Instruction Following1347—

Long Context Not comparable

Granite 4.2 30b: 41.4 (#140), Qwen3.7 Flash: —

Long Context benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
LMArena Longer Query1359—

Writing & Preference Not comparable

Granite 4.2 30b: 53.8 (#156), Qwen3.7 Flash: —

Writing & Preference benchmarks
BenchmarkGranite 4.2 30bQwen3.7 Flash
LMArena Text1361—
LMArena Creative Writing1288—
LMArena Multi-Turn1339—

Frequently asked questions

Is Granite 4.2 30b better than Qwen3.7 Flash?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 39.9 on the Noometry Index.

How many benchmarks do Granite 4.2 30b and Qwen3.7 Flash share?

0 benchmarks have published results for both models. Granite 4.2 30b has 11 scored results on Noometry and Qwen3.7 Flash has 7.

Related comparisons

Go deeper