Model comparison

GPT-5.2 Codex vs Granite 4.2 30b

GPT-5.2 Codex and Granite 4.2 30b score almost the same on the Noometry Index (42.6 vs 41.8), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

GPT-5.2 Codex OpenAI

42.6

Rank #111 Reported

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Summary

  • The widest gap is in coding, where GPT-5.2 Codex leads 45.5 to 41.0.
  • Granite 4.2 30b has downloadable open weights; the other is API-only.

Side by side

GPT-5.2 Codex and Granite 4.2 30b specifications
GPT-5.2 CodexGranite 4.2 30b
ProviderOpenAIIBM
Noometry Index42.641.8
Released2025-12-18—
WeightsProprietaryOpen
Context window400K—
Max output128K—
Input $ / M tokens$1.75—
Output $ / M tokens$14—
Results tracked511

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-5.2 Codex leads

GPT-5.2 Codex: 45.5 (#71), Granite 4.2 30b: 41.0 (#126)

Coding benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
SWE-bench Verified (bash only)72.8%—
LMArena WebDev1339—
SWE-bench Multilingual66.3%—
LMArena Coding—1396
ALE-Bench1,300—

Agentic & Tool Use Not comparable

GPT-5.2 Codex: 41.0 (#22), Granite 4.2 30b: —

Agentic & Tool Use benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
Terminal-Bench66.5%—

Reasoning Not comparable

GPT-5.2 Codex: —, Granite 4.2 30b: 27.8 (#112)

Reasoning benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
LMArena Hard Prompts—1374

Knowledge Not comparable

GPT-5.2 Codex: —, Granite 4.2 30b: 39.1 (#138)

Knowledge benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
LMArena Expert—1406

Multilingual Not comparable

GPT-5.2 Codex: —, Granite 4.2 30b: 47.3 (#151)

Multilingual benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
LMArena Non-English—1340
LMArena Chinese—1414
LMArena Russian—1343

Instruction Following Not comparable

GPT-5.2 Codex: —, Granite 4.2 30b: 71.2 (#155)

Instruction Following benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
LMArena Instruction Following—1347

Long Context Not comparable

GPT-5.2 Codex: —, Granite 4.2 30b: 41.4 (#140)

Long Context benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
LMArena Longer Query—1359

Writing & Preference Not comparable

GPT-5.2 Codex: —, Granite 4.2 30b: 53.8 (#156)

Writing & Preference benchmarks
BenchmarkGPT-5.2 CodexGranite 4.2 30b
LMArena Text—1361
LMArena Creative Writing—1288
LMArena Multi-Turn—1339

Frequently asked questions

Is GPT-5.2 Codex better than Granite 4.2 30b?

GPT-5.2 Codex and Granite 4.2 30b score almost the same on the Noometry Index (42.6 vs 41.8), so choose on price, context window or the category you care about most.

Is GPT-5.2 Codex or Granite 4.2 30b better for coding?

GPT-5.2 Codex scores higher on coding benchmarks: 45.5 versus 41.0 in the Noometry coding category.

How many benchmarks do GPT-5.2 Codex and Granite 4.2 30b share?

0 benchmarks have published results for both models. GPT-5.2 Codex has 5 scored results on Noometry and Granite 4.2 30b has 11.

Related comparisons

Go deeper