Model comparison

Chatgpt 4o Latest 20250326 vs Gemma 4 26B A4B IT

Chatgpt 4o Latest 20250326 and Gemma 4 26B A4B IT score almost the same on the Noometry Index (43.8 vs 43.5), so choose on price, context window or the category you care about most.

Last verified . 16 shared benchmarks.

Chatgpt 4o Latest 20250326 OpenAI

43.8

Rank #82 Confirmed

Gemma 4 26B A4B IT Google

43.5

Rank #92 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Chatgpt 4o Latest 20250326 scores higher in 3 categories and Gemma 4 26B A4B IT in 6 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Chatgpt 4o Latest 20250326 leads 33.8 to 21.8.
  • Gemma 4 26B A4B IT has downloadable open weights; the other is API-only.

Side by side

Chatgpt 4o Latest 20250326 and Gemma 4 26B A4B IT specifications
Chatgpt 4o Latest 20250326Gemma 4 26B A4B IT
ProviderOpenAIGoogle
Noometry Index43.843.5
Released—2026-04-02
WeightsProprietaryOpen
Context window—262K
Max output—33K
Input $ / M tokens—$0.0675
Output $ / M tokens—$0.23
Results tracked2128

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 41.6 (#122), Gemma 4 26B A4B IT: 39.0 (#164)

Coding benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Coding14131447
LMArena WebDev—1359
SciCode—40%
WeirdML—35.2%
ALE-Bench—927.17

Reasoning Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 33.8 (#71), Gemma 4 26B A4B IT: 21.8 (#213)

Reasoning benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Hard Prompts14241439
Kagi LLM Benchmark75%—
CritPt—0%
Chess Puzzles—6%
DTBench—74.9%
LMCA—29.7%
Epoch Capabilities Index—141.85

Math Gemma 4 26B A4B IT leads

Chatgpt 4o Latest 20250326: 38.6 (#134), Gemma 4 26B A4B IT: 47.6 (#67)

Math benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Math14071470
OTIS Mock AIME 2024-2025—82.2%

Knowledge Gemma 4 26B A4B IT leads

Chatgpt 4o Latest 20250326: 39.2 (#137), Gemma 4 26B A4B IT: 45.7 (#85)

Knowledge benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Expert14011447
GPQA Diamond—73.2%
Confabulations16.6%—
Vectara Hallucination Rate—5.2%

Multimodal Gemma 4 26B A4B IT leads

Chatgpt 4o Latest 20250326: 39.6 (#58), Gemma 4 26B A4B IT: 40.6 (#46)

Multimodal benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Vision12431260

Multilingual Too close to call

Chatgpt 4o Latest 20250326: 52.9 (#76), Gemma 4 26B A4B IT: 53.1 (#71)

Multilingual benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Non-English14191421
LMArena Chinese14571495
LMArena French14461460
LMArena Russian14291434
LMArena Spanish14341417
LMArena German1423—
LMArena Japanese1405—
LMArena Korean1396—

Instruction Following Too close to call

Chatgpt 4o Latest 20250326: 74.0 (#107), Gemma 4 26B A4B IT: 74.8 (#82)

Instruction Following benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Instruction Following14031420

Long Context Too close to call

Chatgpt 4o Latest 20250326: 43.1 (#107), Gemma 4 26B A4B IT: 43.6 (#91)

Long Context benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Longer Query14131428

Writing & Preference Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 62.6 (#72), Gemma 4 26B A4B IT: 58.6 (#115)

Writing & Preference benchmarks
BenchmarkChatgpt 4o Latest 20250326Gemma 4 26B A4B IT
LMArena Text14291434
LMArena Creative Writing14051402
EQ-Bench Creative Writing15011305
LMArena Multi-Turn14541441

Frequently asked questions

Is Chatgpt 4o Latest 20250326 better than Gemma 4 26B A4B IT?

Chatgpt 4o Latest 20250326 and Gemma 4 26B A4B IT score almost the same on the Noometry Index (43.8 vs 43.5), so choose on price, context window or the category you care about most.

Is Chatgpt 4o Latest 20250326 or Gemma 4 26B A4B IT better for coding?

Chatgpt 4o Latest 20250326 scores higher on coding benchmarks: 41.6 versus 39.0 in the Noometry coding category.

How many benchmarks do Chatgpt 4o Latest 20250326 and Gemma 4 26B A4B IT share?

16 benchmarks have published results for both models. Chatgpt 4o Latest 20250326 has 21 scored results on Noometry and Gemma 4 26B A4B IT has 28.

Related comparisons

Go deeper