Model comparison

Gemma 4 26B A4B IT vs o1-pro

Gemma 4 26B A4B IT is the stronger model overall, scoring 43.5 to 31.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemma 4 26B A4B IT Google

43.5

Rank #92 Confirmed

o1-pro OpenAI

31.5

Rank #271 Reported

Summary

  • The widest gap is in knowledge, where Gemma 4 26B A4B IT leads 45.7 to 29.7.
  • Gemma 4 26B A4B IT is cheaper at $0.0675 / $0.23 per million input/output tokens, against $150 / $600 for o1-pro.
  • Gemma 4 26B A4B IT accepts more context: 262K tokens versus 200K.
  • Gemma 4 26B A4B IT has downloadable open weights; the other is API-only.

Side by side

Gemma 4 26B A4B IT and o1-pro specifications
Gemma 4 26B A4B ITo1-pro
ProviderGoogleOpenAI
Noometry Index43.531.5
Released2026-04-022025-03-19
WeightsOpenProprietary
Context window262K200K
Max output33K100K
Input $ / M tokens$0.0675$150
Output $ / M tokens$0.23$600
Results tracked283

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 4 26B A4B IT: 39.0 (#164), o1-pro: —

Coding benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
LMArena WebDev1359—
SciCode40%—
WeirdML35.2%—
LMArena Coding1447—
ALE-Bench927.17—

Reasoning Gemma 4 26B A4B IT leads

Gemma 4 26B A4B IT: 21.8 (#213), o1-pro: 20.4 (#239)

Reasoning benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
ARC-AGI-1—23.3%
CritPt0%—
Chess Puzzles6%—
EnigmaEval—6.1%
LMArena Hard Prompts1439—
DTBench74.9%—
LMCA29.7%—
Epoch Capabilities Index141.85—

Math Not comparable

Gemma 4 26B A4B IT: 47.6 (#67), o1-pro: —

Math benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
OTIS Mock AIME 2024-202582.2%—
LMArena Math1470—

Knowledge Gemma 4 26B A4B IT leads

Gemma 4 26B A4B IT: 45.7 (#85), o1-pro: 29.7 (#234)

Knowledge benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
GPQA Diamond73.2%—
Humanity's Last Exam—8.1%
Vectara Hallucination Rate5.2%—
LMArena Expert1447—

Multimodal Not comparable

Gemma 4 26B A4B IT: 40.6 (#46), o1-pro: —

Multimodal benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
LMArena Vision1260—

Multilingual Not comparable

Gemma 4 26B A4B IT: 53.1 (#71), o1-pro: —

Multilingual benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
LMArena Non-English1421—
LMArena Chinese1495—
LMArena French1460—
LMArena Russian1434—
LMArena Spanish1417—

Instruction Following Not comparable

Gemma 4 26B A4B IT: 74.8 (#82), o1-pro: —

Instruction Following benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
LMArena Instruction Following1420—

Long Context Not comparable

Gemma 4 26B A4B IT: 43.6 (#91), o1-pro: —

Long Context benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
LMArena Longer Query1428—

Writing & Preference Not comparable

Gemma 4 26B A4B IT: 58.6 (#115), o1-pro: —

Writing & Preference benchmarks
BenchmarkGemma 4 26B A4B ITo1-pro
LMArena Text1434—
LMArena Creative Writing1402—
EQ-Bench Creative Writing1305—
LMArena Multi-Turn1441—

Frequently asked questions

Is Gemma 4 26B A4B IT better than o1-pro?

Gemma 4 26B A4B IT is the stronger model overall, scoring 43.5 to 31.5 on the Noometry Index.

Which is cheaper, Gemma 4 26B A4B IT or o1-pro?

Gemma 4 26B A4B IT is cheaper. It lists at $0.0675 per million input tokens and $0.23 per million output tokens; o1-pro lists at $150 and $600.

Which has the bigger context window?

Gemma 4 26B A4B IT does, with 262K tokens against 200K.

How many benchmarks do Gemma 4 26B A4B IT and o1-pro share?

0 benchmarks have published results for both models. Gemma 4 26B A4B IT has 28 scored results on Noometry and o1-pro has 3.

Related comparisons

Go deeper