Model comparison

Gemma 1.1 7b IT vs Inkling

Inkling is the stronger model overall, scoring 44.1 to 31.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Inkling Thinking Machines Lab

44.1

Rank #80 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 1 category and Inkling in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Inkling leads 65.2 to 30.4.

Side by side

Gemma 1.1 7b IT and Inkling specifications
Gemma 1.1 7b ITInkling
ProviderGoogleThinking Machines Lab
Noometry Index31.344.1
Released—2026-07-15
WeightsOpenOpen
Context window—66K
Max output—66K
Input $ / M tokens—$1.87
Output $ / M tokens—$4.68
Results tracked1941

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Inkling leads

Gemma 1.1 7b IT: 31.5 (#284), Inkling: 34.5 (#234)

Coding benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Coding10841464
FrontierCode—14%
LMArena WebDev—1413
FrontierSWE—4.1%
SciCode—47%
WeirdML—32.3%
ALE-Bench—946
HumanEval+35.4%—
MBPP+45%—

Agentic & Tool Use Not comparable

Gemma 1.1 7b IT: —, Inkling: 29.6 (#85)

Agentic & Tool Use benchmarks
BenchmarkGemma 1.1 7b ITInkling
APEX-Agents—33.8%
τ²-bench Banking—25%

Reasoning Inkling leads

Gemma 1.1 7b IT: 20.5 (#238), Inkling: 40.4 (#56)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Hard Prompts10711451
ARC-AGI-2—36.5%
SimpleBench—50%
ARC-AGI-1—79.5%
CritPt—5.4%
Chess Puzzles—21%
DTBench—87.5%
LMCA—37.6%
Epoch Capabilities Index—148.54

Math Too close to call

Gemma 1.1 7b IT: 32.0 (#220), Inkling: 31.3 (#225)

Math benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Math11071479
FrontierMath (Tiers 1-3)—33.3%
FrontierMath Tier 4—4.9%
OTIS Mock AIME 2024-2025—88.9%
ProofBench—0%

Knowledge Inkling leads

Gemma 1.1 7b IT: 28.3 (#247), Inkling: 55.1 (#49)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Expert10391465
GPQA Diamond—88.3%
SimpleQA Verified—40.3%

Multilingual Inkling leads

Gemma 1.1 7b IT: 28.1 (#273), Inkling: 54.0 (#52)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Non-English10521434
LMArena Chinese10611490
LMArena French10651458
LMArena German10541446
LMArena Japanese9711429
LMArena Korean9881404
LMArena Russian10461429
LMArena Spanish10491448

Instruction Following Inkling leads

Gemma 1.1 7b IT: 54.0 (#283), Inkling: 75.1 (#71)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Instruction Following10571426

Long Context Inkling leads

Gemma 1.1 7b IT: 32.1 (#272), Inkling: 43.8 (#86)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Longer Query10561434

Writing & Preference Inkling leads

Gemma 1.1 7b IT: 30.4 (#288), Inkling: 65.2 (#51)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITInkling
LMArena Text10941441
LMArena Creative Writing10601387
LMArena Multi-Turn10401436
EQ-Bench Creative Writing—1611
EQ-Bench 4—1226

Frequently asked questions

Is Gemma 1.1 7b IT better than Inkling?

Inkling is the stronger model overall, scoring 44.1 to 31.3 on the Noometry Index.

Is Gemma 1.1 7b IT or Inkling better for coding?

Inkling scores higher on coding benchmarks: 34.5 versus 31.5 in the Noometry coding category.

How many benchmarks do Gemma 1.1 7b IT and Inkling share?

17 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Inkling has 41.

Related comparisons

Go deeper