Model comparison

Gemma 1.1 7b IT vs Qwen2.5 Plus 1127

Qwen2.5 Plus 1127 is the stronger model overall, scoring 38.8 to 31.3 on the Noometry Index.

Last verified . 14 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Qwen2.5 Plus 1127 Alibaba (Qwen)

38.8

Rank #181 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 0 categories and Qwen2.5 Plus 1127 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen2.5 Plus 1127 leads 49.4 to 30.4.
  • Gemma 1.1 7b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 1.1 7b IT and Qwen2.5 Plus 1127 specifications
Gemma 1.1 7b ITQwen2.5 Plus 1127
ProviderGoogleAlibaba (Qwen)
Noometry Index31.338.8
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1914

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 31.5 (#284), Qwen2.5 Plus 1127: 38.5 (#175)

Coding benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Coding10841314
HumanEval+35.4%—
MBPP+45%—

Reasoning Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 20.5 (#238), Qwen2.5 Plus 1127: 25.9 (#141)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Hard Prompts10711299

Math Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 32.0 (#220), Qwen2.5 Plus 1127: 36.1 (#174)

Math benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Math11071298

Knowledge Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 28.3 (#247), Qwen2.5 Plus 1127: 35.5 (#183)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Expert10391289

Multilingual Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 28.1 (#273), Qwen2.5 Plus 1127: 41.9 (#201)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Non-English10521265
LMArena Chinese10611314
LMArena German10541231
LMArena Japanese9711207
LMArena Russian10461271
LMArena French1065—
LMArena Korean988—
LMArena Spanish1049—

Instruction Following Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 54.0 (#283), Qwen2.5 Plus 1127: 67.2 (#199)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Instruction Following10571275

Long Context Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 32.1 (#272), Qwen2.5 Plus 1127: 39.2 (#184)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Longer Query10561292

Writing & Preference Qwen2.5 Plus 1127 leads

Gemma 1.1 7b IT: 30.4 (#288), Qwen2.5 Plus 1127: 49.4 (#192)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITQwen2.5 Plus 1127
LMArena Text10941299
LMArena Creative Writing10601262
LMArena Multi-Turn10401299

Frequently asked questions

Is Gemma 1.1 7b IT better than Qwen2.5 Plus 1127?

Qwen2.5 Plus 1127 is the stronger model overall, scoring 38.8 to 31.3 on the Noometry Index.

Is Gemma 1.1 7b IT or Qwen2.5 Plus 1127 better for coding?

Qwen2.5 Plus 1127 scores higher on coding benchmarks: 38.5 versus 31.5 in the Noometry coding category.

How many benchmarks do Gemma 1.1 7b IT and Qwen2.5 Plus 1127 share?

14 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Qwen2.5 Plus 1127 has 14.

Related comparisons

Go deeper