Model comparison

Gemma 1.1 7b IT vs Qwen Max

Qwen Max is the stronger model overall, scoring 34.7 to 31.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Qwen Max Alibaba (Qwen)

34.7

Rank #230 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 2 categories and Qwen Max in 6 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen Max leads 47.8 to 30.4.
  • Gemma 1.1 7b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 1.1 7b IT and Qwen Max specifications
Gemma 1.1 7b ITQwen Max
ProviderGoogleAlibaba (Qwen)
Noometry Index31.334.7
Released—2024-04-03
WeightsOpenProprietary
Context window—33K
Max output—8K
Input $ / M tokens—$1.60
Output $ / M tokens—$6.40
Results tracked1923

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemma 1.1 7b IT: 31.5 (#284), Qwen Max: 30.7 (#292)

Coding benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Coding10841288
Aider Polyglot—21.8%
HumanEval+35.4%—
MBPP+45%—

Reasoning Qwen Max leads

Gemma 1.1 7b IT: 20.5 (#238), Qwen Max: 25.1 (#151)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Hard Prompts10711269

Math Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 32.0 (#220), Qwen Max: 22.3 (#276)

Math benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Math11071275
OTIS Mock AIME 2024-2025—16.1%
MATH Level 5—67.2%
FrontierMath (Feb 2025 set)—1%

Knowledge Qwen Max leads

Gemma 1.1 7b IT: 28.3 (#247), Qwen Max: 30.3 (#228)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Expert10391248
GPQA Diamond—56.1%

Multilingual Qwen Max leads

Gemma 1.1 7b IT: 28.1 (#273), Qwen Max: 41.8 (#202)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Non-English10521263
LMArena Chinese10611254
LMArena French10651330
LMArena German10541254
LMArena Japanese9711205
LMArena Korean9881142
LMArena Russian10461274
LMArena Spanish10491290

Instruction Following Qwen Max leads

Gemma 1.1 7b IT: 54.0 (#283), Qwen Max: 66.5 (#208)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Instruction Following10571262

Long Context Qwen Max leads

Gemma 1.1 7b IT: 32.1 (#272), Qwen Max: 39.4 (#180)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Longer Query10561288
Fiction.LiveBench—66.7%

Writing & Preference Qwen Max leads

Gemma 1.1 7b IT: 30.4 (#288), Qwen Max: 47.8 (#205)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITQwen Max
LMArena Text10941282
LMArena Creative Writing10601248
LMArena Multi-Turn10401277

Frequently asked questions

Is Gemma 1.1 7b IT better than Qwen Max?

Qwen Max is the stronger model overall, scoring 34.7 to 31.3 on the Noometry Index.

Is Gemma 1.1 7b IT or Qwen Max better for coding?

They score almost the same on coding (31.5 vs 30.7); test both on your own repository before choosing.

How many benchmarks do Gemma 1.1 7b IT and Qwen Max share?

17 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Qwen Max has 23.

Related comparisons

Go deeper