Model comparison

Gemma 1.1 7b IT vs Qwen1.5-72B

Gemma 1.1 7b IT and Qwen1.5-72B score almost the same on the Noometry Index (31.3 vs 30.8), so choose on price, context window or the category you care about most.

Last verified . 19 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Qwen1.5-72B Alibaba (Qwen)

30.8

Rank #285 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 1 category and Qwen1.5-72B in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemma 1.1 7b IT leads 28.3 to 11.5.

Side by side

Gemma 1.1 7b IT and Qwen1.5-72B specifications
Gemma 1.1 7b ITQwen1.5-72B
ProviderGoogleAlibaba (Qwen)
Noometry Index31.330.8
Released—2024-02-04
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1922

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemma 1.1 7b IT: 31.5 (#284), Qwen1.5-72B: 31.9 (#277)

Coding benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Coding10841165
HumanEval+35.4%59.1%
MBPP+45%61.6%
BigCodeBench Instruct—33.2%
BigCodeBench Complete—40.3%

Reasoning Qwen1.5-72B leads

Gemma 1.1 7b IT: 20.5 (#238), Qwen1.5-72B: 22.2 (#203)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Hard Prompts10711148

Math Qwen1.5-72B leads

Gemma 1.1 7b IT: 32.0 (#220), Qwen1.5-72B: 33.2 (#205)

Math benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Math11071164

Knowledge Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 28.3 (#247), Qwen1.5-72B: 11.5 (#300)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Expert10391136
GPQA Diamond—28.8%

Multilingual Qwen1.5-72B leads

Gemma 1.1 7b IT: 28.1 (#273), Qwen1.5-72B: 33.2 (#253)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Non-English10521135
LMArena Chinese10611186
LMArena French10651159
LMArena German10541084
LMArena Japanese9711061
LMArena Korean9881050
LMArena Russian10461104
LMArena Spanish10491110

Instruction Following Qwen1.5-72B leads

Gemma 1.1 7b IT: 54.0 (#283), Qwen1.5-72B: 59.3 (#256)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Instruction Following10571141

Long Context Qwen1.5-72B leads

Gemma 1.1 7b IT: 32.1 (#272), Qwen1.5-72B: 35.1 (#243)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Longer Query10561157

Writing & Preference Qwen1.5-72B leads

Gemma 1.1 7b IT: 30.4 (#288), Qwen1.5-72B: 37.3 (#258)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-72B
LMArena Text10941166
LMArena Creative Writing10601137
LMArena Multi-Turn10401160

Frequently asked questions

Is Gemma 1.1 7b IT better than Qwen1.5-72B?

Gemma 1.1 7b IT and Qwen1.5-72B score almost the same on the Noometry Index (31.3 vs 30.8), so choose on price, context window or the category you care about most.

Is Gemma 1.1 7b IT or Qwen1.5-72B better for coding?

They score almost the same on coding (31.5 vs 31.9); test both on your own repository before choosing.

How many benchmarks do Gemma 1.1 7b IT and Qwen1.5-72B share?

19 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Qwen1.5-72B has 22.

Related comparisons

Go deeper