Model comparison

Gemma 1.1 2b IT vs Qwen Plus

Qwen Plus is the stronger model overall, scoring 37.1 to 29.3 on the Noometry Index.

Last verified . 12 shared benchmarks.

Gemma 1.1 2b IT Google

29.3

Rank #313 Confirmed

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Gemma 1.1 2b IT scores higher in 1 category and Qwen Plus in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen Plus leads 52.2 to 25.1.
  • Gemma 1.1 2b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 1.1 2b IT and Qwen Plus specifications
Gemma 1.1 2b ITQwen Plus
ProviderGoogleAlibaba (Qwen)
Noometry Index29.337.1
Released—2024-01-25
WeightsOpenProprietary
Context window—1M
Max output—33K
Input $ / M tokens—$0.40
Output $ / M tokens—$1.20
Results tracked1620

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen Plus leads

Gemma 1.1 2b IT: 30.1 (#299), Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Coding10341328
HumanEval+17.7%—
MBPP+23.3%—

Reasoning Qwen Plus leads

Gemma 1.1 2b IT: 19.1 (#270), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Hard Prompts10051317
Kagi LLM Benchmark—63.3%
DTBench—81.1%
LMCA—24%

Math Gemma 1.1 2b IT leads

Gemma 1.1 2b IT: 30.8 (#232), Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Math10471326
OTIS Mock AIME 2024-2025—17.8%
MATH Level 5—65.3%
FrontierMath (Feb 2025 set)—1.7%

Knowledge Too close to call

Gemma 1.1 2b IT: 26.5 (#258), Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Expert9701328
GPQA Diamond—48.1%

Multilingual Qwen Plus leads

Gemma 1.1 2b IT: 24.6 (#289), Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Non-English9881310
LMArena Chinese10121347
LMArena Russian9901323
LMArena German944—
LMArena Japanese—1251
LMArena Korean899—

Instruction Following Qwen Plus leads

Gemma 1.1 2b IT: 49.9 (#299), Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Instruction Following9921303

Long Context Qwen Plus leads

Gemma 1.1 2b IT: 30.6 (#286), Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Longer Query10031324

Writing & Preference Qwen Plus leads

Gemma 1.1 2b IT: 25.1 (#306), Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkGemma 1.1 2b ITQwen Plus
LMArena Text10221326
LMArena Creative Writing9981293
LMArena Multi-Turn9591336

Frequently asked questions

Is Gemma 1.1 2b IT better than Qwen Plus?

Qwen Plus is the stronger model overall, scoring 37.1 to 29.3 on the Noometry Index.

Is Gemma 1.1 2b IT or Qwen Plus better for coding?

Qwen Plus scores higher on coding benchmarks: 38.9 versus 30.1 in the Noometry coding category.

How many benchmarks do Gemma 1.1 2b IT and Qwen Plus share?

12 benchmarks have published results for both models. Gemma 1.1 2b IT has 16 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper