Model comparison

Gemma 7B vs Qwen Plus

Qwen Plus is the stronger model overall, scoring 37.1 to 30.0 on the Noometry Index.

Last verified . 12 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Gemma 7B scores higher in 1 category and Qwen Plus in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen Plus leads 52.2 to 27.1.
  • Gemma 7B has downloadable open weights; the other is API-only.

Side by side

Gemma 7B and Qwen Plus specifications
Gemma 7BQwen Plus
ProviderGoogleAlibaba (Qwen)
Noometry Index30.037.1
Released2024-02-212024-01-25
WeightsOpenProprietary
Context window—1M
Max output—33K
Input $ / M tokens—$0.40
Output $ / M tokens—$1.20
Results tracked2720

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen Plus leads

Gemma 7B: 30.5 (#294), Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Coding10481328
HumanEval+28.7%—
MBPP+43.4%—

Reasoning Qwen Plus leads

Gemma 7B: 19.9 (#249), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Hard Prompts10421317
Kagi LLM Benchmark—63.3%
DTBench—81.1%
LMCA—24%
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Gemma 7B leads

Gemma 7B: 31.2 (#228), Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Math10661326
OTIS Mock AIME 2024-2025—17.8%
MATH Level 5—65.3%
FrontierMath (Feb 2025 set)—1.7%
GSM8K46.4%—

Knowledge Too close to call

Gemma 7B: 27.3 (#252), Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Expert10011328
GPQA Diamond—48.1%
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multilingual Qwen Plus leads

Gemma 7B: 25.1 (#287), Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Non-English9991310
LMArena Chinese10351347
LMArena Russian9931323
LMArena French1025—
LMArena Japanese—1251

Instruction Following Qwen Plus leads

Gemma 7B: 51.5 (#295), Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Instruction Following10171303

Long Context Qwen Plus leads

Gemma 7B: 31.1 (#282), Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Longer Query10221324

Writing & Preference Qwen Plus leads

Gemma 7B: 27.1 (#302), Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkGemma 7BQwen Plus
LMArena Text10561326
LMArena Creative Writing10241293
LMArena Multi-Turn9631336

Frequently asked questions

Is Gemma 7B better than Qwen Plus?

Qwen Plus is the stronger model overall, scoring 37.1 to 30.0 on the Noometry Index.

Is Gemma 7B or Qwen Plus better for coding?

Qwen Plus scores higher on coding benchmarks: 38.9 versus 30.5 in the Noometry coding category.

How many benchmarks do Gemma 7B and Qwen Plus share?

12 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper