Model comparison

Gemma 7B vs Qwen Turbo

Gemma 7B is the stronger model overall, scoring 30.0 to 27.1 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Qwen Turbo Alibaba (Qwen)

27.1

Rank #335 Reported

Summary

  • The widest gap is in math, where Gemma 7B leads 31.2 to 15.3.
  • Gemma 7B has downloadable open weights; the other is API-only.

Side by side

Gemma 7B and Qwen Turbo specifications
Gemma 7BQwen Turbo
ProviderGoogleAlibaba (Qwen)
Noometry Index30.027.1
Released2024-02-212024-11-01
WeightsOpenProprietary
Context window—1M
Max output—16K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.20
Results tracked273

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 7B: 30.5 (#294), Qwen Turbo: —

Coding benchmarks
BenchmarkGemma 7BQwen Turbo
LMArena Coding1048—
HumanEval+28.7%—
MBPP+43.4%—

Reasoning Not comparable

Gemma 7B: 19.9 (#249), Qwen Turbo: —

Reasoning benchmarks
BenchmarkGemma 7BQwen Turbo
LMArena Hard Prompts1042—
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Gemma 7B leads

Gemma 7B: 31.2 (#228), Qwen Turbo: 15.3 (#297)

Math benchmarks
BenchmarkGemma 7BQwen Turbo
OTIS Mock AIME 2024-2025—6.1%
LMArena Math1066—
MATH Level 5—56.2%
GSM8K46.4%—

Knowledge Gemma 7B leads

Gemma 7B: 27.3 (#252), Qwen Turbo: 22.2 (#272)

Knowledge benchmarks
BenchmarkGemma 7BQwen Turbo
GPQA Diamond—41.8%
LMArena Expert1001—
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multilingual Not comparable

Gemma 7B: 25.1 (#287), Qwen Turbo: —

Multilingual benchmarks
BenchmarkGemma 7BQwen Turbo
LMArena Non-English999—
LMArena Chinese1035—
LMArena French1025—
LMArena Russian993—

Instruction Following Not comparable

Gemma 7B: 51.5 (#295), Qwen Turbo: —

Instruction Following benchmarks
BenchmarkGemma 7BQwen Turbo
LMArena Instruction Following1017—

Long Context Not comparable

Gemma 7B: 31.1 (#282), Qwen Turbo: —

Long Context benchmarks
BenchmarkGemma 7BQwen Turbo
LMArena Longer Query1022—

Writing & Preference Not comparable

Gemma 7B: 27.1 (#302), Qwen Turbo: —

Writing & Preference benchmarks
BenchmarkGemma 7BQwen Turbo
LMArena Text1056—
LMArena Creative Writing1024—
LMArena Multi-Turn963—

Frequently asked questions

Is Gemma 7B better than Qwen Turbo?

Gemma 7B is the stronger model overall, scoring 30.0 to 27.1 on the Noometry Index.

How many benchmarks do Gemma 7B and Qwen Turbo share?

0 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Qwen Turbo has 3.

Related comparisons

Go deeper