Model comparison

Gemma 2 9B vs Yi-1.5-34B

Yi-1.5-34B is the stronger model overall, scoring 30.6 to 25.9 on the Noometry Index.

Last verified . 21 shared benchmarks.

Gemma 2 9B Google

25.9

Rank #341 Confirmed

Yi-1.5-34B 01.AI

30.6

Rank #289 Confirmed

Summary

  • They share 21 benchmarks with published results for both. Gemma 2 9B scores higher in 2 categories and Yi-1.5-34B in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Yi-1.5-34B leads 27.5 to 9.9.

Side by side

Gemma 2 9B and Yi-1.5-34B specifications
Gemma 2 9BYi-1.5-34B
ProviderGoogle01.AI
Noometry Index25.930.6
Released2024-06-242024-05-13
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked3521

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Yi-1.5-34B leads

Gemma 2 9B: 29.4 (#304), Yi-1.5-34B: 32.4 (#272)

Coding benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
BigCodeBench Instruct34.7%33.9%
LMArena Coding11731169
BigCodeBench Complete40.6%43.8%
LiveBench Coding22.5%—

Reasoning Yi-1.5-34B leads

Gemma 2 9B: 15.9 (#309), Yi-1.5-34B: 22.5 (#191)

Reasoning benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
LMArena Hard Prompts11711160
LiveBench Reasoning15.2%—
LiveBench Data Analysis36.4%—
Epoch Capabilities Index119.83—
LiveBench28.7%—
PIQA83.7%—

Math Yi-1.5-34B leads

Gemma 2 9B: 9.9 (#318), Yi-1.5-34B: 27.5 (#249)

Math benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
LMArena Math11831182
MATH Level 521%25.5%
OTIS Mock AIME 2024-20250.6%—
LiveBench Math19.8%—
GSM8K84.9%—

Knowledge Yi-1.5-34B leads

Gemma 2 9B: 9.7 (#305), Yi-1.5-34B: 14.8 (#295)

Knowledge benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
GPQA Diamond27.5%32%
LMArena Expert11471144
BoolQ85.7%—
MMLU72.1%—

Multilingual Gemma 2 9B leads

Gemma 2 9B: 36.6 (#238), Yi-1.5-34B: 32.3 (#256)

Multilingual benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
LMArena Non-English11881121
LMArena Chinese11851213
LMArena French11901156
LMArena German11861111
LMArena Japanese11441021
LMArena Korean11371005
LMArena Russian12001091
LMArena Spanish12001121

Instruction Following Yi-1.5-34B leads

Gemma 2 9B: 57.6 (#269), Yi-1.5-34B: 59.2 (#257)

Instruction Following benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
LMArena Instruction Following11781139
LiveBench Instruction Following52.6%—

Long Context Gemma 2 9B leads

Gemma 2 9B: 36.3 (#233), Yi-1.5-34B: 34.6 (#248)

Long Context benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
LMArena Longer Query11971143

Writing & Preference Yi-1.5-34B leads

Gemma 2 9B: 32.1 (#281), Yi-1.5-34B: 37.4 (#257)

Writing & Preference benchmarks
BenchmarkGemma 2 9BYi-1.5-34B
LMArena Text12071173
LMArena Creative Writing12061135
LMArena Multi-Turn11931153
EQ-Bench Creative Writing841—
LiveBench Language25.5%—

Frequently asked questions

Is Gemma 2 9B better than Yi-1.5-34B?

Yi-1.5-34B is the stronger model overall, scoring 30.6 to 25.9 on the Noometry Index.

Is Gemma 2 9B or Yi-1.5-34B better for coding?

Yi-1.5-34B scores higher on coding benchmarks: 32.4 versus 29.4 in the Noometry coding category.

How many benchmarks do Gemma 2 9B and Yi-1.5-34B share?

21 benchmarks have published results for both models. Gemma 2 9B has 35 scored results on Noometry and Yi-1.5-34B has 21.

Related comparisons

Go deeper