Model comparison

Gemma 1.1 2b IT vs Yi-Large

Gemma 1.1 2b IT has enough public results to be ranked (#313); Yi-Large does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Gemma 1.1 2b IT Google

29.3

Rank #313 Confirmed

Yi-Large 01.AI

38.1

Unranked Sparse

Summary

  • The widest gap is in coding, where Yi-Large leads 36.7 to 30.1.
  • Gemma 1.1 2b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 1.1 2b IT and Yi-Large specifications
Gemma 1.1 2b ITYi-Large
ProviderGoogle01.AI
Noometry Index29.338.1
Released—2024-05-13
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked163

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Yi-Large leads

Gemma 1.1 2b IT: 30.1 (#299), Yi-Large: 36.7 (#203)

Coding benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
BigCodeBench Instruct—37.7%
LMArena Coding1034—
BigCodeBench Complete—47.2%
HumanEval+17.7%—
MBPP+23.3%—

Reasoning Not comparable

Gemma 1.1 2b IT: 19.1 (#270), Yi-Large: —

Reasoning benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Hard Prompts1005—

Math Not comparable

Gemma 1.1 2b IT: 30.8 (#232), Yi-Large: —

Math benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Math1047—

Knowledge Not comparable

Gemma 1.1 2b IT: 26.5 (#258), Yi-Large: —

Knowledge benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Expert970—
MMLU—79.3%

Multilingual Not comparable

Gemma 1.1 2b IT: 24.6 (#289), Yi-Large: —

Multilingual benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Non-English988—
LMArena Chinese1012—
LMArena German944—
LMArena Korean899—
LMArena Russian990—

Instruction Following Not comparable

Gemma 1.1 2b IT: 49.9 (#299), Yi-Large: —

Instruction Following benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Instruction Following992—

Long Context Not comparable

Gemma 1.1 2b IT: 30.6 (#286), Yi-Large: —

Long Context benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Longer Query1003—

Writing & Preference Not comparable

Gemma 1.1 2b IT: 25.1 (#306), Yi-Large: —

Writing & Preference benchmarks
BenchmarkGemma 1.1 2b ITYi-Large
LMArena Text1022—
LMArena Creative Writing998—
LMArena Multi-Turn959—

Frequently asked questions

Is Gemma 1.1 2b IT better than Yi-Large?

Gemma 1.1 2b IT has enough public results to be ranked (#313); Yi-Large does not yet, so treat this comparison as directional.

Is Gemma 1.1 2b IT or Yi-Large better for coding?

Yi-Large scores higher on coding benchmarks: 36.7 versus 30.1 in the Noometry coding category.

How many benchmarks do Gemma 1.1 2b IT and Yi-Large share?

0 benchmarks have published results for both models. Gemma 1.1 2b IT has 16 scored results on Noometry and Yi-Large has 3.

Related comparisons

Go deeper