Model comparison

DeepSeek LLM 67B vs Gemma 1.1 7b IT

Gemma 1.1 7b IT is the stronger model overall, scoring 31.3 to 24.9 on the Noometry Index.

Last verified . 10 shared benchmarks.

DeepSeek LLM 67B DeepSeek

24.9

Rank #347 Confirmed

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Summary

  • They share 10 benchmarks with published results for both. DeepSeek LLM 67B scores higher in 5 categories and Gemma 1.1 7b IT in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in math, where Gemma 1.1 7b IT leads 32.0 to 8.7.

Side by side

DeepSeek LLM 67B and Gemma 1.1 7b IT specifications
DeepSeek LLM 67BGemma 1.1 7b IT
ProviderDeepSeekGoogle
Noometry Index24.931.3
Released2023-11-29—
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1519

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

DeepSeek LLM 67B: 31.9 (#278), Gemma 1.1 7b IT: 31.5 (#284)

Coding benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Coding10961084
HumanEval+—35.4%
MBPP+—45%

Reasoning Gemma 1.1 7b IT leads

DeepSeek LLM 67B: 16.5 (#304), Gemma 1.1 7b IT: 20.5 (#238)

Reasoning benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Hard Prompts10701071
Chess Puzzles0%—
Epoch Capabilities Index110.5—

Math Gemma 1.1 7b IT leads

DeepSeek LLM 67B: 8.7 (#324), Gemma 1.1 7b IT: 32.0 (#220)

Math benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Math11081107
OTIS Mock AIME 2024-20250.8%—
MATH Level 56.4%—

Knowledge Gemma 1.1 7b IT leads

DeepSeek LLM 67B: 7.0 (#313), Gemma 1.1 7b IT: 28.3 (#247)

Knowledge benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
GPQA Diamond24.6%—
LMArena Expert—1039

Multilingual DeepSeek LLM 67B leads

DeepSeek LLM 67B: 29.4 (#267), Gemma 1.1 7b IT: 28.1 (#273)

Multilingual benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Non-English10731052
LMArena Chinese11321061
LMArena French—1065
LMArena German—1054
LMArena Japanese—971
LMArena Korean—988
LMArena Russian—1046
LMArena Spanish—1049

Instruction Following DeepSeek LLM 67B leads

DeepSeek LLM 67B: 55.4 (#277), Gemma 1.1 7b IT: 54.0 (#283)

Instruction Following benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Instruction Following10791057

Long Context DeepSeek LLM 67B leads

DeepSeek LLM 67B: 33.1 (#265), Gemma 1.1 7b IT: 32.1 (#272)

Long Context benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Longer Query10921056

Writing & Preference DeepSeek LLM 67B leads

DeepSeek LLM 67B: 31.6 (#282), Gemma 1.1 7b IT: 30.4 (#288)

Writing & Preference benchmarks
BenchmarkDeepSeek LLM 67BGemma 1.1 7b IT
LMArena Text11051094
LMArena Creative Writing10671060
LMArena Multi-Turn10821040

Frequently asked questions

Is DeepSeek LLM 67B better than Gemma 1.1 7b IT?

Gemma 1.1 7b IT is the stronger model overall, scoring 31.3 to 24.9 on the Noometry Index.

Is DeepSeek LLM 67B or Gemma 1.1 7b IT better for coding?

They score almost the same on coding (31.9 vs 31.5); test both on your own repository before choosing.

How many benchmarks do DeepSeek LLM 67B and Gemma 1.1 7b IT share?

10 benchmarks have published results for both models. DeepSeek LLM 67B has 15 scored results on Noometry and Gemma 1.1 7b IT has 19.

Related comparisons

Go deeper