Model comparison

DeepSeek Coder 33B vs Gemini 1.5 Flash 8B

Gemini 1.5 Flash 8B has enough public results to be ranked (#301); DeepSeek Coder 33B does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

DeepSeek Coder 33B DeepSeek

38.9

Unranked Sparse

Gemini 1.5 Flash 8B Google

29.9

Rank #301 Confirmed

Summary

  • DeepSeek Coder 33B has downloadable open weights; the other is API-only.

Side by side

DeepSeek Coder 33B and Gemini 1.5 Flash 8B specifications
DeepSeek Coder 33BGemini 1.5 Flash 8B
ProviderDeepSeekGoogle
Noometry Index38.929.9
Released2023-11-022024-10-03
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked921

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek Coder 33B leads

DeepSeek Coder 33B: 38.0 (#184), Gemini 1.5 Flash 8B: 35.5 (#225)

Coding benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
BigCodeBench Instruct42%—
LMArena Coding—1218
BigCodeBench Complete51.1%—
HumanEval+75%—
MBPP+70.1%—

Reasoning Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 20.0 (#244)

Reasoning benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
LMArena Hard Prompts—1209
DTBench—50%
Epoch Capabilities Index96.32—
WinoGrande62%—

Math Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 14.2 (#302)

Math benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
OTIS Mock AIME 2024-2025—4.6%
LMArena Math—1207
GSM8K35.4%—

Knowledge Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 16.0 (#289)

Knowledge benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
GPQA Diamond—33%
LMArena Expert—1185
ARC (AI2) Challenge42.2%—
MMLU39.4%—

Multimodal Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 28.2 (#115)

Multimodal benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
LMArena Vision—1044

Multilingual Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 38.5 (#229)

Multilingual benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
LMArena Non-English—1215
LMArena Chinese—1231
LMArena French—1234
LMArena German—1206
LMArena Japanese—1150
LMArena Korean—1140
LMArena Russian—1236
LMArena Spanish—1212

Instruction Following Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 62.8 (#236)

Instruction Following benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
LMArena Instruction Following—1199

Long Context Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 37.0 (#225)

Long Context benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
LMArena Longer Query—1219

Writing & Preference Not comparable

DeepSeek Coder 33B: —, Gemini 1.5 Flash 8B: 42.8 (#232)

Writing & Preference benchmarks
BenchmarkDeepSeek Coder 33BGemini 1.5 Flash 8B
LMArena Text—1226
LMArena Creative Writing—1218
LMArena Multi-Turn—1185

Frequently asked questions

Is DeepSeek Coder 33B better than Gemini 1.5 Flash 8B?

Gemini 1.5 Flash 8B has enough public results to be ranked (#301); DeepSeek Coder 33B does not yet, so treat this comparison as directional.

Is DeepSeek Coder 33B or Gemini 1.5 Flash 8B better for coding?

DeepSeek Coder 33B scores higher on coding benchmarks: 38.0 versus 35.5 in the Noometry coding category.

How many benchmarks do DeepSeek Coder 33B and Gemini 1.5 Flash 8B share?

0 benchmarks have published results for both models. DeepSeek Coder 33B has 9 scored results on Noometry and Gemini 1.5 Flash 8B has 21.

Related comparisons

Go deeper