Model comparison

DeepSeek-R1 vs Gemini 3.5 Flash Lite

DeepSeek-R1 and Gemini 3.5 Flash Lite score almost the same on the Noometry Index (42.3 vs 41.5), so choose on price, context window or the category you care about most.

Last verified . 27 shared benchmarks.

DeepSeek-R1 DeepSeek

42.3

Rank #115 Confirmed

Gemini 3.5 Flash Lite Google

41.5

Rank #133 Confirmed

Summary

  • They share 27 benchmarks with published results for both. DeepSeek-R1 scores higher in 4 categories and Gemini 3.5 Flash Lite in 5 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in math, where DeepSeek-R1 leads 43.8 to 25.9.
  • The biggest single-benchmark swing is ARC-AGI-1: 21.2% for DeepSeek-R1 and 53.5% for Gemini 3.5 Flash Lite.
  • Gemini 3.5 Flash Lite is cheaper at $0.30 / $2.50 per million input/output tokens, against $0.50 / $2.15 for DeepSeek-R1.
  • Gemini 3.5 Flash Lite accepts more context: 1.05M tokens versus 164K.

Side by side

DeepSeek-R1 and Gemini 3.5 Flash Lite specifications
DeepSeek-R1Gemini 3.5 Flash Lite
ProviderDeepSeekGoogle
Noometry Index42.341.5
Released2025-01-202026-07-21
WeightsProprietaryProprietary
Context window164K1.05M
Max output64K66K
Input $ / M tokens$0.50$0.30
Output $ / M tokens$2.15$2.50
Results tracked5239

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek-R1 leads

DeepSeek-R1: 46.3 (#68), Gemini 3.5 Flash Lite: 42.4 (#104)

Coding benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
SciCode35.7%41.3%
WeirdML41.6%39%
LMArena Coding14271453
ALE-Bench804.12765.27
Aider Polyglot71.4%—
LMArena WebDev—1440
LiveBench Coding66.7%—
AlgoTune1.7—

Agentic & Tool Use DeepSeek-R1 leads

DeepSeek-R1: 30.7 (#75), Gemini 3.5 Flash Lite: 25.7 (#107)

Agentic & Tool Use benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
APEX-Agents—29.3%
DeepResearch Bench35.1%—
BALROG34.9%—
GDP.pdf—10%
METR Time Horizons53.8%—

Reasoning Gemini 3.5 Flash Lite leads

DeepSeek-R1: 18.6 (#278), Gemini 3.5 Flash Lite: 27.8 (#114)

Reasoning benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
ARC-AGI-21.3%10.3%
ARC-AGI-121.2%53.5%
CritPt1.1%0%
LMArena Hard Prompts14161441
Epoch Capabilities Index141.29145.13
SimpleBench40.8%—
Kagi LLM Benchmark69.4%—
NYT Connections (extended)—60.4%
Chess Puzzles—22%
LiveBench Reasoning83.2%—
Mystery Game Puzzles—19%
DTBench—83.5%
LiveBench Data Analysis69.8%—
LMCA—37.6%
ForecastBench60—
LiveBench71.6%—

Math DeepSeek-R1 leads

DeepSeek-R1: 43.8 (#79), Gemini 3.5 Flash Lite: 25.9 (#264)

Math benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
OTIS Mock AIME 2024-202566.4%71.1%
LMArena Math14001433
FrontierMath (Tiers 1-3)—26%
FrontierMath Tier 4—0%
ProofBench—13%
Omni-MATH42.4%—
LiveBench Math80.7%—
MATH Level 596.6%—

Knowledge Gemini 3.5 Flash Lite leads

DeepSeek-R1: 44.5 (#87), Gemini 3.5 Flash Lite: 50.1 (#72)

Knowledge benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
GPQA Diamond76.3%83.3%
LMArena Expert13941435
MMLU-Pro79.3%—
Confabulations12.7%—
Vectara Hallucination Rate11.3%—
GPQA (HELM)66.6%—

Multimodal Not comparable

DeepSeek-R1: —, Gemini 3.5 Flash Lite: 41.1 (#40)

Multimodal benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
LMArena Vision—1268

Multilingual Gemini 3.5 Flash Lite leads

DeepSeek-R1: 52.4 (#85), Gemini 3.5 Flash Lite: 53.5 (#63)

Multilingual benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
LMArena Non-English14121427
LMArena Chinese14421469
LMArena French14171451
LMArena German14041454
LMArena Japanese13911427
LMArena Korean13601407
LMArena Russian14231443
LMArena Spanish14111442

Instruction Following Gemini 3.5 Flash Lite leads

DeepSeek-R1: 72.0 (#143), Gemini 3.5 Flash Lite: 74.8 (#85)

Instruction Following benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
LMArena Instruction Following13821420
LiveBench Instruction Following80.5%—
IFEval78.4%—

Long Context DeepSeek-R1 leads

DeepSeek-R1: 45.4 (#36), Gemini 3.5 Flash Lite: 43.9 (#84)

Long Context benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
LMArena Longer Query13911435
Fiction.LiveBench75%—

Writing & Preference Gemini 3.5 Flash Lite leads

DeepSeek-R1: 61.4 (#88), Gemini 3.5 Flash Lite: 64.2 (#57)

Writing & Preference benchmarks
BenchmarkDeepSeek-R1Gemini 3.5 Flash Lite
LMArena Text14281435
LMArena Creative Writing14051420
EQ-Bench Creative Writing15001559
LMArena Multi-Turn14051445
Short-Story Creative Writing83%—
WildBench82.8%—
LiveBench Language48.5%—

Frequently asked questions

Is DeepSeek-R1 better than Gemini 3.5 Flash Lite?

DeepSeek-R1 and Gemini 3.5 Flash Lite score almost the same on the Noometry Index (42.3 vs 41.5), so choose on price, context window or the category you care about most.

Which is cheaper, DeepSeek-R1 or Gemini 3.5 Flash Lite?

Gemini 3.5 Flash Lite is cheaper. It lists at $0.30 per million input tokens and $2.50 per million output tokens; DeepSeek-R1 lists at $0.50 and $2.15.

Is DeepSeek-R1 or Gemini 3.5 Flash Lite better for coding?

DeepSeek-R1 scores higher on coding benchmarks: 46.3 versus 42.4 in the Noometry coding category.

Which has the bigger context window?

Gemini 3.5 Flash Lite does, with 1.05M tokens against 164K.

How many benchmarks do DeepSeek-R1 and Gemini 3.5 Flash Lite share?

27 benchmarks have published results for both models. DeepSeek-R1 has 52 scored results on Noometry and Gemini 3.5 Flash Lite has 39.

Related comparisons

Go deeper