Model comparison

Gemini 2.0 Flash-Lite vs Kimi K2.7 Code

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 37.8 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemini 2.0 Flash-Lite Google

37.8

Rank #194 Confirmed

Kimi K2.7 Code Moonshot AI

43.3

Rank #94 Confirmed

Summary

  • The widest gap is in math, where Kimi K2.7 Code leads 52.9 to 34.1.
  • Kimi K2.7 Code has downloadable open weights; the other is API-only.

Side by side

Gemini 2.0 Flash-Lite and Kimi K2.7 Code specifications
Gemini 2.0 Flash-LiteKimi K2.7 Code
ProviderGoogleMoonshot AI
Noometry Index37.843.3
Released2025-02-052026-06-12
WeightsProprietaryOpen
Context window—262K
Max output—262K
Input $ / M tokens—$0.95
Output $ / M tokens—$4
Results tracked3219

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.7 Code leads

Gemini 2.0 Flash-Lite: 37.9 (#185), Kimi K2.7 Code: 42.9 (#95)

Coding benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
DeepSWE—30.5%
FrontierCode—30.1%
LMArena WebDev—1473
SciCode—47.5%
WeirdML—54.1%
LiveBench Coding47.1%—
LMArena Coding1322—
ALE-Bench—886.23

Agentic & Tool Use Not comparable

Gemini 2.0 Flash-Lite: —, Kimi K2.7 Code: 24.0 (#122)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
APEX-Agents—37.6%
GBAEval—0.9%
Vending-Bench 2—5,083

Reasoning Kimi K2.7 Code leads

Gemini 2.0 Flash-Lite: 22.0 (#210), Kimi K2.7 Code: 39.0 (#61)

Reasoning benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
SimpleBench—57.9%
CritPt—10%
Chess Puzzles—21%
LiveBench Reasoning50.1%—
LMArena Hard Prompts1324—
DTBench52.5%—
LiveBench Data Analysis65.5%—
Surface Evolver Bench—48.8%
Epoch Capabilities Index—149.97
ForecastBench57.1—
LiveBench54.3%—

Math Kimi K2.7 Code leads

Gemini 2.0 Flash-Lite: 34.1 (#196), Kimi K2.7 Code: 52.9 (#48)

Math benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
FrontierMath (Tiers 1-3)—54%
FrontierMath Tier 4—12.2%
OTIS Mock AIME 2024-2025—95.6%
Omni-MATH37.4%—
LiveBench Math58.1%—
LMArena Math1309—

Knowledge Kimi K2.7 Code leads

Gemini 2.0 Flash-Lite: 35.0 (#189), Kimi K2.7 Code: 53.5 (#57)

Knowledge benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
GPQA Diamond—87.9%
SimpleQA Verified—36.5%
MMLU-Pro72%—
GPQA (HELM)50%—
LMArena Expert1305—

Multimodal Not comparable

Gemini 2.0 Flash-Lite: 31.2 (#109), Kimi K2.7 Code: —

Multimodal benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
LMArena Vision1100—

Multilingual Not comparable

Gemini 2.0 Flash-Lite: 46.0 (#161), Kimi K2.7 Code: —

Multilingual benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
LMArena Non-English1323—
LMArena Chinese1339—
LMArena French1347—
LMArena German1306—
LMArena Japanese1301—
LMArena Korean1325—
LMArena Russian1328—
LMArena Spanish1313—

Instruction Following Not comparable

Gemini 2.0 Flash-Lite: 70.4 (#163), Kimi K2.7 Code: —

Instruction Following benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
LiveBench Instruction Following78.3%—
IFEval82.4%—
LMArena Instruction Following1305—

Long Context Not comparable

Gemini 2.0 Flash-Lite: 40.1 (#160), Kimi K2.7 Code: —

Long Context benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
LMArena Longer Query1320—

Writing & Preference Not comparable

Gemini 2.0 Flash-Lite: 51.7 (#177), Kimi K2.7 Code: —

Writing & Preference benchmarks
BenchmarkGemini 2.0 Flash-LiteKimi K2.7 Code
LMArena Text1330—
LMArena Creative Writing1319—
WildBench79%—
LMArena Multi-Turn1307—
LiveBench Language34.3%—

Frequently asked questions

Is Gemini 2.0 Flash-Lite better than Kimi K2.7 Code?

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 37.8 on the Noometry Index.

Is Gemini 2.0 Flash-Lite or Kimi K2.7 Code better for coding?

Kimi K2.7 Code scores higher on coding benchmarks: 42.9 versus 37.9 in the Noometry coding category.

How many benchmarks do Gemini 2.0 Flash-Lite and Kimi K2.7 Code share?

0 benchmarks have published results for both models. Gemini 2.0 Flash-Lite has 32 scored results on Noometry and Kimi K2.7 Code has 19.

Related comparisons

Go deeper