Model comparison

Codestral vs Gemini 2.0 Flash-Lite

Gemini 2.0 Flash-Lite is the stronger model overall, scoring 37.8 to 30.6 on the Noometry Index.

Last verified . 0 shared benchmarks.

Codestral Mistral AI

30.6

Rank #290 Reported

Gemini 2.0 Flash-Lite Google

37.8

Rank #194 Confirmed

Summary

  • The widest gap is in coding, where Gemini 2.0 Flash-Lite leads 37.9 to 27.3.

Side by side

Codestral and Gemini 2.0 Flash-Lite specifications
CodestralGemini 2.0 Flash-Lite
ProviderMistral AIGoogle
Noometry Index30.637.8
Released2024-05-292025-02-05
WeightsProprietaryProprietary
Context window256K—
Max output8K—
Input $ / M tokens$0.30—
Output $ / M tokens$0.90—
Results tracked732

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 2.0 Flash-Lite leads

Codestral: 27.3 (#321), Gemini 2.0 Flash-Lite: 37.9 (#185)

Coding benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
Aider Polyglot11.1%—
BigCodeBench Instruct41.8%—
LiveBench Coding—47.1%
LMArena Coding—1322
BigCodeBench Complete52.5%—
ALE-Bench137.78—
HumanEval+73.8%—
MBPP+61.9%—

Reasoning Gemini 2.0 Flash-Lite leads

Codestral: 19.8 (#251), Gemini 2.0 Flash-Lite: 22.0 (#210)

Reasoning benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
Kagi LLM Benchmark32.5%—
LiveBench Reasoning—50.1%
LMArena Hard Prompts—1324
DTBench—52.5%
LiveBench Data Analysis—65.5%
ForecastBench—57.1
LiveBench—54.3%

Math Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 34.1 (#196)

Math benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
Omni-MATH—37.4%
LiveBench Math—58.1%
LMArena Math—1309

Knowledge Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 35.0 (#189)

Knowledge benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
MMLU-Pro—72%
GPQA (HELM)—50%
LMArena Expert—1305

Multimodal Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 31.2 (#109)

Multimodal benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
LMArena Vision—1100

Multilingual Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 46.0 (#161)

Multilingual benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
LMArena Non-English—1323
LMArena Chinese—1339
LMArena French—1347
LMArena German—1306
LMArena Japanese—1301
LMArena Korean—1325
LMArena Russian—1328
LMArena Spanish—1313

Instruction Following Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 70.4 (#163)

Instruction Following benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
LiveBench Instruction Following—78.3%
IFEval—82.4%
LMArena Instruction Following—1305

Long Context Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 40.1 (#160)

Long Context benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
LMArena Longer Query—1320

Writing & Preference Not comparable

Codestral: —, Gemini 2.0 Flash-Lite: 51.7 (#177)

Writing & Preference benchmarks
BenchmarkCodestralGemini 2.0 Flash-Lite
LMArena Text—1330
LMArena Creative Writing—1319
WildBench—79%
LMArena Multi-Turn—1307
LiveBench Language—34.3%

Frequently asked questions

Is Codestral better than Gemini 2.0 Flash-Lite?

Gemini 2.0 Flash-Lite is the stronger model overall, scoring 37.8 to 30.6 on the Noometry Index.

Is Codestral or Gemini 2.0 Flash-Lite better for coding?

Gemini 2.0 Flash-Lite scores higher on coding benchmarks: 37.9 versus 27.3 in the Noometry coding category.

How many benchmarks do Codestral and Gemini 2.0 Flash-Lite share?

0 benchmarks have published results for both models. Codestral has 7 scored results on Noometry and Gemini 2.0 Flash-Lite has 32.

Related comparisons

Go deeper