Google, proprietary

Gemini 1.5 Flash (May 2024)

Gemini 1.5 Flash (May 2024) by Google ranks 246th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.2. Its strongest category is multimodal, where it ranks 81st.

Last verified

Specifications

Noometry rank
#246 of 354
Index score
33.2
Evidence
Confirmed 42 results
Provider
Google
Released
May 14, 2024
Weights
Proprietary
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Gemini 1.5 Flash (May 2024) category scores
  1. Coding 34.4
  2. Agentic & Tool Use 26.6
  3. Reasoning 21.7
  4. Math 22.1
  5. Knowledge 26.2
  6. Multimodal 36.0
  7. Multilingual 42.9
  8. Instruction Following 66.8
  9. Long Context 39.0
  10. Writing & Preference 48.7
Gemini 1.5 Flash (May 2024) category ranks
CategoryScoreRankResults
Coding34.4#2364
Agentic & Tool Use26.6#1021
Reasoning21.7#2152
Math22.1#2814
Knowledge26.2#2604
Multimodal36.0#813
Multilingual42.9#1891
Instruction Following66.8#2052
Long Context39.0#1871
Writing & Preference48.7#1964

Strengths and weaknesses

Categories where Gemini 1.5 Flash (May 2024) places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Gemini 1.5 Flash (May 2024): strongest categories
CategoryScorevs medianRank
Reasoning21.7−1.9#215 of 350, top 62%
Writing & Preference48.7−5.1#196 of 312, top 63%
Long Context39.0−1.9#187 of 296, top 64%

Weakest categories

Gemini 1.5 Flash (May 2024): weakest categories
CategoryScorevs medianRank
Math22.1−14.5#281 of 327, top 86%
Knowledge26.2−11.1#260 of 314, top 83%
Coding34.4−4.3#236 of 340, top 70%

Closest competitors

The models ranked just above and below Gemini 1.5 Flash (May 2024). When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Gemini 1.5 Flash (May 2024)
ModelRankScoreBlended $/MSpeed
Mercury 2.5#24233.5$0.0675—Compare
Mistral Small#24333.4$0.26120Compare
Nova 2.0 Pro Preview#24433.4——Compare
Qwen2.5-Coder-32B#24533.4$0.74—Compare
Granite 3.1 2b Instruct#24733.2——Compare
Gemma 2 2b IT#24833.1——Compare
Wizardlm 70b#24933.0——Compare
Phi 3 Medium 4k Instruct#25033.0——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Gemini 1.5 Flash (May 2024) Coding benchmark results
BenchmarkScorePositionSettingSourceDate
WeirdML24.9%#100 of 119, top 85%Epoch AI
BigCodeBench Instruct43.5%#26 of 64, top 41%BigCodeBench2024-05-14
LMArena Coding1236LMArena2026-10-08
LMArena Coding1261#221 of 294, top 76%LMArena2026-10-08
BigCodeBench Complete55.1%#18 of 66, top 28%BigCodeBench2024-05-14
HumanEval+75.6%#14 of 45, top 32%EvalPlus
MBPP+67.5%#18 of 38, top 48%EvalPlus

Agentic & Tool Use

Gemini 1.5 Flash (May 2024) Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
BALROG14.6%#31 of 35, top 89%Epoch AI

Reasoning

Gemini 1.5 Flash (May 2024) Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1220LMArena2026-10-08
LMArena Hard Prompts1257#213 of 297, top 72%LMArena2026-10-08
DTBench53.8%#126 of 151, top 84%Epoch AI
DTBench51.4%Epoch AI
Epoch Capabilities Index122.59Epoch AI2024-05-23
Epoch Capabilities Index129.36#141 of 213, top 67%Epoch AI2024-09-24
ForecastBench53.9#69 of 72, top 96%Epoch AI
PIQA87.5%#3 of 27, top 12%Epoch AI

Math

Gemini 1.5 Flash (May 2024) Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20253.9%Epoch AI2025-02-25
OTIS Mock AIME 2024-202516.3%#129 of 173, top 75%Epoch AI2025-02-25
Omni-MATH30.4%#37 of 57, top 65%HELM Capabilities
LMArena Math1229LMArena2026-10-08
LMArena Math1269#203 of 285, top 72%LMArena2026-10-08
MATH Level 561.9%#39 of 79, top 50%Epoch AI2025-01-27
MATH Level 525.1%Epoch AI2025-01-27
FrontierMath (Feb 2025 set)0%#67 of 68, top 99%Epoch AI2025-03-07
GSM8K82.4%#11 of 38, top 29%Epoch AI

Knowledge

Gemini 1.5 Flash (May 2024) Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond40.4%Epoch AI2025-01-27
GPQA Diamond47.3%#136 of 186, top 74%Epoch AI2025-01-27
MMLU-Pro67.8%#36 of 58, top 63%HELM Capabilities
GPQA (HELM)43.7%#38 of 57, top 67%HELM Capabilities
LMArena Expert1199LMArena2026-10-08
LMArena Expert1233#207 of 273, top 76%LMArena2026-10-08
BoolQ85.8%#7 of 23, top 31%Epoch AI
MMLU77.9%#26 of 81, top 33%Epoch AI
MMLU73.9%Epoch AI
MMLU77.8%Epoch AI

Multimodal

Gemini 1.5 Flash (May 2024) Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1141#101 of 122, top 83%LMArena2026-10-09
LMArena Vision1008LMArena2026-10-09
Video-MME70.3%#8 of 15, top 54%Epoch AI
GeoBench76%#7 of 25, top 29%Epoch AI

Multilingual

Gemini 1.5 Flash (May 2024) Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1231LMArena2026-10-08
LMArena Non-English1278#189 of 297, top 64%LMArena2026-10-08
LMArena Chinese1295#190 of 285, top 67%LMArena2026-10-08
LMArena Chinese1234LMArena2026-10-08
LMArena French1241LMArena2026-10-08
LMArena French1258#182 of 223, top 82%LMArena2026-10-08
LMArena German1262#169 of 231, top 74%LMArena2026-10-08
LMArena German1216LMArena2026-10-08
LMArena Japanese1186LMArena2026-10-08
LMArena Japanese1252#139 of 211, top 66%LMArena2026-10-08
LMArena Korean1201LMArena2026-10-08
LMArena Korean1221#154 of 213, top 73%LMArena2026-10-08
LMArena Russian1239LMArena2026-10-08
LMArena Russian1288#183 of 283, top 65%LMArena2026-10-08
LMArena Spanish1243#185 of 226, top 82%LMArena2026-10-08
LMArena Spanish1224LMArena2026-10-08

Instruction Following

Gemini 1.5 Flash (May 2024) Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
IFEval83.1%#31 of 57, top 55%HELM Capabilities
LMArena Instruction Following1212LMArena2026-10-08
LMArena Instruction Following1258#205 of 298, top 69%LMArena2026-10-08

Long Context

Gemini 1.5 Flash (May 2024) Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1284#200 of 291, top 69%LMArena2026-10-08
LMArena Longer Query1252LMArena2026-10-08

Writing & Preference

Gemini 1.5 Flash (May 2024) Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1287#202 of 297, top 69%LMArena2026-10-08
LMArena Text1239LMArena2026-10-08
LMArena Creative Writing1222LMArena2026-10-08
LMArena Creative Writing1285#182 of 295, top 62%LMArena2026-10-08
WildBench79.2%#34 of 57, top 60%HELM Capabilities
LMArena Multi-Turn1253#222 of 295, top 76%LMArena2026-10-08
LMArena Multi-Turn1228LMArena2026-10-08

Compare Gemini 1.5 Flash (May 2024)

Other Google models

Frequently asked questions

How good is Gemini 1.5 Flash (May 2024)?

Gemini 1.5 Flash (May 2024) by Google ranks 246th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.2. Its strongest category is multimodal, where it ranks 81st.

Is Gemini 1.5 Flash (May 2024) open source?

No. Gemini 1.5 Flash (May 2024) is proprietary and available only through Google's API and partner platforms.

What are Gemini 1.5 Flash (May 2024)'s strengths and weaknesses?

Relative to other ranked models, Gemini 1.5 Flash (May 2024) places best in reasoning, writing & preference, long context and lowest in math, knowledge, coding.

What is Gemini 1.5 Flash (May 2024) best at?

Its best category is multimodal, where it ranks 81st on Noometry.