Google, proprietary
Gemini 1.5 Pro (May 2024)
Gemini 1.5 Pro (May 2024) by Google ranks 261st of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.1. Its strongest category is multimodal, where it ranks 77th.
Last verified
Specifications
- Noometry rank
- #261 of 354
- Index score
- 32.1
- Evidence
- Confirmed 45 results
- Provider
Google
- Released
- February 15, 2024
- Weights
- Proprietary
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 34.2
- Agentic & Tool Use 17.9
- Reasoning 12.3
- Math 25.8
- Knowledge 29.4
- Multimodal 36.8
- Multilingual 45.3
- Instruction Following 68.6
- Long Context 39.8
- Writing & Preference 52.4
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 34.2 | #241 | 5 |
| Agentic & Tool Use | 17.9 | #145 | 3 |
| Reasoning | 12.3 | #338 | 4 |
| Math | 25.8 | #266 | 4 |
| Knowledge | 29.4 | #239 | 6 |
| Multimodal | 36.8 | #77 | 2 |
| Multilingual | 45.3 | #174 | 1 |
| Instruction Following | 68.6 | #185 | 2 |
| Long Context | 39.8 | #169 | 1 |
| Writing & Preference | 52.4 | #172 | 4 |
Strengths and weaknesses
Categories where Gemini 1.5 Pro (May 2024) places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Writing & Preference | 52.4 | −1.4 | #172 of 312, top 56% |
| Long Context | 39.8 | −1.2 | #169 of 296, top 58% |
| Multilingual | 45.3 | −2.1 | #174 of 297, top 59% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 12.3 | −11.4 | #338 of 350, top 97% |
| Agentic & Tool Use | 17.9 | −12.4 | #145 of 154, top 95% |
| Math | 25.8 | −10.7 | #266 of 327, top 82% |
Closest competitors
The models ranked just above and below Gemini 1.5 Pro (May 2024). When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Command R+ | #257 | 32.4 | $4.38 | — | Compare |
| Granite 3.1 8b Instruct | #258 | 32.4 | — | — | Compare |
| Pixtral Large | #259 | 32.2 | $3 | — | Compare |
| Falcon-180B | #260 | 32.2 | — | — | Compare |
| Gemma 3 12B | #262 | 32.1 | $0.075 | — | Compare |
| Mistral Large | #263 | 31.9 | $3 | — | Compare |
| Qwen3-4B | #264 | 31.9 | — | — | Compare |
| Amazon Nova Lite | #265 | 31.9 | $0.10 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| WeirdML | 22.2% | #103 of 119, top 87% | Epoch AI | ||
| BigCodeBench Instruct | 43.8% | #24 of 64, top 38% | BigCodeBench | 2024-05-14 | |
| LMArena Coding | 1266 | LMArena | 2026-10-08 | ||
| LMArena Coding | 1294 | #201 of 294, top 69% | LMArena | 2026-10-08 | |
| BigCodeBench Complete | 57.5% | #11 of 66, top 17% | BigCodeBench | 2024-05-14 | |
| CadEval | 34% | #9 of 14, top 65% | Epoch AI | ||
| HumanEval+ | 61% | EvalPlus | |||
| HumanEval+ | 79.3% | #11 of 45, top 25% | EvalPlus | ||
| MBPP+ | 74.6% | #5 of 38, top 14% | EvalPlus |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| TheAgentCompany | 3.4% | #12 of 14, top 86% | Epoch AI | ||
| Cybench | 7.5% | #18 of 21, top 86% | Epoch AI | ||
| BALROG | 21% | #23 of 35, top 66% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| ARC-AGI-2 | 0.8% | #71 of 83, top 86% | Epoch AI | ||
| SimpleBench | 27.1% | #63 of 77, top 82% | Epoch AI | ||
| LMArena Hard Prompts | 1255 | LMArena | 2026-10-08 | ||
| LMArena Hard Prompts | 1296 | #191 of 297, top 65% | LMArena | 2026-10-08 | |
| DTBench | 56.5% | Epoch AI | |||
| DTBench | 59% | #119 of 151, top 79% | Epoch AI | ||
| BIG-Bench Hard | 89.2% | Best of 27 | Epoch AI | ||
| BIG-Bench Hard | 84% | Epoch AI | |||
| Epoch Capabilities Index | 131.73 | #135 of 213, top 64% | Epoch AI | 2024-09-24 | |
| Epoch Capabilities Index | 126.91 | Epoch AI | 2024-05-14 | ||
| ForecastBench | 58.4 | #49 of 72, top 69% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 6.8% | Epoch AI | 2025-02-25 | ||
| OTIS Mock AIME 2024-2025 | 23.1% | #123 of 173, top 72% | Epoch AI | 2025-02-25 | |
| Omni-MATH | 36.4% | #32 of 57, top 57% | HELM Capabilities | ||
| LMArena Math | 1315 | #174 of 285, top 62% | LMArena | 2026-10-08 | |
| LMArena Math | 1270 | LMArena | 2026-10-08 | ||
| MATH Level 5 | 40.7% | Epoch AI | 2025-01-27 | ||
| MATH Level 5 | 70.4% | #30 of 79, top 38% | Epoch AI | 2025-01-27 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 45.9% | Epoch AI | 2025-01-27 | ||
| GPQA Diamond | 57.2% | #116 of 186, top 63% | Epoch AI | 2025-01-27 | |
| Humanity's Last Exam | 4.6% | #36 of 41, top 88% | Epoch AI | ||
| MMLU-Pro | 73.7% | #29 of 58, top 50% | HELM Capabilities | ||
| Confabulations (lower is better) | 13.5% | #12 of 51, top 24% | sept | Lech Mazur benchmarks | |
| GPQA (HELM) | 53.4% | #29 of 57, top 51% | HELM Capabilities | ||
| LMArena Expert | 1279 | #184 of 273, top 68% | LMArena | 2026-10-08 | |
| LMArena Expert | 1245 | LMArena | 2026-10-08 | ||
| MMLU | 86.9% | #4 of 81, top 5% | Epoch AI | ||
| MMLU | 82.7% | Epoch AI | |||
| MMLU | 85.9% | Epoch AI |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1161 | #95 of 122, top 78% | LMArena | 2026-10-09 | |
| LMArena Vision | 1088 | LMArena | 2026-10-09 | ||
| Video-MME | 75% | #2 of 15, top 14% | Epoch AI | ||
| Video-MME | 75% | #2 of 15, top 14% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1264 | LMArena | 2026-10-08 | ||
| LMArena Non-English | 1312 | #174 of 297, top 59% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1331 | #176 of 285, top 62% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1275 | LMArena | 2026-10-08 | ||
| LMArena French | 1272 | LMArena | 2026-10-08 | ||
| LMArena French | 1302 | #162 of 223, top 73% | LMArena | 2026-10-08 | |
| LMArena German | 1286 | #156 of 231, top 68% | LMArena | 2026-10-08 | |
| LMArena German | 1249 | LMArena | 2026-10-08 | ||
| LMArena Japanese | 1292 | #125 of 211, top 60% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1239 | LMArena | 2026-10-08 | ||
| LMArena Korean | 1298 | #133 of 213, top 63% | LMArena | 2026-10-08 | |
| LMArena Korean | 1225 | LMArena | 2026-10-08 | ||
| LMArena Russian | 1320 | #168 of 283, top 60% | LMArena | 2026-10-08 | |
| LMArena Russian | 1276 | LMArena | 2026-10-08 | ||
| LMArena Spanish | 1311 | #158 of 226, top 70% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1243 | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| IFEval | 83.7% | #26 of 57, top 46% | HELM Capabilities | ||
| LMArena Instruction Following | 1297 | #184 of 298, top 62% | LMArena | 2026-10-08 | |
| LMArena Instruction Following | 1254 | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1291 | LMArena | 2026-10-08 | ||
| LMArena Longer Query | 1308 | #183 of 291, top 63% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1319 | #181 of 297, top 61% | LMArena | 2026-10-08 | |
| LMArena Text | 1274 | LMArena | 2026-10-08 | ||
| LMArena Creative Writing | 1273 | LMArena | 2026-10-08 | ||
| LMArena Creative Writing | 1333 | #143 of 295, top 49% | LMArena | 2026-10-08 | |
| WildBench | 81.3% | #24 of 57, top 43% | HELM Capabilities | ||
| LMArena Multi-Turn | 1268 | LMArena | 2026-10-08 | ||
| LMArena Multi-Turn | 1296 | #191 of 295, top 65% | LMArena | 2026-10-08 |
Compare Gemini 1.5 Pro (May 2024)
- Gemini 1.5 Pro (May 2024) vs Falcon-180B
- Gemini 1.5 Pro (May 2024) vs Gemma 3 12B
- Gemini 1.5 Pro (May 2024) vs Pixtral Large
- Gemini 1.5 Pro (May 2024) vs Mistral Large
- Gemini 1.5 Pro (May 2024) vs Granite 3.1 8b Instruct
- Gemini 1.5 Pro (May 2024) vs Qwen3-4B
- Gemini 1.5 Pro (May 2024) vs GPT-6 Astra
- Gemini 1.5 Pro (May 2024) vs Claude Fable 5.1
- Gemini 1.5 Pro (May 2024) vs Kimi K3
- Gemini 1.5 Pro (May 2024) vs Grok 4.6
- Gemini 1.5 Pro (May 2024) vs Qwen3.8 Max
- Gemini 1.5 Pro (May 2024) vs GLM-5.3
- Gemini 1.5 Pro (May 2024) vs Muse Spark 1.3
- Gemini 1.5 Pro (May 2024) vs DeepSeek V4 Pro
Other Google models
Frequently asked questions
How good is Gemini 1.5 Pro (May 2024)?
Gemini 1.5 Pro (May 2024) by Google ranks 261st of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.1. Its strongest category is multimodal, where it ranks 77th.
Is Gemini 1.5 Pro (May 2024) open source?
No. Gemini 1.5 Pro (May 2024) is proprietary and available only through Google's API and partner platforms.
What are Gemini 1.5 Pro (May 2024)'s strengths and weaknesses?
Relative to other ranked models, Gemini 1.5 Pro (May 2024) places best in writing & preference, long context, multilingual and lowest in reasoning, agentic & tool use, math.
What is Gemini 1.5 Pro (May 2024) best at?
Its best category is multimodal, where it ranks 77th on Noometry.