Google, proprietary

Gemini 3 Pro

Gemini 3 Pro by Google ranks 28th of 354 ranked models on the Noometry Index as of October 2026, with a score of 54.8. Its strongest category is multimodal, where it ranks 2nd.

Last verified

Specifications

Noometry rank
#28 of 354
Index score
54.8
Evidence
Confirmed 67 results
Provider
Google
Released
November 18, 2025
Weights
Proprietary
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
1 tokens/s Kagi
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Gemini 3 Pro category scores
  1. Coding 51.6
  2. Agentic & Tool Use 40.6
  3. Reasoning 52.5
  4. Math 49.9
  5. Knowledge 64.4
  6. Multimodal 57.6
  7. Multilingual 56.9
  8. Instruction Following 76.3
  9. Long Context 44.0
  10. Writing & Preference 66.4
Gemini 3 Pro category ranks
CategoryScoreRankResults
Coding51.6#397
Agentic & Tool Use40.6#2310
Reasoning52.5#319
Math49.9#595
Knowledge64.4#166
Multimodal57.6#23
Multilingual56.9#161
Instruction Following76.3#452
Long Context44.0#792
Writing & Preference66.4#355

Strengths and weaknesses

Categories where Gemini 3 Pro places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Gemini 3 Pro: strongest categories
CategoryScorevs medianRank
Multimodal57.6+19.0#2 of 128, top 2%
Knowledge64.4+27.1#16 of 314, top 6%
Multilingual56.9+9.5#16 of 297, top 6%

Weakest categories

Gemini 3 Pro: weakest categories
CategoryScorevs medianRank
Long Context44.0+3.1#79 of 296, top 27%
Math49.9+13.3#59 of 327, top 19%
Agentic & Tool Use40.6+10.3#23 of 154, top 15%

Closest competitors

The models ranked just above and below Gemini 3 Pro. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Gemini 3 Pro
ModelRankScoreBlended $/MSpeed
Gemini 4 Argon#2456.5——Compare
Grok 4.5#2555.0$34Compare
GLM-5.3#2654.8$2.15—Compare
Muse Spark 1.3#2754.8$2—Compare
Claude Sonnet 5#2954.6$4—Compare
GPT-5.6 Luna#3054.6$0.4512Compare
DeepSeek V4 Pro#3154.3$0.9916Compare
Gemini 3.5 Flash#3254.2$3.38—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Gemini 3 Pro Coding benchmark results
BenchmarkScorePositionSettingSourceDate
SWE-bench Verified72.9%#21 of 32, top 66%Epoch AI2026-02-13
SWE-bench Verified (bash only)74.2%#5 of 39, top 13%SWE-bench2025-11-18
LMArena WebDev1440#58 of 113, top 52%LMArena2026-10-08
SWE-bench Multilingual68.7%#5 of 13, top 39%SWE-bench2026-02-13
GSO18.6%#14 of 31, top 46%Epoch AI
WeirdML69.9%#19 of 119, top 16%Epoch AI
LMArena Coding1481#45 of 294, top 16%LMArena2026-10-08
ALE-Bench1,177#33 of 105, top 32%Epoch AI
AlgoTune1.83#4 of 18, top 23%Epoch AI

Agentic & Tool Use

Gemini 3 Pro Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Terminal-Bench69.4%#7 of 41, top 18%Epoch AI
Berkeley Function Calling Leaderboard72.5%#3 of 49, top 7%promptBerkeley Function Calling Leaderboard
GDPval40.3%#5 of 11, top 46%Epoch AI
Remote Labor Index1.3%#13 of 14, top 93%Epoch AI
τ²-bench Airline80.5%#6 of 7, top 86%highτ²-bench2026-03-02
τ²-bench Banking18%#20 of 26, top 77%highτ²-bench2026-03-02
τ²-bench Retail75.9%#5 of 7, top 72%highτ²-bench2026-03-02
τ²-bench Telecom91%#4 of 7, top 58%highτ²-bench2026-03-02
DeepResearch Bench46.3%#14 of 24, top 59%lowEpoch AI
BALROG58.1%#4 of 35, top 12%Epoch AI
LMArena Search1207#10 of 32, top 32%LMArena2026-08-24
METR Time Horizons71%#8 of 32, top 25%Epoch AI
Vending-Bench 25,478#24 of 60, top 40%Epoch AI

Reasoning

Gemini 3 Pro Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
ARC-AGI-231.1%#40 of 83, top 49%Epoch AI
SimpleBench76.4%#6 of 77, top 8%Epoch AI
Kagi LLM Benchmark80.1%#8 of 99, top 9%Kagi LLM Benchmark
NYT Connections (extended)94.4%#6 of 91, top 7%Lech Mazur benchmarks
ARC-AGI-175%#41 of 83, top 50%Epoch AI
CritPt6.9%#55 of 134, top 42%Epoch AI
Chess Puzzles31%#33 of 129, top 26%Epoch AI2025-12-08
EnigmaEval18.2%#8 of 38, top 22%Epoch AI
LMArena Hard Prompts1480#33 of 297, top 12%LMArena2026-10-08
Epoch Capabilities Index152.92#42 of 213, top 20%Epoch AI2025-11-18
ForecastBench61.2#14 of 72, top 20%Epoch AI

Math

Gemini 3 Pro Math benchmark results
BenchmarkScorePositionSettingSourceDate
MathArena Final-Answer Competitions67%#18 of 29, top 63%previewMathArena
OTIS Mock AIME 2024-202591.4%#47 of 173, top 28%Epoch AI2025-11-19
ProofBench20%#49 of 77, top 64%Epoch AI
Omni-MATH55.5%#13 of 57, top 23%HELM Capabilities
LMArena Math1476#31 of 285, top 11%LMArena2026-10-08
FrontierMath (Feb 2025 set)37.6%#12 of 68, top 18%Epoch AI2025-11-21
FrontierMath Tier 4 (v1)18.8%#11 of 55, top 20%Epoch AI2025-11-21

Knowledge

Gemini 3 Pro Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond92.6%#22 of 186, top 12%Epoch AI2025-11-19
Humanity's Last Exam37.5%#7 of 41, top 18%Epoch AI
MMLU-Pro90.3%Best of 58HELM Capabilities
Vectara Hallucination Rate (lower is better)13.6%#82 of 96, top 86%Vectara Hallucination Leaderboard
GPQA (HELM)80.3%Best of 57HELM Capabilities
LMArena Expert1475#47 of 273, top 18%LMArena2026-10-08

Multimodal

Gemini 3 Pro Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1305#13 of 122, top 11%LMArena2026-10-09
GeoBench84%#3 of 25, top 12%Epoch AI
VPCT91%Best of 24Epoch AI
LMArena Document1434#28 of 38, top 74%LMArena2026-09-13

Multilingual

Gemini 3 Pro Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1474#16 of 297, top 6%LMArena2026-10-08
LMArena Chinese1523#23 of 285, top 9%LMArena2026-10-08
LMArena French1492#18 of 223, top 9%LMArena2026-10-08
LMArena German1515#2 of 231, top 1%LMArena2026-10-08
LMArena Japanese1510#4 of 211, top 2%LMArena2026-10-08
LMArena Korean1448#18 of 213, top 9%LMArena2026-10-08
LMArena Russian1493#12 of 283, top 5%LMArena2026-10-08
LMArena Spanish1470#25 of 226, top 12%LMArena2026-10-08

Instruction Following

Gemini 3 Pro Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
IFEval87.7%#13 of 57, top 23%HELM Capabilities
LMArena Instruction Following1458#36 of 298, top 13%LMArena2026-10-08

Long Context

Gemini 3 Pro Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
CL-bench15.8%#15 of 19, top 79%Epoch AI
LMArena Longer Query1471#33 of 291, top 12%LMArena2026-10-08

Writing & Preference

Gemini 3 Pro Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1479#16 of 297, top 6%LMArena2026-10-08
LMArena Creative Writing1482#10 of 295, top 4%LMArena2026-10-08
EQ-Bench Creative Writing1525#52 of 115, top 46%EQ-Bench
WildBench85.9%#5 of 57, top 9%HELM Capabilities
LMArena Multi-Turn1484#15 of 295, top 6%LMArena2026-10-08

Compare Gemini 3 Pro

Other Google models

Frequently asked questions

How good is Gemini 3 Pro?

Gemini 3 Pro by Google ranks 28th of 354 ranked models on the Noometry Index as of October 2026, with a score of 54.8. Its strongest category is multimodal, where it ranks 2nd.

Is Gemini 3 Pro open source?

No. Gemini 3 Pro is proprietary and available only through Google's API and partner platforms.

How fast is Gemini 3 Pro?

Gemini 3 Pro generated about 1 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are Gemini 3 Pro's strengths and weaknesses?

Relative to other ranked models, Gemini 3 Pro places best in multimodal, knowledge, multilingual and lowest in long context, math, agentic & tool use.

What is Gemini 3 Pro best at?

Its best category is multimodal, where it ranks 2nd on Noometry.