Math benchmark
Omni-MATH leaderboard
As of October 2026, GPT-5 Mini has the highest published Omni-MATH score on Noometry at 72.2%, out of 57 models with results.
Last verified
About Omni-MATH
Olympiad-level problems across 33 sub-domains and ten difficulty levels. HELM Capabilities run.
- Category
- Math
- Introduced
- 2024
- Size
- 4,428 problems
- Format
- Final answer
- Unit
- Percent (random guessing ≈ 0%)
- Official site
- crfm.stanford.edu
Top 15 models
- GPT-5 Mini 72.2%
- o4-mini 72%
- Qwen3 235B-A22B 71.8%
- o3 71.4%
- gpt-oss-120b 68.8%
- Kimi K2 (Jul 2025) 65.4%
- GPT-5 64.7%
- Claude Opus 4 61.6%
- Grok 4 60.3%
- Claude Sonnet 4 60.2%
- gpt-oss-20b 56.5%
- Claude Haiku 4.5 56.1%
- Gemini 3 Pro 55.5%
- Claude Sonnet 4.5 55.3%
- GPT-5 Nano 54.6%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Other math benchmarks
Frequently asked questions
What does Omni-MATH measure?
Olympiad-level problems across 33 sub-domains and ten difficulty levels. HELM Capabilities run.
Which model has the highest Omni-MATH score?
As of October 2026, GPT-5 Mini has the highest published Omni-MATH score on Noometry at 72.2%, out of 57 models with results.
What is the best open-weight model on Omni-MATH?
Qwen3 235B-A22B has the highest Omni-MATH accuracy among open-weight models at 71.8%, ranking 3 of 57 overall.