Math benchmark
FrontierMath (Tiers 1-3) leaderboard
As of October 2026, GPT-6.1 Sol has the highest published FrontierMath (Tiers 1-3) score on Noometry at 93.7%, out of 81 models with results.
Last verified
About FrontierMath (Tiers 1-3)
Unpublished, expert-written mathematics problems ranging from hard undergraduate to research level, with automatically checkable answers. Run privately by Epoch AI.
Top 15 models
- GPT-6.1 Sol 93.7%
- GPT-6 Astra 93.7%
- Claude Opus 5.5 91.2%
- Claude Fable 5.1 90.2%
- GPT-6 Sol 89.8%
- GPT-5.6 Sol 89.1%
- Claude Sonnet 5.5 88.8%
- GPT-5.5 Pro 87.7%
- Claude Fable 5 87%
- GPT-5.6 Terra 86%
- Claude Opus 5 85.6%
- GPT-5.5 85.3%
- GPT-5.4 Pro 82.5%
- GPT-5.6 Luna 82.1%
- Claude Opus 4.8 80%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Other math benchmarks
Frequently asked questions
What does FrontierMath (Tiers 1-3) measure?
Unpublished, expert-written mathematics problems ranging from hard undergraduate to research level, with automatically checkable answers. Run privately by Epoch AI.
Which model has the highest FrontierMath (Tiers 1-3) score?
As of October 2026, GPT-6.1 Sol has the highest published FrontierMath (Tiers 1-3) score on Noometry at 93.7%, out of 81 models with results.
What is the best open-weight model on FrontierMath (Tiers 1-3)?
Kimi K3 has the highest FrontierMath (Tiers 1-3) accuracy among open-weight models at 72.2%, ranking 22 of 81 overall.