Math benchmark
FrontierMath Tier 4 leaderboard
As of October 2026, GPT-6.1 Sol has the highest published FrontierMath Tier 4 score on Noometry at 100%, out of 63 models with results.
Last verified
About FrontierMath Tier 4
The hardest FrontierMath tier: research-level problems written by professional mathematicians.
Top 15 models
- GPT-6.1 Sol 100%
- GPT-6 Astra 97.6%
- Claude Opus 5.5 95%
- Claude Fable 5 90.2%
- GPT-6 Sol 90%
- Claude Fable 5.1 87.8%
- GPT-5.6 Sol 82.9%
- Claude Sonnet 5.5 80.5%
- GPT-5.5 Pro 78%
- AI Co-Mathematician 75.6%
- Claude Opus 5 73.2%
- GPT-5.5 72.5%
- GPT-5.6 Terra 70.7%
- GPT-5.6 Luna 61%
- GPT-5.4 Pro 58.5%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Other math benchmarks
Frequently asked questions
What does FrontierMath Tier 4 measure?
The hardest FrontierMath tier: research-level problems written by professional mathematicians.
Which model has the highest FrontierMath Tier 4 score?
As of October 2026, GPT-6.1 Sol has the highest published FrontierMath Tier 4 score on Noometry at 100%, out of 63 models with results.
What is the best open-weight model on FrontierMath Tier 4?
Kimi K3 has the highest FrontierMath Tier 4 accuracy among open-weight models at 39%, ranking 23 of 63 overall.