Coding benchmark
GSO leaderboard
As of October 2026, Claude Fable 5.1 has the highest published GSO score on Noometry at 88.2%, out of 31 models with results.
Last verified
About GSO
Software optimization tasks: the model must speed up real code bases while keeping them correct.
- Category
- Coding
- Introduced
- 2025
- Format
- Performance patch
- Unit
- Percent (random guessing ≈ 0%)
- Official site
- gso-bench.github.io
Top 15 models
- Claude Fable 5.1 88.2%
- GPT-6 Astra 79.4%
- Claude Fable 5 78.4%
- GPT-5.6 Sol 76.5%
- Claude Opus 4.8 47.1%
- Claude Opus 4.7 44.1%
- Claude Opus 4.6 41.2%
- GPT-5.5 40.2%
- Claude Sonnet 5 37.3%
- GPT-5.4 31.4%
- GPT-5.2 27.4%
- Claude Opus 4.5 26.5%
- Gemini 3.1 Pro Preview 22.6%
- Gemini 3 Pro 18.6%
- Claude Sonnet 4.5 14.7%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Frequently asked questions
What does GSO measure?
Software optimization tasks: the model must speed up real code bases while keeping them correct.
Which model has the highest GSO score?
As of October 2026, Claude Fable 5.1 has the highest published GSO score on Noometry at 88.2%, out of 31 models with results.
What is the best open-weight model on GSO?
Kimi K2 (Jul 2025) has the highest GSO accuracy among open-weight models at 4.9%, ranking 22 of 31 overall.