Writing & Preference benchmark
EQ-Bench 4 leaderboard
As of October 2026, Claude Opus 5 has the highest published EQ-Bench 4 score on Noometry at 1385, out of 28 models with results.
Last verified
About EQ-Bench 4
Emotional intelligence in multi-turn role-play conversations, such as building rapport and handling conflict, rated by pairwise LLM-judged comparisons.
- Category
- Writing & Preference
- Introduced
- 2026
- Format
- Pairwise LLM-judged conversations
- Unit
- Arena rating (Bradley–Terry)
- Official site
- eqbench.com
Top 15 models
- Claude Opus 5 1385
- Claude Fable 5 1340
- Kimi K3 1339
- GPT-5.5 1315
- Claude Opus 4.7 1311
- Claude Opus 4.8 1281
- GPT-5.4 1272
- Muse Spark 1.1 1260
- GPT-5.6 Sol 1250
- Claude Sonnet 5 1236
- GPT-5.6 Terra 1234
- Inkling 1226
- Claude Opus 4.6 1223
- GLM-5.2 1222
- MiMo-V2.5-Pro 1208
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Other writing & preference benchmarks
Frequently asked questions
What does EQ-Bench 4 measure?
Emotional intelligence in multi-turn role-play conversations, such as building rapport and handling conflict, rated by pairwise LLM-judged comparisons.
Which model has the highest EQ-Bench 4 score?
As of October 2026, Claude Opus 5 has the highest published EQ-Bench 4 score on Noometry at 1385, out of 28 models with results.
What is the best open-weight model on EQ-Bench 4?
Kimi K3 has the highest EQ-Bench 4 rating among open-weight models at 1339, ranking 3 of 28 overall.