Writing & Preference benchmark

EQ-Bench 4 leaderboard

As of October 2026, Claude Opus 5 has the highest published EQ-Bench 4 score on Noometry at 1385, out of 28 models with results.

Last verified

About EQ-Bench 4

Emotional intelligence in multi-turn role-play conversations, such as building rapport and handling conflict, rated by pairwise LLM-judged comparisons.

Category
Writing & Preference
Introduced
2026
Format
Pairwise LLM-judged conversations
Unit
Arena rating (Bradley–Terry)
Official site
eqbench.com

Top 15 models

Top models on EQ-Bench 4
  1. Claude Opus 5 1385
  2. Claude Fable 5 1340
  3. Kimi K3 1339
  4. GPT-5.5 1315
  5. Claude Opus 4.7 1311
  6. Claude Opus 4.8 1281
  7. GPT-5.4 1272
  8. Muse Spark 1.1 1260
  9. GPT-5.6 Sol 1250
  10. Claude Sonnet 5 1236
  11. GPT-5.6 Terra 1234
  12. Inkling 1226
  13. Claude Opus 4.6 1223
  14. GLM-5.2 1222
  15. MiMo-V2.5-Pro 1208

Sponsored placements are available on pages like this one. Advertise on Noometry

All results

Compare the leaders

Other writing & preference benchmarks

Frequently asked questions

What does EQ-Bench 4 measure?

Emotional intelligence in multi-turn role-play conversations, such as building rapport and handling conflict, rated by pairwise LLM-judged comparisons.

Which model has the highest EQ-Bench 4 score?

As of October 2026, Claude Opus 5 has the highest published EQ-Bench 4 score on Noometry at 1385, out of 28 models with results.

What is the best open-weight model on EQ-Bench 4?

Kimi K3 has the highest EQ-Bench 4 rating among open-weight models at 1339, ranking 3 of 28 overall.