all the models — AI benchmark observatory
← Benchmarks

EQ-Bench

EQ-Bench is an LLM-judged test evaluating active emotional intelligence abilities, understanding, insight, empathy, and interpersonal skills. The test set contains 45 challenging roleplay scenarios, most of which constitute pre-written prompts spanning 3 turns. The benchmark evaluates the performance of models by validating responses against several criteria and conducts pairwise comparisons to report a normalized Elo computation for each model.

id eq-bench · max 2000 · 2 models reported

#ModelScore
1Grok-4.1 Thinking
xAI
1586.00
2Grok-4.1
xAI
1585.00
EQ-Bench Leaderboard · all the models