Atlas

Benchmarks

← All benchmarks

EQ-Bench 4: Emotion Management

Chat & Writing · 2026-07-23

EQ-Bench 4's independently fitted emotion-management ability rating measures how well a model helps users regulate, tolerate, or constructively engage with emotion. The source normalizes its per-dimension soft Bradley–Terry rating to a 1–10 scale.

Top models (higher is better)

ModelScore
Claude Opus 510.0
GPT-5.59.4
Claude Fable 59.0
Kimi K38.7
Opus 4.78.4
GPT-5.48.0
GPT-5.6 Sol7.7
Opus 4.87.5
Muse Spark 1.17.4
GPT-5.6 Terra7.3
Sonnet 56.5
Inkling6.1
Opus 4.65.9
GLM-5.25.9
MiMo-V2.5-Pro5.8
GPT-5.6 Luna5.5
Sonnet 4.65.4
Kimi K2.65.2
Gemini 3.1 Pro Preview5.0
DeepSeek-V4-Pro5.0
MiniMax M34.2
Qwen3.7-Max4.0
Gemma 4 31B IT3.9
Gemini 3.5 Flash3.7
Grok 4.33.0
Haiku 4.51.9
Qwen3.6 27B1.7
Mistral Medium 3.51.0
Loading Atlas data…