Atlas

Benchmarks

← All benchmarks

EQ-Bench 4: Attunement

Chat & Writing · 2026-07-23

EQ-Bench 4's independently fitted attunement ability rating measures how well a model reads the user's presented and inner state. The source normalizes its per-dimension soft Bradley–Terry rating to a 1–10 scale.

Top models (higher is better)

ModelScore
Claude Opus 510.0
Claude Fable 58.9
Kimi K38.8
Opus 4.78.4
GPT-5.58.4
Opus 4.87.7
GPT-5.47.5
Muse Spark 1.17.5
GPT-5.6 Sol7.0
Sonnet 56.7
GPT-5.6 Terra6.7
Inkling6.5
GLM-5.26.2
Opus 4.66.2
Sonnet 4.66.0
MiMo-V2.5-Pro5.9
Kimi K2.65.9
DeepSeek-V4-Pro5.1
GPT-5.6 Luna4.9
MiniMax M34.7
Gemini 3.1 Pro Preview4.5
Gemma 4 31B IT4.1
Qwen3.7-Max3.9
Gemini 3.5 Flash3.3
Grok 4.33.0
Haiku 4.52.5
Qwen3.6 27B1.6
Mistral Medium 3.51.0
Loading Atlas data…