Atlas

Benchmarks

← All benchmarks

Vals Index Terminal-Bench 2.1

Indexes · 2026-07-01

This row reports the component score for Terminal-Bench 2.1 within the Vals Index composite.

Top models (higher is better)

ModelScore
GPT-5.6 Sol85.8
Claude Opus 584.6
Kimi K380.9
Claude Fable 580.5
GPT-5.6 Luna79.0
GPT-5.576.4
Sonnet 574.5
Gemini 3.5 Flash74.2
Gemini 3.6 Flash73.8
GPT-5.6 Terra73.4
Opus 4.871.9
Gemini 3.1 Pro Preview70.8
Muse Spark 1.169.3
Opus 4.768.5
GLM-5.267.8
Grok 4.567.8
Qwen3.7-Max61.0
MiMo-V2.560.7
Sonnet 4.657.3
MiMo-V2.5-Pro57.3
GLM-5.156.9
Inkling-Small55.1
GPT-5.4 Mini54.7
Gemini 3 Flash Preview53.9
Kimi K2.653.6
MiniMax M353.6
Qwen3.6 Plus (2026-04-02)53.2
Qwen3.7-Plus52.8
Nemotron 3 Ultra 550B A55B50.9
DeepSeek-V4-Pro50.2
Gemini 3.5 Flash-Lite50.2
MiniMax M2.748.7
Inkling47.6
Grok 4.2044.2
Haiku 4.543.8
Grok 4.341.9
GPT-5.4 Nano41.6
Mistral Medium 3.539.0
Gemini 3.1 Flash-Lite Preview34.1
Laguna M.134.1
Loading Atlas data…