Atlas

Benchmarks

← All benchmarks

KernelBench Hard Normalized (RTX PRO 6000)

Code · 2026-09-30

KernelBench Hard on RTX PRO 6000: 100 times mean cell score divided by the current board-best valid score for each of six problems. Missing, failed and unaudited cells contribute zero. The publisher selects best cells across multiple harness routes/settings for each model, so this is a portfolio aggregate, not a single run. Distinct from the earlier mean raw roofline fraction.

Top models (higher is better)

ModelScore
Kimi K376.5
DeepSeek V4.1 Flash69.6
Claude Fable 565.5
GLM-5.263.7
Grok 4.660.4
GPT-5.6 Sol57.8
Grok 4.547.5
GLM-5.347.1
DeepSeek-V4-Flash-073140.4
Opus 4.837.4
Claude Opus 533.2
LongCat-2.021.3
MiniMax M319.8
Opus 4.70.0
Sonnet 50.0
Composer 2.50.0
DeepSeek-V4-Flash0.0
Fugu Ultra v1.00.0
Gemini 3.5 Flash0.0
GLM-5.10.0
GPT-5.50.0
Kimi K2.60.0
Kimi K2.7 Code0.0
MiMo-V2.5-Pro0.0
Qwen3.6-Max-Preview0.0
Qwen3.6 Plus (2026-04-02)0.0
Qwen3.7-Max0.0
Loading Atlas data…