Atlas

Benchmarks

← All benchmarks

KernelBench Mega Normalized (RTX PRO 6000)

Code · 2026-09-30

KernelBench Mega RTX PRO 6000 board: best audited Kimi linear-decode cell relative to the current board-best valid cell, scaled to 0–100. Failed/invalid cells contribute zero. The retired RL-grid problem is absent. Best cells can combine harness routes/settings; distinct from the earlier absolute speedup metric.

Top models (higher is better)

ModelScore
Claude Opus 5.5100.0
Claude Sonnet 5.590.6
GPT-6 Astra Pro70.0
Claude Fable 569.4
Claude Fable 5.164.7
GLM-5.354.8
Kimi K351.0
DeepSeek V4.1 Flash48.2
Opus 4.840.6
GLM-5.231.4
Grok 4.718.0
GPT-5.512.2
Sonnet 511.4
Gemini 3.8 Flash7.7
GPT-5.6 Sol7.4
MiniMax M37.4
Kimi K2.7 Code7.3
GPT-6 Luna7.0
Composer 2.57.0
Gemini 3.5 Flash6.4
Muse Spark 1.36.2
GPT-6 Sol2.7
Grok 4.52.3
DeepSeek-V4-Flash-07310.0
Grok 4.60.0
LongCat-2.00.0
Loading Atlas data…