Atlas

Models

← All models

Grok 4.20

xAI · 2026-03-10

Grok 4.20 is xAI's high-performance model released in March 2026, with a 1 million-token context window and agentic tool-calling capabilities. The current xAI docs identify the reasoning model as grok-4.20-0309-reasoning and list aliases including grok-4.20, while xAI's release notes state Grok 4.20 and Grok 4.20 Multi-agent went live on March 10, 2026.

Benchmark scores

BenchmarkScore
A-Fantasia73.0
A-Fantasia Backwards Spelling82.0
A-Fantasia Chess63.0
A-Fantasia Cube Rotation74.0
AA-LCR59.0
AA-Omniscience Index15.3
Agent Red Teaming k=100 (CAIS)84.6
AIME 2024 (avg@8)98.8
AIME 2024-202596.5
AIME 2025 (avg@8)94.2
ALE-Bench1388
Antidote: Everyday Edition1000
APEX-Agents-AA14.2
ARC-AGI-189.5
ARC-AGI-3 (Semi-Private)0.1
Artificial Analysis Intelligence Index37.0
Artificial Analysis Omniscience Accuracy28.9
Artificial Analysis Omniscience Hallucination Rate16.7
Bench to the Future 20.2
Bench to the Future 2 — Calibration Error0.0
Bench to the Future 2 — Refinement0.0
BioSecBench-Refusal3.1
BioSecBench-Surveillance13.7
BioTIER Refusal71.6
Blueprint-Bench 20.0
BullshitBench v2 Leaderboard67.0
BuseyBench SVG3.8
Buyout Game Bradley-Terry Rating1549
CadQueryEval24.0
CadQueryEval — Bounding Box36.0
Loading Atlas data…