Atlas

Models

← All models

Grok 4.3

xAI · 2026-04-30

Grok 4.3 is xAI's flagship model released in April 2026, with a 1 million-token context window and configurable reasoning. xAI's current docs describe it as their most advanced flagship model, focused on low hallucination rates, agentic tool calling, and instruction following, with text and image inputs and text output.

Benchmark scores

BenchmarkScore
AA-Briefcase Elo759
AA-LCR65.0
AA-Omniscience Index18.3
Agent Red Teaming k=100 (CAIS)90.1
Agents' Last Exam (ALE)6.6
Agents' Last Exam ALE-CLI Pass Rate7.6
Agents' Last Exam ALE-CLI Score24.3
Agents' Last Exam Full-Spectrum Pass Rate7.3
Agents' Last Exam Full-Spectrum Score17.0
Agents' Last Exam Last-Exam Pass Rate0.0
Agents' Last Exam Last-Exam Score2.3
Agents' Last Exam Near-Term Pass Rate9.0
Agents' Last Exam Near-Term Score30.4
Agents' Last Exam Overall Score20.1
Agents' Last Exam Unlicensed Full-Spectrum Pass Rate8.0
Agents' Last Exam Unlicensed Full-Spectrum Score18.5
Agents' Last Exam Unlicensed Last-Exam Pass Rate0.0
Agents' Last Exam Unlicensed Last-Exam Score3.0
Agents' Last Exam Unlicensed Near-Term Pass Rate9.0
Agents' Last Exam Unlicensed Near-Term Score30.4
Agents' Last Exam Unlicensed Overall Pass Rate7.1
Agents' Last Exam Unlicensed Overall Score21.6
AHK-Eval86.1
AHK-Eval Algorithms83.3
AHK-Eval Data Structures50.0
AHK-Eval Date and Time83.3
AHK-Eval Easy91.7
AHK-Eval Hard58.3
AHK-Eval Hidden Cases80.1
AHK-Eval Mid83.3
Loading Atlas data…