Atlas

Benchmarks

← All benchmarks

Harvey HLAB - Criteria Pass Rate - Immigration

Professional Work · 2026-05-06

The Immigration practice-area split of Harvey's Legal Agent Benchmark held-out evaluation criteria pass rate. This reports pooled rubric criteria passed within that split and is distinct from the all-pass task-score split.

Top models (higher is better)

ModelScore
Muse Spark 1.193.2
Kimi K390.9
Claude Fable 590.2
Claude Opus 590.2
Grok 4.589.1
GLM-5.288.0
Sonnet 586.9
Sonnet 4.686.5
Opus 4.885.8
GLM-5.185.5
DeepSeek-V4-Pro84.4
GPT-5.6 Sol84.4
Inkling84.4
MiniMax M382.2
Gemini 3.6 Flash81.8
GPT-5.581.8
Qwen3.7-Max80.0
GPT-5.6 Luna78.5
Kimi K2.678.2
Gemini 3.5 Flash76.7
GPT-5.6 Terra73.8
GPT-5.472.4
Gemini 3.5 Flash-Lite70.5
Grok 4.366.2
Laguna M.162.9
Qwen3.7-Plus62.5
Laguna XS.235.1
Loading Atlas data…