Atlas

Benchmarks

← All benchmarks

Harvey HLAB - Criteria Pass Rate - Employment Labor

Professional Work · 2026-05-06

The Employment Labor practice-area split of Harvey's Legal Agent Benchmark held-out evaluation criteria pass rate. This reports pooled rubric criteria passed within that split and is distinct from the all-pass task-score split.

Top models (higher is better)

ModelScore
Claude Fable 587.6
Kimi K387.6
Muse Spark 1.187.3
Grok 4.585.4
Opus 4.882.8
Claude Opus 582.4
Sonnet 582.0
GLM-5.281.3
Sonnet 4.678.7
MiniMax M377.9
Gemini 3.6 Flash76.8
GPT-5.6 Luna76.8
Qwen3.7-Max76.8
GPT-5.576.4
DeepSeek-V4-Pro74.9
GPT-5.6 Terra73.4
Kimi K2.673.0
GLM-5.171.9
Gemini 3.5 Flash70.4
GPT-5.6 Sol67.0
Inkling67.0
Gemini 3.5 Flash-Lite65.5
GPT-5.464.8
Laguna XS.261.5
Qwen3.7-Plus56.2
Grok 4.355.4
Laguna M.149.4
Loading Atlas data…