Atlas

Benchmarks

← All benchmarks

FrontierFinance v3.1 — Screening and Discovery

Professional Work · 2026-07-06

Screening and Discovery use-case slice of Samaya AI's FrontierFinance v3.1, containing 17 of the 220 public queries. It reports rubric qualification rate macro-averaged across queries in the slice.

Top models (higher is better)

ModelScore
GPT-5.6 Sol36.5
GPT-5.534.8
Claude Fable 533.3
Opus 4.830.2
DeepSeek-V4-Pro28.1
Kimi K2.627.8
Gemini 3.1 Pro Preview27.2
GLM-5.224.9
Loading Atlas data…