Atlas

Benchmarks

← All benchmarks

FrontierFinance v3.1 — Financial Data and Modeling

Professional Work · 2026-07-06

Financial Data and Modeling use-case slice of Samaya AI's FrontierFinance v3.1, containing 70 of the 220 public queries. It reports rubric qualification rate macro-averaged across queries in the slice.

Top models (higher is better)

ModelScore
Claude Fable 555.6
Opus 4.851.0
GPT-5.6 Sol49.0
GPT-5.547.6
GLM-5.247.5
DeepSeek-V4-Pro45.3
Gemini 3.1 Pro Preview31.4
Kimi K2.628.8
Loading Atlas data…