Atlas

Benchmarks

← All benchmarks

AutomationBench (Support)

Agents · 2026-04-20

This AutomationBench subtrack evaluates agents on end-to-end support workflows built with Zapier applications. Scores report the percentage of tasks completed successfully.

Top models (higher is better)

ModelScore
Claude Opus 522.0
Claude Fable 517.0
Loading Atlas data…