Atlas

Benchmarks

← All benchmarks

BoolQ

Classic NLP · 2019-05-24

A yes/no reading comprehension benchmark where models answer naturally occurring questions given a short supporting passage.

Top models (higher is better)

ModelScore
Gemini Nano 279.3
Gemini Nano 171.6
Loading Atlas data…