BoolQ
Classic NLP · 2019-05-24
A yes/no reading comprehension benchmark where models answer naturally occurring questions given a short supporting passage.
Top models (higher is better)
| Model | Score |
|---|---|
| Gemini Nano 2 | 79.3 |
| Gemini Nano 1 | 71.6 |
Classic NLP · 2019-05-24
A yes/no reading comprehension benchmark where models answer naturally occurring questions given a short supporting passage.
| Model | Score |
|---|---|
| Gemini Nano 2 | 79.3 |
| Gemini Nano 1 | 71.6 |