IMO 2025
Math · 2025-07-17
This MathArena track evaluates models on the IMO 2025 problem set. Scores report the percentage of problems answered correctly.
Top models (higher is better)
| Model | Score |
|---|---|
| Gemini 2.5 Deep Think | 60.7 |
| GPT-5 | 38.1 |
| Gemini 2.5 Pro | 31.6 |
| Grok 4 | 21.4 |
| o3 | 16.7 |
| O4 Mini | 14.3 |
| DeepSeek-R1-0528 | 6.8 |