Arena Text Overall — 25% Factuality
Chat & Writing · 2026-07-15
Arena's Text rating with the site's default 25% factuality adjustment applied to the overall raw leaderboard without style control. It is kept separate because factuality weighting changes both ratings and ranks.
Top models (higher is better)
| Model | Score |
|---|---|
| Claude Opus 5 | 1493 |
| Opus 4.6 | 1493 |
| Claude Fable 5 | 1487 |
| GPT-5.4 | 1485 |
| GPT-5.5 | 1485 |
| Gemini 3 Pro Preview | 1481 |
| Qwen3.7 Max Preview | 1480 |
| Gemini 3.5 Flash | 1477 |
| Opus 4.7 | 1476 |
| Gemini 3.6 Flash | 1473 |
| Opus 4.8 | 1472 |
| Kimi K3 | 1472 |
| GPT-5.6 Sol | 1471 |
| Gemini 3.1 Pro Preview | 1471 |
| Gemini 3 Flash Preview | 1467 |
| Grok 4.5 | 1466 |
| Muse Spark | 1466 |
| Muse Spark 1.1 | 1466 |
| GPT-5.6 Terra | 1466 |
| MiMo-V2.5-Pro | 1463 |
| Gemini 2.5 Pro | 1462 |
| Sonnet 4.6 | 1460 |
| GLM-5.1 | 1460 |
| GLM-5.2 | 1459 |
| GPT-5.1 | 1458 |
| Opus 4.5 | 1458 |
| Kimi K2.6 | 1455 |
| Qwen3.7-Plus | 1455 |
| DeepSeek-V4-Pro | 1454 |
| ERNIE 5.1 | 1454 |
| GPT-5.6 Luna | 1454 |
| Sonnet 5 | 1454 |
| GLM-5 | 1451 |
| Inkling | 1451 |
| GPT-5.4 Mini | 1449 |
| Gemma 4 31B | 1449 |
| Hy3 | 1449 |
| Sonnet 4.5 | 1448 |
| GLM 4.6 | 1448 |
| DeepSeek-V3.2-Exp | 1447 |