Hermes 3 Llama 3.1 405B
Nous Research · 2024-08-13 · 405.9B parameters
Hermes 3 Llama 3.1 405B is Nous Research's first full-parameter fine-tune of Meta's Llama 3.1 405B, released in August 2024. It was fine-tuned on an approximately 390-million-token instruction dataset with primarily synthetically generated responses, retains the base model's 131,072-token context window, and adds agentic function calling and structured outputs. Because the dataset size is approximate and the base model's cumulative pretraining exposure is not exact, no exact training-token count is recorded.
Benchmark scores
| Benchmark | Score |
|---|---|
| BuseyBench SVG | 1.5 |
| Capability | 124.6 |
| EQ-Bench v2 | 82.8 |
| EQ-Bench v2 + MAGI-Hard Combined | 79.5 |
| MAGI-Hard | 76.2 |
| SnakeBench Average Apples | 2.0 |
| SnakeBench Best Apples | 5.0 |
| SnakeBench Rating | 20.7 |
| SnakeBench Total Apples | 36.0 |
| SnakeBench Win Rate | 43.8 |