Llama-3.1-Tülu-3-8B
Allen Institute for AI (Ai2) · 2024-11-21 · 8B parameters
Llama-3.1-Tulu-3-8B is Ai2's post-trained 8B Llama 3.1 model. It was produced with the fully open Tülu 3 recipe, combining supervised fine-tuning, Direct Preference Optimization, and reinforcement learning with verifiable rewards.