OLMo 2 32B Instruct
Allen Institute for AI (Ai2) · 2025-03-13 · 32.2B parameters
OLMo 2 32B Instruct is Ai2's fully open post-trained variant of OLMo 2 32B, produced through supervised fine-tuning, DPO, and final reinforcement learning with verifiable rewards using a Tülu 3.1-style recipe. Ai2 released its weights, training data and code, logs, and intermediate RLVR checkpoints.