Olmo 3 32B Think
Allen Institute for AI (Ai2) · 2025-11-20 · 32.2B parameters
Olmo 3 32B Think is Ai2's fully open 32B reasoning checkpoint, built from Olmo 3 Base through thinking SFT, DPO, and 750 steps of reinforcement learning with verifiable rewards. It generates intermediate thinking traces for math, coding, and general problem solving, and Ai2 publishes the weights, data, code, and checkpoints across the model flow.