Olmo 3 7B Instruct
Allen Institute for AI (Ai2) · 2025-11-20 · 7.3B parameters
Olmo 3 7B Instruct is Ai2's fully open 7B chat and quick-response checkpoint, produced from Olmo 3 Base through SFT, DPO, and reinforcement learning with verifiable rewards for multi-turn instruction following and tool use. Its base was explicitly extended from 8,192-token to 65,536-token sequences using YaRN on the full-attention layers, and Ai2 releases the weights, data, code, and intermediate checkpoints across the model flow.