StableLM 2 Zephyr 1.6B
Stability AI · 2024-01-19 · 1.6B parameters
StableLM 2 Zephyr 1.6B is Stability AI's compact 1.6B instruction-tuned language model from the StableLM 2 series, trained using Direct Preference Optimization (DPO) on a combination of public and synthetic datasets following the Zephyr training recipe. Released in January 2024, it targets strong instruction-following performance at a very small model size suitable for on-device or resource-constrained deployment. Despite its size, the model achieves competitive performance on chat benchmarks relative to other sub-2B models.