Llama 4 Maverick Instruct FP8
Meta · 2025-04-05 · 401.6B parameters
Llama 4 Maverick Instruct FP8 is Meta's separately released FP8-quantized checkpoint of the post-trained, natively multimodal Llama 4 Maverick Instruct model. It preserves the 128-expert architecture, 1,048,576-token context window, and text-and-image inputs of the BF16 checkpoint; the released safetensors expose 401,649,841,664 tensor elements across BF16 and FP8 weights.