
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
Modalities
Context
128K
Released
May 14, 2025
Knowledge Cutoff
Dec 2023
Token volume and request traffic to this model over time.
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
Llama 3.3 8B Instruct has a 128,000 token context window.
Llama Guard 4 12B, Llama 4 Maverick, Llama 4 Scout and 5 more are other text models from Meta Llama.
Llama 3.3 8B Instruct was released on May 14, 2025. Its knowledge cutoff is December 31, 2023.