
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
Token volume and request traffic to this model over time.
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
Llama 3.3 8B Instruct has a 128,000 token context window.
Llama Guard 4 12B, Llama 4 Maverick, Llama 4 Scout and 5 more are other text models from Meta Llama.
Llama 3.3 8B Instruct was released on May 14, 2025. Its knowledge cutoff is December 31, 2023.