
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model is optimized for multilingual dialogue use cases and outperforms many of the available open source and closed chat models on common industry benchmarks.
Supported languages: English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai.
Modalities
Price
Free
Context
131K
Released
Dec 6, 2024
Knowledge Cutoff
Dec 2023
Token volume and request traffic to this model over time.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model is optimized for multilingual dialogue use cases and outperforms many of the available open source and closed chat models on common industry benchmarks.
Yes. The pricing shown on this page for Llama 3.3 70B Instruct is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
Llama 3.3 70B Instruct has a 131,072 token context window.
Llama Guard 4 12B, Llama 4 Maverick, Llama 4 Scout and 4 more are other text models from Meta Llama.
Llama 3.3 70B Instruct was released on December 6, 2024. Its knowledge cutoff is December 31, 2023.