
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it supports eight languages, including English, Spanish, and Hindi, and is adaptable for additional languages.
Trained on 9 trillion tokens, the Llama 3.2 3B model excels in instruction-following, complex reasoning, and tool use. Its balanced performance makes it ideal for applications needing accuracy and efficiency in text generation across multilingual settings.
Click here for the original model card(opens in new tab).
Usage of this model is subject to Meta's Acceptable Use Policy(opens in new tab).
Modalities
Price
Free
Context
131K
Released
Sep 25, 2024
Knowledge Cutoff
Dec 2023
Token volume and request traffic to this model over time.
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it supports eight languages, including English, Spanish, and Hindi, and is adaptable for additional languages.
Yes. The pricing shown on this page for Llama 3.2 3B Instruct is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
Llama 3.2 3B Instruct has a 131,072 token context window.
Llama Guard 4 12B, Llama 4 Maverick, Llama 4 Scout and 4 more are other text models from Meta Llama.
Llama 3.2 3B Instruct was released on September 25, 2024. Its knowledge cutoff is December 31, 2023.