Llama 3.2 3B Instruct
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it supports eight languages, including English, Spanish, and Hindi, and is adaptable for additional languages. Trained on 9 trillion tokens, the Llama 3.2 3B model excels in instruction-following, complex reasoning, and tool use. Its balanced performance makes it ideal for applications needing accuracy and efficiency in text generation across multilingual settings. Click here for the [original model card](https://github.com/meta-llama/llama-models/blob/main/models/llama3_2/MODEL_CARD.md). Usage of this model is subject to [Meta's Acceptable Use Policy](https://www.llama.com/llama3/use-policy/).
Specifications
Modalities
Supported Parameters
Max Output Tokens
16,384Frequently asked questions
- Is Llama 3.2 3B Instruct available on FastMetal?
- Not at the moment. GLM 5.3 Flash, from the same lab, is available on the FastMetal API today.
- What is the context window of Llama 3.2 3B Instruct?
- 131,072 tokens, shared between the prompt and the response.