Qwen3 30B A3B
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique ability to switch seamlessly between a thinking mode for complex reasoning and a non-thinking mode for efficient dialogue ensures versatile, high-quality performance. Significantly outperforming prior models like QwQ and Qwen2.5, Qwen3 delivers superior mathematics, coding, commonsense reasoning, creative writing, and interactive dialogue capabilities. The Qwen3-30B-A3B variant includes 30.5 billion parameters (3.3 billion activated), 48 layers, 128 experts (8 activated per task), and supports up to 131K token contexts with YaRN, setting a new standard among open-source models.
Specifications
Modalities
Supported Parameters
Max Output Tokens
40,960Reasoning Configuration
- Default
- Thinking on
- Turning thinking off
- Possible (reasoning.enabled: false)
Frequently asked questions
- Is Qwen3 30B A3B available on FastMetal?
- Not at the moment. Qwen3.8 Max (0902), from the same lab, is available on the FastMetal API today.
- What is the context window of Qwen3 30B A3B?
- 40,960 tokens, shared between the prompt and the response.
- How does Qwen3 30B A3B rank?
- #176 on the public arena's Japanese board (ELO 1,256). Ranks move as the leaderboard is updated.