Back to Models
qwen/qwen3-235b-a22b-2507
Not Available
Qwen3 235B A22B Instruct 2507
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage. The model supports a native 262K context length and does not implement "thinking mode" (<think> blocks). Compared to its base variant, this version delivers significant gains in knowledge coverage, long-context reasoning, coding benchmarks, and alignment with open-ended tasks. It is particularly strong on multilingual understanding, math reasoning (e.g., AIME, HMMT), and alignment evaluations like Arena-Hard and WritingBench.
7/21/2025
262,144 tokens
#78 Text (Japanese)
Specifications
Modalities
Input
text
Output
text
Supported Parameters
frequency_penalty
include_reasoning
logit_bias
logprobs
max_tokens
min_p
presence_penalty
reasoning
reasoning_effort
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_logprobs
top_p
Frequently asked questions
- Is Qwen3 235B A22B Instruct 2507 available on FastMetal?
- Not at the moment. Qwen3.8 Max (0902), from the same lab, is available on the FastMetal API today.
- What is the context window of Qwen3 235B A22B Instruct 2507?
- 262,144 tokens, shared between the prompt and the response.
- How does Qwen3 235B A22B Instruct 2507 rank?
- #78 on the public arena's Japanese board (ELO 1,393). Ranks move as the leaderboard is updated.
Leaderboard
Text
OverallELO: 1,423
#112JapaneseELO: 1,393
#78ChineseELO: 1,469
#96KoreanELO: 1,383
#82EnglishELO: 1,430
#118frenchELO: 1,454
#84germanELO: 1,424
#89spanishELO: 1,426
#91russianELO: 1,418
#105CodingELO: 1,472
#109MathELO: 1,418
#113Creative WritingELO: 1,379
#130Instruction FollowingELO: 1,416
#106Hard PromptsELO: 1,448
#100Multi-TurnELO: 1,437
#98