Back to Models
qwen logo
qwen/qwen3-235b-a22b-2507
Not Available

Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage. The model supports a native 262K context length and does not implement "thinking mode" (<think> blocks). Compared to its base variant, this version delivers significant gains in knowledge coverage, long-context reasoning, coding benchmarks, and alignment with open-ended tasks. It is particularly strong on multilingual understanding, math reasoning (e.g., AIME, HMMT), and alignment evaluations like Arena-Hard and WritingBench.

7/21/2025
262,144 tokens
#78 Text (Japanese)

Specifications

Modalities

Input
text
Output
text

Supported Parameters

frequency_penalty
include_reasoning
logit_bias
logprobs
max_tokens
min_p
presence_penalty
reasoning
reasoning_effort
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_logprobs
top_p

Frequently asked questions

Is Qwen3 235B A22B Instruct 2507 available on FastMetal?
Not at the moment. Qwen3.8 Max (0902), from the same lab, is available on the FastMetal API today.
What is the context window of Qwen3 235B A22B Instruct 2507?
262,144 tokens, shared between the prompt and the response.
How does Qwen3 235B A22B Instruct 2507 rank?
#78 on the public arena's Japanese board (ELO 1,393). Ranks move as the leaderboard is updated.

Leaderboard

Text
OverallELO: 1,423
#112
JapaneseELO: 1,393
#78
ChineseELO: 1,469
#96
KoreanELO: 1,383
#82
EnglishELO: 1,430
#118
frenchELO: 1,454
#84
germanELO: 1,424
#89
spanishELO: 1,426
#91
russianELO: 1,418
#105
CodingELO: 1,472
#109
MathELO: 1,418
#113
Creative WritingELO: 1,379
#130
Instruction FollowingELO: 1,416
#106
Hard PromptsELO: 1,448
#100
Multi-TurnELO: 1,437
#98