Back to Models
qwen/qwen3-235b-a22b-2507
Not Available
Qwen3 235B A22B Instruct 2507
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage. The model supports a native 262K context length and does not implement "thinking mode" (<think> blocks). Compared to its base variant, this version delivers significant gains in knowledge coverage, long-context reasoning, coding benchmarks, and alignment with open-ended tasks. It is particularly strong on multilingual understanding, math reasoning (e.g., AIME, HMMT), and alignment evaluations like Arena-Hard and WritingBench.
7/21/2025
262,144 tokens
#70 Text (Korean)
Specifications
Modalities
Input
text
Output
text
Supported Parameters
frequency_penalty
include_reasoning
logit_bias
logprobs
max_tokens
min_p
presence_penalty
reasoning
reasoning_effort
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_logprobs
top_p
Leaderboard
Text
๐OverallELO: 1,423
#101๐ฏ๐ตJapaneseELO: 1,382
#73๐จ๐ณChineseELO: 1,471
#88๐ฐ๐ทKoreanELO: 1,385
#70๐ฌ๐งEnglishELO: 1,430
#108frenchELO: 1,459
#72germanELO: 1,428
#75spanishELO: 1,423
#81russianELO: 1,417
#97๐ปCodingELO: 1,473
#96๐งฎMathELO: 1,418
#103โ๏ธCreative WritingELO: 1,380
#122๐Instruction FollowingELO: 1,416
#96๐ถ๏ธHard PromptsELO: 1,448
#90๐ฌMulti-TurnELO: 1,438
#86