Back to Models
qwen logo
qwen/qwen3-235b-a22b
Not Available

Qwen3 235B A22B

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and code tasks, and a "non-thinking" mode for general conversational efficiency. The model demonstrates strong reasoning ability, multilingual support (100+ languages and dialects), advanced instruction-following, and agent tool-calling capabilities. It natively handles a 32K token context window and extends up to 131K tokens using YaRN-based scaling.

4/28/2025
131,072 tokens
#120 Text (german)
Specifications

Modalities

Input
text
Output
text

Supported Parameters

include_reasoning
max_tokens
presence_penalty
reasoning
response_format
seed
temperature
tool_choice
tools
top_p

Max Output Tokens

8,192
Leaderboard
Text
๐Ÿ†OverallELO: 1,375
#163
๐Ÿ‡ฏ๐Ÿ‡ตJapaneseELO: 1,311
#130
๐Ÿ‡จ๐Ÿ‡ณChineseELO: 1,392
#166
๐Ÿ‡ฐ๐Ÿ‡ทKoreanELO: 1,318
#145
๐Ÿ‡ฌ๐Ÿ‡งEnglishELO: 1,387
#163
frenchELO: 1,374
#160
germanELO: 1,382
#120
spanishELO: 1,370
#151
russianELO: 1,354
#172
๐Ÿ’ปCodingELO: 1,433
#153
๐ŸงฎMathELO: 1,393
#145
โœ๏ธCreative WritingELO: 1,324
#182
๐Ÿ“Instruction FollowingELO: 1,357
#172
๐ŸŒถ๏ธHard PromptsELO: 1,392
#164
๐Ÿ’ฌMulti-TurnELO: 1,372
#166