Back to Models
qwen/qwen3-235b-a22b
Not Available
Qwen3 235B A22B
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and code tasks, and a "non-thinking" mode for general conversational efficiency. The model demonstrates strong reasoning ability, multilingual support (100+ languages and dialects), advanced instruction-following, and agent tool-calling capabilities. It natively handles a 32K token context window and extends up to 131K tokens using YaRN-based scaling.
4/28/2025
131,072 tokens
#120 Text (german)
Specifications
Modalities
Input
text
Output
text
Supported Parameters
include_reasoning
max_tokens
presence_penalty
reasoning
response_format
seed
temperature
tool_choice
tools
top_p
Max Output Tokens
8,192Leaderboard
Text
๐OverallELO: 1,375
#163๐ฏ๐ตJapaneseELO: 1,311
#130๐จ๐ณChineseELO: 1,392
#166๐ฐ๐ทKoreanELO: 1,318
#145๐ฌ๐งEnglishELO: 1,387
#163frenchELO: 1,374
#160germanELO: 1,382
#120spanishELO: 1,370
#151russianELO: 1,354
#172๐ปCodingELO: 1,433
#153๐งฎMathELO: 1,393
#145โ๏ธCreative WritingELO: 1,324
#182๐Instruction FollowingELO: 1,357
#172๐ถ๏ธHard PromptsELO: 1,392
#164๐ฌMulti-TurnELO: 1,372
#166