Qwen3.5-35B-A3B
Qwen馃嚚馃嚦
qwen/qwen3.5-35b-a3bQwen3.5 35B-A3B is a sparse mixture-of-experts model from Alibaba's Qwen3.5 line (35B total, 3B active parameters), built for efficient general-purpose and agentic tasks at low inference cost.
Cost rate
0.2x
Context
262K
Released
Feb 25, 2026
Input
TextImageVideo
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
Medicine & Healthcare
#141 路 top 38%
Mathematical
#147 路 top 39%
Business, Management, & Finance
#148 路 top 37%
Software & IT Services
#151 路 top 38%
Performance
Median latency and throughput measured across recent requests.
Throughput165 tok/s
Latency