GLM 5.3 Flash (batch)
Z.AI馃嚚馃嚦
z-ai/glm-5.3-flash:batchThis is the batch-processing variant of GLM-5.3-Flash, Z.ai's open-weight natively multimodal mixture-of-experts model (320B total parameters, 18B active, MIT license, 1M-token context, hybrid sparse-plus-linear attention), offering the same coding and agentic capabilities at a lower cost via OpenRouter's asynchronous batch endpoint. It is best for high-volume, non-latency-sensitive coding or agent workloads where cost efficiency matters more than immediate response time.
Cost rate
0.1x
Context
1M
Released
Aug 26, 2026
Input
TextImageVideo
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
Mathematical
#12 路 top 3%
Software & IT Services
#19 路 top 5%
Image understanding
#30 路 top 20%
Life, Physical, & Social Science
#31 路 top 8%