GLM 5.3 Prime
Z.AI🇨🇳
z-ai/glm-5.3-primeGLM-5.3-Prime is the high-speed serving tier of Z.ai's open-weight GLM-5.3 flagship, running the same weights (1M-token context, up to 128K output) behind an accelerated stack for 1.5-2x the output throughput at $2.80/$8.80 per million input/output tokens. Reasoning is always on (low, high, or max effort), and it targets latency-sensitive coding, streaming code generation, and long-horizon multi-turn agent orchestration.
Mức chi phí
2x
Ngữ cảnh
1M
Phát hành
23 thg 9, 2026
Đầu vào
Text
Đầu ra
Text
Hỗ trợ
Suy luậnGọi công cụĐầu ra có cấu trúc