Z.ai
Latency—
Throughput—
Upload rate—
Context length—
Max output—
Input price$1.4per 1M tokens
Output price$4.4per 1M tokens
Cache read$0.364per 1M tokens
Cache write$0per 1M tokens
/z.ai/glm-5.3
The public release of GLM-5.3 was documented in 2026 as part of Z.ai’s GLM-5 model generation. It belongs to the GLM-5 family and is used for general reasoning, coding, and agent-oriented workloads.
Replace <YOUR_API_KEY> with the API key generated on your token management page.
All requests must include the Authorization: Bearer <TOKEN> header.
RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.