Models
Enterprise
Subscribe
Resource
Documentation
Console

z.ai/glm-5.3

zhipu

Z.ai: glm-5.3

The public release of GLM-5.3 was documented in 2026 as part of Z.ai’s GLM-5 model generation. It belongs to the GLM-5 family and is used for general reasoning, coding, and agent-oriented workloads.

Modetext → text
Input price$1.4 per M tokens
Output price$4.4 per M tokens
Context length
Weekly usage22.895K tokens
Listed at

Z.ai

Latency
Throughput
Upload rate
Context length
Max output
Input price$1.4per 1M tokens
Output price$4.4per 1M tokens
Cache read$0.364per 1M tokens
Cache write$0per 1M tokens

Performance

Avg TPS
Avg latency
Avg success rate
TPS
TTFT
Latency
Success rate

Success rate

Speed

Code Example

Replace <YOUR_API_KEY> with the API key generated on your token management page.

Authentication

All requests must include the Authorization: Bearer <TOKEN> header.

Endpoint Path

POST

Supported Parameters

ParameterTypeDefaultDescription

Rate Limits

GroupRPMTPMRPD

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.

Pricing

GroupBilling typePrice summary
Glm0%
Per call
Input price1.4per 1M tokens
Completion price4.4per 1M tokens
Cache read price0.364per 1M tokens
Cache write price0per 1M tokens
Glm0%
Per call
Input price1.4per 1M tokens
Completion price4.4per 1M tokens
Cache read price0.364per 1M tokens
Cache write price0per 1M tokens