GLM 5.2 Pricing
Z.AI · GLM 5 · Live
GLM 5.2 costs $1.4 per million input tokens and $4.4 per million output tokens on the Z.AI API. Cached input reads are billed at $0.26 per million tokens. It supports a 1,048,576-token context window and up to 32,768 output tokens per request. The model supports extended reasoning, image input, tool use, long context. Released June 16, 2026. Pricing verified July 29, 2026.
GLM 5.2 API pricing
- Input (per 1M tokens)
- $1.4
- Output (per 1M tokens)
- $4.4
- Cached input (per 1M tokens)
- $0.26
- Context window
- 1,048,576 tokens
- Max output
- 32,768 tokens
- Released
- June 16, 2026
Capabilities
Price history
Price-stable since June 28, 2026.
Frequently asked questions
Notes
Chinese frontier open-source model from Z.AI. Standalone pay-per-token API launched June 16, 2026 at $1.40/$4.40 per million. 744B parameters, 1M context window, MIT license (fully open). Highest scoring open-weight model on artificial analysis intelligence index (51). Beats GPT-5.5 on Frontier SWE coding benchmark; trails Claude Opus 4.8 by less than 1 percentage point. Cached input at $0.26/M cuts costs ~80%. Z.AI founder told Elon Musk that open-weight Fable-tier capability will arrive before Q1 2027.
