Z
Zhipu: glm-5.3
glm-5.3
1M context131.1K outReasoningToolsCacheStructured
Released Aug 18, 2026Updated Aug 19, 2026
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency. Reasoning is always on and cannot be disabled. Reasoning efforts low, high, and max are supported; max is the default.
Mode chatTokenizer OtherQuantization fp8Expiration Dec 2098
Pricing
Input price
$2.80/ 1M tokens
Output price
$8.96/ 1M tokens
Compatible endpoints openaiVendor Zhipu
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| include_reasoning | All providers | - |
| max_tokens | All providers | - |
| reasoning | All providers | - |
| reasoning_effort | All providers | - |
| response_format | All providers | - |
| temperature | All providers | 1 |
| tool_choice | All providers | - |
| tools | All providers | - |
| top_k | All providers | - |
| top_p | All providers | 0.95 |