Z
Zhipu: glm-5-turbo
glm-5-turbo
128K context131.1K outReasoningToolsCacheStructured
Released Jan 15, 2026Updated Jul 22, 2026
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows involving long execution chains, with improved complex instruction decomposition, tool use, scheduled and persistent execution, and overall stability across extended tasks.
Mode chatTokenizer OtherExpiration Dec 2098
Pricing
Input price
$0.80/ 1M tokens$1.2033% off
Output price
$2.67/ 1M tokens$4.0033% off
Context window 200K tokensCompatible endpoints openai, anthropicVendor Zhipu
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| frequency_penalty | Some providers | Not sent by default |
| include_reasoning | All providers | - |
| max_tokens | All providers | - |
| presence_penalty | Some providers | Not sent by default |
| reasoning | All providers | - |
| repetition_penalty | Some providers | Not sent by default |
| response_format | All providers | - |
| temperature | All providers | 1 |
| tool_choice | All providers | - |
| tools | All providers | - |
| top_k | All providers | Not sent by default |
| top_p | All providers | 0.95 |