Z

Zhipu: glm-5-turbo

glm-5-turbo
128K context131.1K outReasoningToolsCacheStructured
Released Jan 15, 2026Updated Jul 22, 2026

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows involving long execution chains, with improved complex instruction decomposition, tool use, scheduled and persistent execution, and overall stability across extended tasks.

Mode chatTokenizer OtherExpiration Dec 2098

Pricing

Input price
$0.80/ 1M tokens$1.2033% off
Output price
$2.67/ 1M tokens$4.0033% off
Context window 200K tokensCompatible endpoints openai, anthropicVendor Zhipu

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltySome providersNot sent by default
include_reasoningAll providers-
max_tokensAll providers-
presence_penaltySome providersNot sent by default
reasoningAll providers-
repetition_penaltySome providersNot sent by default
response_formatAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providersNot sent by default
top_pAll providers0.95

Frequently asked questions

Similar models