Z

Zhipu: glm-5:free

glm-5:free
204.8K context128K outReasoningToolsCacheStructured
Released Feb 11, 2026Updated Jun 20, 2026

GLM-5.2 is the flagship model for the era of long tasks. It supports truly usable 1M context and has been tested to be capable of handling project-level engineering contexts. Long-term tasks are executed more stably and engineering standards are followed more reliably, resulting in a further increase in the success rate of development scenarios. A single task can complete the entire development chain from requirements to multi-platform deployable products.

Mode chatTokenizer OtherQuantization fp8zai-org/GLM-5

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 204.8K tokensCompatible endpoints openaiVendor Zhipu

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providersNot sent by default
include_reasoningAll providers-
logit_biasSome providers-
logprobsSome providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltySome providers-
reasoningAll providers-
repetition_penaltySome providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_logprobsSome providers-
top_pAll providers0.95

Frequently asked questions

Similar models