Z

Zhipu: glm-5.2-thinking:free

glm-5.2-thinking:free
1M context262.1K outReasoningToolsParallel toolsCacheStructured
Released Jun 16, 2026Updated Jul 15, 2026

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering, and complex multi-step automation. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is particularly strong at coding and tool use across long-running tasks, able to maintain engineering context and follow standards consistently through a full development workflow, from requirements to multi-platform deployment, in a single task.

Mode chatTokenizer OtherQuantization fp8zai-org/GLM-5.2

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Compatible endpoints openaiVendor Zhipu

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providersNot sent by default
include_reasoningAll providers-
logit_biasSome providers-
logprobsAll providers-
max_tokensAll providers-
min_pSome providers-
parallel_tool_callsSome providers-
presence_penaltyAll providersNot sent by default
reasoningAll providers-
reasoning_effortAll providers-
repetition_penaltyAll providersNot sent by default
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providersNot sent by default
top_logprobsAll providers-
top_pAll providers0.95

Frequently asked questions

Similar models