Z

Zhipu: glm-5.3-think-search:free

glm-5.3-think-search:free
1M context131.1K outReasoningToolsCacheStructured
Released Aug 18, 2026Updated Aug 19, 2026

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency. Reasoning is always on and cannot be disabled. Reasoning efforts low, high, and max are supported; max is the default.

Tokenizer OtherQuantization fp8Expiration Dec 2098

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Compatible endpoints openaiVendor Zhipu

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
include_reasoningAll providers-
max_tokensAll providers-
reasoningAll providers-
reasoning_effortAll providers-
response_formatAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_pAll providers0.95

Frequently asked questions

Similar models