Z

Zhipu: glm-4.7-thinking

glm-4.7-thinking
204.8K context200K outReasoningToolsCacheStructured
Released Dec 22, 2025Knowledge cutoff Nov 2025Updated Aug 24, 2026

GLM-4.7 is Zhipu AI's 2026 flagship model, featuring a 355B parameter Mixture-of-Experts (MoE) architecture. Its signature innovation is the "Interleaved Thinking" system, which enables the model to reason before every response and tool call, ensuring unparalleled instruction adherence. It has gained fame as the premier engine for "Vibe Coding," capable of translating vague creative descriptions into aesthetically superior, production-ready UI/UX. Ranking top among open-weight models on SWE-bench, it rivals proprietary giants like Claude 3.5 Sonnet in autonomous software engineering and complex agentic workflows.

Mode chatTokenizer OtherExpiration Dec 2026zai-org/GLM-4.7

Pricing

Input price
$0.95/ 1M tokens
Output price
$3.50/ 1M tokens
Context window 204.8K tokensCompatible endpoints openaiVendor Zhipu

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providersNot sent by default
include_reasoningAll providers-
logit_biasSome providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providersNot sent by default
reasoningAll providers-
repetition_penaltyAll providersNot sent by default
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_aSome providers-
top_kAll providersNot sent by default
top_pAll providers0.95

Frequently asked questions

Similar models