Z
Zhipu: glm-4.7-flash-heretic:free
glm-4.7-flash-heretic:free
200K context131.1K outReasoningToolsCacheStructured
Released Jan 19, 2026Updated Jun 20, 2026
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards.
Pricing
Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 204.8K tokensCompatible endpoints openaiVendor Zhipu
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| frequency_penalty | All providers | Not sent by default |
| include_reasoning | All providers | - |
| logit_bias | All providers | - |
| logprobs | Some providers | - |
| max_tokens | All providers | - |
| min_p | All providers | - |
| presence_penalty | All providers | - |
| reasoning | All providers | - |
| repetition_penalty | All providers | - |
| response_format | All providers | - |
| seed | All providers | - |
| stop | All providers | - |
| structured_outputs | All providers | - |
| temperature | All providers | 1 |
| tool_choice | All providers | - |
| tools | All providers | - |
| top_k | All providers | - |
| top_logprobs | Some providers | - |
| top_p | All providers | 0.95 |