Z
Zhipu: glm-4-flash:free
glm-4-flash:free
128K context4.1K outToolsVisionFiles
Released Aug 26, 2025Knowledge cutoff Sep 2025Updated Jul 27, 2026
A high-speed, lightweight flagship model from Zhipu AI. As the most cost-effective member of the GLM-4 family, it is deeply optimized for high-concurrency and low-latency scenarios. It maintains a leading position in Chinese semantic understanding while excelling in agent task scheduling, long-form summarization, and real-time dialogue.
All providers for this model are busy right now
Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.
Request this model on DiscordPricing
Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 128K tokensCompatible endpoints openaiVendor Zhipu
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...