Z

Zhipu: glm-4-flash:free

glm-4-flash:free
128K context4.1K outToolsVisionFiles
Released Aug 26, 2025Knowledge cutoff Sep 2025Updated Jul 27, 2026

A high-speed, lightweight flagship model from Zhipu AI. As the most cost-effective member of the GLM-4 family, it is deeply optimized for high-concurrency and low-latency scenarios. It maintains a leading position in Chinese semantic understanding while excelling in agent task scheduling, long-form summarization, and real-time dialogue.

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Context window 128K tokensCompatible endpoints openaiVendor Zhipu

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Frequently asked questions

Similar models