Z
Zhipu: glm-5.3-flash
glm-5.3-flash
1M contesto131.1K outputRagionamentoStrumentiVideoCacheStrutturato
Rilasciato Aug 26, 2026Aggiornato Aug 26, 2026
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.
Prezzi
Prezzo di input
$0.15/ 1 M token
Prezzo di output
$0.50/ 1 M token
Endpoint compatibili openaiProvider Zhipu
Disponibilità
Performance
Caricamento dati di performance...
Utilizzo e classifica
Caricamento utilizzo...
Parametri supportati
Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.
| Parametro | Provider | Predefinito |
|---|---|---|
| frequency_penalty | Tutti i provider | - |
| include_reasoning | Tutti i provider | - |
| max_tokens | Tutti i provider | - |
| presence_penalty | Tutti i provider | - |
| reasoning | Tutti i provider | - |
| reasoning_effort | Tutti i provider | - |
| repetition_penalty | Tutti i provider | - |
| response_format | Tutti i provider | - |
| seed | Tutti i provider | - |
| stop | Tutti i provider | - |
| temperature | Tutti i provider | 1 |
| tool_choice | Tutti i provider | - |
| tools | Tutti i provider | - |
| top_k | Tutti i provider | - |
| top_p | Tutti i provider | 0.95 |