Zhipu AI/ GLM
GLM-4.5-Flash
zhipuai/glm/4-5-flash
Efficient GLM model for fast reasoning, coding, and agent workflows
Released
2025-07-28
Context window
131K
Max output
98K
Weights
Open weight
Modalities
Input
Text
Output
Text
Identity
- Version
- glm-4.5-flash
- Knowledge cutoff
- 2025-04
- Aliases
- GLM-4.5-Flash
Resources
Providers
| Provider | Provider model ID | Protocol | Context | Max output | Input | Output | Capabilities | Notes | Pricing | Streaming |
|---|---|---|---|---|---|---|---|---|---|---|
| Zhipu AI | glm-4.5-flash | custom | 131K | 98K | Text | Text | ReasoningTool callTemperature | - | input $0/M · cache read $0/M output $0/M · cache write $0/M | Yes |
| UnoRouter | glm-4.5-flash:free | openai_chat | 131K | 98K | Text | Text | ReasoningTool callTemperature | - | input $0/M output $0/M | Yes |
| EmpirioLabs AI | glm-4-5-flash | openai_chat | 131K | 98K | Text | Text | ReasoningTool callTemperature | - | input $0/M output $0/M | Yes |
| Z.AI | glm-4.5-flash | custom | 131K | 98K | Text | Text | ReasoningTool callTemperature | - | input $0/M · cache read $0/M output $0/M · cache write $0/M | Yes |
Capabilities
ReasoningTool callTemperature
Official price
input $0/M · cache read $0/M
output $0/M · cache write $0/M