Model information
Released
2026-04-07
Context window
200K
Max output
131K
Model parameters
754B
Weights
Open
Modalities
Input
Text
Output
Text
Identity
- Version
- glm-5.1
- License
- MIT
- Aliases
- GLM-5.1
Resources
Capabilities
Open weightReasoningStructuredTemperatureTool call
Official price
input $1.4/M · output $4.4/M
cache read $0.26/M · cache write $0/M
Providers
| Provider | Provider model ID | Protocol | Context | Max output | Input | Output | Capabilities | Notes | Pricing | Streaming |
|---|---|---|---|---|---|---|---|---|---|---|
| Zhipu AI | glm-5.1 | custom | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.4/M · output $4.4/M cache read $0.26/M · cache write $0/M | Yes | ||
| GreenPT | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.756/M · output $5.518/M | Yes | ||
| OpenCode Go | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callTemperature | - | input $1.4/M · output $4.4/M cache read $0.26/M | Yes | ||
| Zhipu AI Coding Plan | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0/M · output $0/M cache read $0/M · cache write $0/M | Yes | ||
| CrofAI | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.45/M · output $2.15/M cache read $0.08/M · cache write $0/M | Yes | ||
| LLM Gateway | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.931/M · output $2.93/M cache read $0.173/M · cache write $0/M | Yes | ||
| Kenari | glm-5-1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0/M · output $0/M | Yes | ||
| OpenCode Zen | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callTemperature | - | input $1.4/M · output $4.4/M cache read $0.26/M | Yes | ||
| 302.AI | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.86/M · output $3.5/M | Yes | ||
| DInference | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.25/M · output $3.89/M | Yes | ||
| EmpirioLabs AI | glm-5-1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.825/M · output $3.301/M cache read $0.165/M | Yes | ||
| Alibaba Token Plan | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0/M · output $0/M cache read $0/M · cache write $0/M | Yes | ||
| Wafer | GLM-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1/M · output $3.2/M cache read $0.1/M · cache write $0/M | Yes | ||
| DigitalOcean | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.975/M · output $4.3/M cache read $0.26/M | Yes | ||
| Auriko | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.4/M · output $4.4/M cache read $0.26/M | Yes | ||
| Charm Hyper | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.524/M · output $4.791/M cache read $0.283/M | Yes | ||
| Alibaba (China) | glm-5.1 | custom | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.87/M · output $3.48/M cache read $0.17/M | Yes | ||
| Alibaba Token Plan (China) | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0/M · output $0/M cache read $0/M · cache write $0/M | Yes | ||
| Ollama Cloud | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool call | - | - | Yes | ||
| Z.AI | glm-5.1 | custom | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.4/M · output $4.4/M cache read $0.26/M · cache write $0/M | Yes | ||
| Cortecs | glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.31/M · output $4.1/M cache read $0.24/M | Yes | ||
| EBCloud | GLM-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.857/M · output $3.429/M | Yes | ||
| OpenRouter | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | Hugging Face: zai-org/GLM-5.1 | input $0.966/M · output $3.036/M cache read $0.179/M | Yes | ||
| CrossModel | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1/M · output $3.8/M cache read $0.2/M · cache write $1/M | Yes | ||
| FastRouter | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.05/M · output $3.5/M | Yes | ||
| ZenMux | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $0.878/M · output $3.513/M cache read $0.19/M | Yes | ||
| OrcaRouter | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.4/M · output $4.4/M cache read $0.26/M · cache write $0/M | Yes | ||
| TensorX | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.4/M · output $4.4/M cache read $0.35/M · cache write $1.75/M | Yes | ||
| Kilo Gateway | z-ai/glm-5.1 | openai_chat | 200K | 131K | ReasoningTool callStructuredTemperature | - | input $1.38/M · output $4.4/M cache read $0.26/M | Yes |