← model directory
~z-ai
Z.ai: GLM Flash Latest
~z-ai/glm-flash-latest
↓ runs free on your own hardware⚙ tool calling
This model always redirects to the latest model in the GLM Flash family.
specs & pricing
- Type
- text
- Provider
- ~z-ai
- Model ID
- ~z-ai/glm-flash-latest
- Capabilities
- vision, tools, reasoning
- Context window
- 1.3M tokens
- Self-hostable
- Yes — runs on your own GPU
- Input price
- $0.083 / 1M tokens
- Output price
- $0.28 / 1M tokens
Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.