← model directory

~z-ai

Z.ai: GLM Flash Latest

~z-ai/glm-flash-latest

↓ runs free on your own hardware⚙ tool calling

This model always redirects to the latest model in the GLM Flash family.

specs & pricing
Type
text
Provider
~z-ai
Model ID
~z-ai/glm-flash-latest
Capabilities
vision, tools, reasoning
Context window
1.3M tokens
Self-hostable
Yes — runs on your own GPU
Input price
$0.083 / 1M tokens
Output price
$0.28 / 1M tokens

Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.

Z.ai: GLM Flash Latest — pricing, context & specs | Wide Area Intelligence