glm-5.1
GLM-5.1 is Z.AI’s latest flagship model, designed for long-horizon tasks. It can work continuously and autonomously on a single task for up to 8 hours, completing the full loop from planning and execution to iterative optimization and delivering production-grade results.
Context
200K
Input / 1M
$1.40
Output / 1M
$4.40
Modalities
Protocol
2- chat.completions
- messages
Status & performance
Last 7 days
Loading model performance
Capabilities
TextLong contextCache
Pricing
Input tokens
$1.40/1M
Output tokens
$4.40/1M
Cache read
$0.26/1M
Cached input
$0.26/1M
Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.
