glm-5.1

GLM-5.1 is Z.AI’s latest flagship model, designed for long-horizon tasks. It can work continuously and autonomously on a single task for up to 8 hours, completing the full loop from planning and execution to iterative optimization and delivering production-grade results.

LiveZhipu AI2 protocols200K contextStream cancellation unsupported
Context
200K
Input / 1M
$1.40
Output / 1M
$4.40
Modalities
Protocol
2
  • chat.completions
  • messages

Status & performance

Last 7 days
Loading model performance

Capabilities

TextLong contextCache

Pricing

Input tokens
$1.40/1M
Output tokens
$4.40/1M
Cache read
$0.26/1M
Cached input
$0.26/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.