glm-5.2

GLM-5.2 is a flagship model built for the era of long-horizon tasks. With truly usable 1M-token context, it has been tested to handle project-scale engineering context, delivering more stable long-task execution, more reliable adherence to engineering standards, and higher success rates in development scenarios. A single task can complete the full development workflow—from requirements to deployable products across multiple platforms.

LiveZhipu AI2 protocols1M contextStream cancellation unsupported
Context
1M
Input / 1M
$1.40
Output / 1M
$4.40
Modalities
Protocol
2
  • chat.completions
  • messages

Status & performance

Last 7 days
Loading model performance

Capabilities

TextLong contextCache

Pricing

Input tokens
$1.40/1M
Output tokens
$4.40/1M
Cache read
$0.28/1M
Cached input
$0.28/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.