qwen3.7-max

The Max model, the largest and most capable in the Qwen3.7 series, currently offers a pure‑text‑only interface for public experimentation. Qwen3.7 is a next‑generation flagship model designed for the agent‑centric era, with its core strengths lying in the breadth and depth of its agent‑level capabilities: it excels at programming, office and productivity tasks, and long‑term autonomous execution.

LiveAlibaba3 protocols1M contextStream cancellation unsupported
Context
1M
Input / 1M
$2.50
Output / 1M
$7.50
Modalities
Protocol
3
  • chat.completions
  • messages
  • responses

Status & performance

Last 7 days
Loading model performance

Capabilities

TextLong contextCache

Pricing

Input tokens
$2.50/1M
Output tokens
$7.50/1M
Cache read
$0.5/1M
Cached input
$0.5/1M
Cache write (5m)
$3.13/1M
Cache write (1h)
$3.13/1M
Cache write
$3.13/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.