qwen3.8-max

A 2.4-trillion-parameter MoE flagship with a major leap in coding and office productivity, able to autonomously code for over ten days to deliver complete projects. It handles hundreds of professional tasks across law, finance, design, and more, delivering production-grade results end-to-end in a single conversation. Native visual understanding runs through the entire planning, execution, and verification pipeline, enabling deep semantic parsing of ultra-long documents and long videos. It plans

LiveAlibaba3 protocols1M contextStream cancellation unsupported
Context
1M
Input / 1M
$2.00
Output / 1M
$6.00
Modalities
Protocol
3
  • chat.completions
  • messages
  • responses

Status & performance

Last 7 days
Loading model performance

Capabilities

TextVisionLong contextCache

Pricing

Input tokens
$2.00/1M
Output tokens
$6.00/1M
Cache read
$0.25/1M
Cached input
$0.25/1M
Cache write (5m)
$2.50/1M
Cache write (1h)
$2.50/1M
Cache write
$2.50/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.