qwen3.8-max
A 2.4-trillion-parameter MoE flagship with a major leap in coding and office productivity, able to autonomously code for over ten days to deliver complete projects. It handles hundreds of professional tasks across law, finance, design, and more, delivering production-grade results end-to-end in a single conversation. Native visual understanding runs through the entire planning, execution, and verification pipeline, enabling deep semantic parsing of ultra-long documents and long videos. It plans
Context
1M
Input / 1M
$2.00
Output / 1M
$6.00
Modalities
Protocol
3- chat.completions
- messages
- responses
Status & performance
Last 7 days
Loading model performance
Capabilities
TextVisionLong contextCache
Pricing
Input tokens
$2.00/1M
Output tokens
$6.00/1M
Cache read
$0.25/1M
Cached input
$0.25/1M
Cache write (5m)
$2.50/1M
Cache write (1h)
$2.50/1M
Cache write
$2.50/1M
Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.
