qwen3.8-flash

Qwen3.8-Flash is the latest multimodal model from the Qwen family, combining powerful reasoning and generation with remarkable speed. It natively supports a million-token context window, allowing it to process lengthy documents, entire codebases, and complex conversations in a single pass. It shines in coding assistance, agentic workflows, and visual understanding — whether it's fixing code autonomously, operating desktop applications, or analyzing charts and long videos. Fully compatible with b

LiveAlibaba3 个协议1M 上下文不支持流式取消
上下文
1M
输入 / 1M
$0.15
输出 / 1M
$0.47
Modalities
协议
3
  • chat.completions
  • messages
  • responses

可用性与性能

近 7 天
正在加载模型性能

能力

文本视觉长上下文缓存

计费明细

输入 Tokens
$0.15/1M
输出 Tokens
$0.47/1M
缓存命中
$0.016/1M
缓存输入
$0.016/1M
缓存写入(5m)
$0.2/1M
缓存写入(1h)
$0.2/1M
缓存写入
$0.2/1M

价格以 USD / 1M tokens 展示(除非另注单位)。Batch API 与缓存命中会享有额外折扣,最终以工作空间实时费率为准。