qwen3.8-flash
Qwen3.8-Flash is the latest multimodal model from the Qwen family, combining powerful reasoning and generation with remarkable speed. It natively supports a million-token context window, allowing it to process lengthy documents, entire codebases, and complex conversations in a single pass. It shines in coding assistance, agentic workflows, and visual understanding — whether it's fixing code autonomously, operating desktop applications, or analyzing charts and long videos. Fully compatible with b
上下文
1M
输入 / 1M
$0.15
输出 / 1M
$0.47
Modalities
协议
3- chat.completions
- messages
- responses
可用性与性能
近 7 天
正在加载模型性能
能力
文本视觉长上下文缓存
计费明细
输入 Tokens
$0.15/1M
输出 Tokens
$0.47/1M
缓存命中
$0.016/1M
缓存输入
$0.016/1M
缓存写入(5m)
$0.2/1M
缓存写入(1h)
$0.2/1M
缓存写入
$0.2/1M
价格以 USD / 1M tokens 展示(除非另注单位)。Batch API 与缓存命中会享有额外折扣,最终以工作空间实时费率为准。
