qwen3.7-plus

Among the Qwen3.7 series, the cost-effective Plus model builds on its robust text capabilities while delivering a comprehensive upgrade to its vision‑language abilities, all while preserving its full‑stack agent‑level intelligence for coding, tool use, and productivity workflows. Its key distinguishing feature is multi‑modal interactive hybrid agent capabilities, enabling it to perceive real‑world scenes, read screens and interact with GUIs, generate code based on visual references, and perform

LiveAlibaba3 protocols1M contextStream cancellation unsupported
Context
1M
Input / 1M
$0.40
Output / 1M
$1.60
Modalities
Protocol
3
  • chat.completions
  • messages
  • responses

Status & performance

Last 7 days
Loading model performance

Capabilities

TextVisionLong contextCache

Pricing

Input tokens
$0.4/1M
Output tokens
$1.60/1M
Cache read
$0.08/1M
Cached input
$0.08/1M
Cache write (5m)
$0.5/1M
Cache write (1h)
$0.5/1M
Cache write
$0.5/1M
input_above_threshold
$1.20/1M
cached_input_above_threshold
$0.24/1M
output_above_threshold
$4.80/1M
input_image_above_threshold
$1.20/1M
Cache write (long context)
$1.50/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.