gpt-5.6-luna

GPT-5.6 model optimized for cost-sensitive workloads

LiveOpenAI2 protocols1M contextStream cancellation unsupported
Context
1M
Input / 1M
$0.20
Output / 1M
$1.20
Modalities
Protocol
2
  • chat.completions
  • responses

Status & performance

Last 7 days
Loading model performance

Capabilities

TextReasoningVisionLong contextCache

Pricing

Input tokens
$0.2/1M
Output tokens
$1.20/1M
Cache read
$0.02/1M
Cached input
$0.02/1M
Cache write
$0.25/1M
Reasoning tokens
$1.20/1M
input_text
$0.2/1M
Image
$0.2/1M
output_text
$1.20/1M
input_above_threshold
$0.4/1M
cached_input_above_threshold
$0.04/1M
output_above_threshold
$1.80/1M
input_image_above_threshold
$0.4/1M
Cache write (long context)
$0.5/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.