gpt-5.6-luna
GPT-5.6 model optimized for cost-sensitive workloads
Context
1M
Input / 1M
$0.20
Output / 1M
$1.20
Modalities
Protocol
2- chat.completions
- responses
Status & performance
Last 7 days
Loading model performance
Capabilities
TextReasoningVisionLong contextCache
Pricing
Input tokens
$0.2/1M
Output tokens
$1.20/1M
Cache read
$0.02/1M
Cached input
$0.02/1M
Cache write
$0.25/1M
Reasoning tokens
$1.20/1M
input_text
$0.2/1M
Image
$0.2/1M
output_text
$1.20/1M
input_above_threshold
$0.4/1M
cached_input_above_threshold
$0.04/1M
output_above_threshold
$1.80/1M
input_image_above_threshold
$0.4/1M
Cache write (long context)
$0.5/1M
Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.
