Capability support
Context window1.05Mtokens
Official pricing
The model's published list prices, used only as the comparison baseline for route pricing and discounts
Input$0.201M
Output$1.201M
Cache read$0.02001M
Cache write$0.251M
Official context-tier pricing
Context range
Input
Output
Cache read
Cache write
Input tokens ≤ 272KStandard context
$0.20
$1.20
$0.0200
$0.25
Input tokens > 272KLong context
$0.40
$1.80
$0.0400
$0.50
Available routes
Provider, public group, live pricing, latency, and uptime per active route
Provider / groupInput1M tokensOutput1M tokensCache read1M tokensCache write1M tokensLatencyThroughputUptime
OpenAIGroup: gpt-5.6-luna50% OFF
Input$0.10$0.20
Output$0.60$1.20
Cache read$0.0100$0.0200
Cache write$0.13$0.25
Latency1.0sThroughput7.2 t/s
Uptime66.7%

