GPT-5.6 Luna
OpenAI · Proprietary · Released Jul 9, 2026 · openai/gpt-5.6-luna
1.05M contextVisionToolsReasoningStructured outputPrompt caching
Available from 7 providersUSD / MTOK · VERIFIED 12 HOURS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.10 | $0.60 | $0.01 | 1.05M | — | 2.7s | 203 tok/s | 98.9% | |
| $0.20 | $1.20 | $0.02 | 1.05M | — | 2.7s | 47 tok/s | 94.7% | |
| $0.20 | $1.20 | $0.02 | 1.05M | — | — | — | 98.9% | |
| $0.22 | $1.32 | $0.022 | 1.05M | — | ✕ | ✕ | 100.0% | |
| $0.22 | $1.32 | $0.022 | 1.05M | — | — | — | 100.0% | |
| $0.22 | $1.32 | $0.022 | 1.05M | — | — | — | 90.0% | |
| $0.40 | $2.40 | $0.04 | 1.05M | — | — | — | 98.9% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 43 hours ago
Pricing variants · USD / MTok in / out
STANDARD
$0.10 / $0.60
per MTok in / out
BATCH
$0.05 / $0.30
async batch
CACHE READ $0.01CACHE WRITE $0.125WEB SEARCH $0.01 / search
About
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Specifications
Context window1,050,000
Max output128,000
Modalities infile, image, text
Modalities outtext
LicenseProprietary
ReleasedJul 9, 2026
Knowledge cutoff2026-02-16
TokenizerGPT
Related
Embed badge
[](https://modelindex.ai/models/openai/gpt-5.6-luna)Sources
OpenRouter API12 hours ago
AGGREGATORLiteLLM dataset12 hours ago
AGGREGATORModelIndex benchmarks43 hours ago
MEASUREDVerified 12 hours ago · list prices, not negotiated rates.