- Good as a general-purpose model for text generation and everyday reasoning.
- Good when you want package quota first, with pay-as-you-go enabled by key configuration.
GPT-5.6 Luna API Pricing and Package Multipliers
Efficient GPT-5.6 model for cost-sensitive agent and long-context workloads.
gpt-5.6-luna GPT-5.6 Luna API Pricing
| Channel | SU8 price | Billing unit | Quota multiplier | Effective multiplier |
|---|---|---|---|---|
| Azure | Input $0.57 Output $3.42 Cache read $0.057 Cache write $0.7125 | USD / 1M | 6.384x | 0.95x |
Cache hits are billed at the cache read price. Cache writes use the cache write price, and the Console usage detail shows the final cost breakdown.
GPT-5.6 Luna available packages and multipliers
| Package | Channel | SU8 price | Billing unit | Quota multiplier | Effective multiplier |
|---|---|---|---|---|---|
| Super Ultra | Azure | Input $0.5384 Output $3.2301 Cache read $0.0538 Cache write $0.6729 | USD / 1M | 63.479x | 0.8973x |
| Ultra | Azure | Input $0.5502 Output $3.3014 Cache read $0.055 Cache write $0.6878 | USD / 1M | 63.479x | 0.9171x |
| SU500 | Azure | Input $0.5078 Output $3.047 Cache read $0.0508 Cache write $0.6348 | USD / 1M | 63.479x | 0.8464x |
| SU750 | Azure | Input $0.5012 Output $3.0069 Cache read $0.0501 Cache write $0.6264 | USD / 1M | 63.479x | 0.8353x |
| SU1000 | Azure | Input $0.5078 Output $3.047 Cache read $0.0508 Cache write $0.6348 | USD / 1M | 63.479x | 0.8464x |
| Max | Azure | Input $0.6603 Output $3.9619 Cache read $0.066 Cache write $0.8254 | USD / 1M | 63.479x | 1.1005x |
GPT-5.6 Luna OpenAI-compatible API access
After creating a key in the console, call this model with the unified model name gpt-5.6-luna. When package and pay-as-you-go balance are both available, package quota is used first.
- Not ideal for legacy integrations that only accept OpenAI-compatible APIs.
- Not ideal when you need dedicated SLA terms, private capacity, or a fixed-cost enterprise commitment.
This model is available through the public catalog and unified gateway pricing configuration.
The public model page is for discovery and comparison. Actual API access, key creation, and model scopes are managed after sign-in.
Models worth comparing
GPT-5.6 flagship model for frontier reasoning, coding agents, and long-context production workloads.
Balanced GPT-5.6 model for high-quality agent and long-context workloads.
Lightweight Codex-oriented GPT model for quick code Q&A, small edits, and lower-cost development tasks.
