- Good as a general-purpose model for text generation and everyday reasoning.
- Good when you want package quota first, with pay-as-you-go enabled by key configuration.
GPT-5.6 Luna API Pricing and Package Multipliers
Efficient GPT-5.6 model for cost-sensitive agent and long-context workloads.
gpt-5.6-luna GPT-5.6 Luna API Pricing
Cache hits are billed at the cache read price. Cache writes use the cache write price, and the Console usage detail shows the final cost breakdown.
GPT-5.6 Luna available packages and multipliers
GPT-5.6 Luna OpenAI-compatible API access
After creating a key in the console, call this model with the unified model name gpt-5.6-luna. When package and pay-as-you-go balance are both available, package quota is used first.
- Not ideal for legacy integrations that only accept OpenAI-compatible APIs.
- Not ideal when you need dedicated SLA terms, private capacity, or a fixed-cost enterprise commitment.
This model is available through the public catalog and unified gateway pricing configuration.
The public model page is for discovery and comparison. Actual API access, key creation, and model scopes are managed after sign-in.
Models worth comparing
GPT-5.6 flagship model for frontier reasoning, coding agents, and long-context production workloads.
Balanced GPT-5.6 model for high-quality agent and long-context workloads.
Lightweight Codex-oriented GPT model for quick code Q&A, small edits, and lower-cost development tasks.
