Details
- OpenAI announces major pricing changes across its GPT-5.6 lineup, focused on cost efficiency and speed.
- GPT-5.6 Luna API pricing is reduced by 80%, and GPT-5.6 Terra by 20%, significantly lowering token costs for high-volume and mid-tier workloads.
- OpenAI introduces a Fast mode for GPT-5.6 Sol in the API, delivering up to 2.5x the speed of Standard processing at 2x the Standard price.
- Fast mode for Sol is positioned as a latency-focused option, providing substantially faster responses with no change in model intelligence or capability.
- Auto-review in the ChatGPT app and Codex CLI is upgraded from GPT-5.4 to GPT-5.6 Luna, aligning internal tools with the newer, cheaper model tier.
- OpenAI expects the Auto-review feature to be about 10x less expensive under Luna, making agentic and automation-heavy workflows far more cost-efficient.
- The company frames these changes as passing recent efficiency gains from GPT-5.6 Sol on to customers via lower Luna and Terra prices and new performance options.
- OpenAI reiterates its mission of making advanced intelligence more abundant and affordable as part of a broader effort to ensure AGI benefits all of humanity.
- The announcement underscores a pricing strategy that differentiates Sol as the high-capability flagship, Terra as the balanced default, and Luna as the fast, low-cost production workhorse.
Impact
By sharply cutting Luna and Terra prices while adding a fast, premium-latency mode for Sol, OpenAI deepens its tiered strategy against rivals that also segment models by cost and capability. The move lowers barriers for large-scale, agentic workflows and encourages developers to standardize on GPT-5.6, potentially accelerating migration from older GPT tiers and intensifying price competition in frontier-model APIs.