Sam Altman Slashes GPT-5.6 Prices by Up to 80%

Share:
Audio Loading voice…
Sam Altman Slashes GPT-5.6 Prices by Up to 80%

Synopsis

OpenAI CEO Sam Altman announced up to 80% price cuts across the GPT-5.6 model family on July 30, 2026. GPT-5.6 Luna drops to $0.20 per million input tokens, Terra falls 20%, and Sol gains a 2.5x-speed Fast mode at twice the price — reshaping cost calculus for API developers worldwide.

Key Takeaways

GPT-5.6 Luna receives an 80% price cut , now priced at $0.20 per million input tokens and $1.20 per million output tokens .
GPT-5.6 Terra drops 20% to $2 per million input tokens and $12 per million output tokens .
GPT-5.6 Sol gains a new Fast mode in the API offering up to 2.5x the speed at 2x the price , with no change in intelligence.
The cuts follow OpenAI's established pattern of reducing API pricing alongside model optimisations to deepen developer adoption.
The Luna tier's new pricing makes it directly competitive with the most aggressively priced large language model alternatives on the market.
The cost of building with OpenAI just fell off a cliff. OpenAI chief executive Sam Altman announced sweeping price cuts across the GPT-5.6 model family on Thursday, July 30, 2026 — with the sharpest reduction hitting GPT-5.6 Luna, which now costs just $0.20 per million input tokens and $1.20 per million output tokens, an 80% drop from previous rates.

Three models, three different cuts

The announcement covers the entire GPT-5.6 tier. GPT-5.6 Luna, the lightest of the three, takes the headline cut — 80% cheaper, making it one of the most affordable frontier-model options available via API. GPT-5.6 Terra, the mid-range variant, drops 20% to $2 per million input tokens and $12 per million output tokens. The third variant, GPT-5.6 Sol, does not get a price cut — it gets speed. Altman announced a new Fast mode for Sol in the API, delivering up to 2.5 times the throughput at twice the price, with no change to intelligence. That trade-off is aimed squarely at latency-sensitive production workloads where speed, not cost, is the bottleneck.

OpenAI's pattern: cut prices, chase the market

This is not OpenAI's first move of this kind. Since GPT-4's API launch in March 2023, the company has repeatedly reduced per-token pricing alongside model optimisations — a playbook designed to deepen developer adoption and defend market share as competitors close the capability gap. Each price reduction effectively lowers the floor for what developers can build and still turn a profit. For Indian API developers and enterprise AI teams, the Luna cut is particularly significant. At $0.20 per million input tokens, high-volume applications — chatbots, document processing pipelines, search augmentation — become dramatically cheaper to run at scale. The Luna tier now competes directly with some of the most aggressively priced alternatives in the market. The Sol Fast mode, meanwhile, signals a maturing API strategy: OpenAI is no longer selling only intelligence, it is selling a menu of intelligence-plus-performance combinations. Developers can now choose their trade-off explicitly — cheaper and slower, or faster and pricier, with the same underlying model quality. The cuts land as the broader AI infrastructure market faces intensifying pressure on margins. Cheaper tokens mean more experiments, more integrations, and more lock-in — which is precisely the point.

Point of View

While the Sol Fast mode opens a premium lane for latency-critical enterprise use cases. For the broader AI industry, this accelerates the commoditisation of inference, compressing margins across the stack. Developers who built cost models on last month's rates will need to reprice — and that friction is an opportunity OpenAI is betting on.
NationPress
31 Jul 2026

Frequently Asked Questions

What is the new price of GPT-5.6 Luna after the cut?
GPT-5.6 Luna now costs $0.20 per million input tokens and $1.20 per million output tokens following an 80% price reduction announced by OpenAI CEO Sam Altman on July 30, 2026.
How much did GPT-5.6 Terra prices drop?
GPT-5.6 Terra received a 20% price cut , bringing it to $2 per million input tokens and $12 per million output tokens .
What is GPT-5.6 Sol Fast mode?
GPT-5.6 Sol Fast mode is a new API option that delivers up to 2.5 times the processing speed at twice the standard price , with no reduction in model intelligence — designed for latency-sensitive applications.
Why is OpenAI cutting API prices?
OpenAI has a documented pattern of reducing API pricing alongside model optimisations to expand developer adoption and remain competitive as other AI providers offer similar large language models at lower costs.
How do OpenAI GPT-5.6 price cuts affect Indian developers?
Indian API developers and enterprise teams building high-volume applications — such as chatbots, document pipelines, and search tools — stand to benefit significantly, as the Luna tier's new pricing makes large-scale deployments considerably more affordable.
Nation Press
The Trail

Connected Dots

Tracing the thread behind this story — newest first.

8 Dots
  1. Latest 1 hour ago
  2. 2 weeks ago
  3. 2 weeks ago
  4. 2 weeks ago
  5. 3 weeks ago
  6. 1 month ago
  7. 1 month ago
  8. 1 month ago
Google Prefer NP
On Google