Editorial illustration for OpenAI Slashes GPT-5.6 Luna AI Model Price by 80%
OpenAI Cuts GPT-5.6 Luna Price 80% This Week
OpenAI Slashes GPT-5.6 Luna AI Model Price by 80%
OpenAI cut the price of GPT-5.6 Luna by 80% this week, dropping the smallest model in its frontier lineup to $0.20 per million input tokens and $1.20 per million output tokens. GPT-5.6 Terra, the mid-tier option, got a 20% cut to $2 and $12 per million tokens respectively. Sol, the flagship, keeps its $5/$30 pricing but gains a new Fast mode at double the cost, $10 per million input tokens and $60 per million output tokens, promising up to 2.5 times the throughput with no change to the underlying model.
The timing isn't incidental. Anthropic just released Claude Opus 5 at the same price point as Opus 4.8, and Google rolled out Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, both built for cheaper inference and faster agent workloads. OpenAI's Luna cut lands squarely in that low-cost tier Google has been courting, while the Sol Fast option targets a different kind of buyer entirely: one already inside Anthropic's ecosystem who might not blink at a higher bill if the response comes back quicker.
Three companies, three models, one week. Here's how OpenAI is framing the move.
Why this matters
For developers building on GPT-5.6, the math just changed overnight. Luna at $1.40 combined per million tokens instead of $7 makes it a viable default for high-volume tasks, classification, summarization, chat routing, where model quality per dollar matters more than raw capability. That's an 80% cut, not a rounding error, and it lands days after Anthropic's own pricing move, which tells us OpenAI isn't setting prices in a vacuum anymore.
Founders running lean should treat this as a signal to re-audit their model stack: if you're still paying Sol-tier prices for Luna-tier work, you're burning margin for no reason. The Fast mode addition on Sol is worth watching too. It suggests OpenAI is segmenting more aggressively, cheap and fast at the bottom, premium and fast at the top, squeezing the middle.
For researchers tracking the economics of frontier AI, this is the clearest evidence yet that inference costs, not just benchmark scores, are becoming the competitive battleground. Expect Anthropic and Google to respond in kind. Whether Terra's smaller 20% cut holds up against that pressure is the next thing to watch.
Common Questions Answered
What are the new pricing tiers for OpenAI's GPT-5.6 models after the recent price cuts?
GPT-5.6 Luna, the smallest model, was cut by 80% to $0.20 per million input tokens and $1.20 per million output tokens. GPT-5.6 Terra, the mid-tier option, received a 20% reduction to $2 and $12 per million tokens respectively. The flagship Sol model maintains its original $5/$30 pricing but now offers a new Fast mode at $10/$60 per million tokens with up to 2.5 times the throughput.
How does OpenAI's Luna pricing compare to competitors after the 80% price reduction?
While OpenAI is still not the lowest-priced provider on a pure token basis, Luna's 80% reduction has moved it from the middle of the market into a pricing tier now populated by smaller models from Google, Xiaomi, DeepSeek, MiniMax and other vendors. This pricing change materially improves OpenAI's competitive position in the cost-sensitive segment of the market.
What use cases make GPT-5.6 Luna viable for developers after the price cut?
Luna's new combined pricing of $1.40 per million tokens makes it a viable default for high-volume tasks including classification, summarization, chat routing, and other applications where model quality per dollar matters more than raw capability. The 80% price reduction fundamentally changes the economics for developers building on GPT-5.6 Luna, especially for those running lean operations.
What is the new Fast mode feature available for GPT-5.6 Sol?
GPT-5.6 Sol's new Fast mode is priced at double the standard rate ($10 per million input tokens and $60 per million output tokens) and promises up to 2.5 times the throughput compared to the standard mode. This option allows developers to trade higher costs for significantly improved processing speed when needed.
Further Reading
- OpenAI slashes API prices for GPT-5.6 lineup as cost worries mount - CNBC
- OpenAI cuts prices on smaller models as businesses push back on AI costs - Reuters via Yahoo Finance
- AI price wars: OpenAI cuts GPT-5.6 Luna prices by 80% as model competition shifts toward cost - VentureBeat
- OpenAI slashes Luna pricing - Axios
- OpenAI just made its cheapest AI model dramatically cheaper - Notebookcheck