Skip to main content
OpenAI GPT-5.6 Luna AI model price reduction by 80%, with a futuristic AI chip and data streams.

Editorial illustration for OpenAI Slashes GPT-5.6 Luna AI Model Price by 80%

OpenAI Cuts GPT-5.6 Luna Price 80% This Week

OpenAI Slashes GPT-5.6 Luna AI Model Price by 80%

4 min read

OpenAI cut the price of GPT-5.6 Luna by 80% this week, dropping the smallest model in its frontier lineup to $0.20 per million input tokens and $1.20 per million output tokens. GPT-5.6 Terra, the mid-tier option, got a 20% cut to $2 and $12 per million tokens respectively. Sol, the flagship, keeps its $5/$30 pricing but gains a new Fast mode at double the cost, $10 per million input tokens and $60 per million output tokens, promising up to 2.5 times the throughput with no change to the underlying model.

The timing isn't incidental. Anthropic just released Claude Opus 5 at the same price point as Opus 4.8, and Google rolled out Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, both built for cheaper inference and faster agent workloads. OpenAI's Luna cut lands squarely in that low-cost tier Google has been courting, while the Sol Fast option targets a different kind of buyer entirely: one already inside Anthropic's ecosystem who might not blink at a higher bill if the response comes back quicker.

Three companies, three models, one week. Here's how OpenAI is framing the move.

OpenAI is still not the lowest-priced provider on a pure token basis. But Luna’s 80% reduction materially changes its position, moving it from the middle of the market into a pricing tier populated by smaller models from Google, Xiaomi, DeepSeek, MiniMax and other vendors.

Why this matters

For developers building on GPT-5.6, the math just changed overnight. Luna at $1.40 combined per million tokens instead of $7 makes it a viable default for high-volume tasks, classification, summarization, chat routing, where model quality per dollar matters more than raw capability. That's an 80% cut, not a rounding error, and it lands days after Anthropic's own pricing move, which tells us OpenAI isn't setting prices in a vacuum anymore.

Founders running lean should treat this as a signal to re-audit their model stack: if you're still paying Sol-tier prices for Luna-tier work, you're burning margin for no reason. The Fast mode addition on Sol is worth watching too. It suggests OpenAI is segmenting more aggressively, cheap and fast at the bottom, premium and fast at the top, squeezing the middle.

For researchers tracking the economics of frontier AI, this is the clearest evidence yet that inference costs, not just benchmark scores, are becoming the competitive battleground. Expect Anthropic and Google to respond in kind. Whether Terra's smaller 20% cut holds up against that pressure is the next thing to watch.

Common Questions Answered

What are the new pricing tiers for OpenAI's GPT-5.6 models after the recent price cuts?

GPT-5.6 Luna, the smallest model, was cut by 80% to $0.20 per million input tokens and $1.20 per million output tokens. GPT-5.6 Terra, the mid-tier option, received a 20% reduction to $2 and $12 per million tokens respectively. The flagship Sol model maintains its original $5/$30 pricing but now offers a new Fast mode at $10/$60 per million tokens with up to 2.5 times the throughput.

How does OpenAI's Luna pricing compare to competitors after the 80% price reduction?

While OpenAI is still not the lowest-priced provider on a pure token basis, Luna's 80% reduction has moved it from the middle of the market into a pricing tier now populated by smaller models from Google, Xiaomi, DeepSeek, MiniMax and other vendors. This pricing change materially improves OpenAI's competitive position in the cost-sensitive segment of the market.

What use cases make GPT-5.6 Luna viable for developers after the price cut?

Luna's new combined pricing of $1.40 per million tokens makes it a viable default for high-volume tasks including classification, summarization, chat routing, and other applications where model quality per dollar matters more than raw capability. The 80% price reduction fundamentally changes the economics for developers building on GPT-5.6 Luna, especially for those running lean operations.

What is the new Fast mode feature available for GPT-5.6 Sol?

GPT-5.6 Sol's new Fast mode is priced at double the standard rate ($10 per million input tokens and $60 per million output tokens) and promises up to 2.5 times the throughput compared to the standard mode. This option allows developers to trade higher costs for significantly improved processing speed when needed.

LIVE01:40Frozen CNN Feature Extractors Show Task-Dependent Sparsity in Reinforcement Learning