
OpenAI goes full China pricing mode with an 80 percent cut to its most affordable GPT-5.6 model
Quick Answer
OpenAI has slashed prices for its GPT-5.6 Luna model by 80% to $0.20 per million tokens, while Terra sees a 20% reduction.
Quick Take
This price drop, attributed to improved infrastructure efficiency, positions Luna as a cost-effective alternative, potentially impacting revenue growth for competitors amid a price war driven by low-cost Chinese providers.
Key Points
- Luna now costs $0.20 per million input tokens and $1.20 for output tokens.
- Terra prices reduced to $2 for input and $12 for output tokens.
- Luna matches last year's leading models, costing 6 cents for tasks that were $1.
- OpenAI's infrastructure improvements cut deployment costs by 20% and boosted token generation by 15%.
- Microsoft promotes its MAI models as cheaper alternatives, intensifying market competition.
📖 Reader Mode
~1 min readOpenAI is cutting GPT-5.6 Luna prices by 80 percent and Terra by 20 percent, effective July 30. Luna drops to $0.20 per million input tokens and $1.20 per million output tokens, while Terra falls to $2 and $12. Sol pricing stays the same. OpenAI says Luna matches the performance of leading models from a year ago, but a task that cost a dollar with those models now runs about 6 cents on Luna, nearly nine times faster. All models are available through ChatGPT Work, Codex, and the OpenAI API.

OpenAI says the cuts are possible because GPT-5.6 Sol made the company's own infrastructure more efficient. The model allegedly optimized GPU software on its own, cutting deployment costs by 20 percent. It also improved token generation by more than 15 percent through speculative decoding.
Growing price pressure across the AI market likely played a role too, especially from low-cost Chinese providers. Microsoft is now openly promoting its own MAI models as cheaper alternatives to OpenAI. The price war could hurt the broader market if it slows revenue growth at frontier labs whose balance sheets are tied to massive infrastructure investments.
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

