We're performing brief scheduled maintenance — saving, following, and sign-up are paused for a few minutes. You can keep reading as normal.

OpenAI slashes GPT-5.6 prices by up to 80% as businesses push back on AI costs

Reviewed byNidhi Govil

10 Sources

Share

OpenAI has cut prices on its GPT-5.6 Luna model by 80% and Terra by 20%, leaving its flagship Sol unchanged. The aggressive pricing strategy aims to counter cheaper Chinese alternatives and address growing cost sensitivity among enterprises. The move intensifies competition with Anthropic and Google while potentially straining finances ahead of anticipated IPOs.

OpenAI Cuts Prices on Smaller Models Amid Growing Cost Pressure

OpenAI has announced dramatic OpenAI price cuts across its GPT-5.6 lineup, slashing the cost of its smallest model by 80% and its mid-tier offering by 20% as businesses scrutinize AI spend with increasing intensity

1

. The ChatGPT maker reduced GPT-5.6 Luna pricing to $0.20 per million input tokens and $1.20 per million output tokens, down from $1 and $6 respectively, while GPT-5.6 Terra now costs $2 per million input tokens and $12 per million output tokens, compared to its previous $2.50 and $15 pricing

2

. The flagship Sol model remains unchanged, though OpenAI introduced a premium Fast mode at double the standard price.

Source: VentureBeat

Source: VentureBeat

The aggressive pricing strategy arrives roughly three weeks after the GPT-5.6 series public release and reflects mounting pressure from enterprises demanding clearer return on investment before deploying expensive AI models

2

. Companies have grown increasingly wary of unpredictable and often higher bills as AI firms shift from flat subscriptions to usage-based token pricing, making cost per task harder to estimate

4

.

AI Price Wars Intensify With Chinese Open-Source Alternatives

The OpenAI price cuts signal the beginning of full-scale AI price wars as American labs battle cheaper Chinese rivals for cost-sensitive customers

3

. OpenAI now faces direct competition from Chinese open-source alternatives like Z.ai's GLM-5.2 and DeepSeek's models that nearly match performance at lower costs

1

. With approximately 3GW of dedicated inference-related compute infrastructure running on NVIDIA GPUs, OpenAI had substantial room to implement these cuts, while China's AI labs operate with an estimated 500MW of inference compute on older hardware

5

.

Source: Reuters

Source: Reuters

The new Luna pricing places it below Google's Gemini 3.5 Flash-Lite at $2.80 combined per million tokens and far below Gemini 3.6 Flash at $9

3

. Terra now matches Google's Gemini 3.1 Pro Preview pricing at $14 combined per million tokens for context windows under 200,000 tokens. These competitive pricing strategies come just days after Anthropic released Claude Opus 5 at unchanged pricing and Google introduced two models focused on lower inference costs

3

.

Cost Efficiency Gains Enable Lower Token Pricing

OpenAI attributes the reduced pricing partly to efficiency gains from the GPT-5.6 architecture, including the model's ability to improve code and optimize performance during internal development

1

. The company maintains its strategy remains focused on advancing both capability and cost efficiency so each generation of AI models can accomplish more work at lower cost

2

. This approach benefits businesses broadly as smaller models can now handle work that recently required top-tier systems at significantly reduced expense.

The pricing adjustments turn up heat on Anthropic, whose Claude models dominate enterprise and developer use but sit at the costlier end of the market. Anthropic's mid-tier Claude Sonnet 4.6 costs $3 per million input tokens and $15 per million output tokens, above Terra's new rates

1

. Analysts note that cutting prices could boost usage of OpenAI's technology but may strain finances ahead of highly anticipated initial public offerings for both companies

4

.

Model Competition Shifts Toward Cost Sensitivity Among Enterprises

The dramatic reductions reflect a fundamental shift in model competition as the era of unlimited AI spending gives way to rigorous cost-benefit analysis. Many tech CEOs have emphasized in recent months that cheaper AI options are essential to the technology's widespread adoption

1

. OpenAI CEO Sam Altman announced the changes as "major price cuts" on social media, signaling the company's commitment to addressing enterprise concerns about AI spending.

Source: Wccftech

Source: Wccftech

The GPT-5.6 series offers distinct tradeoffs among intelligence, latency, and cost. Sol targets complex reasoning-heavy and agentic workloads including advanced coding and multi-step planning. Terra serves general production use requiring balanced capability and efficiency. Luna handles high-throughput, low-latency tasks like summarization, classification, and lightweight real-time assistants where cost per request is the primary constraint

3

. Industry observers now watch for responses from Chinese labs, which have historically proven formidable competitors in pricing battles

5

.

Today's Top Stories

© 2026 TheOutpost.AI All rights reserved