OpenAI Slashes AI Prices Up to 80% and Launches Ultra-Fast API Mode as Efficiency Gains Transform Enterprise Costs
Summary
OpenAI slashes prices on select GPT-5.6 models by up to 80% and launches a new ultra-fast API mode delivering 2.5× faster processing speeds, as major efficiency gains across model architecture and inference systems dramatically reduce enterprise AI costs.
Key Points
- OpenAI is slashing prices on GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20%, making high-volume AI workloads significantly more affordable for businesses and enterprise customers.
- A new Fast mode is launching in the API for GPT-5.6 Sol, delivering up to 2.5× faster processing speeds than Standard mode at twice the price, replacing the previous Priority Processing offering.
- Efficiency gains are being driven by improvements across model architecture, inference systems, and agentic tooling, with GPT-5.6 Sol autonomously optimizing production kernels and cutting end-to-end serving costs by 20% while boosting token-generation efficiency by over 15%.