AI Pulse by Inblix

AI Costs Slashed

The Decoder · Jun 30, 2026 · 1 min read · Read original article →

Curated by the Inblix editorial team


Featured image for article: AI Costs Slashed

OpenAI has successfully reduced the costs of running its AI models by over half, specifically for guest users of ChatGPT. This was achieved through new optimizations that decreased the number of required Nvidia GPUs to just a few hundred. The techniques used to accomplish this are unclear, and it’s uncertain whether these gains can be applied to the full ChatGPT product. This development could lead to improved services, faster responses, or increased profitability. Why it matters: these cost savings could give AI labs the breathing room they need to scale and improve their services without being held back by high operational costs.

💡 Key Takeaways

  1. OpenAI cut inference costs for guest ChatGPT users by over half through new optimizations
  2. The number of Nvidia GPUs required to serve guest users decreased to just a few hundred
  3. The cost savings could be used to improve services, increase profitability, or scale operations

Keep reading: See related articles below for more coverage on this topic.

Get smarter about AI

The sharpest AI news, curated daily. Delivered free to your inbox.

Learn more

Glossary terms

← Back to all articles