OpenAI has reportedly achieved a major technical breakthrough, discovering a software-level optimization method that cuts its AI model inference costs by over 50%. By improving the efficiency of existing server resources rather than relying on new hardware, the technique significantly reduces the number of GPUs required to run models like ChatGPT. This development comes amid OpenAI's broader push to lower compute expenses, which includes the June 2026 unveiling of its custom Jalapeño chip co-developed with Broadcom, and intensifies competition with rivals like Anthropic and Meta.
Aug 7, 2026 · 5 sources
Aug 10, 2026 · 1 source
Aug 9, 2026 · 3 sources
Aug 8, 2026 · 7 sources
Story comments
Loading comments…