According to a report from The Information, OpenAI engineers have developed a series of new optimization techniques that reduce model inference costs by more than 50% and decrease reliance on Nvidia GPUs.
The report says OpenAI may pass on some of those savings to customers by lowering API pricing or increasing usage limits for products like ChatGPT.