New Delhi: OpenAI is leading in the AI race in terms of user base. As people grow increasingly exhausted by screens, the CEO, Sam Altman, led the company to focus on advancing its audio-based AI models. According to some reports, OpenAI’s upcoming device, which is being developed in collaboration with the former Apple’s ex-design chief John Ive, is expected to be largely audio-based. ChatGPT’s audio capabilities are pretty impressive. The AI models that power the AI chatbot’s following responses are different from the ones used to power its speaking capabilities.
OpenAI believe the current audio models lag behind the text-based models in the accuracy of their responses and how quickly they answer questions. Citing people familiar with the matter, the publication stated OpenAI is now working on unifying its engineering, product and research teams to improve its audio models for future devices. OpenAI’s efforts to improve on the audio front seem to be paying off. The company has also reported developing a new audio-model architecture which gives more natural, accurate and in-depth answers and speaks at the same time as the user, something which the current generation of audio AI models are unable to do.
The much-speculated model is expected to come out sometime in the first quarter of 2026. Researchers working for the ChatGPT maker have been focusing on audio AI models because the company’s upcoming device is stated to interact with users through speech instead of a screen. Jony Ive, Apple’s ex-design chief, has previously said that OpenAI’s forthcoming product is a priority for him. Rumours have it that the mysterious device could be in the shape of an AI-powered pen of some kind and may include two-way communication with ChatGPT.


