OpenAI Prepares a Voice Interface Revolution: New Device and Voice Mode in ChatGPT
OpenAI
OpenAI, in collaboration with Sam Altman and Jony Ive, is developing a new voice-centric device, likely a smart speaker. Meanwhile, the company has released Voice Mode in ChatGPT, which enables real-time, natural conversation with minimal latency, potentially replacing traditional input methods like keyboards and mice.
Recent reports indicate that Sam Altman and Jony Ive have been working on a joint project involving a new hardware device, now revealed to be a smart speaker focused on voice interaction. Although the smart speaker market is already crowded, the key differentiator is OpenAI's new Voice Mode in ChatGPT. Unlike simple dictation, Voice Mode listens continuously, responds almost instantly, and can even speak over the user, creating a natural conversational flow. Users have found it useful as a language tutor and translator. The combination of this device and Voice Mode suggests OpenAI is aiming to make voice the primary input method for computers, potentially replacing keyboards and mice. This shift would require AI to adapt to human speech rather than humans learning to use input devices. Ultimately, voice could be the first step toward a more natural, multimodal human-computer interface involving cameras, gestures, and context awareness. Siri is credited as a pioneer in voice interaction, but modern LLMs have advanced significantly. If this trend continues, keyboards may become specialized tools for professionals, and voice interfaces will become the next big wave for startups.
- Сокращения
- LLM = Large Language Model — Большая языковая модель
Source: Habr — хаб ИИ —
original
