OpenAI's latest voice model invites conversation from users.

OpenAI's latest voice model invites conversation from users.
Summary
OpenAI launched GPT-Live, allowing ChatGPT Voice to listen and respond simultaneously.
ChatGPT now acknowledges users with short phrases while processing multiple requests concurrently.
Competitors like Thinking Machines aim to enhance AI interactions beyond traditional turn-based systems.

Share

Bookmark

Newsletter

OpenAI is transforming the way we engage with AI through its latest innovation, GPT-Live, which powers the new ChatGPT Voice feature. During a livestream event held on Wednesday, the company emphasized its goal to create a more human-like conversational experience with artificial intelligence.

With the new update, the assistant no longer waits for users to finish their sentences before responding. Instead, it can simultaneously listen and speak, allowing it to interact in a more dynamic and fluid manner. This means that as a person talks, the assistant can interject with acknowledgments such as "mmhm," "yeah," or "got it," enhancing the feel of natural dialogue.

In a demonstration, a user asked ChatGPT to verify the date of an upcoming meeting while also inquiring about the weather and traffic conditions. The assistant responded with brief confirmations like "hmm" and "sure," maintaining its focus and smoothly processing additional requests from the user in real-time.

Moreover, OpenAI showcased ChatGPT's upgraded capability for real-time language translation. Unlike previous assistants that required users to complete their thoughts before translating, this full-duplex model can listen and respond simultaneously, which aligns well with the pace of natural conversations, as highlighted by the company.

Greg Brockman, OpenAI’s president, characterized this update as a significant leap toward more intuitive interactions with technology, stating that it represents a "much more natural way of interacting with your computer."

OpenAI is not alone in this initiative; other companies are also striving to make AI voice assistants more engaging. In May, Thinking Machines—a lab led by former OpenAI CTO Mira Murati—revealed similar advancements. They expressed that their models are designed to facilitate ongoing interaction across various formats, including audio, video, and text, rather than following the usual stop-and-start pattern of conventional chatbots.

This development comes at a time when AI language models often find themselves the subject of comedic commentary online. Many creators poke fun at the overly enthusiastic tones, the repetitive use of phrases like "awesome," and the often awkward pauses that characterize AI responses.

Loading comments...