
OpenAI has introduced its latest conversational models, GPT-Live-1 and GPT-Live-1 mini, which promise a more natural and fluid interaction experience. These full-duplex models are designed to allow simultaneous speaking and listening, enabling users to interrupt seamlessly, a feature that enhances live translation capabilities. The company plans to replace its existing Advanced Voice Mode in ChatGPT with the GPT-Live-1 mini for all users, while those on paid plans will have access to the more advanced GPT-Live-1 model. Previously, the system relied on a combination of speech-to-text for transcription, a large language model for generating responses, and text-to-speech for output. However, the new models address common issues, such as the difficulty of interrupting ongoing speech and the limitations in contextual understanding. During a recent press briefing, OpenAI highlighted how these new models interact with their latest text generation models, like GPT-5.5, to enhance search and reasoning capabilities while maintaining the flow of conversation. The models can also remain silent to process ongoing dialogue and respond when prompted. Additionally, they leverage newer GPT capabilities to present information visually. In the competitive landscape, startups like Monogram, which recently secured $40 million in funding, are also exploring visual responses to make virtual assistants more engaging. OpenAI's voice mode is tailored for extended conversations, with product lead Atty Eleti sharing that he has experienced discussions lasting between 30 to 40 minutes while on the move. Looking ahead, OpenAI envisions voice technology becoming a primary interface for complex tasks. Although there are rumors of upcoming AI-powered earbuds, the company has not confirmed any hardware developments. Eleti stated, "We believe voice will evolve into an essential tool for managing intricate and prolonged agentic work, akin to the remarkable applications users have found for Codex and ChatGPT." Over recent years, OpenAI has focused on enhancing the naturalness of ChatGPT's voice features, with over 150 million users engaging through voice and dictation. Competitors such as Apple and Amazon are also striving to improve their assistants' conversational abilities. Meanwhile, startups like Sesame are launching AI assistants that prioritize natural dialogue while performing tasks in the background. OpenAI aims to facilitate hands-free communication with its assistant, focusing on providing age-appropriate responses and offering resources for sensitive topics. However, challenges remain; during a demonstration of the live translation feature in Hindi, the assistant's accent and delivery were noted to be less than optimal. The company acknowledged that while the new voice mode is designed for various languages, specific details regarding language support were not disclosed.
Prentis, a cutting-edge AI research laboratory co-founded by Ritankar Das, Reid Hoffman, and Marc Pincus, is currently i...
TechCrunch | Jul 24, 2026, 22:45
Waymo is reportedly exploring options to exit its partnership with Uber, which has allowed the Alphabet-owned firm to de...
TechCrunch | Jul 24, 2026, 21:00
Vietnam is contemplating a distinctive approach to youth social media regulations, diverging from the more common outrig...
TechCrunch | Jul 24, 2026, 21:25
Last week, OpenAI made its debut in the hardware landscape with the launch of Micro, a stylish keypad designed to integr...
TechCrunch | Jul 25, 2026, 24:40
It has been a challenging week for Elon Musk, as both Tesla and SpaceX experienced substantial stock declines. Tesla sha...
CNBC | Jul 24, 2026, 20:40