Tag
Sam Altman tweets that he now talks to ChatGPT more than he types, noting the new voice model has crossed a threshold in quality.
OpenAI's new GPT voice model enables highly realistic, real-time voice conversations with low latency and emotional expression, marking a significant leap in AI voice interaction.
OpenAI has fully rolled out GPT-Live, a new generation of voice models for natural human-AI interaction, to ChatGPT users on Go, Plus, and Pro plans, with free user rollout in progress.
Simba 3.2, claimed as the world's #1 voice model, now powers new voice agents.
OpenAI announces GPT-Live, a new full-duplex voice model that enables more natural, real-time conversations by allowing simultaneous listening and speaking, with GPT-5.5 as the backend model.
OpenAI's new voice model Bidi 1 first test exposure, supports bidirectional voice design, real-time translation, and stronger context memory, currently being pushed to a small group on ChatGPT.
Multiple AI model releases are delayed: GPT-5.6 now expected mid-July, DeepMind's 3.5 Pro postponed, while OpenAI's Bidi voice model and Claude Sonnet 5 for enterprises see progress.
OpenAI plans to release GPT-Bidi-1, its next-generation voice model that can listen and speak simultaneously, handle interruptions, and enable more natural conversations.
DramaBox is a highly expressive voice model based on LTX 2.3, released by Resemble AI with open-source code and models on GitHub and Hugging Face.
OpenAI released the GPT-Realtime-2 voice model, featuring GPT-5-level reasoning capabilities and a 128,000 token context window. It supports real-time translation from over 70 input languages to 13 output languages, achieving 96.6% accuracy on the Big Bench Audio Intelligence benchmark. Greg Brockman called it a milestone in voice translation.
OpenAI has released a new voice model, GPT-Live, capable of natural multi-task conversations and real-time execution of complex planning, such as arranging a one-day trip from Tokyo to Dubai to Hawaii.
OpenAI demonstrates the background robustness of its new voice model in noisy environments, accurately identifying conversation partners, understanding context, allowing users to interrupt naturally, and achieving smooth multi-person interaction.
OpenAI demonstrates the real-time simultaneous interpretation and conversation capability of the new voice model GPT-Live, which can listen and speak simultaneously and seamlessly switch between translation and chatting.
OpenAI showcases breakthrough in personalization and naturalness of the new voice model, capable of natural brainstorming conversations, displaying empathy and timely interjection, approaching human conversation rhythm.