Representative Image
OpenAI is also bringing Voice to the ChatGPT desktop experience on macOS and Windows. Users can use their voice to start tasks, check progress, ask questions about agents and coordinate multiple agents through a conversation. The feature is rolling out to Plus, Pro, Business, Edu and Enterprise plans.
What users can do with GPT-Live
GPT-Live-1 and GPT-Live-1 mini are already part of ChatGPT Voice, with availability and usage limits varying by plan. OpenAI has launched GPT-Live.
OpenAI says this allows Voice to move beyond conversations and help users get work done while their hands or attention are occupied. OpenAI says the models are designed to make spoken interaction more natural while allowing more complex work to be delegated to other models, tools or agents when required. This new generation of voice models is designed to make conversations between people and AI more natural, while also letting ChatGPT handle tasks during a voice conversation. The technology powers ChatGPT Voice and enables it to listen and speak at the same time, making interruptions and back-and-forth conversations more fluid. OpenAI is now extending these capabilities beyond simple voice chats, allowing users to access tools, connected apps and agentic features through voice on the web, mobile and desktop. GPT-Live is designed around what OpenAI calls a full-duplex voice experience. Instead of waiting for a person to finish speaking before responding, the model can listen while it speaks, follow pauses and interruptions, and adjust when the direction of a conversation changes. In an X post, the company wrote, “We’re bringing ChatGPT’s agentic capabilities to the voice experience on web and mobile for all users. For users with access to ChatGPT Work, Voice can also be used to create documents, presentations and spreadsheets, work with connected apps and use a browser. A Work task can continue in text if the user ends the voice conversation while the task is still running.

