2 Sources
[1]
Agentic coding goes hands free as OpenAI brings GPT-Live's full duplex voice control to Codex and ChatGPT on the desktop
Two weeks after debuting its more naturalistic GPT-Live audio AI model with full-duplex capabilities (listening and speaking at the same time), OpenAI is bringing it directly into developer workflows. The company announced that GPT-Live now powers the ChatGPT desktop application on macOS and Windows, integrating directly with agentic systems like Codex and ChatGPT Work (which are separate experiences available in the ChatGPT desktop app). When OpenAI initially launched GPT-Live on July 8, 2026, it introduced a continuous audio model capable of listening and speaking simultaneously -- eliminating rigid turn-taking while delegating complex reasoning to background models like GPT-5.5. Today's release expands that conversational layer to technical tasks, enabling software engineers to orchestrate multi-threaded coding jobs, review pull requests, and debug applications using natural voice commands. As such, it could usher in a new era of "hands free" software development and even live, in-person group coding parties for Codex's more than 5 million weekly active users. Codex, of course, is the name given to OpenAI's models and harness focused on coding, but which the company has this year expanded into a more general productivity platform. An OpenAI spokesperson told VentureBeat this is the first time voice activation has been included natively with Codex on the desktop. OpenAI posted a promotional video showing some of its employees, Codex developer experience engineer Jason Liu and Codex technical staffer Guinness Chen, speaking to the same ChatGPT desktop app session in the same room, each issuing different instructions and conversing with the same model. New capabilities unlocked At its core, this integration relies on decoupling the real-time voice layer from the underlying execution engines. While GPT-Live maintains fluid conversation -- inserting natural verbal acknowledgments like "got it" without interrupting the user -- it passes heavy computational workloads to background reasoning models. On macOS, the desktop application incorporates "Appshots" and screen context features, allowing ChatGPT Voice to analyze the frontmost window alongside local files, codebase structures, and active plugins. This architecture creates a pair-programming dynamic where developers talk through problems conversationally while agents execute tasks asynchronously. Rather than manually stopping coding sessions to type detailed instructions or switch windows, developers direct the system hands-free. The full-duplex engine dynamically decides when to speak, pause, or invoke tools, maintaining conversational state even as background agents process complex code modifications. Directing coding and complex builds with your voice alone The central operational capability in this update centers on multi-task execution across Codex and ChatGPT Work environments. Software engineers can initiate multiple concurrent task threads from a single spoken prompt. For instance, a developer preparing to ship a feature can instruct the system to investigate an open authentication bug, review a pending API migration pull request, and generate missing unit tests simultaneously. The desktop application coordinates these actions across disparate contexts, tracing issues through Slack conversations, GitHub repositories, and local codebases. Developers can also verbally convert design mockups into working code, splitting tasks across frontend, backend, and testing layers. With support for multi-folder projects (build 26.715) and remote execution via iOS, engineers can check task progress, answer agent prompts, and redirect active jobs without switching applications or managing individual processes line by line. Proprietary license OpenAI's voice-enabled desktop release operates under a proprietary, commercial enterprise model. Access is restricted to paid subscribers across Plus, Pro, Business, Enterprise, and Education plans. For individual developers and corporate engineering departments, this commercial structure means the model weights, voice processing pipelines, and agent state architectures remain fully closed. Organizations cannot modify or self-host the underlying systems. Furthermore, tasks initiated via ChatGPT Voice consume standard usage allocations directly from existing Codex and ChatGPT Work plan quotas, treating voice-triggered actions identically to standard agentic workloads. Community reactions Developer communities immediately noted the implications of bringing continuous full-duplex voice to autonomous coding workflows. Reacting to the build 26.715 release announcement -- which details voice integration and multi-folder project support -- AI Insider journalist @ChrisGPT noted on X: "Today OpenAI will release voice and remote guidance for codex ! One step closer to personal AGI". Early technical feedback highlights widespread enthusiasm for orchestrating complex agentic tasks hands-free, particularly when stepping away from the workstation or managing build pipelines remotely.
[2]
OpenAI debuts ChatGPT Voice so you can have ongoing conversations, ask the AI to complete tasks while hands-free | Fortune
OpenAI is adding its latest, most-advanced voice capabilities to the ChatGPT desktop app, allowing users to talk to their computers and ask the AI to complete tasks in a stream-of-consciousness, hands-free manner. The move extends the GPT Live voice technology -- which OpenAI released earlier this month on mobile devices -- to desktop and laptop PCs. That makes the voice technology accessible to software programmers, the AI power users who consume vast amounts of AI tokens in their jobs writing computer code and who are considered one of the most lucrative segments of the AI market. OpenAI said that ChatGPT Voice, the name of its new desktop voice feature, will work while coding through Codex, OpenAI's programming product which competes with Anthropic's Claude Code. Users can program a hotkey to launch ChatGPT Voice to talk to it while they're working in other apps, making it readily accessible, or use a designated Voice button within the ChatGPT desktop app. OpenAI is leaning into voice technology, viewing it as a competitive advantage in a crowded field of AI model makers. The company first rolled out voice recognition capabilities for its chatbot in 2024, but the newly released GPT Live technology provides more advanced capabilities that make conversations feel more natural. GPT-Live may also power the company's forthcoming hardware device, which Bloomberg reports will be an in-home speaker and conversation partner. "Today's launch is a step toward this future: talk through what you need, and ChatGPT can start moving multiple tasks forward at the speed of your thought," OpenAI said in a statement. According to the company, ChatGPT Voice will allow a user to kick off multiple work streams from a single conversation, and continue talking while agents complete the tasks in the background. The company gives the example of planning a work trip, with the user asking ChatGPT to check their calendar for conflicts, comb their inbox for any changes to the flight, and prepare notes for meetings -- all "while you make coffee or work on something else." The voice feature also lets users also talk through an idea with ChatGPT, while it creates a new document outlining the concept and next steps. OpenAI merged Codex into ChatGPT on July 9, forming a new product called ChatGPT Work. At the time, it said five million people use Codex every week. Following the launch of ChatGPT Work, Thibault Sottiaux, who oversees Codex, tweeted that the product had hit 10 million users.
Share
Copy Link
OpenAI has integrated its advanced GPT-Live voice technology into the ChatGPT desktop app, allowing developers to code hands-free through natural voice commands. The update brings full-duplex capabilities to Codex and ChatGPT Work, enabling 5 million weekly users to orchestrate multi-threaded coding tasks, debug applications, and manage complex workflows through ongoing conversations with AI.
OpenAI has brought ChatGPT Voice to its desktop application, marking a significant shift in how developers interact with AI coding assistants. Just two weeks after launching GPT-Live—its naturalistic audio AI model with full-duplex voice control—the company has integrated this technology directly into developer workflows on both macOS and Windows
1
. The integration powers the ChatGPT desktop application and works seamlessly with agentic systems like Codex and ChatGPT Work, enabling what OpenAI describes as hands-free software development for its more than 5 million weekly active users1
.This release extends voice technology to software programmers, considered AI power users who consume vast amounts of AI tokens in their jobs writing computer code—one of the most lucrative segments of the AI market . According to OpenAI, users can program a hotkey to launch ChatGPT Voice while working in other apps or use a designated Voice button within the desktop application
2
.
Source: VentureBeat
The core innovation behind this update lies in GPT-Live's ability to listen and speak simultaneously, eliminating rigid turn-taking that characterized earlier voice interfaces. OpenAI initially launched GPT-Live on July 8, 2026, as a continuous audio model that delegates complex reasoning to background models like GPT-5.5 while maintaining fluid conversation
1
. This architecture allows the system to insert natural verbal acknowledgments like "got it" without interrupting the user, creating a more human-like interaction pattern.The technology enables stream-of-consciousness commands, allowing developers to talk through problems conversationally while agents execute tasks asynchronously. Rather than manually stopping coding sessions to type detailed instructions or switch windows, engineers can direct the system entirely hands-free
1
. "Today's launch is a step toward this future: talk through what you need, and ChatGPT can start moving multiple tasks forward at the speed of your thought," OpenAI stated2
.The Codex integration unlocks powerful capabilities for hands-free coding workflows. Software engineers can initiate multiple concurrent task threads from a single spoken prompt, orchestrating complex operations across different contexts. For instance, a developer preparing to ship a feature can instruct the system to investigate an open authentication bug, review a pending API migration pull request, and generate missing unit tests simultaneously
1
.The desktop application coordinates these actions across disparate contexts, tracing issues through Slack conversations, GitHub repositories, and local codebases. On macOS, the application incorporates "Appshots" and screen context features, allowing ChatGPT Voice to analyze the frontmost window alongside local files, codebase structures, and active plugins
1
. This creates a pair-programming dynamic where background reasoning models handle heavy computational workloads while the voice layer maintains conversational flow.OpenAI posted a promotional video showing employees Jason Liu and Guinness Chen speaking to the same ChatGPT desktop app session in the same room, each issuing different instructions and conversing with the same model—suggesting potential for live, in-person group coding sessions
1
.ChatGPT Voice allows users to kick off multiple work streams from a single conversation and continue talking while agents complete tasks in the background. The company provides the example of planning a work trip, with users asking ChatGPT to check their calendar for conflicts, comb their inbox for flight changes, and prepare meeting notes—all "while you make coffee or work on something else"
2
.With support for multi-folder projects in build 26.715 and remote execution via iOS, engineers can check task progress, answer agent prompts, and redirect active jobs without switching applications or managing individual processes line by line
1
. The voice feature also enables users to talk through an idea with ChatGPT while it creates a new document outlining the concept and next steps2
.Related Stories
OpenAI's voice-enabled desktop release operates under a proprietary license model, with access restricted to paid subscribers across Plus, Pro, Business, Enterprise, and Education plans. The model weights, voice processing pipelines, and agent state architectures remain fully closed, meaning organizations cannot modify or self-host the underlying systems
1
. Tasks initiated via ChatGPT Voice consume standard usage allocations from existing Codex and ChatGPT Work plan quotas.OpenAI is leaning into voice technology as a competitive advantage in a crowded field of AI model makers. The company first rolled out voice recognition capabilities in 2024, but GPT-Live provides more advanced capabilities that make conversations feel more natural
2
. GPT-Live may also power the company's forthcoming hardware device, which Bloomberg reports will be an in-home speaker and conversation partner2
.Developer communities immediately noted the implications of bringing continuous full-duplex voice to autonomous coding workflows. Reacting to the build 26.715 release announcement, AI Insider journalist @ChrisGPT noted on X: "Today OpenAI will release voice and remote guidance for codex! One step closer to personal AGI"
1
. Early technical feedback highlights widespread enthusiasm for orchestrating complex agentic tasks hands-free.OpenAI merged Codex into ChatGPT on July 9, forming ChatGPT Work. Following that launch, Thibault Sottiaux, who oversees Codex, tweeted that the product had hit 10 million users
2
. An OpenAI spokesperson confirmed this is the first time voice activation has been included natively with Codex on the desktop1
.Summarized by
Navi
[1]
1
Technology

2
Policy and Regulation

3
Science and Research
