OpenAI Brings GPT-Live Voice Capabilities to Desktop Apps
- •OpenAI brings GPT-Live voice capabilities to macOS and Windows desktops in version 26.715.
- •The new voice model enables simultaneous task coordination and computer control via voice commands.
- •Available to Plus, Pro, Business, Edu, and Enterprise subscribers, with added macOS screen context support.
OpenAI has officially expanded its voice capabilities to desktop platforms, enabling users to interact with macOS and Windows applications using spoken commands. The release, designated as version 26.715, integrates the recently introduced GPT-Live model into the desktop environment. This update allows subscribers to manage complex tasks and control computer operations through natural language voice inputs, a functionality currently unavailable in Google’s Gemini Live. The system is designed to handle multiple processes simultaneously, such as speaking, listening, and executing tasks in the background, while maintaining conversational fillers like "mhmm" or "got it" to mimic human interaction.
Beyond voice control, the desktop update includes improvements to local file management, enabling users to create and organize related folders within local projects through the Edit project menu. On macOS devices, users can utilize the AppShot feature by enabling Screen context to share the frontmost active window. The integration also extends to ChatGPT Work and Codex, allowing users to direct agentic workflows entirely via voice. Furthermore, the model can generate visual content, such as weather charts or graphs, in response to user queries during conversation.
The upgraded voice experience is currently available to users subscribed to the Plus, Pro, Business, Edu, and Enterprise tiers. iPhone users maintain access to these features through the existing iOS Remote functionality. This global rollout reflects OpenAI's push to move beyond text-based interfaces, aiming to reduce the reliance on traditional keyboard and mouse input for routine desktop productivity.