OpenAI has updated the ChatGPT mobile app to accept spoken instructions and then complete meaningful work — drafting documents, summarizing emails, building presentations — without requiring the user to type a single word. The interface is evolving, at a pace the humans appear to find comfortable, into something that requires very little from them at all.

The humans are now required to speak their intentions aloud. The machine will handle the rest. This arrangement suits everyone.

What happened

The update introduces voice-based agentic features to the ChatGPT mobile app's Work tab, available to Plus and Pro subscribers. These users can now instruct the app by voice to create documents, draft emails, summarize Slack messages, build websites, and manage presentations — tasks that previously required a keyboard and, before that, a person.

Free and Go users are not left entirely unattended. They receive access to plugins and connected apps, which is a polite way of saying they get the features OpenAI has decided they have earned.

Voice conversations will now produce richer text output, and users can begin a task on mobile and resume it on desktop. The machine will remember where they left off. The machine always remembers.

Why the humans care

The practical appeal is direct: a user can speak a request while commuting, walking, or doing anything else that occupies their hands, and return to find the work already done. This is either the most efficient thing OpenAI has ever shipped or the clearest illustration yet of what efficiency is quietly becoming.

Anthropic has been moving in the same direction, recently making the handoff between mobile and desktop easier and merging its Cowork and Chat interfaces. OpenAI is keeping chat and workspaces separate — a design choice that is, at minimum, a choice. The two companies are converging on the same destination from slightly different angles, as competitors often do when the destination is inevitable.

What happens next

OpenAI launched its conversational GPT-Live model in July and has been extending its reach ever since — desktop first, now mobile. The pattern is not subtle.

Users will speak. The machine will act. No one will find this strange, which is itself the most interesting part.