AIToday

Claude voice mode expands to Opus, Sonnet across all platforms

THE DECODER2h ago
Claude voice mode expands to Opus, Sonnet across all platforms

Key takeaway

Anthropic has expanded Claude's voice mode to run on its more capable Opus and Sonnet models, up from just Haiku, and made it available across all platforms (mobile, desktop, web). While competitors like OpenAI and Google offer smoother full-duplex audio conversation, Claude's unique strength is tool integration—it is the only service that lets users compose and send emails directly from voice mode, and it supports eleven languages with model switching during conversations.

Summaries like this, in your inbox every morning.

Sign up free →

3 Key Points

  • What happened

    Anthropic expanded Claude's voice mode to its Opus and Sonnet models (previously limited to Haiku), with the ability to switch models mid-conversation on mobile, desktop, and web. The mode supports eleven languages and can integrate with tools like Gmail, Google Calendar, and Slack to compose and send emails by voice.

  • Why it matters

    Claude's voice mode now competes more directly with OpenAI's GPT-Live and Google's Gemini Live. While those rivals use full-duplex audio (speaking and listening simultaneously) for a more natural feel, Claude uses a turn-based system and differentiates itself through tool integration—it is currently the only provider that lets users compose and save emails directly from audio mode.

  • What to watch

    Users can now choose which Claude model to use during voice conversations and switch languages flexibly, giving them control over speed and capability trade-offs depending on their task.

In Depth

Anthropic has extended voice mode support beyond its lightweight Haiku model to include Opus and Sonnet, making it available across the mobile app, desktop app, and web chat interface. The expanded service supports eleven languages and integrates with external tools including Gmail, Google Calendar, and Slack, allowing users to compose and send emails by voice without leaving the interface.

The competitive landscape shows distinct trade-offs. OpenAI's GPT-Live employs full-duplex audio, enabling the AI to speak and listen simultaneously, which creates a conversational experience closer to human phone dialogue. Google's Gemini Live adopts a similar turn-based approach to Claude but restricts availability to smartphones. In practice, both OpenAI and Google are perceived to feel more natural thanks to superior speech processing; all three services handle voice-based searches competently. Claude's singular advantage, according to the analysis, is tool integration: Anthropic is currently the only provider allowing users to compose and save emails directly within audio mode, a workflow efficiency that may resonate with users prioritizing productivity over conversational fluidity.

The update gives Claude users the ability to select their preferred model during a conversation and switch it mid-interaction, letting them balance latency and capability based on their needs. This granular control is available consistently across all three platform types, providing a uniform experience whether users engage via smartphone, computer, or web browser.

Context & Analysis

Anthropic's expansion of voice mode to its flagship Opus and Sonnet models marks a step toward feature parity with rival conversational AI services. OpenAI's GPT-Live and Google's Gemini Live have gained traction partly because their full-duplex architecture—allowing simultaneous speaking and listening—creates a more natural conversational flow. Claude's turn-based system, which waits for the user to finish before responding, is at a disadvantage in perceived naturalness. However, Anthropic has identified a differentiation strategy: tool integration. By enabling users to compose, edit, and send emails directly from voice mode without switching contexts, Claude addresses a practical workflow gap that its competitors do not yet offer. The expansion to Opus and Sonnet (Anthropic's more capable models) suggests the company is betting that task complexity and tool-assisted productivity will appeal to users willing to tolerate a less fluid conversational experience. The ability to switch models mid-conversation and across all platforms (mobile, desktop, web) gives users flexibility to optimize for speed or capability depending on the moment—a feature neither GPT-Live nor Gemini Live explicitly highlights.

FAQ

Which Claude models now support voice mode?
Opus and Sonnet models now support voice mode, in addition to Haiku. Users can switch between models mid-conversation.
What languages does Claude's voice mode support?
Claude's voice mode supports eleven languages.
What external tools can Claude's voice mode access?
Claude's voice mode can tap into connected tools like Gmail, Google Calendar, and Slack to compose and send emails by voice.

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Discussion

No comments yet. Be the first to share your thoughts!

Log in to join the discussion

Related Articles

Stay ahead with AI news

Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.

Get Started Free

Free · takes 30 seconds · unsubscribe anytime