
Anthropic has expanded Claude's voice mode from its lightweight Haiku model to its more powerful Opus and Sonnet models, which can handle complex business reasoning and take actions like drafting pitches or rescheduling meetings. Voice mode is now also available in nine languages including Japanese, and integrated into apps like Gmail, Slack, and Canva—letting users switch between text and voice, and between models, mid-conversation depending on task complexity.
Summaries like this, in your inbox every morning.
Sign up free →What happened
Anthropic has made its Opus and Sonnet models available in voice mode—until now limited to the lighter Haiku model. The company is also bringing voice mode to apps like Gmail, Slack, and Canva, and expanding language support to French, German, Spanish, Hindi, Indonesian, Italian, Japanese, Korean, and Portuguese.
Why it matters
Haiku was designed for quick answers and kept conversations brief, but users started applying voice mode to deeper work—analyzing business problems and generating complex responses. Opus and Sonnet can handle that load: they can deliver detailed analysis, turn conversations into one-page pitches, or adjust your calendar if your train is late. For businesses and developers in Japan and other markets, this opens voice interaction to tasks that require real reasoning, not just fast replies.
What to watch
Users can now switch between text and voice mid-conversation and shift between models on the fly—so a quick Haiku chat can seamlessly escalate to Opus if the problem gets harder. Language support now includes Japanese, removing a barrier for native speakers.
When Anthropic launched voice mode last year, it was architected for simplicity: fast responses to quick questions, powered by Claude Haiku, the company's faster but less capable model. But real-world adoption quickly exceeded that design. Users began using voice mode not for casual queries but for deeper work—thinking through business problems, exploring ideas, and requesting actions that only more powerful reasoning could handle. Haiku kept those conversations quick, but not always deep. In response, Anthropic is now rolling out voice mode for Claude Opus and Sonnet, its more advanced models designed for what the company describes as "hard problem-solving." These models can deliver more complex responses and analysis, and crucially, can take action on a user's behalf. A user might dictate a business idea to Opus and ask it to turn the conversation into a one-page pitch, or mention that their train is running late and have the model adjust their calendar appointments accordingly. To make this flexible, Anthropic is allowing users to shift between text and voice mode within a single conversation, and to switch models mid-discussion. A quick question posed to Haiku might spark a deeper idea that a user wants to explore further—they can now seamlessly upgrade to Opus without starting a new thread. The company is also expanding the geographic and linguistic reach of voice mode. Until now, languages beyond English were available only in beta. Voice mode is now live in French, German, Spanish, Hindi, Indonesian, Italian, Japanese, Korean, and Portuguese—removing a significant friction point for non-English speakers who prefer speaking to typing.
Anthropic's original voice mode launch in 2025 centered on speed: quick answers with minimal delay through the lightweight Haiku model. However, the company discovered that real users were applying voice interaction to harder problems—working through business challenges that demanded sustained reasoning and follow-up actions. Haiku, designed to keep conversations quick, was not equipped for this deeper work. By expanding to Opus and Sonnet, Anthropic is responding to demonstrated user behavior rather than anticipated use. The models can now not only generate complex analysis but also take actions: drafting a one-page pitch from a conversation, or reshuffling calendar appointments based on a user's stated need. The mid-conversation model switching feature amplifies this: users can start lightweight and escalate only when necessary, avoiding the cost and latency of running the most powerful model for every query. The expansion of language support—especially the addition of Japanese alongside eight others—signals an intentional move into non-English markets where voice interaction removes typing friction.
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
No comments yet. Be the first to share your thoughts!
Log in to join the discussion





Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.
Get Started FreeFree · takes 30 seconds · unsubscribe anytime
1 minute a day. The AI essentials.
200+ sources · Email / LINE / Slack