Anthropic's Claude Opus Finally Powers Voice Mode With Real Tool Access

Claude's voice mode now runs on Opus and Sonnet, can call your connected tools mid-conversation, and speaks 18+ languages on every plan

·
·
Read4 min
TypeNews
TopicLlms · Api
  • Smarter models in voice: Claude voice mode now runs on Opus and Sonnet instead of Haiku, enabling real reasoning during spoken conversations.
  • Tool access mid-conversation: Claude can now reach connected tools like email and calendar while you're talking, without switching to chat.
  • 18+ languages on all plans: Multilingual support including Spanish, French, Hindi, and Japanese is now available to Free users too.
  • Rolling out now: The update is live in public beta on mobile, desktop, and web today.
  • Text-to-speech pipeline: Voice still runs on a conventional TTS stack (via ElevenLabs), not a speech-native model, so audio nuance is limited.
  • Quota counts apply: Voice conversations on Opus or Sonnet consume your standard message quota; Free users get roughly 20-30 conversations per day.

Claude's voice mode just got a meaningful upgrade. Anthropic is rolling out a public beta that lets voice conversations run on its more capable models, Opus and Sonnet, reach tools already connected to Claude, and speak in 18 languages beyond English. The update is live today on mobile, desktop, and web.

The model upgrade that actually changes things

Until now, voice mode ran on Claude Haiku 4.5 regardless of what the model selector showed. That selector had sat inside the voice interface for roughly three weeks, but the choice was cosmetic: whichever option you picked, the session still ran on Haiku 4.5. That changes with this release.

Selecting Opus or Sonnet now routes voice conversations through those models rather than silently falling back to Haiku. Haiku is Claude's fastest, lightest model, optimized for speed and low cost. Opus and Sonnet sit above it in the reasoning hierarchy, handling longer chains of thought, multi-step tasks, and complex instructions significantly better.

In testing, interruption handling held up well. The model pauses the moment it hears you and resumes sensibly, tolerating long silences without cutting in early. Routing Opus or Sonnet into voice opens tasks that were awkward before: longer reasoning chains, tool-heavy requests handled entirely by speech.

Your existing tools, now reachable by voice

Claude can now reach the integrations you've already set up in chat, mid-conversation. If you've connected Gmail, Google Calendar, or Google Drive, you can ask about them out loud and get a spoken answer back. No keyboard required.

For paying users, this makes voice mode feel like a spoken assistant rather than a novelty. Ask it to summarize your schedule, review a document, or pull details from a file while you're cooking or commuting. The same integrations you've configured in chat work immediately in voice, with no additional setup.

18 languages, out of beta

Anthropic has officially moved multilingual support out of beta, now covering 18 languages alongside push-to-talk. The announcement specifically calls out Spanish, French, Hindi, and Japanese. The update applies to every plan, including Free.

Both ChatGPT and Gemini have had multilingual voice support for some time. Claude's voice feature had stayed English-only until now. That gap is closed.

How the pipeline works

This is a text-to-speech pipeline, not a speech-native model like what OpenAI has been building. Claude handles the reasoning and drives the conversation; spoken output appears to run through ElevenLabs in the background, the same provider listed since the feature launched.

Where OpenAI is investing in a bidirectional, speech-native model, Anthropic is betting that putting its strongest reasoning models on a conventional voice stack delivers more value than chasing voice-specific capabilities. The tradeoff is real: you get Claude's full reasoning power spoken aloud, but without the emotional nuance or real-time audio understanding of a natively speech-trained system.

What this is good for

  • Talking through hard problems: Opus-level reasoning spoken back to you is useful for brainstorming, debugging logic, or working through a decision out loud.
  • Hands-free task management: Check your calendar or summarize an email thread while cooking, commuting, or between meetings.
  • Non-English workflows: Teams working primarily in Spanish, French, Hindi, Japanese, or other supported languages can now use voice natively.
  • Accessibility: Voice-first interaction with a capable model lowers the barrier for users who find typing slow or difficult.

Current limitations

Each voice interaction on Sonnet or Opus counts against your standard message quota. Free plan users should expect roughly 20 to 30 voice conversations per day before hitting limits.

The text-to-speech pipeline also means Claude won't pick up tone, emotion, or audio context the way a speech-native model could. Claude Opus 4.5 (Anthropic's top-tier model) is absent from voice for now.

How to get it

The update is rolling out today in public beta. Download the Claude app for iOS or Android, or open the desktop or web app, and tap the sound wave icon to start a voice session. The model selector in the voice interface now routes to the model you actually pick.

Comments

avatar