CRM & AutomationNew Release11 min readPublished July 25, 2026

Opus & Sonnet reach voice · 4 connectors named by Anthropic · turn-based by design

Claude Voice Mode Grows Up: Opus, Sonnet, and Connectors

On July 23, 2026, Anthropic lifted Claude voice mode out of its Haiku-only era. Paid plans can now talk to Opus and Sonnet, switch models mid-conversation, and — the part that matters for operations teams — have Claude read and act on Gmail, Calendar, Docs, and Slack aloud. Still beta, still turn-based, and Fable remains out of reach.

DA
Digital Applied Team
Senior strategists · Published Jul 25, 2026
PublishedJul 25, 2026
Read time11 min
SourcesAnthropic docs + 6 outlets
Voice models · paid plans
3
Opus · Sonnet · Haiku
was Haiku-only
Connectors in Anthropic’s doc
4
Gmail · Calendar · Docs · Slack
+Canva, Notion per press
Language coverage
~10langs
11 locale options
Free-tier connectors
1
connected tool in voice
Haiku-only model

Claude voice mode just had its most consequential update since launch. Per Anthropic’s support documentation and press briefings on July 23, 2026, the feature that had settled into running exclusively on Haiku — Anthropic’s smallest, fastest model — now lets paid plans talk directly to Claude Opus and Claude Sonnet, and lets voice conversations reach connected tools like Gmail, Google Calendar, Google Docs, and Slack.

Why this matters is less about novelty and more about what a voice interface is now allowed to touch. Before this update, talking to Claude meant talking to its least capable model, walled off from your actual working data. After it, a Pro or Max subscriber can ask Claude — out loud, hands busy — to check this afternoon’s calendar, summarize an email thread, or draft a Slack reply, with the same model quality they’d get in text chat. That shifts voice from a demo feature to a plausible operations surface.

This guide covers what actually shipped, the odd 14-month history most coverage skipped, a plan-by-plan entitlement matrix no single source has published, the connector details that matter for CRM and automation teams, and the architectural trade-off between Claude’s turn-based approach and the full-duplex direction OpenAI shipped the very same day.

Key takeaways
  1. 01
    Voice mode dropped its Haiku-only ceiling.Paid plans (Pro, Max, Team, Enterprise) can now select Claude Opus or Claude Sonnet in voice mode, switch models mid-conversation, and move between text and voice in the same chat without losing context. Free stays Haiku-only.
  2. 02
    Connectors now work while speaking.Anthropic’s own help-center doc names Gmail, Google Calendar, Google Docs, and Slack as connected tools reachable in voice; press coverage additionally demonstrates Canva and Notion. Claude asks permission before acting on any of them.
  3. 03
    The defaults are smart, not surprising.Voice mode defaults to whichever model you last used in text chat and runs that model’s fastest variant so exchanges stay real-time. Voice conversations draw from the same usage limits as text — no separate quota.
  4. 04
    It is still deliberately turn-based, and still beta.Claude listens, pauses to think, then responds — not the full-duplex simultaneous listening-and-speaking design of OpenAI’s GPT-Live. Anthropic says this release focused on intelligence and tool access, with no stated timeline for anything more.
  5. 05
    Fable is the one model voice still can’t reach.Anthropic’s flagship Mythos-class model, Claude Fable, is explicitly excluded from voice mode. The smartest model in the lineup remains text-only — a gap worth remembering before routing hard reasoning through the microphone.

01What ShippedSmarter models, spoken tool access, same beta label.

The update, announced on Anthropic’s product blog and detailed in its help-center article on voice mode (both dated July 23–24, 2026), has three load-bearing parts. First, model choice: paid subscribers can select Claude Opus or Claude Sonnet as their voice model, alongside the existing Haiku option, and switch between them mid-conversation via a model selector. Second, connector access: voice conversations can now reach the connected tools a user has authorized — more on that in Section 04. Third, continuity: you can move between text and voice inside the same chat without losing context, so a spoken commute conversation picks up where a typed desk session left off.

The defaults are worth noting because they reveal the design intent. Voice mode starts with whichever model you last used in text chat, and — per TechCrunch and Anthropic’s support doc — runs that model’s fastest variant so the exchange still flows in real time. Anthropic isn’t asking users to choose between speed and intelligence per-reply; it’s routing to the quickest version of the smartest model you already prefer.

Deep reasoning
Claude Opus
Paid plans · new to voice

Anthropic’s heavyweight workhorse model, now selectable in voice. The choice for thinking-out-loud sessions where you want follow-up questions and real analytical depth rather than quick answers.

Opus 4.8 · $5/$25 per Mtok in API
Balanced default
Claude Sonnet
Paid plans · new to voice

The mid-tier model most text users already default to. Strong enough for connector-driven work — calendar triage, email summaries — while staying responsive in conversation.

Sonnet 5 · $2/$10 intro to Aug 31
Speed floor
Claude Haiku
All plans · the incumbent

The smallest, fastest model — the only one voice mode ran on immediately before this update, chosen to minimize latency. Remains the sole option on the free plan.

Free tier’s only voice model
Sourcing note
Unusually for an Anthropic feature launch, there is no anthropic.com newsroom post for this update. The primary sources are Anthropic’s claude.com product blog (“Think through hard problems in voice mode”) and its help-center article on voice mode, both updated July 23–24, 2026 — corroborated same-week by TechCrunch, Engadget, 9to5Mac, MacRumors, and Dataconomy. Everything in this post traces to those sources.

02The Back StoryThe 14-month round trip from Sonnet to Haiku and back.

The history most of this week’s coverage skipped: Claude voice mode is not new, and it did not always run on Haiku. The feature launched in beta on May 27, 2025 — mobile-only, five voice options, Google Workspace integration for paid tiers — and per TechCrunch’s original launch coverage it defaulted to Claude Sonnet 4, not Haiku. At the time, Anthropic pitched it plainly: “Voice mode enables you to speak to Claude and hear responses through voice, making it easier to use Claude when your hands are busy but your mind isn’t.”

Somewhere between that May 2025 beta and the 2026 rollout, voice mode had settled into Haiku-only — an apparent cost-and-latency optimization that Anthropic’s own July 2026 announcement language acknowledges: Haiku handled quick exchanges well but struggled with complex requests. So the July 23 update is less a first than a reclamation. Roughly 14 months after launching on a mid-tier model, voice mode finally catches back up to the model quality text users take for granted — with one conspicuous exception.

That exception is Fable. Anthropic’s help center states it directly: “Claude Fable is not currently available in voice mode.” Fable 5 is the company’s flagship Mythos-class model, launched June 9, 2026 and positioned above Opus 4.8 in capability — which makes its exclusion the most telling detail in the release. Voice has graduated from the smallest model to the strong ones, but the smartest model in the lineup still cannot be talked to. If your workflow needs Fable-grade reasoning, it needs a keyboard.

"This release is focused on intelligence and tool access. We're continuing to invest in voice."— Anthropic, statement reported by Engadget, July 23, 2026

03EntitlementsWhat each plan actually gets in voice mode.

Every outlet covering this update describes model access, connector limits, and language support as separate bullet points scattered across articles. Nobody has assembled the full plan-by-plan grid — so here it is, built from Anthropic’s help-center article (the plan, model, and connector-count baseline) cross-checked against TechCrunch, 9to5Mac, and Engadget for the connector examples and language beta status.

Claude voice mode plan entitlement matrix, July 2026: voice models, connectors, languages, and platforms by plan tier
PlanVoice modelsConnectors in voiceLanguagesPlatforms
FreeHaiku only1 connected tool~10 languages, 11 locale options; non-English in beta, manual selectioniOS, Android, Desktop, Web
ProOpus, Sonnet, Haiku — switchable mid-conversationMultiple; Anthropic names Gmail, Calendar, Docs, SlackSame ~10 / 11 setiOS, Android, Desktop, Web
MaxOpus, Sonnet, Haiku — switchable mid-conversationMultiple; Canva and Notion demonstrated in press examplesSame ~10 / 11 setiOS, Android, Desktop, Web
Team / EnterpriseOpus, Sonnet, Haiku — switchable mid-conversationMultiple; same consent model as text chatSame ~10 / 11 setiOS, Android, Desktop, Web
Every planBeta feature on all tiers · Fable excluded everywhere · voice draws from the same usage limits as text — no separate voice quota

Two cells deserve a second look. The free tier’s single connected tool is a genuinely usable on-ramp — one Gmail or Calendar connection covers the highest-value spoken use case for most people. And the shared usage pool means a long Opus voice session burns the same allotment as a long Opus text session; teams already bumping against plan limits should treat voice as another drain on the same tank, not free capacity.

04ConnectorsYour inbox, calendar, and Slack — read aloud.

The model upgrade got the headlines, but connectors-in-voice is the change that matters for an operations audience. Anthropic’s own help-center doc names four connected tools explicitly: Gmail, Google Calendar, Google Docs, and Slack. Press coverage — TechCrunch among others — additionally demonstrates Canva and Notion in worked examples, though those two are press-reported rather than named in Anthropic’s own documentation. The worked examples across coverage are concrete: rescheduling a calendar meeting by voice, drafting or summarizing an email thread aloud, creating a document in Notion, updating a Canva design — all spoken, not typed.

The consent model carries over from text chat: Claude requires user permission before acting on any connected tool in voice. That matters more here than in text, because the failure mode of a spoken interface is acting on a misheard instruction — an authorization gate between “Claude heard something” and “Claude sent something” is the difference between a useful assistant and a liability.

For CRM and automation work, this is the shape of thing we’ve been building toward with clients: an AI layer that sits over real operational data — pipeline, inbox, calendar — and answers in whatever modality the moment demands. A spoken interface over connectors is essentially a hands-free front end to the same CRM automation patterns that already run in text: triage, summarize, draft, schedule. If Claude’s connector list covers your stack, the marginal cost of adding voice to an existing workflow is close to zero.

Anthropic-named
Connectors in the official doc
4

Gmail, Google Calendar, Google Docs, and Slack are the four tools Anthropic’s help center names for voice. Treat these as the confirmed floor of what voice can reach.

support.claude.com
Press-reported
Canva & Notion in demos
+2

TechCrunch and wire coverage demonstrate updating a Canva design and creating a Notion doc by voice. Attributed to press examples, not Anthropic’s own documentation — verify in-app before relying on them.

TechCrunch, Jul 23
Consent gate
Permission before action
1ask

The same connector-consent model as text chat: Claude requests authorization before acting on any connected tool. No spoken command silently mutates your inbox or calendar.

Same model as text
The ops angle
The practical hook isn’t “talking to an AI” — it’s that Claude can now function as a spoken interface over real operational data. Asking for this afternoon’s calendar or a summary of an email thread while driving between client sites is the use case; the novelty conversation is not. That’s why this lands in our CRM & Automation category rather than consumer news.

05ArchitectureTurn-based vs full-duplex: what each side trades away.

Most coverage treated this week’s two voice stories as unrelated news landing coincidentally together. They’re better read as one story: two labs making opposite architectural bets on the same day. Claude’s voice mode remains deliberately turn-based — it listens, pauses to think, then responds, as Engadget’s coverage lays out. OpenAI, meanwhile, brought its full-duplex GPT-Live voice family to the ChatGPT desktop app on July 23 — the same family that first shipped for mobile on July 8 — and we cover that whole launch, including the hands-free Codex angle, in our companion piece on ChatGPT’s desktop voice and hands-free agentic coding.

Turn-based versus full-duplex voice architectures compared: Claude voice mode against ChatGPT GPT-Live, July 2026
DimensionClaude voice mode (turn-based)ChatGPT / GPT-Live (full-duplex)
How it conversesListens, pauses to think, then responds — discrete turnsListens and speaks simultaneously — interruptible, continuous
Conversational feelLess natural; the pause is audible and by designCloser to human conversation rhythm
Model strength reachableOpus and Sonnet on paid plans (Fable excluded)The GPT-Live voice-model family itself
Tool / connector depthGmail, Calendar, Docs, Slack per Anthropic’s doc, with consent gatingVoice-commands ChatGPT, “Work,” and Codex on desktop per TechCrunch
MaturityBeta, all plans, all platformsRolling out on macOS and Windows for paid plans, Jul 23
What’s next, per vendor“Continuing to invest in voice” — no timeline statedExtending an existing full-duplex family across surfaces

Laid side by side, the trade is clean: today, from either vendor, you can have deep reasoning plus real connector access in voice, or natural simultaneous speech — not both. Anthropic acknowledged the naturalness gap directly rather than papering over it, telling Engadget the release focused on intelligence and tool access while it continues to invest in voice. Read that as a roadmap signal that something more speech-native is a direction, not a shipped or dated commitment — no source gives a timeline, and we won’t invent one.

Our interpretation: the two labs are optimizing for different jobs. OpenAI is betting voice wins on feel — the interface that disappears. Anthropic is betting voice wins on what it can legitimately do with your working data while your hands are elsewhere. For business workflows, the second bet is currently the more useful one; for ambient consumer assistants, the first. The coordinated timing — GPT-Live mobile on July 8, then both the Claude voice-models update and ChatGPT desktop voice on July 23 — says both labs now treat voice as a work surface worth fighting over, not a novelty checkbox.

06Fine PrintLanguages, limits, and the beta caveats.

Language support is the one spec where the coverage disagrees — instructively. TechCrunch and Engadget report “10 languages”; 9to5Mac and Anthropic’s own blog enumerate 11 locale entries. The gap is an accounting artifact: Spanish is counted once as a language but twice as a locale (Latin America and Spain). The honest phrasing is about 10 languages, 11 locale options — English, French, German, Hindi, Indonesian, Italian, Japanese, Korean, Brazilian Portuguese, and the two Spanish variants. Non-English languages remain in beta and must be selected manually; there is no automatic language-switch detection mid-conversation.

The other operational fine print, quickly: voice mode is described by Anthropic as a beta feature available on all plans — Free, Pro, Max, Team, and Enterprise — across iOS, Android, desktop, and web, though Anthropic says it’s built to work best from your phone. Voice conversations count toward the same usage limits as text. And the pricing context for the two models voice just gained: in the API, Opus 4.8 runs $5/$25 per million tokens, while Sonnet 5 carries an introductory $2/$10 through August 31, 2026, after which it moves to $3/$15 — a reminder that the “smarter voice” you’re consuming on a subscription is metered compute underneath.

07The PlaybookWhere voice-plus-connectors earns a place in your workflow.

Treat this like any other beta capability: find the workflows where the modality is the bottleneck, not the intelligence. Four decision lanes we’d give a client evaluating it this quarter:

Field & commute time
Spoken triage over Gmail + Calendar

The clearest win. Sales reps, field managers, and founders between meetings can have Sonnet summarize threads and check schedules hands-free. One connector on the free tier is enough to pilot this.

Adopt now, in beta
Thinking sessions
Opus as a talking partner

Anthropic’s own framing for this release is thinking through hard problems aloud — Claude asking follow-ups rather than dispensing answers. Useful for strategy walks; just remember Fable-grade reasoning still requires text.

Adopt for ideation
Regulated / high-stakes actions
Spoken writes to CRM & email

The consent gate helps, but beta voice plus irreversible actions is a risky pairing. Keep spoken interactions read-mostly — summaries, lookups, drafts for later review — until the feature exits beta.

Read-only for now
Customer-facing voice
Don’t confuse this with a voice agent

This is a personal assistant surface, not infrastructure for customer-facing voice experiences. For that lane — full-duplex models answering your customers — see our GPT-Live analysis and scope it as a build.

Different tool entirely

Looking forward: the pattern across both labs is that voice is becoming a front end to agents rather than a feature of chat apps. Claude’s version already reads your inbox aloud; OpenAI’s GPT-Live voice models are being positioned for customer-experience deployments; and Anthropic’s broader agent platform — effort controls, webhooks, skills — keeps maturing in parallel, as we covered in Claude’s managed-agents update. The reasonable projection is that within a few quarters the question stops being “can I talk to my AI” and becomes “which of my systems is it allowed to touch when I do.” Teams that wire up connectors, permissions, and clean CRM data now will be the ones for whom each new modality is a free upgrade — a big part of why we push AI transformation engagements toward data plumbing before interface novelty.

08ConclusionVoice mode finally catches up — except for the flagship.

The shape of AI voice, July 2026

The interesting question is no longer how natural the voice sounds — it’s what the voice is allowed to touch.

Fourteen months after launching, Claude voice mode has grown into something an operations team can take seriously: Opus and Sonnet on paid plans, mid-conversation model switching, and connectors that let a spoken request read and act on real Gmail, Calendar, Docs, and Slack data behind a consent gate. The beta label, the turn-based pauses, and the Fable exclusion are real limits — but they’re the limits of a feature being grown carefully, not one being neglected.

The same-day split-screen with OpenAI made the strategic picture unusually legible. One lab shipped conversational naturalness; the other shipped intelligence and tool access; neither offers both yet, and neither has committed to a date for closing its gap. For business workflows, we’d take the connector depth today and the naturalness later — a spoken summary of your actual pipeline beats a beautifully fluid conversation about nothing.

The practical move this week is small: connect one tool, pick Sonnet, and spend a commute triaging your inbox by voice. If that sticks, the wiring you do next — connectors, permissions, clean data — is the same wiring every future interface will sit on. Modalities keep changing; the plumbing compounds.

Put AI over your operational data

The teams winning with AI voice wired up their data first.

Our team designs and builds AI layers over real operational data — CRM, inbox, calendar, pipeline — so every new interface Anthropic or OpenAI ships becomes a free upgrade, not a new project.

Free consultationExpert guidanceTailored solutions
What we work on

AI-over-your-data engagements

  • CRM automation with AI triage, summaries & drafting
  • Connector & permission architecture for AI assistants
  • Voice and chat front ends over pipeline data
  • Multi-vendor AI routing — Claude, GPT, Gemini
  • Data plumbing that survives interface churn
FAQ · Claude voice mode

The questions we get every week.

Per Anthropic’s support documentation and press briefings, three things: paid plans (Pro, Max, Team, Enterprise) can now select Claude Opus or Claude Sonnet as their voice model instead of being limited to Haiku; voice conversations can reach connected tools — Anthropic’s own doc names Gmail, Google Calendar, Google Docs, and Slack; and users can switch models mid-conversation and move between text and voice in the same chat without losing context. The feature remains a beta available on all plans, including free, across iOS, Android, desktop, and web. Anthropic announced the update on its claude.com product blog rather than its main newsroom.
Related dispatches

Continue exploring AI at work.