CRM & AutomationNew Release11 min readPublished July 22, 2026

Narrated screen recording → rerunnable skill · Claude Desktop only · five weeks after Codex

Claude’s Record a Skill: Train Agents by Demonstration

Anthropic shipped Record a Skill inside Claude Cowork on July 21, 2026. Record your screen, narrate the task as you do it, and Claude turns the demonstration into a reusable skill it can run again — no prompt writing. It lands five weeks after OpenAI put the same bet into Codex, and it changes how CRM and back-office automation gets built.

DA
Digital Applied Team
Senior strategists · Published Jul 22, 2026
PublishedJuly 22, 2026
Read time11 min
SourcesAnthropic + 6 outlets
Launch gap vs Codex
33days
Jun 18 → Jul 21, 2026
Plan floor
$20/mo
Claude Pro · Free excluded
Ways to create a Skill
5
recording is the newest
Codex session cap
30min
Record & Replay limit

Claude’s Record a Skill, shipped inside Claude Cowork on July 21, 2026, replaces prompt writing with demonstration: you record your screen while doing a task, talk through your reasoning as you go, and Claude converts the walkthrough into a reusable skill it can run again on demand. No instruction files, no prompt engineering — the demo is the spec.

The stakes are bigger than one feature. OpenAI shipped the same idea — Record & Replay in the Codex desktop app — five weeks earlier, on June 18. Two frontier labs converging on demonstration-based skill capture within a single product cycle is a signal about where office automation is heading: away from typed instructions, toward watching how your team actually works.

This guide covers what launched and for whom, how the recording flow works, a structured head-to-head against Codex Record & Replay that no outlet has published, and the lens the general coverage missed entirely: which CRM and back-office tasks demo well, which need judgment, and how to govern recordings that inevitably point a camera at business data.

Key takeaways
  1. 01
    Recording replaces prompt writing for repeatable work.Record a Skill captures screen activity, mouse clicks, keystrokes, and spoken narration, then converts the demonstration into a structured skill saved to your library — rerunnable without repeating the walkthrough.
  2. 02
    Paid plans and the Desktop app only.Pro ($20/mo), Max ($100–$200/mo), and Team ($20–$125/seat/mo) users get it from the ‘+’ menu in the Claude Desktop app. The Free plan is excluded; the web app and mobile don’t have it.
  3. 03
    It’s Anthropic’s answer to Codex Record & Replay.OpenAI shipped the directly comparable feature on June 18, 2026 — 33 days earlier. Codex is macOS-only and blocked in the EEA, UK, and Switzerland; Claude’s version ships on Windows and macOS with no disclosed regional exclusions.
  4. 04
    CRM triage and report pulls demo well; judgment doesn’t.Rule-based work — lead triage, report pulls, data normalization — records cleanly. Exception handling, pricing calls, and anything touching credentials or customer PII belong on the do-not-record list.
  5. 05
    Documentation is still thin — govern before you scale.Anthropic had not published a dedicated help page covering recording retention, editing, or admin controls at launch. Treat every recording as potentially containing sensitive data until that documentation exists.

01What ShippedA new option in the “+” menu — paid plans, Desktop only.

Anthropic announced Record a Skill on July 21, 2026 via its official @claudeai account, and coverage from The Decoder, CyberSecurity News, and Search Engine Journal corroborated the details the same week. The feature lives in one place: the “+” menu of the Claude Desktop app, inside a Cowork session. It is not in the web app and not on mobile.

Access is gated to paid tiers. Every paid individual and team plan gets it; the Free plan is excluded entirely. That makes the practical entry price $20 per month — the same floor OpenAI set for Codex Record & Replay with its Plus tier.

Entry tier
Claude Pro
$20/mo

The cheapest plan with Record a Skill access. Free-plan users don’t get the feature at all — demonstration-based training is a paid capability from day one.

Free plan excluded
Power tier
Claude Max
$100–200/mo

The 5x and 20x usage tiers carry the feature with the higher Cowork limits heavy recording-and-rerunning workflows will actually consume.

5x / 20x usage
Team tier
Claude Team
$20–125/seat/mo

Standard and Premium seats both include it, with a 5-seat minimum. For ops teams, this is the tier where a shared library of recorded skills starts to compound.

5-seat minimum
Release snapshot
Record a Skill shipped July 21, 2026 in the Claude Desktop app’s “+” menu for Pro, Max, and Team subscribers. The recording captures screen activity, mouse clicks, keystrokes, and spoken narration, and the resulting skill is saved to your skill library for reuse without repeating the walkthrough. It extends Cowork’s existing automation layer — file access, persistent memory, connectors, and Cowork Projects — rather than introducing a new product surface.

02How It WorksThe demo is the spec.

The workflow Anthropic recommends, as summarized by Coursiv’s launch coverage, runs in eight steps: update and open the Claude Desktop app, start a Cowork task, open the “+” menu, choose Record a skill, perform the workflow yourself, narrate your choices, rules, and exceptions aloud as you work, let Claude convert the demonstration into a Skill, and then review and test that Skill on a fresh example before relying on it.

The narration is the part teams underestimate. The recorder captures what you click and type, but the spoken track is where the rules live — “we always check the branch field before assigning,” “if the amount is over the threshold, stop and escalate.” A silent demo produces a skill that mimics motions; a narrated demo produces one that carries reasoning.

“Record your screen while you do a task, talk through it as you go, and Claude turns it into a skill it can run again.”— @claudeai, Anthropic’s launch announcement, July 21, 2026

Under the hood, the output is a Claude Skill — a pre-existing product concept Anthropic describes in its Skills tutorial as a reusable package of task-specific instructions Claude loads automatically when the context matches. Recording doesn’t change what a Skill is; it changes how one gets made. Before July 21 there were already four creation paths — recording is the fifth, and the first that requires no writing at all.

Method 01
Write it manually
Files → Skill

Author the instruction files and supporting assets yourself. Maximum control, highest effort — the original path for teams with documented SOPs.

Hand-authored
Method 02
Describe it to Claude
Description → Skill

Explain the task in writing and ask Claude to build the Skill from your description. Fast, but the quality depends on how well you articulate edge cases.

Prompt-built
Method 03
Package a Cowork workflow
Session → Skill

Complete a task in Cowork once, then ask Claude to package what it just did into a reusable Skill. Good when Claude already executed the workflow well.

Session-derived
Method 04
Upload a prepared package
Upload → Skill

Import a Skill package built elsewhere — the distribution path for teams standardizing skills across seats and workspaces.

Imported
Method 05 · New
Record a demonstration
Demo + narration → Skill

Show Claude the task on screen while narrating the rules. The only method that requires no writing — and the only one that captures how work actually happens rather than how it’s documented.

Shipped Jul 21, 2026

Coursiv’s comparison of Claude’s customization primitives is a useful mental model here: Instructions set broad rules, Project knowledge holds per-project context, Memory carries facts across conversations — and Skills are loadable, task-specific procedures. Record a Skill is the newest way to create that last category, not a replacement for the other three.

03Head to HeadSame bet, 33 days apart.

OpenAI got there first. Codex Record & Replay shipped on June 18, 2026 in the Codex desktop app (v26.616) — 33 days, just under five weeks, before Anthropic’s July 21 launch. The Decoder put it plainly: “OpenAI launched a comparable Record and Replay feature in June, making demonstration-based skill capture an industry standard within a single product cycle.”

Existing coverage namechecks the parallel in a sentence and moves on. The structured comparison matters more than the namecheck, because the two implementations differ on exactly the dimensions an operations team has to care about: platform, regions, session limits, and — critically — what each vendor has actually disclosed about how the recording is represented.

Head-to-head comparison of Claude Cowork’s Record a Skill (launched July 21, 2026) and OpenAI Codex’s Record & Replay (launched June 18, 2026) across launch date, platform, plan floor, regional exclusions, recording limits, output artifact, internal representation, and documentation depth. Compiled from Anthropic’s launch announcement as corroborated by The Decoder, CyberSecurity News, and Coursiv, and from OpenAI’s Record & Replay documentation as covered in our Codex guide.
DimensionClaude Cowork · Record a SkillOpenAI Codex · Record & Replay
Launch dateJuly 21, 2026June 18, 2026 (desktop app v26.616) — 33 days earlier
PlatformClaude Desktop app (Windows + macOS); not the web app, not mobileCodex desktop app, macOS only at launch; requires Computer Use
Plan floorPro at $20/mo; also Max ($100–$200/mo) and Team ($20–$125/seat/mo, 5-seat minimum). Free plan excludedPlus at $20/mo; also Pro, Team, and Enterprise tiers. Paid plans only
Regional exclusionsNone disclosed at launchUnavailable in the EEA, UK, and Switzerland
Max recording lengthNot disclosed30 minutes per session
Output artifactA Skill in your library — conceptually a folder with a SKILL.md file Claude loads when relevant, per The DecoderA SKILL.md file — inspectable, editable markdown adopted across several major AI coding tools
How the demo is representedNot disclosed — Anthropic has published no technical description of the internal representationSemantic-intent JSON of actions and window content rather than raw pixel coordinates, per OpenAI’s docs — the stated reason replays survive small UI changes
Documentation depthThin — no dedicated help page at launch; retention and admin controls undocumentedOpenAI’s own docs describe the recording mechanics, limits, and regional availability

Two asymmetries stand out. First, reach: Claude’s version ships on Windows and macOS with no disclosed regional exclusions, while Codex is macOS-only and dark across the EEA, UK, and Switzerland — for a European or mixed-fleet operations team, Claude is currently the only one of the two you can actually roll out. Second, transparency runs the other way: OpenAI has documented how its recorder represents a demo — semantic intent as JSON, not pixel coordinates — while Anthropic has disclosed nothing about its internal representation. Whether Claude’s recorded skills survive UI changes the way Codex claims its replays do is, for now, an open question you can only answer by testing.

04The CRM LensWhat to record, what to keep human.

Every outlet covered this as a general productivity story. The more useful frame for operations teams is narrower: which CRM and back-office tasks actually demo well? Early coverage converges on a pattern — strong fits are repeatable, rule-based, and verifiable (weekly reporting, spreadsheet normalization, support-triage classification, publishing checklists); weak fits are one-off tasks, workflows driven by un-narrated intuition, unverifiable outputs, and anything touching money transfers, account recovery, contracts, passwords, or customer PII.

Mapped onto the workflows we build in CRM automation engagements, that pattern turns into a decision table. The dividing line isn’t task complexity — it’s whether the rules can be spoken aloud completely during a single demonstration.

Decision table mapping common CRM and back-office tasks to whether they are good candidates for Claude’s Record a Skill, why, and the governance note for each. Our framework, built on the strong-fit and weak-fit use cases and the pre-recording privacy checklist published in Coursiv’s launch coverage.
CRM / back-office taskRecord it?WhyGovernance note
Strong candidates
Lead-routing triageYesRule-based classification — status fields, branch assignment, source tagging — narrates cleanly in one passDemo on sample records with fake names and account numbers, never live customer data
Scheduled report pullsYesSame steps, same screens, every week — the archetypal repeatable, verifiable workflowStrip internal URLs, API keys, and tokens from anything visible on screen before recording
Data entry & spreadsheet normalizationYesDeterministic transforms with checkable outputs — errors surface immediately on reviewUse a sample dataset; inspect the generated Skill for accidentally copied identifiers before sharing
Conditional — record with a human gate
Support first-response draftingPartialTemplate selection and classification demo well; tone and edge-case judgment don’tKeep a human review step before anything customer-facing sends
Do not record
Exception handling & escalationsNoDriven by un-narrated intuition — the judgment that makes the call never makes it into the recordingKeep humans on judgment calls; a skill that mimics an escalation without the reasoning is a liability
Pricing & contract decisionsNoRadically input-dependent, and contracts sit explicitly on the flagged do-not-record listAutomate the paperwork around the decision, never the decision itself
Credentials, payments, customer PIINeverMoney transfers, account recovery, passwords, and medical records are all explicitly flagged as unsuitableClose password managers, email, and finance apps; disable notification previews before every recording

The narration test is the practical heuristic: if a competent new hire could do the task correctly after watching you do it once while you talked through the rules, it’s a candidate. If they’d need three months of context to know when the rules bend, it isn’t — and no recording feature from any vendor changes that.

05GovernanceThe documentation is thin — govern accordingly.

Here’s the honest gap in this launch: as of the announcement window, Anthropic had not published a dedicated technical or help page for Record a Skill — nothing covering recording retention, editing controls, or full rollout detail. Coursiv flagged the absence explicitly, and it means every governance decision right now has to lean on Anthropic’s pre-existing Cowork guidance rather than feature-specific policy.

That pre-existing guidance is worth reading before your first recording. Anthropic’s Use Claude Cowork safely page notes that Cowork can read files, browse the web, run code, and use connected apps, and that remote-session work is processed on Anthropic’s servers. Its computer-use guidance separately warns that information visible on screen during agentic use can include personal or confidential data. A feature whose entire input is your screen and your voice inherits both warnings at full strength.

AI Weekly’s launch analysis lists the questions the reporting doesn’t answer: what happens when the underlying app’s UI changes and a replay breaks, how the recording handles credentials or sensitive data scrolling past on screen, and whether skills can be shared or edited across a Team plan’s users. Until Anthropic documents those answers, the conservative reading is the right one — the same posture we argued for around enterprise agent-deployment governance: capability ships first, controls catch up, and the gap is where incidents happen.

Pre-recording checklist
The closest thing to a governance checklist anyone has published is third-party best practice from Coursiv, not Anthropic policy: close email, chat, password managers, finance and health apps, and unrelated customer records; disable notification previews; demo on a sample dataset with fake names and account numbers; strip API keys, tokens, and internal URLs from view; narrate the rule without speaking private facts aloud; and inspect the generated Skill for accidentally copied identifiers before sharing it. Treat it as a floor, not a ceiling.

06Strategic ReadThe desktop is the new battleground.

Why did both labs ship the same feature within a single product cycle? Because demonstration solves the cold-start problem that has throttled office-work automation since the first prompt was typed: most operational knowledge is tacit. The people who run your back office can’t write a complete spec of what they do — but they can show it, and they narrate the exceptions naturally while doing it. Recording converts tacit knowledge into an executable artifact at the cost of a single walkthrough, which is the same show-don’t-tell logic behind Claude’s shift toward shareable, show-me interfaces elsewhere in the product line.

The moat argument
AI Weekly’s editorial synthesis of the launch: “What is really being tested is whose desktop client owns the surface where knowledge workers demonstrate their tasks” — arguing that a shared library of company-specific skills is a stickier moat than any single benchmark win. That’s an analyst judgment, not a vendor claim, but it explains the five-week sprint: neither lab can afford to let the other become the default place work gets demonstrated.

Our read: the skill library is the real product. Once a team has recorded twenty workflows into one vendor’s desktop app — on the same Cowork surface where Anthropic’s flagship Fable 5 model already runs — switching costs stop being about model quality and start being about re-demonstrating institutional knowledge. Expect both vendors to push team-level sharing, versioning, and admin controls next, because that’s what converts a personal convenience into an organizational asset the org can’t easily leave.

Temper the projection with the reliability record, though. As we noted in our Codex coverage, the 2026 AI Index put average AI-agent success on the OSWorld computer-use benchmark at 66% — demonstration narrows the gap between what you meant and what the agent does, but it doesn’t make replays reliable on its own. And the anxiety is already public: Search Engine Journal’s coverage quoted one user reaction — “This has to be the easiest way to be replaced and lose your job man.” — which is sentiment, not data, but it’s the sentiment ops leaders will be managing in every rollout conversation. The teams that win this transition will frame recorded skills as removing the work people complain about, and keep the judgment calls visibly human.

07PlaybookA four-lane adoption plan.

If you run a CRM-centric operation, this feature is worth a pilot this quarter — scoped tightly. Sort your candidate workflows into four lanes and move one lane at a time.

Lane 1
Repeatable, rule-based back office

Report pulls, spreadsheet normalization, triage classification. Record on sample data, test on a fresh example per Anthropic’s own guidance, then run supervised for two weeks before trusting outputs.

Record now
Lane 2
Templated, customer-facing drafting

Support first responses, publishing checklists, brand-template cleanup. The mechanics record well; keep a named human approving anything that leaves the building.

Record + human gate
Lane 3
Judgment, exceptions, escalations

Pricing calls, contract terms, escalation decisions. The reasoning that matters is exactly what a screen recording can’t capture. Revisit only if your rules become fully explicit.

Keep human-run
Lane 4
Credentials, payments, PII

Money movement, account recovery, medical records, passwords. Explicitly flagged as unsuitable in every serious treatment of this launch — and undocumented retention makes the risk unquantifiable today.

Do not record

Two disciplines make the pilot real. First, review every generated skill like code: read it, test it on a fresh example, and check it for copied identifiers before it enters a shared library. Second, assign an owner — recorded skills are process documentation that executes, and unowned executable documentation rots into risk. If you want help scoping which workflows clear the bar and building the governance around them, that’s the assessment we run at the start of AI transformation engagements.

08ConclusionDemonstration just became table stakes.

The bottom line, July 2026

Show, don’t prompt — but govern what the camera sees.

Record a Skill is a genuinely lower floor for agent training: record a narrated walkthrough, get a rerunnable skill, no prompt writing. With Codex Record & Replay shipping the same bet 33 days earlier, demonstration-based skill capture went from novel to industry default inside a single product cycle — and Claude’s version is the one you can deploy on Windows and in Europe today.

The winners will be disciplined, not fast. The tasks that record well — triage, report pulls, normalization — are real capacity back for CRM and operations teams. The tasks that don’t record well haven’t changed, and Anthropic’s documentation on retention and admin controls hasn’t shipped yet. Pilot in the strong lanes, keep judgment human, and treat every recording as sensitive until the vendor tells you otherwise.

Put demonstration-based automation to work

Your best workflows deserve better than a silent recording.

Our team maps which of your CRM and back-office workflows are safe, high-ROI candidates for demonstration-based automation — and builds the governance so recorded skills help your team instead of exposing it.

Free consultationExpert guidanceTailored solutions
What we work on

CRM & automation engagements

  • Workflow audits — what to record vs keep human
  • Recorded-skill pilots with review-and-test gates
  • Governance checklists for screen-recorded automation
  • Zoho & custom-CRM triage and reporting automation
  • Claude vs Codex platform evaluation for your stack
FAQ · Claude Record a Skill

The questions we get every week.

Record a Skill is a feature Anthropic shipped inside Claude Cowork on July 21, 2026. Instead of writing prompts or instruction files, you record your screen while performing a task and narrate your reasoning out loud; Claude processes the screen activity, mouse clicks, keystrokes, and spoken narration into a structured, reusable skill. The skill is saved to your skill library and can be triggered again without repeating the walkthrough. Anthropic announced it via its official @claudeai account, and the details were corroborated the same week by outlets including The Decoder, CyberSecurity News, Search Engine Journal, and Android Authority.
Related dispatches

Continue exploring AI-powered automation.

CRM & Automation

Claude for Small Business: QuickBooks, HubSpot, Square

Anthropic launched Claude for Small Business on May 13 with QuickBooks, PayPal, HubSpot, and Slack integrations plus a 10-city US workshop tour.

May 16, 2026 · 14 minRead
CRM & Automation

Claude Cowork Enterprise: Tasks, Plugins Guide

Anthropic expands Cowork with Pro user access, recurring tasks, and Google Drive, Gmail, DocuSign, FactSet connectors. Complete enterprise deployment guide.

March 2, 2026 · 8 minRead
CRM & Automation

Baidu Unlimited-OCR 3B: Single-Pass PDFs, MIT-Licensed

Baidu's MIT-licensed Unlimited-OCR parses documents in one pass and beats DeepSeek-OCR on OmniDocBench. A free, self-hostable option for ops document teams.

July 23, 2026 · 12 minRead
CRM & Automation

Qwen-Audio-3.0-TTS Tops the Arena at a Third of the Price

Qwen-Audio-3.0-TTS-Plus tops the Artificial Analysis TTS arena (~1,236 Elo) at ~$27.59/1M chars, a third of ElevenLabs — but it is hosted-only, no weights.

July 22, 2026 · 10 minRead
CRM & Automation

Basis AI $100M: Agentic Accounting Tax and Audit

Basis AI raises $100M Series B at $1.15B valuation for agentic accounting. How AI agents transform tax preparation, audit workflows, and financial compliance.

March 1, 2026 · 11 minRead
CRM & Automation

AI Operations Management: Planning and Costing Guide

AI transforms operations management with demand forecasting, resource allocation, and automated costing. Practical guide for operations leaders in any industry.

February 21, 2026 · 14 minRead