Claude’s Record a Skill, shipped inside Claude Cowork on July 21, 2026, replaces prompt writing with demonstration: you record your screen while doing a task, talk through your reasoning as you go, and Claude converts the walkthrough into a reusable skill it can run again on demand. No instruction files, no prompt engineering — the demo is the spec.
The stakes are bigger than one feature. OpenAI shipped the same idea — Record & Replay in the Codex desktop app — five weeks earlier, on June 18. Two frontier labs converging on demonstration-based skill capture within a single product cycle is a signal about where office automation is heading: away from typed instructions, toward watching how your team actually works.
This guide covers what launched and for whom, how the recording flow works, a structured head-to-head against Codex Record & Replay that no outlet has published, and the lens the general coverage missed entirely: which CRM and back-office tasks demo well, which need judgment, and how to govern recordings that inevitably point a camera at business data.
- 01Recording replaces prompt writing for repeatable work.Record a Skill captures screen activity, mouse clicks, keystrokes, and spoken narration, then converts the demonstration into a structured skill saved to your library — rerunnable without repeating the walkthrough.
- 02Paid plans and the Desktop app only.Pro ($20/mo), Max ($100–$200/mo), and Team ($20–$125/seat/mo) users get it from the ‘+’ menu in the Claude Desktop app. The Free plan is excluded; the web app and mobile don’t have it.
- 03It’s Anthropic’s answer to Codex Record & Replay.OpenAI shipped the directly comparable feature on June 18, 2026 — 33 days earlier. Codex is macOS-only and blocked in the EEA, UK, and Switzerland; Claude’s version ships on Windows and macOS with no disclosed regional exclusions.
- 04CRM triage and report pulls demo well; judgment doesn’t.Rule-based work — lead triage, report pulls, data normalization — records cleanly. Exception handling, pricing calls, and anything touching credentials or customer PII belong on the do-not-record list.
- 05Documentation is still thin — govern before you scale.Anthropic had not published a dedicated help page covering recording retention, editing, or admin controls at launch. Treat every recording as potentially containing sensitive data until that documentation exists.
01 — What ShippedA new option in the “+” menu — paid plans, Desktop only.
Anthropic announced Record a Skill on July 21, 2026 via its official @claudeai account, and coverage from The Decoder, CyberSecurity News, and Search Engine Journal corroborated the details the same week. The feature lives in one place: the “+” menu of the Claude Desktop app, inside a Cowork session. It is not in the web app and not on mobile.
Access is gated to paid tiers. Every paid individual and team plan gets it; the Free plan is excluded entirely. That makes the practical entry price $20 per month — the same floor OpenAI set for Codex Record & Replay with its Plus tier.
Claude Pro
The cheapest plan with Record a Skill access. Free-plan users don’t get the feature at all — demonstration-based training is a paid capability from day one.
Claude Max
The 5x and 20x usage tiers carry the feature with the higher Cowork limits heavy recording-and-rerunning workflows will actually consume.
Claude Team
Standard and Premium seats both include it, with a 5-seat minimum. For ops teams, this is the tier where a shared library of recorded skills starts to compound.
02 — How It WorksThe demo is the spec.
The workflow Anthropic recommends, as summarized by Coursiv’s launch coverage, runs in eight steps: update and open the Claude Desktop app, start a Cowork task, open the “+” menu, choose Record a skill, perform the workflow yourself, narrate your choices, rules, and exceptions aloud as you work, let Claude convert the demonstration into a Skill, and then review and test that Skill on a fresh example before relying on it.
The narration is the part teams underestimate. The recorder captures what you click and type, but the spoken track is where the rules live — “we always check the branch field before assigning,” “if the amount is over the threshold, stop and escalate.” A silent demo produces a skill that mimics motions; a narrated demo produces one that carries reasoning.
“Record your screen while you do a task, talk through it as you go, and Claude turns it into a skill it can run again.”— @claudeai, Anthropic’s launch announcement, July 21, 2026
Under the hood, the output is a Claude Skill — a pre-existing product concept Anthropic describes in its Skills tutorial as a reusable package of task-specific instructions Claude loads automatically when the context matches. Recording doesn’t change what a Skill is; it changes how one gets made. Before July 21 there were already four creation paths — recording is the fifth, and the first that requires no writing at all.
Write it manually
Author the instruction files and supporting assets yourself. Maximum control, highest effort — the original path for teams with documented SOPs.
Describe it to Claude
Explain the task in writing and ask Claude to build the Skill from your description. Fast, but the quality depends on how well you articulate edge cases.
Package a Cowork workflow
Complete a task in Cowork once, then ask Claude to package what it just did into a reusable Skill. Good when Claude already executed the workflow well.
Upload a prepared package
Import a Skill package built elsewhere — the distribution path for teams standardizing skills across seats and workspaces.
Record a demonstration
Show Claude the task on screen while narrating the rules. The only method that requires no writing — and the only one that captures how work actually happens rather than how it’s documented.
Coursiv’s comparison of Claude’s customization primitives is a useful mental model here: Instructions set broad rules, Project knowledge holds per-project context, Memory carries facts across conversations — and Skills are loadable, task-specific procedures. Record a Skill is the newest way to create that last category, not a replacement for the other three.
03 — Head to HeadSame bet, 33 days apart.
OpenAI got there first. Codex Record & Replay shipped on June 18, 2026 in the Codex desktop app (v26.616) — 33 days, just under five weeks, before Anthropic’s July 21 launch. The Decoder put it plainly: “OpenAI launched a comparable Record and Replay feature in June, making demonstration-based skill capture an industry standard within a single product cycle.”
Existing coverage namechecks the parallel in a sentence and moves on. The structured comparison matters more than the namecheck, because the two implementations differ on exactly the dimensions an operations team has to care about: platform, regions, session limits, and — critically — what each vendor has actually disclosed about how the recording is represented.
| Dimension | Claude Cowork · Record a Skill | OpenAI Codex · Record & Replay |
|---|---|---|
| Launch date | July 21, 2026 | June 18, 2026 (desktop app v26.616) — 33 days earlier |
| Platform | Claude Desktop app (Windows + macOS); not the web app, not mobile | Codex desktop app, macOS only at launch; requires Computer Use |
| Plan floor | Pro at $20/mo; also Max ($100–$200/mo) and Team ($20–$125/seat/mo, 5-seat minimum). Free plan excluded | Plus at $20/mo; also Pro, Team, and Enterprise tiers. Paid plans only |
| Regional exclusions | None disclosed at launch | Unavailable in the EEA, UK, and Switzerland |
| Max recording length | Not disclosed | 30 minutes per session |
| Output artifact | A Skill in your library — conceptually a folder with a SKILL.md file Claude loads when relevant, per The Decoder | A SKILL.md file — inspectable, editable markdown adopted across several major AI coding tools |
| How the demo is represented | Not disclosed — Anthropic has published no technical description of the internal representation | Semantic-intent JSON of actions and window content rather than raw pixel coordinates, per OpenAI’s docs — the stated reason replays survive small UI changes |
| Documentation depth | Thin — no dedicated help page at launch; retention and admin controls undocumented | OpenAI’s own docs describe the recording mechanics, limits, and regional availability |
Two asymmetries stand out. First, reach: Claude’s version ships on Windows and macOS with no disclosed regional exclusions, while Codex is macOS-only and dark across the EEA, UK, and Switzerland — for a European or mixed-fleet operations team, Claude is currently the only one of the two you can actually roll out. Second, transparency runs the other way: OpenAI has documented how its recorder represents a demo — semantic intent as JSON, not pixel coordinates — while Anthropic has disclosed nothing about its internal representation. Whether Claude’s recorded skills survive UI changes the way Codex claims its replays do is, for now, an open question you can only answer by testing.
04 — The CRM LensWhat to record, what to keep human.
Every outlet covered this as a general productivity story. The more useful frame for operations teams is narrower: which CRM and back-office tasks actually demo well? Early coverage converges on a pattern — strong fits are repeatable, rule-based, and verifiable (weekly reporting, spreadsheet normalization, support-triage classification, publishing checklists); weak fits are one-off tasks, workflows driven by un-narrated intuition, unverifiable outputs, and anything touching money transfers, account recovery, contracts, passwords, or customer PII.
Mapped onto the workflows we build in CRM automation engagements, that pattern turns into a decision table. The dividing line isn’t task complexity — it’s whether the rules can be spoken aloud completely during a single demonstration.
| CRM / back-office task | Record it? | Why | Governance note |
|---|---|---|---|
| Strong candidates | |||
| Lead-routing triage | Yes | Rule-based classification — status fields, branch assignment, source tagging — narrates cleanly in one pass | Demo on sample records with fake names and account numbers, never live customer data |
| Scheduled report pulls | Yes | Same steps, same screens, every week — the archetypal repeatable, verifiable workflow | Strip internal URLs, API keys, and tokens from anything visible on screen before recording |
| Data entry & spreadsheet normalization | Yes | Deterministic transforms with checkable outputs — errors surface immediately on review | Use a sample dataset; inspect the generated Skill for accidentally copied identifiers before sharing |
| Conditional — record with a human gate | |||
| Support first-response drafting | Partial | Template selection and classification demo well; tone and edge-case judgment don’t | Keep a human review step before anything customer-facing sends |
| Do not record | |||
| Exception handling & escalations | No | Driven by un-narrated intuition — the judgment that makes the call never makes it into the recording | Keep humans on judgment calls; a skill that mimics an escalation without the reasoning is a liability |
| Pricing & contract decisions | No | Radically input-dependent, and contracts sit explicitly on the flagged do-not-record list | Automate the paperwork around the decision, never the decision itself |
| Credentials, payments, customer PII | Never | Money transfers, account recovery, passwords, and medical records are all explicitly flagged as unsuitable | Close password managers, email, and finance apps; disable notification previews before every recording |
The narration test is the practical heuristic: if a competent new hire could do the task correctly after watching you do it once while you talked through the rules, it’s a candidate. If they’d need three months of context to know when the rules bend, it isn’t — and no recording feature from any vendor changes that.
05 — GovernanceThe documentation is thin — govern accordingly.
Here’s the honest gap in this launch: as of the announcement window, Anthropic had not published a dedicated technical or help page for Record a Skill — nothing covering recording retention, editing controls, or full rollout detail. Coursiv flagged the absence explicitly, and it means every governance decision right now has to lean on Anthropic’s pre-existing Cowork guidance rather than feature-specific policy.
That pre-existing guidance is worth reading before your first recording. Anthropic’s Use Claude Cowork safely page notes that Cowork can read files, browse the web, run code, and use connected apps, and that remote-session work is processed on Anthropic’s servers. Its computer-use guidance separately warns that information visible on screen during agentic use can include personal or confidential data. A feature whose entire input is your screen and your voice inherits both warnings at full strength.
AI Weekly’s launch analysis lists the questions the reporting doesn’t answer: what happens when the underlying app’s UI changes and a replay breaks, how the recording handles credentials or sensitive data scrolling past on screen, and whether skills can be shared or edited across a Team plan’s users. Until Anthropic documents those answers, the conservative reading is the right one — the same posture we argued for around enterprise agent-deployment governance: capability ships first, controls catch up, and the gap is where incidents happen.
06 — Strategic ReadThe desktop is the new battleground.
Why did both labs ship the same feature within a single product cycle? Because demonstration solves the cold-start problem that has throttled office-work automation since the first prompt was typed: most operational knowledge is tacit. The people who run your back office can’t write a complete spec of what they do — but they can show it, and they narrate the exceptions naturally while doing it. Recording converts tacit knowledge into an executable artifact at the cost of a single walkthrough, which is the same show-don’t-tell logic behind Claude’s shift toward shareable, show-me interfaces elsewhere in the product line.
Our read: the skill library is the real product. Once a team has recorded twenty workflows into one vendor’s desktop app — on the same Cowork surface where Anthropic’s flagship Fable 5 model already runs — switching costs stop being about model quality and start being about re-demonstrating institutional knowledge. Expect both vendors to push team-level sharing, versioning, and admin controls next, because that’s what converts a personal convenience into an organizational asset the org can’t easily leave.
Temper the projection with the reliability record, though. As we noted in our Codex coverage, the 2026 AI Index put average AI-agent success on the OSWorld computer-use benchmark at 66% — demonstration narrows the gap between what you meant and what the agent does, but it doesn’t make replays reliable on its own. And the anxiety is already public: Search Engine Journal’s coverage quoted one user reaction — “This has to be the easiest way to be replaced and lose your job man.” — which is sentiment, not data, but it’s the sentiment ops leaders will be managing in every rollout conversation. The teams that win this transition will frame recorded skills as removing the work people complain about, and keep the judgment calls visibly human.
07 — PlaybookA four-lane adoption plan.
If you run a CRM-centric operation, this feature is worth a pilot this quarter — scoped tightly. Sort your candidate workflows into four lanes and move one lane at a time.
Repeatable, rule-based back office
Report pulls, spreadsheet normalization, triage classification. Record on sample data, test on a fresh example per Anthropic’s own guidance, then run supervised for two weeks before trusting outputs.
Templated, customer-facing drafting
Support first responses, publishing checklists, brand-template cleanup. The mechanics record well; keep a named human approving anything that leaves the building.
Judgment, exceptions, escalations
Pricing calls, contract terms, escalation decisions. The reasoning that matters is exactly what a screen recording can’t capture. Revisit only if your rules become fully explicit.
Credentials, payments, PII
Money movement, account recovery, medical records, passwords. Explicitly flagged as unsuitable in every serious treatment of this launch — and undocumented retention makes the risk unquantifiable today.
Two disciplines make the pilot real. First, review every generated skill like code: read it, test it on a fresh example, and check it for copied identifiers before it enters a shared library. Second, assign an owner — recorded skills are process documentation that executes, and unowned executable documentation rots into risk. If you want help scoping which workflows clear the bar and building the governance around them, that’s the assessment we run at the start of AI transformation engagements.
08 — ConclusionDemonstration just became table stakes.
Show, don’t prompt — but govern what the camera sees.
Record a Skill is a genuinely lower floor for agent training: record a narrated walkthrough, get a rerunnable skill, no prompt writing. With Codex Record & Replay shipping the same bet 33 days earlier, demonstration-based skill capture went from novel to industry default inside a single product cycle — and Claude’s version is the one you can deploy on Windows and in Europe today.
The winners will be disciplined, not fast. The tasks that record well — triage, report pulls, normalization — are real capacity back for CRM and operations teams. The tasks that don’t record well haven’t changed, and Anthropic’s documentation on retention and admin controls hasn’t shipped yet. Pilot in the strong lanes, keep judgment human, and treat every recording as sensitive until the vendor tells you otherwise.