Agent skill

Voice Agent Design

by mohitagw15856 in mohitagw15856/pm-claude-skills

Design a voice AI agent for phone or in-app conversations — call flows, interruption handling, escalation to humans, and the metrics that catch a bad voice experience.

MITAuto-check passedAI & LLM Engineering

Install Voice Agent Design

skills CLI
$ npx skills add mohitagw15856/pm-claude-skills --skill voice-agent-design -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install mohitagw15856/pm-claude-skills voice-agent-design --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/mohitagw15856/pm-claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/voice-agent-design .claude/skills/voice-agent-design && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
voice-agent-design
GitHub stars
1.4k
Token cost
~1.6k tokens
SKILL.md length
790 words
Files
1
Skills in repo
1,348
Repo updated
First seen
Licence
MIT

At a glance

Design a voice AI agent for phone or in-app conversations — call flows, interruption handling, escalation to humans, and the metrics that catch a bad voice experience.

  • Works in 6 steps: Scope by intent, ruthlessly. From the… → Disclose and set the frame in the first… → Design turns for ears, not eyes. One… → …
  • Asked to design a voice agent
  • SKILL.md covers What This Skill Produces, Required Inputs, Design Method and Output Format, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Voice Agent Design is an agent skill from mohitagw15856/pm-claude-skills. Design a voice AI agent for phone or in-app conversations — call flows, interruption handling, escalation to humans, and the metrics that catch a bad voice experience. Use when asked to design a voice agent, automate a phone line, spec an IVR replacement, or review why callers hate an existing voice bot. Produces a voice agent spec: persona and disclosure policy, conversation architecture, barge-in and repair behaviour, human-handoff rules, and a launch scorecard.

Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering, covering Speech recognition and synthesis. The repository describes itself as: 1255 professional Agent Skills for Claude, ChatGPT, Gemini, Cursor & Codex — PRDs, postmortems, leases, medical bills, layoffs, go-bags, new countries. Plain markdown, MIT, in… The licence is MIT.

When your agent uses it

  • Asked to design a voice agent
  • Automate a phone line
  • Spec an IVR replacement
  • Review why callers hate an existing voice bot

Example prompts

  • “/voice-agent-design”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Scope by intent, ruthlessly. From the intent list, the agent owns only intents that are (a) high-volume, (b) completable with its actual…
  2. Disclose and set the frame in the first five seconds. The agent says it's an AI (increasingly required by law; always required by trust)…
  3. Design turns for ears, not eyes. One question per turn · ≤2 sentences before yielding · numbers and options in threes at most ("I can do…
  4. Engineer the mechanics that make it feel alive
  5. Make the handoff a feature. Triggers: caller asks (always, instantly) · two failed repairs on one slot · negative-emotion cues · any…
  6. Score what callers feel, not what dashboards flatter. Containment alone is gameable (trap callers and containment "improves"). The…

What it can do on your machine

Read from SKILL.md and the folder at commit 1cbf1f0. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Voice Agent Design loads about 1.6k tokens when it runs. Until then it costs about 122 tokens; SKILL.md has 790 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~122
When it runs · the whole SKILL.md, loaded when a task matches
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from mohitagw15856/pm-claude-skills at commit 1cbf1f0, republished under its MIT licence (© mohitagw15856). 790 words, ~1,552 tokens.

Download SKILL.mdSave it as .claude/skills/voice-agent-design/SKILL.md (or your agent's skills folder).
name
voice-agent-design
description
Design a voice AI agent for phone or in-app conversations — call flows, interruption handling, escalation to humans, and the metrics that catch a bad voice experience. Use when asked to design a voice agent, automate a phone line, spec an IVR replacement, or review why callers hate an existing voice bot. Produces a voice agent spec: persona and disclosure policy, conversation architecture, barge-in and repair behaviour, human-handoff rules, and a launch scorecard.

Voice Agent Design Skill

Voice is the least forgiving agent surface: no screen to fall back on, dead air reads as failure within two seconds, and the caller is often already annoyed. This skill designs voice agents around the medium's real constraints — turn-taking, interruption, repair — instead of shipping a chatbot with a text-to-speech voice.

What This Skill Produces

  • A scope decision: which call intents the agent owns end-to-end, which it triages, which go straight to humans
  • A conversation architecture: openings, turn design, confirmation strategy, repair loops
  • Barge-in, silence, and error behaviour — the mechanics that decide whether it feels alive or infuriating
  • Human-handoff rules with context transfer, and a launch scorecard

Required Inputs

Ask for (if not already provided):

  • The line and its traffic: what people call about (top intents with rough volumes), current handle times
  • What the agent may actually do — which systems it can read/write, what it can promise
  • The escalation reality: human hours, queue lengths, what happens after-hours
  • Compliance context: recording consent, disclosure requirements, regulated statements in this domain

Design Method

  1. Scope by intent, ruthlessly. From the intent list, the agent owns only intents that are (a) high-volume, (b) completable with its actual system access, and (c) low-stakes-if-wrong. It triages everything it can identify but not complete. It immediately passes anything emotional, legal, or high-value — a furious caller is a human's job on the first turn, not after three failed bot turns.
  2. Disclose and set the frame in the first five seconds. The agent says it's an AI (increasingly required by law; always required by trust), what it can do, and how to reach a human ("say 'agent' anytime"). Hiding the escape hatch inflates containment metrics and rage in equal measure.
  3. Design turns for ears, not eyes. One question per turn · ≤2 sentences before yielding · numbers and options in threes at most ("I can do A, B, or C — which one?") · never read a paragraph. Anything long ("your options are…") gets offered as SMS/email instead of spoken.
  4. Engineer the mechanics that make it feel alive:
    • Barge-in on: the caller can interrupt any utterance; the agent stops mid-sentence and processes.
    • Latency masked: acknowledge within ~1s ("let me check that…") whenever a lookup exceeds it; dead air past 2s is where trust dies.
    • Confirmation proportional to stakes: implicit for low stakes ("okay, Tuesday…"), explicit read-back for money, addresses, and anything irreversible.
    • Repair, not repeat: on a misunderstanding, change strategy — rephrase, offer options, or fall to keypad — never re-ask the same question the same way twice.
  5. Make the handoff a feature. Triggers: caller asks (always, instantly) · two failed repairs on one slot · negative-emotion cues · any regulated topic. The transfer carries a whisper summary (who, what they want, what's been tried, account pulled up) — the caller never repeats themselves; that single property beats every other quality bar in perceived experience.
  6. Score what callers feel, not what dashboards flatter. Containment alone is gameable (trap callers and containment "improves"). The scorecard pairs it with: task success as the caller defines it (post-call yes/no), escapes-requested rate, repair rate, silent-transfer rate, and hang-ups mid-flow. Set launch gates on the pairs.
Show full SKILL.md (270 more words)Show less

Output Format

Voice Agent Spec: [line/product]

Intent scope

IntentVolumeOwn / Triage / PassWhy

Opening script: [verbatim — disclosure, capability, escape hatch]

Conversation architecture: [turn rules · confirmation strategy by stakes · the repair ladder (rephrase → options → keypad → human)]

Mechanics: [barge-in behaviour · latency masking thresholds · silence handling]

Handoff: [triggers · whisper-summary fields · after-hours behaviour]

Compliance: [disclosure line · recording consent flow · statements the agent must never make]

Launch scorecard

MetricGateWhy paired
Containment + caller-scored successcontainment alone is gameable
Escape-request ratemeasures trapped callers
Repair rate / hang-ups mid-flowfrustration signals

Quality Checks

  • Every owned intent is completable with the agent's actual system access — no "owns refunds" without refund API access
  • The opening discloses AI status and the escape hatch, verbatim in the spec
  • No designed utterance exceeds two sentences before yielding
  • The repair ladder changes strategy at each rung — no repeat-louder step
  • Handoff carries the whisper summary; "please hold while I transfer you" to a cold human fails the spec
  • The scorecard pairs containment with caller-scored success

Anti-Patterns

  • Do not port the chatbot script to voice — text tolerates paragraphs and menus; ears don't
  • Do not hide the human escape hatch to protect containment metrics — callers find the exit anyway, angrier
  • Do not let the agent bluff on regulated topics (medical, legal, financial advice) — pass or read the approved statement
  • Do not re-ask a failed question unchanged — the caller heard you; the strategy failed, not their ears
  • Do not launch without the mid-flow hang-up metric — it's where voice agents quietly hemorrhage trust

Example Trigger Phrases

  • "Design a voice agent."
  • "Automate a phone line."
  • "Spec an IVR replacement."
  • "Review why callers hate an existing voice bot."

© mohitagw15856, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/voice-agent-design of mohitagw15856/pm-claude-skills.

Open the folder on GitHubat commit 1cbf1f0

Compare with similar skills

Voice Agent Design next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Voice Agent Design compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Voice Agent Design this skillmohitagw15856/pm-claude-skills1.4k—~1.6kAutomated safety check: PassMIT
TriageTalAter/annyang6.8k—~810Automated safety check: NotesMIT
Yichen Asrmcncarl/yichen-skills4.4k—~780Automated safety check: PassCustom licence
Dingtalk MinutesDingTalk-Real-AI/dingtalk-workspace-cli3.2k—~2.3kAutomated safety check: PassApache-2.0
Youtube FetcherJimmySadek/youtube-fetcher-to-markdown485—~3.1kAutomated safety check: PassMIT
Yichen Web Researchmcncarl/yichen-skills4.4k—~1.9kAutomated safety check: PassCustom licence

Similar skills

  • Triage

    TalAter/annyang

    Triage and close GitHub issues on TalAter/annyang. An agent skill from TalAter/annyang.

    6.8k GitHub stars~810 tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check: notes
  • Yichen Asr

    mcncarl/yichen-skills

    逸尘自用的统一音视频转写入口,在 StepFun Step ASR 与火山引擎豆包 ASR 之间按输出需求、安全边界和可用状态路由。用于本地音频或视频的纯文本转写、时间戳、SRT 字幕、口播粗剪,以及转写前体检;用户明确指定服务商时不得静默切换。Use when a local audio or video file needs transcription and the correct…

    4.4k GitHub stars~780 tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check passed
  • Dingtalk Minutes

    DingTalk-Real-AI/dingtalk-workspace-cli

    钉钉 AI 听记。Use when 查询或修改听记摘要、完整逐字稿、关键词、标签、行动项、录音、上传、思维导图、发言人洞察、ASR 热词/识别词配置或分享权限。写文档走 dingtalk-doc;建待办走 dingtalk-todo;日程走 dingtalk-calendar。命令前缀:dws minutes。

    3.2k GitHub stars~2.3k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Youtube Fetcher

    JimmySadek/youtube-fetcher-to-markdown

    Retrieve transcripts from YouTube, Instagram, TikTok, X, Vimeo and other video sites, summarize or analyze what was said (and shown on screen), or save an Obsidian-ready Markdown knowledge-base note…

    485 GitHub stars~3.1k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Yichen Web Research

    mcncarl/yichen-skills

    逸尘自用的互联网研究总入口。用于跨平台且跨阶段、用户尚未确定工具,或明确要求对公司、产品、人物、技术、行业和领域做横纵分析、发展史加现状对比或有来源约束的系统深度研究;先生成有截止日期和证据闸门的计划,再把搜索发现、候选核验、有限归档、按需转写和证据综合路由到…

    4.4k GitHub stars~1.9k tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check passed
  • Volcengine Asr

    ysyecust/lecture-to-notes

    Transcribe local audio or video with Volcengine Doubao file ASR, including BigASR 1.0 Turbo direct upload and asynchronous 1.0 standard, 1.0 idle, or 2.0 standard jobs through TOS.

    273 GitHub stars~783 tokensUpdated 8 days ago
    AI & LLM EngineeringAuto-check passed

More from mohitagw15856/pm-claude-skills

All 1,348 skills in this repo
  • Car Tco

    mohitagw15856/pm-claude-skills

    Compare the total cost of car ownership across buy-new, buy-used, lease, and keep-your-current-car — depreciation, insurance, maintenance ramp, and fuel over a real horizon, not just the monthly…

    1.4k GitHub stars~1.1k tokensUpdated 2 days ago
    Auto-check passed
  • Cs Health Scorecard

    mohitagw15856/pm-claude-skills

    Build a customer health scorecard for a specific account. An agent skill from mohitagw15856/pm-claude-skills.

    1.4k GitHub stars~2.4k tokensUpdated 2 days ago
    Auto-check passed
  • Exit Waterfall

    mohitagw15856/pm-claude-skills

    Compute who gets what at each exit price from a cap table — liquidation preferences, conversion points, and where the founders' share collapses.

    1.4k GitHub stars~1.1k tokensUpdated 2 days ago
    Auto-check passed
  • Feature Prioritisation

    mohitagw15856/pm-claude-skills

    Apply prioritisation frameworks (RICE, MoSCoW, Kano, ICE, Opportunity Scoring) to rank features and backlog items.

    1.4k GitHub stars~2k tokensUpdated 2 days ago
    Auto-check passed
  • Fire Number

    mohitagw15856/pm-claude-skills

    Compute a financial-independence (FIRE) target and years-to-reach with every assumption labeled as an assumption — plus a sensitivity table instead of a single false-precision answer.

    1.4k GitHub stars~1.1k tokensUpdated 2 days ago
    Auto-check passed
  • Freelance Rate

    mohitagw15856/pm-claude-skills

    Derive a freelance day/hourly rate backwards from target income, honest billable utilization, overhead, and the self-employment tax premium — the arithmetic that proves a rate is not salary÷2000.

    1.4k GitHub stars~1.2k tokensUpdated 2 days ago
    Auto-check passed

Questions about Voice Agent Design

What does Voice Agent Design do?

Design a voice AI agent for phone or in-app conversations — call flows, interruption handling, escalation to humans, and the metrics that catch a bad voice experience. Voice Agent Design is an agent skill from mohitagw15856/pm-claude-skills. Design a voice AI agent for phone or in-app conversations — call flows, interruption handling, escalation to humans, and the metrics that catch a bad voice experience.

When should I use Voice Agent Design?

Voice Agent Design fits situations like: asked to design a voice agent; automate a phone line; spec an IVR replacement; review why callers hate an existing voice bot.

How do I install Voice Agent Design in Claude Code?

Run `npx skills add mohitagw15856/pm-claude-skills --skill voice-agent-design -a claude-code`. Or copy the skill folder (skills/voice-agent-design in mohitagw15856/pm-claude-skills) into .claude/skills/voice-agent-design in your project. Claude Code loads it when a task matches its description.

How do I install Voice Agent Design in Codex?

Run `npx skills add mohitagw15856/pm-claude-skills --skill voice-agent-design -a codex`. Or copy the skill folder (skills/voice-agent-design in mohitagw15856/pm-claude-skills) into .agents/skills/voice-agent-design in your project. Codex loads it when a task matches its description.

Can I use Voice Agent Design in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mohitagw15856/pm-claude-skills --skill voice-agent-design -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-agent-design, .gemini/skills/voice-agent-design, .github/skills/voice-agent-design and .opencode/skills/voice-agent-design in your project.

What does Voice Agent Design need to run?

SKILL.md names no scripts, command-line tools or credentials: Voice Agent Design is instructions for the agent only.

Does Voice Agent Design access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Voice Agent Design safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Voice Agent Design use?

Voice Agent Design is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Voice Agent Design use?

About 1.6k tokens (SKILL.md is roughly 6.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Voice Agent Design?

Skills that share tags, products or a category with Voice Agent Design: Triage (TalAter/annyang, 6.8k stars), Yichen Asr (mcncarl/yichen-skills, 4.4k stars), Dingtalk Minutes (DingTalk-Real-AI/dingtalk-workspace-cli, 3.2k stars) and Youtube Fetcher (JimmySadek/youtube-fetcher-to-markdown, 485 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Voice Agent Design?

mohitagw15856 (a GitHub user) maintains it in mohitagw15856/pm-claude-skills, which has 1,434 GitHub stars. The repository holds 1,348 skills in this directory. The repository was last updated on October 9, 2026.

Source: mohitagw15856/pm-claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.