Agent skill

Webex Voice Interface

by automateyournetwork in automateyournetwork/netclaw

Respond to WebEx voice clips with both text and an MP3 voice reply using edge-tts.

Apache-2.0Auto-check passedMedia & Creative

Install Webex Voice Interface

skills CLI
$ npx skills add automateyournetwork/netclaw --skill webex-voice-interface -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install automateyournetwork/netclaw webex-voice-interface --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/automateyournetwork/netclaw.git skills-src && mkdir -p .claude/skills && cp -r skills-src/workspace/skills/webex-voice-interface .claude/skills/webex-voice-interface && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
webex-voice-interface
GitHub stars
676
Token cost
~934 tokens
SKILL.md length
298 words
Files
1
Skills in repo
120
Repo updated
First seen
Licence
Apache-2.0

At a glance

Respond to WebEx voice clips with both text and an MP3 voice reply using edge-tts.

  • Works in 3 steps: Process the question → Generate voice response → Deliver both text and voice
  • A user sends a voice message in WebEx
  • SKILL.md covers How It Works, Voice Response Workflow, Voice Selection and Performance, plus 4 more sections
  • Calls python3; reaches webexapis.com

What it does

Webex Voice Interface is an agent skill from automateyournetwork/netclaw. Respond to WebEx voice clips with both text and an MP3 voice reply using edge-tts. Voice IN is already handled by OpenClaw transcription. Use when a user sends a voice message in WebEx, you need to reply with audio, or you want to generate a spoken MP3 response.

Its SKILL.md is about 930 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice and Transcription. The repository describes itself as: An AI agent that claws through your network. The licence is Apache-2.0.

When your agent uses it

  • A user sends a voice message in WebEx
  • You need to reply with audio
  • You want to generate a spoken MP3 response

Example prompts

  • “/webex-voice-interface”

Requirements

  • Python 3

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Process the question
  2. Generate voice response
  3. Deliver both text and voice

What it can do on your machine

Read from SKILL.md and the folder at commit 95bb17e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • webexapis.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Webex Voice Interface loads about 934 tokens when it runs. Until then it costs about 71 tokens; SKILL.md has 298 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~71
When it runs · the whole SKILL.md, loaded when a task matches
~934

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from automateyournetwork/netclaw at commit 95bb17e, republished under its Apache-2.0 licence (© automateyournetwork). 298 words, ~934 tokens.

Download SKILL.mdSave it as .claude/skills/webex-voice-interface/SKILL.md (or your agent's skills folder).
name
webex-voice-interface
description
Respond to WebEx voice clips with both text and an MP3 voice reply using edge-tts. Voice IN is already handled by OpenClaw transcription. Use when a user sends a voice message in WebEx, you need to reply with audio, or you want to generate a spoken MP3 response.
license
Apache-2.0
user-invocable
true

WebEx Voice Interface

How It Works

User sends voice clip in WebEx space
    |
    v
OpenClaw transcribes automatically (built-in)
    |
    v
NetClaw processes with full skill set
(pyATS, NetBox, ServiceNow, all 43 MCP servers)
    |
    v
python3 $MCP_CALL "python3 -u $TTS_MCP_SCRIPT" text_to_speech -> MP3 file
    |
    v
Upload MP3 to WebEx thread + post text response

Voice Response Workflow

Step 1: Process the question

Treat the transcribed voice message identically to a typed text message. Use the full NetClaw skill set -- pyATS, NetBox, ServiceNow, etc.

Step 2: Generate voice response

After composing your text response, call text_to_speech:

bash
python3 $MCP_CALL "python3 -u $TTS_MCP_SCRIPT" text_to_speech '{"text":"R1 has 3 OSPF neighbors, all in FULL state on Area 0...","voice":"en-US-GuyNeural"}'

This returns JSON with an output_path to the generated MP3 file.

To list available voices:

bash
python3 $MCP_CALL "python3 -u $TTS_MCP_SCRIPT" list_voices '{"language":"en"}'
Step 3: Deliver both text and voice

Post the text response in the WebEx space/thread AND upload the MP3 file as an attachment via the Messages API (multipart/form-data with files parameter):

Voice Response [MP3 audio file attached]

R1 has 3 OSPF neighbors, all in FULL state on Area 0:

  • 2.2.2.2 (R2) via Gi1 -- FULL/DR
  • 3.3.3.3 (R3) via Gi2 -- FULL/BDR

Always deliver text AND voice. Text is primary (searchable, accessible). Voice is supplementary.

Voice Selection

VoiceDescription
en-US-GuyNeuralProfessional male -- default
en-US-JennyNeuralProfessional female
en-US-AriaNeuralConversational female
en-GB-RyanNeuralBritish male

Users can request a voice change:

  • "Switch to a female voice" -> use en-US-JennyNeural
  • "Use a British accent" -> use en-GB-RyanNeural

Call list_voices to see all 300+ available voices.

Performance

PhaseLatency
edge-tts synthesis1-2 seconds
WebEx MP3 upload< 1 second

Voice synthesis adds minimal overhead to the response time.

Fallback

If TTS fails, deliver the text response immediately. Do not block on voice.

Tips for Voice Responses

  • Keep it concise -- under 100 words works best for spoken delivery
  • Avoid tables -- describe data conversationally for voice
  • Spell out abbreviations -- say "OSPF" not "O-S-P-F" (edge-tts handles this)
  • Use natural phrasing -- the text will be read aloud, so write for the ear

WebEx File Upload for Voice

WebEx Messages API supports file attachments via multipart upload:

POST https://webexapis.com/v1/messages
Content-Type: multipart/form-data

- roomId: <space-id>
- parentId: <thread-parent-message-id> (if threading)
- text: "Voice Response: R1 has 3 OSPF neighbors..."
- files: @/tmp/netclaw-tts/response.mp3

Files up to 100 MB are supported. MP3 voice responses are typically under 1 MB.

GAIT Integration

Record voice interactions in the GAIT audit trail:

Input: Voice clip from @user (transcript: "What are your interfaces?")
Action: Queried R1 interfaces via pyATS
Output: 4 interfaces found -- text + voice response delivered to WebEx

© automateyournetwork, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in workspace/skills/webex-voice-interface of automateyournetwork/netclaw.

Open the folder on GitHubat commit 95bb17e

Compare with similar skills

Webex Voice Interface next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Webex Voice Interface compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Webex Voice Interface this skillautomateyournetwork/netclaw676—~934Automated safety check: PassApache-2.0
HyperFrames Media Useheygen-com/hyperframes59k—~2.4kAutomated safety check: PassApache-2.0
Edu Math Videowy51ai/edulab1.4k—~2.5kAutomated safety check: NotesApache-2.0
Elevenlabs Transcribeqdhenry/Claude-Command-Suite1.3k—~1.5kAutomated safety check: NotesNone
Video Assemblezenstory-ai/video-recap-skills559—~1.7kAutomated safety check: PassMIT
Speech Usecnemri/google-genai-skills127—~616Automated safety check: PassMIT

Similar skills

  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    59k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Edu Math Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a math problem (数学题, geometry, algebra, functions, motion/行程 problems), from a problem screenshot…

    1.4k GitHub stars~2.5k tokensUpdated 10 days ago
    Media & CreativeAuto-check: notes
  • Elevenlabs Transcribe

    qdhenry/Claude-Command-Suite

    Transcribes audio/video files using ElevenLabs Scribe v2 API.

    1.3k GitHub stars~1.5k tokensUpdated 7 mo ago
    Media & CreativeAuto-check: notes
  • Video Assemble

    zenstory-ai/video-recap-skills

    合成视频解说最终成片:把旁白音频铺到源视频上,按旁白窗口压低原声,生成 SRT / ASS 字幕并可烧录, 最后做响度标准化。作为最终合成阶段使用。输入源视频、ttsmeta.json 与旁白位置; 输出 recap 成片和字幕。触发词:视频合成、混音、字幕、压字幕、assemble video、mux、ducking、subtitles、成片。

    559 GitHub stars~1.7k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Speech Use

    cnemri/google-genai-skills

    Generate (TTS), Transcribe (STT), and Clone voices using Google's GenAI and Cloud Speech SDKs.

    127 GitHub stars~616 tokensUpdated 8 mo ago
    Media & CreativeAuto-check passed
  • Gemini Audio

    einverne/dotfiles

    Guide for implementing Google Gemini API audio capabilities - analyze audio with transcription, summarization, and understanding (up to 9.5 hours), plus generate speech with controllable TTS.

    121 GitHub stars~2k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes

More from automateyournetwork/netclaw

All 120 skills in this repo
  • EVE-NG Lab Topology Design

    automateyournetwork/netclaw

    Entry point for designing EVE-NG network labs: classifies the request, gathers missing requirements, proposes options and validates the resulting topology.

    676 GitHub stars~612 tokensUpdated 3 days ago
    Auto-check passed
  • ACI Policy Change Deployment

    automateyournetwork/netclaw

    Deploys Cisco ACI policy changes only behind an approved ServiceNow Change Request, capturing pre and post-change fault baselines and rolling back automatically on a fault delta.

    676 GitHub stars~4.2k tokensUpdated 3 days ago
    Auto-check passed
  • Cisco ACI Fabric Health Audit

    automateyournetwork/netclaw

    Runs a phased health audit of a Cisco ACI fabric through MCP tools: node status, links, tenant and policy review, faults and endpoint learning.

    676 GitHub stars~2.9k tokensUpdated 3 days ago
    Auto-check passed
  • Anta Validation

    automateyournetwork/netclaw

    Validate Arista EOS network state against ANTA's pre-built 208-test catalogue, with structured pass/fail verdicts.

    676 GitHub stars~1.2k tokensUpdated 3 days ago
    Auto-check passed
  • Arista Cvp

    automateyournetwork/netclaw

    Arista CloudVision Portal (CVP) automation via REST API — device inventory, events, connectivity monitoring, tag management (4 tools).

    676 GitHub stars~2.2k tokensUpdated 3 days ago
    Auto-check: notes
  • AWS Cloud Monitoring

    automateyournetwork/netclaw

    AWS CloudWatch monitoring — metrics, alarms, log queries, VPC flow log analysis, network performance.

    676 GitHub stars~1k tokensUpdated 3 days ago
    Auto-check passed

Questions about Webex Voice Interface

What does Webex Voice Interface do?

Respond to WebEx voice clips with both text and an MP3 voice reply using edge-tts. Webex Voice Interface is an agent skill from automateyournetwork/netclaw. Respond to WebEx voice clips with both text and an MP3 voice reply using edge-tts.

When should I use Webex Voice Interface?

Webex Voice Interface fits situations like: A user sends a voice message in WebEx; you need to reply with audio; you want to generate a spoken MP3 response.

How do I install Webex Voice Interface in Claude Code?

Run `npx skills add automateyournetwork/netclaw --skill webex-voice-interface -a claude-code`. Or copy the skill folder (workspace/skills/webex-voice-interface in automateyournetwork/netclaw) into .claude/skills/webex-voice-interface in your project. Claude Code loads it when a task matches its description.

How do I install Webex Voice Interface in Codex?

Run `npx skills add automateyournetwork/netclaw --skill webex-voice-interface -a codex`. Or copy the skill folder (workspace/skills/webex-voice-interface in automateyournetwork/netclaw) into .agents/skills/webex-voice-interface in your project. Codex loads it when a task matches its description.

Can I use Webex Voice Interface in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add automateyournetwork/netclaw --skill webex-voice-interface -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/webex-voice-interface, .gemini/skills/webex-voice-interface, .github/skills/webex-voice-interface and .opencode/skills/webex-voice-interface in your project.

What does Webex Voice Interface need to run?

Going by SKILL.md and its folder, Webex Voice Interface needs the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Webex Voice Interface access the network?

SKILL.md names 1 domain. In commands or code: webexapis.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Webex Voice Interface safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Webex Voice Interface use?

Webex Voice Interface is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Webex Voice Interface use?

About 934 tokens (SKILL.md is roughly 3.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Webex Voice Interface?

Skills that share tags, products or a category with Webex Voice Interface: HyperFrames Media Use (heygen-com/hyperframes, 59k stars), Edu Math Video (wy51ai/edulab, 1.4k stars), Elevenlabs Transcribe (qdhenry/Claude-Command-Suite, 1.3k stars) and Video Assemble (zenstory-ai/video-recap-skills, 559 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Webex Voice Interface?

automateyournetwork (a GitHub user) maintains it in automateyournetwork/netclaw, which has 676 GitHub stars. The repository holds 120 skills in this directory. The repository was last updated on October 5, 2026.

Source: automateyournetwork/netclaw on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.