Agent skill

Ops Voice

by davepoon in davepoon/buildwithclaude

Voice operations — make phone calls (Bland AI), text-to-speech (ElevenLabs), transcribe audio (Whisper/Groq).

MITAuto-check: notesMedia & Creative

Install Ops Voice

skills CLI
$ npx skills add davepoon/buildwithclaude --skill ops-voice -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install davepoon/buildwithclaude ops-voice --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/claude-ops/skills/ops-voice .claude/skills/ops-voice && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
ops-voice
GitHub stars
3.6k
Token cost
~1.5k tokens
SKILL.md length
231 words
Files
1
Skills in repo
245
Repo updated
First seen
Licence
MIT

At a glance

Voice operations — make phone calls (Bland AI), text-to-speech (ElevenLabs), transcribe audio (Whisper/Groq).

  • Works in 3 steps: Bland AI: curl -s -H "authorization:… → ElevenLabs: curl -s -H "xi-api-key:… → Groq: curl -s -H "Authorization: Bearer…
  • Tasks that involve Text to speech and voice
  • SKILL.md covers Sub-commands and Execution
  • Calls curl, jq and python3; reaches api.bland.ai and api.elevenlabs.io; needs BLAND_AI_API_KEY and ELEVENLABS_API_KEY

What it does

Ops Voice is an agent skill from davepoon/buildwithclaude. Voice operations — make phone calls (Bland AI), text-to-speech (ElevenLabs), transcribe audio (Whisper/Groq). Replace OpenClaw voice capabilities.

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice and Transcription. It works with ElevenLabs. The repository describes itself as: A single hub to find Claude Skills, Agents, Commands, Hooks, Plugins, and Marketplace collections to extend Claude Code, Claude Desktop, Agent SDK and OpenClaw. The licence is MIT.

When your agent uses it

  • Tasks that involve Text to speech and voice
  • Tasks that involve Transcription

Example prompts

  • “/ops-voice”

Requirements

  • Python 3
  • A credential in BLAND_AI_API_KEY
  • A credential in BLAND_KEY
  • Pre-approved tools (allowed-tools): Bash, Read, Write, AskUserQuestion, WebFetch

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Bland AI: curl -s -H "authorization: $KEY" https://api.bland.ai/v1/me — check balance
  2. ElevenLabs: curl -s -H "xi-api-key: $KEY" https://api.elevenlabs.io/v1/voices?page_size=1 — list voices
  3. Groq: curl -s -H "Authorization: Bearer $KEY" https://api.groq.com/openai/v1/models — list models

What it can do on your machine

Read from SKILL.md and the folder at commit 10bfc43. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash
    • Read
    • Write
    • AskUserQuestion
    • WebFetch

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl
    • jq
    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.bland.ai
    • api.elevenlabs.io
    • api.groq.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • BLAND_AI_API_KEY
    • ELEVENLABS_API_KEY
    • GROQ_API_KEY
    • BLAND_KEY
    • EL_KEY
    • GROQ_KEY
    • BLAND_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Ops Voice loads about 1.5k tokens when it runs. Until then it costs about 39 tokens; SKILL.md has 231 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~39
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash, Read, Write, AskUserQuestion, WebFetch

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from davepoon/buildwithclaude at commit 10bfc43, republished under its MIT licence (© davepoon). 231 words, ~1,493 tokens.

Download SKILL.mdSave it as .claude/skills/ops-voice/SKILL.md (or your agent's skills folder).
name
ops-voice
description
Voice operations — make phone calls (Bland AI), text-to-speech (ElevenLabs), transcribe audio (Whisper/Groq). Replace OpenClaw voice capabilities.
allowed-tools
Bash, Read, Write, AskUserQuestion, WebFetch
argument-hint
[call|tts|transcribe|setup]

OPS:VOICE — Voice Operations

Voice interface commands. All API calls via curl — no SDK dependencies.

Credential resolution order: userConfig → env vars → Doppler MCP tools (mcp__doppler__*) → Doppler CLI fallback (doppler secrets get <KEY> --plain) → password manager


Sub-commands

Parse $ARGUMENTS for the command keyword, then execute:


call [phone] [prompt] — Bland AI phone call

Requires: bland_ai_api_key in userConfig or BLAND_AI_API_KEY env or Doppler.

bash
BLAND_KEY="${BLAND_AI_API_KEY:-$(doppler secrets get BLAND_AI_API_KEY --plain 2>/dev/null || true)}"
PHONE="<extracted from $ARGUMENTS>"
PROMPT="<extracted from $ARGUMENTS or ask user>"
MAX_DURATION="${BLAND_MAX_DURATION:-300}"  # seconds
VOICE="${BLAND_VOICE:-male}"

# Make the call
RESPONSE=$(curl -s -X POST "https://api.bland.ai/v1/calls" \
  -H "authorization: $BLAND_KEY" \
  -H "Content-Type: application/json" \
  -d "{
    \"phone_number\": \"$PHONE\",
    \"task\": \"$PROMPT\",
    \"voice\": \"$VOICE\",
    \"max_duration\": $MAX_DURATION,
    \"record\": true
  }")

CALL_ID=$(echo "$RESPONSE" | python3 -c "import json,sys; print(json.load(sys.stdin).get('call_id',''))" 2>/dev/null)

# Poll for completion (up to 5 min)
if [ -n "$CALL_ID" ]; then
  echo "Call initiated: $CALL_ID"
  for i in $(seq 1 30); do
    sleep 10
    STATUS=$(curl -s "https://api.bland.ai/v1/calls/$CALL_ID" \
      -H "authorization: $BLAND_KEY" | \
      python3 -c "import json,sys; d=json.load(sys.stdin); print(d.get('status',''), d.get('transcripts','')[-1].get('text','') if d.get('transcripts') else '')" 2>/dev/null)
    echo "Status: $STATUS"
    [[ "$STATUS" == completed* ]] && break
  done
fi

Output: Call ID, live status, transcript when complete.


tts [text] [--voice voice_id] [--out file.mp3] — ElevenLabs text-to-speech

Requires: elevenlabs_api_key in userConfig or ELEVENLABS_API_KEY env or Doppler.

bash
EL_KEY="${ELEVENLABS_API_KEY:-$(doppler secrets get ELEVENLABS_API_KEY --plain 2>/dev/null || true)}"
VOICE_ID="${ELEVENLABS_VOICE_ID:-21m00Tcm4TlvDq8ikWAM}"  # Rachel (default)
TEXT="<extracted from $ARGUMENTS>"
OUT_FILE="${OUT_FILE:-/tmp/ops-tts-$(date +%s).mp3}"

# List voices if voice name provided (not an ID)
# Synthesize
curl -s -X POST "https://api.elevenlabs.io/v1/text-to-speech/${VOICE_ID}" \
  -H "xi-api-key: $EL_KEY" \
  -H "Content-Type: application/json" \
  -d "{
    \"text\": \"$TEXT\",
    \"model_id\": \"eleven_monolingual_v1\",
    \"voice_settings\": {\"stability\": 0.5, \"similarity_boost\": 0.75}
  }" \
  --output "$OUT_FILE"

echo "Audio saved to: $OUT_FILE"
# Auto-play on macOS
command -v afplay &>/dev/null && afplay "$OUT_FILE" &

Output: Audio file path. Auto-plays on macOS via afplay.


transcribe [file_path] — Groq Whisper transcription

Requires: groq_api_key in userConfig or GROQ_API_KEY env or Doppler.

bash
GROQ_KEY="${GROQ_API_KEY:-$(doppler secrets get GROQ_API_KEY --plain 2>/dev/null || true)}"
AUDIO_FILE="<extracted from $ARGUMENTS>"

if [ ! -f "$AUDIO_FILE" ]; then
  echo "ERROR: File not found: $AUDIO_FILE"
  exit 1
fi

TRANSCRIPT=$(curl -s -X POST "https://api.groq.com/openai/v1/audio/transcriptions" \
  -H "Authorization: Bearer $GROQ_KEY" \
  -F "file=@$AUDIO_FILE" \
  -F "model=whisper-large-v3" \
  -F "response_format=json" | \
  python3 -c "import json,sys; print(json.load(sys.stdin).get('text',''))" 2>/dev/null)

echo "$TRANSCRIPT"

Output: Transcript text printed to stdout.


setup — Configure voice API keys

Before asking for anything, auto-scan ALL sources in a single background batch:

bash
# Env vars
printenv BLAND_AI_API_KEY BLAND_API_KEY ELEVENLABS_API_KEY GROQ_API_KEY 2>/dev/null

# Shell profiles
grep -h 'BLAND\|ELEVENLABS\|GROQ' ~/.zshrc ~/.bashrc ~/.zprofile ~/.envrc 2>/dev/null | grep -v '^#'

# Doppler — ALL projects
for proj in $(doppler projects --json 2>/dev/null | jq -r '.[].slug'); do
  for cfg in dev stg prd; do
    doppler secrets --project "$proj" --config "$cfg" --json 2>/dev/null | \
      jq -r --arg proj "$proj" --arg cfg "$cfg" 'to_entries[] | select(.key | test("BLAND|ELEVENLABS|GROQ"; "i")) | "\(.key)=\(.value.computed | .[0:12])... (doppler:\($proj)/\($cfg))"'
  done
done

# Dashlane
dcli password bland --output json 2>/dev/null | jq -r '.[] | select(.password != null) | "\(.title): key found"'
dcli password elevenlabs --output json 2>/dev/null | jq -r '.[] | select(.password != null) | "\(.title): key found"'
dcli password groq --output json 2>/dev/null | jq -r '.[] | select(.password != null) | "\(.title): key found"'

# Keychain
security find-generic-password -s "bland-ai-api-key" -w 2>/dev/null
security find-generic-password -s "elevenlabs-api-key" -w 2>/dev/null
security find-generic-password -s "groq-api-key" -w 2>/dev/null

Present all findings. Only prompt for keys NOT found in any source. Then validate each found key in background:

  1. Bland AI: curl -s -H "authorization: $KEY" https://api.bland.ai/v1/me — check balance
  2. ElevenLabs: curl -s -H "xi-api-key: $KEY" https://api.elevenlabs.io/v1/voices?page_size=1 — list voices
  3. Groq: curl -s -H "Authorization: Bearer $KEY" https://api.groq.com/openai/v1/models — list models

Report: [service] ✓ connected or [service] ✗ invalid key — [error]


Execution

  1. Resolve the sub-command from $ARGUMENTS (first word: call / tts / transcribe / setup)
  2. Resolve credentials in order: env → Doppler
  3. Execute the matching curl block above
  4. If a required key is missing and setup was not invoked, suggest /ops:ops-voice setup

© davepoon, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in plugins/claude-ops/skills/ops-voice of davepoon/buildwithclaude.

Open the folder on GitHubat commit 10bfc43

Compare with similar skills

Ops Voice next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Ops Voice compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Ops Voice this skilldavepoon/buildwithclaude3.6k—~1.5kAutomated safety check: NotesMIT
Elevenlabs Transcribeqdhenry/Claude-Command-Suite1.3k—~1.5kAutomated safety check: NotesNone
Elevenlabszapier/connectors176—~3.7kAutomated safety check: PassElastic-2.0
Speech To Texttadaspetra/loop2963 repos~2kAutomated safety check: PassMIT
Video Translatorshang-zhu/violin1.1k—~1kAutomated safety check: NotesMIT
Hyperframes Mediachmonitor/chmonitor2991 repos~2.8kAutomated safety check: NotesGPL-3.0

Similar skills

  • Elevenlabs Transcribe

    qdhenry/Claude-Command-Suite

    Transcribes audio/video files using ElevenLabs Scribe v2 API.

    1.3k GitHub stars~1.5k tokensUpdated 7 mo ago
    Media & CreativeAuto-check: notes
  • Elevenlabs

    zapier/connectors

    Official

    Agent-callable ElevenLabs tools — generate spoken audio from text, create sound effects and multi-speaker dialogue, re-voice and clean up audio, transcribe audio and video, design synthetic voices…

    176 GitHub stars~3.7k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Speech To Text

    tadaspetra/loop

    Transcribe audio to text using ElevenLabs Scribe v2. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 3 repos~2k tokens
    Media & CreativeAuto-check passed
  • Video Translator

    shang-zhu/violin

    Dub a video into another language and generate subtitles using the default Together + Cartesia stack.

    1.1k GitHub stars~1k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes
  • Hyperframes Media

    chmonitor/chmonitor

    Audio and media assets for HyperFrames compositions, produced by one shared audio engine (scripts/audio.mjs) — multi-provider TTS (HeyGen / ElevenLabs / Kokoro local), background music + sound…

    299 GitHub starsUsed in 1 repo~2.8k tokens
    Media & CreativeAuto-check: notes
  • Local AI Use

    amd/skills

    Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.

    398 GitHub stars~5k tokensUpdated yesterday
    Media & CreativeAuto-check: notes

More from davepoon/buildwithclaude

All 245 skills in this repo
  • Qwen Vision

    davepoon/buildwithclaude

    A skill your agent uses when the user asks to "analyze video", "watch this video", "what happens in this video", "describe this clip", "review this footage", "classify these videos", "compare…

    3.6k GitHub starsUsed in 1 repo~1.2k tokens
    Auto-check passed
  • Hard Predict Future

    davepoon/buildwithclaude

    Activate this agent for any future-oriented question that requires deep quantitative analysis, historical precedents, and structured scenario planning.

    3.6k GitHub starsUsed in 1 repo~4.2k tokens
    Auto-check passed
  • iOS Hig Design Guide

    davepoon/buildwithclaude

    Build, update, and apply iOS design specifications using Apple Human Interface Guidelines (HIG) source data.

    3.6k GitHub stars~735 tokensUpdated 2 days ago
    Auto-check passed
  • Video Downloader

    davepoon/buildwithclaude

    Download YouTube videos with customizable quality and format options.

    3.6k GitHub starsUsed in 1 repo~871 tokens
    Auto-check passed
  • Atlas Cloud Media

    davepoon/buildwithclaude

    Discover Atlas Cloud image and video models, inspect their live schemas, and submit one confirmed media generation request with bounded GET polling.

    3.6k GitHub stars~852 tokensUpdated 2 days ago
    Auto-check passed
  • Slack Gif Creator

    davepoon/buildwithclaude

    Toolkit for creating animated GIFs optimized for Slack, with validators for size constraints and composable animation primitives.

    3.6k GitHub starsUsed in 12 repos~4.3k tokens
    Auto-check passed

Works with

Questions about Ops Voice

What does Ops Voice do?

Voice operations — make phone calls (Bland AI), text-to-speech (ElevenLabs), transcribe audio (Whisper/Groq). Ops Voice is an agent skill from davepoon/buildwithclaude. Voice operations — make phone calls (Bland AI), text-to-speech (ElevenLabs), transcribe audio (Whisper/Groq).

When should I use Ops Voice?

Ops Voice fits situations like: tasks that involve Text to speech and voice; tasks that involve Transcription.

How do I install Ops Voice in Claude Code?

Run `npx skills add davepoon/buildwithclaude --skill ops-voice -a claude-code`. Or copy the skill folder (plugins/claude-ops/skills/ops-voice in davepoon/buildwithclaude) into .claude/skills/ops-voice in your project. Claude Code loads it when a task matches its description.

How do I install Ops Voice in Codex?

Run `npx skills add davepoon/buildwithclaude --skill ops-voice -a codex`. Or copy the skill folder (plugins/claude-ops/skills/ops-voice in davepoon/buildwithclaude) into .agents/skills/ops-voice in your project. Codex loads it when a task matches its description.

Can I use Ops Voice in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add davepoon/buildwithclaude --skill ops-voice -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/ops-voice, .gemini/skills/ops-voice, .github/skills/ops-voice and .opencode/skills/ops-voice in your project.

What does Ops Voice need to run?

Going by SKILL.md and its folder, Ops Voice needs the command-line tools its instructions call (curl, jq and python3) and credentials named BLAND_AI_API_KEY, ELEVENLABS_API_KEY, GROQ_API_KEY and BLAND_KEY. Our summary lists: Python 3; A credential in BLAND_AI_API_KEY; A credential in BLAND_KEY. Its frontmatter pre-approves these tools: Bash, Read, Write, AskUserQuestion, WebFetch.

Does Ops Voice access the network?

SKILL.md names 3 domains. In commands or code: api.bland.ai, api.elevenlabs.io and api.groq.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Ops Voice safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Ops Voice use?

Ops Voice is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Ops Voice use?

About 1.5k tokens (SKILL.md is roughly 6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Ops Voice?

Skills that share tags, products or a category with Ops Voice: Elevenlabs Transcribe (qdhenry/Claude-Command-Suite, 1.3k stars), Elevenlabs (zapier/connectors, 176 stars), Speech To Text (tadaspetra/loop, 296 stars) and Video Translator (shang-zhu/violin, 1.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Ops Voice?

davepoon (a GitHub user) maintains it in davepoon/buildwithclaude, which has 3,604 GitHub stars. The repository holds 245 skills in this directory. The repository was last updated on October 6, 2026.

Source: davepoon/buildwithclaude on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.