Agent skill

Voice Memo

by letta-ai in letta-ai/lettabot

Reply with voice memos using text-to-speech. An agent skill from letta-ai/lettabot.

Apache-2.0Auto-check passedMedia & Creative

Install Voice Memo

skills CLI
$ npx skills add letta-ai/lettabot --skill voice-memo -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install letta-ai/lettabot voice-memo --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/letta-ai/lettabot.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/voice-memo .claude/skills/voice-memo && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
voice-memo
GitHub stars
327
Token cost
~487 tokens
SKILL.md length
218 words
Files
2
Skills in repo
7
Repo updated
First seen
Licence
Apache-2.0

At a glance

Reply with voice memos using text-to-speech. An agent skill from letta-ai/lettabot.

  • The user sends a voice message
  • SKILL.md covers Usage, When to Use Voice, When NOT to Use Voice and Notes
  • Needs ELEVENLABS_API_KEY
  • Asks for an audio reply

What it does

Voice Memo is an agent skill from letta-ai/lettabot. Reply with voice memos using text-to-speech. Use when the user sends a voice message, asks for an audio reply, or when a voice response would be more natural.

Its SKILL.md is about 490 tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file.

It sits in Media & Creative, covering Transcription and Text to speech and voice. It works with Telegram. The repository describes itself as: Archived - has been replaced by Letta Code channels/schedules! The licence is Apache-2.0.

When your agent uses it

  • The user sends a voice message
  • Asks for an audio reply
  • A voice response would be more natural

Example prompts

  • “/voice-memo”

Requirements

  • A credential in ELEVENLABS_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit 99c3b5d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • ELEVENLABS_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Voice Memo loads about 487 tokens when it runs. Until then it costs about 42 tokens; SKILL.md has 218 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~42
When it runs · the whole SKILL.md, loaded when a task matches
~487

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from letta-ai/lettabot at commit 99c3b5d, republished under its Apache-2.0 licence (© letta-ai). 218 words, ~487 tokens.

Download SKILL.mdSave it as .claude/skills/voice-memo/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
voice-memo
description
Reply with voice memos using text-to-speech. Use when the user sends a voice message, asks for an audio reply, or when a voice response would be more natural.

Voice Memo Responses

Generate voice memos using TTS and send them as native voice notes.

Usage

Use the <voice> directive to send voice memos. No tool calls needed:

<actions>
  <voice>Hey, here's a quick update on that thing we discussed.</voice>
</actions>

With accompanying text:

<actions>
  <voice>Here's the summary as audio.</voice>
</actions>
And here it is in text form too!
Silent mode (heartbeats, cron)

For background tasks that need to send voice without a user message context:

bash
OUTPUT=$(lettabot-tts "Your message here") || exit 1
lettabot-message send --file "$OUTPUT" --voice

When to Use Voice

  • User sent a voice message and a voice reply feels natural
  • User explicitly asks for a voice/audio response
  • Short, conversational responses (voice is awkward for long technical content)

When NOT to Use Voice

  • Code snippets, file paths, URLs, or structured data (these should be text)
  • Long responses -- keep voice memos under ~30 seconds of speech
  • When the user has indicated a preference for text
  • When ELEVENLABS_API_KEY is not set

Notes

  • Audio format is OGG Opus, which renders as native voice bubbles on Telegram and WhatsApp
  • Discord and Slack will show it as a playable audio attachment
  • Use cleanup="true" to delete the audio file after sending
  • The data/outbound/ directory is the default allowed path for send-file directives
  • The script uses $LETTABOT_WORKING_DIR to output files to the correct directory
  • On Telegram, if the user has voice message privacy enabled (Telegram Premium), the bot falls back to sending as an audio file instead of a voice bubble. Users can allow voice messages via Settings > Privacy and Security > Voice Messages.

© letta-ai, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/voice-memo of letta-ai/lettabot.

  • SKILL.md
  • lettabot-tts

Open the folder on GitHubat commit 99c3b5d

Compare with similar skills

Voice Memo next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Voice Memo compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Voice Memo this skillletta-ai/lettabot327—~487Automated safety check: PassApache-2.0
Voice Personadavepoon/buildwithclaude3.6k—~947Automated safety check: PassMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0
Edu Chem Videowy51ai/edulab1.4k—~2.1kAutomated safety check: NotesApache-2.0
Edu Math Videowy51ai/edulab1.4k—~2.5kAutomated safety check: NotesApache-2.0
Edu Physics Videowy51ai/edulab1.4k—~2.3kAutomated safety check: NotesApache-2.0

Similar skills

  • Voice Persona

    davepoon/buildwithclaude

    让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl.

    3.6k GitHub stars~947 tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Edu Chem Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a chemistry problem (化学题: 氧化还原配平 双线桥 电子守恒, 物质的量计算, 化学平衡 三段式 平衡常数 转化率 反应速率, 离子反应, 电化学, 溶液 滴定…

    1.4k GitHub stars~2.1k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Edu Math Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a math problem (数学题, geometry, algebra, functions, motion/行程 problems), from a problem screenshot…

    1.4k GitHub stars~2.5k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Edu Physics Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a physics problem (物理题: mechanics/力学 受力分析 牛顿定律 斜面 传送带 板块 平抛 圆周 能量 动量, optics/光学 折射…

    1.4k GitHub stars~2.3k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Elevenlabs Transcribe

    qdhenry/Claude-Command-Suite

    Transcribes audio/video files using ElevenLabs Scribe v2 API.

    1.3k GitHub stars~1.5k tokensUpdated 7 mo ago
    Media & CreativeAuto-check: notes

More from letta-ai/lettabot

  • Linear

    letta-ai/lettabot

    Manage Linear issues via GraphQL API. An agent skill from letta-ai/lettabot.

    327 GitHub stars~583 tokensUpdated 4 mo ago
    Auto-check passed
  • Lettabot

    letta-ai/lettabot

    Set up and run LettaBot - a multi-channel AI assistant for Telegram, Slack, Discord, WhatsApp, and Signal.

    327 GitHub stars~3.4k tokensUpdated 4 mo ago
    Auto-check passed
  • Scheduling

    letta-ai/lettabot

    Create scheduled tasks and one-off reminders. An agent skill from letta-ai/lettabot.

    327 GitHub stars~852 tokensUpdated 4 mo ago
    Auto-check passed
  • Bluesky

    letta-ai/lettabot

    Post, reply, like, and repost on Bluesky using the lettabot-bluesky CLI.

    327 GitHub stars~954 tokensUpdated 4 mo ago
    Auto-check passed
  • Cron

    letta-ai/lettabot

    Create and manage scheduled tasks (cron jobs) that send you messages at specified times.

    327 GitHub stars~814 tokensUpdated 4 mo ago
    Auto-check passed
  • Google

    letta-ai/lettabot

    Google Workspace CLI (gog) for Gmail, Calendar, Drive, Contacts, Sheets, and Docs.

    327 GitHub stars~774 tokensUpdated 4 mo ago
    Auto-check: notes

Works with

Questions about Voice Memo

What does Voice Memo do?

Reply with voice memos using text-to-speech. An agent skill from letta-ai/lettabot. Voice Memo is an agent skill from letta-ai/lettabot. Reply with voice memos using text-to-speech.

When should I use Voice Memo?

Voice Memo fits situations like: the user sends a voice message; asks for an audio reply; A voice response would be more natural.

How do I install Voice Memo in Claude Code?

Run `npx skills add letta-ai/lettabot --skill voice-memo -a claude-code`. Or copy the skill folder (skills/voice-memo in letta-ai/lettabot) into .claude/skills/voice-memo in your project. Claude Code loads it when a task matches its description.

How do I install Voice Memo in Codex?

Run `npx skills add letta-ai/lettabot --skill voice-memo -a codex`. Or copy the skill folder (skills/voice-memo in letta-ai/lettabot) into .agents/skills/voice-memo in your project. Codex loads it when a task matches its description.

Can I use Voice Memo in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add letta-ai/lettabot --skill voice-memo -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-memo, .gemini/skills/voice-memo, .github/skills/voice-memo and .opencode/skills/voice-memo in your project.

What does Voice Memo need to run?

Going by SKILL.md and its folder, Voice Memo needs credentials named ELEVENLABS_API_KEY. Our summary lists: A credential in ELEVENLABS_API_KEY.

Does Voice Memo access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Voice Memo safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Voice Memo use?

Voice Memo is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Voice Memo use?

About 487 tokens (SKILL.md is roughly 1.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Voice Memo?

Skills that share tags, products or a category with Voice Memo: Voice Persona (davepoon/buildwithclaude, 3.6k stars), HyperFrames Media Use (heygen-com/hyperframes, 60k stars), Edu Chem Video (wy51ai/edulab, 1.4k stars) and Edu Math Video (wy51ai/edulab, 1.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Voice Memo?

letta-ai (a GitHub organization) maintains it in letta-ai/lettabot, which has 327 GitHub stars. The repository holds 7 skills in this directory. The repository was last updated on May 25, 2026.

Source: letta-ai/lettabot on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.