Install the "voice-memo-sync" agent skill from https://github.com/LeoYeAI/openclaw-master-skills/tree/main/skills/voice-memo-sync into .claude/skills/voice-memo-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-memo-sync", then confirm the skill loads.
Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Type this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a codex
Project install goes to .agents/skills/; add -g for ~/.codex/skills/.
Install the "voice-memo-sync" agent skill from https://github.com/LeoYeAI/openclaw-master-skills/tree/main/skills/voice-memo-sync into .agents/skills/voice-memo-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-memo-sync", then confirm the skill loads.
Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a cursor
Project install goes to .agents/skills/; add -g for ~/.cursor/skills/.
Install the "voice-memo-sync" agent skill from https://github.com/LeoYeAI/openclaw-master-skills/tree/main/skills/voice-memo-sync into .cursor/skills/voice-memo-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-memo-sync", then confirm the skill loads.
Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a gemini-cli
Project install goes to .agents/skills/; add -g for ~/.gemini/skills/.
Install the "voice-memo-sync" agent skill from https://github.com/LeoYeAI/openclaw-master-skills/tree/main/skills/voice-memo-sync into .gemini/skills/voice-memo-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-memo-sync", then confirm the skill loads.
Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Installs for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a github-copilot
Project install goes to .agents/skills/; add -g for ~/.copilot/skills/.
Install the "voice-memo-sync" agent skill from https://github.com/LeoYeAI/openclaw-master-skills/tree/main/skills/voice-memo-sync into .github/skills/voice-memo-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-memo-sync", then confirm the skill loads.
GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a opencode
OpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
Install the "voice-memo-sync" agent skill from https://github.com/LeoYeAI/openclaw-master-skills/tree/main/skills/voice-memo-sync into .opencode/skills/voice-memo-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-memo-sync", then confirm the skill loads.
OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Facts
Skill name
voice-memo-sync
GitHub stars
2.2k
Token cost
~5.1k tokens
SKILL.md length
798 words
Files
12 (incl. scripts)
Skills in repo
1,235
Repo updated
First seen
Licence
MIT
At a glance
Sync, transcribe, and intelligently organize voice memos, audio/video files, and URLs.
Works in 8 steps: Detect Input Type / 识别输入类型 → Save Source Info / 保存源信息 → Get/Save Transcript / 获取保存转录 → …
Tasks that involve Transcription
SKILL.md covers Quick Start / 快速开始, When to Use / 何时使用, Supported Formats / 支持格式 and Processing Pipeline / 处理流程, plus 8 more sections
Runs Shell and Python scripts from its folder; calls python3, brew and osascript
What it does
Voice Memo Sync is an agent skill from LeoYeAI/openclaw-master-skills. Sync, transcribe, and intelligently organize voice memos, audio/video files, and URLs. 同步、转录、智能整理语音备忘录、音视频文件和视频链接。
Its SKILL.md is about 5.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 14 other files, including scripts (for example `README.md`, `README_CN.md` and `_meta.json`).
It sits in Media & Creative, covering Transcription. It works with Whisper and YouTube. The repository describes itself as: 🧠 Curated collection of 1209+ best OpenClaw skills — weekly updated by MyClaw.ai. The licence is MIT.
When your agent uses it
Tasks that involve Transcription
Example prompts
“/voice-memo-sync”
Requirements
Python 3
A Bash shell
Workflow steps
8 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit e5199b5. It shows what the files ask for, not the result of running them.
Tool permissions
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Runs code
Ships 6 files in scripts/ (Shell and Python), which the agent can run.
Shell commands in SKILL.md call:
python3
brew
osascript
make
pandoc
From the folder's file list and the shell code blocks in SKILL.md.
Network
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Credentials
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Context cost
Voice Memo Sync loads about 5.1k tokens when it runs. Until then it costs about 33 tokens; SKILL.md has 798 words of instructions outside code blocks.
Always· name and description, kept in context so the agent knows when to use it
~33
When it runs· the whole SKILL.md, loaded when a task matches
~5.1k
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
Safety
Auto-check passed
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
Download SKILL.mdSave it as .claude/skills/voice-memo-sync/SKILL.md (or your agent's skills folder). This skill also uses 11 other files; get the full folder from GitHub.
name
voice-memo-sync
description
Sync, transcribe, and intelligently organize voice memos, audio/video files, and URLs.
同步、转录、智能整理语音备忘录、音视频文件和视频链接。
version
1.6.1
author
Ying Wen
homepage
https://github.com/ying-wen/voice-memo-sync
license
MIT
Voice Memo Sync 🎙️
Intelligent voice/video transcription and organization system. 智能语音/视频转录与整理系统。
Quick Start / 快速开始
bash
# Run installation script / 运行安装脚本
cd ~/.openclaw/workspace/skills/voice-memo-sync
./scripts/install.sh
What it does / 安装内容:
Creates data directory memory/voice-memos/ / 创建数据目录
# Record to memory/voice-memos/sources/
echo '{"input":"...", "type":"...", "date":"YYYY-MM-DD"}' > sources/xxx.json
Step 3: Get/Save Transcript / 获取保存转录
bash
# Save to memory/voice-memos/transcripts/YYYY-MM-DD_source_title.md
# Include: source info + full raw transcript
Step 4: LLM Deep Processing / LLM深度整理
Read USER.md and MEMORY.md, combining user context.
**MODE SELECTION (Auto-detect or Manual Override) / 模式选择:**
┌─────────────────────────────────────────────────────────────────┐
│ Mode A: Solo Memo (Default) / 短语音 │
│ Trigger: < 5 min, single speaker, casual │
│ Output: Clean text + Key points + TODOs + Connections │
└─────────────────────────────────────────────────────────────────┘
┌─────────────────────────────────────────────────────────────────┐
│ Mode B: Deep Meeting / 深度会议 │
│ Trigger: 15-60 min, multi-speaker with labels │
│ Output: │
│ 1. Executive Summary (1 paragraph) │
│ 2. Chronological Detail by time blocks │
│ 3. Debate Flow (who said what, conflicts) │
│ 4. Decision Matrix (Issue → Decision → Rationale) │
│ 5. Action Items with owners │
│ 6. Vital Quotes (preserve Voice) │
└─────────────────────────────────────────────────────────────────┘
┌─────────────────────────────────────────────────────────────────┐
│ Mode C: Lecture / Talk / 讲座模式 (NEW) │
│ Trigger: Single speaker, 30min-3hr, structured presentation │
│ Output: │
│ 1. Executive Summary (1 paragraph) │
│ 2. **Argument Structure (论点层级)**: │
│ - Core Thesis (核心论点) │
│ - Supporting Arguments (分论点 1, 2, 3...) │
│ - Key Evidence/Examples for each argument │
│ - Counter-arguments addressed (if any) │
│ 3. Key Definitions (关键定义/概念) │
│ 4. Notable Quotes (金句, with timestamps if available) │
│ 5. Connections to User's Work (个人关联) │
│ 6. Questions Raised / Gaps (讲座未解决的问题) │
└─────────────────────────────────────────────────────────────────┘
┌─────────────────────────────────────────────────────────────────┐
│ Mode D: Lecture + Q&A / 讲座+问答 (NEW) │
│ Trigger: First part monologue, second part Q&A │
│ Output: │
│ **Part I: Lecture Section** (use Mode C structure) │
│ **Part II: Q&A Section** │
│ - Group questions by theme/topic (not chronological) │
│ - Format: Q1 → A1 (summary), Q2 → A2... │
│ - Highlight: Best Questions, Surprising Answers │
└─────────────────────────────────────────────────────────────────┘
┌─────────────────────────────────────────────────────────────────┐
│ Mode E: Long-form No-Speaker-Label / 超长无标注会议 (NEW) │
│ Trigger: > 90 min, NO speaker diarization (text is a blob) │
│ Strategy: │
│ 1. **Chunking**: Split into ~30min segments for processing │
│ 2. **Topic Detection**: Identify topic shift points │
│ (Don't force time blocks; use semantic breaks) │
│ 3. **Abandon Attribution**: Don't guess who said what │
│ Output: │
│ 1. Executive Summary │
│ 2. **Topic Blocks** (not time blocks): │
│ - Topic 1: [Summary] + [Key points] + [Quotes] │
│ - Topic 2: ... │
│ 3. Unresolved Issues / Open Questions │
│ 4. Action Items (may lack owners) │
│ 5. Full Cleaned Transcript (appended or linked) │
└─────────────────────────────────────────────────────────────────┘
**TWO-PASS PROCESSING for Long Content (> 60 min):**
- Pass 1 (Quick Scan): Identify structure type, speaker presence, topic shifts
- Pass 2 (Deep Process): Apply appropriate mode to each segment
**OUTPUT DENSITY LEVELS (User can request):**
- Level 1: Executive Only (1 page, for busy stakeholders)
- Level 2: Structured Summary (5-10 pages, default)
- Level 3: Full Annotated Transcript (everything, with margin notes)
Step 5: Save Processed Result / 保存处理结果
bash
# Save to memory/voice-memos/processed/YYYY-MM-DD_source_title.md
Step 6: Sync to Apple Notes (MANDATORY) / 同步到Apple Notes(必须执行)
⚠️ CRITICAL: This step is MANDATORY. Never skip it. ⚠️ 关键:此步骤必须执行,不可跳过。
⚠️ Apple Notes requires HTML format, NOT Markdown! ⚠️ Apple Notes 需要 HTML 格式,不能直接用 Markdown!
Correct workflow / 正确流程:
bash
# 1. Convert Markdown to HTML using pandoc (REQUIRED)
pandoc /path/to/processed.md -f markdown -t html -o /tmp/note-content.html
# 2. Create note with HTML content via AppleScript
osascript <<'EOF'
set htmlContent to do shell script "cat /tmp/note-content.html"
set noteTitle to "🎙️ Note Title"
tell application "Notes"
set folderName to "Voice Memos"
set targetFolder to missing value
repeat with f in folders
if name of f is folderName then
set targetFolder to f
exit repeat
end if
end repeat
if targetFolder is missing value then
make new folder with properties {name:folderName}
delay 1
set targetFolder to folder folderName
end if
tell targetFolder
make new note with properties {name:noteTitle, body:htmlContent}
end tell
end tell
EOF
All transcription runs locally by default / 所有转录默认在本地完成
Apple native transcripts extracted from local files / Apple原生转录从本地文件提取
Whisper runs locally / Whisper在本地运行
No data sent to external servers (unless user explicitly configures external API)
User data stored only in local memory directory
Troubleshooting / 故障排除
Whisper not found
bash
brew install openai-whisper
yt-dlp download fails
bash
# Update yt-dlp
brew upgrade yt-dlp
# Or use proxy
export ALL_PROXY=http://127.0.0.1:7890
Apple Notes folder not created
bash
# Manually create via AppleScript
osascript -e 'tell application "Notes" to tell account "iCloud" to make new folder with properties {name:"Voice Memos"}'
Transcription quality issues
bash
# Use larger model for better accuracy
# Edit config: whisper_model: "medium" or "large"
Changelog / 更新日志
v1.6.1 (2026-03-09)
CRITICAL FIX: Apple Notes sync step marked as MANDATORY (不可跳过).
FORMAT FIX: Explicit requirement to convert Markdown → HTML via pandoc before syncing.
Added complete AppleScript template with folder creation.
Common mistakes checklist to prevent format issues.
v1.6.0 (2026-03-09)
QTA Format Documentation: Added detailed technical reference for Apple's QTA file format.
Voice Memo Sync next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
Voice Memo Sync compared with similar skills
Skill
Stars
Used in
Tokens
Auto-check
Licence
Repo updated
Voice Memo Sync this skillLeoYeAI/openclaw-master-skills
A skill your agent uses when user provides a short video platform URL or local video/audio file and wants subtitles/AI summary, or when user asks to list their own AI Douyin historical tasks.
A skill your agent uses when users provide YouTube, Bilibili, or X/Twitter lecture URLs and want reader-first Chinese LaTeX/PDF notes with source-faithful claims, fluent authored prose, and verified…
The user shared a video URL, a YouTube/TikTok/stream link, a local video file, a screen recording, a meeting recording, or a playlist/folder of videos — "watch this", "summarize this video", "what's…
Manages pipelines on a DevOps quality and efficiency platform through its OpenAPI: list workspaces and templates, create, update, run and cancel pipelines, and read run records.
Patches OpenClaw's Feishu extension so an edited document triggers an isolated agent session that reads the doc and replies inline, turning it into a live chat space.
Multi-context memory management system for OpenClaw agents with group-isolated storage, global shared memory, workspace organization, and group-specific skills isolation.
Runs a brand's AI-search visibility work end to end: diagnosing how AI platforms represent it, repositioning it, producing AI-optimized content and monitoring ongoing mentions.
Installs and authenticates the gws CLI, then automates Gmail, Drive, Sheets, Calendar, Docs, Chat and Tasks with ready-made recipes, persona bundles and security audits.
Runs four advisor roles, a fitness coach, nutritionist, data analyst and TCM practitioner, to build a health profile and track workouts, diet and wellness over time.
Sync, transcribe, and intelligently organize voice memos, audio/video files, and URLs. Voice Memo Sync is an agent skill from LeoYeAI/openclaw-master-skills. Sync, transcribe, and intelligently organize voice memos, audio/video files, and URLs.
When should I use Voice Memo Sync?
Voice Memo Sync fits situations like: tasks that involve Transcription.
How do I install Voice Memo Sync in Claude Code?
Run `npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a claude-code`. Or copy the skill folder (skills/voice-memo-sync in LeoYeAI/openclaw-master-skills) into .claude/skills/voice-memo-sync in your project. Claude Code loads it when a task matches its description.
How do I install Voice Memo Sync in Codex?
Run `npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a codex`. Or copy the skill folder (skills/voice-memo-sync in LeoYeAI/openclaw-master-skills) into .agents/skills/voice-memo-sync in your project. Codex loads it when a task matches its description.
Can I use Voice Memo Sync in Cursor, Gemini CLI or GitHub Copilot?
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add LeoYeAI/openclaw-master-skills --skill voice-memo-sync -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-memo-sync, .gemini/skills/voice-memo-sync, .github/skills/voice-memo-sync and .opencode/skills/voice-memo-sync in your project.
What does Voice Memo Sync need to run?
Going by SKILL.md and its folder, Voice Memo Sync needs a shell and Python for the scripts in its folder and the command-line tools its instructions call (python3, brew, osascript, make and pandoc). Our summary lists: Python 3; A Bash shell.
Does Voice Memo Sync access the network?
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Is Voice Memo Sync safe to install?
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
What licence does Voice Memo Sync use?
Voice Memo Sync is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
How many tokens does Voice Memo Sync use?
About 5.1k tokens (SKILL.md is roughly 20k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
What are the alternatives to Voice Memo Sync?
Skills that share tags, products or a category with Voice Memo Sync: Video To Subtitle Summary (imlewc/video-to-subtitle-summary-skill, 218 stars), Watch Video (coreyhaines31/makerskills, 851 stars), Whisper Transcription (guia-matthieu/clawfu-skills, 150 stars) and Transcribe Md (hrescak/transcribe-md, 104 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Who maintains Voice Memo Sync?
LeoYeAI (a GitHub user) maintains it in LeoYeAI/openclaw-master-skills, which has 2,161 GitHub stars. The repository holds 1,235 skills in this directory. The repository was last updated on July 20, 2026.