Yichen Asr
mcncarl/yichen-skills
逸尘自用的统一音视频转写入口,在 StepFun Step ASR 与火山引擎豆包 ASR 之间按输出需求、安全边界和可用状态路由。用于本地音频或视频的纯文本转写、时间戳、SRT 字幕、口播粗剪,以及转写前体检;用户明确指定服务商时不得静默切换。Use when a local audio or video file needs transcription and the correct…
Design voice interactions and speech interfaces that work for people with diverse speech patterns, accents, and communication styles.
$ npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Owl-Listener/inclusive-design-skills voice-interaction --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/inclusive-interaction/skills/voice-interaction .claude/skills/voice-interaction && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "voice-interaction" agent skill from https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interaction into .claude/skills/voice-interaction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-interaction", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interactionType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Owl-Listener/inclusive-design-skills voice-interaction --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/inclusive-interaction/skills/voice-interaction .agents/skills/voice-interaction && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "voice-interaction" agent skill from https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interaction into .agents/skills/voice-interaction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-interaction", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Owl-Listener/inclusive-design-skills voice-interaction --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/inclusive-interaction/skills/voice-interaction .cursor/skills/voice-interaction && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "voice-interaction" agent skill from https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interaction into .cursor/skills/voice-interaction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-interaction", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Owl-Listener/inclusive-design-skills.git --path inclusive-interaction/skills/voice-interaction--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Owl-Listener/inclusive-design-skills voice-interaction --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/inclusive-interaction/skills/voice-interaction .gemini/skills/voice-interaction && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "voice-interaction" agent skill from https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interaction into .gemini/skills/voice-interaction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-interaction", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Owl-Listener/inclusive-design-skills voice-interactionInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/inclusive-interaction/skills/voice-interaction .github/skills/voice-interaction && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "voice-interaction" agent skill from https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interaction into .github/skills/voice-interaction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-interaction", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Owl-Listener/inclusive-design-skills voice-interaction --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/inclusive-interaction/skills/voice-interaction .opencode/skills/voice-interaction && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "voice-interaction" agent skill from https://github.com/Owl-Listener/inclusive-design-skills/tree/main/inclusive-interaction/skills/voice-interaction into .opencode/skills/voice-interaction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-interaction", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
voice-interactionDesign voice interactions and speech interfaces that work for people with diverse speech patterns, accents, and communication styles.
Voice Interaction is an agent skill from Owl-Listener/inclusive-design-skills. Design voice interactions and speech interfaces that work for people with diverse speech patterns, accents, and communication styles. Use when designing voice commands, voice search, dictation, voice assistants, or any interface that accepts speech input. Triggers on: voice, speech, dictation, voice command, voice search, speech recognition, accent, stutter, speech disability, non-verbal, AAC, voice assistant, talk to type.
Its SKILL.md is about 730 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in AI & LLM Engineering, covering Speech recognition and synthesis and Transcription. The repository describes itself as: Inclusive design skills for AI coding agents — from cognitive accessibility to adaptive interfaces, inclusive research, and accessibility decision-making. The licence is MIT.
5 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 6e0740f. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Voice Interaction loads about 730 tokens when it runs. Until then it costs about 111 tokens; SKILL.md has 341 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Owl-Listener/inclusive-design-skills at commit 6e0740f, republished under its MIT licence (© Owl-Listener). 341 words, ~730 tokens.
.claude/skills/voice-interaction/SKILL.md (or your agent's skills folder).Design voice interfaces that work for the full range of human speech — including accents, speech disabilities, non-native speakers, and people in noisy or quiet environments.
© Owl-Listener, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in inclusive-interaction/skills/voice-interaction of Owl-Listener/inclusive-design-skills.
Open the folder on GitHubat commit 6e0740f
Voice Interaction next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Voice Interaction this skillOwl-Listener/inclusive-design-skills | 104 | — | ~730 | Automated safety check: Pass | MIT | |
| Yichen Asrmcncarl/yichen-skills | 4.4k | — | ~780 | Automated safety check: Pass | Custom licence | |
| Youtube FetcherJimmySadek/youtube-fetcher-to-markdown | 485 | — | ~3.1k | Automated safety check: Pass | MIT | |
| Volcengine Asrysyecust/lecture-to-notes | 273 | — | ~783 | Automated safety check: Pass | Custom licence | |
| Openai Whisperhuangruiteng/CS-Notes | 4k | 18 repos | ~228 | Automated safety check: Pass | MIT | |
| Ax Audiodosco/aithy | 107 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 |
mcncarl/yichen-skills
逸尘自用的统一音视频转写入口,在 StepFun Step ASR 与火山引擎豆包 ASR 之间按输出需求、安全边界和可用状态路由。用于本地音频或视频的纯文本转写、时间戳、SRT 字幕、口播粗剪,以及转写前体检;用户明确指定服务商时不得静默切换。Use when a local audio or video file needs transcription and the correct…
JimmySadek/youtube-fetcher-to-markdown
Retrieve transcripts from YouTube, Instagram, TikTok, X, Vimeo and other video sites, summarize or analyze what was said (and shown on screen), or save an Obsidian-ready Markdown knowledge-base note…
ysyecust/lecture-to-notes
Transcribe local audio or video with Volcengine Doubao file ASR, including BigASR 1.0 Turbo direct upload and asynchronous 1.0 standard, 1.0 idle, or 2.0 standard jobs through TOS.
huangruiteng/CS-Notes
Local speech-to-text with the Whisper CLI (no API key). An agent skill from huangruiteng/CS-Notes.
dosco/aithy
This skill helps an LLM generate correct audio code with @ax-llm/ax.
rapidaai/voice-ai
Add or modify speech-to-text providers in assistant-api with transport-aware ingestion (WS/SDK/HTTP), transcript packet correctness, and UI/provider wiring.
Owl-Listener/inclusive-design-skills
Writes usage scenarios, use cases and storyboards that show real people using screen readers, switches, voice control and other assistive technology to finish tasks.
Owl-Listener/inclusive-design-skills
Designs forgiving forms and flows: prevent input errors, write messages that say what happened and what to do, and add undo, confirmation and recovery paths.
Owl-Listener/inclusive-design-skills
Designs heading hierarchies for screen reader navigation and cognitive accessibility on pages, articles, dashboards and forms.
Owl-Listener/inclusive-design-skills
Maps a feature across a range of vision, hearing, motor and cognitive ability to find where the design starts to fail, and to explain accessibility scope to stakeholders.
Owl-Listener/inclusive-design-skills
Track and manage accessibility debt — known accessibility issues that have been deferred.
Owl-Listener/inclusive-design-skills
Plan what to test, how to test, and who should test for accessibility.
Categories
Design voice interactions and speech interfaces that work for people with diverse speech patterns, accents, and communication styles. Voice Interaction is an agent skill from Owl-Listener/inclusive-design-skills. Design voice interactions and speech interfaces that work for people with diverse speech patterns, accents, and communication styles.
Voice Interaction fits situations like: designing voice commands; voice assistants; any interface that accepts speech input; speech recognition.
Run `npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a claude-code`. Or copy the skill folder (inclusive-interaction/skills/voice-interaction in Owl-Listener/inclusive-design-skills) into .claude/skills/voice-interaction in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a codex`. Or copy the skill folder (inclusive-interaction/skills/voice-interaction in Owl-Listener/inclusive-design-skills) into .agents/skills/voice-interaction in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Owl-Listener/inclusive-design-skills --skill voice-interaction -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-interaction, .gemini/skills/voice-interaction, .github/skills/voice-interaction and .opencode/skills/voice-interaction in your project.
SKILL.md names no scripts, command-line tools or credentials: Voice Interaction is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Voice Interaction is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 730 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Voice Interaction: Yichen Asr (mcncarl/yichen-skills, 4.4k stars), Youtube Fetcher (JimmySadek/youtube-fetcher-to-markdown, 485 stars), Volcengine Asr (ysyecust/lecture-to-notes, 273 stars) and Openai Whisper (huangruiteng/CS-Notes, 4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Owl-Listener (a GitHub user) maintains it in Owl-Listener/inclusive-design-skills, which has 104 GitHub stars. The repository holds 55 skills in this directory. The repository was last updated on June 9, 2026.
Source: Owl-Listener/inclusive-design-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.