Voxtype Install
pchalasani/claude-code-tools
Guide the user through installing, configuring, and launching voxtype — local on-device voice dictation (speech-to-text that types wherever the cursor is).
A skill your agent uses whenever the user wants to transcribe audio to text, convert speech to text, or get a transcript from an audio or video file.
$ npx skills add NoizAI/skills --skill speech-to-text -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install NoizAI/skills speech-to-text --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/NoizAI/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/speech-to-text .claude/skills/speech-to-text && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "speech-to-text" agent skill from https://github.com/NoizAI/skills/tree/main/skills/speech-to-text into .claude/skills/speech-to-text/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "speech-to-text", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/NoizAI/skills/tree/main/skills/speech-to-textType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add NoizAI/skills --skill speech-to-text -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install NoizAI/skills speech-to-text --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/NoizAI/skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/speech-to-text .agents/skills/speech-to-text && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "speech-to-text" agent skill from https://github.com/NoizAI/skills/tree/main/skills/speech-to-text into .agents/skills/speech-to-text/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "speech-to-text", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add NoizAI/skills --skill speech-to-text -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install NoizAI/skills speech-to-text --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/NoizAI/skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/speech-to-text .cursor/skills/speech-to-text && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "speech-to-text" agent skill from https://github.com/NoizAI/skills/tree/main/skills/speech-to-text into .cursor/skills/speech-to-text/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "speech-to-text", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/NoizAI/skills.git --path skills/speech-to-text--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add NoizAI/skills --skill speech-to-text -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install NoizAI/skills speech-to-text --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/NoizAI/skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/speech-to-text .gemini/skills/speech-to-text && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "speech-to-text" agent skill from https://github.com/NoizAI/skills/tree/main/skills/speech-to-text into .gemini/skills/speech-to-text/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "speech-to-text", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install NoizAI/skills speech-to-textInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add NoizAI/skills --skill speech-to-text -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/NoizAI/skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/speech-to-text .github/skills/speech-to-text && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "speech-to-text" agent skill from https://github.com/NoizAI/skills/tree/main/skills/speech-to-text into .github/skills/speech-to-text/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "speech-to-text", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add NoizAI/skills --skill speech-to-text -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install NoizAI/skills speech-to-text --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/NoizAI/skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/speech-to-text .opencode/skills/speech-to-text && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "speech-to-text" agent skill from https://github.com/NoizAI/skills/tree/main/skills/speech-to-text into .opencode/skills/speech-to-text/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "speech-to-text", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
speech-to-textA skill your agent uses whenever the user wants to transcribe audio to text, convert speech to text, or get a transcript from an audio or video file.
Speech To Text is an agent skill from NoizAI/skills. Use this skill whenever the user wants to transcribe audio to text, convert speech to text, or get a transcript from an audio or video file. Triggers include: any mention of 'transcribe', 'transcription', 'speech to text', 'STT', 'convert audio to text', 'what does this audio say', 'get transcript', 'subtitle generation', or requests to extract spoken words from a file. Also use when the user wants speaker identification from audio, timestamps for captions, or multilingual transcription.
Its SKILL.md is about 920 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/stt.py`).
It sits in Media & Creative, covering Transcription and Speech recognition and synthesis. The repository describes itself as: Allow your 🦞 bot to Shout, Speak, with "human" vibe.
Read from SKILL.md and the folder at commit 779f30a. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
python3pipFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
noiz.aiAlso links to:
developers.noiz.aiFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
NOIZ_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Speech To Text loads about 917 tokens when it runs. Until then it costs about 127 tokens; SKILL.md has 249 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
Without a licence we can't republish the file, so here is its outline and opening line. It has 249 words (~917 tokens).
“Transcribe any audio file to text. Supports multilingual auto-detection, timestamps, and speaker labels.”
SKILL.md and 1 other file (scripts) in skills/speech-to-text of NoizAI/skills.
Open the folder on GitHubat commit 779f30a
Speech To Text next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Speech To Text this skillNoizAI/skills | 526 | — | ~917 | Automated safety check: Pass | None | |
| Voxtype Installpchalasani/claude-code-tools | 2k | — | ~857 | Automated safety check: Notes | MIT | |
| Podcast Transcript FetcherVarnan-Tech/opendirectory | 674 | — | ~1.9k | Automated safety check: Notes | MIT | |
| Audio Transcriptionmitsuhiko/agent-stuff | 3.2k | — | ~1k | Automated safety check: Pass | Apache-2.0 | |
| Speech Recognitiondpearson2699/swift-ios-skills | 1.2k | — | ~3.7k | Automated safety check: Pass | Custom licence | |
| Speech Buildcnemri/google-genai-skills | 127 | — | ~430 | Automated safety check: Pass | MIT |
pchalasani/claude-code-tools
Guide the user through installing, configuring, and launching voxtype — local on-device voice dictation (speech-to-text that types wherever the cursor is).
Varnan-Tech/opendirectory
A skill your agent uses when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast.
mitsuhiko/agent-stuff
Transcribe local audio/video and Apple Voice Memos quickly with cached MLX Whisper models, including bad/low-quality audio.
dpearson2699/swift-ios-skills
Transcribe speech to text using Apple's Speech framework. An agent skill from dpearson2699/swift-ios-skills.
cnemri/google-genai-skills
Generate and transcribe speech using Google's Gemini-TTS and Chirp 3 models.
cat-xierluo/legal-skills
转录稿纠错与轻度优化。本技能应在用户需要按用户词典纠正 ASR 转录稿同音字与英文专有名称漂移时使用。不要用于:重写为课程章节、报告、总结,或完全空白的素材创作。
NoizAI/skills
Build a music video from an existing song and lyrics, either as p5.js/p5.brush animation or by assembling local images and clips.
NoizAI/skills
A skill your agent uses whenever the user wants speech to sound more human, companion-like, or emotionally expressive.
NoizAI/skills
Chat with any real person or fictional character in their own voice by automatically finding their speech online, extracting a clean reference sample, and generating audio replies.
NoizAI/skills
A skill your agent uses whenever the user wants to generate sound effects, ambient audio, or short audio clips from a text description.
NoizAI/skills
A skill your agent uses whenever the user wants to generate a song or music with vocals from a text description and/or lyrics, or cover an existing song in a new style.
NoizAI/skills
A skill your agent uses whenever the user wants to convert text into speech, generate audio from text, or produce voiceovers.
Categories
A skill your agent uses whenever the user wants to transcribe audio to text, convert speech to text, or get a transcript from an audio or video file. Speech To Text is an agent skill from NoizAI/skills. Use this skill whenever the user wants to transcribe audio to text, convert speech to text, or get a transcript from an audio or video file.
Speech To Text fits situations like: the user wants to transcribe audio to text; convert speech to text; get a transcript from an audio; include: any mention of transcribe.
Run `npx skills add NoizAI/skills --skill speech-to-text -a claude-code`. Or copy the skill folder (skills/speech-to-text in NoizAI/skills) into .claude/skills/speech-to-text in your project. Claude Code loads it when a task matches its description.
Run `npx skills add NoizAI/skills --skill speech-to-text -a codex`. Or copy the skill folder (skills/speech-to-text in NoizAI/skills) into .agents/skills/speech-to-text in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add NoizAI/skills --skill speech-to-text -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/speech-to-text, .gemini/skills/speech-to-text, .github/skills/speech-to-text and .opencode/skills/speech-to-text in your project.
Going by SKILL.md and its folder, Speech To Text needs Python for the scripts in its folder, the command-line tools its instructions call (python3 and pip) and credentials named NOIZ_API_KEY. Our summary lists: Python 3; A credential in NOIZ_API_KEY; A credential in YOUR_KEY.
SKILL.md names 2 domains. In commands or code: noiz.ai; the agent is likely to contact it when it follows the instructions. As links in the text: developers.noiz.ai. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
No licence was found for Speech To Text or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.
About 917 tokens (SKILL.md is roughly 3.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Speech To Text: Voxtype Install (pchalasani/claude-code-tools, 2k stars), Podcast Transcript Fetcher (Varnan-Tech/opendirectory, 674 stars), Audio Transcription (mitsuhiko/agent-stuff, 3.2k stars) and Speech Recognition (dpearson2699/swift-ios-skills, 1.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
NoizAI (a GitHub organization) maintains it in NoizAI/skills, which has 526 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on September 28, 2026.
Source: NoizAI/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.