Agent skill

Voice Dictation Setup

by witnesstodark in witnesstodark/mr-mak-workspace

Help configure optional local Whisper dictation into terminals and other apps, with language, hotkey and CPU or GPU checks.

MITAuto-check passedMedia & Creative

Install Voice Dictation Setup

skills CLI
$ npx skills add witnesstodark/mr-mak-workspace --skill voice-dictation-setup -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install witnesstodark/mr-mak-workspace voice-dictation-setup --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/witnesstodark/mr-mak-workspace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/voice-dictation-setup .claude/skills/voice-dictation-setup && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
voice-dictation-setup
GitHub stars
336
Token cost
~333 tokens
SKILL.md length
171 words
Files
1
Skills in repo
19
Repo updated
First seen
Licence
MIT

At a glance

Help configure optional local Whisper dictation into terminals and other apps, with language, hotkey and CPU or GPU checks.

  • Tasks that involve Transcription
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Tasks that involve Speech recognition and synthesis

What it does

Voice Dictation Setup is an agent skill from witnesstodark/mr-mak-workspace. Help configure optional local Whisper dictation into terminals and other apps, with language, hotkey and CPU or GPU checks.

Its SKILL.md is about 330 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Transcription and Speech recognition and synthesis. The repository describes itself as: A local desktop workspace for Codex and Claude Code, with project reports, creative skills and optional voice. The licence is MIT.

When your agent uses it

  • Tasks that involve Transcription
  • Tasks that involve Speech recognition and synthesis

Example prompts

  • “/voice-dictation-setup”

What it can do on your machine

Read from SKILL.md and the folder at commit 1e0c7c3. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Voice Dictation Setup loads about 333 tokens when it runs. Until then it costs about 36 tokens; SKILL.md has 171 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~36
When it runs · the whole SKILL.md, loaded when a task matches
~333

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from witnesstodark/mr-mak-workspace at commit 1e0c7c3, republished under its MIT licence (© witnesstodark). 171 words, ~333 tokens.

Download SKILL.mdSave it as .claude/skills/voice-dictation-setup/SKILL.md (or your agent's skills folder).
name
voice-dictation-setup
description
Help configure optional local Whisper dictation into terminals and other apps, with language, hotkey and CPU or GPU checks.

Voice dictation setup

Read the setup note and the chosen tool's current upstream instructions. Dictation is independent of Mr. Mak's conversational voice assistant. Keep the two workflows distinct.

Check the operating system, microphone, language and available hardware. Choose a local model that fits the machine. Verify the tool's defaults: a configured language or GPU build flag may need changing for this user. Do not promise that a build compiled with CUDA will run without the required runtime.

Choose a hotkey with the user before rebinding an existing shortcut. Test a short sentence in a plain text field, then in a terminal without submitting it. Check mixed technical vocabulary, silence and quick repeated recordings. Confirm that the next recording starts reliably after a failure.

Leave optional cloud formatting disabled unless requested. If enabled, explain which text goes to the selected provider and use the user's own credentials. Store only reusable setup notes in this project, not voice history or account settings. Do not restart Mr. Mak to configure a separate dictation application.

© witnesstodark, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/voice-dictation-setup of witnesstodark/mr-mak-workspace.

Open the folder on GitHubat commit 1e0c7c3

Compare with similar skills

Voice Dictation Setup next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Voice Dictation Setup compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Voice Dictation Setup this skillwitnesstodark/mr-mak-workspace336—~333Automated safety check: PassMIT
Audio Transcriptionmitsuhiko/agent-stuff3.2k—~1kAutomated safety check: PassApache-2.0
Speech Recognitiondpearson2699/swift-ios-skills1.2k—~3.7kAutomated safety check: PassCustom licence
Stepfun Asrdaymade/claude-code-skills1.4k—~3kAutomated safety check: PassMIT
Sttmikeyobrien/rho372—~149Automated safety check: PassMIT
Groq Core Workflow Bjeremylongshore/tons-of-skills-marketplace2.8k—~1.4kAutomated safety check: PassMIT

Similar skills

  • Audio Transcription

    mitsuhiko/agent-stuff

    Transcribe local audio/video and Apple Voice Memos quickly with cached MLX Whisper models, including bad/low-quality audio.

    3.2k GitHub stars~1k tokensUpdated 11 days ago
    Media & CreativeAuto-check passed
  • Speech Recognition

    dpearson2699/swift-ios-skills

    Transcribe speech to text using Apple's Speech framework. An agent skill from dpearson2699/swift-ios-skills.

    1.2k GitHub stars~3.7k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Stepfun Asr

    daymade/claude-code-skills

    Transcribes Chinese/English audio with StepFun's stepaudio-3-asr-max via its SSE endpoint (not /v1/audio/transcriptions) — one call handles long-form audio with no chunking.

    1.4k GitHub stars~3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Stt

    mikeyobrien/rho

    Speech-to-text — transcribe voice to text using device microphone.

    372 GitHub stars~149 tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Groq Core Workflow B

    jeremylongshore/tons-of-skills-marketplace

    A skill your agent uses when you need Groq's non-chat endpoints — transcribing or translating audio with Whisper, understanding images with Llama 4 vision, generating speech (TTS), or benchmarking…

    2.8k GitHub stars~1.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Whisper

    AlexAI-MCP/hermes-CCC

    OpenAI Whisper for speech recognition and transcription — local inference, multiple model sizes, language detection, and subtitle generation.

    135 GitHub stars~1.9k tokensUpdated 6 mo ago
    Media & CreativeAuto-check passed

More from witnesstodark/mr-mak-workspace

All 19 skills in this repo
  • Fal AI Generation

    witnesstodark/mr-mak-workspace

    Generate and edit images, create video, audio, 3D assets or material maps through fal.ai MCP or the Python queue client.

    336 GitHub stars~1k tokensUpdated 3 days ago
    Auto-check: notes
  • Game Vfx Workflow

    witnesstodark/mr-mak-workspace

    Develop game effects from visual references and motion studies through an engine handoff, with event timing, lifecycle, readability and native review.

    336 GitHub stars~896 tokensUpdated 3 days ago
    Auto-check passed
  • Game Audio Workflow

    witnesstodark/mr-mak-workspace

    Build and review game sound banks, map chosen takes to real events, and prepare or verify runtime playback with timing, repetition and mix limits.

    336 GitHub stars~811 tokensUpdated 3 days ago
    Auto-check passed
  • Game Level Design

    witnesstodark/mr-mak-workspace

    Turn a game-level brief or existing scene snapshot into a reviewable spatial plan, then apply approved changes with coordinate, navigation and scene checks.

    336 GitHub stars~735 tokensUpdated 3 days ago
    Auto-check passed
  • Game UI Workflow

    witnesstodark/mr-mak-workspace

    Design and implement game HUDs, menus, icons and portrait systems from visual proposals through reusable engine assets and input checks.

    336 GitHub stars~867 tokensUpdated 3 days ago
    Auto-check passed
  • Gameplay Visual Review

    witnesstodark/mr-mak-workspace

    Verify a visual game change using controlled native captures, targeted behavioral checks and a reviewable evidence record.

    336 GitHub stars~720 tokensUpdated 3 days ago
    Auto-check passed

Questions about Voice Dictation Setup

What does Voice Dictation Setup do?

Help configure optional local Whisper dictation into terminals and other apps, with language, hotkey and CPU or GPU checks. Voice Dictation Setup is an agent skill from witnesstodark/mr-mak-workspace. Help configure optional local Whisper dictation into terminals and other apps, with language, hotkey and CPU or GPU checks.

When should I use Voice Dictation Setup?

Voice Dictation Setup fits situations like: tasks that involve Transcription; tasks that involve Speech recognition and synthesis.

How do I install Voice Dictation Setup in Claude Code?

Run `npx skills add witnesstodark/mr-mak-workspace --skill voice-dictation-setup -a claude-code`. Or copy the skill folder (.agents/skills/voice-dictation-setup in witnesstodark/mr-mak-workspace) into .claude/skills/voice-dictation-setup in your project. Claude Code loads it when a task matches its description.

How do I install Voice Dictation Setup in Codex?

Run `npx skills add witnesstodark/mr-mak-workspace --skill voice-dictation-setup -a codex`. Or copy the skill folder (.agents/skills/voice-dictation-setup in witnesstodark/mr-mak-workspace) into .agents/skills/voice-dictation-setup in your project. Codex loads it when a task matches its description.

Can I use Voice Dictation Setup in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add witnesstodark/mr-mak-workspace --skill voice-dictation-setup -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-dictation-setup, .gemini/skills/voice-dictation-setup, .github/skills/voice-dictation-setup and .opencode/skills/voice-dictation-setup in your project.

What does Voice Dictation Setup need to run?

SKILL.md names no scripts, command-line tools or credentials: Voice Dictation Setup is instructions for the agent only.

Does Voice Dictation Setup access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Voice Dictation Setup safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Voice Dictation Setup use?

Voice Dictation Setup is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Voice Dictation Setup use?

About 333 tokens (SKILL.md is roughly 1.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Voice Dictation Setup?

Skills that share tags, products or a category with Voice Dictation Setup: Audio Transcription (mitsuhiko/agent-stuff, 3.2k stars), Speech Recognition (dpearson2699/swift-ios-skills, 1.2k stars), Stepfun Asr (daymade/claude-code-skills, 1.4k stars) and Stt (mikeyobrien/rho, 372 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Voice Dictation Setup?

witnesstodark (a GitHub user) maintains it in witnesstodark/mr-mak-workspace, which has 336 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on October 5, 2026.

Source: witnesstodark/mr-mak-workspace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.