Agent skill

Oma Voice

by first-fluke in first-fluke/oh-my-agent

Generate speech or transcribe audio locally with Voicebox. An agent skill from first-fluke/oh-my-agent.

MITAuto-check passedMedia & Creative

Install Oma Voice

skills CLI
$ npx skills add first-fluke/oh-my-agent --skill oma-voice -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install first-fluke/oh-my-agent oma-voice --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/first-fluke/oh-my-agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/oma-voice .claude/skills/oma-voice && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
oma-voice
GitHub stars
1.3k
Token cost
~3.3k tokens
SKILL.md length
1,598 words
Files
28 (incl. references)
Skills in repo
57
Repo updated
First seen
Licence
MIT

At a glance

Generate speech or transcribe audio locally with Voicebox. An agent skill from first-fluke/oh-my-agent.

  • Works in 4 steps: Detect the requested mode: notification,… → Verify Voicebox is reachable via MCP… → On the first run only, call MCP… → …
  • Meeting transcription
  • SKILL.md covers Scheduling, Structural Flow, Logical Operations and References
  • Calls claude

What it does

Oma Voice is an agent skill from first-fluke/oh-my-agent. Generate speech or transcribe audio locally with Voicebox. Use for narration, voice assets, dictation, and meeting transcription.

Its SKILL.md is about 3.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 33 other files, including reference files (for example `config/voice-config.yaml`, `references/_shared/conditional/experiment-ledger.md` and `references/_shared/conditional/exploration-loop.md`).

It sits in Media & Creative, covering Transcription and Text to speech and voice. The repository describes itself as: Mechanical verification for AI coding agents — skills pack or full harness (stop-hook gates, artifact checks, independent judges). The licence is MIT.

When your agent uses it

  • Meeting transcription
  • Tasks that involve Transcription
  • Tasks that involve Text to speech and voice

Example prompts

  • “/oma-voice”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Detect the requested mode: notification, asset TTS, or transcription.
  2. Verify Voicebox is reachable via MCP handshake or GET /health.
  3. On the first run only, call MCP tools/list and cache the resolved tool names.
  4. For notification or asset TTS, resolve the target voice profile id. For transcription,

What it can do on your machine

Read from SKILL.md and the folder at commit 268bb4a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • claude

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Oma Voice loads about 3.3k tokens when it runs, and up to ~18k if it reads all its reference files. Until then it costs about 35 tokens; SKILL.md has 1,598 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~35
When it runs · the whole SKILL.md, loaded when a task matches
~3.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~18k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from first-fluke/oh-my-agent at commit 268bb4a, republished under its MIT licence (© first-fluke). 1,598 words, ~3,329 tokens.

Download SKILL.mdSave it as .claude/skills/oma-voice/SKILL.md (or your agent's skills folder). This skill also uses 27 other files; get the full folder from GitHub.
name
oma-voice
description
Generate speech or transcribe audio locally with Voicebox. Use for narration, voice assets, dictation, and meeting transcription.

Voice Skill - Local TTS and STT via Voicebox

Scheduling

Goal

Drive the Voicebox local app through its MCP server so any MCP-aware agent can speak (TTS) or listen (STT) without invoking cloud vendors. The skill standardizes intent routing, voice profile resolution, output layout, and guardrails while voicebox itself owns the engines, voice cloning UI, captures archive, and stories editor.

Intent signature
  • User asks to generate speech, narrate text, produce a voiceover, create an mp3 or wav from text.
  • User wants an audio file transcribed into text, meeting notes, or a transcript.
  • User asks for a voice notification when a long task completes or a workflow step is blocked.
  • Another skill needs local audio generation infrastructure.
When to use
  • Generating short notification audio for agent task completion or blockers.
  • Producing voiceover, narration, or audio assets (mp3 or wav) for apps and content.
  • Transcribing local audio files (mp3, wav, m4a, webm, flac) to Markdown.
  • Comparing voice profiles by re-running the same text against different profile ids.
When NOT to use
  • Cloud TTS or high-fidelity multilingual cloud voices -> out of scope; future multi-vendor extension.
  • Real-time microphone dictation loop in the terminal -> use Voicebox app's built-in hotkey dictation.
  • Voice cloning sample upload and profile creation -> done in the Voicebox desktop app UI.
  • Video synthesis, music, sound design -> out of scope.
  • Stories Editor multi-voice timeline composition -> use the Voicebox app UI.
Expected inputs
  • TTS: text (<= 5000 chars per call), optional profile id, optional engine, optional language, optional output path.
  • STT: audio file path (absolute or relative to $CWD), optional language hint.
  • Notification: short message (<= 240 chars), profile id resolved from config.
Expected outputs
  • TTS: audio file (wav — Voicebox's native TTS output format; mp3 optional via a local ffmpeg transcode) at .agents/results/voice/{timestamp}-{shortid}/output.{wav|mp3} plus manifest.json.
  • STT: transcript.md at .agents/results/voice/transcripts/{timestamp}-{shortid}/ plus manifest.json.
  • Notification: ephemeral playback through Voicebox; no disk write by default.
Dependencies
  • Voicebox desktop app installed and running locally.
  • Voicebox MCP registered (claude mcp add --transport http voicebox http://127.0.0.1:17493/mcp).
  • TTS only: at least one voice profile created in the Voicebox app UI.
  • TTS only: optionally pre-downloaded engine models for the selected profile.
Control-flow features
  • Branches by mode (notify, asset, transcribe), language, and profile availability.
  • Calls voicebox via MCP tools, with REST GET /health as the handshake probe.
  • Reads input audio files and writes generated audio plus manifests.
  • Caches discovered MCP tool names after the first successful tools/list.

Structural Flow

Entry
  1. Detect the requested mode: notification, asset TTS, or transcription.
  2. Verify Voicebox is reachable via MCP handshake or GET /health.
  3. On the first run only, call MCP tools/list and cache the resolved tool names.
  4. For notification or asset TTS, resolve the target voice profile id. For transcription, validate the audio input and continue without a profile.
Scenes
  1. PREPARE: Validate text length, audio duration, language, output path, and profile id.
  2. ACQUIRE: If a required signal is missing, run the clarification protocol once.
  3. ACT: Invoke the appropriate MCP tool (TTS or STT) with the resolved parameters.
  4. VERIFY: Confirm the response carries audio output or transcript content. Validate manifest fields.
  5. FINALIZE: Write manifest.json alongside the output. Report the path or transcript to the user.
Transitions
  • If voicebox is unreachable, surface the install or launch hint and exit. Do not attempt auto-relaunch.
  • If a TTS request has no usable profile, point the user at the Voicebox app UI to create a profile, then exit. A transcription request never needs a profile.
  • If a TTS request exceeds 5000 chars, ask whether to truncate or split. Do not auto-chunk in v1.
  • If an STT input exceeds 30 minutes, ask whether to proceed. Do not auto-split.
  • If the selected engine model is not loaded, ask the user before triggering a download.
Failure and recovery
FailureRecovery
Voicebox app not runningPrint install/launch hint, exit code 5
No voice profile for TTSPrint "create a profile in Voicebox" hint, exit code 3
Engine model missingAsk before triggering download
Output path outside $PWDUse an explicitly requested path; ask only if the destination is ambiguous or overwrites unrelated data
TTS over 5000 charsAsk the user to split or truncate
STT over 30 minutesConfirm only if the requested duration or resource cost is unresolved
MCP tool name driftRe-run tools/list and update the cache
SIGINTAbort the MCP call, write no partial output
Exit
  • Success: audio file or transcript exists with a complete manifest, and the path is reported.
  • Partial success: output exists but a guardrail warning is surfaced (length, disk, model fallback).
  • Failure: no output, the blocker (auth, profile, engine, network) is explicit.

Logical Operations

Actions
ActionSSL primitiveEvidence
Validate mode and inputsVALIDATEClarification protocol in execution-protocol.md
Resolve TTS voice profileSELECTvoicebox_list_profiles + config defaults
Health checkREADMCP handshake or GET /health
Generate speechCALL_TOOLMCP voicebox_speak
Transcribe audioCALL_TOOLMCP voicebox_transcribe
Write output and manifestWRITEAudio or transcript plus manifest.json
Inspect resultVALIDATEOutput presence, duration, manifest fields
Report resultNOTIFYFinal user-facing summary
Tools and instruments
  • Voicebox MCP server at http://127.0.0.1:17493/mcp.
  • REST surface for health and audio retrieval (GET /health, GET /audio/{generation_id}).
  • Resource references: voice matrix, prompt tips, execution protocol, checklist.
Canonical command path
text
# 1. MCP handshake or REST health
GET http://127.0.0.1:17493/health  ->  200 OK

# 2. Discover tool names on first run
MCP tools/list                      ->  cache real names

# 3. TTS only: resolve profile
MCP voicebox_list_profiles          ->  pick profile by name or config default

# 4. Generate or transcribe (STT skips profile lookup and model-status check)
MCP voicebox_speak     { text, profile, language?, engine?, personality? }
MCP voicebox_transcribe { audio_path | audio_base64, language?, model? }

# 5. Fetch the generated audio (MCP has no save-to-disk; TTS output is wav)
GET http://127.0.0.1:17493/audio/{generation_id}  ->  wav bytes

# 6. Persist output + manifest
.agents/results/voice/<timestamp>-<shortid>/output.wav + manifest.json
.agents/results/voice/transcripts/<timestamp>-<shortid>/transcript.md + manifest.json
MCP tool mapping (verified against Voicebox 0.5.0)
Use caseMCP toolREST backing
TTS generationvoicebox_speakPOST /speak
STT transcriptionvoicebox_transcribePOST /transcribe
Profile listingvoicebox_list_profilesGET /profiles
Captures listingvoicebox_list_capturesGET /history (captures view)

Tools not exposed via MCP (REST only): model status (GET /models/status), audio file serving (GET /audio/{generation_id}), per-version audio (GET /audio/version/{version_id}). The skill calls those over loopback HTTP when needed.

Notes on voicebox_speak:

  • Required: text. Optional: profile, engine, language, personality (bool).
  • Audio plays on the user speakers and is saved to the Captures / History panel automatically. There is no save_to_disk toggle on the MCP tool itself; to persist a local copy, fetch GET /audio/{generation_id} (Voicebox TTS always stores wav).
  • Without a default profile set in Voicebox Settings, profile= is required.

Notes on voicebox_transcribe:

  • Accepts exactly one of audio_base64 or audio_path (loopback only). Optional language, model.
Show full SKILL.md (604 more words)Show less
Resource scope
ScopeResource target
LOCAL_FSInput audio, generated audio, transcripts, manifests
PROCESSLocal Voicebox app subprocess (managed by the user)
NETWORKLoopback HTTP to 127.0.0.1:17493 only
MEMORYCached MCP tool names, resolved profile metadata
CREDENTIALSNone. Voicebox is local and key-free.
Preconditions
  • Voicebox app is running and the MCP handshake succeeds.
  • TTS only: at least one voice profile exists.
  • TTS only: the selected engine model is loaded or the user approves a download.
  • Output directory is inside $PWD unless explicitly allowed.
Effects and side effects
  • Creates audio files, transcripts, and manifests under .agents/results/voice/.
  • Triggers local Voicebox generation, which consumes CPU or GPU.
  • May trigger an engine model download when the user approves.
  • Does not call any cloud service. No external network traffic.
Guardrails
  1. Voicebox required: if the MCP handshake or GET /health fails, exit with a one-shot install or launch hint. Do not retry, do not auto-relaunch.
  2. Profile required for TTS only: resolve a profile and check the TTS engine only for notification or asset mode. Transcription proceeds with a valid audio input and optional STT model even when no profile exists.
  3. Tool-name discovery: on first invocation, call MCP tools/list and cache the resolved names. Reuse the cache for subsequent calls in the same session.
  4. Length limits: TTS calls cap at 5000 chars per call; warn at 2000. STT inputs cap at 30 minutes. v1 does not auto-chunk or auto-split.
  5. Auto-invocation transparency: notifications fire automatically only when the active task exceeds auto_notify_after_sec (default 60s). This threshold is agent-enforced guidance — no hook measures task duration — so apply it by judgment when a long task completes or blocks. Always announce intent in one short line before generating audio.
  6. Path safety: an explicitly requested output path authorizes writing there. Resolve ambiguity or unrelated-data replacement before the dependent write; preserve required CLI path flags.
  7. Cancellation: SIGINT aborts the MCP call and writes no partial output.
  8. Manifest required for persisted output: asset TTS and transcription modes write manifest.json with at minimum: skill, mode, voicebox_generation_id, text (or transcript_preview), profile, engine, language, format (TTS only), created_at. Notification mode is exempt because Voicebox Captures is its system of record and no disk output is written by default.
  9. Out of scope: voice cloning UI, captures archive, stories editor, microphone dictation loop, and cloud vendors are intentionally not exposed.
  10. No cost guard: Voicebox is free. The cost guardrail from oma-image does not apply.
Clarification protocol

Before invoking a TTS or STT call, the agent checks the following. If any required signal is missing, clarify with the user first.

TTS (asset mode) required:

  • Text content provided?
  • Voice profile id or tone description provided?

TTS strongly recommended:

  • Language explicit or detectable from the text?
  • Output format (wav default; mp3 requires a local ffmpeg transcode)?

STT required:

  • Audio path provided and the file exists?
  • Duration within 30 minutes, or user approves splitting?

Notification mode skips clarification. It uses notification_profile from config and language is auto-detected from the message.

Invocation
Standalone
text
/oma-voice "build succeeded, 4 minor warnings"
/oma-voice transcribe ~/Downloads/standup.m4a
/oma-voice --profile prof_warm_korean "다음 단계 진행 준비됐어요"
Shared infrastructure (other skills)

Other skills can request audio output by calling the same MCP tools directly, or by invoking /oma-voice with their text. There is no separate CLI; the skill is MCP-native.

References

  • Voice engine matrix: resources/voice-matrix.md
  • Prompt writing rules: resources/prompt-tips.md
  • Execution protocol: resources/execution-protocol.md
  • Pre-flight checklist: resources/checklist.md
  • Configuration: read the voice: section of .agents/oma-config.yaml first, then fall back to config/voice-config.yaml for any key it does not set. Both profiles ship as null and must be set per machine — write them to .agents/oma-config.yaml, since oma update overwrites the skill config.
  • Context loading: references/_shared/core/context-loading.md
  • Quality principles: references/_shared/core/quality-principles.md
  • Design reference: ../../../docs/plans/designs/012-oma-voice.md (source repo only; absent in global-mode installs)

© first-fluke, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 27 other files (references) in skills/oma-voice of first-fluke/oh-my-agent.

  • SKILL.md
  • config/voice-config.yaml
  • references/_shared/conditional/experiment-ledger.md
  • references/_shared/conditional/exploration-loop.md
  • references/_shared/conditional/quality-score.md
  • references/_shared/core/api-contracts/README.md
  • references/_shared/core/api-contracts/template.md
  • references/_shared/core/clarification-protocol.md
  • references/_shared/core/common-checklist.md
  • references/_shared/core/context-budget.md
  • references/_shared/core/context-loading.md
  • references/_shared/core/difficulty-guide.md
  • references/_shared/core/execution-policy.md
  • references/_shared/core/lessons-learned.md
  • references/_shared/core/prompt-structure.md
  • … and 13 more

Open the folder on GitHubat commit 268bb4a

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders. This page covers the copy in first-fluke/oh-my-agent, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Oma Voice next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Oma Voice compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Oma Voice this skillfirst-fluke/oh-my-agent1.3k—~3.3kAutomated safety check: PassMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0
Edu Chem Videowy51ai/edulab1.4k—~2.1kAutomated safety check: NotesApache-2.0
Edu Math Videowy51ai/edulab1.4k—~2.5kAutomated safety check: NotesApache-2.0
Edu Physics Videowy51ai/edulab1.4k—~2.3kAutomated safety check: NotesApache-2.0
Elevenlabs Transcribeqdhenry/Claude-Command-Suite1.3k—~1.5kAutomated safety check: NotesNone

Similar skills

  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Edu Chem Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a chemistry problem (化学题: 氧化还原配平 双线桥 电子守恒, 物质的量计算, 化学平衡 三段式 平衡常数 转化率 反应速率, 离子反应, 电化学, 溶液 滴定…

    1.4k GitHub stars~2.1k tokensUpdated today
    Media & CreativeAuto-check: notes
  • Edu Math Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a math problem (数学题, geometry, algebra, functions, motion/行程 problems), from a problem screenshot…

    1.4k GitHub stars~2.5k tokensUpdated today
    Media & CreativeAuto-check: notes
  • Edu Physics Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a physics problem (物理题: mechanics/力学 受力分析 牛顿定律 斜面 传送带 板块 平抛 圆周 能量 动量, optics/光学 折射…

    1.4k GitHub stars~2.3k tokensUpdated today
    Media & CreativeAuto-check: notes
  • Elevenlabs Transcribe

    qdhenry/Claude-Command-Suite

    Transcribes audio/video files using ElevenLabs Scribe v2 API.

    1.3k GitHub stars~1.5k tokensUpdated 7 mo ago
    Media & CreativeAuto-check: notes
  • Video Assemble

    zenstory-ai/video-recap-skills

    合成视频解说最终成片:把旁白音频铺到源视频上,按旁白窗口压低原声,生成 SRT / ASS 字幕并可烧录, 最后做响度标准化。作为最终合成阶段使用。输入源视频、ttsmeta.json 与旁白位置; 输出 recap 成片和字幕。触发词:视频合成、混音、字幕、压字幕、assemble video、mux、ducking、subtitles、成片。

    561 GitHub stars~1.7k tokensUpdated 6 days ago
    Media & CreativeAuto-check passed

More from first-fluke/oh-my-agent

All 57 skills in this repo
  • OMA Multi-Agent Orchestration

    first-fluke/oh-my-agent

    Decomposes a complex feature into tasks, dispatches parallel specialist agents with durable state, and supervises verification, QA review and retries.

    1.3k GitHub stars~4.1k tokensUpdated today
    Auto-check passed
  • OMA Multi-Agent Orchestrator

    first-fluke/oh-my-agent

    Splits a complex feature into prioritized tasks, spawns specialist CLI subagents in parallel, tracks them through shared memory and verifies each result.

    1.3k GitHub stars~3.1k tokensUpdated today
    Auto-check passed
  • Architecture Decisions and ADRs

    first-fluke/oh-my-agent

    Evaluates system boundaries and tradeoffs and writes architecture recommendations, option comparisons or ADRs, with a Mermaid diagram when structure changes.

    1.3k GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • OMA Brainstorm

    first-fluke/oh-my-agent

    Explores goals, constraints and alternative designs one question at a time and saves an approved design document before any planning or coding starts.

    1.3k GitHub stars~2.8k tokensUpdated today
    Auto-check passed
  • Oma Coordination

    first-fluke/oh-my-agent

    Coordinate assigned specialist tasks and handoffs manually. An agent skill from first-fluke/oh-my-agent.

    1.3k GitHub stars~1.9k tokensUpdated today
    Auto-check passed
  • Oma Image

    first-fluke/oh-my-agent

    Generate raster images or reference-guided variations through the OMA image CLI.

    1.3k GitHub stars~2k tokensUpdated today
    Auto-check passed

Questions about Oma Voice

What does Oma Voice do?

Generate speech or transcribe audio locally with Voicebox. An agent skill from first-fluke/oh-my-agent. Oma Voice is an agent skill from first-fluke/oh-my-agent. Generate speech or transcribe audio locally with Voicebox.

When should I use Oma Voice?

Oma Voice fits situations like: meeting transcription; tasks that involve Transcription; tasks that involve Text to speech and voice.

How do I install Oma Voice in Claude Code?

Run `npx skills add first-fluke/oh-my-agent --skill oma-voice -a claude-code`. Or copy the skill folder (skills/oma-voice in first-fluke/oh-my-agent) into .claude/skills/oma-voice in your project. Claude Code loads it when a task matches its description.

How do I install Oma Voice in Codex?

Run `npx skills add first-fluke/oh-my-agent --skill oma-voice -a codex`. Or copy the skill folder (skills/oma-voice in first-fluke/oh-my-agent) into .agents/skills/oma-voice in your project. Codex loads it when a task matches its description.

Can I use Oma Voice in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add first-fluke/oh-my-agent --skill oma-voice -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/oma-voice, .gemini/skills/oma-voice, .github/skills/oma-voice and .opencode/skills/oma-voice in your project.

What does Oma Voice need to run?

Going by SKILL.md and its folder, Oma Voice needs the command-line tools its instructions call (claude).

Does Oma Voice access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Oma Voice safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Oma Voice use?

Oma Voice is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Oma Voice use?

About 3.3k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 15k tokens, read only when the agent opens those files.

What are the alternatives to Oma Voice?

Skills that share tags, products or a category with Oma Voice: HyperFrames Media Use (heygen-com/hyperframes, 60k stars), Edu Chem Video (wy51ai/edulab, 1.4k stars), Edu Math Video (wy51ai/edulab, 1.4k stars) and Edu Physics Video (wy51ai/edulab, 1.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Oma Voice?

first-fluke (a GitHub organization) maintains it in first-fluke/oh-my-agent, which has 1,336 GitHub stars. The repository holds 57 skills in this directory. The repository was last updated on October 10, 2026.

Source: first-fluke/oh-my-agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.