Hyperframes Media
chmonitor/chmonitor
Audio and media assets for HyperFrames compositions, produced by one shared audio engine (scripts/audio.mjs) — multi-provider TTS (HeyGen / ElevenLabs / Kokoro local), background music + sound…
Convert local audio/video files or public media URLs into subtitle files by uploading local files to Cloudflare R2 and calling Volcengine AI MediaKit ASR subtitles API.
$ npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install bozhouDev/video-skills-toolkit audio-to-subtitles --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/audio-to-subtitles .claude/skills/audio-to-subtitles && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "audio-to-subtitles" agent skill from https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitles into .claude/skills/audio-to-subtitles/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-to-subtitles", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitlesType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install bozhouDev/video-skills-toolkit audio-to-subtitles --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/audio-to-subtitles .agents/skills/audio-to-subtitles && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "audio-to-subtitles" agent skill from https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitles into .agents/skills/audio-to-subtitles/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-to-subtitles", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install bozhouDev/video-skills-toolkit audio-to-subtitles --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/audio-to-subtitles .cursor/skills/audio-to-subtitles && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "audio-to-subtitles" agent skill from https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitles into .cursor/skills/audio-to-subtitles/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-to-subtitles", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/bozhouDev/video-skills-toolkit.git --path skills/audio-to-subtitles--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install bozhouDev/video-skills-toolkit audio-to-subtitles --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/audio-to-subtitles .gemini/skills/audio-to-subtitles && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "audio-to-subtitles" agent skill from https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitles into .gemini/skills/audio-to-subtitles/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-to-subtitles", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install bozhouDev/video-skills-toolkit audio-to-subtitlesInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/audio-to-subtitles .github/skills/audio-to-subtitles && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "audio-to-subtitles" agent skill from https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitles into .github/skills/audio-to-subtitles/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-to-subtitles", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install bozhouDev/video-skills-toolkit audio-to-subtitles --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bozhouDev/video-skills-toolkit.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/audio-to-subtitles .opencode/skills/audio-to-subtitles && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "audio-to-subtitles" agent skill from https://github.com/bozhouDev/video-skills-toolkit/tree/main/skills/audio-to-subtitles into .opencode/skills/audio-to-subtitles/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-to-subtitles", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
audio-to-subtitlesConvert local audio/video files or public media URLs into subtitle files by uploading local files to Cloudflare R2 and calling Volcengine AI MediaKit ASR subtitles API.
Audio To Subtitles is an agent skill from bozhouDev/video-skills-toolkit. Convert local audio/video files or public media URLs into subtitle files by uploading local files to Cloudflare R2 and calling Volcengine AI MediaKit ASR subtitles API. Also supports MiniMax text-to-speech generation from text or text files, with optional subtitle generation from the produced audio. Supports batch processing, SRT/VTT/JSON output, language selection, speaker labels, configurable MiniMax voice ID and model. Use when the user asks to turn audio/video into subtitles, generate ASR subtitles, create…
Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including scripts (for example `scripts/main.ts` and `scripts/r2.ts`).
It sits in Media & Creative, covering Transcription and Text to speech and voice. It works with MiniMax, Cloudflare R2 and HeyGen. The repository describes itself as: Video skills toolkit for Remotion talking-head, sketch story, and audio-to-subtitles workflows. The licence is MIT.
4 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 4766a16. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 2 files in scripts/ (TypeScript), which the agent can run.
Shell commands in SKILL.md call:
npxFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
api.minimax.ioFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
MINIMAX_API_KEYMINIMAX_TTS_API_KEYMEDIAKIT_API_KEYAI_MEDIAKIT_API_KEYVOLCENGINE_MEDIAKIT_API_KEYR2_ACCESS_KEY_IDR2_SECRET_ACCESS_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Audio To Subtitles loads about 1.9k tokens when it runs. Until then it costs about 153 tokens; SKILL.md has 657 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
onment files in this order. The nearest `.env.r2` is loaded last and overrides global R2 defaults for this vault.1. `~/.skills/.env`2. `~/.baoyu-skills/.env`3. nearest `.env.r2` from current directory upward4. nearest `.skills/.env` from current directory upward5. nearest `.baoyu-skills/.env` from current directory upward- Do not print API keys or commit real `.env.r2` values.Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from bozhouDev/video-skills-toolkit at commit 4766a16, republished under its MIT licence (© bozhouDev). 657 words, ~1,892 tokens.
.claude/skills/audio-to-subtitles/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.Use this skill to:
For media inputs:
audio_url or video_url to AI MediaKit ASR subtitles..json, .srt, and .vtt outputs.For local video-engine projects:
talking-head-hyperframes project under <VIDEO_WORKSPACE>, save subtitles under that project's work/captions/.captions.srt and captions.vtt as the main cleaned files, keep raw ASR backups as captions.raw.srt and captions.raw.vtt, and keep the MediaKit payload as asr-result.json.captions_aligned.json parsed from the cleaned SRT so video-script, talking-head-hyperframes, and hyperframes-scene-animator share the same locked timeline.<VIDEO_WORKSPACE>/<project>/work/captionsFor text inputs:
--subtitles is set, upload the audio to R2 and run the same AI MediaKit subtitle workflow.# Set this to the directory containing this installed skill (not the current working directory).
SKILL_DIR="<SKILL_ROOT>/audio-to-subtitles"
# Local audio/video: upload to R2, then transcribe
npx -y bun "$SKILL_DIR/scripts/main.ts" audio.mp3 --language zh-CN --out-dir subtitles
# Existing public URL: skip R2 upload
npx -y bun "$SKILL_DIR/scripts/main.ts" "https://example.com/audio.mp3" --out-dir subtitles
# Batch files/URLs
npx -y bun "$SKILL_DIR/scripts/main.ts" a.mp3 b.m4a "https://example.com/video.mp4" --out-dir subtitles
# Batch from manifest
npx -y bun "$SKILL_DIR/scripts/main.ts" --manifest inputs.txt --out-dir subtitles
# Speaker diarization
npx -y bun "$SKILL_DIR/scripts/main.ts" interview.mp3 --speaker --language zh-CN
# Text to audio only
npx -y bun "$SKILL_DIR/scripts/main.ts" --text "你好,欢迎收听。" --voice-id female-shaonv --out-dir audio
# Text file to audio + subtitles
npx -y bun "$SKILL_DIR/scripts/main.ts" --text-file script.md --subtitles --language zh-CN --out-dir audio| Option | Description |
|---|---|
--text <text> | Generate MiniMax TTS audio from inline text. Can be repeated. |
--text-file <path> | Generate MiniMax TTS audio from a text file. Can be repeated. |
--subtitles | With text input, also generate subtitles from the generated audio. |
--manifest <path> | Batch input manifest. Text files use one input per line; JSON supports an array of strings or objects. |
--out-dir <path> | Output directory. Default: subtitles. |
| `--format <all | srt |
| `--language <cmn-Hans-CN | zh-CN |
--speaker | Enable speaker info and prefix subtitles with speaker labels when returned. |
--voice-id <id> | MiniMax voice ID. Default: MINIMAX_VOICE_ID, then MINIMAX_TTS_VOICE_ID, then female-shaonv. |
--tts-model <model> | MiniMax TTS model. Default: MINIMAX_TTS_MODEL, then speech-02-hd. |
--tts-speed <n> | TTS speed. Default: MINIMAX_TTS_SPEED, then 1.0. |
--tts-vol <n> | TTS volume. Default: MINIMAX_TTS_VOL, then 1.0. |
--tts-pitch <n> | TTS pitch. Default: MINIMAX_TTS_PITCH, then 0. |
--tts-emotion <emotion> | TTS emotion. Default: MINIMAX_TTS_EMOTION, then happy. |
| `--tts-format <mp3 | wav |
--tts-sample-rate <n> | TTS sample rate. Default: MINIMAX_TTS_SAMPLE_RATE, then 32000. |
--tts-bitrate <n> | TTS bitrate. Default: MINIMAX_TTS_BITRATE, then 128000. |
| `--media-kind <auto | audio |
--r2-prefix <prefix> | R2 object key prefix for uploaded local audio/video. Default: R2_AUDIO_KEY_PREFIX, then audio/YYYY-MM-DD. |
--concurrency <n> | Batch concurrency. Default: 1. |
--poll-interval <seconds> | Poll interval. Default: 5. |
--timeout <seconds> | Per-task timeout. Default: 7200. |
--json | Print machine-readable run summary to stdout. |
Text manifests (.txt) are one media input per line.
JSON manifests can mix media and text jobs:
[
"local-audio.mp3",
{ "url": "https://example.com/video.mp4", "mediaKind": "video" },
{ "text": "你好,欢迎收听。", "outputName": "intro", "subtitles": true, "voiceId": "female-shaonv" },
{ "textFile": "script.md", "outputName": "script-audio", "subtitles": true }
]The script loads environment files in this order. The nearest .env.r2 is loaded last and overrides global R2 defaults for this vault.
~/.skills/.env~/.baoyu-skills/.env.env.r2 from current directory upward.skills/.env from current directory upward.baoyu-skills/.env from current directory upwardRequired for MiniMax TTS text input:
| Variable | Description |
|---|---|
MINIMAX_API_KEY | MiniMax API key. |
MINIMAX_TTS_API_KEY | Also accepted. |
MINIMAX_API_HOST | Optional. Default: https://api.minimax.io. |
MINIMAX_VOICE_ID | Optional default voice ID. |
MINIMAX_TTS_MODEL | Optional default model. |
Required for MediaKit subtitle generation:
| Variable | Description |
|---|---|
MEDIAKIT_API_KEY | AI MediaKit API key. |
AI_MEDIAKIT_API_KEY | Also accepted. |
VOLCENGINE_MEDIAKIT_API_KEY | Also accepted. |
Required for local-file uploads:
| Variable | Description |
|---|---|
R2_ACCESS_KEY_ID | Cloudflare R2 access key. |
R2_SECRET_ACCESS_KEY | Cloudflare R2 secret key. |
R2_ACCOUNT_ID | Cloudflare account ID. |
R2_BUCKET | R2 bucket name. |
R2_PUBLIC_BASE_URL | Public base URL. R2_PUBLIC_URL is also accepted. |
R2_AUDIO_KEY_PREFIX | Optional R2 prefix for uploaded audio/video. Default: audio/YYYY-MM-DD. |
.env.r2 values.<VIDEO_WORKSPACE>/<project>/work/captions/ directory.© bozhouDev, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 2 other files (scripts) in skills/audio-to-subtitles of bozhouDev/video-skills-toolkit.
Open the folder on GitHubat commit 4766a16
Audio To Subtitles next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Audio To Subtitles this skillbozhouDev/video-skills-toolkit | 150 | — | ~1.9k | Automated safety check: Notes | MIT | |
| Hyperframes Mediachmonitor/chmonitor | 301 | 1 repos | ~2.8k | Automated safety check: Notes | GPL-3.0 | |
| Hyperframes CLInateherkai/hyperframes-student-kit | 1.3k | 3 repos | ~1.2k | Automated safety check: Pass | Custom licence | |
| Content To Videoarchitectds/modeldock | 117 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | |
| Vox ExplainerCK42BB/vox-explainer-skill | 109 | — | ~2.6k | Automated safety check: Pass | MIT | |
| Hyperframes CLIcoleam00/hyperframes-ai-video-generation | 149 | — | ~1.6k | Automated safety check: Pass | None |
chmonitor/chmonitor
Audio and media assets for HyperFrames compositions, produced by one shared audio engine (scripts/audio.mjs) — multi-provider TTS (HeyGen / ElevenLabs / Kokoro local), background music + sound…
nateherkai/hyperframes-student-kit
HyperFrames CLI tool — hyperframes init, lint, preview, render, transcribe, tts, doctor, browser, info, upgrade, compositions, docs, benchmark.
architectds/modeldock
Turn arbitrary source content (README, article, story, slides, deck, data/report, product description, tutorial text, audio/transcript, or a bare topic) into a finished, high-quality MP4 video.
CK42BB/vox-explainer-skill
End-to-end pipeline for producing Vox-style explainer videos from a single topic prompt.
coleam00/hyperframes-ai-video-generation
HyperFrames CLI tool — hyperframes init, lint, inspect, preview, render, transcribe, tts, doctor, browser, info, upgrade, compositions, docs, benchmark.
cacity/VideoHub
把长视频或已有字幕转成有完整叙事的几分钟短片。先基于原文字幕和画面证据理解、选段与重排,再对最终时间轴重新翻译和可选润色;既可输出保留原声的双语字幕版,也可把原声降到 30% 并用 MiniMax 或豆包 TTS 生成影视解说、短剧混剪、播客串讲或知识解读版。已有项目可进入本地五轨时间线继续调整切点、旁白、原声窗口、字幕、音量和转场,并按修订版本渲染。用于“把长视频讲成短故事”“按字幕自动剪辑”…
bozhouDev/video-skills-toolkit
Convert audio/video URLs or local media into corrected Markdown transcripts through Volcengine recording-file ASR 2.0.
bozhouDev/video-skills-toolkit
为 HyperFrames 口播或旁白项目创建、修复并验证固定舞台,锁定数字人 PIP 的区域、裁切、人物安全区和不透明背景,归档输入,生成 manifest 与 template handoff,并在就绪后按“字幕驱动的全镜头静态审核→动效”门禁路由到 hyperframes-scene-animator。适用于“新建 HyperFrames…
bozhouDev/video-skills-toolkit
Generate music using ElevenLabs Music API. An agent skill from bozhouDev/video-skills-toolkit.
bozhouDev/video-skills-toolkit
判断、扫描、拆解并归档抖音视频、小红书图文或小红书视频。实时读取用户同平台粉丝数并划分主对标池/跨级灵感池,用已登录浏览器读取目标作品和作者主页公开指标,再用确定性代码判定普通、小爆、爆款、现象级并扫描作者近 20 条候选;只对用户选中的爆款和现象级先构建可追溯证据包,再调用子 Agent…
bozhouDev/video-skills-toolkit
生成抖音、视频号、小红书等短视频封面图、视频标题图和合集封面,也能诊断和改版已有封面。用户说做封面、生成封面、抖音封面、视频封面、标题图、合集封面、3:4、4:3、1:1、短视频首图、动态封面首帧、给这期视频做图、这封面为什么没人点、帮我改封面、封面点击率怎么提升、诊断封面时都应使用。小白学AI系列封面除外:遇到“小白学AI封面/小白学AI第N集封面”时优先使用…
bozhouDev/video-skills-toolkit
用 MiniMax 云端为视频制作可审批的声音导演稿,再生成、挑选和验收人声,最后以定稿音频产生字幕。用于用户明确选择 MiniMax 配音、继续已有 MiniMax 视频配音项目,或明确请求 MiniMax Voice ID/克隆/设计。泛指本地 TTS 或 IndexTTS 不使用本 skill;音乐、BGM、歌曲使用同级 music Skill。
Works with
Categories
Convert local audio/video files or public media URLs into subtitle files by uploading local files to Cloudflare R2 and calling Volcengine AI MediaKit ASR subtitles API. Audio To Subtitles is an agent skill from bozhouDev/video-skills-toolkit. Convert local audio/video files or public media URLs into subtitle files by uploading local files to Cloudflare R2 and calling Volcengine AI MediaKit ASR subtitles API.
Audio To Subtitles fits situations like: the user asks to turn audio/video into subtitles; generate ASR subtitles; batch transcription; generate TTS audio with optional subtitles.
Run `npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a claude-code`. Or copy the skill folder (skills/audio-to-subtitles in bozhouDev/video-skills-toolkit) into .claude/skills/audio-to-subtitles in your project. Claude Code loads it when a task matches its description.
Run `npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a codex`. Or copy the skill folder (skills/audio-to-subtitles in bozhouDev/video-skills-toolkit) into .agents/skills/audio-to-subtitles in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add bozhouDev/video-skills-toolkit --skill audio-to-subtitles -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audio-to-subtitles, .gemini/skills/audio-to-subtitles, .github/skills/audio-to-subtitles and .opencode/skills/audio-to-subtitles in your project.
Going by SKILL.md and its folder, Audio To Subtitles needs TypeScript for the scripts in its folder, the command-line tools its instructions call (npx) and credentials named MINIMAX_API_KEY, MINIMAX_TTS_API_KEY, MEDIAKIT_API_KEY and AI_MEDIAKIT_API_KEY. Our summary lists: Node.js; A credential in MINIMAX_API_KEY; A credential in MINIMAX_TTS_API_KEY.
SKILL.md names 1 domain. In commands or code: api.minimax.io; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Audio To Subtitles is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.9k tokens (SKILL.md is roughly 7.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Audio To Subtitles: Hyperframes Media (chmonitor/chmonitor, 301 stars), Hyperframes CLI (nateherkai/hyperframes-student-kit, 1.3k stars), Content To Video (architectds/modeldock, 117 stars) and Vox Explainer (CK42BB/vox-explainer-skill, 109 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
bozhouDev (a GitHub user) maintains it in bozhouDev/video-skills-toolkit, which has 150 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on July 27, 2026.
Source: bozhouDev/video-skills-toolkit on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.