Video Downloader
kangarooking/kangarooking-skills
Download or open videos and recover platform captions, audio transcripts, keyframes, screen text, visual facts, and editing observations as a plain multimodaltranscript.md.
A skill your agent uses when the user gives a topic and wants an automated topic-driven narrated explainer, podcast, or knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat…
$ npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Agents365-ai/video-podcast-maker video-podcast-maker --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Agents365-ai/video-podcast-maker.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/video-podcast-maker .claude/skills/video-podcast-maker && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "video-podcast-maker" agent skill from https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-maker into .claude/skills/video-podcast-maker/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-podcast-maker", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-makerType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Agents365-ai/video-podcast-maker video-podcast-maker --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Agents365-ai/video-podcast-maker.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/video-podcast-maker .agents/skills/video-podcast-maker && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "video-podcast-maker" agent skill from https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-maker into .agents/skills/video-podcast-maker/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-podcast-maker", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Agents365-ai/video-podcast-maker video-podcast-maker --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Agents365-ai/video-podcast-maker.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/video-podcast-maker .cursor/skills/video-podcast-maker && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "video-podcast-maker" agent skill from https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-maker into .cursor/skills/video-podcast-maker/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-podcast-maker", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Agents365-ai/video-podcast-maker.git --path skills/video-podcast-maker--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Agents365-ai/video-podcast-maker video-podcast-maker --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Agents365-ai/video-podcast-maker.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/video-podcast-maker .gemini/skills/video-podcast-maker && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "video-podcast-maker" agent skill from https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-maker into .gemini/skills/video-podcast-maker/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-podcast-maker", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Agents365-ai/video-podcast-maker video-podcast-makerInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Agents365-ai/video-podcast-maker.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/video-podcast-maker .github/skills/video-podcast-maker && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "video-podcast-maker" agent skill from https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-maker into .github/skills/video-podcast-maker/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-podcast-maker", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Agents365-ai/video-podcast-maker video-podcast-maker --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Agents365-ai/video-podcast-maker.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/video-podcast-maker .opencode/skills/video-podcast-maker && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "video-podcast-maker" agent skill from https://github.com/Agents365-ai/video-podcast-maker/tree/main/skills/video-podcast-maker into .opencode/skills/video-podcast-maker/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-podcast-maker", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
video-podcast-makerA skill your agent uses when the user gives a topic and wants an automated topic-driven narrated explainer, podcast, or knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat…
Video Podcast Maker is an agent skill from Agents365-ai/video-podcast-maker. Use when the user gives a topic and wants an automated topic-driven narrated explainer, podcast, or knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat Channels), or asks to learn visual design patterns from a reference video/image. Trigger when the user mentions creating a knowledge video, narrated explainer, video podcast, or animated infographic-style video from a topic — even if they don't say "video podcast" explicitly. Also trigger when the user wants to regenerate, re-render…
Its SKILL.md is about 4.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 110 other files, including scripts, reference files and assets (for example `.skillspector-baseline.yaml`, `AGENTS.md` and `package.json`).
It sits in Media & Creative, covering Video production and Text to speech and voice. It works with Remotion, Bilibili, Douyin and WeChat. The repository describes itself as: Topic → 4K narrated video for coding agents. v5.3.0: local TTS (edge free + azure, no external engine), manifest-based Asset Engine, Remotion composition, cost-gated AI…. The licence is MIT.
4 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 33b8078. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/, which the agent can run.
Shell commands in SKILL.md call:
npxpython3gitnpmFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
github.comremotion.devFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
AZURE_SPEECH_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Video Podcast Maker loads about 4.9k tokens when it runs, and up to ~43k if it reads all its reference files. Until then it costs about 248 tokens; SKILL.md has 1,610 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from Agents365-ai/video-podcast-maker at commit 33b8078, republished under its MIT licence (© Agents365-ai). 1,610 words, ~4,870 tokens.
.claude/skills/video-podcast-maker/SKILL.md (or your agent's skills folder). This skill also uses 108 other files; get the full folder from GitHub.Recommended: Load Remotion Best Practices
This skill benefits from
remotion-best-practices(not bundled) for the full Remotion pattern library. It is optional — minimum rules are below if absent.
- Pi: read the loaded skill at
remotion-best-practices(listed in available skills).- Claude Code: invoke
remotion-best-practicesskill/tool before proceeding.Not installed? Get it from remotion-dev/skills (docs: remotion.dev/docs/ai/skills).
If
remotion-best-practicesis not installed, minimum rules: chromium must be available, always wrap 4K content in<Scale4K>, use<TransitionSeries>withlinearTiming, and treat audio as the master clock.
Automated pipeline for 4K Bilibili horizontal knowledge videos from a topic. Coding agent + TTS backend + Remotion + FFmpeg.
references/ fileResolve SKILL_DIR to the directory containing this SKILL.md:
SKILL_DIR to that directory before running commands.${CLAUDE_SKILL_DIR} is auto-populated.SKILL_DIR="${SKILL_DIR:-${CLAUDE_SKILL_DIR}}"
# Prerequisites (CLIs + backend env vars)
python3 "${SKILL_DIR}/scripts/check_prereqs.py"Updates flow through the skills CLI (npx skills update video-podcast-maker -g); direct git-clone installs use git pull per the README. This skill performs no update checks.
Prereqs failures — see README.md for setup. The check is backend-aware (resolves TTS_BACKEND env → user_prefs.json global.tts.backend → edge default), so only env vars required by the active backend are validated.
First video in a new project? Prefer reusing an existing Remotion project with node_modules/ already installed — creating a fresh project downloads ~2.2 GB of npm packages plus a 90 MB Chrome headless shell (one-time per project). If the user has a project from a previous video, use it. If a fresh project is necessary, run npm install in the background while you do Steps 1-4 (topic research and script writing).
All rendering goes into videos/{name}/ — every output.mp4, final_video.mp4, and thumbnail_*.png lands directly in the per-video directory. Never render to an out/ or dist/ directory; the --public-dir videos/{name}/ convention keeps everything self-contained.
TTS engine — two local backends, no external skill:
edge (default) — free, no key, via edge-tts.azure — needs AZURE_SPEECH_KEY + AZURE_SPEECH_REGION (Microsoft Speech SDK).Each synthesizes in-house (scripts/tts/backends/native.py) — pronunciation
(display → spoken → back to display for subtitles) and phoneme application are
built in. check_prereqs.py validates the active backend's env vars.
Multi-platform TTS? The former ttscn component skill that provided the 9-backend matrix is no longer a dependency of this skill. If you want those platforms, install Agents365-ai/ttsCN separately and call it directly — this skill ships only edge + azure.
Design Learning shortcut: If the user provides a reference video/image or asks to save/list/delete style profiles, see references/design-learning.md instead of running the workflow below.
Detect Auto Mode (default) vs Interactive Mode at workflow start — the Auto-default decision table and per-request overrides are in references/workflow-script.md.
If videos/{name}/ already exists and the user is iterating on a finished or in-progress video, reuse that directory. Do NOT start a new project or a new videos/{newname}/.
Pick the smallest re-run for what actually changed:
| Changed | Re-run | Reuses (don't redo) |
|---|---|---|
Narration script (podcast.txt) | Step 7 (TTS) → Step 8 preview → render+mix | topic research + section design |
| Visuals only (components, layout, colors) | Step 8 preview → render+mix | audio (podcast_audio.wav / timing.json) |
| Background music only | Re-mix BGM | output.mp4 (no re-render) |
| Subtitles only | Step 10.1 finalize | output.mp4 / video_with_bgm.mp4 |
Any re-run that changes what the viewer sees or hears re-enters the Step 8 gate: apply the change, let Studio hot-reload, and wait for a fresh explicit "render 4K" — the previous confirmation does not carry over. A script change shifts every downstream timestamp, so always regenerate timing.json through TTS — never hand-edit it. After any re-run, re-verify:
python3 ${SKILL_DIR}/scripts/verify_output.py videos/{name}/Iterating on a finished video? If
videos/{name}/already exists, see Regenerating an Existing Video above for the minimal re-run — do NOT start at Step 1.
At Step 1 start, create one task per step in your agent's tracker. Mark in_progress on start, completed on finish. Files in videos/{name}/ are the durable record — if interrupted, inspect the directory to determine where to resume.
| # | Step | Output | Phase file |
|---|---|---|---|
| 1 | Define topic direction | topic_definition.md | workflow-script.md |
| 2 | Research topic | topic_research.md | workflow-script.md |
| 3 | Design 5-7 sections | (in-memory) | workflow-script.md |
| 4 | Write narration script | podcast.txt | workflow-script.md |
| 4.5 | Pronunciation pre-flight (zh-CN) | phonemes.json | workflow-script.md |
| 5 | Asset plan & resolve | assets/manifest.json | workflow-assets.md |
| 6 | Generate thumbnails (16:9 + 4:3) | thumbnail_*.png | workflow-production.md |
| 7 | Generate TTS audio | podcast_audio.wav, timing.json | workflow-production.md |
| 8 | Remotion composition + Studio preview | — | workflow-production.md |
| 9 | Render 4K + mix BGM | output.mp4, video_with_bgm.mp4 | workflow-production.md |
| 10 | Publish info + verify output | publish_info.md, final_video.mp4 | workflow-publish.md |
| 11 | Generate vertical shorts (optional) | shorts/ | workflow-publish.md |
Mandatory stops (bold rows above):
npx remotion studio and wait for user feedback before rendering. NEVER render 4K until the user explicitly confirms ("render 4K" / "render final"). A reply containing adjustment requests is not confirmation — apply the changes, let Studio hot-reload, and ask again. Every round of adjustments needs its own fresh confirmation before Step 9.verify_output.py. MUST pass before declaring the video done. Exit 0 = green; exit 2 = warnings still publishable. Auto-fixes common omissions (creates final_video.mp4 if missing). Validates publish info (title, description, tags, chapters) against the platform matrix — generate it in Steps 5.5 and 10.2. For machine-readable output add --format json.Pre-render audit (recommended) — before Step 8:
python3 ${SKILL_DIR}/scripts/audit_beat_sync.py <Video.tsx> <timing.json>Flags beats that drift > 1.5s from narration.
Auto Mode: visual self-review. When running in Auto Mode (no user watching Studio), render 3-5 key frame stills before asking for render confirmation:
npx remotion still src/remotion/index.ts <CompositionId> videos/{name}/_review_001.png --public-dir videos/{name}/ --frame=<midpoint_frame>Pick frames at: hero title (~10% in), a dense section midpoint, and the outro. Read the stills back as images and run the design-guide.md and visual-taste.md checklists against actual rendered output. Catch overflow, contrast, and layout regressions before the 4K render. Delete _review_*.png after review.
| After Step | Check |
|---|---|
| 7 (TTS) | podcast_audio.wav plays · timing.json covers all sections · SRT is UTF-8 |
| 9 (Render) | output.mp4 is 3840×2160 · audio-video sync · no black frames |
| 10 (Verify) | verify_output.py exits 0 (or 2 with reviewed warnings) |
| Rule | Requirement |
|---|---|
| Single Project | All videos under videos/{name}/ in user's Remotion project. NEVER create a new project per video. |
| 4K Output | 3840×2160 (or 2160×3840 vertical), use scale(2) wrapper over 1920×1080 design space |
| Audio Sync | Audio (podcast_audio.wav + podcast_audio.srt) is the master clock. timing.json MUST be generated from the real TTS output, never hand-estimated. Before rendering, final video duration must match audio within ±0.5s. See Audio-Master Clock & Sync. |
| Thumbnail | MUST generate both 16:9 (1920×1080) AND 4:3 (1200×900) — see design-guide.md |
| Studio Before Render | MUST launch remotion studio for review. NEVER render 4K until user explicitly confirms. Adjustment feedback ≠ confirmation — apply, hot-reload, ask again. |
--public-dir | Every Remotion command uses --public-dir videos/{name}/. All output files (output.mp4, final_video.mp4, thumbnails) go directly into videos/{name}/ — never an out/ or dist/ dir. |
Visual minimums (text sizes, content width, safe zones, animation safety) live in references/design-guide.md. MUST load before Step 8.
podcast_audio.wav and podcast_audio.srt.podcast.txt → generate_tts.py → podcast_audio.wav + podcast_audio.srt + timing.json → composition → render.timing.json before audio exists. If you already have curated slides, run align_timing_from_srt.py to anchor them to the real SRT.TransitionSeries renders sum(section.duration_frames) - (N-1) * transitionFrames frames. Scale every section proportionally to keep the rendered length equal to timing.total_frames. Do not stuff all overlap frames into the first section. The corrected pattern is in templates/Video.tsx.| When | Check |
|---|---|
| After Step 7 (TTS) | timing.json.total_duration matches podcast_audio.wav within ±0.5s |
| Before render | Video.tsx scales all sections for transition overlap |
| After render | final_video.mp4 duration matches podcast_audio.wav within ±0.5s |
| Step 10 (verify) | verify_output.py exits 0 and reports green on audio/timing |
If any checkpoint fails, stop. Do not publish.
| Parameter | Horizontal (16:9) | Vertical (9:16) |
|---|---|---|
| Resolution | 3840×2160 (4K) | 2160×3840 (4K) |
| Frame rate | 30 fps | 30 fps |
| Encoding | H.264, 16Mbps | H.264, 16Mbps |
| Audio | AAC, 192kbps | AAC, 192kbps |
| Duration | 1-15 min | 60-90s (highlight) |
project-root/ # Remotion project root
├── src/remotion/ # Remotion source (Root.tsx, compositions, index.ts)
├── videos/{video-name}/ # Per-video directory
│ ├── topic_definition.md # Step 1
│ ├── topic_research.md # Step 2
│ ├── podcast.txt # Step 4: narration script
│ ├── phonemes.json # Step 4.5: zh-CN pronunciation overrides
│ ├── assets/manifest.json # Step 5: per-section asset registry
│ ├── publish_info.md # Step 10: title/description/tags
│ ├── podcast_audio.wav # Step 7: TTS audio
│ ├── podcast_audio.srt # Step 7: subtitles
│ ├── timing.json # Step 7: timeline (drives animations)
│ ├── thumbnail_*.png # Step 6
│ ├── output.mp4 # Step 9: 4K render
│ ├── video_with_bgm.mp4 # Step 9: with BGM
│ ├── final_video.mp4 # Step 10: final output
│ └── bgm.mp3 # Background music
└── remotion.config.ts--public-dir per videoEvery Remotion command uses --public-dir videos/{name}/ — each video's assets stay in its own directory, enabling parallel renders:
npx remotion studio src/remotion/index.ts --public-dir videos/{name}/
npx remotion render ... videos/{name}/output.mp4 --public-dir videos/{name}/ --video-bitrate 16M
npx remotion still ... videos/{name}/thumbnail.png --public-dir videos/{name}/{video-name}: lowercase English, hyphen-separated (e.g. reference-manager-comparison){section}: lowercase English, underscore-separated, matches [SECTION:xxx]thumbnail_remotion_16x9.png + thumbnail_remotion_4x3.png (or _ai_ prefix for AI-generated)Load on demand — do NOT load all at once:
| File | Load when |
|---|---|
| references/workflow-script.md | Steps 1-4 (topic → script) + Execution Modes (Auto vs Interactive) |
| references/natural-narration.md | Load before Step 4 script writing — anti-slop rules for spoken narration (kill list, structural tells, checklist) |
| references/script-polish.md | Load after Step 4 draft is written — deep editing toolkit with 24 EN+ZH before/after patterns, evidence boundaries, quality rubrics |
| references/workflow-assets.md | Step 5, or when the user supplies images/clips or wants stock/AI media |
| references/workflow-assets.md | A section needs a data-chart/infographic animation beyond the component library (transparent overlay via Hyperframes) |
| references/workflow-production.md | Steps 5.5-9.5 (publish info draft → thumbnails → TTS → Remotion → render → BGM mix) |
| references/workflow-publish.md | Steps 10-11 (publish info, verify, shorts) |
| references/platform-matrix.md | Platform-specific behavior (thumbnails, chapters, outro, publish info, shorts) |
| references/design-guide.md | MUST load before Step 8 — visual minimums, typography, animation safety |
| references/visual-taste.md | Load before Step 8 alongside design-guide — design dials, anti-default rules, visual modes, section rhythm |
| references/design-learning.md | User provides a reference video/image, or manages style profiles |
| references/troubleshooting.md | Choosing Azure voice/style, debugging hoarse/glitchy audio |
| references/troubleshooting.md | On error, script/CLI discovery, or user asks about preferences/BGM |
| templates/presets/kinetic-typography/ | Bold type-driven preset (opinion / argument / declaration videos) |
All scripts are reachable through one dispatcher — start with python3 ${SKILL_DIR}/scripts/cli.py --help; full routes and envelope error codes: references/troubleshooting.md.
Mutable state (user_prefs.json, phonemes.json) lives in ~/.video-podcast-maker/ — safe from skill updates. Auto-migrated from the skill directory on first run. Run "show preferences" to view, or "set X Y" to change. Full commands: references/troubleshooting.md.
See references/troubleshooting.md on errors, BGM options, preference learning, design-learning issues.
© Agents365-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 108 other files (scripts, references, assets) in skills/video-podcast-maker of Agents365-ai/video-podcast-maker.
Open the folder on GitHubat commit 33b8078
Video Podcast Maker next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Video Podcast Maker this skillAgents365-ai/video-podcast-maker | 1.7k | — | ~4.9k | Automated safety check: Pass | MIT | |
| Video Downloaderkangarooking/kangarooking-skills | 662 | — | ~8.3k | Automated safety check: Pass | None | |
| Ra Video Wash PipelinePluviobyte/rnskill | 1.6k | — | ~2.4k | Automated safety check: Pass | Custom licence | |
| Media To TranscriptbozhouDev/video-skills-toolkit | 150 | — | ~1.8k | Automated safety check: Notes | MIT | |
| Learning Notes Automationchubbyguan/chubbyskills | 1.2k | — | ~898 | Automated safety check: Pass | MIT | |
| Suggest Sfxhassancs91/claude-youtube-editor | 328 | — | ~3k | Automated safety check: Pass | MIT |
kangarooking/kangarooking-skills
Download or open videos and recover platform captions, audio transcripts, keyframes, screen text, visual facts, and editing observations as a plain multimodaltranscript.md.
Pluviobyte/rnskill
End-to-end Chinese video washing pipeline. An agent skill from Pluviobyte/rnskill.
bozhouDev/video-skills-toolkit
Convert audio/video URLs or local media into corrected Markdown transcripts through Volcengine recording-file ASR 2.0.
chubbyguan/chubbyskills
学习笔记自动化:视频/播客转录 → 知识点提取 → 闪卡生成 → 知识图谱更新。触发词:学习笔记、闪卡、Anki、知识提取、视频学习
hassancs91/claude-youtube-editor
Step 4 of the AI Video Editor pipeline — the SFX pass. An agent skill from hassancs91/claude-youtube-editor.
idiotLeoLYJ/Daliu-Awesome-Skills
Download videos from Douyin, Kuaishou, Xiaohongshu, and Bilibili by sharing link.
Agents365-ai/video-podcast-maker
Minimal personal narrated-video pipeline — a topic becomes a talking-head-free explainer MP4 (1080p or 4K) via script → Azure TTS (SSML) → Remotion.
Agents365-ai/video-podcast-maker
Smallest personal narrated-explainer-video pipeline (spoken narration over visuals, not an audio podcast), fully tool-agnostic and autonomous by default — topic → research ∥ asset collection →…
Categories
A skill your agent uses when the user gives a topic and wants an automated topic-driven narrated explainer, podcast, or knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat…. Video Podcast Maker is an agent skill from Agents365-ai/video-podcast-maker. Use when the user gives a topic and wants an automated topic-driven narrated explainer, podcast, or knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat Channels), or asks to learn visual design patterns from a reference video/image.
Video Podcast Maker fits situations like: the user gives a topic and wants an automated topic-driven narrated explainer; knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat Channels); asks to learn visual design patterns from a reference video/image; the user mentions creating a knowledge video.
Run `npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a claude-code`. Or copy the skill folder (skills/video-podcast-maker in Agents365-ai/video-podcast-maker) into .claude/skills/video-podcast-maker in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a codex`. Or copy the skill folder (skills/video-podcast-maker in Agents365-ai/video-podcast-maker) into .agents/skills/video-podcast-maker in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Agents365-ai/video-podcast-maker --skill video-podcast-maker -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video-podcast-maker, .gemini/skills/video-podcast-maker, .github/skills/video-podcast-maker and .opencode/skills/video-podcast-maker in your project.
Going by SKILL.md and its folder, Video Podcast Maker needs the command-line tools its instructions call (npx, python3, git and npm) and credentials named AZURE_SPEECH_KEY. Our summary lists: Python 3; Node.js; A credential in AZURE_SPEECH_KEY.
SKILL.md names 2 domains. As links in the text: github.com and remotion.dev. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Video Podcast Maker is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.
About 4.9k tokens (SKILL.md is roughly 19k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 39k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Video Podcast Maker: Video Downloader (kangarooking/kangarooking-skills, 662 stars), Ra Video Wash Pipeline (Pluviobyte/rnskill, 1.6k stars), Media To Transcript (bozhouDev/video-skills-toolkit, 150 stars) and Learning Notes Automation (chubbyguan/chubbyskills, 1.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Agents365-ai (a GitHub user) maintains it in Agents365-ai/video-podcast-maker, which has 1,670 GitHub stars. The repository holds 3 skills in this directory. The repository was last updated on October 1, 2026.
Source: Agents365-ai/video-podcast-maker on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.