Agent skill

Audio Visualizer

by ZJU-REAL in ZJU-REAL/Easel

音频可视化视频:把纯音频(播客片段、音乐、口播金句、电台)渲染成带动态波形/频谱的视频,配封面和标题,好发到抖音/B站/视频号等只收视频的平台。当用户说 音频可视化、音频转视频、播客做成视频、音频波形视频、音乐可视化、给音频配画面、声波视频、频谱视频、把音频发到视频平台、电台切片视频 时使用。基于 shared/scripts/audioviz.py(ffmpeg…

Apache-2.0Auto-check passedMedia & Creative

Install Audio Visualizer

skills CLI
$ npx skills add ZJU-REAL/Easel --skill audio-visualizer -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ZJU-REAL/Easel audio-visualizer --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/openclaw/audio-visualizer .claude/skills/audio-visualizer && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
audio-visualizer
GitHub stars
3.3k
Token cost
~542 tokens
SKILL.md length
89 words
Files
1
Skills in repo
113
Repo updated
First seen
Licence
Apache-2.0

At a glance

音频可视化视频:把纯音频(播客片段、音乐、口播金句、电台)渲染成带动态波形/频谱的视频,配封面和标题,好发到抖音/B站/视频号等只收视频的平台。当用户说 音频可视化、音频转视频、播客做成视频、音频波形视频、音乐可视化、给音频配画面、声波视频、频谱视频、把音频发到视频平台、电台切片视频 时使用。基于 shared/scripts/audioviz.py(ffmpeg…

  • Works in 4 steps: 长音频先用 audio-editing/text-condenser… → 音频原声完整嵌入输出,不重采样丢质量。 → 封面图会等比缩放居中,标题自动描边保证可读。 → …
  • Tasks that involve Video production
  • SKILL.md covers 输入, 输出(outputs/主题名/), 执行步骤 and 模式怎么选, plus 3 more sections
  • Calls python

What it does

Audio Visualizer is an agent skill from ZJU-REAL/Easel. 音频可视化视频:把纯音频(播客片段、音乐、口播金句、电台)渲染成带动态波形/频谱的视频,配封面和标题,好发到抖音/B站/视频号等只收视频的平台。当用户说 音频可视化、音频转视频、播客做成视频、音频波形视频、音乐可视化、给音频配画面、声波视频、频谱视频、把音频发到视频平台、电台切片视频 时使用。基于 shared/scripts/audioviz.py(ffmpeg showwaves/showcqt/showspectrum)。与 audio-mix 区别:那个输出音频,本 SKILL 输出视频;与 slideshow-video 区别:那个用图片,本 SKILL 用音频驱动画面。

Its SKILL.md is about 540 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Video production. It works with FFmpeg. The repository describes itself as: An open-source AI agent for social media — discover trends, create content, publish everywhere, and learn what works across Xiaohongshu, Douyin, Zhihu, Bilibili, and more.🎨一个开源的… The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Video production

Example prompts

  • “/audio-visualizer”

Requirements

  • Python 3

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. 长音频先用 audio-editing/text-condenser 截出金句片段再可视化,别整集渲染。
  2. 音频原声完整嵌入输出,不重采样丢质量。
  3. 封面图会等比缩放居中,标题自动描边保证可读。
  4. 产物统一进 outputs/主题名/。

What it can do on your machine

Read from SKILL.md and the folder at commit 5e0ccc1. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Audio Visualizer loads about 542 tokens when it runs. Until then it costs about 78 tokens; SKILL.md has 89 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~78
When it runs · the whole SKILL.md, loaded when a task matches
~542

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ZJU-REAL/Easel at commit 5e0ccc1, republished under its Apache-2.0 licence (© ZJU-REAL). 89 words, ~542 tokens.

Download SKILL.mdSave it as .claude/skills/audio-visualizer/SKILL.md (or your agent's skills folder).
name
audio-visualizer
description
音频可视化视频:把纯音频(播客片段、音乐、口播金句、电台)渲染成带动态波形/频谱的视频,配封面和标题,好发到抖音/B站/视频号等只收视频的平台。当用户说 音频可视化、音频转视频、播客做成视频、音频波形视频、音乐可视化、给音频配画面、声波视频、频谱视频、把音频发到视频平台、电台切片视频 时使用。基于 shared/scripts/audio_viz.py(ffmpeg showwaves/showcqt/showspectrum)。与 audio-mix 区别:那个输出音频,本 SKILL 输出视频;与 slideshow-video 区别:那个用图片,本 SKILL 用音频驱动画面。
layer
produce

音频可视化视频

把音频渲染成带动态波形/频谱的视频,配封面+标题,让纯音频能发到视频平台。全部走 skills/shared/scripts/audio_viz.py,不要手拼 showwaves/showcqt 滤镜。

输出音频(混音)见 audio-mix;用图片做视频见 slideshow-video; 给已有视频加字幕见 auto-subtitle。

输入

字段必填说明
音频文件是播客/音乐/口播片段(没给就问)
画幅是用户或上游任务未明确横版/竖版(或具体分辨率)时,制作前必须追问并等确认;不得按平台、Profile 或默认值静默推断,已明确则不重复问
模式否cqt(默认,音乐最好看)/ bars / waves / spectrum
封面否居中封面图(专辑封面/头像/主题图)
标题否顶部标题文字

输出(outputs/主题名/)

  • 可视化视频(*.mp4,音频已嵌入)
  • 报告:模式、时长、画幅

执行步骤

脚本路径(相对项目根):skills/shared/scripts/audio_viz.py(render -h 看参数)。

bash
# 音乐/金句:CQT 音乐频谱(随音符跳动,最好看)
python skills/shared/scripts/audio_viz.py render -i clip.mp3 \
  -o outputs/主题名/out.mp4 --mode cqt --title "本期金句" --cover cover.jpg

# 播客口播:底部波形条 + 封面
python skills/shared/scripts/audio_viz.py render -i podcast.mp3 \
  -o outputs/主题名/out.mp4 --mode waves --cover avatar.png --size 1080x1920

# 律动柱状 / 滚动声谱
python skills/shared/scripts/audio_viz.py render -i song.mp3 -o out.mp4 --mode bars
python skills/shared/scripts/audio_viz.py render -i song.mp3 -o out.mp4 --mode spectrum

模式怎么选

模式观感适用
cqt全屏音符频谱,随旋律跳动音乐、有旋律的内容(默认)
bars底部频谱柱,律动感强音乐、卡点、电台
waves底部波形线,简洁干净播客、口播、访谈
spectrum全屏滚动声谱图,科技感电子/科技类、氛围

--bg-image 换背景图,--color 换背景色,--wave-color 换波形颜色。

Profile 感知

  • 有 Profile:platforms.md 只用于给出画幅建议,仍须用户确认;标题/封面风格贴合账号; 播客/口播账号默认 waves,音乐账号默认 cqt/bars。
  • 无 Profile:先确认横版/竖版;默认 cqt 模式。

规则

  1. 长音频先用 audio-editing/text-condenser 截出金句片段再可视化,别整集渲染。
  2. 音频原声完整嵌入输出,不重采样丢质量。
  3. 封面图会等比缩放居中,标题自动描边保证可读。
  4. 产物统一进 outputs/主题名/。

参考来源

音频波形/频谱可视化用 ffmpeg showwaves/showfreqs/showspectrum/showcqt,是播客/音频号 上视频平台的标准做法。把各可视化滤镜与封面/标题合成封装成确定性脚本。

© ZJU-REAL, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/openclaw/audio-visualizer of ZJU-REAL/Easel.

Open the folder on GitHubat commit 5e0ccc1

Compare with similar skills

Audio Visualizer next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Audio Visualizer compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Audio Visualizer this skillZJU-REAL/Easel3.3k—~542Automated safety check: PassApache-2.0
Video Understandcalesthio/OpenMontage66k—~841Automated safety check: PassAGPL-3.0
HyperFrames Video Entry Pointheygen-com/hyperframes59k3 repos~5.2kAutomated safety check: PassApache-2.0
Video Shotseternityspring/reelbench-skills8721 repos~1.8kAutomated safety check: NotesApache-2.0
Mobile Demo Film Editorsuperset-sh/superset15k—~1.8kAutomated safety check: PassCustom licence
Video Editcalesthio/OpenMontage66k—~855Automated safety check: NotesAGPL-3.0

Similar skills

  • Video Understand

    calesthio/OpenMontage

    Understand video content locally using ffmpeg frame extraction and Whisper transcription.

    66k GitHub stars~841 tokensUpdated 6 days ago
    Media & CreativeAuto-check passed
  • HyperFrames Video Entry Point

    heygen-com/hyperframes

    Entry point for making, editing and rendering videos from HTML compositions with HyperFrames, routing each request to the right workflow.

    59k GitHub starsUsed in 3 repos~5.2k tokens
    Media & CreativeAuto-check passed
  • Video Shots

    eternityspring/reelbench-skills

    拉片:把一条成片拆成逐镜头的分析表——每个镜头的时长、景别、类别、运镜、画面. An agent skill from eternityspring/reelbench-skills.

    872 GitHub starsUsed in 1 repo~1.8k tokens
    Media & CreativeAuto-check: notes
  • Mobile Demo Film Editor

    superset-sh/superset

    Edits real mobile screen recordings into a configurable demo video with phone framing, title cards, cutaways and an end card, using a bundled renderer.

    15k GitHub stars~1.8k tokensUpdated today
    Media & CreativeAuto-check passed
  • Video Edit

    calesthio/OpenMontage

    Edit videos locally using ffmpeg. An agent skill from calesthio/OpenMontage.

    66k GitHub stars~855 tokensUpdated 6 days ago
    Media & CreativeAuto-check: notes
  • Lemo-Opuscar Short Film Director

    lemomo-ai/lemo-opuscar

    Directs a short film made entirely in code in one of the Lemo-Opuscar library's named visual styles, from fetching the library through production and delivery.

    1.4k GitHub stars~596 tokensUpdated yesterday
    Media & CreativeAuto-check passed

More from ZJU-REAL/Easel

All 113 skills in this repo
  • 微信公众号文章自动创作与发布工具。给定参考文章、文字或文档,自动搜索整理全网相关信息、生成图文并茂的公众号文章,并发布到微信公众号草稿箱。特别强调反 AI 检测写作。

    3.3k GitHub stars~1.8k tokensUpdated today
    Auto-check passed
  • Card Design

    ZJU-REAL/Easel

    社媒卡片视觉设计系统:提供配色、中文字体层级、满画幅布局、品类骨架和死空白/密度质检,避免模板化 PPT 与廉价 AI 感。

    3.3k GitHub stars~657 tokensUpdated today
    Auto-check passed
  • Ecom Details Image

    ZJU-REAL/Easel

    生成电商商品视觉方案:主图概念、场景图、详情页视觉方向和 AI 生图 Prompt. An agent skill from ZJU-REAL/Easel.

    3.3k GitHub stars~1.1k tokensUpdated today
    Auto-check: notes
  • Infographic

    ZJU-REAL/Easel

    将数据或文字内容转化为可视化信息图,支持静态(AntV)和动画 GIF 两种模式。当用户需要制作信息图、数据可视化、流程图、对比图、动画图表、GIF 图表、思维导图、SWOT 分析图时调用。本地渲染信息图/GIF 动画;要单张静态图片 URL 用 chart-visualization,要 CSV/JSON→整页报告用 data-report

    3.3k GitHub stars~643 tokensUpdated today
    Auto-check passed
  • Novel Writer

    ZJU-REAL/Easel

    长篇小说/网文连载创作:从世界观、人设和三级大纲写到逐章正文,并用文件化状态维护伏笔、前情和跨章一致性. An agent skill from ZJU-REAL/Easel.

    3.3k GitHub stars~1k tokensUpdated today
    Auto-check passed
  • Paper Explainer

    ZJU-REAL/Easel

    科研论文解读:解析 arXiv/PDF 的公式与图表,提炼问题、贡献、方法、关键图和结论,再产出 B站/视频号解读视频或知乎/公众号图文。

    3.3k GitHub stars~1.4k tokensUpdated today
    Auto-check passed

Works with

Questions about Audio Visualizer

What does Audio Visualizer do?

音频可视化视频:把纯音频(播客片段、音乐、口播金句、电台)渲染成带动态波形/频谱的视频,配封面和标题,好发到抖音/B站/视频号等只收视频的平台。当用户说 音频可视化、音频转视频、播客做成视频、音频波形视频、音乐可视化、给音频配画面、声波视频、频谱视频、把音频发到视频平台、电台切片视频 时使用。基于 shared/scripts/audioviz.py(ffmpeg…. Audio Visualizer is an agent skill from ZJU-REAL/Easel.

When should I use Audio Visualizer?

Audio Visualizer fits situations like: tasks that involve Video production.

How do I install Audio Visualizer in Claude Code?

Run `npx skills add ZJU-REAL/Easel --skill audio-visualizer -a claude-code`. Or copy the skill folder (skills/openclaw/audio-visualizer in ZJU-REAL/Easel) into .claude/skills/audio-visualizer in your project. Claude Code loads it when a task matches its description.

How do I install Audio Visualizer in Codex?

Run `npx skills add ZJU-REAL/Easel --skill audio-visualizer -a codex`. Or copy the skill folder (skills/openclaw/audio-visualizer in ZJU-REAL/Easel) into .agents/skills/audio-visualizer in your project. Codex loads it when a task matches its description.

Can I use Audio Visualizer in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ZJU-REAL/Easel --skill audio-visualizer -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audio-visualizer, .gemini/skills/audio-visualizer, .github/skills/audio-visualizer and .opencode/skills/audio-visualizer in your project.

What does Audio Visualizer need to run?

Going by SKILL.md and its folder, Audio Visualizer needs the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Audio Visualizer access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Audio Visualizer safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Audio Visualizer use?

Audio Visualizer is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Audio Visualizer use?

About 542 tokens (SKILL.md is roughly 2.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Audio Visualizer?

Skills that share tags, products or a category with Audio Visualizer: Video Understand (calesthio/OpenMontage, 66k stars), HyperFrames Video Entry Point (heygen-com/hyperframes, 59k stars), Video Shots (eternityspring/reelbench-skills, 872 stars) and Mobile Demo Film Editor (superset-sh/superset, 15k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Audio Visualizer?

ZJU-REAL (a GitHub organization) maintains it in ZJU-REAL/Easel, which has 3,310 GitHub stars. The repository holds 113 skills in this directory. The repository was last updated on October 9, 2026.

Source: ZJU-REAL/Easel on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.