Agent skill

Audio Editing

by ZJU-REAL in ZJU-REAL/Easel

通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing…

Apache-2.0Auto-check passedMedia & Creative

Install Audio Editing

skills CLI
$ npx skills add ZJU-REAL/Easel --skill audio-editing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ZJU-REAL/Easel audio-editing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/openclaw/audio-editing .claude/skills/audio-editing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
audio-editing
GitHub stars
3.4k
Token cost
~753 tokens
SKILL.md length
143 words
Files
2
Skills in repo
114
Repo updated
First seen
Licence
Apache-2.0

At a glance

通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing…

  • Works in 9 steps: 环境检查 + 探测 → 裁剪 trim → 转码 convert → …
  • Tasks that involve Music and audio generation
  • SKILL.md covers 输入, 输出, 执行步骤 and 与 audio-denoise 的边界, plus 2 more sections
  • Calls python

What it does

Audio Editing is an agent skill from ZJU-REAL/Easel. 通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing 做除降噪外的通用音频操作(也内置 denoise 作为兜底)。

Its SKILL.md is about 750 tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `EASEL-META.md`).

It sits in Media & Creative, covering Music and audio generation. It works with FFmpeg. The repository describes itself as: An open-source AI agent for social media — discover trends, create content, publish everywhere, and learn what works across Xiaohongshu, Douyin, Zhihu, Bilibili, and more.🎨一个开源的… The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Music and audio generation

Example prompts

  • “转成 mp3”
  • “调音量/响度”
  • “提取音轨/扒音频”
  • “/audio-editing”

Requirements

  • Python 3

Workflow steps

9 steps, taken from the step headings in SKILL.md.

  1. 环境检查 + 探测
  2. 裁剪 trim
  3. 转码 convert
  4. 音量归一化 normalize
  5. 提取音轨 extract
  6. 拼接 concat
  7. 淡入淡出 fade
  8. 变速 speed(保持音高)
  9. 降噪 denoise(兜底,专项请用 audio-denoise)

What it can do on your machine

Read from SKILL.md and the folder at commit 278f420. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Audio Editing loads about 753 tokens when it runs. Until then it costs about 60 tokens; SKILL.md has 143 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~60
When it runs · the whole SKILL.md, loaded when a task matches
~753

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ZJU-REAL/Easel at commit 278f420, republished under its Apache-2.0 licence (© ZJU-REAL). 143 words, ~753 tokens.

Download SKILL.mdSave it as .claude/skills/audio-editing/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
audio-editing
description
通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing 做除降噪外的通用音频操作(也内置 denoise 作为兜底)。
layer
produce

通用音频处理

除降噪外的通用音频操作:裁剪、转码、音量归一化、提取音轨、拼接、淡入淡出、变速。全部通过共享脚本 skills/shared/scripts/audio_ops.py 封装 ffmpeg,参数确定、可复现,不现场手拼命令。

输入

字段必填说明
input_file是音频或视频文件路径
operation是trim / convert / normalize / extract / concat / fade / speed / denoise / info
output_file否默认 outputs/主题名/{name}-{op}.{ext}

支持格式:wav / mp3 / m4a / aac / flac,视频容器 mp4 / mkv / mov(提取音轨)。

输出

  • 处理后的音频文件(放入 outputs/主题名/)
  • 每次操作打印实际执行的 ffmpeg 命令 + 输出文件的时长/码率/声道/采样率

执行步骤

脚本路径(相对项目根):skills/shared/scripts/audio_ops.py。每个子命令都支持 -h。

0. 环境检查 + 探测
bash
python skills/shared/scripts/audio_ops.py info input.mp3

脚本自身会检查 ffmpeg/ffprobe,缺失时给安装提示。先 info 展示文件时长/码率/声道再动手。

1. 裁剪 trim
bash
# 起止时间
python skills/shared/scripts/audio_ops.py trim in.mp3 -o outputs/主题名/clip.mp3 --start 00:00:05 --end 00:00:20
# 起点 + 时长
python skills/shared/scripts/audio_ops.py trim in.mp3 -o clip.mp3 --start 5 --duration 15
2. 转码 convert
bash
python skills/shared/scripts/audio_ops.py convert in.wav -o out.mp3 --bitrate 192k
python skills/shared/scripts/audio_ops.py convert in.m4a -o out.wav --sample-rate 44100 --channels 2

输出格式由扩展名决定(mp3/wav/m4a/aac)。

3. 音量归一化 normalize
bash
python skills/shared/scripts/audio_ops.py normalize in.mp3 -o out.mp3

默认 loudnorm 到 -14 LUFS / -1.5 dBTP(社媒/播客通用响度)。可用 --i --tp --lra 覆盖。

4. 提取音轨 extract
bash
python skills/shared/scripts/audio_ops.py extract video.mp4 -o audio.m4a
python skills/shared/scripts/audio_ops.py extract video.mp4 -o audio.aac --copy   # 不重编码,最快
5. 拼接 concat
bash
python skills/shared/scripts/audio_ops.py concat a.mp3 b.mp3 c.mp3 -o all.mp3

按参数顺序拼接,重编码方式兼容不同采样率/容器。

6. 淡入淡出 fade
bash
python skills/shared/scripts/audio_ops.py fade in.mp3 -o out.mp3 --fade-in 2 --fade-out 3

--fade-out 自动定位到结尾前 N 秒。

7. 变速 speed(保持音高)
bash
python skills/shared/scripts/audio_ops.py speed in.mp3 -o out.mp3 --factor 1.5   # 1.5 倍速
python skills/shared/scripts/audio_ops.py speed in.mp3 -o out.mp3 --factor 0.8   # 放慢

基于 atempo 保音高,超 0.5-2.0 范围自动级联。

8. 降噪 denoise(兜底,专项请用 audio-denoise)
bash
python skills/shared/scripts/audio_ops.py denoise in.wav -o out.wav --tier 2

与 audio-denoise 的边界

  • audio-denoise:降噪专项 SKILL,负责选级、模型下载、降噪报告等完整流程——需要认真降噪时用它。
  • audio-editing:通用音频处理。denoise 子命令仅作为顺手兜底,两者调用同一个 audio_ops.py denoise,能力一致,定位不同。

规则

  1. 绝不删除原始文件 — 只写新文件到 outputs/主题名/。
  2. 先 info 再处理 — 展示文件信息,避免误操作。
  3. 视频输入只动音频 — denoise 对视频自动 -c:v copy;纯抽音用 extract。
  4. 透明执行 — 脚本会打印实际 ffmpeg 命令。
  5. 无 Profile 也能用 — 音频处理不依赖账号画像。

Profile 感知

有 Profile 时可据平台偏好选默认参数(如短视频响度、目标码率);无 Profile 退到通用默认(-14 LUFS、192k)。

© ZJU-REAL, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/openclaw/audio-editing of ZJU-REAL/Easel.

  • SKILL.md
  • EASEL-META.md

Open the folder on GitHubat commit 278f420

Compare with similar skills

Audio Editing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Audio Editing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Audio Editing this skillZJU-REAL/Easel3.4k—~753Automated safety check: PassApache-2.0
ElevenLabs Voiceover Generatordigitalsamba/claude-code-video-toolkit2.2k1 repos~2.7kAutomated safety check: NotesMIT
Qiaomu Cutjoeseesun/qiaomu-cut-skill372—~6.8kAutomated safety check: NotesMIT
ShowtimeFavioVazquez/showtime220—~3kAutomated safety check: PassMIT
Release Videohuytieu/COG-second-brain1.3k—~1.7kAutomated safety check: PassMIT
Noti Tiktok Vnnotivn/AIEV127—~5.4kAutomated safety check: PassMIT

Similar skills

  • ElevenLabs Voiceover Generator

    digitalsamba/claude-code-video-toolkit

    Generates narration, sound effects and cloned voices through the ElevenLabs API, with model and setting choices tuned to the content's style.

    2.2k GitHub starsUsed in 1 repo~2.7k tokens
    Media & CreativeAuto-check: notes
  • Qiaomu Cut

    joeseesun/qiaomu-cut-skill

    把一句话需求转成可复现、可验收视频工程的乔木智能剪辑导演。Use when the user asks to create, plan, edit, remix, explain, narrate, subtitle, animate, composite, or render a video—including one-line requests such as “制作一个科普视频:介绍…

    372 GitHub stars~6.8k tokensUpdated 12 days ago
    Media & CreativeAuto-check: notes
  • Showtime

    FavioVazquez/showtime

    A skill your agent uses when the user wants a video made, edited or finished: a launch or promo, product demo, explainer, trailer or teaser, tutorial or walkthrough, a screen recording turned into a…

    220 GitHub stars~3k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Release Video

    huytieu/COG-second-brain

    Turn a product release (the list of shipped items plus real screen recordings) into a motion recap video and one explained demo per feature, with sound effects tied to on-screen motion and a…

    1.3k GitHub stars~1.7k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Noti Tiktok Vn

    notivn/AIEV

    Edit a Vietnamese vertical TikTok video (9:16) with HyperFrames following the Noti.vn/GĐT standard - talking-head + kinetic typography + karaoke captions + zoom/punch-in camera + timestamp-synced…

    127 GitHub stars~5.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Scenario Seedance Music Video

    scenario-labs/skills

    A skill your agent uses when turning a song, track, or audio master into a finished music video with Scenario and Seedance: planning shots against beats and sections, transcribing lyrics, generating…

    946 GitHub stars~1.8k tokensUpdated today
    Media & CreativeAuto-check passed

More from ZJU-REAL/Easel

All 114 skills in this repo
  • Gzh Design

    ZJU-REAL/Easel

    微信公众号文章排版引擎:把 Markdown / Word(.docx) / PDF / 纯文本转成可直接粘贴进公众号编辑器的 HTML,自动章节编号、关键词标记、引言卡、目录、代码块、图片/GIF、作者签名;主题从 references/theme-index.md…

    3.4k GitHub stars~2.1k tokensUpdated yesterday
    Auto-check passed
  • 微信公众号文章自动创作与发布工具。给定参考文章、文字或文档,自动搜索整理全网相关信息、生成图文并茂的公众号文章,并发布到微信公众号草稿箱。特别强调反 AI 检测写作。

    3.4k GitHub stars~1.8k tokensUpdated yesterday
    Auto-check passed
  • Card Design

    ZJU-REAL/Easel

    社媒卡片视觉设计系统:提供配色、中文字体层级、满画幅布局、品类骨架和死空白/密度质检,避免模板化 PPT 与廉价 AI 感。

    3.4k GitHub stars~657 tokensUpdated yesterday
    Auto-check passed
  • Ecom Details Image

    ZJU-REAL/Easel

    生成电商商品视觉方案:主图概念、场景图、详情页视觉方向和 AI 生图 Prompt. An agent skill from ZJU-REAL/Easel.

    3.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check: notes
  • Infographic

    ZJU-REAL/Easel

    将数据或文字内容转化为可视化信息图,支持静态(AntV)和动画 GIF 两种模式。当用户需要制作信息图、数据可视化、流程图、对比图、动画图表、GIF 图表、思维导图、SWOT 分析图时调用。本地渲染信息图/GIF 动画;要单张静态图片 URL 用 chart-visualization,要 CSV/JSON→整页报告用 data-report

    3.4k GitHub stars~643 tokensUpdated yesterday
    Auto-check passed
  • Novel Writer

    ZJU-REAL/Easel

    长篇小说/网文连载创作:从世界观、人设和三级大纲写到逐章正文,并用文件化状态维护伏笔、前情和跨章一致性. An agent skill from ZJU-REAL/Easel.

    3.4k GitHub stars~1k tokensUpdated yesterday
    Auto-check passed

Works with

Questions about Audio Editing

What does Audio Editing do?

通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing…. Audio Editing is an agent skill from ZJU-REAL/Easel.

When should I use Audio Editing?

Audio Editing fits situations like: tasks that involve Music and audio generation.

How do I install Audio Editing in Claude Code?

Run `npx skills add ZJU-REAL/Easel --skill audio-editing -a claude-code`. Or copy the skill folder (skills/openclaw/audio-editing in ZJU-REAL/Easel) into .claude/skills/audio-editing in your project. Claude Code loads it when a task matches its description.

How do I install Audio Editing in Codex?

Run `npx skills add ZJU-REAL/Easel --skill audio-editing -a codex`. Or copy the skill folder (skills/openclaw/audio-editing in ZJU-REAL/Easel) into .agents/skills/audio-editing in your project. Codex loads it when a task matches its description.

Can I use Audio Editing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ZJU-REAL/Easel --skill audio-editing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audio-editing, .gemini/skills/audio-editing, .github/skills/audio-editing and .opencode/skills/audio-editing in your project.

What does Audio Editing need to run?

Going by SKILL.md and its folder, Audio Editing needs the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Audio Editing access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Audio Editing safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Audio Editing use?

Audio Editing is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Audio Editing use?

About 753 tokens (SKILL.md is roughly 3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Audio Editing?

Skills that share tags, products or a category with Audio Editing: ElevenLabs Voiceover Generator (digitalsamba/claude-code-video-toolkit, 2.2k stars), Qiaomu Cut (joeseesun/qiaomu-cut-skill, 372 stars), Showtime (FavioVazquez/showtime, 220 stars) and Release Video (huytieu/COG-second-brain, 1.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Audio Editing?

ZJU-REAL (a GitHub organization) maintains it in ZJU-REAL/Easel, which has 3,376 GitHub stars. The repository holds 114 skills in this directory. The repository was last updated on October 9, 2026.

Source: ZJU-REAL/Easel on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.