Multimedia Accessibility
Owl-Listener/inclusive-design-skills
Design accessible video, audio, and multimedia content with captions, transcripts, and audio descriptions.
Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social…
$ npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install calesthio/generative-media-skills audio-mixing-mastering --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/production/audio-craft/audio-mixing-mastering .claude/skills/audio-mixing-mastering && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "audio-mixing-mastering" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-mastering into .claude/skills/audio-mixing-mastering/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-mixing-mastering", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-masteringType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install calesthio/generative-media-skills audio-mixing-mastering --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/production/audio-craft/audio-mixing-mastering .agents/skills/audio-mixing-mastering && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "audio-mixing-mastering" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-mastering into .agents/skills/audio-mixing-mastering/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-mixing-mastering", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install calesthio/generative-media-skills audio-mixing-mastering --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/production/audio-craft/audio-mixing-mastering .cursor/skills/audio-mixing-mastering && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "audio-mixing-mastering" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-mastering into .cursor/skills/audio-mixing-mastering/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-mixing-mastering", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/calesthio/generative-media-skills.git --path skills/production/audio-craft/audio-mixing-mastering--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install calesthio/generative-media-skills audio-mixing-mastering --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/production/audio-craft/audio-mixing-mastering .gemini/skills/audio-mixing-mastering && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "audio-mixing-mastering" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-mastering into .gemini/skills/audio-mixing-mastering/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-mixing-mastering", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install calesthio/generative-media-skills audio-mixing-masteringInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/production/audio-craft/audio-mixing-mastering .github/skills/audio-mixing-mastering && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "audio-mixing-mastering" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-mastering into .github/skills/audio-mixing-mastering/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-mixing-mastering", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install calesthio/generative-media-skills audio-mixing-mastering --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/production/audio-craft/audio-mixing-mastering .opencode/skills/audio-mixing-mastering && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "audio-mixing-mastering" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/audio-craft/audio-mixing-mastering into .opencode/skills/audio-mixing-mastering/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audio-mixing-mastering", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
audio-mixing-masteringProvider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social…
Audio Mixing Mastering is an agent skill from calesthio/generative-media-skills. Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social content. Use when planning, mixing, repairing, mastering, QCing, or delivering dialogue, music, ambience, and sound effects, including loudness/true-peak targets, intelligibility, accessibility, stems, stereo/immersive decisions, platform/client specs, and final audio QA.
Its SKILL.md is about 7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts (for example `EVAL.md`, `scripts/measure_loudness.py` and `tests/test_measure_loudness.py`).
It sits in Media & Creative, covering Podcasting, Social media posts and Music and audio generation. The repository describes itself as: Research-backed agent skills and tools for premium image, video, audio, voice, and generative media production across AI coding assistants. The licence is MIT.
7 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 8c85352. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
pythonFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
itu.inttech.ebu.chatsc.orgaes.orgsupport.spotify.compodcasters.apple.comsupport.google.compartnerhelp.netflixstudios.comw3.orgffmpeg.orgFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Audio Mixing Mastering loads about 7k tokens when it runs. Until then it costs about 124 tokens; SKILL.md has 3,730 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from calesthio/generative-media-skills at commit 8c85352, republished under its MIT licence (© calesthio). 3,730 words, ~7,030 tokens.
.claude/skills/audio-mixing-mastering/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.Use this skill when audio must survive real playback: phone speakers, earbuds, laptops, TV soundbars, cinema-style trailers, podcasts, client review links, broadcast deliveries, and social feeds. Treat the mix as a production decision, not as a last-minute normalization step.
The job is to make the audience understand the foreground, feel the intended energy, avoid fatigue or distortion, and pass the declared delivery spec.
Before touching levels, write a short audio contract for the project:
Do not master blindly to the loudest reference. Streaming and broadcast systems may normalize loudness, and over-limiting can reduce clarity while gaining little or nothing at playback.
Documented facts:
loudnorm filter implements EBU R128 loudness normalization, supports single- and double-pass operation, can target integrated loudness, loudness range, and true peak, and upsamples to 192 kHz in dynamic mode for true-peak detection. Verified 2026-07-10: https://ffmpeg.org/ffmpeg-filters.html#loudnormUse scripts/measure_loudness.py when an agent needs a repeatable local measurement before making mix or delivery decisions. It requires Python 3.11+ and an ffmpeg executable with the loudnorm filter. The script invokes FFmpeg with an argument array rather than a shell, selects the first audio stream, and emits stable JSON. It never normalizes, rewrites, or replaces the input.
Measure without assuming a target:
python scripts/measure_loudness.py final_mix.wav --prettyCompare against an explicitly selected project target:
python scripts/measure_loudness.py final_mix.wav \
--target-lufs -16 --lufs-tolerance 1 \
--target-true-peak -1 --prettyFor mono files intended to play from both speakers as dual mono, request FFmpeg's dual-mono compensation explicitly:
python scripts/measure_loudness.py narration_mono.wav --dual-mono --prettyThe JSON report includes:
Exit codes are:
0: measurement completed and every requested target check passed;2: measurement completed but one or more requested target checks failed;3: operational failure such as a missing input, missing FFmpeg, timeout, absent audio stream, or malformed loudnorm output.This tool is deterministic assistance, not final mix approval. Listen to the complete deliverable, confirm the selected target is authoritative, inspect channel layout and codec behavior, and remeasure encoded outputs when delivery risk warrants it. Use a proper calibrated meter or the receiver's mandated QC system when the specification requires one.
Empirical observations:
Production heuristics:
Make a reproducible session before processing:
If receiving only a final exported video, demux audio for analysis, but keep picture reference locked. If receiving multitrack/stems, avoid mastering the stereo mix until stem balance is correct.
Build the mix from the foreground outward:
Do not solve a balance problem only with a limiter. If the voice disappears under music, reduce or carve the music. If SFX feel too sharp, shape the SFX. If the master misses target loudness after the mix feels right, then use bus gain/limiting carefully.
For spoken-word content:
If AI voice artifacts remain audible, prefer localized repairs: replace the line, regenerate the phrase, use spectral repair, trim clicks, redraw fades, or edit breaths. Broadband noise reduction on the full VO can make artifacts worse.
Music must be licensed, appropriate, and technically controlled:
SFX should clarify action, scale, or emotion:
Use processing to solve named problems:
Limiter safety:
Prefer this order:
Useful target map, verified 2026-07-10:
| Destination | Documented or working target | True peak | Notes |
|---|---|---|---|
| EBU R 128 broadcast/program exchange | -23 LUFS | -1 dBTP | Use compliant meter; measure full programme. Practical target tolerance may be +/-1 LU; QC measurement tolerance may be +/-0.2 LU. |
| ATSC A/85 TV delivery/exchange without metadata | -24 LKFS | -2 dBTP | Long-form dialogue loudness; short-form full-program mix loudness per A/85 quick reference. |
| ATSC A/85 streaming services | -23 to -27 LKFS | -2 dBTP | Unless prior arrangement says otherwise. |
| Netflix branded content | -27 LKFS +/-2 LU dialog-gated | -2 dBTP | Requires full spec compliance and stems/format rules. |
| Apple Podcasts | about -16 dB LKFS, +/-1 dB | -1 dB FS true peak | Precondition before encoding. |
| Spotify music | -14 dB integrated LUFS | below -1 dBTP, or below -2 dBTP if louder than -14 LUFS | Platform-specific music guidance. |
| AES internet audio examples | speech/assorted around -18 LUFS; music -16 LUFS track-normalized; album loudest track -14 LUFS | -1 dBTP at lossy codec input | For distributors; useful context for producers. |
| YouTube upload | no official loudness target found on upload help page | use project heuristic | Official page lists audio bitrates, not a loudness target. |
| General speech-first social/web video | -16 to -14 LUFS heuristic | -1 to -2 dBTP heuristic | Check speech clarity on phone/laptop/earbuds. |
Never convert these targets into universal rules. If the client asks for a spec, obey the client. If the destination will normalize loudness, prioritize clean transients, dialogue clarity, and codec-safe peaks.
For stereo/social:
For 5.1/7.1/Atmos/immersive:
Treat clear speech as an accessibility requirement:
Plan deliverables before final mix export:
Stems must sum cleanly to the full mix unless the client explicitly requests differently. Check that stem exports start at the same timecode, same sample rate, same length, same channel layout, and no hidden master-bus processing is missing. If master-bus compression or limiting changes the balance, either print processed stems through the bus in solo-safe groups or document that stems are pre-master elements.
When a mix fails, diagnose by symptom:
Iterate in this order: source repair, edit cleanup, balance automation, track processing, bus processing, loudness normalization, encode check, final QC. Late mastering should not hide known source or edit defects.
Before delivery, create a concise QA note:
If the project has multiple outputs, QA each exported file, not just the first master.
Example: speech-first social explainer
Production intent: 60-second vertical explainer with AI narration, light music bed, UI clicks, and subtitles for Instagram/TikTok/YouTube Shorts. No formal platform loudness spec supplied.
Direction:
Expected result: narration is intelligible on phone speakers, music feels present but not competitive, no clipping after encode, and client can revise music without rebuilding VO.
Likely failure modes: too much high-mid EQ on AI voice, pumping music duck, clicks from VO edits, SFX peaks causing limiter gain reduction.
Example: podcast/video recut from noisy remote interview
Production intent: 18-minute podcast highlight video assembled from two remote speakers, intro music, outro ad tag, and occasional lower-thirds.
Direction:
Expected result: speakers feel matched, background noise is reduced but natural, intro/outro are not startling, and listeners do not ride the volume.
Likely failure modes: over-denoising, clipping preserved from source, one speaker much brighter/louder, music intro hitting the limiter and lowering subsequent speech.
Example: trailer/ad with VO, music, hits, and broadcast-style client review
Production intent: 30-second product trailer with cinematic music, VO, whooshes, impacts, and final logo sting. Client may place it on paid social and possibly broadcast later.
Direction:
Expected result: the trailer feels big without masking the message, has safe peaks for lossy platforms, and can be re-versioned for broadcast with stems.
Likely failure modes: SFX hits forcing limiter pumping, music masking final CTA, mono fold-down losing wide risers, broadcast version made by only lowering the social master instead of remixing to spec.
© calesthio, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 3 other files (scripts) in skills/production/audio-craft/audio-mixing-mastering of calesthio/generative-media-skills.
Open the folder on GitHubat commit 8c85352
Audio Mixing Mastering next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Audio Mixing Mastering this skillcalesthio/generative-media-skills | 197 | — | ~7k | Automated safety check: Pass | MIT | |
| Multimedia AccessibilityOwl-Listener/inclusive-design-skills | 105 | — | ~894 | Automated safety check: Pass | MIT | |
| ContentclawLeoYeAI/openclaw-master-skills | 2.2k | — | ~6.8k | Automated safety check: Notes | MIT | |
| AI Fomovincelele/ai-fomo-skills | 248 | — | ~1.2k | Automated safety check: Pass | MIT | |
| Z Qwen Audio Studiotjxj/z-skills | 548 | — | ~716 | Automated safety check: Pass | None | |
| Podcastteam-attention/plugins-for-claude-natives | 827 | — | ~1.5k | Automated safety check: Pass | MIT |
Owl-Listener/inclusive-design-skills
Design accessible video, audio, and multimedia content with captions, transcripts, and audio descriptions.
LeoYeAI/openclaw-master-skills
Automated content generation engine. An agent skill from LeoYeAI/openclaw-master-skills.
vincelele/ai-fomo-skills
Judge, summarize, create podcast learning notes, and file AI-related information through the user's Personal Alignment Layer.
tjxj/z-skills
A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…
team-attention/plugins-for-claude-natives
Generate Korean podcast episodes from any source (URLs, tweets, articles, PDFs) — analyzes content, writes a script, generates audio via OpenAI TTS, converts to MP4, and auto-uploads to YouTube.
NeverSight/learn-skills.dev
Create AI-powered podcasts with text-to-speech, music, and audio editing.
calesthio/generative-media-skills
A skill your agent uses to turn generated, captured, scanned, or modeled 3D output into production-ready standalone assets for DCC, real-time engine, web, or interchange delivery.
calesthio/generative-media-skills
Provider-independent captions and media accessibility direction for AI agents producing or finishing generated videos, ads, social clips, explainers, avatar videos, documentaries, podcasts/video…
calesthio/generative-media-skills
Provider-independent production workflow for agents assembling, auditing, executing, and handing off ComfyUI node-graph workflows for image, video, upscale, inpaint, conditioning, and batch media…
calesthio/generative-media-skills
Provider-independent FFmpeg finishing workflow for AI agents preparing generated or edited media deliverables.
calesthio/generative-media-skills
Provider-independent quality assurance for AI-generated and AI-assisted media.
calesthio/generative-media-skills
Provider-independent production workflow for AI agents assembling generated or source media into HyperFrames HTML/CSS/JS videos.
Categories
Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social…. Audio Mixing Mastering is an agent skill from calesthio/generative-media-skills. Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social content.
Audio Mixing Mastering fits situations like: delivering dialogue; including loudness/true-peak targets; intelligibility; stereo/immersive decisions.
Run `npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a claude-code`. Or copy the skill folder (skills/production/audio-craft/audio-mixing-mastering in calesthio/generative-media-skills) into .claude/skills/audio-mixing-mastering in your project. Claude Code loads it when a task matches its description.
Run `npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a codex`. Or copy the skill folder (skills/production/audio-craft/audio-mixing-mastering in calesthio/generative-media-skills) into .agents/skills/audio-mixing-mastering in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add calesthio/generative-media-skills --skill audio-mixing-mastering -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audio-mixing-mastering, .gemini/skills/audio-mixing-mastering, .github/skills/audio-mixing-mastering and .opencode/skills/audio-mixing-mastering in your project.
Going by SKILL.md and its folder, Audio Mixing Mastering needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.
SKILL.md names 10 domains. As links in the text: itu.int, tech.ebu.ch, atsc.org, aes.org, support.spotify.com, podcasters.apple.com, support.google.com, partnerhelp.netflixstudios.com, w3.org and ffmpeg.org. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Audio Mixing Mastering is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 7k tokens (SKILL.md is roughly 28k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Audio Mixing Mastering: Multimedia Accessibility (Owl-Listener/inclusive-design-skills, 105 stars), Contentclaw (LeoYeAI/openclaw-master-skills, 2.2k stars), AI Fomo (vincelele/ai-fomo-skills, 248 stars) and Z Qwen Audio Studio (tjxj/z-skills, 548 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
calesthio (a GitHub user) maintains it in calesthio/generative-media-skills, which has 197 GitHub stars. The repository holds 26 skills in this directory. The repository was last updated on July 14, 2026.
Source: calesthio/generative-media-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.