HyperFrames Media Use
heygen-com/hyperframes
Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.
Listen to generated music by measuring it and by reading it as sheet music: render Strudel code or record a Web Audio page through the real engine in headless Chromium, then read spectrograms, pYIN…
$ npx skills add oaustegard/claude-skills --skill listening-to-music -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install oaustegard/claude-skills listening-to-music --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/oaustegard/claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/listening-to-music .claude/skills/listening-to-music && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "listening-to-music" agent skill from https://github.com/oaustegard/claude-skills/tree/main/listening-to-music into .claude/skills/listening-to-music/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "listening-to-music", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/oaustegard/claude-skills/tree/main/listening-to-musicType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add oaustegard/claude-skills --skill listening-to-music -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install oaustegard/claude-skills listening-to-music --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/oaustegard/claude-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/listening-to-music .agents/skills/listening-to-music && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "listening-to-music" agent skill from https://github.com/oaustegard/claude-skills/tree/main/listening-to-music into .agents/skills/listening-to-music/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "listening-to-music", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add oaustegard/claude-skills --skill listening-to-music -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install oaustegard/claude-skills listening-to-music --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/oaustegard/claude-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/listening-to-music .cursor/skills/listening-to-music && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "listening-to-music" agent skill from https://github.com/oaustegard/claude-skills/tree/main/listening-to-music into .cursor/skills/listening-to-music/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "listening-to-music", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/oaustegard/claude-skills.git --path listening-to-music--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add oaustegard/claude-skills --skill listening-to-music -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install oaustegard/claude-skills listening-to-music --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/oaustegard/claude-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/listening-to-music .gemini/skills/listening-to-music && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "listening-to-music" agent skill from https://github.com/oaustegard/claude-skills/tree/main/listening-to-music into .gemini/skills/listening-to-music/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "listening-to-music", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install oaustegard/claude-skills listening-to-musicInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add oaustegard/claude-skills --skill listening-to-music -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/oaustegard/claude-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/listening-to-music .github/skills/listening-to-music && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "listening-to-music" agent skill from https://github.com/oaustegard/claude-skills/tree/main/listening-to-music into .github/skills/listening-to-music/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "listening-to-music", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add oaustegard/claude-skills --skill listening-to-music -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install oaustegard/claude-skills listening-to-music --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/oaustegard/claude-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/listening-to-music .opencode/skills/listening-to-music && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "listening-to-music" agent skill from https://github.com/oaustegard/claude-skills/tree/main/listening-to-music into .opencode/skills/listening-to-music/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "listening-to-music", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
listening-to-musicListen to generated music by measuring it and by reading it as sheet music: render Strudel code or record a Web Audio page through the real engine in headless Chromium, then read spectrograms, pYIN…
Listening To Music is an agent skill from oaustegard/claude-skills. Listen to generated music by measuring it and by reading it as sheet music: render Strudel code or record a Web Audio page through the real engine in headless Chromium, then read spectrograms, pYIN pitch lines, chroma, roughness and semitone-rub meters; engrave notes as a score with beat-by-beat harmony; transcribe a melody from audio; read MusicXML/MIDI/ABC or a score image into notes and into Strudel code. Use to check whether generated music sounds right or better (before/after a change), to match a reference…
Its SKILL.md is about 2.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 17 other files, including scripts (for example `CHANGELOG.md`, `scripts/clashes.py` and `scripts/compare_notes.py`).
It sits in Media & Creative, covering Transcription. The repository describes itself as: My collection of Claude skills. The licence is MIT.
4 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 90b0f1b. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 15 files in scripts/ (Python and JavaScript), which the agent can run.
Shell commands in SKILL.md call:
python3nodepipFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Listening To Music loads about 2.5k tokens when it runs. Until then it costs about 255 tokens; SKILL.md has 1,232 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from oaustegard/claude-skills at commit 90b0f1b, republished under its MIT licence (© oaustegard). 1,232 words, ~2,500 tokens.
.claude/skills/listening-to-music/SKILL.md (or your agent's skills folder). This skill also uses 16 other files; get the full folder from GitHub.Claude hears nothing, but it can render audio in the real engine and measure it. Set up the loop render → look → measure → change one thing → re-render, and keep a table of the numbers per version. A version is better when a number you chose beforehand moves by more than the take-to-take noise, and the spectrogram shows the same thing.
Requirements: Node with playwright and Chromium (preinstalled in Claude Code on the web),
Python with librosa soundfile scipy matplotlib pillow, and for sheet music
music21 verovio mir_eval cairosvg
(pip install --break-system-packages librosa music21 verovio mir_eval cairosvg). The first render installs @strudel/web into
~/.cache/listening-to-music/. All scripts live in scripts/; run them with --help or read the
docstring for options.
S=/path/to/listening-to-music/scripts
node $S/render_strudel.mjs loop.js --out v1.wav --seconds 40 --warmup 4 # Strudel code
node $S/record_page.mjs page.html --click "#power" --out v1.wav --seconds 40 --warmup 8 \
--query "station=house" --eval "window.__fm.setSeed(7)" # any Web Audio pageBoth print peak level and flag clipping. A peak above 1.0 means the listener hears hard clipping; fix levels before judging anything else.
For a before/after comparison, fix the randomness (a seed hook via --eval) so both versions
play the same material, and record each version twice.
python3 $S/spectrogram.py v1.wav --out v1.png --from-onset --pyin-floor G3 --label v1Then open v1.png with the image viewer. The top panel is a note-axis CQT spectrogram with the
pYIN lead line, and the bottom panel is chroma. Compare it with the reference picture or the
previous version side by side. The .npz next to it holds the numbers for step 3.
| Question | Tool | Reads |
|---|---|---|
| Do simultaneous notes rub? (discordant chords) | render_strudel.mjs --events → clashes.py | minor 2nd/9th overlap per cycle, by layer pair; worst moments |
| Do they rub in the recording? | rubs.py, --pair for replicated takes | share of tonal peak energy in semitone/minor-9th pairs |
| Is it rougher or smoother overall? | roughness.py, --pair for replicated takes | Sethares roughness median/p90 |
| Does it match a reference image? | reference.py ticks/decode/compare | per-semitone profile, melody agreement, overtone offsets, chroma agreement |
| Is the melody audible over the mix? | spectrogram.py pYIN voiced %, reference.py compare melody % |
The clash scan is exact for the notes as written. rubs.py confirms them in the audio, where
release tails and reverb add overlaps the note data does not show. Record harmony checks with
percussion and noise textures muted in both versions: drum partials form their own peaks and hold
a full mix at a rub share near 0.28 whatever the chords do. Use --warmup, 40 s or more, and two
takes per version; --pair prints the take-to-take spread and says when a change is inside it.
Summed roughness (roughness.py) is a coarse overall measure. On Strudel FM (2026-09-24) it
could not tell the harmony fix apart from take-to-take noise in sleep and house. On the same
takes, the rub meter measured sleep −51%, ambient −33%, lofi −20% and house −13%, with spreads of
0.006–0.015. A loud bass drone dominates the roughness normalisation and hides a pad's rubs.
Don't treat an unmoved roughness figure as proof that nothing changed.
Notation is the other way to hear. A score shows voicings, clusters (noteheads pushed sideways are seconds), register and rhythm at a glance, and Claude reads clean engraving accurately. All these scripts share one events JSON (layer, t0, t1, midi; times in bars).
node $S/render_strudel.mjs loop.js --events ev.json --cycles 8 # notes as written
python3 $S/score.py ev.json --out score.png --layers LEAD,KEYS,BASS --harmony # engrave + analyse
python3 $S/transcribe.py take.wav --out mel.json --bpm 115 # notes as heard (melody)
python3 $S/compare_notes.py ev.json mel.json --ref-layer LEAD --cand-layer melody --align 1
python3 $S/read_score.py tune.musicxml --out ref.json # .mxl .mid .abc .krn, corpus:
python3 $S/to_strudel.py ref.json --out tune.js # score -> Strudel codescore.py --harmony prints each chord change with its pitches, music21's chord name, a Roman
numeral in the detected key, and a rub flag. On Strudel FM the same four lofi bars went from 15
of 28 chord changes with a rub to 0 of 28 after the voicing fix, and the engraving showed why:
the old Fmaj7 had its E and F side by side.
Measured on 2026-09-25 (compare_notes.py, onset within 0.05 bar and pitch within 50 cents):
| Path | Test | F1 |
|---|---|---|
| read_score → to_strudel → render events | Bach BWV 66.6, 4 parts, 163 notes | 100% every part |
| transcribe.py, one line | same chorale, soprano alone, piano, not used for tuning | 87% |
transcribe.py --no-split | synth lead with delay echoes | 82% (65% with splitting) |
| transcribe.py | full four-part chorale | 0%: monophonic only |
| OMR (oemer 0.1.8) | clean engraved soprano line | 6%; key read as flats, quarters as whole notes |
| Claude reading the image → ABC | unseen chorale, soprano + bass, 38 notes | 100% (one clean engraving) |
So:
K:, M:, L:, one V: per staff), convert it with
read_score.py, re-engrave with score.py, and compare the two pictures bar by bar before
using the notes. Use oemer only as a rough first pass, and never trust it unchecked.--no-split for leads with delay or echo.compare_notes.py --align N. It
searches shifts up to N bars and prints the one used.Change one parameter or rule at a time, and after each change render and rerun the same measurements. A metric can reward the wrong thing, so check it against the spectrogram each round.
.delaytime("3/8") means "3, over 8
cycles", so the delay was 3 s (clamped to 1 s). Pass a number: .delaytime(0.39).
.gain("0.9, 0.3") on a stacked pattern is itself a stack, so every hit fired twice and the
drums peaked at 2.9× full scale. Give each layer its own scalar gain.supersaw, …) need a secure context. Pages served from
http:// get no audioWorklet, and those synths are silent while everything else plays. The
scripts serve local pages from https://local.test/ for this reason. initStrudel also loads
the worklets only on the first document click, so the scripts click.reference.py compare (the overtone rows against the control rows) before voicing a pad
from them. A misread gave a Gmaj7 that the source never played.chordify() discards pitch spelling, and a Key's inferred tonic
transposes to the wrong letter (a leading tone of F instead of E♯ in F♯ minor). notes.spell
re-spells after every chordify from the key's own scale plus the raised 6th and 7th. Without
that, the score reads as wrong notes.reference.py ticks rather than by eye; a half-bin error splits every note across two
semitones.© oaustegard, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 16 other files (scripts) in listening-to-music of oaustegard/claude-skills.
Open the folder on GitHubat commit 90b0f1b
Listening To Music next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Listening To Music this skilloaustegard/claude-skills | 150 | — | ~2.5k | Automated safety check: Pass | MIT | |
| HyperFrames Media Useheygen-com/hyperframes | 60k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | |
| Native Subtitle Quote Imagechengyi-ai/native-subtitle-quote-image | 2.6k | — | ~2.4k | Automated safety check: Pass | MIT | |
| Edu Chem Videowy51ai/edulab | 1.4k | — | ~2.1k | Automated safety check: Notes | Apache-2.0 | |
| Transcription Memory ReconstructionNxcoreAI/EverRoom | 3k | — | ~714 | Automated safety check: Pass | Custom licence | |
| Edu Math Videowy51ai/edulab | 1.4k | — | ~2.5k | Automated safety check: Notes | Apache-2.0 |
heygen-com/hyperframes
Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.
chengyi-ai/native-subtitle-quote-image
将本地视频或用户有权处理的在线视频,经过来源获取、文字稿定位、选题选句、精确取帧、紧凑裁切、拼图和逐张质检,制作成 3:4 或保留画面原比例的视频字幕长图。支持两种明确分开的输出:保留画面内已烧录字幕的原生字幕模式,以及把已审核的时间点与台词绘制到真实视频帧上的脚本字幕模式。用户要求原生字幕截图、字幕帧拼图、YouTube…
wy51ai/edulab
A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a chemistry problem (化学题: 氧化还原配平 双线桥 电子守恒, 物质的量计算, 化学平衡 三段式 平衡常数 转化率 反应速率, 离子反应, 电化学, 溶液 滴定…
NxcoreAI/EverRoom
Reconstruct a complete, searchable memory from an untrusted meeting or conversation transcript.
wy51ai/edulab
A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a math problem (数学题, geometry, algebra, functions, motion/行程 problems), from a problem screenshot…
JetBrains/skills
Transcribe audio files to text with optional diarization and known-speaker hints.
oaustegard/claude-skills
Builds interactive Vega-Lite charts from uploaded data: analyzes the fields, picks five to ten fitting chart types, and produces a React artifact with the data embedded inline.
oaustegard/claude-skills
Builds self-contained single-file HTML pages such as reports, decks, postmortems, flowcharts and prototypes from a small spec using a bundled Python composer and templates.
oaustegard/claude-skills
Routes, triages, flags and rates a piece of text with a probability for every option: which department or queue a ticket goes to, which intent a message expresses, whether a yes/no condition holds…
oaustegard/claude-skills
Rewrites model-sounding prose into plain technical writing and checks that every claim survives, for PR text, docs, commit messages and similar drafts.
oaustegard/claude-skills
Guides building standards-based Preact apps with native-first choices, HTM syntax, import maps and vendored ESM, from single-file demos to larger builds.
oaustegard/claude-skills
Deprecated sampler that captures short windows of the Bluesky firehose, clusters trending terms and builds an HTML report; replaced by the browsing-bluesky skill.
Categories
Listen to generated music by measuring it and by reading it as sheet music: render Strudel code or record a Web Audio page through the real engine in headless Chromium, then read spectrograms, pYIN…. Listening To Music is an agent skill from oaustegard/claude-skills. Listen to generated music by measuring it and by reading it as sheet music: render Strudel code or record a Web Audio page through the real engine in headless Chromium, then read spectrograms, pYIN pitch lines, chroma, roughness and semitone-rub meters; engrave notes as a score with beat-by-beat harmony; transcribe a melody from audio; read MusicXML/MIDI/ABC or a score image into notes and into Strudel code.
Listening To Music fits situations like: check whether generated music sounds right; better (before/after a change); match a reference track; spectrogram image.
Run `npx skills add oaustegard/claude-skills --skill listening-to-music -a claude-code`. Or copy the skill folder (listening-to-music in oaustegard/claude-skills) into .claude/skills/listening-to-music in your project. Claude Code loads it when a task matches its description.
Run `npx skills add oaustegard/claude-skills --skill listening-to-music -a codex`. Or copy the skill folder (listening-to-music in oaustegard/claude-skills) into .agents/skills/listening-to-music in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add oaustegard/claude-skills --skill listening-to-music -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/listening-to-music, .gemini/skills/listening-to-music, .github/skills/listening-to-music and .opencode/skills/listening-to-music in your project.
Going by SKILL.md and its folder, Listening To Music needs Python and JavaScript for the scripts in its folder and the command-line tools its instructions call (python3, node and pip). Our summary lists: Python 3; Node.js.
SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Listening To Music is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.5k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Listening To Music: HyperFrames Media Use (heygen-com/hyperframes, 60k stars), Native Subtitle Quote Image (chengyi-ai/native-subtitle-quote-image, 2.6k stars), Edu Chem Video (wy51ai/edulab, 1.4k stars) and Transcription Memory Reconstruction (NxcoreAI/EverRoom, 3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
oaustegard (a GitHub user) maintains it in oaustegard/claude-skills, which has 150 GitHub stars. The repository holds 67 skills in this directory. The repository was last updated on October 9, 2026.
Source: oaustegard/claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.