AI Marketing Videos
NeverSight/learn-skills.dev
Create AI marketing videos for ads, promos, product launches, and brand content.
Generate podcast clip visualization video prompts for Seedance 2.0 on Higgsfield.
$ npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install rediumvex/ai-video-generator-claude seedance-podcast-visual --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/10-podcast-visual .claude/skills/seedance-podcast-visual && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "seedance-podcast-visual" agent skill from https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visual into .claude/skills/seedance-podcast-visual/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seedance-podcast-visual", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visualType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install rediumvex/ai-video-generator-claude seedance-podcast-visual --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/10-podcast-visual .agents/skills/seedance-podcast-visual && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "seedance-podcast-visual" agent skill from https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visual into .agents/skills/seedance-podcast-visual/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seedance-podcast-visual", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install rediumvex/ai-video-generator-claude seedance-podcast-visual --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/10-podcast-visual .cursor/skills/seedance-podcast-visual && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "seedance-podcast-visual" agent skill from https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visual into .cursor/skills/seedance-podcast-visual/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seedance-podcast-visual", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/rediumvex/ai-video-generator-claude.git --path skills/10-podcast-visual--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install rediumvex/ai-video-generator-claude seedance-podcast-visual --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/10-podcast-visual .gemini/skills/seedance-podcast-visual && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "seedance-podcast-visual" agent skill from https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visual into .gemini/skills/seedance-podcast-visual/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seedance-podcast-visual", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install rediumvex/ai-video-generator-claude seedance-podcast-visualInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/10-podcast-visual .github/skills/seedance-podcast-visual && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "seedance-podcast-visual" agent skill from https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visual into .github/skills/seedance-podcast-visual/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seedance-podcast-visual", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install rediumvex/ai-video-generator-claude seedance-podcast-visual --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/10-podcast-visual .opencode/skills/seedance-podcast-visual && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "seedance-podcast-visual" agent skill from https://github.com/rediumvex/ai-video-generator-claude/tree/main/skills/10-podcast-visual into .opencode/skills/seedance-podcast-visual/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seedance-podcast-visual", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
seedance-podcast-visualGenerate podcast clip visualization video prompts for Seedance 2.0 on Higgsfield.
Seedance Podcast Visual is an agent skill from rediumvex/ai-video-generator-claude. Generate podcast clip visualization video prompts for Seedance 2.0 on Higgsfield. Use for podcast clip videos, audio-to-visual content, audiogram alternatives, podcast highlight reels, interview clip visuals, or any video that transforms audio content into engaging visual format. Triggers on podcast, audio clip, audiogram, interview clip, sound bite, audio visual, podcast video, episode highlight, podcast clip.
Its SKILL.md is about 6.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Media & Creative, covering AI video generation. It works with Seedance, Instagram and TikTok. The repository describes itself as: 10 Claude skills that generate studio-quality AI video prompts for Seedance 2.0 on Higgsfield. Viral hooks, SaaS demos, personal brand, faceless content, luxury aesthetic & more. The licence is MIT.
10 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit ffdad7d. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Seedance Podcast Visual loads about 6.9k tokens when it runs. Until then it costs about 110 tokens; SKILL.md has 3,911 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from rediumvex/ai-video-generator-claude at commit ffdad7d, republished under its MIT licence (© rediumvex). 3,911 words, ~6,869 tokens.
.claude/skills/seedance-podcast-visual/SKILL.md (or your agent's skills folder).Transform podcast audio into cinematic visual content using Seedance 2.0 on Higgsfield. This skill produces video prompts that replace static audiograms with storytelling-driven visual experiences built entirely from constructed imagery.
Primary inputs:
Audio file handling:
What you extract from audio before writing prompts:
| Old model (audiogram) | New model (podcast visual) |
|---|---|
| Show the waveform | Show what the words feel like |
| Static background image | Constructed cinematic environment |
| Speaker photo as thumbnail | Speaker reconstructed in scene |
| Generic brand colors | Lighting and atmosphere matched to tone |
| Passive viewing | Active emotional engagement |
| Optimized for "audio on" | Compelling even on mute |
The hook is the opening frame that stops the scroll. It must communicate emotion, intrigue, or tension before a single word is heard. Four proven structures:
Display the most provocative line from the clip as large kinetic text before audio begins. The text arrives with weight — not a gentle fade, but a hard cut or a push-in. The visual behind it is blurred or dark, forcing the text into full focus.
When to use: clips with a single devastating sentence, contrarian takes, counterintuitive statistics, direct challenges to conventional wisdom.
Visual execution in prompt: specify "bold white sans-serif typography slams onto dark background, camera holds for 1.5 seconds, then cuts to speaker close-up, shallow depth of field, background softly bokeh'd."
Open on the speaker's face at the moment of peak emotional expression — surprise, laughter, conviction, vulnerability — before any context is given. This creates a curiosity gap: the viewer needs to hear what caused that expression.
When to use: interview moments where a genuine reaction occurs, storytelling clips where the speaker relives something visceral, moments of realization or revelation.
Visual execution in prompt: specify "extreme close-up on speaker's face, caught mid-expression, eyes slightly wide, ambient room sound implied by environment, camera slowly eases back over 3 seconds to reveal setting."
Instead of showing the speaker at all, open with an environmental or abstract image that represents the core concept of the clip. A podcast about burnout opens on dying embers. A clip about compounding returns opens on a single drop rippling outward. The metaphor does expository work so the audio can focus on depth.
When to use: concept-heavy clips, philosophical discussions, any clip where the idea is more powerful than the person delivering it.
Visual execution in prompt: specify the metaphor object explicitly, its lighting, its motion quality, and a precise camera behavior (slow push, orbital, static hold with foreground element drifting through).
Not a functional audiogram waveform — instead, an artistic rendering of sound as visual sculpture. Particles forming and dissolving in rhythm with imagined speech cadence. Light bending through air as if vibrated by voice. Sound made beautiful, not informational.
When to use: music-adjacent podcasts, high-production brand content, moments where you want to foreground the craft of the medium itself.
Visual execution in prompt: specify particle behavior, color palette tied to the emotional register of the clip, and whether motion is rhythmic/predictable or fluid/organic. Avoid the word "waveform" — describe it as "acoustic particle field" or "resonant light diffusion."
The audio inspires a visual world that does not contain the speaker at all. Instead, abstract imagery — light, texture, particle systems, color gradients, fluid dynamics — evolves in response to the imagined emotional arc of the audio.
Core parameters:
Prompt elements to always include: dominant color palette, motion behavior (fluid, particle, crystalline, liquid, smoke), camera behavior (static, slow push, orbital), and whether the environment is finite (a room implied by light edges) or infinite (void space)
Construct a series of visuals that would, in a traditional documentary, accompany the audio as b-roll. Except here every frame is generated — no stock footage, no compromises. The b-roll tells the story of the words.
Core parameters:
Prompt elements to always include: specific environment (time of day, weather, geography implied), one or two key objects in frame, camera move, lighting source, color grade direction (film noir, golden hour, overcast flat light, neon-saturated).
Reconstruct the podcast conversation as if it were a filmed interview, split-screen between two constructed environments. Each speaker occupies a distinct visual space — differentiated by lighting color temperature, depth of field, and environmental detail — while remaining in visual dialogue with each other.
Core parameters:
Prompt elements to always include: panel ratio (50/50, 60/40, or dynamic shift), description of each environment, lighting scheme for each, camera behavior for each, whether there is any visual bleed or hard line between panels.
The words themselves become the visual. The transcript animates — letters forming, words scaling, phrases colliding, key terms expanding to fill frame. The background is subordinate to type.
Core parameters:
Prompt elements to always include: primary font style (condensed sans, geometric sans, slab serif, display), animation behavior per word type (verbs get thrust/motion, nouns get scale, punctuation gets pause), background treatment (solid, gradient, subtle texture, barely-there environmental photo), and transition behavior between lines.
Seedance 2.0 generates video from text, not from audio input directly. Syncing visual rhythm to audio content is a craft challenge solved at the prompting stage.
Map the cadence of the clip before writing the prompt. Fast-talking speakers with high information density need visuals with shorter hold times and more cuts implied. Slower, deliberate speakers — the kind who pause for effect — need visuals that can breathe with them.
Practical approach:
Identify the 2-3 words or phrases in the clip that carry the most weight. These become visual events — a push-in, a light shift, a particle burst, a text scale event. The prompt should describe these precisely.
Example language for prompts:
Silence in a podcast clip is intentional. Great speakers use pauses like punctuation. These moments are visual opportunities: a held frame, a slow-motion exhale, a breath of empty space in the composition. Do not cut away during a pause — hold, and let the tension live.
Prompt language for pauses: "camera holds completely static," "particle motion slows to near-still," "background light dims 15% and recovers," "single dust mote drifts through foreground."
Everything in a Seedance 2.0 podcast visual is constructed, not captured. This is not a limitation — it is the entire point. Camera behavior must be specified deliberately.
Core camera vocabulary for podcast visuals:
Composition principles:
Lighting is the single most powerful emotional signal in a constructed podcast visual. Specify it precisely.
A clean, professional environment that reads as controlled but not sterile. Two-point or three-point lighting setup. The subject is clearly lit; the background is separated but not overlit.
Prompt language: "soft key light from camera-left, warm 4800K, catchlight visible in eyes; subtle fill from camera-right at 1/3 power; dark grey background with faint rim separation from practical studio light; no hard shadows on face."
The subject or object exists in pure darkness, lit by a single motivated source. Extreme contrast. The background gives nothing — all attention is on the subject. This style works for any clip where the words are the only thing that matters.
Prompt language: "single motivated light source from above-left, 3200K tungsten warmth, subject is lit in isolation against black void, no ambient fill, deep shadow on right side of face, high contrast ratio."
Intimate, inviting, lived-in. Suggests a real conversation happening in a real space — a kitchen, a living room, a café booth. Light sources are practical: a window, a lamp, candles. Color temperature is warm (2700K–3200K). This is the lighting of trust.
Prompt language: "warm practical lamp at frame-right, 2700K, soft pool of light; secondary fill from large window implied off-frame left, cool daylight, low intensity; environment reads as domestic interior, bookshelves soft in background, overall low-contrast warmth."
One hard light. Everything else is darkness or very deep shadow. This is the lighting of revelation, of stakes, of honesty forced into the open. Use for clips where the speaker says something they'd normally not say publicly.
Prompt language: "single hard light source from above, 5500K, 45-degree angle to subject, no fill, deep shadows across lower face and neck, light falls off sharply, background completely unlit, high-drama chiaroscuro treatment."
When text appears in the video — quotes, speaker IDs, timestamps — it must be designed with the same care as the visual.
Full-frame typographic hold: the quote appears as the primary visual. Large, bold, white or cream text on dark background. The text itself is the frame. Used for the most powerful single sentence.
Lower-third persistent text: the quote runs as a lower-third during the speaker visual. Not subtitles — the quote is selected for impact, not transcription accuracy. Use a font weight heavier than body text. Maximum 10-12 words.
Word-by-word reveal: each word appears individually, timed to the rhythm of delivery. Prompt must specify animation behavior: "words appear with a hard cut at each syllable boundary, no fades, duration of each word matches speech pacing."
Floating quote with depth: the text appears to exist in three-dimensional space within the scene, as if the words are physically present. Prompt: "quote text appears to float in mid-air within the environment, slight parallax offset from background movement, letters cast soft shadow onto the space behind them."
Speaker name and show/platform title appear within the first 5 seconds and again at the end. Keep it subordinate to the quote unless brand-building is the primary goal.
Placement options:
Typography rule: speaker ID font should be same family as quote text, lower weight (regular vs. bold). Never two competing display fonts.
For podcast clips on TikTok/Reels, burned-in subtitles are expected:
For longer clips with multiple audio files, timestamp markers orient the viewer within the episode. These are subtle — a small badge in the upper-right, episode time in small sans-serif. They add credibility (this is a real episode worth finding) without competing with the primary visual.
Context: A founder describing the moment she decided to quit her corporate job. Single audio clip, 45 seconds. Quote: "I wasn't jumping. I was finally standing still."
Seedance 2.0 video prompt, 9:16 vertical, 45 seconds:
Opening frame: pure black void, single tungsten-warm spotlight descends from directly above, illuminating only the empty center of frame. Dust particles drift slowly through the beam. Camera is completely static. Hold 2 seconds.
At second 2: bold white condensed sans-serif text appears with a hard cut — "I wasn't jumping." — centered at mid-frame, massive scale, filling 60% of horizontal width. Text holds for 1.8 seconds with absolute stillness. No animation, no fade. The words have weight.
Hard cut to black. 0.5 second black.
Text reappears: "I was finally standing still." — same font, same scale, same position. Camera begins an imperceptibly slow push-in toward the text, barely perceptible over 3 seconds. Text dissolves.
At second 8: transition to extreme close-up on a woman's hands resting palms-down on a dark wooden desk, ring light catchlight visible on skin, warm 3200K key light from camera-left. Hands are still, completely still. Camera holds.
Slow pull-back over 12 seconds reveals the desk, a single sheet of paper face-down on the surface, pen beside it. The environment implies a home office — dark bookshelf edge visible, one small practical lamp providing ambient fill. The scene is intimate and decisive.
At second 28: the hands turn the paper over. Slow motion, 50% speed. We do not see what is written. Camera slowly rises from hands to the window behind the desk, soft late-afternoon light, city skyline slightly out of focus.
Final 8 seconds: frame holds on the window with golden-hour light. Speaker name appears lower-left in clean white regular-weight sans-serif: "— [Speaker Name]." below it in smaller type: "@showname | Episode 47." Fade to black.
Color grade: high contrast, warm shadows, teal-pulled highlights. Cinematic aspect, no vignette.
Context: Two-person podcast. Host and guest discussing why most startup pivots fail. Two audio clips — host question and guest response. Combined 75 seconds.
Seedance 2.0 video prompt, 16:9 widescreen, 75 seconds:
Opening: split-screen, hard vertical line at 50% horizontal. Left panel: host environment — modern home office, cool-toned ambient light, bookshelves visible in background out of focus, camera at eye level, static hold, subject framing leaves slight negative space on their right for visual breathing room. Right panel: guest environment — warmer, slightly softer ambience, practical desk lamp at frame-right, 2800K warmth, camera position slightly lower than eye-level suggesting the guest is more relaxed.
Both panels are simultaneously visible. The subjects are constructed, not real footage — described as follows:
Left panel subject: man, early 40s, lean posture, dark shirt, slightly leaning forward toward camera, expression attentive, one hand resting on desk, neutral background with soft depth separation. Light: key light from camera-left 4500K, minimal fill, clean catchlight.
Right panel subject: woman, mid-30s, natural posture, light-colored top, relaxed but engaged expression, hands occasionally gesture just below frame. Light: warm key from camera-right 3000K, softer shadow wrap, environment feels more inhabited.
At second 8: left panel pushes to 60% of screen width as host speaks. Right panel reduces to 40%, subject slightly defocused. No title card yet.
At second 22: guest begins speaking. Panels return to 50/50. Left panel camera begins barely perceptible slow zoom-in on host's listening expression. Right panel camera holds static on guest.
At second 35: key quote from guest appears as lower-third overlay on the right panel, white bold condensed sans-serif on semi-transparent dark strip: "Most pivots fail because the founders pivot away from pain instead of toward insight." Text persists for 6 seconds, then dissolves.
Final 15 seconds: both panels hold in conversation mode, cameras hold static. Full-screen end frame fades in from black: episode title centered in clean display type, both names and handles below in smaller weight, podcast logo upper-right corner. Dark background, minimal, professional.
Color grade: both panels intentionally slightly different — left cooler, right warmer — reinforcing their distinct perspectives. No filter effect, cinematic naturalism.
Context: Solo podcast host reflecting on burnout. Single audio clip, 60 seconds. Emotional, introspective delivery. Key line: "I kept shipping things I didn't believe in anymore."
Seedance 2.0 video prompt, 9:16 vertical, 60 seconds:
Opening 3 seconds: extreme close-up on a laptop keyboard, single finger resting on the Enter key, not pressing. Shallow depth of field, bokeh'd background implies a dim room. Camera holds static. The finger does not move. No other motion in frame.
Cut at second 3: interior of a home office at night. Blue-grey ambient light from a monitor out of frame to the left. Desk covered in post-it notes, cups, cables — organized chaos of overwork. Camera positioned low, looking up at the desk from below-counter angle. Slow dolly-left movement, 5-second duration, revealing more of the desk's surface and a wall of notes behind. One note reads "SHIP IT" in visible handwriting but slightly out of focus.
Cut at second 14: back to the keyboard close-up, now slightly wider — we see both hands, neither moving. Camera begins an extremely slow push-in toward the hands over 8 seconds.
Cut at second 22: new environment. A coffee cup, half-full, cold. Steam: none. Morning light implied from window light on the cup's surface — but the light is pale, low-energy overcast. Camera static. Cup is centered in frame. 4 second hold.
Cut at second 26: corridor of a generic office building, late at night, fluorescent lights, empty. Camera at end of corridor looking down its length. Camera does not move for 5 seconds, then begins an extremely slow push forward down the corridor.
At second 31: kinetic text overlaid on the corridor shot: words appear one at a time, white heavy-weight sans-serif, centered: "things" — "I" — "didn't" — "believe" — "in" — "anymore." Each word appears at the rhythm of the imagined speech delivery (slow, deliberate). Final word "anymore." holds for 2 seconds, then the entire corridor shot and text fades to a held black.
From second 45: return to laptop close-up, tightest framing yet — just the trackpad, one thumb resting on it. Stillness. Over 8 seconds, a single small light — implied notification — pulses twice on the screen reflection in the trackpad surface. The thumb does not move.
Final 7 seconds: fade to deep navy. Speaker handle appears in small white regular-weight type, center-frame. Below it: episode name. Below that: "full episode linked in bio." All text fades in slowly, no animation.
Color grade: desaturated, cool blue-green shadows, low contrast in highlights. The visual language of exhaustion. No warmth until there is something to be warm about — and in this clip, there isn't.
Every visual element must serve the emotional content of the audio — no decorative shots without emotional purpose.
Never use the word "waveform" in a prompt. A waveform is a technical representation of audio data. A Seedance 2.0 podcast visual is a cinematic translation of meaning. Use language like "acoustic particle field," "resonant light behavior," "sound-reactive atmospheric texture" — or, better, describe what the emotion looks like rather than what the sound looks like.
Specify timing in seconds. Seedance 2.0 prompts that include "at second X" guidance produce more coherent visual sequences than prompts that describe only static states. Map every visual event to a time anchor.
One dominant visual per scene. Avoid prompts that describe three competing visual elements in a single shot. The viewer can hold one thing at a time. Complexity comes from sequencing, not from stuffing each frame.
Write the pause. Every description of motion must be balanced by a description of stillness. "Camera holds completely static for 4 seconds" is as important as any camera movement. Stillness communicates confidence.
Match color temperature to emotional register. This is not a suggestion — it is a requirement. Cool temperatures (5000K+) for analytical, distanced, or revelatory content. Warm temperatures (under 3200K) for intimate, vulnerable, nostalgic content. Mixed temperatures for conflict or unresolved tension.
Typography is visual, not verbal. When including quote text in prompts, describe its physical appearance (scale, weight, position, animation behavior) as precisely as you describe a person's expression. The font is a costume; the animation is a performance.
Audio files inform the prompt; they do not dictate it. The visual narrative should be coherent on mute. If the video requires the audio to make sense, the prompt has not done its job. Test by asking: "would a viewer who cannot hear this clip understand its emotional content within 5 seconds?"
End on identity. The last 5-10 seconds of every podcast visual should establish the creator's identity (name, handle, show, episode). This is not vanity — it is the conversion moment. The viewer has just felt something; tell them where to find more.
Multi-clip sequencing. When using 2 or 3 audio files, each clip gets its own visual arc (opening hook, development, resolution), but all arcs share a unified color grade and typographic system. Consistency across clips signals craft; variation within clips signals dynamism.
© rediumvex, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/10-podcast-visual of rediumvex/ai-video-generator-claude.
Open the folder on GitHubat commit ffdad7d
Seedance Podcast Visual next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Seedance Podcast Visual this skillrediumvex/ai-video-generator-claude | 409 | — | ~6.9k | Automated safety check: Pass | MIT | |
| AI Marketing VideosNeverSight/learn-skills.dev | 217 | 3 repos | ~2.1k | Automated safety check: Pass | None | |
| Pexo Agentpexoai/pexo-skills | 804 | — | ~5k | Automated safety check: Pass | MIT-0 | |
| Seedance Social Hookbeshuaxian/higgsfield-seedance2-jineng | 952 | — | ~21k | Automated safety check: Pass | None | |
| Seedance Social Hookbeshuaxian/higgsfield-seedance2-jineng | 952 | — | ~20k | Automated safety check: Pass | None | |
| Pexo Agentaiskillstore/marketplace | 433 | — | ~3.6k | Automated safety check: Pass | None |
NeverSight/learn-skills.dev
Create AI marketing videos for ads, promos, product launches, and brand content.
pexoai/pexo-skills
AI video generation skill with auto model selection across Seedance 2, Kling 3.0, HappyHorse, and 10+ models.
beshuaxian/higgsfield-seedance2-jineng
Generate viral social media hook video prompts for TikTok, Instagram Reels, and YouTube Shorts using Seedance 2.0 on Higgsfield.
beshuaxian/higgsfield-seedance2-jineng
使用 Seedance 2.0(Higgsfield)为 TikTok、Instagram Reels 和 YouTube Shorts 生成病毒式社交媒体钩子视频提示。在用户想要创建令人停下滚动的钩子、病毒式短视频、引人注目的开场白、TikTok 内容、Reels 内容、Shorts 内容或任何社交媒体优化视频时使用。在以下情况下触发:社交媒体视频、TikTok、Instagram…
aiskillstore/marketplace
AI video generation skill with auto model selection across Seedance 2, Kling 3.0, HappyHorse, and 10+ models.
beshuaxian/higgsfield-seedance2-jineng
使用 Seedance 2.0(Higgsfield)为电子商务产品生成广告视频提示。在用户想要产品广告、电子商务视频、产品展示、开箱、产品演示、购物广告、时尚广告、美容广告、食品广告或任何用于在线销售的商业产品视频时使用。触发条件:产品广告、电子商务、产品展示、Amazon视频、Shopify广告、Instagram店铺、TikTok店铺、产品商业广告、时尚视频、美容广告、食品广告、产品演示、…
rediumvex/ai-video-generator-claude
Generate before-and-after transformation video prompts for Seedance 2.0 on Higgsfield.
rediumvex/ai-video-generator-claude
Generate faceless content video prompts for Seedance 2.0 on Higgsfield.
rediumvex/ai-video-generator-claude
Generate personal brand and founder story video prompts for Seedance 2.0 on Higgsfield.
rediumvex/ai-video-generator-claude
Generate SaaS product launch and software demo video prompts for Seedance 2.0 on Higgsfield.
rediumvex/ai-video-generator-claude
Generate scroll-stopping viral hook video prompts for Seedance 2.0 on Higgsfield.
rediumvex/ai-video-generator-claude
Generate AI avatar and digital persona video prompts for Seedance 2.0 on Higgsfield.
Categories
Generate podcast clip visualization video prompts for Seedance 2.0 on Higgsfield. Seedance Podcast Visual is an agent skill from rediumvex/ai-video-generator-claude.0 on Higgsfield.
Seedance Podcast Visual fits situations like: podcast clip videos; audio-to-visual content; audiogram alternatives; podcast highlight reels.
Run `npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a claude-code`. Or copy the skill folder (skills/10-podcast-visual in rediumvex/ai-video-generator-claude) into .claude/skills/seedance-podcast-visual in your project. Claude Code loads it when a task matches its description.
Run `npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a codex`. Or copy the skill folder (skills/10-podcast-visual in rediumvex/ai-video-generator-claude) into .agents/skills/seedance-podcast-visual in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add rediumvex/ai-video-generator-claude --skill seedance-podcast-visual -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/seedance-podcast-visual, .gemini/skills/seedance-podcast-visual, .github/skills/seedance-podcast-visual and .opencode/skills/seedance-podcast-visual in your project.
SKILL.md names no scripts, command-line tools or credentials: Seedance Podcast Visual is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Seedance Podcast Visual is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 6.9k tokens (SKILL.md is roughly 27k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Seedance Podcast Visual: AI Marketing Videos (NeverSight/learn-skills.dev, 217 stars), Pexo Agent (pexoai/pexo-skills, 804 stars), Seedance Social Hook (beshuaxian/higgsfield-seedance2-jineng, 952 stars) and Seedance Social Hook (beshuaxian/higgsfield-seedance2-jineng, 952 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
rediumvex (a GitHub user) maintains it in rediumvex/ai-video-generator-claude, which has 409 GitHub stars. The repository holds 10 skills in this directory. The repository was last updated on August 11, 2026.
Source: rediumvex/ai-video-generator-claude on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.