Agent skill

Talking Head And Piece To Camera

by social-media-skills in social-media-skills/skills

The on-camera delivery craft — helping a real human film themselves talking to a lens and look like themselves doing it.

MITAuto-check passedMedia & Creative

Install Talking Head And Piece To Camera

skills CLI
$ npx skills add social-media-skills/skills --skill talking-head-and-piece-to-camera -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install social-media-skills/skills talking-head-and-piece-to-camera --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/social-media-skills/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/talking-head-and-piece-to-camera .claude/skills/talking-head-and-piece-to-camera && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
talking-head-and-piece-to-camera
GitHub stars
134
Token cost
~2.2k tokens
SKILL.md length
1,021 words
Files
6 (incl. references)
Skills in repo
106
Repo updated
First seen
Licence
MIT

At a glance

The on-camera delivery craft — helping a real human film themselves talking to a lens and look like themselves doing it.

  • Works in 2 steps: brand-profile + voice-builder — who's… → short-form-video-script (or…
  • Someone wants a talking head video
  • SKILL.md covers The POV: presence beats…, Read these first, The framework: TAKES and The reality (verify-quarterly), plus 5 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Talking Head And Piece To Camera is an agent skill from social-media-skills/skills. The on-camera delivery craft — helping a real human film themselves talking to a lens and look like themselves doing it. Use when someone wants a "talking head video" or "piece to camera," says "film myself" or "I look stiff on camera," asks about a teleprompter, framing, lighting, audio, or retakes, or wants to batch-film videos. Uses the TAKES framework. Phone-first: gear is almost never the bottleneck. Reads brand-profile + voice-builder first; takes its script from short-form-video-script (that writes it…

Its SKILL.md is about 2.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including reference files (for example `evals/evals.json`, `references/batch-filming-and-recipes.md` and `references/scope-and-connections.md`).

It sits in Media & Creative, covering Video production. It works with HeyGen. The repository describes itself as: 106 social media skills for AI agents - strategy, writing, video, design, platform growth, publishing, and analytics. Works with Claude, Cursor, OpenClaw, Hermes & 40+ agents. The licence is MIT.

When your agent uses it

  • Someone wants a talking head video
  • Piece to camera
  • Says film myself
  • I look stiff on camera

Example prompts

  • “talking head video”
  • “piece to camera,”
  • “film myself”
  • “/talking-head-and-piece-to-camera”

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. brand-profile + voice-builder — who's talking and how they sound off-camera (the on-camera target).
  2. short-form-video-script (or youtube-long-form for long pieces) — the script/beats being delivered;

What it can do on your machine

Read from SKILL.md and the folder at commit 6e30eeb. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Talking Head And Piece To Camera loads about 2.2k tokens when it runs, and up to ~6.4k if it reads all its reference files. Until then it costs about 255 tokens; SKILL.md has 1,021 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~255
When it runs · the whole SKILL.md, loaded when a task matches
~2.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~6.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from social-media-skills/skills at commit 6e30eeb, republished under its MIT licence (© social-media-skills). 1,021 words, ~2,170 tokens.

Download SKILL.mdSave it as .claude/skills/talking-head-and-piece-to-camera/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
talking-head-and-piece-to-camera
description
The on-camera delivery craft — helping a real human film themselves talking to a lens and look like themselves doing it. Use when someone wants a "talking head video" or "piece to camera," says "film myself" or "I look stiff on camera," asks about a teleprompter, framing, lighting, audio, or retakes, or wants to batch-film videos. Uses the TAKES framework. Phone-first: gear is almost never the bottleneck. Reads brand-profile + voice-builder first; takes its script from short-form-video-script (that writes it, this delivers it). The agent coaches setup + delivery, formats prompter/beat-map scripts, and plans batch days; the HUMAN films and picks the take (the agent cannot see footage); WoopSocial publishes the finished file. Camera-shy? Route honestly to heygen/synthesia or faceless formats. Never fabricates "that take looks great." Distinct from scripting-and-storyboarding (the shoot plan), heygen/synthesia (avatars), and captions-and-clipping/capcut/descript (the edit).
version
1.0.0

talking-head-and-piece-to-camera

The on-camera delivery craft — tape the setup, anchor the map (not the lines), kick the first 3 seconds, embrace the retake rules, stack the batch. The script comes from short-form-video-script; the human films and picks the take; WoopSocial publishes the finished file.

The POV: presence beats polish, and the phone in your pocket is enough

A talking head works because a real face builds parasocial trust an avatar can't (that's exactly why synthesia routes trust-led founder content here). Three truths most first-timers get backwards. First, gear is not the bottleneck — a phone at eye level, facing a window, with a cheap lav mic outperforms an expensive camera set up wrong; viewers forgive soft video and never forgive bad audio. Second, reading kills it — memorize the map (the beats), not the lines; a word-for-word read shows in the eyes, and a slightly imperfect riff reads as human. Third, the good-enough take ships — take 4 is usually worse than take 2 because energy decays faster than delivery improves; perfectionism is a retention strategy for exactly nobody. Deliver 20% more energy than feels natural, talk to one person, and publish the take where you sound like yourself.

Read these first

  1. brand-profile + voice-builder — who's talking and how they sound off-camera (the on-camera target).
  2. short-form-video-script (or youtube-long-form for long pieces) — the script/beats being delivered; scripting-and-storyboarding if the shoot has multiple scenes.

The framework: TAKES

(Depth: references/the-takes-framework.md.)

  • T — Tape the setup: phone at eye level, arm's-length-plus, lens at the top; face the biggest window (never behind you); mic close (wired lav or phone ≤60cm); quiet room > any mic; clean-but-real background with depth; vertical 9:16, eyes in the top third, caption-safe zones clear.
  • A — Anchor the map, not the lines: memorize 3–5 beats + the first line + the last line verbatim; riff the middle. Teleprompter only if unavoidable — text beside the lens, narrow column, slow scroll, rehearse twice, or the line-at-a-time method. Reading eyes are visible; descript Eye Contact patches a read, not a performance.
  • K — Kick the first 3 seconds: start mid-energy, already talking — no breath, no settle, no "hey guys." Say the hook fresh, first, every session. Smile-then-speak; hands visible; deliver to ONE person behind the lens.
  • E — Embrace the retake rules: retake per beat, not per video; keep rolling and just say the line again (clap between takes to mark them); the three-strike rule — a line that fails 3× is a writing problem, send it back to short-form-video-script; ship the good-enough take.
  • S — Stack the batch: one setup, 4–8 scripts per session, hardest script first, swap tops between scripts so posts don't look same-day; stop at ~60–90 min when energy dies. Plan with batch-content-plan / content-calendar.

The reality (verify-quarterly)

Any recent phone shoots 4K that out-resolves every social feed; audio drives perceived quality more than image (creator consensus — attribute); a below-eye lens reads as looming, backlit windows silhouette you; on-camera energy reads ~20% flatter than it feels (broadcast coaching convention); take quality typically peaks by take 2–3 then decays with energy; batch sessions fade after ~60–90 minutes — directional, attribute, verify-quarterly. Full figures + phone-first setup specifics: references/talking-head-2026-reality.md. Batch-day recipe, setup recipes (desk / walking / car), and camera-shy on-ramps: references/batch-filming-and-recipes.md.

Honest scope (never violate)

  • The agent coaches setup and delivery, formats the script as a beat map or prompter text, writes shot lists and batch plans, and gives a self-review checklist. The human films, performs, and picks the take. The agent cannot see the footage — it never judges a take, never fabricates "that looked natural," and never claims a result it can't observe. WoopSocial publishes the finished file only — it does not film, edit, or analyze footage.
  • Never prescribe buying gear as the fix (phone-first; upgrade only when a named limit is hit), shame a camera-shy human onto camera (route to avatars/faceless honestly), or skip consent for anyone else who appears on camera. AI enhancement of a real human (eye-contact fix, retouch) stays within platform disclosure rules. (Full scope: references/scope-and-connections.md.)
Show full SKILL.md (367 more words)Show less

Edge cases (handle honestly)

  • Camera-shy / won't film: legitimate. Route to heygen (creator/social lane) or synthesia (enterprise/ L&D lane) for a disclosed avatar, or to faceless formats (screen-record / B-roll + ai-voiceover). Offer the gentle on-ramp — voice-only first, then hands/desk shots, then face — but never pressure.
  • Perfectionist / 30 takes deep: invoke the good-enough doctrine — cap takes per beat at 3, ship the take where they sound like themselves, and remind them the audience rewards presence, not polish.
  • "Watch my take and tell me it's good": can't — no eyes on footage. Hand over the self-review checklist (hook lands on mute? energy? eyes on lens? audio clean?) and let the human verdict stand.

Distinct from its siblings (route correctly)

talking-head-and-piece-to-camera (this) = the human filming/delivery craft · short-form-video-script = the script this delivers (pair) · scripting-and-storyboarding = the multi-scene shoot plan (this is the shoot-day performance) · heygen / synthesia = synthetic presenters when the human can't/won't film · captions-and-clipping / capcut / descript = the edit after the shoot (descript's Eye Contact patches a read; it doesn't replace delivery) · livestream-and-realtime = live to-camera (no retakes) · ai-voiceover = voice without a face.

Where this connects

Reads first: brand-profile + voice-builder. Takes the script from: short-form-video-script (or youtube-long-form), the plan from scripting-and-storyboarding, batch slots from batch-content-plan + content-calendar. Feeds: captions-and-clipping / capcut / descript (the edit), opus-clip (clipping long pieces), cross-platform-repurposing. Routes away: avatars → heygen / synthesia. Publishes via: edited file → scheduling-and-queue → WoopSocial. Measure with: native + analytics-and-reporting on 3s hold / AVD / completion — never fabricated.

Definition of done

A filmed piece to camera delivered from a beat map (first + last lines verbatim, middle riffed), shot phone-first at eye level facing the light with clean close audio and a caption-safe 9:16 frame, opening mid-energy on the hook with no wind-up, retaken per beat under the three-strike rule and shipped at good-enough rather than sanded lifeless, batched (4–8 scripts, top swaps, ≤90 min) when volume is the goal; camera-shy humans routed honestly to heygen/synthesia or faceless formats; the human filmed and picked the take (the agent never judged footage it can't see, never fabricated praise, never prescribed gear as the fix); consent handled for anyone else in frame; the file edited via captions-and-clipping/capcut/descript and published via scheduling-and-queue → WoopSocial; measured on 3s hold / AVD / completion; and correctly distinguished from short-form-video-script, scripting-and-storyboarding, heygen/synthesia, and the editing skills.

© social-media-skills, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (references) in skills/talking-head-and-piece-to-camera of social-media-skills/skills.

  • SKILL.md
  • evals/evals.json
  • references/batch-filming-and-recipes.md
  • references/scope-and-connections.md
  • references/talking-head-2026-reality.md
  • references/the-takes-framework.md

Open the folder on GitHubat commit 6e30eeb

Compare with similar skills

Talking Head And Piece To Camera next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Talking Head And Piece To Camera compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Talking Head And Piece To Camera this skillsocial-media-skills/skills134—~2.2kAutomated safety check: PassMIT
HyperFrames Animationheygen-com/hyperframes60k3 repos~2.1kAutomated safety check: PassApache-2.0
Faceless Explainer Videoheygen-com/hyperframes60k3 repos~7.7kAutomated safety check: NotesApache-2.0
Figma to HyperFramesheygen-com/hyperframes60k3 repos~4.5kAutomated safety check: NotesApache-2.0
HyperFrames Video Entry Pointheygen-com/hyperframes60k3 repos~5.2kAutomated safety check: PassApache-2.0
Music to Videoheygen-com/hyperframes60k3 repos~4.7kAutomated safety check: NotesApache-2.0

Similar skills

  • HyperFrames Animation

    heygen-com/hyperframes

    Collects motion rules, scene blueprints, transitions and runtime adapters for HyperFrames video compositions, with GSAP as the default animation runtime.

    60k GitHub starsUsed in 3 repos~2.1k tokens
    Media & CreativeAuto-check passed
  • Faceless Explainer Video

    heygen-com/hyperframes

    Turns an article, notes or a topic brief into an explainer video whose visuals are invented per scene, built frame by frame in HyperFrames with no footage.

    60k GitHub starsUsed in 3 repos~7.7k tokens
    Media & CreativeAuto-check: notes
  • Figma to HyperFrames

    heygen-com/hyperframes

    Imports Figma assets, brand tokens, components and motion into a HyperFrames video composition, using the Figma REST API with a connector or native export for shaders.

    60k GitHub starsUsed in 3 repos~4.5k tokens
    Media & CreativeAuto-check: notes
  • HyperFrames Video Entry Point

    heygen-com/hyperframes

    Entry point for making, editing and rendering videos from HTML compositions with HyperFrames, routing each request to the right workflow.

    60k GitHub starsUsed in 3 repos~5.2k tokens
    Media & CreativeAuto-check passed
  • Music to Video

    heygen-com/hyperframes

    Turns a music track into a beat-synced HyperFrames video such as a lyric video, slideshow or kinetic promo, with any supplied images or clips cut onto the beat grid.

    60k GitHub starsUsed in 3 repos~4.7k tokens
    Media & CreativeAuto-check: notes
  • Short Form Edit

    nateherkai/hyperframes-student-kit

    Turn talking-head footage into a finished reel, YouTube Short, or short advertisement with curiosity-led openings, earned payoffs, story-driven cuts, transcript-synced motion graphics, moving…

    1.3k GitHub stars~5.3k tokensUpdated 12 days ago
    Media & CreativeAuto-check passed

More from social-media-skills/skills

All 106 skills in this repo
  • AI Image Editing

    social-media-skills/skills

    The AI image-editing router — inpainting/object removal, background removal, upscaling, outpainting, old-photo restoration, and retouch, routed task-first to the right engine.

    134 GitHub stars~2.1k tokensUpdated 9 days ago
    Auto-check passed
  • AI Music And Sound

    social-media-skills/skills

    The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos.

    134 GitHub stars~1.9k tokensUpdated 9 days ago
    Auto-check passed
  • AI Search Optimization

    social-media-skills/skills

    A skill your agent uses to get a brand and its content CITED and RECOMMENDED by AI answer engines — the GEO (Generative Engine Optimization) / AI-search-visibility skill.

    134 GitHub stars~2k tokensUpdated 9 days ago
    Auto-check passed
  • AI Video

    social-media-skills/skills

    The model-agnostic AI-video router and brief — the counterpart to image-prompt.

    134 GitHub stars~1.3k tokensUpdated 9 days ago
    Auto-check passed
  • AI Voiceover

    social-media-skills/skills

    The AI narration / voiceover mini-skill (ElevenLabs-led). An agent skill from social-media-skills/skills.

    134 GitHub stars~1.1k tokensUpdated 9 days ago
    Auto-check passed
  • Analytics And Reporting

    social-media-skills/skills

    Social media analytics and reporting — read native platform data honestly and turn it into next actions.

    134 GitHub stars~1.4k tokensUpdated 9 days ago
    Auto-check passed

Works with

Questions about Talking Head And Piece To Camera

What does Talking Head And Piece To Camera do?

The on-camera delivery craft — helping a real human film themselves talking to a lens and look like themselves doing it. Talking Head And Piece To Camera is an agent skill from social-media-skills/skills. The on-camera delivery craft — helping a real human film themselves talking to a lens and look like themselves doing it.

When should I use Talking Head And Piece To Camera?

Talking Head And Piece To Camera fits situations like: someone wants a talking head video; piece to camera; says film myself; I look stiff on camera.

How do I install Talking Head And Piece To Camera in Claude Code?

Run `npx skills add social-media-skills/skills --skill talking-head-and-piece-to-camera -a claude-code`. Or copy the skill folder (skills/talking-head-and-piece-to-camera in social-media-skills/skills) into .claude/skills/talking-head-and-piece-to-camera in your project. Claude Code loads it when a task matches its description.

How do I install Talking Head And Piece To Camera in Codex?

Run `npx skills add social-media-skills/skills --skill talking-head-and-piece-to-camera -a codex`. Or copy the skill folder (skills/talking-head-and-piece-to-camera in social-media-skills/skills) into .agents/skills/talking-head-and-piece-to-camera in your project. Codex loads it when a task matches its description.

Can I use Talking Head And Piece To Camera in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add social-media-skills/skills --skill talking-head-and-piece-to-camera -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/talking-head-and-piece-to-camera, .gemini/skills/talking-head-and-piece-to-camera, .github/skills/talking-head-and-piece-to-camera and .opencode/skills/talking-head-and-piece-to-camera in your project.

What does Talking Head And Piece To Camera need to run?

SKILL.md names no scripts, command-line tools or credentials: Talking Head And Piece To Camera is instructions for the agent only.

Does Talking Head And Piece To Camera access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Talking Head And Piece To Camera safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Talking Head And Piece To Camera use?

Talking Head And Piece To Camera is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Talking Head And Piece To Camera use?

About 2.2k tokens (SKILL.md is roughly 8.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 4.3k tokens, read only when the agent opens those files.

What are the alternatives to Talking Head And Piece To Camera?

Skills that share tags, products or a category with Talking Head And Piece To Camera: HyperFrames Animation (heygen-com/hyperframes, 60k stars), Faceless Explainer Video (heygen-com/hyperframes, 60k stars), Figma to HyperFrames (heygen-com/hyperframes, 60k stars) and HyperFrames Video Entry Point (heygen-com/hyperframes, 60k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Talking Head And Piece To Camera?

social-media-skills (a GitHub organization) maintains it in social-media-skills/skills, which has 134 GitHub stars. The repository holds 106 skills in this directory. The repository was last updated on October 1, 2026.

Source: social-media-skills/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.