Scientific Writer
gaasher/Agent-Loop-Skills
A skill your agent uses when the user has a scientific draft (with its dataset, figures, and optional analysis code) and wants it iteratively revised until it clears a quality bar.
Provider-independent governance workflow for verifying and correcting human- or model-generated video descriptions.
$ npx skills add calesthio/generative-media-skills --skill video-description-oversight -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install calesthio/generative-media-skills video-description-oversight --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/production/governance-delivery/video-description-oversight .claude/skills/video-description-oversight && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "video-description-oversight" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversight into .claude/skills/video-description-oversight/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-description-oversight", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversightType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add calesthio/generative-media-skills --skill video-description-oversight -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install calesthio/generative-media-skills video-description-oversight --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/production/governance-delivery/video-description-oversight .agents/skills/video-description-oversight && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "video-description-oversight" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversight into .agents/skills/video-description-oversight/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-description-oversight", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add calesthio/generative-media-skills --skill video-description-oversight -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install calesthio/generative-media-skills video-description-oversight --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/production/governance-delivery/video-description-oversight .cursor/skills/video-description-oversight && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "video-description-oversight" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversight into .cursor/skills/video-description-oversight/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-description-oversight", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/calesthio/generative-media-skills.git --path skills/production/governance-delivery/video-description-oversight--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add calesthio/generative-media-skills --skill video-description-oversight -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install calesthio/generative-media-skills video-description-oversight --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/production/governance-delivery/video-description-oversight .gemini/skills/video-description-oversight && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "video-description-oversight" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversight into .gemini/skills/video-description-oversight/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-description-oversight", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install calesthio/generative-media-skills video-description-oversightInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add calesthio/generative-media-skills --skill video-description-oversight -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/production/governance-delivery/video-description-oversight .github/skills/video-description-oversight && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "video-description-oversight" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversight into .github/skills/video-description-oversight/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-description-oversight", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add calesthio/generative-media-skills --skill video-description-oversight -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install calesthio/generative-media-skills video-description-oversight --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/calesthio/generative-media-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/production/governance-delivery/video-description-oversight .opencode/skills/video-description-oversight && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "video-description-oversight" agent skill from https://github.com/calesthio/generative-media-skills/tree/main/skills/production/governance-delivery/video-description-oversight into .opencode/skills/video-description-oversight/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "video-description-oversight", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
video-description-oversightProvider-independent governance workflow for verifying and correcting human- or model-generated video descriptions.
Video Description Oversight is an agent skill from calesthio/generative-media-skills. Provider-independent governance workflow for verifying and correcting human- or model-generated video descriptions. Use for pre-caption critique and post-caption revision, aspect-by-aspect factual review, critique precision/recall/constructiveness, second-stage peer review, calibration, appeals, versioned triplets, provenance, and acceptance reporting; not for pixel/audio QA, accessibility captioning, model training, or creative shot direction.
Its SKILL.md is about 3.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `EVAL.md`).
It sits in Research & Science, covering Performance reviews, Fine-tuning and Peer review. The repository describes itself as: Research-backed agent skills and tools for premium image, video, audio, voice, and generative media production across AI coding assistants. The licence is MIT.
6 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 8c85352. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md (its code samples are json).
From the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
arxiv.orghuggingface.colinzhiqiu.github.iogithub.comdoi.orgspec.c2pa.orgFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Video Description Oversight loads about 3.5k tokens when it runs. Until then it costs about 119 tokens; SKILL.md has 1,536 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from calesthio/generative-media-skills at commit 8c85352, republished under its MIT licence (© calesthio). 1,536 words, ~3,450 tokens.
.claude/skills/video-description-oversight/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.Use this skill when language about a video must be accepted as evidence-quality metadata rather than plausible prose. It governs a correction loop:
source video + specification + pre-caption
-> evidence-backed critique
-> revised post-caption
-> independent acceptance review
-> versioned triplet and decision recordThe reviewer verifies what language says about media. Technical file integrity, visual artifact QA, accessibility captions, creative intent, and model post-training belong elsewhere.
The workflow is informed by Lin et al.'s CHAI framework, which reports that critiques used to revise precise video captions are most useful when accurate, complete, and constructive. Sources were verified 2026-07-14. CHAI's performance numbers are first-party research results on its data and annotator program, not universal service-level guarantees.
Use this skill for descriptions used in:
Route elsewhere for:
The workflow requires a description specification. If none exists, establish one through precise-video-description before grading completeness or terminology.
Collect:
Do not let a reviewer critique video they cannot access. A text-only “blind critique” cannot establish visual accuracy.
Before detailed review, decide:
Record the reason. Do not force a critique workflow onto an unusable source/caption pair.
Use the project's precise-description contract rather than intuition.
Check entity count, stable identifiers, visible attributes, pose, relationships, entry/exit, and unverified identity/demographic inference.
Check setting, time/weather evidence, overlays versus physical objects, transitions, and unsupported mood/theme claims.
Check action verbs, actors/targets, temporal order, simultaneous activity, contact, direction, and causal overstatement.
Check shot size, frame position, depth, overlap/occlusion, start/end framing, and subject-relative versus frame-relative direction.
Check translation versus rotation, zoom versus movement, angle/height/roll, focus changes, steadiness, playback effects, edit transitions, and fabricated technical settings.
Review at normal speed and frame-level around disputed events. Each finding must include a locator or bounded interval where possible.
Every critique must pass three independent gates:
Every finding is real and supported by visible evidence/specification. Do not add a plausible correction that the video does not establish.
All consequential errors and omissions within the agreed scope are covered. Recall is not maximal verbosity; trivial details outside the contract can remain omitted.
Each finding says how to repair the caption: replace, add, delete, reorder, qualify, or mark uncertain. “This is wrong” is not a usable correction.
Recommended finding shape:
{
"aspect": "camera",
"time_range": ["00:02.100", "00:04.800"],
"error_type": "incorrect-term",
"caption_claim": "the camera zooms in",
"evidence": "foreground/background parallax increases while field of view appears stable",
"correction": "replace with 'the camera moves forward'",
"severity": "major",
"reviewer_id": "reviewer-17",
"spec_version": "video-language-2.1"
}This is an example schema. It must not contain invented confidence scores. Separate observed evidence from the proposed wording.
Do not invent feedback to prove that review occurred. If no corrections are required, record an explicit sentinel such as:
The description matches the reviewed source and specification; no edits are required.An “accurate” decision still needs review scope, asset/version, reviewer, date, glossary/specification version, and any unreviewed lanes.
The author/model revises from the accepted critique. Preserve all versions; never overwrite the pre-caption.
Second-stage acceptance asks:
The accepting reviewer should be different from the first reviewer for high-risk or dataset use. Lower-risk internal work may use calibrated spot review according to a documented policy.
Define severity from downstream consequence, not word count:
Possible dispositions: accepted, accepted_with_limitations, revise, reject, escalate, blocked.
Do not mechanically waive privacy, consent, safety, or material factual errors.
Use gold examples and periodic blind calibration. Track agreement by aspect because a team may agree on Subject/Scene while failing on Spatial/Camera.
When reviewers disagree:
Do not impose a universal kappa threshold, daily review quota, compensation scheme, or expertise ladder. CHAI used a highly trained professional pipeline; teams with different content or reviewers must establish their own validated calibration criteria.
Preserve:
The (pre-caption, critique, post-caption) triplet can become a valuable dataset asset, but production permission does not automatically grant model-training permission. Create a datasheet and confirm source-video, caption, reviewer, and derivative-data rights before reuse.
C2PA or signed records can support provenance; they do not prove that a caption is true.
Detailed video descriptions may expose identity, location, private behavior, screens, health/financial details, or copyrighted story content. Minimize access, use pseudonymous reviewer IDs, define retention/deletion, separate public output from private review metadata, and audit exports.
Match reviewers to domain/language complexity. Cinematography expertise does not imply medical or cultural expertise; language fluency does not guarantee camera-motion discrimination. CHAI's findings came from trained professional creators and predominantly professional video domains. Do not assume equivalent results from untrained crowdworkers, multilingual auto-translation, long-form footage, or specialized domains.
Track trends rather than optimizing one number:
High critique volume can indicate either poor sources or overcritical reviewers. Audit evidence before drawing conclusions.
This is a complete example, not a mandatory formula.
Source: six-second kitchen shot. Pre-caption claim: “The camera zooms in on the woman as she glances to the right.”
Evidence review: background/foreground parallax changes, supporting forward camera translation; the woman looks toward frame-right, but whether that is her left/right is not needed. Focus remains on her eye.
Critique: “Camera, 00:02.0-00:05.0: replace ‘zooms in’ with ‘moves forward’ because the viewpoint translates and parallax changes rather than only the field of view. Spatial/Motion: keep ‘looks toward frame-right’; do not convert it to the subject's right without evidence. The remaining subject and scene description is accurate.”
Post-caption: “In a dim kitchen, the camera moves smoothly forward toward a woman as she raises her gaze toward frame-right; focus remains on her near eye.”
Acceptance: second reviewer confirms correction, no new claims, and records the exact source/spec versions.
Failure to avoid: “The caption is wrong about camera motion.” It lacks the supported replacement and is non-constructive.
This is a complete example, not a mandatory formula.
Pre-caption: “A knight runs through a level while the camera follows.”
Reviewed scope: all five aspects for dataset metadata.
Findings:
Constructive critique: provides each replacement/addition and preserves uncertainty. Post-caption: incorporates all five without claiming world speed or exact lens. Disposition: accepted after second review.
Failure to avoid: adding speculative game title, character name, player intent, or “dynamic exciting atmosphere.”
Verified 2026-07-14:
© calesthio, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file in skills/production/governance-delivery/video-description-oversight of calesthio/generative-media-skills.
Open the folder on GitHubat commit 8c85352
Video Description Oversight next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Video Description Oversight this skillcalesthio/generative-media-skills | 197 | — | ~3.5k | Automated safety check: Pass | MIT | |
| Scientific Writergaasher/Agent-Loop-Skills | 174 | — | ~3.3k | Automated safety check: Pass | MIT | |
| External Model Validationaipoch/medical-research-skills | 1.9k | — | ~3.2k | Automated safety check: Pass | MIT | |
| Jeg Rebuttalbrycewang-stanford/Awesome-Journal-Skills | 1.2k | — | ~1.2k | Automated safety check: Pass | MIT | |
| Quaxnstarman/quax | 143 | — | ~5.5k | Automated safety check: Pass | Apache-2.0 | |
| Paper To HTMLysyecust/lecture-to-notes | 273 | — | ~1.3k | Automated safety check: Pass | Custom licence |
gaasher/Agent-Loop-Skills
A skill your agent uses when the user has a scientific draft (with its dataset, figures, and optional analysis code) and wants it iteratively revised until it clears a quality bar.
aipoch/medical-research-skills
A skill your agent uses when validating an existing prognostic risk signature on an external bulk expression cohort with survival outcomes, producing risk scores, Kaplan-Meier curves, risk…
brycewang-stanford/Awesome-Journal-Skills
A skill your agent uses when responding after a Journal of Economic Growth revise-and-resubmit to organize responses about growth mechanisms, model assumptions, empirical identification, calibration…
nstarman/quax
A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify…
ysyecust/lecture-to-notes
Generate a self-contained, beautifully styled HTML analysis of an academic paper.
maziyarpanahi/openmed
Assign negation, temporality, and uncertainty (the ConText axes) to clinical entities extracted by OpenMed, so "denies chest pain" is not counted as chest pain and "history of MI" is not counted as…
calesthio/generative-media-skills
A skill your agent uses to turn generated, captured, scanned, or modeled 3D output into production-ready standalone assets for DCC, real-time engine, web, or interchange delivery.
calesthio/generative-media-skills
Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social…
calesthio/generative-media-skills
Provider-independent captions and media accessibility direction for AI agents producing or finishing generated videos, ads, social clips, explainers, avatar videos, documentaries, podcasts/video…
calesthio/generative-media-skills
Provider-independent production workflow for agents assembling, auditing, executing, and handing off ComfyUI node-graph workflows for image, video, upscale, inpaint, conditioning, and batch media…
calesthio/generative-media-skills
Provider-independent FFmpeg finishing workflow for AI agents preparing generated or edited media deliverables.
calesthio/generative-media-skills
Provider-independent quality assurance for AI-generated and AI-assisted media.
Provider-independent governance workflow for verifying and correcting human- or model-generated video descriptions. Video Description Oversight is an agent skill from calesthio/generative-media-skills. Provider-independent governance workflow for verifying and correcting human- or model-generated video descriptions.
Video Description Oversight fits situations like: pre-caption critique and post-caption revision; aspect-by-aspect factual review; critique precision/recall/constructiveness; second-stage peer review.
Run `npx skills add calesthio/generative-media-skills --skill video-description-oversight -a claude-code`. Or copy the skill folder (skills/production/governance-delivery/video-description-oversight in calesthio/generative-media-skills) into .claude/skills/video-description-oversight in your project. Claude Code loads it when a task matches its description.
Run `npx skills add calesthio/generative-media-skills --skill video-description-oversight -a codex`. Or copy the skill folder (skills/production/governance-delivery/video-description-oversight in calesthio/generative-media-skills) into .agents/skills/video-description-oversight in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add calesthio/generative-media-skills --skill video-description-oversight -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video-description-oversight, .gemini/skills/video-description-oversight, .github/skills/video-description-oversight and .opencode/skills/video-description-oversight in your project.
SKILL.md names no scripts, command-line tools or credentials: Video Description Oversight is instructions for the agent only.
SKILL.md names 6 domains. As links in the text: arxiv.org, huggingface.co, linzhiqiu.github.io, github.com, doi.org and spec.c2pa.org. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Video Description Oversight is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.5k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Video Description Oversight: Scientific Writer (gaasher/Agent-Loop-Skills, 174 stars), External Model Validation (aipoch/medical-research-skills, 1.9k stars), Jeg Rebuttal (brycewang-stanford/Awesome-Journal-Skills, 1.2k stars) and Quax (nstarman/quax, 143 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
calesthio (a GitHub user) maintains it in calesthio/generative-media-skills, which has 197 GitHub stars. The repository holds 26 skills in this directory. The repository was last updated on July 14, 2026.
Source: calesthio/generative-media-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.