Agent skill

Meeting Video Grounding

by gaotiexinqu in gaotiexinqu/OneResearchClaw

Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

MITAuto-check passedMedia & Creative

Install Meeting Video Grounding

skills CLI
$ npx skills add gaotiexinqu/OneResearchClaw --skill meeting-video-grounding -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gaotiexinqu/OneResearchClaw meeting-video-grounding --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gaotiexinqu/OneResearchClaw.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.cursor/skills/meeting-video-grounding .claude/skills/meeting-video-grounding && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
meeting-video-grounding
GitHub stars
450
Token cost
~1.2k tokens
SKILL.md length
548 words
Files
3 (incl. scripts)
Skills in repo
15
Repo updated
First seen
Licence
MIT

At a glance

Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

  • Works in 3 steps: extracting audio from the video → reusing the existing audio_structuring… → reusing the existing meeting-grounding…
  • Media & Creative work in your project
  • SKILL.md covers When to Use, Input, Output Bundle and Important separation of…, plus 6 more sections
  • Runs Python and Shell scripts from its folder

What it does

Meeting Video Grounding is an agent skill from gaotiexinqu/OneResearchClaw. Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including scripts (for example `scripts/prepare_bundle.py` and `scripts/run.sh`).

It sits in Media & Creative. The repository describes itself as: Any research. One Claw. 🦞 From any materials to research with fully autonomous & skill-driven researcher. The licence is MIT.

When your agent uses it

  • Media & Creative work in your project

Example prompts

  • “/meeting-video-grounding”

Requirements

  • Python 3
  • A Bash shell

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. extracting audio from the video
  2. reusing the existing audio_structuring skill to produce meeting_transcript.txt
  3. reusing the existing meeting-grounding skill to turn that transcript into meeting grounding outputs

What it can do on your machine

Read from SKILL.md and the folder at commit 37e86c6. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 2 files in scripts/ (Python and Shell), which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Meeting Video Grounding loads about 1.2k tokens when it runs. Until then it costs about 41 tokens; SKILL.md has 548 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~41
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from gaotiexinqu/OneResearchClaw at commit 37e86c6, republished under its MIT licence (© gaotiexinqu). 548 words, ~1,204 tokens.

Download SKILL.mdSave it as .claude/skills/meeting-video-grounding/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
meeting-video-grounding
description
Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

Meeting Video Grounding

Convert a meeting video into structured meeting grounding outputs by:

  1. extracting audio from the video
  2. reusing the existing audio_structuring skill to produce meeting_transcript.txt
  3. reusing the existing meeting-grounding skill to turn that transcript into meeting grounding outputs

This skill is for meeting videos where the primary information comes from speech. It is intentionally audio-first. It does not attempt full visual understanding of the video.

When to Use

Use this skill when:

  • the input is a meeting video or discussion video
  • the main information is expected to come from spoken content
  • you want to reuse the existing audio transcription and meeting grounding workflow

Do not use this skill when:

  • the task requires visual analysis of slides, demos, whiteboards, or screen content as first-class evidence
  • the video has little or no speech
  • the task is to write a polished final report directly from the raw video

Input

A single meeting video file.

Typical examples:

  • .mp4
  • .mkv
  • .mov
  • .webm

Optional input:

  • transcription_language: language code such as en or zh

Output Bundle

For each input video, create one bundle directory:

text
data/grounded_notes/<ground_id>/

Inside that bundle, the expected outputs are always:

text
<bundle_dir>/
├─ extracted.md
├─ extracted_meta.json
├─ grounded.md
├─ audio/
│  └─ meeting_audio.wav
└─ transcript/
   └─ meeting_transcript.txt

If the meeting contains multiple independent topics that should be researched separately downstream, the bundle may also contain:

text
<bundle_dir>/
├─ topic_manifest.json
└─ child_outputs/
   ├─ topic_01/
   │  └─ grounded.md
   ├─ topic_02/
   │  └─ grounded.md
   └─ ...

Important separation of responsibilities

  • scripts/run.sh is responsible for:
    • extracting audio from the input video
    • calling the existing audio_structuring skill
    • creating the bundle files:
      • audio/meeting_audio.wav
      • transcript/meeting_transcript.txt
      • extracted.md
      • extracted_meta.json
  • The agent is responsible for:
    • reading the transcript bundle
    • applying the existing meeting-grounding skill
    • always writing the meeting-level:
      • grounded.md
    • and, when appropriate, also writing:
      • topic_manifest.json
      • child_outputs/topic_xx/grounded.md

grounded.md must be a real grounding note. It must not remain a placeholder scaffold.

Required Agent Workflow

  1. Run the existing scripts/run.sh entrypoint for this skill.
  2. Confirm that the bundle exists and that these files are present:
    • extracted.md
    • extracted_meta.json
    • audio/meeting_audio.wav
    • transcript/meeting_transcript.txt
  3. Read transcript/meeting_transcript.txt as the primary grounding evidence.
  4. Reuse the existing meeting-grounding skill on that transcript.
  5. Always save the meeting-level structured note to grounded.md inside the same bundle directory.
  6. If the transcript clearly contains multiple independent topics, also save:
    • topic_manifest.json
    • child_outputs/topic_xx/grounded.md
  7. Do not stop after confirming that the transcript bundle exists.
Show full SKILL.md (191 more words)Show less

Important scope rule

This skill currently treats meeting videos as audio-first inputs.

That means:

  • the extracted transcript is the primary evidence for grounding
  • absence of visual analysis is not, by itself, a failure
  • do not invent slide content, visual details, or screen evidence that were not captured in the transcript

Output Format

The final grounded.md must follow the existing meeting-grounding schema exactly.

If topic children are created, each child grounded note must also follow the same meeting-grounding schema.

Do not invent a new schema here.

Instructions

  • Do not implement a new ASR pipeline.
  • Do not implement a new meeting summarizer.
  • Do not directly summarize the raw video without first running the existing workflow.
  • Reuse the existing audio_structuring skill for transcription.
  • Reuse the existing meeting-grounding skill for transcript grounding.
  • Treat the transcript as the primary evidence.
  • Keep the workflow simple and stable.

Failure Handling

  • If the input video file does not exist, fail clearly.
  • If the video has no audio stream, fail clearly.
  • If audio extraction fails, fail clearly.
  • If meeting_transcript.txt is not produced, fail clearly.
  • Do not pretend the task succeeded if only part of the workflow completed.

Example Invocation

/meeting-video-grounding

© gaotiexinqu, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (scripts) in .cursor/skills/meeting-video-grounding of gaotiexinqu/OneResearchClaw.

  • SKILL.md
  • scripts/prepare_bundle.py
  • scripts/run.sh

Open the folder on GitHubat commit 37e86c6

Compare with similar skills

Meeting Video Grounding next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Meeting Video Grounding compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Meeting Video Grounding this skillgaotiexinqu/OneResearchClaw450—~1.2kAutomated safety check: PassMIT
Guizang Social Cardsop7418/guizang-social-card-skill7.4k1 repos~7.8kAutomated safety check: PassAGPL-3.0
Weekly Changelog Videoheygen-com/hyperframes60k—~3.3kAutomated safety check: PassApache-2.0
Anthropic Brand Stylinganthropics/skills180k30 repos~559Automated safety check: PassApache-2.0
MoneyPrinterTurbo Video Generatorharry0703/MoneyPrinterTurbo130k—~2.1kAutomated safety check: WarnMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0

Similar skills

  • Guizang Social Cards

    op7418/guizang-social-card-skill

    Produces social card sets for Xiaohongshu and WeChat: carousels, Live Photo motion cards and puzzle layouts, and WeChat cover pairs, rendered from single-file HTML.

    7.4k GitHub starsUsed in 1 repo~7.8k tokens
    Media & CreativeAuto-check passed
  • Weekly Changelog Video

    heygen-com/hyperframes

    Turns a weekly changelog markdown file into a branded HyperFrames video with voiceover, animated mock-UI scenes and captions, using fonts, background and scripts bundled in the skill.

    60k GitHub stars~3.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Anthropic Brand Styling

    anthropics/skills

    Official

    Applies Anthropic's brand colors and fonts to artifacts such as PowerPoint slides, using fixed hex values for text and accents, Poppins headings and Lora body text.

    180k GitHub starsUsed in 30 repos~559 tokens
    Media & CreativeAuto-check passed
  • MoneyPrinterTurbo Video Generator

    harry0703/MoneyPrinterTurbo

    Installs and runs MoneyPrinterTurbo to turn a topic or script into a finished short video with voice-over, subtitles, stock footage and music.

    130k GitHub stars~2.1k tokensUpdated yesterday
    Media & CreativeAuto-check: warnings
  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Holo Card Studio

    EverettFish/holo-card-studio

    Create collectible holographic foil cards and two-image lenticular flip cards with AI-generated full-color ukiyo-e and colored sumi-e anime artwork, layered Blender scenes, renders, GLB export, and…

    1.9k GitHub stars~1.4k tokensUpdated 19 days ago
    Media & CreativeAuto-check passed

More from gaotiexinqu/OneResearchClaw

All 15 skills in this repo
  • Remote Input

    gaotiexinqu/OneResearchClaw

    Download remote content (arxiv papers, YouTube videos, Bilibili videos) to local storage and route to downstream grounding pipeline.

    450 GitHub stars~2.8k tokensUpdated 5 mo ago
    Auto-check passed
  • Grounded Research Lit

    gaotiexinqu/OneResearchClaw

    Run focused literature and web research from a grounded note.

    450 GitHub stars~11k tokensUpdated 5 mo ago
    Auto-check passed
  • Archive Grounding

    gaotiexinqu/OneResearchClaw

    Unpack a ZIP archive, inventory its files, run the corresponding child grounding skill for each supported child file, and then write a real archive-level grounded.md.

    450 GitHub stars~2.1k tokensUpdated 5 mo ago
    Auto-check: notes
  • Document Grounding

    gaotiexinqu/OneResearchClaw

    Convert a raw document into a structured grounding note for downstream research and summarization.

    450 GitHub stars~2k tokensUpdated 5 mo ago
    Auto-check passed
  • Meeting Audio Grounding

    gaotiexinqu/OneResearchClaw

    Convert a meeting audio file into a transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

    450 GitHub stars~1.1k tokensUpdated 5 mo ago
    Auto-check passed
  • PPTX Grounding

    gaotiexinqu/OneResearchClaw

    Extract a structured evidence bundle from a .pptx deck, then write a real grounded.md from the bundle.

    450 GitHub stars~2.1k tokensUpdated 5 mo ago
    Auto-check: notes

Questions about Meeting Video Grounding

What does Meeting Video Grounding do?

Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs. Meeting Video Grounding is an agent skill from gaotiexinqu/OneResearchClaw. Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

When should I use Meeting Video Grounding?

Meeting Video Grounding fits situations like: media & Creative work in your project.

How do I install Meeting Video Grounding in Claude Code?

Run `npx skills add gaotiexinqu/OneResearchClaw --skill meeting-video-grounding -a claude-code`. Or copy the skill folder (.cursor/skills/meeting-video-grounding in gaotiexinqu/OneResearchClaw) into .claude/skills/meeting-video-grounding in your project. Claude Code loads it when a task matches its description.

How do I install Meeting Video Grounding in Codex?

Run `npx skills add gaotiexinqu/OneResearchClaw --skill meeting-video-grounding -a codex`. Or copy the skill folder (.cursor/skills/meeting-video-grounding in gaotiexinqu/OneResearchClaw) into .agents/skills/meeting-video-grounding in your project. Codex loads it when a task matches its description.

Can I use Meeting Video Grounding in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gaotiexinqu/OneResearchClaw --skill meeting-video-grounding -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/meeting-video-grounding, .gemini/skills/meeting-video-grounding, .github/skills/meeting-video-grounding and .opencode/skills/meeting-video-grounding in your project.

What does Meeting Video Grounding need to run?

Going by SKILL.md and its folder, Meeting Video Grounding needs Python and a shell for the scripts in its folder. Our summary lists: Python 3; A Bash shell.

Does Meeting Video Grounding access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Meeting Video Grounding safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Meeting Video Grounding use?

Meeting Video Grounding is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Meeting Video Grounding use?

About 1.2k tokens (SKILL.md is roughly 4.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Meeting Video Grounding?

Skills that share tags, products or a category with Meeting Video Grounding: Guizang Social Cards (op7418/guizang-social-card-skill, 7.4k stars), Weekly Changelog Video (heygen-com/hyperframes, 60k stars), Anthropic Brand Styling (anthropics/skills, 180k stars) and MoneyPrinterTurbo Video Generator (harry0703/MoneyPrinterTurbo, 130k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Meeting Video Grounding?

gaotiexinqu (a GitHub user) maintains it in gaotiexinqu/OneResearchClaw, which has 450 GitHub stars. The repository holds 15 skills in this directory. The repository was last updated on May 9, 2026.

Source: gaotiexinqu/OneResearchClaw on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.