Agent skill

Media Ops Skill

by OpenLoaf in OpenLoaf/OpenLoaf

Process local image / video / audio files, or download videos from the web.

AGPL-3.0Auto-check passedMedia & Creative

Install Media Ops Skill

skills CLI
$ npx skills add OpenLoaf/OpenLoaf --skill media-ops-skill -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install OpenLoaf/OpenLoaf media-ops-skill --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/OpenLoaf/OpenLoaf.git skills-src && mkdir -p .claude/skills && cp -r skills-src/apps/server/src/ai/builtin-skills/media-ops/en .claude/skills/media-ops-skill && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
media-ops-skill
GitHub stars
108
Token cost
~1.4k tokens
SKILL.md length
460 words
Files
1
Skills in repo
33
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Process local image / video / audio files, or download videos from the web.

  • Tasks that involve Video production
  • SKILL.md covers Tool Inventory, Decision Tree, ImageProcess — Image Processing and VideoConvert — Video/Audio…, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Media Ops Skill is an agent skill from OpenLoaf/OpenLoaf. Process local image / video / audio files, or download videos from the web. Images: resize, crop, rotate, convert format (→webp/png/jpeg), apply filters, inspect dimensions. Video: convert format, extract audio, inspect metadata. Download: public videos from YouTube / Bilibili etc. Also the go-to skill for post-processing AI-generated images (format conversion, resizing). Not for: AI generating images/video/voice from scratch (→cloud-media-skill), or rendering charts inline (→visualization-ops-skill).

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Video production. It works with Bilibili and YouTube. The repository describes itself as: 🍞Open-source, local-first AI workspace with Agents, multi-model chat (GPT/Claude/Gemini/DeepSeek), Notion-like docs, AI image & video generation, email, calendar & terminal…. The licence is AGPL-3.0.

When your agent uses it

  • Tasks that involve Video production

Example prompts

  • “/media-ops-skill”

What it can do on your machine

Read from SKILL.md and the folder at commit f7eccf6. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Media Ops Skill loads about 1.4k tokens when it runs. Until then it costs about 132 tokens; SKILL.md has 460 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~132
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from OpenLoaf/OpenLoaf at commit f7eccf6, republished under its AGPL-3.0 licence (© OpenLoaf). 460 words, ~1,352 tokens.

Download SKILL.mdSave it as .claude/skills/media-ops-skill/SKILL.md (or your agent's skills folder).
name
media-ops-skill
description
Process local image / video / audio files, or download videos from the web. Images: resize, crop, rotate, convert format (→webp/png/jpeg), apply filters, inspect dimensions. Video: convert format, extract audio, inspect metadata. Download: public videos from YouTube / Bilibili etc. Also the go-to skill for post-processing AI-generated images (format conversion, resizing). **Not for**: AI generating images/video/voice from scratch (→cloud-media-skill), or rendering charts inline (→visualization-ops-skill).
tools
ImageProcess, VideoConvert, VideoDownload

Image / Video / Audio Processing

This skill covers three local media tools. Image processing is the primary capability.

⚡ Act immediately upon reading this — do not stop to reply to the user. Right now, call ToolSearch(names: "ImageProcess") to activate the tool schema, then call ImageProcess to complete the operation. All three steps must happen within the same response.

Tool Inventory

ToolResponsibility
ImageProcessImage processing: resize / crop / rotate / flip / format conversion / grayscale / blur / sharpen / tint / metadata
VideoConvertVideo format conversion / audio extraction / metadata
VideoDownloadDownload public videos from YouTube / Bilibili etc.

Loading: all are deferred tools; run ToolSearch(names: "ImageProcess,VideoConvert,VideoDownload") to activate their schemas before calling.

Decision Tree

text
User needs a media operation
├── Process an image? (resize / crop / rotate / convert / filter / inspect)
│   └── ImageProcess
├── Process a video?
│   ├── Format conversion / resolution change → VideoConvert (action: convert)
│   └── Extract audio → VideoConvert (action: extract-audio)
├── Inspect file info?
│   ├── Image (dimensions / format / DPI) → ImageProcess (action: get-info)
│   └── Video (duration / resolution / codec) → VideoConvert (action: get-info)
└── Download a video from the web?
    └── VideoDownload

ImageProcess — Image Processing

Powered by sharp. Also applies to AI-generated images.

actionPurposeKey parameters
get-infoRead dimensions, format, DPI, file sizefilePath only
resizeScalewidth/height/fit
cropRectangular cropleft/top/width/height
rotateRotateangle
flipFlipdirection
convertFormat conversionformat (png/jpeg/webp/avif/tiff)
grayscale/blur/sharpen/tintFilter effectsEach has its own parameters

Run get-info first: know the original dimensions and format before operating — for example, if the user says "shrink by half", you need the original width/height to compute the target.

Output path: by default, append a suffix (e.g. photo_resized.jpg). Overwriting in place is risky — the user may need the original for comparison or rollback. Only overwrite when explicitly requested.

VideoConvert — Video/Audio Conversion

Powered by FFmpeg; handles format conversion and audio extraction on existing files.

actionPurposeKey parameters
get-infoRead duration, resolution, codec, stream infofilePath only
convertVideo format conversionformat, resolution
extract-audioExtract audio from videoaudioFormat (mp3/wav/aac/flac)

Format guidelines:

  • Universal: MP4 (H.264) — plays on virtually any device
  • Efficient: WebM (VP9) — smaller files, slightly weaker compatibility
  • Audio: MP3 (universal), FLAC (lossless)

Large files: transcoding time scales with file size. Warn the user upfront for files over 500 MB.

Show full SKILL.md (166 more words)Show less

VideoDownload — Video Download

Downloads public video URLs via yt-dlp.

Use when: the user provides a public video link and needs it locally for further processing.

Not applicable:

  • Generating a new video → cloud-media-skill
  • Converting an existing local video → VideoConvert
  • Private / login-required videos → tell the user yt-dlp only handles public content

Common follow-ups: proactively ask whether the user needs audio extraction, format conversion, or trimming after download.

Common Workflows

Post-processing AI-generated images
CloudImageGenerate(…) → get absolutePath
ImageProcess(action: "get-info", filePath: "…")   # confirm original dimensions
ImageProcess(action: "convert", filePath: "…", format: "webp", outputPath: "…")
Video download → audio extraction
VideoDownload(url: "...") → get filePath
VideoConvert(action: "extract-audio", filePath: "...", outputPath: "output.mp3")
Batch image format conversion

Call ImageProcess(action: "convert", format: "webp") for each image. WebP maintains visual quality at roughly 70% of JPEG file size — ideal for web delivery.

Common Pitfalls

Unsupported format → check whether the extension matches the actual encoding. Confirm the real format with get-info first.

Output path matches input path → some operations don't support in-place overwrite. Use a different outputPath and notify the user if replacement is needed afterward.

"Compress the image" is ambiguous → may mean: reduce dimensions (resize), lower quality (adjust quality in convert), or switch format (jpeg→webp). Clarify proactively.

© OpenLoaf, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in apps/server/src/ai/builtin-skills/media-ops/en of OpenLoaf/OpenLoaf.

Open the folder on GitHubat commit f7eccf6

Compare with similar skills

Media Ops Skill next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Media Ops Skill compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Media Ops Skill this skillOpenLoaf/OpenLoaf108—~1.4kAutomated safety check: PassAGPL-3.0
Yt Dlp DownloaderMapleShaw/yt-dlp-downloader-skill2091 repos~1.5kAutomated safety check: PassNone
Video Downloaderisjiamu/jiamu-skills139—~1.2kAutomated safety check: PassMIT
Videohub Youtubecacity/VideoHub168—~550Automated safety check: PassMIT
Video Podcast MakerAgents365-ai/video-podcast-maker1.7k—~4.9kAutomated safety check: PassMIT
Video Clip Extractorlinzzzzzz/openclip569—~2.8kAutomated safety check: WarnMIT

Similar skills

  • Yt Dlp Downloader

    MapleShaw/yt-dlp-downloader-skill

    Download videos from YouTube, Bilibili, Twitter, and thousands of other sites using yt-dlp.

    209 GitHub starsUsed in 1 repo~1.5k tokens
    Media & CreativeAuto-check passed
  • Video Downloader

    isjiamu/jiamu-skills

    Download videos from 1000+ websites (YouTube, Bilibili, Twitter/X, TikTok, etc.) using yt-dlp.

    139 GitHub stars~1.2k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Videohub Youtube

    cacity/VideoHub

    处理 YouTube、Twitter(X)、Bilibili 和本地音视频/文本的转写、字幕、翻译与总结。优先复用 src/youtubetranscriber.py 现有 CLI。

    168 GitHub stars~550 tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • Video Podcast Maker

    Agents365-ai/video-podcast-maker

    A skill your agent uses when the user gives a topic and wants an automated topic-driven narrated explainer, podcast, or knowledge-summary video (Bilibili / YouTube / Xiaohongshu / Douyin / WeChat…

    1.7k GitHub stars~4.9k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Video Clip Extractor

    linzzzzzz/openclip

    Processes videos to identify engaging moments, generate transcripts, and create highlight clips with artistic titles and custom cover images.

    569 GitHub stars~2.8k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: warnings
  • Generates SRT, VTT and caption JSON for final narration audio or a merged video using Volcengine Doubao ASR word timestamps, with a quality gate before delivery.

    1.6k GitHub stars~1.1k tokensUpdated 18 days ago
    Media & CreativeAuto-check: notes

More from OpenLoaf/OpenLoaf

All 33 skills in this repo
  • Agent Orchestration Skill

    OpenLoaf/OpenLoaf

    Triggers when the master Agent faces a multi-step complex task and is deciding whether / how to outsource sub-tasks to built-in subagents (browser / doc-editor / data-analyst / extractor /…

    108 GitHub stars~1.9k tokensUpdated 4 mo ago
    Auto-check passed
  • Browser Ops Skill

    OpenLoaf/OpenLoaf

    Triggered when the user asks for page-level interaction with a specific webpage: login, form filling, button clicks, pagination scraping, screenshots, downloading page images, handling CAPTCHAs or…

    108 GitHub stars~1.5k tokensUpdated 4 mo ago
    Auto-check passed
  • Canvas Ops Skill

    OpenLoaf/OpenLoaf

    Triggered when the user wants lifecycle management of OpenLoaf canvases / whiteboards: create, open, filter, duplicate, delete, rename, or change ownership.

    108 GitHub stars~1.4k tokensUpdated 4 mo ago
    Auto-check passed
  • Reads, edits, converts and reviews Word documents through three dedicated tools, covering tracked changes, comments, tables, images and format conversion.

    108 GitHub stars~1.9k tokensUpdated 4 mo ago
    Auto-check passed
  • Email Operations

    OpenLoaf/OpenLoaf

    Handles a real email account through query and mutate tools: check the inbox, read, search, reply, forward, compose and organize, with sending always confirmed first.

    108 GitHub stars~1.8k tokensUpdated 4 mo ago
    Auto-check passed
  • macOS Desktop Control

    OpenLoaf/OpenLoaf

    Guides an agent to operate native macOS apps by surveying an app first, acting through intents, menus or keystrokes, and verifying each step, in OpenLoaf Desktop only.

    108 GitHub stars~2.4k tokensUpdated 4 mo ago
    Auto-check passed

Works with

Questions about Media Ops Skill

What does Media Ops Skill do?

Process local image / video / audio files, or download videos from the web. Media Ops Skill is an agent skill from OpenLoaf/OpenLoaf. Process local image / video / audio files, or download videos from the web.

When should I use Media Ops Skill?

Media Ops Skill fits situations like: tasks that involve Video production.

How do I install Media Ops Skill in Claude Code?

Run `npx skills add OpenLoaf/OpenLoaf --skill media-ops-skill -a claude-code`. Or copy the skill folder (apps/server/src/ai/builtin-skills/media-ops/en in OpenLoaf/OpenLoaf) into .claude/skills/media-ops-skill in your project. Claude Code loads it when a task matches its description.

How do I install Media Ops Skill in Codex?

Run `npx skills add OpenLoaf/OpenLoaf --skill media-ops-skill -a codex`. Or copy the skill folder (apps/server/src/ai/builtin-skills/media-ops/en in OpenLoaf/OpenLoaf) into .agents/skills/media-ops-skill in your project. Codex loads it when a task matches its description.

Can I use Media Ops Skill in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add OpenLoaf/OpenLoaf --skill media-ops-skill -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/media-ops-skill, .gemini/skills/media-ops-skill, .github/skills/media-ops-skill and .opencode/skills/media-ops-skill in your project.

What does Media Ops Skill need to run?

SKILL.md names no scripts, command-line tools or credentials: Media Ops Skill is instructions for the agent only.

Does Media Ops Skill access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Media Ops Skill safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Media Ops Skill use?

Media Ops Skill is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Media Ops Skill use?

About 1.4k tokens (SKILL.md is roughly 5.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Media Ops Skill?

Skills that share tags, products or a category with Media Ops Skill: Yt Dlp Downloader (MapleShaw/yt-dlp-downloader-skill, 209 stars), Video Downloader (isjiamu/jiamu-skills, 139 stars), Videohub Youtube (cacity/VideoHub, 168 stars) and Video Podcast Maker (Agents365-ai/video-podcast-maker, 1.7k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Media Ops Skill?

OpenLoaf (a GitHub organization) maintains it in OpenLoaf/OpenLoaf, which has 108 GitHub stars. The repository holds 33 skills in this directory. The repository was last updated on May 14, 2026.

Source: OpenLoaf/OpenLoaf on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.