Agent skill

Online Content Collector

by cafe3310 in cafe3310/public-agent-skills

对 Obsidian 仓库进行自动素材媒体剪藏,本地化特定 tag 标注的网页、视频及附件. An agent skill from cafe3310/public-agent-skills.

Apache-2.0Auto-check passed

Install Online Content Collector

skills CLI
$ npx skills add cafe3310/public-agent-skills --skill online-content-collector -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cafe3310/public-agent-skills online-content-collector --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cafe3310/public-agent-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/online-content-collector .claude/skills/online-content-collector && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
online-content-collector
GitHub stars
255
Token cost
~455 tokens
SKILL.md length
137 words
Files
21 (incl. scripts, assets)
Skills in repo
29
Repo updated
First seen
Licence
Apache-2.0

At a glance

对 Obsidian 仓库进行自动素材媒体剪藏,本地化特定 tag 标注的网页、视频及附件. An agent skill from cafe3310/public-agent-skills.

  • Works in 3 steps: 配置读取: Agent 读取 AGENTS.md,确定目标路径。 → 执行脚本: 调用 scripts/collect_links.py,传入… → 元数据提取与更新: 脚本扫描包含 #Marker-待下载 的文件,提取 time…
  • SKILL.md covers 概述, 核心配置解析, 核心工作流 and 依赖工具, plus 1 more section

What it does

Online Content Collector is an agent skill from cafe3310/public-agent-skills. 对 Obsidian 仓库进行自动素材媒体剪藏,本地化特定 tag 标注的网页、视频及附件

Its SKILL.md is about 460 tokens, which your agent loads only when the skill is triggered. The skill folder holds 29 other files, including scripts and assets (for example `2026-04-28-20-55-online-collector-requirements.md`, `assets/example_vault/Archives/Clips/[2026-05-02-15] 演示 dQw4w9WgXcQ/[2026-05-02-15] 演示 dQw4w9WgXcQ.md` and `assets/example_vault/Archives/Clips/[2026-05-02-18] 演示 2036268335927796152/[2026-05-02-18] 演示 2036268335927796152.md`).

It works with Obsidian. The repository describes itself as: personal agent skills for better QoL. The licence is Apache-2.0.

Example prompts

  • “/online-content-collector”

Requirements

  • Python 3

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. 配置读取: Agent 读取 AGENTS.md,确定目标路径。
  2. 执行脚本: 调用 scripts/collect_links.py,传入 --vault-path 和 --list-dir。
  3. 元数据提取与更新: 脚本扫描包含 #Marker-待下载 的文件,提取 time 和 source,生成 YAML 格式的任务列表 [yyyy-mm-dd-hh 下载列表整理.md],并将原始文件标签更新为 #Marker-下载中-YYYYMMDD。

What it can do on your machine

Read from SKILL.md and the folder at commit 6c45501. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Online Content Collector loads about 455 tokens when it runs. Until then it costs about 18 tokens; SKILL.md has 137 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~18
When it runs · the whole SKILL.md, loaded when a task matches
~455

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from cafe3310/public-agent-skills at commit 6c45501, republished under its Apache-2.0 licence (© cafe3310). 137 words, ~455 tokens.

Download SKILL.mdSave it as .claude/skills/online-content-collector/SKILL.md (or your agent's skills folder). This skill also uses 20 other files; get the full folder from GitHub.
name
online-content-collector
description
对 Obsidian 仓库进行自动素材媒体剪藏,本地化特定 tag 标注的网页、视频及附件
license
Apache-2.0
author
github/cafe3310
depends_on_binary
yt-dlp, pandoc, ffmpeg, python3

技能:线上素材收集器 (Online Content Collector)

概述

此技能旨在实现从“发现链接”到“本地化存档”的完全自动化。它扫描 Obsidian 仓库中带有特定标签的链接,将其汇总并下载为包含文本、图片、视频及附件的完整本地 Markdown 存档。

核心配置解析

此技能依赖用户 Vault 根目录下的 AGENTS.md。Agent 在执行脚本前,必须首先读取此文件,并解析出以下路径:

  • 下载列表目录位于: 用于传递给 --list-dir。
  • 下载内容目录位于: 用于传递给 --archive-dir。

核心工作流

第一阶段:扫描与汇总 (Discovery & Aggregation)
  1. 配置读取: Agent 读取 AGENTS.md,确定目标路径。
  2. 执行脚本: 调用 scripts/collect_links.py,传入 --vault-path 和 --list-dir。
  3. 元数据提取与更新: 脚本扫描包含 #Marker-待下载 的文件,提取 time 和 source,生成 YAML 格式的任务列表 [yyyy-mm-dd-hh 下载列表整理.md],并将原始文件标签更新为 #Marker-下载中-YYYYMMDD。
第二阶段:用户确认 (User Confirmation)
  1. 停止并检查: Agent 输出 YAML 列表文件路径,等待用户确认。
第三阶段:执行下载与剪藏 (Execution & Archival)
  1. 执行脚本: 调用 scripts/process_downloads.py,传入 --list-file 和 --archive-dir。
  2. 任务处理与分发:
    • YouTube / X (Twitter): 使用 yt-dlp 下载。请求最高画质,必须下载并保留全量 JSON 元数据(--write-info-json)。
    • 未知站点: 如果无法识别域名或未配置下载方式,则直接标记为“下载失败(未识别站点)”,不进行尝试。
  3. 隔离目录创建: 为每个下载任务创建独立目录,命名规范:[YYYY-MM-DD-HH] {分类} {描述/ID}。
  4. 内容本地化:
    • 主文档: 在目录下创建一个同名的 .md 文件。
    • 资产存放: 所有的 .mp4, .json, .jpg 等资产全部存放在该任务目录下。
    • 引用关联: Markdown 文件中使用本地相对路径链接同目录下的视频。
第四阶段:状态汇报与闭环 (Reporting & Closing)
  1. 列表回写: 脚本在 YAML 列表中更新状态为“下载完成”或“下载失败”。
  2. 标签同步: Agent 根据脚本输出,将原始文件中的链接标签更新为 #Marker-已下载-YYYYMMDD。

依赖工具

  • yt-dlp: 视频抓取。
  • MarkItDown / Pandoc: 网页转 Markdown。
  • ffmpeg: 视频合并。

最佳实践

  • 路径对齐: 始终从 agents.md 读取路径,不要硬编码。
  • 元数据保留: 在剪藏的 Markdown 头部记录原始 URL 和收集时间。
  • 异常容错: 下载失败时记录错误原因,不中断后续任务。

© cafe3310, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 20 other files (scripts, assets) in skills/online-content-collector of cafe3310/public-agent-skills.

  • SKILL.md
  • 2026-04-28-20-55-online-collector-requirements.md
  • assets/example_vault/Archives/Clips/[2026-05-02-15] 演示 dQw4w9WgXcQ/[2026-05-02-15] 演示 dQw4w9WgXcQ.md
  • assets/example_vault/Archives/Clips/[2026-05-02-15] 演示 dQw4w9WgXcQ/metadata.json.example
  • assets/example_vault/Archives/Clips/[2026-05-02-15] 演示 dQw4w9WgXcQ/video.mp4.example
  • assets/example_vault/Archives/Clips/[2026-05-02-18] 演示 2036268335927796152/[2026-05-02-18] 演示 2036268335927796152.md
  • assets/example_vault/Archives/Clips/[2026-05-02-18] 演示 2036268335927796152/image_1.jpg.example
  • assets/example_vault/Archives/Clips/[2026-05-02-18] 演示 2036268335927796152/metadata.json.example
  • assets/example_vault/DailyNotes/2026-05-01.md
  • assets/example_vault/DailyNotes/2026-05-02.md
  • assets/example_vault/DailyNotes/failure_test.md
  • assets/example_vault/Workflows/DownloadLists/2026-05-02-01 下载列表整理.md
  • … and 9 more

Open the folder on GitHubat commit 6c45501

Compare with similar skills

Online Content Collector next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Online Content Collector compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Online Content Collector this skillcafe3310/public-agent-skills255—~455Automated safety check: PassApache-2.0
Obsidian BasesAtmosphere/atmosphere3.8k22 repos~3.2kAutomated safety check: PassApache-2.0
Knap Markdown Templateskepano/obsidian-skills49k2 repos~986Automated safety check: PassMIT
JSON Canvasheyitsnoah/claudesidian2.6k18 repos~3.5kAutomated safety check: PassMIT
Obsidian MarkdownAtmosphere/atmosphere3.8k20 repos~1.3kAutomated safety check: PassApache-2.0
Obsidian CLIAtmosphere/atmosphere3.8k13 repos~795Automated safety check: PassApache-2.0

Similar skills

  • Obsidian Bases

    Atmosphere/atmosphere

    Create and edit Obsidian Bases (.base files) with views, filters, formulas, and summaries.

    3.8k GitHub starsUsed in 22 repos~3.2k tokens
    Backend & APIsAuto-check passed
  • Knap Markdown Templates

    kepano/obsidian-skills

    Renders Markdown notes from Knap templates and JSON data on the command line, including notes built from Defuddle web page output.

    49k GitHub starsUsed in 2 repos~986 tokens
    Documents & OfficeAuto-check passed
  • JSON Canvas

    heyitsnoah/claudesidian

    Create and edit JSON Canvas files (.canvas) with nodes, edges, groups, and connections.

    2.6k GitHub starsUsed in 18 repos~3.5k tokens
    DevelopmentAuto-check passed
  • Obsidian Markdown

    Atmosphere/atmosphere

    Create and edit Obsidian Flavored Markdown with wikilinks, embeds, callouts, properties, and other Obsidian-specific syntax.

    3.8k GitHub starsUsed in 20 repos~1.3k tokens
    Documents & OfficeAuto-check passed
  • Obsidian CLI

    Atmosphere/atmosphere

    Interact with Obsidian vaults using the Obsidian CLI to read, create, search, and manage notes, tasks, properties, and more.

    3.8k GitHub starsUsed in 13 repos~795 tokens
    Knowledge ManagementAuto-check passed
  • Obsidian Canvas Creator

    axtonliu/axton-obsidian-visual-skills

    Create Obsidian Canvas files from text content, supporting both MindMap and freeform layouts.

    3.6k GitHub stars~1.6k tokensUpdated 3 mo ago
    Auto-check passed

More from cafe3310/public-agent-skills

All 29 skills in this repo
  • Impeccable

    cafe3310/public-agent-skills

    A skill your agent uses when the user wants to design, redesign, shape, critique, audit, polish, clarify, distill, harden, optimize, adapt, animate, colorize, extract, or otherwise improve a…

    255 GitHub stars~4.9k tokensUpdated 3 mo ago
    Auto-check passed
  • Text Watermark Fountain

    cafe3310/public-agent-skills

    A specialized skill for embedding and extracting resilient watermarks in text by manipulating sentence lengths and using Fountain Codes.

    255 GitHub stars~921 tokensUpdated 3 mo ago
    Auto-check passed
  • Obsidian Todo Collector

    cafe3310/public-agent-skills

    从 Obsidian 知识库中扫描指定时间范围内未完成事件,生成/更新未完成事件整理文档. An agent skill from cafe3310/public-agent-skills.

    255 GitHub stars~575 tokensUpdated 3 mo ago
    Auto-check passed
  • Deep Research

    cafe3310/public-agent-skills

    一个全面、自主的深度研究框架。当用户请求对复杂主题、市场调研、技术格局进行深入的多维度调查,或需要大量网页浏览、数据合成和结构化报告的任何任务时,使用此技能。它协调子代理(subagents)并使用基于文件系统的状态管理来防止上下文膨胀。

    255 GitHub stars~996 tokensUpdated 3 mo ago
    Auto-check passed
  • Long Audio To Obsidian

    cafe3310/public-agent-skills

    将语音转写项目输出的复杂文件结构整理合并为适合 Obsidian 归档的 Markdown 文档. An agent skill from cafe3310/public-agent-skills.

    255 GitHub stars~772 tokensUpdated 3 mo ago
    Auto-check passed
  • Markdown New

    cafe3310/public-agent-skills

    通过 markdown.new API 将网页、整站或搜索结果转换为干净的 Markdown. An agent skill from cafe3310/public-agent-skills.

    255 GitHub stars~344 tokensUpdated 3 mo ago
    Auto-check passed

Works with

Questions about Online Content Collector

What does Online Content Collector do?

对 Obsidian 仓库进行自动素材媒体剪藏,本地化特定 tag 标注的网页、视频及附件. An agent skill from cafe3310/public-agent-skills. Online Content Collector is an agent skill from cafe3310/public-agent-skills.

How do I install Online Content Collector in Claude Code?

Run `npx skills add cafe3310/public-agent-skills --skill online-content-collector -a claude-code`. Or copy the skill folder (skills/online-content-collector in cafe3310/public-agent-skills) into .claude/skills/online-content-collector in your project. Claude Code loads it when a task matches its description.

How do I install Online Content Collector in Codex?

Run `npx skills add cafe3310/public-agent-skills --skill online-content-collector -a codex`. Or copy the skill folder (skills/online-content-collector in cafe3310/public-agent-skills) into .agents/skills/online-content-collector in your project. Codex loads it when a task matches its description.

Can I use Online Content Collector in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cafe3310/public-agent-skills --skill online-content-collector -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/online-content-collector, .gemini/skills/online-content-collector, .github/skills/online-content-collector and .opencode/skills/online-content-collector in your project.

What does Online Content Collector need to run?

SKILL.md names no scripts, command-line tools or credentials: Online Content Collector is instructions for the agent only. Our summary lists: Python 3.

Does Online Content Collector access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Online Content Collector safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Online Content Collector use?

Online Content Collector is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Online Content Collector use?

About 455 tokens (SKILL.md is roughly 1.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Online Content Collector?

Skills that share tags, products or a category with Online Content Collector: Obsidian Bases (Atmosphere/atmosphere, 3.8k stars), Knap Markdown Templates (kepano/obsidian-skills, 49k stars), JSON Canvas (heyitsnoah/claudesidian, 2.6k stars) and Obsidian Markdown (Atmosphere/atmosphere, 3.8k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Online Content Collector?

cafe3310 (a GitHub user) maintains it in cafe3310/public-agent-skills, which has 255 GitHub stars. The repository holds 29 skills in this directory. The repository was last updated on June 26, 2026.

Source: cafe3310/public-agent-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.