Agent skill

Sn Da Image Caption

by MichaelYang-lyx in MichaelYang-lyx/AIDABench

图片理解与数据提取 skill。当图片文件(.png/.jpg/.jpeg/.gif/.webp/.bmp)是主要输入且用户需要理解、提取数据或分析图片内容时使用。提供预配置的 caption 脚本(scripts/caption.py),通过 vision 模型将图片转为文本描述,无需额外配置 API Key。覆盖:(1) 通过 scripts/caption.py…

No licenceAuto-check passedDocuments & Office

Install Sn Da Image Caption

skills CLI
$ npx skills add MichaelYang-lyx/AIDABench --skill sn-da-image-caption -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install MichaelYang-lyx/AIDABench sn-da-image-caption --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/MichaelYang-lyx/AIDABench.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/sn-da-image-caption .claude/skills/sn-da-image-caption && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
sn-da-image-caption
GitHub stars
111
Used in
1 other repo
Token cost
~2k tokens
SKILL.md length
310 words
Files
2 (incl. scripts)
Skills in repo
39
Repo updated
First seen
Licence
None found

At a glance

图片理解与数据提取 skill。当图片文件(.png/.jpg/.jpeg/.gif/.webp/.bmp)是主要输入且用户需要理解、提取数据或分析图片内容时使用。提供预配置的 caption 脚本(scripts/caption.py),通过 vision 模型将图片转为文本描述,无需额外配置 API Key。覆盖:(1) 通过 scripts/caption.py…

  • Works in 3 steps: Run scripts/caption.py to get a text… → Parse the description into structured… → Analyze, visualize, or export
  • Tasks that involve Excel spreadsheets
  • SKILL.md covers Overview, scripts/caption.py — Image…, Calling from Python and Prompt Strategy, plus 5 more sections
  • Runs Python scripts from its folder; calls python3; needs SN_API_KEY and SN_VISION_API_KEY

What it does

Sn Da Image Caption is an agent skill from MichaelYang-lyx/AIDABench. 图片理解与数据提取 skill。当图片文件(.png/.jpg/.jpeg/.gif/.webp/.bmp)是主要输入且用户需要理解、提取数据或分析图片内容时使用。提供预配置的 caption 脚本(scripts/caption.py),通过 vision 模型将图片转为文本描述,无需额外配置 API Key。覆盖:(1) 通过 scripts/caption.py 对图表/表格/截图/流程图进行 caption,(2) 将 caption 文本解析为结构化 DataFrame,(3) 基于提取数据重新生成可视化图表,(4) 导出为 Excel/CSV。遇到以下任一情况就主动使用本 skill,不要自行猜测图片内容:①用户出现触发词:图片分析 / 图表提取 / 表格识别 / OCR / 图片描述 / 截图分析 / 图表数据 / 提取图片中的数据 / 图片转表格 / 识别图片 / image caption / extract data from image / chart analysis / table OCR;②用户上传或指定了图片文件(.png / .jpg / .jpeg / .gif / .webp /…

Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/caption.py`).

It sits in Documents & Office, covering Excel spreadsheets, DataFrames and CSV and tabular files. It works with Microsoft Excel. The repository describes itself as: Code for paper AIDABench: AI Data Analytics Benchmark.

When your agent uses it

  • Tasks that involve Excel spreadsheets
  • Tasks that involve DataFrames
  • Tasks that involve CSV and tabular files

Example prompts

  • “/sn-da-image-caption”

Requirements

  • Python 3
  • A credential in SN_API_KEY
  • A credential in SN_VISION_API_KEY

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Run scripts/caption.py to get a text description of the image
  2. Parse the description into structured data (DataFrame, etc.)
  3. Analyze, visualize, or export

What it can do on your machine

Read from SKILL.md and the folder at commit 6dd4206. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • SN_API_KEY
    • SN_VISION_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Sn Da Image Caption loads about 2k tokens when it runs. Until then it costs about 168 tokens; SKILL.md has 310 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~168
When it runs · the whole SKILL.md, loaded when a task matches
~2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 310 words (~1,957 tokens).

“Analyze, extract data from, or understand image files (.png, .jpg, .jpeg, .gif, .webp, .bmp). The core workflow:”

— opening of SKILL.md by MichaelYang-lyx
name
sn-da-image-caption

Read the full SKILL.md on GitHub

Files

SKILL.md and 1 other file (scripts) in skills/sn-da-image-caption of MichaelYang-lyx/AIDABench.

  • SKILL.md
  • scripts/caption.py

Open the folder on GitHubat commit 6dd4206

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in MichaelYang-lyx/AIDABench, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Sn Da Image Caption next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Sn Da Image Caption compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Sn Da Image Caption this skillMichaelYang-lyx/AIDABench1111 repos~2kAutomated safety check: PassNone
Tabular Cleanupgaasher/Agent-Loop-Skills174—~4kAutomated safety check: PassMIT
Codebookbrycewang-stanford/Auto-Empirical-Research-Skills4.5k—~527Automated safety check: NotesCustom licence
Convert Fileduckdb/duckdb-skills6001 repos~720Automated safety check: NotesMIT
Excel ParserHarryoung/efka104—~2.3kAutomated safety check: PassApache-2.0
CSV and Excel MergerOneWave-AI/claude-skills323—~1.6kAutomated safety check: PassMIT

Similar skills

  • Tabular Cleanup

    gaasher/Agent-Loop-Skills

    A skill your agent uses when the user has a messy tabular data dump (CSV/TSV/parquet/Excel/JSON) and wants it iteratively cleaned to an inferred data contract — a checklist of deterministic…

    174 GitHub stars~4k tokensUpdated 3 mo ago
    Documents & OfficeAuto-check passed
  • Codebook

    brycewang-stanford/Auto-Empirical-Research-Skills

    Auto-generates a Markdown codebook from a dataset (CSV, DTA, Excel, Parquet) with types and summary statistics.

    4.5k GitHub stars~527 tokensUpdated 3 days ago
    Documents & OfficeAuto-check: notes
  • Convert File

    duckdb/duckdb-skills

    Official

    Convert any data file to another format: CSV, Parquet, JSON, Excel, GeoJSON, and more.

    600 GitHub starsUsed in 1 repo~720 tokens
    Documents & OfficeAuto-check: notes
  • Excel Parser

    Harryoung/efka

    Smart Excel/CSV file parsing with intelligent routing based on file complexity analysis.

    104 GitHub stars~2.3k tokensUpdated 6 mo ago
    Documents & OfficeAuto-check passed
  • CSV and Excel Merger

    OneWave-AI/claude-skills

    Combines CSV, TSV and Excel files into one verified table with pandas, by stacking or joining, mapping columns, normalizing keys and removing duplicates.

    323 GitHub stars~1.6k tokensUpdated 6 days ago
    Documents & OfficeAuto-check passed
  • Generate Codebook

    Aperivue/medsci-skills

    A skill your agent uses when a tabular dataset (CSV, Excel, Parquet, Stata, SAS) needs a data dictionary.

    329 GitHub stars~1.1k tokensUpdated 3 days ago
    Documents & OfficeAuto-check passed

More from MichaelYang-lyx/AIDABench

All 39 skills in this repo
  • 对Excel数据进行自定义分类统计、交叉分析与可视化,并基于多维度指标(如文本长度、术语密度、正则匹配等)进行综合评分与分级,适用于多类别数据分布统计及文本内容难度/质量评估场景。

    111 GitHub starsUsed in 2 repos~1.6k tokens
    Auto-check passed
  • Excel Bar Chart Visualization

    MichaelYang-lyx/AIDABench

    读取多工作表Excel文件,自动处理合并单元格与数据清洗,进行交叉分组统计并生成带总计行的结果表,最后绘制支持中英文字体的美化柱状图,适用于多维度数据汇总与可视化分析。

    111 GitHub starsUsed in 1 repo~991 tokens
    Auto-check passed
  • 执行全面的异常值检测与数据质量评估,利用 IQR 方法识别异常值并结合偏度、峰度分析数据分布特征,适用于非正态分布数据的预处理阶段。

    111 GitHub starsUsed in 1 repo~1k tokens
    Auto-check passed
  • Sn Da Excel Workflow

    MichaelYang-lyx/AIDABench

    Excel 数据分析多步编排器。覆盖:(1) 读取多 Sheet Excel 文件并统计行数,(2) 大文件检测(≥10k 行自动 Parquet 优化),(3) 数据清洗(缺失值、文本标准化、无效字符),(4) 条件筛选与分类提取,(5) 跨 Sheet 统计聚合,(6) 导出 Excel/CSV 并提供下载链接。覆盖从数据读取到报告生成全流程,按步骤编排 capability 子…

    111 GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed
  • Sn Da Large File Analysis

    MichaelYang-lyx/AIDABench

    万行以上 Excel 数据集的高性能分析引擎。提供 openpyxl readonly 流式读取(iterrows 支持 10 万行以上)、Parquet 转换加速、内存优化、分块处理和大文件写入模式。遇到以下任一情况就主动使用本 skill:①数据行数 ≥ 10k(由 sn-da-excel-workflow 的行数评估步骤触发);②用户出现触发词:大文件 / 大数据量 / 性能优化 /…

    111 GitHub starsUsed in 1 repo~3.1k tokens
    Auto-check passed
  • Single Sheet Reading And Analysis

    MichaelYang-lyx/AIDABench

    读取并解析单个Excel工作表数据,支持合并单元格处理、数据清洗、交叉分析及多维度可视化,适用于需要从单表中提取关键指标并进行趋势模拟与图表生成的场景。

    111 GitHub starsUsed in 2 repos~943 tokens
    Auto-check passed

Works with

Questions about Sn Da Image Caption

What does Sn Da Image Caption do?

图片理解与数据提取 skill。当图片文件(.png/.jpg/.jpeg/.gif/.webp/.bmp)是主要输入且用户需要理解、提取数据或分析图片内容时使用。提供预配置的 caption 脚本(scripts/caption.py),通过 vision 模型将图片转为文本描述,无需额外配置 API Key。覆盖:(1) 通过 scripts/caption.py…. Sn Da Image Caption is an agent skill from MichaelYang-lyx/AIDABench.

When should I use Sn Da Image Caption?

Sn Da Image Caption fits situations like: tasks that involve Excel spreadsheets; tasks that involve DataFrames; tasks that involve CSV and tabular files.

How do I install Sn Da Image Caption in Claude Code?

Run `npx skills add MichaelYang-lyx/AIDABench --skill sn-da-image-caption -a claude-code`. Or copy the skill folder (skills/sn-da-image-caption in MichaelYang-lyx/AIDABench) into .claude/skills/sn-da-image-caption in your project. Claude Code loads it when a task matches its description.

How do I install Sn Da Image Caption in Codex?

Run `npx skills add MichaelYang-lyx/AIDABench --skill sn-da-image-caption -a codex`. Or copy the skill folder (skills/sn-da-image-caption in MichaelYang-lyx/AIDABench) into .agents/skills/sn-da-image-caption in your project. Codex loads it when a task matches its description.

Can I use Sn Da Image Caption in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add MichaelYang-lyx/AIDABench --skill sn-da-image-caption -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/sn-da-image-caption, .gemini/skills/sn-da-image-caption, .github/skills/sn-da-image-caption and .opencode/skills/sn-da-image-caption in your project.

What does Sn Da Image Caption need to run?

Going by SKILL.md and its folder, Sn Da Image Caption needs Python for the scripts in its folder, the command-line tools its instructions call (python3) and credentials named SN_API_KEY and SN_VISION_API_KEY. Our summary lists: Python 3; A credential in SN_API_KEY; A credential in SN_VISION_API_KEY.

Does Sn Da Image Caption access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Sn Da Image Caption safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Sn Da Image Caption use?

No licence was found for Sn Da Image Caption or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Sn Da Image Caption use?

About 2k tokens (SKILL.md is roughly 7.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Sn Da Image Caption?

Skills that share tags, products or a category with Sn Da Image Caption: Tabular Cleanup (gaasher/Agent-Loop-Skills, 174 stars), Codebook (brycewang-stanford/Auto-Empirical-Research-Skills, 4.5k stars), Convert File (duckdb/duckdb-skills, 600 stars) and Excel Parser (Harryoung/efka, 104 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Sn Da Image Caption?

MichaelYang-lyx (a GitHub user) maintains it in MichaelYang-lyx/AIDABench, which has 111 GitHub stars. The repository holds 39 skills in this directory. The repository was last updated on September 28, 2026.

Source: MichaelYang-lyx/AIDABench on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.