Exploratory Data Analysis
spacering-net/codeg
Perform comprehensive exploratory data analysis on scientific data files across 200+ file formats.
Complete data-analysis tasks with bounded inspection, correct data semantics, native artifact handling, complete delivery, and risk-based verification.
$ npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Prism-Shadow/penguin-harness data-analysis --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Prism-Shadow/penguin-harness.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/data-analysis/skills/data-analysis .claude/skills/data-analysis && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "data-analysis" agent skill from https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysis into .claude/skills/data-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-analysis", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysisType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Prism-Shadow/penguin-harness data-analysis --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prism-Shadow/penguin-harness.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/data-analysis/skills/data-analysis .agents/skills/data-analysis && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "data-analysis" agent skill from https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysis into .agents/skills/data-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-analysis", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Prism-Shadow/penguin-harness data-analysis --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prism-Shadow/penguin-harness.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/data-analysis/skills/data-analysis .cursor/skills/data-analysis && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "data-analysis" agent skill from https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysis into .cursor/skills/data-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-analysis", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Prism-Shadow/penguin-harness.git --path plugins/data-analysis/skills/data-analysis--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Prism-Shadow/penguin-harness data-analysis --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prism-Shadow/penguin-harness.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/data-analysis/skills/data-analysis .gemini/skills/data-analysis && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "data-analysis" agent skill from https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysis into .gemini/skills/data-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-analysis", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Prism-Shadow/penguin-harness data-analysisInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Prism-Shadow/penguin-harness.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/data-analysis/skills/data-analysis .github/skills/data-analysis && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "data-analysis" agent skill from https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysis into .github/skills/data-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-analysis", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Prism-Shadow/penguin-harness data-analysis --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prism-Shadow/penguin-harness.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/data-analysis/skills/data-analysis .opencode/skills/data-analysis && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "data-analysis" agent skill from https://github.com/Prism-Shadow/penguin-harness/tree/main/plugins/data-analysis/skills/data-analysis into .opencode/skills/data-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-analysis", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
data-analysisComplete data-analysis tasks with bounded inspection, correct data semantics, native artifact handling, complete delivery, and risk-based verification.
Data Analysis is an agent skill from Prism-Shadow/penguin-harness. Complete data-analysis tasks with bounded inspection, correct data semantics, native artifact handling, complete delivery, and risk-based verification.
Its SKILL.md is about 940 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Data & Analytics, covering Data analysis. The repository describes itself as: 🐧 Unified and Stable RSI Platform. The licence is Apache-2.0.
Read from SKILL.md and the folder at commit d56d9ce. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Data Analysis loads about 935 tokens when it runs. Until then it costs about 41 tokens; SKILL.md has 489 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Prism-Shadow/penguin-harness at commit d56d9ce, republished under its Apache-2.0 licence (© Prism-Shadow). 489 words, ~935 tokens.
.claude/skills/data-analysis/SKILL.md (or your agent's skills folder).Deliver the requested result and artifacts. Do not turn the task into a proof exercise or add evidence, reports, explanations, or intermediate files that were not requested.
Require a concrete data-analysis task, its available inputs, and the requested deliverable, location, and format. Ask only when missing information prevents a defensible result and would materially change the deliverable; otherwise proceed.
Read the task, supplied inputs, and relevant data documentation. Identify every required output path and format, plus only the definitions that can change the result: scope, observation grain, keys, units, operators, ordering, coverage, and explicit formatting rules. Treat examples as illustrative unless the task makes them normative.
If information is incomplete or ambiguous, first resolve it from the supplied materials. Ask only when the missing choice prevents a defensible result and would materially change the deliverable. Otherwise choose the best-supported interpretation and proceed.
For large or unfamiliar inputs, begin with a bounded inventory, schema check, targeted sample, or narrow query. Expand inspection only when it can change a selection, transformation, calculation, or output. Do not exhaustively read or render data merely to increase confidence.
Compute at the correct row or entity grain. Evaluate conjunctive conditions on the same record or entity; do not replace row-level matching with unions of separate field values. Preserve nulls, exclusions, and explicit prohibitions. Enumerated outputs must cover the complete requested universe.
Ground answer-changing choices in the task and supplied data. Preserve documented source semantics, units, mappings, and native workflow behavior when they define the requested result. Do not reproduce an apparent source or tool defect merely for consistency. When plausible methods disagree, compare only the smallest answer-changing difference, choose the best-supported method, and use it consistently.
Preserve the requested artifact type and structure. When correctness depends on spreadsheet formulas, recalculation, formatting, database semantics, document layout, or export behavior, prefer a tool path that preserves and can verify those native properties. Restore temporarily changed inputs or formulas before finalizing. Use intermediate files only when they help produce or verify the requested deliverable.
As soon as a complete best-supported result exists, write every requested artifact at its exact path. For a multi-artifact task, establish a valid version of every artifact before refining any one of them. Do not leave a required artifact missing while pursuing additional certainty, polish, or diagnostics. If later evidence changes the result, update the artifact.
Choose checks in proportion to answer-changing risk. Use the smallest independent check that can falsify each load-bearing assumption or computation. If a check disagrees, isolate and resolve the concrete difference. Do not repeat equivalent searches, calculations, renders, or inspections once remaining uncertainty cannot change the deliverable.
Reopen the actual deliverables and verify their path, format, schema or structure, values, coverage, and openability as applicable. Confirm that every requested artifact exists and reflects the chosen method. Report the output paths concisely and stop.
© Prism-Shadow, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in plugins/data-analysis/skills/data-analysis of Prism-Shadow/penguin-harness.
Open the folder on GitHubat commit d56d9ce
Data Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Data Analysis this skillPrism-Shadow/penguin-harness | 2.5k | — | ~935 | Automated safety check: Pass | Apache-2.0 | |
| Exploratory Data Analysisspacering-net/codeg | 3.8k | 15 repos | ~3.6k | Automated safety check: Pass | MIT | |
| Excel and CSV Data Analysisbytedance/deer-flow | 83k | 4 repos | ~2.2k | Automated safety check: Pass | MIT | |
| Exploratory Data AnalysisOleafly/Oleafly | 205 | 2 repos | ~3.4k | Automated safety check: Notes | MIT | |
| Python Executorcortega26/chile-hub | 113 | 2 repos | ~1.5k | Automated safety check: Pass | MIT | |
| Agentic Kaggle WorkflowFrankS-IntelLab/agentic-kaggle-skill | 188 | — | ~4k | Automated safety check: Pass | MIT |
spacering-net/codeg
Perform comprehensive exploratory data analysis on scientific data files across 200+ file formats.
bytedance/deer-flow
Analyzes uploaded Excel and CSV files with SQL through DuckDB, producing schema inspections, statistical summaries and exports to CSV, JSON or Markdown.
Oleafly/Oleafly
Perform bounded, local exploratory analysis of explicitly supported scientific files.
cortega26/chile-hub
Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh).
FrankS-IntelLab/agentic-kaggle-skill
Takes a Kaggle competition from rules and validation design through baselines, ensembling and notebook architecture to a scored submission.
mcncarl/yichen-skills
Read, decrypt, query, search, and export local WeCom/企业微信 5.x desktop databases on macOS into a private read-only vault.
Prism-Shadow/penguin-harness
Make a reply easier to read and act on with rich blocks inside ordinary Markdown — a choice the user picks from, a form that collects several answers, a procedure as steps with warnings in place, a…
Prism-Shadow/penguin-harness
A skill your agent uses when developing PenguinHarness itself — changing packages/{core,server,web,cli,desktop,landing,docs,skills}, the built-in model catalog, the installers or the release…
Prism-Shadow/penguin-harness
Create and edit Bento presentations — self-contained .bento.html decks whose document is JSON.
Prism-Shadow/penguin-harness
A skill your agent uses when standing PenguinHarness up to try a change by hand — launching the Web App, the desktop shell, the landing page, the docs site or the component gallery to click through…
Prism-Shadow/penguin-harness
A skill your agent uses when changing the PenguinHarness Web App (packages/web) or the shared UI package — adding or restyling any UI, picking a status colour, adding an icon, laying out a row or a…
Prism-Shadow/penguin-harness
Drive the PenguinHarness agent browser — the desktop app's built-in browser or the user's own Chrome — from the shell with penguin browser: open pages, read them as simplified HTML or text, act with…
Categories
Complete data-analysis tasks with bounded inspection, correct data semantics, native artifact handling, complete delivery, and risk-based verification. Data Analysis is an agent skill from Prism-Shadow/penguin-harness. Complete data-analysis tasks with bounded inspection, correct data semantics, native artifact handling, complete delivery, and risk-based verification.
Data Analysis fits situations like: tasks that involve Data analysis.
Run `npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a claude-code`. Or copy the skill folder (plugins/data-analysis/skills/data-analysis in Prism-Shadow/penguin-harness) into .claude/skills/data-analysis in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a codex`. Or copy the skill folder (plugins/data-analysis/skills/data-analysis in Prism-Shadow/penguin-harness) into .agents/skills/data-analysis in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Prism-Shadow/penguin-harness --skill data-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/data-analysis, .gemini/skills/data-analysis, .github/skills/data-analysis and .opencode/skills/data-analysis in your project.
SKILL.md names no scripts, command-line tools or credentials: Data Analysis is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Data Analysis is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 935 tokens (SKILL.md is roughly 3.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Data Analysis: Exploratory Data Analysis (spacering-net/codeg, 3.8k stars), Excel and CSV Data Analysis (bytedance/deer-flow, 83k stars), Exploratory Data Analysis (Oleafly/Oleafly, 205 stars) and Python Executor (cortega26/chile-hub, 113 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Prism-Shadow (a GitHub organization) maintains it in Prism-Shadow/penguin-harness, which has 2,450 GitHub stars. The repository holds 31 skills in this directory. The repository was last updated on October 7, 2026.
Source: Prism-Shadow/penguin-harness on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.