Agent skill

Archive Grounding

by gaotiexinqu in gaotiexinqu/OneResearchClaw

Unpack a ZIP archive, inventory its files, run the corresponding child grounding skill for each supported child file, and then write a real archive-level grounded.md.

MITAuto-check: notesDocuments & Office

Install Archive Grounding

skills CLI
$ npx skills add gaotiexinqu/OneResearchClaw --skill archive-grounding -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gaotiexinqu/OneResearchClaw archive-grounding --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gaotiexinqu/OneResearchClaw.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.cursor/skills/archive-grounding .claude/skills/archive-grounding && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
archive-grounding
GitHub stars
450
Token cost
~2.1k tokens
SKILL.md length
951 words
Files
3 (incl. scripts)
Skills in repo
15
Repo updated
First seen
Licence
MIT

At a glance

Unpack a ZIP archive, inventory its files, run the corresponding child grounding skill for each supported child file, and then write a real archive-level grounded.md.

  • Works in 6 steps: unpack the archive, → build a stable archive bundle, → identify supported child files, → …
  • Documents & Office work in your project
  • SKILL.md covers Position in the pipeline, Supported child file routing, Required Workflow and Input, plus 5 more sections
  • Runs Python and Shell scripts from its folder; calls bash

What it does

Archive Grounding is an agent skill from gaotiexinqu/OneResearchClaw. Unpack a ZIP archive, inventory its files, run the corresponding child grounding skill for each supported child file, and then write a real archive-level grounded.md.

Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including scripts (for example `scripts/ground_archive.py` and `scripts/run.sh`).

It sits in Documents & Office. It works with Microsoft PowerPoint. The repository describes itself as: Any research. One Claw. 🦞 From any materials to research with fully autonomous & skill-driven researcher. The licence is MIT.

When your agent uses it

  • Documents & Office work in your project

Example prompts

  • “/archive-grounding”

Requirements

  • Python 3
  • A Bash shell
  • Pre-approved tools (allowed-tools): Bash, Read, Write, Edit, Grep, Glob

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. unpack the archive,
  2. build a stable archive bundle,
  3. identify supported child files,
  4. run the corresponding child skill for each supported child file,
  5. collect those child bundles under the current archive bundle's child_outputs/ directory,
  6. and only then write a real archive-level grounded.md.

What it can do on your machine

Read from SKILL.md and the folder at commit 37e86c6. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash
    • Read
    • Write
    • Edit
    • Grep
    • Glob

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 2 files in scripts/ (Python and Shell), which the agent can run.

    Shell commands in SKILL.md call:

    • bash

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Archive Grounding loads about 2.1k tokens when it runs. Until then it costs about 46 tokens; SKILL.md has 951 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~46
When it runs · the whole SKILL.md, loaded when a task matches
~2.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash, Read, Write, Edit, Grep, Glob

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from gaotiexinqu/OneResearchClaw at commit 37e86c6, republished under its MIT licence (© gaotiexinqu). 951 words, ~2,075 tokens.

Download SKILL.mdSave it as .claude/skills/archive-grounding/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
archive-grounding
description
Unpack a ZIP archive, inventory its files, run the corresponding child grounding skill for each supported child file, and then write a real archive-level grounded.md.
allowed-tools
Bash, Read, Write, Edit, Grep, Glob

Archive Grounding

This skill handles ZIP archives as container inputs in the any-input -> grounding -> downstream research / summary / report pipeline.

Its job is not to directly pretend that the whole archive has already been deeply grounded after unpacking. Its job is to:

  1. unpack the archive,
  2. build a stable archive bundle,
  3. identify supported child files,
  4. run the corresponding child skill for each supported child file,
  5. collect those child bundles under the current archive bundle's child_outputs/ directory,
  6. and only then write a real archive-level grounded.md.

Position in the pipeline

This skill is for archives that package together multiple materials, such as:

  • project material bundles
  • meeting material bundles
  • report + slides + tables + audio/video attachments
  • mixed research input packages

This skill is not a replacement for the child grounding skills themselves.

  • document-grounding still handles document content
  • table-grounding still handles spreadsheets / CSV tables
  • pptx-grounding still handles PowerPoint decks
  • meeting-audio-grounding handles meeting audio inputs
  • meeting-video-grounding handles meeting videos
  • audio_structuring remains the atomic audio transcription backend reused by the higher-level audio/video meeting entry skills

This skill acts as an orchestrator over packaged files.


Supported child file routing

The first version should route supported child files using a simple extension-based mapping:

  • .pdf, .docx, .md, .txt -> document-grounding
  • .xlsx, .csv -> table-grounding
  • .pptx -> pptx-grounding
  • .mp3, .wav, .m4a -> meeting-audio-grounding
  • .mp4, .mov, .mkv -> meeting-video-grounding

Unsupported files may be recorded and skipped, but they must not be silently treated as grounded.


Required Workflow

When using this skill, you must follow this workflow:

  1. First run the existing script entrypoint:

    bash
    bash .cursor/skills/archive-grounding/scripts/run.sh <input_zip> <output_root>
  2. The script generates an archive bundle skeleton and inventory files such as:

    • extracted.md
    • extracted_meta.json
    • manifest.json
    • routed_items.json
    • unpacked/
    • child_outputs/
  3. After the archive bundle is generated, read:

    • extracted.md
    • extracted_meta.json
    • manifest.json
    • routed_items.json
  4. Then enumerate the supported child files in the archive.

  5. For each supported child file, you must actually run the corresponding child skill. Do not stop after inventory / route planning.

  6. For child files inside the archive, all downstream child grounding outputs must be written inside the current archive bundle's child_outputs/ directory, not into the global grounding root.

  7. Use the recommended child output path recorded in routed_items.json when present.

  8. A child file is only considered completed if one of the following is true:

    • its corresponding child bundle was actually generated under child_outputs/, or
    • a concrete failure reason is explicitly recorded.
  9. Do not pretend that a child file was fully grounded merely because it was detected, inventoried, or assigned a route.

  10. Only after the supported child files have been processed (or explicit failures have been recorded) may you write the archive-level grounded.md.

  11. Do not create a placeholder grounded.md.

  12. The task is not complete if:

  • only the archive bundle skeleton exists,
  • manifest.json / routed_items.json exist but no child skills were actually run,
  • child bundles were written to the global grounding root instead of this archive bundle's child_outputs/, or
  • archive-level grounded.md was written before the child processing step was completed.

Input

bash
bash .cursor/skills/archive-grounding/scripts/run.sh <input_zip> <output_root>
Arguments
  • input_zip: path to a .zip archive
  • output_root: parent directory under which the archive bundle should be created

Example:

bash
bash .cursor/skills/archive-grounding/scripts/run.sh \
  /path/to/materials.zip \
  data/grounded_notes

This will create something like:

text
/data/grounded_notes/archive-materials/

Output bundle contract

The script should create:

text
<archive_bundle>/
├─ extracted.md
├─ extracted_meta.json
├─ manifest.json
├─ routed_items.json
├─ unpacked/
├─ child_outputs/
└─ grounded.md   # written later by the agent, not by the script
File roles
extracted.md

Human-readable archive overview for the agent.

It should include:

  • archive overview
  • detected file inventory summary
  • supported vs skipped items
  • routed child skills
  • the rule that child outputs must live under child_outputs/
extracted_meta.json

Global archive metadata, such as:

  • source archive path
  • archive id
  • total file count
  • supported file count
  • skipped file count
  • unpack directory
  • generation timestamp
Show full SKILL.md (378 more words)Show less
manifest.json

Machine-readable file inventory for unpacked contents.

Each item should include:

  • relative path
  • file name
  • extension
  • size
  • detected type
  • supported / unsupported
  • skip reason if any
routed_items.json

Machine-readable routing plan for supported child files.

Each routed item should include:

  • source relative path
  • detected type
  • routed skill
  • status
  • recommended child output path
  • notes / failure reason if any
child_outputs/

Container directory for all downstream child grounding bundles generated from supported child files in this archive.

Child grounding outputs must be stored here rather than in the global grounding root.


After the child files have been processed and the child bundles exist, the agent should write a stable archive grounding note such as:

markdown
# Archive Grounding

## 1. Archive Overview

## 2. Included Materials

## 3. Successfully Processed Child Items

## 4. Key Signals Across Materials

## 5. Skipped / Unsupported / Failed Items

## 6. Suggested Next Steps

## 7. Search Keywords

This is an archive-level grounding note, not a polished final report.


Important rule on absent file types

Do not treat the absence of a file type as a risk by default.

Examples:

  • if the archive simply does not contain PPTX files, do not mark that as a missing item
  • if the archive simply does not contain audio or video files, do not mark that as a missing item
  • if the archive simply does not contain tables, do not mark that as a missing item

Only report:

  • actual unsupported items,
  • actual skipped items,
  • actual failed child processing steps,
  • or actual inconsistencies between the archive contents and the generated child outputs.

Quality bar

A good result means:

  • the archive bundle exists
  • the file inventory is correct
  • every supported child file was either actually processed by its corresponding child skill or recorded with a concrete failure reason
  • child bundles are stored under this archive bundle's child_outputs/
  • the archive-level grounded.md is written only after child processing is complete
  • the final archive note is useful as a stable intermediate artifact for downstream research / summary / report

A bad result means:

  • the agent only unpacked and inventoried the archive
  • the agent wrote archive-level grounded.md without running the child skills
  • child outputs were scattered into the global grounding root
  • unsupported or skipped files were silently treated as processed
  • placeholder grounded.md was created

Failure handling

If a supported child file cannot be processed successfully, record that explicitly.

Examples:

  • child skill missing
  • child script failed
  • child bundle path not created
  • file appears corrupted
  • environment dependency missing

Do not hide child-processing failures behind a fake “archive success”.

© gaotiexinqu, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (scripts) in .cursor/skills/archive-grounding of gaotiexinqu/OneResearchClaw.

  • SKILL.md
  • scripts/ground_archive.py
  • scripts/run.sh

Open the folder on GitHubat commit 37e86c6

Compare with similar skills

Archive Grounding next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Archive Grounding compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Archive Grounding this skillgaotiexinqu/OneResearchClaw450—~2.1kAutomated safety check: NotesMIT
MarkitdownImCa0/just-laws78114 repos~3.2kAutomated safety check: NotesMIT
Image To Editable Pptningzimu/image-to-editable-ppt-skill2.9k—~4.3kAutomated safety check: PassMIT
GenOffice Document CLIgenspark-ai/genoffice9.2k—~19kAutomated safety check: PassApache-2.0
PPTX PostersK-Dense-AI/claude-scientific-writer2.4k2 repos~2.7kAutomated safety check: NotesMIT
Slidesfcakyon/claude-codex-settings1.2k1 repos~1.1kAutomated safety check: PassMIT

Similar skills

  • Markitdown

    ImCa0/just-laws

    Convert files and office documents to Markdown. An agent skill from ImCa0/just-laws.

    781 GitHub starsUsed in 14 repos~3.2k tokens
    Documents & OfficeAuto-check: notes
  • Image To Editable Ppt

    ningzimu/image-to-editable-ppt-skill

    Rebuild slide images, scanned or image-based PPT/PPTX files, and PDF decks into object-level editable PowerPoint (.pptx), preserving speaker notes when supplied.

    2.9k GitHub stars~4.3k tokensUpdated 25 days ago
    Documents & OfficeAuto-check passed
  • GenOffice Document CLI

    genspark-ai/genoffice

    Creates, converts, reads and edits real pptx, xlsx, docx and PDF files locally through the genoffice command line.

    9.2k GitHub stars~19k tokensUpdated yesterday
    Documents & OfficeAuto-check passed
  • PPTX Posters

    K-Dense-AI/claude-scientific-writer

    Create and audit editable scientific posters in macro-free PowerPoint (.pptx) from author-approved local content and assets.

    2.4k GitHub starsUsed in 2 repos~2.7k tokens
    Documents & OfficeAuto-check: notes
  • Slides

    fcakyon/claude-codex-settings

    Create and edit presentation slide decks (.pptx) with PptxGenJS, bundled layout helpers, and render/validation utilities.

    1.2k GitHub starsUsed in 1 repo~1.1k tokens
    Documents & OfficeAuto-check passed
  • Ppt Image First

    NyxTides/ppt-image-first

    Build presentation plans for PPT / slides / decks through a conversation-first workflow, then propose multiple visual directions with preview images before writing deck specs.

    1.2k GitHub stars~1.6k tokensUpdated 5 mo ago
    Documents & OfficeAuto-check passed

More from gaotiexinqu/OneResearchClaw

All 15 skills in this repo
  • Remote Input

    gaotiexinqu/OneResearchClaw

    Download remote content (arxiv papers, YouTube videos, Bilibili videos) to local storage and route to downstream grounding pipeline.

    450 GitHub stars~2.8k tokensUpdated 5 mo ago
    Auto-check passed
  • Grounded Research Lit

    gaotiexinqu/OneResearchClaw

    Run focused literature and web research from a grounded note.

    450 GitHub stars~11k tokensUpdated 5 mo ago
    Auto-check passed
  • Document Grounding

    gaotiexinqu/OneResearchClaw

    Convert a raw document into a structured grounding note for downstream research and summarization.

    450 GitHub stars~2k tokensUpdated 5 mo ago
    Auto-check passed
  • Meeting Audio Grounding

    gaotiexinqu/OneResearchClaw

    Convert a meeting audio file into a transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

    450 GitHub stars~1.1k tokensUpdated 5 mo ago
    Auto-check passed
  • Meeting Video Grounding

    gaotiexinqu/OneResearchClaw

    Convert a meeting video into an audio-first transcript bundle, then use meeting-grounding to produce structured meeting grounding outputs.

    450 GitHub stars~1.2k tokensUpdated 5 mo ago
    Auto-check passed
  • PPTX Grounding

    gaotiexinqu/OneResearchClaw

    Extract a structured evidence bundle from a .pptx deck, then write a real grounded.md from the bundle.

    450 GitHub stars~2.1k tokensUpdated 5 mo ago
    Auto-check: notes

Questions about Archive Grounding

What does Archive Grounding do?

Unpack a ZIP archive, inventory its files, run the corresponding child grounding skill for each supported child file, and then write a real archive-level grounded.md. Archive Grounding is an agent skill from gaotiexinqu/OneResearchClaw.md.

When should I use Archive Grounding?

Archive Grounding fits situations like: documents & Office work in your project.

How do I install Archive Grounding in Claude Code?

Run `npx skills add gaotiexinqu/OneResearchClaw --skill archive-grounding -a claude-code`. Or copy the skill folder (.cursor/skills/archive-grounding in gaotiexinqu/OneResearchClaw) into .claude/skills/archive-grounding in your project. Claude Code loads it when a task matches its description.

How do I install Archive Grounding in Codex?

Run `npx skills add gaotiexinqu/OneResearchClaw --skill archive-grounding -a codex`. Or copy the skill folder (.cursor/skills/archive-grounding in gaotiexinqu/OneResearchClaw) into .agents/skills/archive-grounding in your project. Codex loads it when a task matches its description.

Can I use Archive Grounding in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gaotiexinqu/OneResearchClaw --skill archive-grounding -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/archive-grounding, .gemini/skills/archive-grounding, .github/skills/archive-grounding and .opencode/skills/archive-grounding in your project.

What does Archive Grounding need to run?

Going by SKILL.md and its folder, Archive Grounding needs Python and a shell for the scripts in its folder and the command-line tools its instructions call (bash). Our summary lists: Python 3; A Bash shell. Its frontmatter pre-approves these tools: Bash, Read, Write, Edit, Grep, Glob.

Does Archive Grounding access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Archive Grounding safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Archive Grounding use?

Archive Grounding is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Archive Grounding use?

About 2.1k tokens (SKILL.md is roughly 8.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Archive Grounding?

Skills that share tags, products or a category with Archive Grounding: Markitdown (ImCa0/just-laws, 781 stars), Image To Editable Ppt (ningzimu/image-to-editable-ppt-skill, 2.9k stars), GenOffice Document CLI (genspark-ai/genoffice, 9.2k stars) and PPTX Posters (K-Dense-AI/claude-scientific-writer, 2.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Archive Grounding?

gaotiexinqu (a GitHub user) maintains it in gaotiexinqu/OneResearchClaw, which has 450 GitHub stars. The repository holds 15 skills in this directory. The repository was last updated on May 9, 2026.

Source: gaotiexinqu/OneResearchClaw on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.