Agent skill

Agent Survey Corpus

by WILLOSCAR in WILLOSCAR/research-units-pipeline-skills

Download a small corpus of open-access arXiv survey/review PDFs about agentic systems and extract text for style learning.

No licenceAuto-check passedResearch & Science

Install Agent Survey Corpus

skills CLI
$ npx skills add WILLOSCAR/research-units-pipeline-skills --skill agent-survey-corpus -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install WILLOSCAR/research-units-pipeline-skills agent-survey-corpus --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/WILLOSCAR/research-units-pipeline-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.codex/skills/agent-survey-corpus .claude/skills/agent-survey-corpus && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
agent-survey-corpus
GitHub stars
513
Token cost
~628 tokens
SKILL.md length
260 words
Files
2 (incl. scripts)
Skills in repo
107
Repo updated
First seen
Licence
None found

At a glance

Download a small corpus of open-access arXiv survey/review PDFs about agentic systems and extract text for style learning.

  • Works in 3 steps: Edit ref/agent-surveys/arxiv_ids.txt… → Run the downloader to fetch PDFs and… → Skim the extracted text under…
  • Tasks that involve Academic paper search
  • SKILL.md covers Triggers & routing, Inputs, Outputs and Workflow, plus 2 more sections
  • Runs Python scripts from its folder; calls uv

What it does

Agent Survey Corpus is an agent skill from WILLOSCAR/research-units-pipeline-skills. Download a small corpus of open-access arXiv survey/review PDFs about agentic systems and extract text for style learning.

Its SKILL.md is about 630 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/run.py`).

It sits in Research & Science, covering Academic paper search. It works with arXiv and Python. The repository describes itself as: Research pipelines as semantic execution units: each skill declares inputs/outputs, acceptance criteria, and guardrails. Evidence-first methodology prevents hollow writing…

When your agent uses it

  • Tasks that involve Academic paper search

Example prompts

  • “/agent-survey-corpus”

Requirements

  • Python 3

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Edit ref/agent-surveys/arxiv_ids.txt (one arXiv id per line).
  2. Run the downloader to fetch PDFs and extract the first N pages to text.
  3. Skim the extracted text under ref/agent-surveys/text/

What it can do on your machine

Read from SKILL.md and the folder at commit c92912a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Agent Survey Corpus loads about 628 tokens when it runs. Until then it costs about 36 tokens; SKILL.md has 260 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~36
When it runs · the whole SKILL.md, loaded when a task matches
~628

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 260 words (~628 tokens).

“Goal: create a small, local reference library so you can learn from real agent surveys when refining:”

— opening of SKILL.md by WILLOSCAR
name
agent-survey-corpus
binding
library

Read the full SKILL.md on GitHub

Files

SKILL.md and 1 other file (scripts) in .codex/skills/agent-survey-corpus of WILLOSCAR/research-units-pipeline-skills.

  • SKILL.md
  • scripts/run.py

Open the folder on GitHubat commit c92912a

Compare with similar skills

Agent Survey Corpus next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Agent Survey Corpus compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Agent Survey Corpus this skillWILLOSCAR/research-units-pipeline-skills513—~628Automated safety check: PassNone
Literature PDF OCR Library BuilderLigphiDonk/Oh-my--paper739—~1.1kAutomated safety check: PassMIT
Ref Downloaderltczding-gif/ref-downloader139—~5.9kAutomated safety check: PassMIT
Papers Skillsickn33/agentic-awesome-skills47k1 repos~2.1kAutomated safety check: PassMIT
Citation ManagementK-Dense-AI/claude-scientific-writer2.4k2 repos~3.9kAutomated safety check: NotesMIT
Paper2codePrathamLearnsToCode/paper2code1.5k—~1.3kAutomated safety check: PassMIT

Similar skills

  • Literature PDF OCR Library Builder

    LigphiDonk/Oh-my--paper

    Searches and downloads legally accessible academic PDFs, OCRs them to Markdown, and organizes the results into a traceable, AI-readable literature library.

    739 GitHub stars~1.1k tokensUpdated 5 mo ago
    Research & ScienceAuto-check passed
  • Ref Downloader

    ltczding-gif/ref-downloader

    A skill your agent uses when the user asks to batch-download academic PDFs with ref-downloader — either ALL references of one paper (Mode A: DOI or PDF input), OR a custom batch of papers (Mode B…

    139 GitHub stars~5.9k tokensUpdated 4 mo ago
    Research & ScienceAuto-check passed
  • Papers Skill

    sickn33/agentic-awesome-skills

    Skill for academic research workflows: search Semantic Scholar (200M+ papers), inspect citations, download arXiv PDFs, and extract PDF text.

    47k GitHub starsUsed in 1 repo~2.1k tokens
    Research & ScienceAuto-check passed
  • Citation Management

    K-Dense-AI/claude-scientific-writer

    Finds papers in OpenAlex, PubMed and Google Scholar, turns DOIs, PMIDs and arXiv IDs into clean BibTeX, and validates citations for a manuscript or thesis.

    2.4k GitHub starsUsed in 2 repos~3.9k tokens
    Research & ScienceAuto-check: notes
  • Paper2code

    PrathamLearnsToCode/paper2code

    Converts an arxiv paper into a minimal, citation-anchored Python implementation.

    1.5k GitHub stars~1.3k tokensUpdated 6 mo ago
    Research & ScienceAuto-check passed
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub stars~1.5k tokensUpdated yesterday
    Research & ScienceAuto-check passed

More from WILLOSCAR/research-units-pipeline-skills

All 107 skills in this repo
  • Appendix Table Writer

    WILLOSCAR/research-units-pipeline-skills

    Curate reader-facing survey tables for the Appendix (clean layout + high information density), using only in-scope evidence and existing citation keys.

    513 GitHub stars~1.8k tokensUpdated 5 days ago
    Auto-check passed
  • Artifact Contract Auditor

    WILLOSCAR/research-units-pipeline-skills

    Audit one research Workspace for declared Unit outputs and Pipeline target Artifacts, writing output/CONTRACTREPORT.md; use for mid-Run coverage snapshots or final delivery completeness, not deep…

    513 GitHub stars~918 tokensUpdated 5 days ago
    Auto-check passed
  • Arxiv Search

    WILLOSCAR/research-units-pipeline-skills

    Retrieve arXiv paper metadata with keyword queries or import an offline arXiv export, and save results as JSONL (papers/papersraw.jsonl).

    513 GitHub stars~1.5k tokensUpdated 5 days ago
    Auto-check passed
  • Chapter Lead Writer

    WILLOSCAR/research-units-pipeline-skills

    Write H2 chapter lead blocks (sections/S<secidlead.md) that preview the chapter's comparison lens and connect its H3 subsections, without adding new facts.

    513 GitHub stars~1.5k tokensUpdated 5 days ago
    Auto-check passed
  • Chapter Skeleton

    WILLOSCAR/research-units-pipeline-skills

    Build a retrieval-informed chapter skeleton (outline/chapterskeleton.yml) from taxonomy/core scope before stable H3 decomposition.

    513 GitHub stars~475 tokensUpdated 5 days ago
    Auto-check passed
  • Dedupe Rank

    WILLOSCAR/research-units-pipeline-skills

    A skill your agent uses when a broad paper candidate pool needs deterministic deduplication and a stable core set.

    513 GitHub stars~438 tokensUpdated 5 days ago
    Auto-check passed

Works with

Questions about Agent Survey Corpus

What does Agent Survey Corpus do?

Download a small corpus of open-access arXiv survey/review PDFs about agentic systems and extract text for style learning. Agent Survey Corpus is an agent skill from WILLOSCAR/research-units-pipeline-skills. Download a small corpus of open-access arXiv survey/review PDFs about agentic systems and extract text for style learning.

When should I use Agent Survey Corpus?

Agent Survey Corpus fits situations like: tasks that involve Academic paper search.

How do I install Agent Survey Corpus in Claude Code?

Run `npx skills add WILLOSCAR/research-units-pipeline-skills --skill agent-survey-corpus -a claude-code`. Or copy the skill folder (.codex/skills/agent-survey-corpus in WILLOSCAR/research-units-pipeline-skills) into .claude/skills/agent-survey-corpus in your project. Claude Code loads it when a task matches its description.

How do I install Agent Survey Corpus in Codex?

Run `npx skills add WILLOSCAR/research-units-pipeline-skills --skill agent-survey-corpus -a codex`. Or copy the skill folder (.codex/skills/agent-survey-corpus in WILLOSCAR/research-units-pipeline-skills) into .agents/skills/agent-survey-corpus in your project. Codex loads it when a task matches its description.

Can I use Agent Survey Corpus in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add WILLOSCAR/research-units-pipeline-skills --skill agent-survey-corpus -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/agent-survey-corpus, .gemini/skills/agent-survey-corpus, .github/skills/agent-survey-corpus and .opencode/skills/agent-survey-corpus in your project.

What does Agent Survey Corpus need to run?

Going by SKILL.md and its folder, Agent Survey Corpus needs Python for the scripts in its folder and the command-line tools its instructions call (uv). Our summary lists: Python 3.

Does Agent Survey Corpus access the network?

SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Agent Survey Corpus safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Agent Survey Corpus use?

No licence was found for Agent Survey Corpus or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Agent Survey Corpus use?

About 628 tokens (SKILL.md is roughly 2.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Agent Survey Corpus?

Skills that share tags, products or a category with Agent Survey Corpus: Literature PDF OCR Library Builder (LigphiDonk/Oh-my--paper, 739 stars), Ref Downloader (ltczding-gif/ref-downloader, 139 stars), Papers Skill (sickn33/agentic-awesome-skills, 47k stars) and Citation Management (K-Dense-AI/claude-scientific-writer, 2.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Agent Survey Corpus?

WILLOSCAR (a GitHub user) maintains it in WILLOSCAR/research-units-pipeline-skills, which has 513 GitHub stars. The repository holds 107 skills in this directory. The repository was last updated on October 5, 2026.

Source: WILLOSCAR/research-units-pipeline-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.