Retrieve arXiv paper metadata with keyword queries or import an offline arXiv export, and save results as JSONL (papers/papersraw.jsonl).

No licenceAuto-check passedResearch & Science

Install Arxiv Search

skills CLI
$ npx skills add WILLOSCAR/research-units-pipeline-skills --skill arxiv-search -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install WILLOSCAR/research-units-pipeline-skills arxiv-search --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/WILLOSCAR/research-units-pipeline-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.codex/skills/arxiv-search .claude/skills/arxiv-search && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
arxiv-search
GitHub stars
513
Token cost
~1.5k tokens
SKILL.md length
661 words
Files
5 (incl. scripts, references, assets)
Skills in repo
107
Repo updated
First seen
Licence
None found

At a glance

Retrieve arXiv paper metadata with keyword queries or import an offline arXiv export, and save results as JSONL (papers/papersraw.jsonl).

  • Works in 5 steps: Read queries.md and expand into concrete… → Retrieve results (online) or import an… → Normalize every record to include at least → …
  • Tasks that involve Academic paper search
  • SKILL.md covers Triggers & routing, Load Order, Script Boundary and Contract-driven behavior, plus 8 more sections
  • Runs Python scripts from its folder; calls uv

What it does

Arxiv Search is an agent skill from WILLOSCAR/research-units-pipeline-skills. Retrieve arXiv paper metadata with keyword queries or import an offline arXiv export, and save results as JSONL (papers/papersraw.jsonl).

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 8 other files, including scripts, reference files and assets (for example `assets/domain_packs/embodied_ai.json`, `assets/domain_packs/llm_agents.json` and `references/domain_pack_overview.md`).

It sits in Research & Science, covering Academic paper search. It works with arXiv. The repository describes itself as: Research pipelines as semantic execution units: each skill declares inputs/outputs, acceptance criteria, and guardrails. Evidence-first methodology prevents hollow writing…

When your agent uses it

  • Tasks that involve Academic paper search

Example prompts

  • “/arxiv-search”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Read queries.md and expand into concrete query strings.
  2. Retrieve results (online) or import an export (offline).
  3. Normalize every record to include at least
  4. Keep the set broad at this stage; dedupe/ranking comes next.
  5. Apply time window and max_results if specified.

What it can do on your machine

Read from SKILL.md and the folder at commit c92912a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Arxiv Search loads about 1.5k tokens when it runs, and up to ~1.8k if it reads all its reference files. Until then it costs about 38 tokens; SKILL.md has 661 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~38
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 661 words (~1,501 tokens).

“Collect an initial paper set with enough metadata to support downstream ranking, taxonomy building, and citation generation.”

— opening of SKILL.md by WILLOSCAR
name
arxiv-search

Read the full SKILL.md on GitHub

Files

SKILL.md and 4 other files (scripts, references, assets) in .codex/skills/arxiv-search of WILLOSCAR/research-units-pipeline-skills.

  • SKILL.md
  • assets/domain_packs/embodied_ai.json
  • assets/domain_packs/llm_agents.json
  • references/domain_pack_overview.md
  • scripts/run.py

Open the folder on GitHubat commit c92912a

Compare with similar skills

Arxiv Search next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Arxiv Search compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Arxiv Search this skillWILLOSCAR/research-units-pipeline-skills513—~1.5kAutomated safety check: PassNone
Read arXiv Paperkarpathy/nanochat59k1 repos~494Automated safety check: PassMIT
Literature Reviewneflibata-feng/MyArxiv-Agent12620 repos~5.9kAutomated safety check: NotesMIT
Openalex Databaseneflibata-feng/MyArxiv-Agent12612 repos~3kAutomated safety check: PassCustom licence
Citation ManagementK-Dense-AI/claude-scientific-writer2.4k2 repos~3.9kAutomated safety check: NotesMIT
Citation Managementneflibata-feng/MyArxiv-Agent12619 repos~8.1kAutomated safety check: NotesMIT

Similar skills

  • Read arXiv Paper

    karpathy/nanochat

    Fetches the TeX source of an arXiv paper from its URL, reads it and writes a markdown summary tied to the nanochat project.

    59k GitHub starsUsed in 1 repo~494 tokens
    Research & ScienceAuto-check passed
  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 20 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Openalex Database

    neflibata-feng/MyArxiv-Agent

    Query and analyze scholarly literature using the OpenAlex database.

    126 GitHub starsUsed in 12 repos~3k tokens
    Research & ScienceAuto-check passed
  • Citation Management

    K-Dense-AI/claude-scientific-writer

    Finds papers in OpenAlex, PubMed and Google Scholar, turns DOIs, PMIDs and arXiv IDs into clean BibTeX, and validates citations for a manuscript or thesis.

    2.4k GitHub starsUsed in 2 repos~3.9k tokens
    Research & ScienceAuto-check: notes
  • Citation Management

    neflibata-feng/MyArxiv-Agent

    Comprehensive citation management for academic research. An agent skill from neflibata-feng/MyArxiv-Agent.

    126 GitHub starsUsed in 19 repos~8.1k tokens
    Research & ScienceAuto-check: notes
  • Searches arXiv across many papers on one topic, extracts each paper's methodology and findings in parallel, and synthesizes a cited literature review.

    84k GitHub starsUsed in 2 repos~4.3k tokens
    Research & ScienceAuto-check passed

More from WILLOSCAR/research-units-pipeline-skills

All 107 skills in this repo
  • Appendix Table Writer

    WILLOSCAR/research-units-pipeline-skills

    Curate reader-facing survey tables for the Appendix (clean layout + high information density), using only in-scope evidence and existing citation keys.

    513 GitHub stars~1.8k tokensUpdated 5 days ago
    Auto-check passed
  • Artifact Contract Auditor

    WILLOSCAR/research-units-pipeline-skills

    Audit one research Workspace for declared Unit outputs and Pipeline target Artifacts, writing output/CONTRACTREPORT.md; use for mid-Run coverage snapshots or final delivery completeness, not deep…

    513 GitHub stars~918 tokensUpdated 5 days ago
    Auto-check passed
  • Chapter Lead Writer

    WILLOSCAR/research-units-pipeline-skills

    Write H2 chapter lead blocks (sections/S<secidlead.md) that preview the chapter's comparison lens and connect its H3 subsections, without adding new facts.

    513 GitHub stars~1.5k tokensUpdated 5 days ago
    Auto-check passed
  • Chapter Skeleton

    WILLOSCAR/research-units-pipeline-skills

    Build a retrieval-informed chapter skeleton (outline/chapterskeleton.yml) from taxonomy/core scope before stable H3 decomposition.

    513 GitHub stars~475 tokensUpdated 5 days ago
    Auto-check passed
  • Dedupe Rank

    WILLOSCAR/research-units-pipeline-skills

    A skill your agent uses when a broad paper candidate pool needs deterministic deduplication and a stable core set.

    513 GitHub stars~438 tokensUpdated 5 days ago
    Auto-check passed
  • Evaluation Anchor Checker

    WILLOSCAR/research-units-pipeline-skills

    Audit and rewrite evaluation/numeric claims to ensure they carry minimal protocol context (task + metric + constraint) and avoid underspecified model naming.

    513 GitHub stars~1.4k tokensUpdated 5 days ago
    Auto-check passed

Works with

Questions about Arxiv Search

What does Arxiv Search do?

Retrieve arXiv paper metadata with keyword queries or import an offline arXiv export, and save results as JSONL (papers/papersraw.jsonl). Arxiv Search is an agent skill from WILLOSCAR/research-units-pipeline-skills.jsonl).

When should I use Arxiv Search?

Arxiv Search fits situations like: tasks that involve Academic paper search.

How do I install Arxiv Search in Claude Code?

Run `npx skills add WILLOSCAR/research-units-pipeline-skills --skill arxiv-search -a claude-code`. Or copy the skill folder (.codex/skills/arxiv-search in WILLOSCAR/research-units-pipeline-skills) into .claude/skills/arxiv-search in your project. Claude Code loads it when a task matches its description.

How do I install Arxiv Search in Codex?

Run `npx skills add WILLOSCAR/research-units-pipeline-skills --skill arxiv-search -a codex`. Or copy the skill folder (.codex/skills/arxiv-search in WILLOSCAR/research-units-pipeline-skills) into .agents/skills/arxiv-search in your project. Codex loads it when a task matches its description.

Can I use Arxiv Search in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add WILLOSCAR/research-units-pipeline-skills --skill arxiv-search -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/arxiv-search, .gemini/skills/arxiv-search, .github/skills/arxiv-search and .opencode/skills/arxiv-search in your project.

What does Arxiv Search need to run?

Going by SKILL.md and its folder, Arxiv Search needs Python for the scripts in its folder and the command-line tools its instructions call (uv). Our summary lists: Python 3.

Does Arxiv Search access the network?

SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Arxiv Search safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Arxiv Search use?

No licence was found for Arxiv Search or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Arxiv Search use?

About 1.5k tokens (SKILL.md is roughly 6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 258 tokens, read only when the agent opens those files.

What are the alternatives to Arxiv Search?

Skills that share tags, products or a category with Arxiv Search: Read arXiv Paper (karpathy/nanochat, 59k stars), Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Openalex Database (neflibata-feng/MyArxiv-Agent, 126 stars) and Citation Management (K-Dense-AI/claude-scientific-writer, 2.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Arxiv Search?

WILLOSCAR (a GitHub user) maintains it in WILLOSCAR/research-units-pipeline-skills, which has 513 GitHub stars. The repository holds 107 skills in this directory. The repository was last updated on October 5, 2026.

Source: WILLOSCAR/research-units-pipeline-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.