Agent skill

Arxiv Database

by aipoch in aipoch/medical-research-skills

Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading.

MITAuto-check passedResearch & Science

Install Arxiv Database

skills CLI
$ npx skills add aipoch/medical-research-skills --skill arxiv-database -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install aipoch/medical-research-skills arxiv-database --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/aipoch/medical-research-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/'scientific-skills/Evidence Insight/arxiv-database' .claude/skills/arxiv-database && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
arxiv-database
GitHub stars
2k
Token cost
~1.5k tokens
SKILL.md length
664 words
Files
4 (incl. scripts, references)
Skills in repo
567
Repo updated
First seen
Licence
MIT

At a glance

Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading.

  • Works in 5 steps: When to Use → Key Features → Dependencies → …
  • You need to find papers by keyword/author/category
  • SKILL.md covers When to Use, Key Features, Dependencies and Example Usage, plus 6 more sections
  • Runs Python scripts from its folder; calls python and pip

What it does

Arxiv Database is an agent skill from aipoch/medical-research-skills. Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading.

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts and reference files (for example `arxiv-database_audit_result_v1.json`, `references/api_docs.md` and `scripts/arxiv_search.py`).

It sits in Research & Science, covering Academic paper search. It works with arXiv. The repository describes itself as: Hundreds of agent skills for medical research, including protocol design, data analysis, evidence insights, and academic writing. The licence is MIT.

When your agent uses it

  • You need to find papers by keyword/author/category
  • Fetch metadata (abstract
  • Download PDFs for offline reading

Example prompts

  • “/arxiv-database”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. When to Use
  2. Key Features
  3. Dependencies
  4. Example Usage
  5. Implementation Details

What it can do on your machine

Read from SKILL.md and the folder at commit 686e09d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python
    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Arxiv Database loads about 1.5k tokens when it runs, and up to ~1.7k if it reads all its reference files. Until then it costs about 53 tokens; SKILL.md has 664 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~53
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from aipoch/medical-research-skills at commit 686e09d, republished under its MIT licence (© aipoch). 664 words, ~1,531 tokens.

Download SKILL.mdSave it as .claude/skills/arxiv-database/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
arxiv-database
description
Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading.
license
MIT
author
AIPOCH

Source: https://github.com/aipoch/medical-research-skills

ArXiv Database Skill

When to Use

  • Use this skill when you need search and retrieve scientific preprints from arxiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, doi, pdf url), or download pdfs for offline reading in a reproducible workflow.
  • Use this skill when a evidence insight task needs a packaged method instead of ad-hoc freeform output.
  • Use this skill when the user expects a concrete deliverable, validation step, or file-based result.
  • Use this skill when scripts/arxiv_search.py is the most direct path to complete the request.
  • Use this skill when you need the arxiv-database package behavior rather than a generic answer.

Key Features

  • Scope-focused workflow aligned to: Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading.
  • Packaged executable path(s): scripts/arxiv_search.py.
  • Reference material available in references/ for task-specific guidance.
  • Structured execution path designed to keep outputs consistent and reviewable.

Dependencies

  • Python: 3.10+. Repository baseline for current packaged skills.
  • Third-party packages: not explicitly version-pinned in this skill package. Add pinned versions if this skill needs stricter environment control.

Example Usage

bash
cd "20260316/scientific-skills/Evidence Insight/arxiv-database"
python -m py_compile scripts/arxiv_search.py
python scripts/arxiv_search.py --help

Example run plan:

  1. Confirm the user input, output path, and any required config values.
  2. Edit the in-file CONFIG block or documented parameters if the script uses fixed settings.
  3. Run python scripts/arxiv_search.py with the validated inputs.
  4. Review the generated output and return the final artifact with any assumptions called out.

Implementation Details

  • Execution model: validate the request, choose the packaged workflow, and produce a bounded deliverable.
  • Input controls: confirm the source files, scope limits, output format, and acceptance criteria before running any script.
  • Primary implementation surface: scripts/arxiv_search.py.
  • Reference guidance: references/ contains supporting rules, prompts, or checklists.
  • Parameters to clarify first: input path, output path, scope filters, thresholds, and any domain-specific constraints.
  • Output discipline: keep results reproducible, identify assumptions explicitly, and avoid undocumented side effects.

1. When to Use

  • You need to quickly find arXiv preprints by keyword, phrase, author, or category (e.g., cs.AI, cs.CL).
  • You want to collect paper metadata (title, authors, publication date, abstract/summary, PDF link) for review or indexing.
  • You need the latest submissions in a topic area (sorted by submission date or last updated date).
  • You want to download one or more PDFs from search results for offline reading or batch processing.
  • You have a known arXiv identifier and want to retrieve the corresponding paper directly.
Show full SKILL.md (257 more words)Show less

2. Key Features

  • arXiv query-based search (supports category filters, author filters, phrases, and ID lookups).
  • Configurable result limits (--max-results).
  • Sort control (--sort-by: Relevance, LastUpdatedDate, SubmittedDate).
  • Metadata output per result (title, authors, published date, abstract/summary, PDF URL; DOI when available via arXiv metadata).
  • Optional PDF download for returned results (--download) with configurable output directory (--dir).

3. Dependencies

  • Python 3.8+
  • arxiv (Python package) — version depends on your environment; install a recent release (e.g., arxiv>=1.4.0)

4. Example Usage

Install dependencies
bash
pip install "arxiv>=1.4.0"
Run searches and downloads

Search for papers in cs.AI about reinforcement learning (top 5 results):

bash
python scripts/arxiv_search.py --query "cat:cs.AI AND reinforcement learning" --max-results 5

Search for “Large Language Models” in cs.CL:

bash
python scripts/arxiv_search.py --query "cat:cs.CL AND \"Large Language Models\""

Get the latest 5 papers on “quantum computing” (sorted by submission date):

bash
python scripts/arxiv_search.py --query "quantum computing" --sort-by SubmittedDate --max-results 5

Download a specific paper by arXiv ID:

bash
python scripts/arxiv_search.py --query "id:2101.12345" --download

Download results into a specific directory:

bash
python scripts/arxiv_search.py --query "cat:cs.LG AND diffusion" --max-results 3 --download --dir ./papers

5. Implementation Details

  • Entry point: scripts/arxiv_search.py wraps the arxiv Python API to execute queries against the arXiv search endpoint.
  • Query syntax: The --query string is passed to arXiv search and can include:
    • Category filters (e.g., cat:cs.AI)
    • Author filters (e.g., au:Smith)
    • Exact phrases using quotes (e.g., "Large Language Models")
    • ID lookup (e.g., id:2101.12345)
    • Boolean operators such as AND
  • Result limiting: --max-results controls how many entries are returned (default: 10).
  • Sorting: --sort-by selects the ordering of results:
    • Relevance (default)
    • LastUpdatedDate
    • SubmittedDate
  • Downloads: When --download is set, the script downloads the PDF for each returned result using the provided PDF URL and saves it to --dir (default: current working directory).
  • Metadata fields: Each result includes core arXiv metadata (title, authors, published date, summary/abstract, PDF URL). DOI is included when present in arXiv’s metadata for that record.

© aipoch, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts, references) in scientific-skills/Evidence Insight/arxiv-database of aipoch/medical-research-skills.

  • SKILL.md
  • arxiv-database_audit_result_v1.json
  • references/api_docs.md
  • scripts/arxiv_search.py

Open the folder on GitHubat commit 686e09d

Compare with similar skills

Arxiv Database next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Arxiv Database compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Arxiv Database this skillaipoch/medical-research-skills2k—~1.5kAutomated safety check: PassMIT
Read arXiv Paperkarpathy/nanochat58k2 repos~494Automated safety check: PassMIT
Literature Reviewneflibata-feng/MyArxiv-Agent12621 repos~5.9kAutomated safety check: NotesMIT
Citation ManagementK-Dense-AI/claude-scientific-writer2.4k3 repos~3.9kAutomated safety check: NotesMIT
Openalex Databaseneflibata-feng/MyArxiv-Agent12613 repos~3kAutomated safety check: PassCustom licence
Citation Managementneflibata-feng/MyArxiv-Agent12620 repos~8.1kAutomated safety check: NotesMIT

Similar skills

  • Read arXiv Paper

    karpathy/nanochat

    Fetches the TeX source of an arXiv paper from its URL, reads it and writes a markdown summary tied to the nanochat project.

    58k GitHub starsUsed in 2 repos~494 tokens
    Research & ScienceAuto-check passed
  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 21 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Citation Management

    K-Dense-AI/claude-scientific-writer

    Finds papers in OpenAlex, PubMed and Google Scholar, turns DOIs, PMIDs and arXiv IDs into clean BibTeX, and validates citations for a manuscript or thesis.

    2.4k GitHub starsUsed in 3 repos~3.9k tokens
    Research & ScienceAuto-check: notes
  • Openalex Database

    neflibata-feng/MyArxiv-Agent

    Query and analyze scholarly literature using the OpenAlex database.

    126 GitHub starsUsed in 13 repos~3k tokens
    Research & ScienceAuto-check passed
  • Citation Management

    neflibata-feng/MyArxiv-Agent

    Comprehensive citation management for academic research. An agent skill from neflibata-feng/MyArxiv-Agent.

    126 GitHub starsUsed in 20 repos~8.1k tokens
    Research & ScienceAuto-check: notes
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub starsUsed in 1 repo~1.5k tokens
    Research & ScienceAuto-check passed

More from aipoch/medical-research-skills

All 567 skills in this repo
  • Academic Poster Generator

    aipoch/medical-research-skills

    Complete workflow for generating academic research posters from PDF literature; use when you need to extract paper content from PDFs and produce a LaTeX-based poster…

    2k GitHub stars~2.2k tokensUpdated 21 days ago
    Auto-check passed
  • Diagnostic Study Quality Assessment Quadas

    aipoch/medical-research-skills

    Analyzes clinical diagnostic accuracy studies for bias using the QUADAS-2 tool.

    2k GitHub stars~1.4k tokensUpdated 21 days ago
    Auto-check passed
  • Exploratory Data Analysis

    aipoch/medical-research-skills

    Perform comprehensive exploratory data analysis on scientific data files across 200+ file formats.

    2k GitHub stars~3.7k tokensUpdated 21 days ago
    Auto-check passed
  • Iso Certification

    aipoch/medical-research-skills

    A toolkit for preparing ISO 13485:2016 certification documentation for medical device QMS.

    2k GitHub stars~1.8k tokensUpdated 21 days ago
    Auto-check passed
  • Journal Skills

    aipoch/medical-research-skills

    Recommends target journals for manuscript submission by analyzing the paper topic/abstract and the journal distribution of similar PubMed literature; use when users ask for journal…

    2k GitHub stars~1.7k tokensUpdated 21 days ago
    Auto-check passed
  • Latex Posters

    aipoch/medical-research-skills

    Creates academic-poster writing packages for LaTeX using beamerposter, tikzposter, or baposter.

    2k GitHub stars~1.3k tokensUpdated 21 days ago
    Auto-check passed

Works with

Questions about Arxiv Database

What does Arxiv Database do?

Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading. Arxiv Database is an agent skill from aipoch/medical-research-skills. Search and retrieve scientific preprints from arXiv; use it when you need to find papers by keyword/author/category, fetch metadata (abstract, DOI, PDF URL), or download PDFs for offline reading.

When should I use Arxiv Database?

Arxiv Database fits situations like: you need to find papers by keyword/author/category; fetch metadata (abstract; download PDFs for offline reading.

How do I install Arxiv Database in Claude Code?

Run `npx skills add aipoch/medical-research-skills --skill arxiv-database -a claude-code`. Or copy the skill folder (scientific-skills/Evidence Insight/arxiv-database in aipoch/medical-research-skills) into .claude/skills/arxiv-database in your project. Claude Code loads it when a task matches its description.

How do I install Arxiv Database in Codex?

Run `npx skills add aipoch/medical-research-skills --skill arxiv-database -a codex`. Or copy the skill folder (scientific-skills/Evidence Insight/arxiv-database in aipoch/medical-research-skills) into .agents/skills/arxiv-database in your project. Codex loads it when a task matches its description.

Can I use Arxiv Database in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add aipoch/medical-research-skills --skill arxiv-database -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/arxiv-database, .gemini/skills/arxiv-database, .github/skills/arxiv-database and .opencode/skills/arxiv-database in your project.

What does Arxiv Database need to run?

Going by SKILL.md and its folder, Arxiv Database needs Python for the scripts in its folder and the command-line tools its instructions call (python and pip). Our summary lists: Python 3.

Does Arxiv Database access the network?

SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Arxiv Database safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Arxiv Database use?

Arxiv Database is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Arxiv Database use?

About 1.5k tokens (SKILL.md is roughly 6.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 207 tokens, read only when the agent opens those files.

What are the alternatives to Arxiv Database?

Skills that share tags, products or a category with Arxiv Database: Read arXiv Paper (karpathy/nanochat, 58k stars), Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Citation Management (K-Dense-AI/claude-scientific-writer, 2.4k stars) and Openalex Database (neflibata-feng/MyArxiv-Agent, 126 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Arxiv Database?

aipoch (a GitHub organization) maintains it in aipoch/medical-research-skills, which has 1,974 GitHub stars. The repository holds 567 skills in this directory. The repository was last updated on September 17, 2026.

Source: aipoch/medical-research-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.