Agent skill

arXiv Paper Search

by taracodlabs in taracodlabs/aiden

Searches arXiv by keyword, category, author or paper ID through its public API and downloads PDFs, with no API key needed.

Apache-2.0Auto-check passedResearch & Science

Install arXiv Paper Search

skills CLI
$ npx skills add taracodlabs/aiden --skill arxiv -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install taracodlabs/aiden arxiv --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/taracodlabs/aiden.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/arxiv .claude/skills/arxiv && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
arxiv
GitHub stars
851
Token cost
~1k tokens
SKILL.md length
237 words
Files
2
Skills in repo
63
Repo updated
First seen
Licence
Apache-2.0

At a glance

Searches arXiv by keyword, category, author or paper ID through its public API and downloads PDFs, with no API key needed.

  • Works in 6 steps: Search papers by keyword → Search by category (cs.AI, cs.LG,… → Fetch a specific paper by arXiv ID → …
  • Finding recent machine learning papers on a topic
  • SKILL.md covers When to Use, How to Use, Examples and Cautions
  • Reaches export.arxiv.org and arxiv.org

What it does

This skill gives the agent working recipes for the free arXiv REST API, which answers in Atom XML. The examples are written mostly in PowerShell with Invoke-RestMethod, plus one Python version built on urllib and ElementTree, and they cover searching by keyword, by category such as cs.LG or cs.AI, by author, or by a known arXiv ID.

A download recipe saves a paper's PDF from arxiv.org/pdf to a local folder, and its path example uses Windows syntax. Worked examples show how to map requests such as finding the five most recent papers on retrieval-augmented generation, reading one paper's abstract, or fetching the Attention Is All You Need paper by its ID. The folder also holds a skill.json file.

When your agent uses it

  • Finding recent machine learning papers on a topic
  • Reading the abstract of a paper when you know its arXiv ID
  • Downloading a paper's PDF to a local folder
  • Listing the latest papers by a particular author

Example prompts

  • “Find the five newest cs.CL papers on retrieval-augmented generation.”
  • “Get me the abstract of arXiv paper 2305.17333.”
  • “Download the Attention Is All You Need paper as a PDF.”
  • “List recent arXiv papers by Andrej Karpathy.”

Requirements

  • PowerShell or Python 3
  • Network access to export.arxiv.org and arxiv.org

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Search papers by keyword
  2. Search by category (cs.AI, cs.LG, stat.ML, etc.)
  3. Fetch a specific paper by arXiv ID
  4. Download a paper PDF
  5. Search papers with Python
  6. Find papers by author

What it can do on your machine

Read from SKILL.md and the folder at commit 3704204. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are powershell and python).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • export.arxiv.org
    • arxiv.org
    • w3.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

arXiv Paper Search loads about 1k tokens when it runs. Until then it costs about 15 tokens; SKILL.md has 237 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~15
When it runs · the whole SKILL.md, loaded when a task matches
~1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from taracodlabs/aiden at commit 3704204, republished under its Apache-2.0 licence (© taracodlabs). 237 words, ~1,044 tokens.

Download SKILL.mdSave it as .claude/skills/arxiv/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
arxiv
description
Search and download arXiv papers (no API key needed)
category
research
version
1.0.0
origin
aiden
license
Apache-2.0
tags
arxiv, research, papers, academic, ai, ml, science, pdf, preprint, citations

arXiv Paper Search and Download

Search and retrieve academic papers from arXiv using the free public REST API. No API key or authentication required.

When to Use

  • User wants to find recent ML/AI/CS papers on a topic
  • User wants to read the abstract of a specific paper
  • User wants to download a paper PDF
  • User wants to find papers by a specific author
  • User wants to stay up-to-date on a research area

How to Use

1. Search papers by keyword

The arXiv API uses Atom XML — parse with PowerShell or Python.

powershell
$query   = [Uri]::EscapeDataString("attention mechanism transformer")
$url     = "http://export.arxiv.org/api/query?search_query=all:$query&start=0&max_results=5&sortBy=lastUpdatedDate&sortOrder=descending"
$resp    = Invoke-RestMethod -Uri $url
$resp.feed.entry | ForEach-Object {
  [PSCustomObject]@{
    Title   = $_.title
    Authors = ($_.author | ForEach-Object { $_.name }) -join ", "
    Date    = $_.published
    Id      = $_.id
  }
} | Format-Table -AutoSize
2. Search by category (cs.AI, cs.LG, stat.ML, etc.)
powershell
$cat   = "cs.LG"
$query = [Uri]::EscapeDataString("large language models")
$url   = "http://export.arxiv.org/api/query?search_query=cat:$cat+AND+all:$query&max_results=10&sortBy=submittedDate&sortOrder=descending"
$resp  = Invoke-RestMethod -Uri $url
$resp.feed.entry | Select-Object title, published
3. Fetch a specific paper by arXiv ID
powershell
$arxivId = "2305.17333"   # from URL: arxiv.org/abs/2305.17333
$url     = "http://export.arxiv.org/api/query?id_list=$arxivId"
$resp    = Invoke-RestMethod -Uri $url
$entry   = $resp.feed.entry
Write-Host "Title:   " $entry.title
Write-Host "Authors: " (($entry.author | ForEach-Object { $_.name }) -join ", ")
Write-Host "Abstract:" $entry.summary
4. Download a paper PDF
powershell
$arxivId = "2305.17333"
$pdfUrl  = "https://arxiv.org/pdf/$arxivId.pdf"
$outPath = "C:\Users\<you>\Downloads\paper_$arxivId.pdf"
Invoke-WebRequest -Uri $pdfUrl -OutFile $outPath
Write-Host "Downloaded to $outPath"
5. Search papers with Python
python
import urllib.request, urllib.parse, xml.etree.ElementTree as ET

def search_arxiv(query, max_results=5, category="cs.LG"):
  params = urllib.parse.urlencode({
    "search_query": f"cat:{category} AND all:{query}",
    "max_results": max_results,
    "sortBy": "submittedDate",
    "sortOrder": "descending"
  })
  url  = f"http://export.arxiv.org/api/query?{params}"
  resp = urllib.request.urlopen(url).read()
  root = ET.fromstring(resp)
  ns   = {"a": "http://www.w3.org/2005/Atom"}
  for entry in root.findall("a:entry", ns):
    print(entry.find("a:title", ns).text.strip())
    print(entry.find("a:id", ns).text.strip())
    print()

search_arxiv("chain of thought reasoning", max_results=5)
6. Find papers by author
powershell
$author  = [Uri]::EscapeDataString("Andrej Karpathy")
$url     = "http://export.arxiv.org/api/query?search_query=au:$author&max_results=10&sortBy=submittedDate&sortOrder=descending"
$resp    = Invoke-RestMethod -Uri $url
$resp.feed.entry | Select-Object title, published | Format-Table

Examples

"Find the 5 most recent papers on retrieval-augmented generation" → Use step 2 with query retrieval augmented generation and category cs.CL or cs.AI.

"Get me the abstract of paper 2305.17333" → Use step 3 with the arXiv ID.

"Download the Attention Is All You Need paper" → arXiv ID is 1706.03762. Use step 4 to download the PDF.

Cautions

  • arXiv API rate limit is 3 requests per second — add a 1-second delay between calls for bulk operations
  • Papers on arXiv are preprints and have not necessarily been peer-reviewed
  • arXiv IDs changed format in 2007 — old IDs use category/YYMMNNN format; new ones use YYMM.NNNNN
  • PDF download uses the standard arxiv.org/pdf/<id>.pdf URL — some papers have versioned PDFs at arxiv.org/pdf/<id>v1.pdf

© taracodlabs, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/arxiv of taracodlabs/aiden.

  • SKILL.md
  • skill.json

Open the folder on GitHubat commit 3704204

Compare with similar skills

arXiv Paper Search next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

arXiv Paper Search compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
arXiv Paper Search this skilltaracodlabs/aiden851—~1kAutomated safety check: PassApache-2.0
Citation ManagementK-Dense-AI/claude-scientific-writer2.4k2 repos~3.9kAutomated safety check: NotesMIT
Paper2codePrathamLearnsToCode/paper2code1.5k—~1.3kAutomated safety check: PassMIT
Paper Research on arXivXiaomiMiMo/MiMo-Code14k—~1.5kAutomated safety check: PassMIT
Arxiv MCP Serverblazickjp/arxiv-mcp-server3.2k—~353Automated safety check: PassApache-2.0
Hugging Face Paper Publisherhuggingface/skills11k4 repos~4.2kAutomated safety check: PassApache-2.0

Similar skills

  • Citation Management

    K-Dense-AI/claude-scientific-writer

    Finds papers in OpenAlex, PubMed and Google Scholar, turns DOIs, PMIDs and arXiv IDs into clean BibTeX, and validates citations for a manuscript or thesis.

    2.4k GitHub starsUsed in 2 repos~3.9k tokens
    Research & ScienceAuto-check: notes
  • Paper2code

    PrathamLearnsToCode/paper2code

    Converts an arxiv paper into a minimal, citation-anchored Python implementation.

    1.5k GitHub stars~1.3k tokensUpdated 6 mo ago
    Research & ScienceAuto-check passed
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub stars~1.5k tokensUpdated yesterday
    Research & ScienceAuto-check passed
  • Arxiv MCP Server

    blazickjp/arxiv-mcp-server

    A skill your agent uses when finding, comparing, reading, or monitoring arXiv papers, including requests for abstracts, citation graphs, original LaTeX, section-level technical details, or…

    3.2k GitHub stars~353 tokensUpdated yesterday
    Research & ScienceAuto-check passed
  • Official

    Indexes research papers on the Hugging Face Hub from arXiv, links them to models and datasets, claims authorship and generates markdown research articles from templates.

    11k GitHub starsUsed in 4 repos~4.2k tokens
    Research & ScienceAuto-check passed
  • PaperSeek Literature Search

    MingfengHong/paperseek

    Routes literature searches through the PaperSeek launcher, picks suitable scholarly sources, parses JSON output and keeps API keys out of the chat.

    200 GitHub stars~1.6k tokensUpdated 4 days ago
    Research & ScienceAuto-check passed

More from taracodlabs/aiden

All 63 skills in this repo
  • Google Flights Search

    taracodlabs/aiden

    Searches Google Flights for prices, schedules and availability through browser automation with URL-based queries, and stops short of booking.

    851 GitHub stars~1.5k tokensUpdated 26 days ago
    Auto-check passed
  • OpenAI Codex CLI Bridge

    taracodlabs/aiden

    Delegates code generation, editing and explanation tasks to the OpenAI Codex CLI, with commands for interactive, auto-edit, question-only and model-specific runs.

    851 GitHub starsUsed in 1 repo~826 tokens
    Auto-check passed
  • Google Hotels Search

    taracodlabs/aiden

    Searches Google Hotels through the agent-browser tool for prices, ratings, amenities and availability, building a search URL from the location and dates and reporting a results table.

    851 GitHub stars~1.5k tokensUpdated 26 days ago
    Auto-check passed
  • Generates dark-themed architecture, component, data-flow and network diagrams as self-contained HTML and SVG files that open in any browser.

    851 GitHub stars~1.3k tokensUpdated 26 days ago
    Auto-check passed
  • Aggregates holdings across Zerodha, Upstox and Angel One and normalizes order parameters into one format, with confirmation required before any routing.

    851 GitHub stars~1.1k tokensUpdated 26 days ago
    Auto-check passed
  • ASCII Art Banners

    taracodlabs/aiden

    Produces ASCII text banners, cowsay-style speech bubbles and bordered boxes for terminal scripts, READMEs and CLI splash screens with pyfiglet, cowsay and boxes.

    851 GitHub stars~985 tokensUpdated 26 days ago
    Auto-check passed

Questions about arXiv Paper Search

What does arXiv Paper Search do?

Searches arXiv by keyword, category, author or paper ID through its public API and downloads PDFs, with no API key needed. This skill gives the agent working recipes for the free arXiv REST API, which answers in Atom XML.AI, by author, or by a known arXiv ID.

When should I use arXiv Paper Search?

arXiv Paper Search fits situations like: finding recent machine learning papers on a topic; reading the abstract of a paper when you know its arXiv ID; downloading a paper's PDF to a local folder; listing the latest papers by a particular author.

How do I install arXiv Paper Search in Claude Code?

Run `npx skills add taracodlabs/aiden --skill arxiv -a claude-code`. Or copy the skill folder (skills/arxiv in taracodlabs/aiden) into .claude/skills/arxiv in your project. Claude Code loads it when a task matches its description.

How do I install arXiv Paper Search in Codex?

Run `npx skills add taracodlabs/aiden --skill arxiv -a codex`. Or copy the skill folder (skills/arxiv in taracodlabs/aiden) into .agents/skills/arxiv in your project. Codex loads it when a task matches its description.

Can I use arXiv Paper Search in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add taracodlabs/aiden --skill arxiv -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/arxiv, .gemini/skills/arxiv, .github/skills/arxiv and .opencode/skills/arxiv in your project.

What does arXiv Paper Search need to run?

SKILL.md names no scripts, command-line tools or credentials: arXiv Paper Search is instructions for the agent only. Our summary lists: PowerShell or Python 3; Network access to export.arxiv.org and arxiv.org.

Does arXiv Paper Search access the network?

SKILL.md names 3 domains. In commands or code: export.arxiv.org, arxiv.org and w3.org; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is arXiv Paper Search safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does arXiv Paper Search use?

arXiv Paper Search is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does arXiv Paper Search use?

About 1k tokens (SKILL.md is roughly 4.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to arXiv Paper Search?

Skills that share tags, products or a category with arXiv Paper Search: Citation Management (K-Dense-AI/claude-scientific-writer, 2.4k stars), Paper2code (PrathamLearnsToCode/paper2code, 1.5k stars), Paper Research on arXiv (XiaomiMiMo/MiMo-Code, 14k stars) and Arxiv MCP Server (blazickjp/arxiv-mcp-server, 3.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains arXiv Paper Search?

taracodlabs (a GitHub organization) maintains it in taracodlabs/aiden, which has 851 GitHub stars. The repository holds 63 skills in this directory. The repository was last updated on September 13, 2026.

Source: taracodlabs/aiden on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.