Agent skill

Literature Search Arxiv

by google-deepmind in google-deepmind/science-skills

Search for scientific papers, preprints, and publications on arXiv.

Apache-2.0Auto-check passedResearch & Science

Install Literature Search Arxiv

skills CLI
$ npx skills add google-deepmind/science-skills --skill literature-search-arxiv -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install google-deepmind/science-skills literature-search-arxiv --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/google-deepmind/science-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/literature_search_arxiv .claude/skills/literature-search-arxiv && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
literature-search-arxiv
GitHub stars
3.2k
Used in
2 other repos
Token cost
~1.2k tokens
SKILL.md length
533 words
Files
6 (incl. scripts, references)
Skills in repo
40
Repo updated
First seen
Licence
Apache-2.0

At a glance

Search for scientific papers, preprints, and publications on arXiv.

  • Works in 2 steps: uv: Read the uv skill and follow its… → User Notification: If…
  • The user asks to find research papers
  • SKILL.md covers Prerequisites, Core Rules, Utility Scripts and Reference, plus 1 more section
  • Runs Python scripts from its folder; calls uv

What it does

Literature Search Arxiv is an agent skill from google-deepmind/science-skills. Search for scientific papers, preprints, and publications on arXiv. Extract metadata, abstracts, and download full-text PDFs or HTML versions of papers. Use when the user asks to find research papers, literature, or specific arXiv IDs.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts and reference files (for example `references/query_syntax.md`, `scripts/download_paper.py` and `scripts/download_paper_source.py`).

It sits in Research & Science, covering Academic paper search and Literature review. It works with arXiv. The repository describes itself as: GDM Science Skills to speed up agentic scientific workflows with better grounding and higher token efficiency. Integrate insights from AlphaGenome, AFDB, UniProt and 30+ other… The licence is Apache-2.0.

When your agent uses it

  • The user asks to find research papers
  • Specific arXiv IDs

Example prompts

  • “/literature-search-arxiv”

Requirements

  • Python 3

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. uv: Read the uv skill and follow its Setup instructions to ensure
  2. User Notification: If .licenses/literature_search_arxiv_LICENSE.txt does

What it can do on your machine

Read from SKILL.md and the folder at commit 8ab7672. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 3 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • info.arxiv.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Literature Search Arxiv loads about 1.2k tokens when it runs, and up to ~1.5k if it reads all its reference files. Until then it costs about 65 tokens; SKILL.md has 533 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~65
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from google-deepmind/science-skills at commit 8ab7672, republished under its Apache-2.0 licence (© google-deepmind). 533 words, ~1,157 tokens.

Download SKILL.mdSave it as .claude/skills/literature-search-arxiv/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
literature-search-arxiv
description
Search for scientific papers, preprints, and publications on arXiv. Extract metadata, abstracts, and download full-text PDFs or HTML versions of papers. Use when the user asks to find research papers, literature, or specific arXiv IDs.

arXiv Search and Retrieval

Prerequisites

  1. uv: Read the uv skill and follow its Setup instructions to ensure uv is installed and on PATH.
  2. User Notification: If .licenses/literature_search_arxiv_LICENSE.txt does not already exist in the workspace root directory then (1) prominently notify the user to check the terms at https://info.arxiv.org/help/api/index.html and to always check the license of the papers retrieved by the skill for any restrictions, then (2) create the file recording the notification text and timestamp.

Core Rules

  • Terms of Use: You MUST respect arXiv's Terms of Use.
    • Maximum 1 request every 3 seconds.
    • The provided utility scripts handle rate limiting automatically. Always use these scripts rather than writing your own curl/python requests.
  • If this skill is used, ensure this is mentioned in the output AND list the URLs of all papers that were used in producing the output.

Utility Scripts

1. Search and Extract Metadata

Search arXiv and return a clean JSON array of matching papers.

bash
uv run scripts/search_arxiv.py --query "au:einstein AND ti:relativity" \
  --max_results 5 2>/dev/null > /tmp/arxiv_search_results.json

Important: The tool outputs a large JSON result to stdout. Requesting 100+ results will produce a massive JSON that might exceed your context length. Limit --max_results (e.g., 5-10) or paginate carefully using --start. Always redirect output to a file and parse it separately, otherwise terminal output will be truncated.

Returned Metadata: JSON results include id, title, summary, published, authors, pdf_url, primary_category, doi, journal_ref, and comment. Note: the doi field only contains DOI information in case the paper has an external DOI and if only an arXiv-issued DOI exists, this is DOI is not returned.

Options:

  • --query: Search string. See references/query_syntax.md for advanced syntax.
  • --id_list: Comma-separated list of arXiv IDs to fetch directly (e.g., 1706.03762v5).
  • --start: Pagination offset (default 0).
  • --max_results: Number of results to return (default 10).
  • --sort_by: relevance, lastUpdatedDate, or submittedDate. (Use --sort_by submittedDate --sort_order descending for the most recent papers).
  • --sort_order: ascending or descending.

2. Download Paper (PDF or HTML)

Download the full text of a paper to your local workspace for reading.

Show full SKILL.md (210 more words)Show less
bash
uv run scripts/download_paper.py --id 1706.03762 --format pdf --output attention.pdf

Options:

  • --id: The arXiv ID (e.g., 1706.03762 or 1706.03762v5).
  • --format: pdf or html. Note: HTML is only available for newer papers.
  • --output: Filepath to save the downloaded document.

Important: when downloading papers, make sure you download them to a location where you do not overwrite other files and do not clutter existing directory structure.

3. Download Paper Source (tar.gz)

Download the LaTeX source files of a paper to your local workspace. Note that not all papers have source available.

bash
uv run scripts/download_paper_source.py --id 2010.11645 --output source.tar.gz

Options:

  • --id: The arXiv ID (e.g., 2010.11645).
  • --output: Filepath to save the downloaded tar.gz file.

Caution: Care should be exercised when untar'ing the downloaded file for security and to avoid cluttering your filesystem, as archives may contain many files or unexpected directory structures.

Safe Extraction Requirements: NEVER extract directly into your working directory! Always extract into a dedicated new directory: bash mkdir paper_source && tar -xzf source.tar.gz -C paper_source

Reference

Workflow

  1. Search for papers using search_arxiv.py. Review the JSON summaries.
  2. If full text is needed, use download_paper.py to fetch the PDF or HTML.
  3. If downloading a PDF, verify the PDF is not empty or corrupted.
  4. Read the downloaded file using standard file reading tools.

© google-deepmind, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (scripts, references) in skills/literature_search_arxiv of google-deepmind/science-skills.

  • SKILL.md
  • references/citation.bib
  • references/query_syntax.md
  • scripts/download_paper.py
  • scripts/download_paper_source.py
  • scripts/search_arxiv.py

Open the folder on GitHubat commit 8ab7672

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in google-deepmind/science-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Literature Search Arxiv next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Literature Search Arxiv compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Literature Search Arxiv this skillgoogle-deepmind/science-skills3.2k2 repos~1.2kAutomated safety check: PassApache-2.0
Literature Reviewneflibata-feng/MyArxiv-Agent12620 repos~5.9kAutomated safety check: NotesMIT
Systematic Literature Review Builderbytedance/deer-flow84k2 repos~4.3kAutomated safety check: PassMIT
Paper Research on arXivXiaomiMiMo/MiMo-Code14k—~1.5kAutomated safety check: PassMIT
Literature Review AgentAr9av/PaperOrchestra6791 repos~5.2kAutomated safety check: PassCustom licence
Arxiv MCP Serverblazickjp/arxiv-mcp-server3.2k—~353Automated safety check: PassApache-2.0

Similar skills

  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 20 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Searches arXiv across many papers on one topic, extracts each paper's methodology and findings in parallel, and synthesizes a cited literature review.

    84k GitHub starsUsed in 2 repos~4.3k tokens
    Research & ScienceAuto-check passed
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub stars~1.5k tokensUpdated 2 days ago
    Research & ScienceAuto-check passed
  • Literature Review Agent

    Ar9av/PaperOrchestra

    Step 3 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~5.2k tokens
    Research & ScienceAuto-check passed
  • Arxiv MCP Server

    blazickjp/arxiv-mcp-server

    A skill your agent uses when finding, comparing, reading, or monitoring arXiv papers, including requests for abstracts, citation graphs, original LaTeX, section-level technical details, or…

    3.2k GitHub stars~353 tokensUpdated 3 days ago
    Research & ScienceAuto-check passed
  • Arxiv Paper Writer

    appautomaton/latex-arxiv-SKILL

    Write LaTeX ML/AI review articles for arXiv using the IEEEtran template and verified BibTeX citations.

    458 GitHub stars~2.3k tokensUpdated 27 days ago
    Research & ScienceAuto-check passed

More from google-deepmind/science-skills

All 40 skills in this repo
  • Alphafold Database Fetch And Analyze

    google-deepmind/science-skills

    Retrieve and analyze AlphaFold predicted structures for a protein.

    3.2k GitHub starsUsed in 2 repos~1.2k tokens
    Auto-check passed
  • Alphagenome Single Variant Analysis

    google-deepmind/science-skills

    Analyzes genetic variant effects on gene expression (RNA-seq), chromatin accessibility (DNASE), histone marks (ChIP), and transcription factors using the AlphaGenome API.

    3.2k GitHub starsUsed in 2 repos~3k tokens
    Auto-check: notes
  • Chembl Database

    google-deepmind/science-skills

    Query the ChEMBL database for bioactive molecules, drug targets, bioactivity data, approved drugs, and chemical structures.

    3.2k GitHub starsUsed in 2 repos~2.9k tokens
    Auto-check passed
  • Clinical Trials Database

    google-deepmind/science-skills

    Query ClinicalTrials.gov via APIv2. An agent skill from google-deepmind/science-skills.

    3.2k GitHub starsUsed in 2 repos~3.2k tokens
    Auto-check passed
  • Clinvar Database

    google-deepmind/science-skills

    A skill your agent uses when needing clinical significance, pathogenicity classifications (e.g., Pathogenic, Benign, VUS), clinical evidence rationales, or finding "hard positive" benchmark controls…

    3.2k GitHub starsUsed in 2 repos~3.9k tokens
    Auto-check: notes
  • Dbsnp Database

    google-deepmind/science-skills

    A skill your agent uses when you want to look up, map, and search for short genetic variants (SNPs, indels) in NCBI's dbSNP database.

    3.2k GitHub starsUsed in 2 repos~3.4k tokens
    Auto-check: notes

Works with

Questions about Literature Search Arxiv

What does Literature Search Arxiv do?

Search for scientific papers, preprints, and publications on arXiv. Literature Search Arxiv is an agent skill from google-deepmind/science-skills. Search for scientific papers, preprints, and publications on arXiv.

When should I use Literature Search Arxiv?

Literature Search Arxiv fits situations like: the user asks to find research papers; specific arXiv IDs.

How do I install Literature Search Arxiv in Claude Code?

Run `npx skills add google-deepmind/science-skills --skill literature-search-arxiv -a claude-code`. Or copy the skill folder (skills/literature_search_arxiv in google-deepmind/science-skills) into .claude/skills/literature-search-arxiv in your project. Claude Code loads it when a task matches its description.

How do I install Literature Search Arxiv in Codex?

Run `npx skills add google-deepmind/science-skills --skill literature-search-arxiv -a codex`. Or copy the skill folder (skills/literature_search_arxiv in google-deepmind/science-skills) into .agents/skills/literature-search-arxiv in your project. Codex loads it when a task matches its description.

Can I use Literature Search Arxiv in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add google-deepmind/science-skills --skill literature-search-arxiv -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/literature-search-arxiv, .gemini/skills/literature-search-arxiv, .github/skills/literature-search-arxiv and .opencode/skills/literature-search-arxiv in your project.

What does Literature Search Arxiv need to run?

Going by SKILL.md and its folder, Literature Search Arxiv needs Python for the scripts in its folder and the command-line tools its instructions call (uv). Our summary lists: Python 3.

Does Literature Search Arxiv access the network?

SKILL.md names 1 domain. As links in the text: info.arxiv.org. This is read from the text; nothing was executed.

Is Literature Search Arxiv safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Literature Search Arxiv use?

Literature Search Arxiv is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Literature Search Arxiv use?

About 1.2k tokens (SKILL.md is roughly 4.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 341 tokens, read only when the agent opens those files.

What are the alternatives to Literature Search Arxiv?

Skills that share tags, products or a category with Literature Search Arxiv: Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Systematic Literature Review Builder (bytedance/deer-flow, 84k stars), Paper Research on arXiv (XiaomiMiMo/MiMo-Code, 14k stars) and Literature Review Agent (Ar9av/PaperOrchestra, 679 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Literature Search Arxiv?

google-deepmind (a GitHub organization) maintains it in google-deepmind/science-skills, which has 3,233 GitHub stars. The repository holds 40 skills in this directory. The repository was last updated on October 9, 2026.

Source: google-deepmind/science-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.