Agent skill

Literature Search Biorxiv

by google-deepmind in google-deepmind/science-skills

Browse, filter, and download life sciences, biology, and medical preprints from bioRxiv and medRxiv.

Apache-2.0Auto-check passedResearch & Science

Install Literature Search Biorxiv

skills CLI
$ npx skills add google-deepmind/science-skills --skill literature-search-biorxiv -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install google-deepmind/science-skills literature-search-biorxiv --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/google-deepmind/science-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/literature_search_biorxiv .claude/skills/literature-search-biorxiv && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
literature-search-biorxiv
GitHub stars
3.2k
Used in
2 other repos
Token cost
~1.9k tokens
SKILL.md length
754 words
Files
4 (incl. scripts, references)
Skills in repo
40
Repo updated
First seen
Licence
Apache-2.0

At a glance

Browse, filter, and download life sciences, biology, and medical preprints from bioRxiv and medRxiv.

  • Works in 2 steps: Search by Dates (search_by_dates.py) → Fetch Metadata by DOI (search_by_doi.py)
  • Tasks that involve Literature review
  • SKILL.md covers Prerequisites, Search Strategy Guide (Read…, Core Rules and Utility Scripts, plus 1 more section
  • Runs Python scripts from its folder; calls uv

What it does

Literature Search Biorxiv is an agent skill from google-deepmind/science-skills. Browse, filter, and download life sciences, biology, and medical preprints from bioRxiv and medRxiv. Supports fetching paper metadata by DOI, and browsing by date range with category and keyword filters. Keyword filtering is local, so date ranges MUST be narrow (1-4 weeks) with a category to prevent timeouts.

Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts and reference files (for example `scripts/search_by_dates.py` and `scripts/search_by_doi.py`).

It sits in Research & Science, covering Literature review and Academic paper search. The repository describes itself as: GDM Science Skills to speed up agentic scientific workflows with better grounding and higher token efficiency. Integrate insights from AlphaGenome, AFDB, UniProt and 30+ other… The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Literature review
  • Tasks that involve Academic paper search

Example prompts

  • “/literature-search-biorxiv”

Requirements

  • Python 3

Workflow steps

2 steps, taken from the step headings in SKILL.md.

  1. Search by Dates (search_by_dates.py)
  2. Fetch Metadata by DOI (search_by_doi.py)

What it can do on your machine

Read from SKILL.md and the folder at commit 8ab7672. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 2 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • api.biorxiv.org
    • biorxiv.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Literature Search Biorxiv loads about 1.9k tokens when it runs, and up to ~2.7k if it reads all its reference files. Until then it costs about 84 tokens; SKILL.md has 754 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~84
When it runs · the whole SKILL.md, loaded when a task matches
~1.9k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~2.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from google-deepmind/science-skills at commit 8ab7672, republished under its Apache-2.0 licence (© google-deepmind). 754 words, ~1,935 tokens.

Download SKILL.mdSave it as .claude/skills/literature-search-biorxiv/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
literature-search-biorxiv
description
Browse, filter, and download life sciences, biology, and medical preprints from bioRxiv and medRxiv. Supports fetching paper metadata by DOI, and browsing by date range with category and keyword filters. Keyword filtering is local, so date ranges MUST be narrow (1-4 weeks) with a category to prevent timeouts.

Prerequisites

  1. uv: Read the uv skill and follow its Setup instructions to ensure uv is installed and on PATH.
  2. User Notification: If .licenses/literature_search_biorxiv_LICENSE.txt does not already exist in the workspace root directory then (1) prominently notify the user to check the terms at https://api.biorxiv.org/ and https://www.biorxiv.org/content/about-biorxiv and to always check the license of the papers retrieved by the skill for any restrictions, then (2) create the file recording the notification text and timestamp.

Search Strategy Guide (Read First)

This skill browses a date-based preprint archive. It is NOT a keyword search engine. Choose your approach based on what you already know:

  • A DOI (e.g., from a citation): Use search_by_doi.py. Fast and reliable.
  • Approximate date + category: Use search_by_dates.py with a 1–4 week range and --category.
  • Only a topic or keywords, no date: Do NOT use this skill for discovery. Use a keyword-capable literature skill first to find relevant DOIs, then return here to fetch metadata.

CRITICAL ANTI-PATTERN — Do NOT do this: Do NOT attempt to search broad date ranges (months or years) with --keywords hoping to find a specific paper. The bioRxiv API does not support server-side keyword search. The script must download ALL metadata for the entire date range and filter locally in Python. Broad ranges will result in thousands of API calls, timeouts, and your request being blocked for API abuse. This is the #1 reason this skill fails.

Core Rules

  • Use the Wrapper: ALWAYS execute the provided helper scripts to query the database rather than accessing the database directly. The scripts automatically enforce the required rate limit gracefully.
  • Local Filtering (CRITICAL WARNING): Unlike arXiv, the bioRxiv API does not support server-side keyword or author searches. Keyword and author filtering is performed locally by the scripts after downloading all metadata for a specified date range. You MUST use narrow date ranges (e.g., 1-4 weeks) AND the --category filter when searching with --keywords or --author.
  • Abstracts Excluded By Default: To save context space in the resulting JSON, abstracts are stripped from the output by default. If you are searching by --keywords and want to read the abstracts of the resulting papers to understand their context, you MUST pass the --include_abstracts flag.
  • Output Redirection: Search commands output JSON arrays to standard output. Always redirect output to a file (e.g., > results.json) and parse the file separately.
  • List Sources If this skill is used, ensure this is mentioned in the output AND list the URLs of all papers that were used in producing the output.
Show full SKILL.md (335 more words)Show less

Utility Scripts

All tools enforce a cross-process rate limits and retry with backoff on failure. To ensure you respect terms-of-service, do NOT write custom curl queries.

Pagination: The bioRxiv API returns results in pages of up to 100 papers. The search_by_dates.py script automatically fetches all pages and reports pagination progress to stderr (e.g., [Page 2] Fetched 200/543 papers...). The JSON output to stdout contains the complete filtered result set across all pages — no manual pagination is needed.

1. Search by Dates (search_by_dates.py)

Search for preprints within an explicit date range, optionally filtering by category, keywords, or author.

bash
# Broad category search over a 2-week period
uv run scripts/search_by_dates.py --server biorxiv \
  --start_date 2024-01-01 --end_date 2024-01-14 \
  --category neuroscience > results.json

# Deep keyword filtering using OR logic and including abstracts
uv run scripts/search_by_dates.py --server medrxiv \
  --start_date 2023-11-01 --end_date 2023-11-30 \
  --category infectious_diseases \
  --keywords "covid" "sars-cov-2" --match_logic OR \
  --include_abstracts > covid_papers.json

# Finding papers by a specific author in a narrow window
uv run scripts/search_by_dates.py \
  --start_date 2024-05-01 --end_date 2024-05-14 \
  --author "Smith" > smith_papers.json

Required Arguments:

  • --start_date: YYYY-MM-DD
  • --end_date: YYYY-MM-DD

Optional Arguments:

  • --server: biorxiv (default) or medrxiv
  • --category: A valid subject category (see below). Highly recommended — dramatically reduces the data the script must download and filter.
  • --keywords: List of strings to search in the title/abstract.
  • --match_logic: AND (default) or OR for keywords.
  • --author: Author name (case-insensitive string match).
  • --include_abstracts: Flag to include full abstracts in the JSON output.
2. Fetch Metadata by DOI (search_by_doi.py)

Retrieve the detailed JSON metadata for a single paper if you already know its DOI. This is the most reliable entry point.

bash
uv run scripts/search_by_doi.py --server biorxiv \
  --doi "10.1101/2023.08.15.551388" \
  --include_abstracts > paper_info.json
Downloading Full-Text PDFs

This skill does NOT support PDF downloads. To download the full-text PDF of a bioRxiv or medRxiv preprint, use the literature-search-europepmc skill. First, use the paper's DOI to look up its PMCID via EuropePMC, then use EuropePMC's PDF retrieval to download the document.

Valid Subject Categories

You can pass these to the --category flag in search_by_dates.py. The script will strictly validate them.

bioRxiv Categories:

animal_behavior_and_cognition, biochemistry, bioengineering, bioinformatics, biophysics, cancer_biology, cell_biology, clinical_trials, developmental_biology, ecology, epidemiology, evolutionary_biology, genetics, genomics, immunology, microbiology, molecular_biology, neuroscience, paleontology, pathology, pharmacology_and_toxicology, physiology, plant_biology, scientific_communication_and_education, synthetic_biology, systems_biology, zoology

medRxiv Categories:

addiction_medicine, allergy_and_immunology, anesthesia, cardiovascular_medicine, dentistry_and_oral_medicine, dermatology, emergency_medicine, endocrinology, epidemiology, forensic_medicine, gastroenterology, genetic_and_genomic_medicine, health_informatics, health_economics_and_outcomes_research, health_policy, health_systems_and_quality_improvement, hematology, hiv_aids, infectious_diseases, intensive_care_and_critical_care_medicine, medical_education, medical_ethics, nephrology, neurology, nursing, nutrition, obstetrics_and_gynecology, occupational_and_environmental_health, oncology, ophthalmology, orthopedics, otolaryngology, pain_medicine, palliative_care, pathology, pediatrics, pharmacology_and_therapeutics, primary_care_research, psychiatry_and_clinical_psychology, public_and_global_health, radiology_and_imaging, rehabilitation_medicine_and_physical_therapy, respiratory_medicine, rheumatology, sexual_and_reproductive_health, sports_medicine, surgery, toxicology, transplantation, urology

© google-deepmind, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts, references) in skills/literature_search_biorxiv of google-deepmind/science-skills.

  • SKILL.md
  • references/citation.bib
  • scripts/search_by_dates.py
  • scripts/search_by_doi.py

Open the folder on GitHubat commit 8ab7672

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in google-deepmind/science-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Literature Search Biorxiv next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Literature Search Biorxiv compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Literature Search Biorxiv this skillgoogle-deepmind/science-skills3.2k2 repos~1.9kAutomated safety check: PassApache-2.0
Literature Reviewneflibata-feng/MyArxiv-Agent12620 repos~5.9kAutomated safety check: NotesMIT
Preprint Search on bioRxivLigphiDonk/Oh-my--paper73912 repos~3.7kAutomated safety check: PassMIT
Systematic Literature Review Builderbytedance/deer-flow84k2 repos~4.3kAutomated safety check: PassMIT
Paper Research on arXivXiaomiMiMo/MiMo-Code14k—~1.5kAutomated safety check: PassMIT
Literature Review AgentAr9av/PaperOrchestra6791 repos~5.2kAutomated safety check: PassCustom licence

Similar skills

  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 20 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Preprint Search on bioRxiv

    LigphiDonk/Oh-my--paper

    Searches bioRxiv life sciences preprints by keyword, author, date range or category with a Python script, returning JSON metadata and optional PDF downloads.

    739 GitHub starsUsed in 12 repos~3.7k tokens
    Research & ScienceAuto-check passed
  • Searches arXiv across many papers on one topic, extracts each paper's methodology and findings in parallel, and synthesizes a cited literature review.

    84k GitHub starsUsed in 2 repos~4.3k tokens
    Research & ScienceAuto-check passed
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub stars~1.5k tokensUpdated 2 days ago
    Research & ScienceAuto-check passed
  • Literature Review Agent

    Ar9av/PaperOrchestra

    Step 3 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~5.2k tokens
    Research & ScienceAuto-check passed
  • Conference Paper Recommender

    juliye2025/evil-read-arxiv

    Searches DBLP for papers from top conferences such as CVPR, ICLR and NeurIPS, adds Semantic Scholar data, scores them and writes a yearly Obsidian note.

    1.7k GitHub stars~2.5k tokensUpdated 26 days ago
    Research & ScienceAuto-check passed

More from google-deepmind/science-skills

All 40 skills in this repo
  • Alphafold Database Fetch And Analyze

    google-deepmind/science-skills

    Retrieve and analyze AlphaFold predicted structures for a protein.

    3.2k GitHub starsUsed in 2 repos~1.2k tokens
    Auto-check passed
  • Alphagenome Single Variant Analysis

    google-deepmind/science-skills

    Analyzes genetic variant effects on gene expression (RNA-seq), chromatin accessibility (DNASE), histone marks (ChIP), and transcription factors using the AlphaGenome API.

    3.2k GitHub starsUsed in 2 repos~3k tokens
    Auto-check: notes
  • Chembl Database

    google-deepmind/science-skills

    Query the ChEMBL database for bioactive molecules, drug targets, bioactivity data, approved drugs, and chemical structures.

    3.2k GitHub starsUsed in 2 repos~2.9k tokens
    Auto-check passed
  • Clinical Trials Database

    google-deepmind/science-skills

    Query ClinicalTrials.gov via APIv2. An agent skill from google-deepmind/science-skills.

    3.2k GitHub starsUsed in 2 repos~3.2k tokens
    Auto-check passed
  • Clinvar Database

    google-deepmind/science-skills

    A skill your agent uses when needing clinical significance, pathogenicity classifications (e.g., Pathogenic, Benign, VUS), clinical evidence rationales, or finding "hard positive" benchmark controls…

    3.2k GitHub starsUsed in 2 repos~3.9k tokens
    Auto-check: notes
  • Dbsnp Database

    google-deepmind/science-skills

    A skill your agent uses when you want to look up, map, and search for short genetic variants (SNPs, indels) in NCBI's dbSNP database.

    3.2k GitHub starsUsed in 2 repos~3.4k tokens
    Auto-check: notes

Questions about Literature Search Biorxiv

What does Literature Search Biorxiv do?

Browse, filter, and download life sciences, biology, and medical preprints from bioRxiv and medRxiv. Literature Search Biorxiv is an agent skill from google-deepmind/science-skills. Browse, filter, and download life sciences, biology, and medical preprints from bioRxiv and medRxiv.

When should I use Literature Search Biorxiv?

Literature Search Biorxiv fits situations like: tasks that involve Literature review; tasks that involve Academic paper search.

How do I install Literature Search Biorxiv in Claude Code?

Run `npx skills add google-deepmind/science-skills --skill literature-search-biorxiv -a claude-code`. Or copy the skill folder (skills/literature_search_biorxiv in google-deepmind/science-skills) into .claude/skills/literature-search-biorxiv in your project. Claude Code loads it when a task matches its description.

How do I install Literature Search Biorxiv in Codex?

Run `npx skills add google-deepmind/science-skills --skill literature-search-biorxiv -a codex`. Or copy the skill folder (skills/literature_search_biorxiv in google-deepmind/science-skills) into .agents/skills/literature-search-biorxiv in your project. Codex loads it when a task matches its description.

Can I use Literature Search Biorxiv in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add google-deepmind/science-skills --skill literature-search-biorxiv -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/literature-search-biorxiv, .gemini/skills/literature-search-biorxiv, .github/skills/literature-search-biorxiv and .opencode/skills/literature-search-biorxiv in your project.

What does Literature Search Biorxiv need to run?

Going by SKILL.md and its folder, Literature Search Biorxiv needs Python for the scripts in its folder and the command-line tools its instructions call (uv). Our summary lists: Python 3.

Does Literature Search Biorxiv access the network?

SKILL.md names 2 domains. As links in the text: api.biorxiv.org and biorxiv.org. This is read from the text; nothing was executed.

Is Literature Search Biorxiv safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Literature Search Biorxiv use?

Literature Search Biorxiv is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Literature Search Biorxiv use?

About 1.9k tokens (SKILL.md is roughly 7.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 755 tokens, read only when the agent opens those files.

What are the alternatives to Literature Search Biorxiv?

Skills that share tags, products or a category with Literature Search Biorxiv: Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Preprint Search on bioRxiv (LigphiDonk/Oh-my--paper, 739 stars), Systematic Literature Review Builder (bytedance/deer-flow, 84k stars) and Paper Research on arXiv (XiaomiMiMo/MiMo-Code, 14k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Literature Search Biorxiv?

google-deepmind (a GitHub organization) maintains it in google-deepmind/science-skills, which has 3,233 GitHub stars. The repository holds 40 skills in this directory. The repository was last updated on October 9, 2026.

Source: google-deepmind/science-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.