Agent skill

Literature Search Openalex

by google-deepmind in google-deepmind/science-skills

Query the OpenAlex scholarly database for research papers, authors, institutions, topics, sources, publishers, funders, geo-locations, and keywords.

Apache-2.0Auto-check: notesResearch & Science

Install Literature Search Openalex

skills CLI
$ npx skills add google-deepmind/science-skills --skill literature-search-openalex -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install google-deepmind/science-skills literature-search-openalex --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/google-deepmind/science-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/literature_search_openalex .claude/skills/literature-search-openalex && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
literature-search-openalex
GitHub stars
3.2k
Used in
2 other repos
Token cost
~1.8k tokens
SKILL.md length
676 words
Files
12 (incl. scripts, references)
Skills in repo
40
Repo updated
First seen
Licence
Apache-2.0

At a glance

Query the OpenAlex scholarly database for research papers, authors, institutions, topics, sources, publishers, funders, geo-locations, and keywords.

  • Works in 4 steps: uv: Read the uv skill and follow its… → User Notification: If… → .env file: Make sure the .env file… → …
  • Searching academic papers
  • SKILL.md covers Prerequisites, Core Rules, Rate Limits and CLI Reference, plus 3 more sections
  • Runs Python scripts from its folder; calls uv and jq; reaches doi.org; needs OPENALEX_API_KEY

What it does

Literature Search Openalex is an agent skill from google-deepmind/science-skills. Query the OpenAlex scholarly database for research papers, authors, institutions, topics, sources, publishers, funders, geo-locations, and keywords. Use when searching academic papers, resolving DOIs, downloading open-access PDFs, finding an author's publications, aggregating bibliometric data (citation counts, h-index, impact factor), exploring the research taxonomies, or performing DOI lookups.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 13 other files, including scripts and reference files (for example `references/authors.md`, `references/geo_and_language.md` and `references/institutions.md`).

It sits in Research & Science, covering Academic paper search and Literature review. The repository describes itself as: GDM Science Skills to speed up agentic scientific workflows with better grounding and higher token efficiency. Integrate insights from AlphaGenome, AFDB, UniProt and 30+ other… The licence is Apache-2.0.

When your agent uses it

  • Searching academic papers
  • Downloading open-access PDFs
  • Finding an authors publications
  • Aggregating bibliometric data (citation counts

Example prompts

  • “/literature-search-openalex”

Requirements

  • Python 3
  • A credential in OPENALEX_API_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. uv: Read the uv skill and follow its Setup instructions to ensure
  2. User Notification: If .licenses/literature_search_openalex_LICENSE.txt
  3. .env file: Make sure the .env file exists in your home directory.
  4. OPENALEX_API_KEY (optional but recommended): Enables the OpenAlex

What it can do on your machine

Read from SKILL.md and the folder at commit 8ab7672. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv
    • jq

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • doi.org

    Also links to:

    • developers.openalex.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENALEX_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Literature Search Openalex loads about 1.8k tokens when it runs, and up to ~12k if it reads all its reference files. Until then it costs about 107 tokens; SKILL.md has 676 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~107
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~12k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:24
    3.  **`.env` file**: Make sure the `.env` file exists in your home directory.
  • NoteMentions a .env fileSKILL.md:47
    their `.env` file.
  • NoteMentions a .env fileSKILL.md:156
    :      :                     : `.env`                        :
  • NoteMentions a .env fileSKILL.md:164
    : user add API key to `.env`    :

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from google-deepmind/science-skills at commit 8ab7672, republished under its Apache-2.0 licence (© google-deepmind). 676 words, ~1,844 tokens.

Download SKILL.mdSave it as .claude/skills/literature-search-openalex/SKILL.md (or your agent's skills folder). This skill also uses 11 other files; get the full folder from GitHub.
name
literature-search-openalex
description
Query the OpenAlex scholarly database for research papers, authors, institutions, topics, sources, publishers, funders, geo-locations, and keywords. Use when searching academic papers, resolving DOIs, downloading open-access PDFs, finding an author's publications, aggregating bibliometric data (citation counts, h-index, impact factor), exploring the research taxonomies, or performing DOI lookups.

OpenAlex Skill

Prerequisites

  1. uv: Read the uv skill and follow its Setup instructions to ensure uv is installed and on PATH.
  2. User Notification: If .licenses/literature_search_openalex_LICENSE.txt does not already exist in the workspace root directory then (1) prominently notify the user to check the terms at https://developers.openalex.org/ and to always check the license of the papers retrieved by the skill for any restrictions, then (2) create the file recording the notification text and timestamp.
  3. .env file: Make sure the .env file exists in your home directory. Create one if it does not exist.
  4. OPENALEX_API_KEY (optional but recommended): Enables the OpenAlex Premium API with higher rate limits. The skill works without it (using the free "polite pool"). You can obtain a key at OpenAlex.org → account settings. You MUST use the safe credentials protocol in the credentials skill to check for and request this key if this skill looks relevant to the user's request.

Core Rules

  1. List Sources. If this skill is used, ensure this is mentioned in the output AND list the URLs of all papers that were used in producing the output.
  2. Resolve before filter. NEVER filter by name. Always resolve a name to an ID first, then use that ID in --filter.
  3. Use the CLI only. Never call the API via curl/urllib. The CLI handles retries and rate limiting.
  4. No fabrication. Never invent OpenAlex IDs or DOIs. Use resolve/get to look them up. Report empty results accurately.
  5. API key. If a command returns 401/429 or you need high-volume queries, you MUST use the safe credentials protocol in the credentials skill to check for and request the OPENALEX_API_KEY to help the user add it to their .env file.
  6. Keep output small. Always use --select and --per-page 5–10 for overview queries. Pipe filter output to a file (> results.json), then slim with jq before reading into context.

Rate Limits

  • With key: ~10 req/s, $1/day free budget.
  • Without key: Very limited, $0.01/day budget.
OperationCost
Singleton getFree
filter$0.0001
--search / resolve$0.001
download-pdf$0.01

CLI Reference

uv run scripts/openalex_cli.py [--api-key KEY] <command> [flags]

Entity types (shared across commands): works, authors, sources, institutions, topics, domains, fields, subfields, sdgs, countries, continents, languages, keywords, publishers, funders, work-types, source-types, institution-types, licenses

Show full SKILL.md (309 more words)Show less
Commands

resolve <entity> <query> — Name → ID candidates. Returns id, display_name, hint. Use --per-page N for more candidates.

get <entity> <id> — Full metadata for one entity. Accepts short ID (W2741809807), full URL, or DOI URL. Use --select to limit fields.

filter <entity> — Search/filter entities. Key flags are:

  • --search <query>: Full-text search (10× cost of --filter)
  • --filter <expr>: Filter expressions. Use , for AND and | for OR.
  • --sort <field:dir>: Sort results (e.g., cited_by_count:desc)
  • --select <fields>: Limit the fields returned in the output.
  • --group-by <field>: Aggregate results by a specific field.
  • --per-page <N>: Number of results per page (default 25, max 100).
  • --page <N>: Specify the page number to retrieve.
  • --sample <N>: Get a random sample of up to 10,000 results.
  • --seed <N>: Seed for reproducible sampling.

download-pdf <work-id> <output-path> — Download PDF (requires API key). Falls back to alternative pdf_url locations if primary fails. Whenever you download a PDF, verify it is not empty or corrupted.

rate-limit — Check current rate limit status (requires API key).

Search Tips
  • If resolve returns no matches, try alternate spellings or abbreviations.
  • If --search returns 0 results, try broader terms (max 3 retries).
  • If resolve returns multiple candidates, present them to the user with display_name and hint for manual selection.

Entity References

Consult references/ for valid filter, sort, and group-by fields per entity:

Common Workflows

bash
# Author's works (resolve → filter)
uv run scripts/openalex_cli.py resolve authors "Geoffrey Hinton"
uv run scripts/openalex_cli.py filter works \
  --filter "authorships.author.id:A5108093963" \
  --sort "cited_by_count:desc" --per-page 10 > papers.json
cat papers.json | jq '[.results[] | {id, title: .display_name, year: .publication_year, citations: .cited_by_count}]'

# DOI lookup
uv run scripts/openalex_cli.py get works "https://doi.org/10.1038/s41586-021-03819-2"

# Bulk DOI lookup (up to 100)
uv run scripts/openalex_cli.py filter works \
  --filter "doi:10.1234/a|10.1234/b|10.1234/c" --per-page 100 > results.json

# Institutional impact by year
uv run scripts/openalex_cli.py resolve institutions "MIT"
uv run scripts/openalex_cli.py filter works \
  --filter "authorships.institutions.id:I63966007" \
  --group-by "publication_year" > mit_by_year.json

# Random sample
uv run scripts/openalex_cli.py filter works \
  --filter "publication_year:2023,is_oa:true" \
  --sample 100 --seed 42 > results.json

Error Handling

CodeMeaningAction
401UnauthorizedYou MUST use safe credentials
: : : protocol in credentials skill :
: : : to help user add API key to :
: : : .env :
403Plan upgrade neededInform user; see
: : : https://openalex.org/pricing :
404Not foundVerify ID; try resolve
: : : first :
429Rate limitedWait and retry; you MUST use
: : : safe credentials protocol in :
: : : credentials skill to help :
: : : user add API key to .env :

Known premium-only filters: from_updated_date, to_updated_date.

Never fabricate results on empty responses — report accurately and suggest alternate search terms.

© google-deepmind, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 11 other files (scripts, references) in skills/literature_search_openalex of google-deepmind/science-skills.

  • SKILL.md
  • references/authors.md
  • references/citation.bib
  • references/geo_and_language.md
  • references/institutions.md
  • references/publishers_funders.md
  • references/sources.md
  • references/taxonomy.md
  • references/topics.md
  • references/type_values.md
  • references/works.md
  • scripts/openalex_cli.py

Open the folder on GitHubat commit 8ab7672

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in google-deepmind/science-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Literature Search Openalex next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Literature Search Openalex compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Literature Search Openalex this skillgoogle-deepmind/science-skills3.2k2 repos~1.8kAutomated safety check: NotesApache-2.0
Literature Reviewneflibata-feng/MyArxiv-Agent12620 repos~5.9kAutomated safety check: NotesMIT
Preprint Search on bioRxivLigphiDonk/Oh-my--paper73912 repos~3.7kAutomated safety check: PassMIT
Systematic Literature Review Builderbytedance/deer-flow84k2 repos~4.3kAutomated safety check: PassMIT
Paper Research on arXivXiaomiMiMo/MiMo-Code14k—~1.5kAutomated safety check: PassMIT
Literature Review AgentAr9av/PaperOrchestra6791 repos~5.2kAutomated safety check: PassCustom licence

Similar skills

  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 20 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Preprint Search on bioRxiv

    LigphiDonk/Oh-my--paper

    Searches bioRxiv life sciences preprints by keyword, author, date range or category with a Python script, returning JSON metadata and optional PDF downloads.

    739 GitHub starsUsed in 12 repos~3.7k tokens
    Research & ScienceAuto-check passed
  • Searches arXiv across many papers on one topic, extracts each paper's methodology and findings in parallel, and synthesizes a cited literature review.

    84k GitHub starsUsed in 2 repos~4.3k tokens
    Research & ScienceAuto-check passed
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub stars~1.5k tokensUpdated 2 days ago
    Research & ScienceAuto-check passed
  • Literature Review Agent

    Ar9av/PaperOrchestra

    Step 3 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~5.2k tokens
    Research & ScienceAuto-check passed
  • Conference Paper Recommender

    juliye2025/evil-read-arxiv

    Searches DBLP for papers from top conferences such as CVPR, ICLR and NeurIPS, adds Semantic Scholar data, scores them and writes a yearly Obsidian note.

    1.7k GitHub stars~2.5k tokensUpdated 26 days ago
    Research & ScienceAuto-check passed

More from google-deepmind/science-skills

All 40 skills in this repo
  • Alphafold Database Fetch And Analyze

    google-deepmind/science-skills

    Retrieve and analyze AlphaFold predicted structures for a protein.

    3.2k GitHub starsUsed in 2 repos~1.2k tokens
    Auto-check passed
  • Alphagenome Single Variant Analysis

    google-deepmind/science-skills

    Analyzes genetic variant effects on gene expression (RNA-seq), chromatin accessibility (DNASE), histone marks (ChIP), and transcription factors using the AlphaGenome API.

    3.2k GitHub starsUsed in 2 repos~3k tokens
    Auto-check: notes
  • Chembl Database

    google-deepmind/science-skills

    Query the ChEMBL database for bioactive molecules, drug targets, bioactivity data, approved drugs, and chemical structures.

    3.2k GitHub starsUsed in 2 repos~2.9k tokens
    Auto-check passed
  • Clinical Trials Database

    google-deepmind/science-skills

    Query ClinicalTrials.gov via APIv2. An agent skill from google-deepmind/science-skills.

    3.2k GitHub starsUsed in 2 repos~3.2k tokens
    Auto-check passed
  • Clinvar Database

    google-deepmind/science-skills

    A skill your agent uses when needing clinical significance, pathogenicity classifications (e.g., Pathogenic, Benign, VUS), clinical evidence rationales, or finding "hard positive" benchmark controls…

    3.2k GitHub starsUsed in 2 repos~3.9k tokens
    Auto-check: notes
  • Dbsnp Database

    google-deepmind/science-skills

    A skill your agent uses when you want to look up, map, and search for short genetic variants (SNPs, indels) in NCBI's dbSNP database.

    3.2k GitHub starsUsed in 2 repos~3.4k tokens
    Auto-check: notes

Questions about Literature Search Openalex

What does Literature Search Openalex do?

Query the OpenAlex scholarly database for research papers, authors, institutions, topics, sources, publishers, funders, geo-locations, and keywords. Literature Search Openalex is an agent skill from google-deepmind/science-skills. Query the OpenAlex scholarly database for research papers, authors, institutions, topics, sources, publishers, funders, geo-locations, and keywords.

When should I use Literature Search Openalex?

Literature Search Openalex fits situations like: searching academic papers; downloading open-access PDFs; finding an authors publications; aggregating bibliometric data (citation counts.

How do I install Literature Search Openalex in Claude Code?

Run `npx skills add google-deepmind/science-skills --skill literature-search-openalex -a claude-code`. Or copy the skill folder (skills/literature_search_openalex in google-deepmind/science-skills) into .claude/skills/literature-search-openalex in your project. Claude Code loads it when a task matches its description.

How do I install Literature Search Openalex in Codex?

Run `npx skills add google-deepmind/science-skills --skill literature-search-openalex -a codex`. Or copy the skill folder (skills/literature_search_openalex in google-deepmind/science-skills) into .agents/skills/literature-search-openalex in your project. Codex loads it when a task matches its description.

Can I use Literature Search Openalex in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add google-deepmind/science-skills --skill literature-search-openalex -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/literature-search-openalex, .gemini/skills/literature-search-openalex, .github/skills/literature-search-openalex and .opencode/skills/literature-search-openalex in your project.

What does Literature Search Openalex need to run?

Going by SKILL.md and its folder, Literature Search Openalex needs Python for the scripts in its folder, the command-line tools its instructions call (uv and jq) and credentials named OPENALEX_API_KEY. Our summary lists: Python 3; A credential in OPENALEX_API_KEY.

Does Literature Search Openalex access the network?

SKILL.md names 2 domains. In commands or code: doi.org; the agent is likely to contact it when it follows the instructions. As links in the text: developers.openalex.org. This is read from the text; nothing was executed.

Is Literature Search Openalex safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Literature Search Openalex use?

Literature Search Openalex is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Literature Search Openalex use?

About 1.8k tokens (SKILL.md is roughly 7.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 10k tokens, read only when the agent opens those files.

What are the alternatives to Literature Search Openalex?

Skills that share tags, products or a category with Literature Search Openalex: Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Preprint Search on bioRxiv (LigphiDonk/Oh-my--paper, 739 stars), Systematic Literature Review Builder (bytedance/deer-flow, 84k stars) and Paper Research on arXiv (XiaomiMiMo/MiMo-Code, 14k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Literature Search Openalex?

google-deepmind (a GitHub organization) maintains it in google-deepmind/science-skills, which has 3,233 GitHub stars. The repository holds 40 skills in this directory. The repository was last updated on October 9, 2026.

Source: google-deepmind/science-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.