Discover, evaluate, and download publicly available datasets from the internet.

No licenceAuto-check: warningsResearch & Science

Install Data Download

The automated check flagged lines worth reading first. See the safety section below.

skills CLI
$ npx skills add GRIND-Lab-Core/night_owl_research_agent --skill data-download -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install GRIND-Lab-Core/night_owl_research_agent data-download --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/GRIND-Lab-Core/night_owl_research_agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/data-download .claude/skills/data-download && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
data-download
GitHub stars
106
Token cost
~8.1k tokens
SKILL.md length
2,709 words
Files
1
Skills in repo
19
Repo updated
First seen
Licence
None found

At a glance

Discover, evaluate, and download publicly available datasets from the internet.

  • Works in 7 steps: Understand the Data Need → Discover Data Sources → Evaluate and Select Sources → …
  • User says download data
  • SKILL.md covers Overview, Constants, Phase 0: Understand the Data… and Phase 1: Discover Data Sources, plus 6 more sections
  • Calls curl and wget; reaches api.census.gov and api.open-meteo.com; needs API_KEY

What it does

Data Download is an agent skill from GRIND-Lab-Core/night_owl_research_agent. Discover, evaluate, and download publicly available datasets from the internet. Infers data needs from a research question or task, selects authoritative sources, downloads reproducibly, validates file integrity, and documents provenance. Pauses for user input when authentication, API keys, or major tradeoffs require a decision. Use when user says "download data", "get data", "find a dataset", "I need boundary files", "download census data", or needs any external dataset for analysis.

Its SKILL.md is about 8.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Research & Science, covering Hypothesis generation. The repository describes itself as: Fully automatic AI research agent for Geoscientists, Remote Sensing researchers, and GIScientists. Features harness engineering, GeoBenchmark (OLS/GWR/MGWR), journal templates…

When your agent uses it

  • User says download data
  • I need boundary files
  • Download census data
  • Needs any external dataset for analysis

Example prompts

  • “download data”
  • “get data”
  • “find a dataset”
  • “/data-download”

Requirements

  • Python 3
  • A credential in API_KEY
  • Pre-approved tools (allowed-tools): Bash(*), Read, Write, Edit, Grep, Glob, WebSearch, WebFetch, Agent

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. Understand the Data Need
  2. Discover Data Sources
  3. Evaluate and Select Sources
  4. Download
  5. 5: Human Checkpoint — Data Synthesis
  6. Validate Downloads
  7. Document Provenance

What it can do on your machine

Read from SKILL.md and the folder at commit 5f7a246. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(*)
    • Read
    • Write
    • Edit
    • Grep
    • Glob
    • WebSearch
    • WebFetch
    • Agent

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl
    • wget

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.census.gov
    • api.open-meteo.com
    • raw.githubusercontent.com
    • planetarycomputer.microsoft.com
    • overpass-api.de
    • www2.census.gov

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Data Download loads about 8.1k tokens when it runs. Until then it costs about 126 tokens; SKILL.md has 2,709 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~126
When it runs · the whole SKILL.md, loaded when a task matches
~8.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: warnings

The automated check found patterns that need a careful read before installing.

  • WarningTells the agent its actions are pre-authorized / not to stop for confirmationSKILL.md:736
    t << 'EOF' > file`) to write in chunks. Do NOT ask the user for permission — just do it silently.
  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash(*), Read, Write, Edit, Grep, Glob, WebSearch, WebFetch, Agent

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 2,709 words (~8,100 tokens).

“Find and download data for: $ARGUMENTS”

— opening of SKILL.md by GRIND-Lab-Core
name
data-download
allowed-tools
Bash(*), Read, Write, Edit, Grep, Glob, WebSearch, WebFetch, Agent
argument-hint
data-need-or-research-question

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills/data-download of GRIND-Lab-Core/night_owl_research_agent.

Open the folder on GitHubat commit 5f7a246

Compare with similar skills

Data Download next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Data Download compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Data Download this skillGRIND-Lab-Core/night_owl_research_agent106—~8.1kAutomated safety check: WarnNone
Hypothesis Generationspacering-net/codeg3.9k14 repos~3.6kAutomated safety check: NotesMIT
Nature Paper CardYuan1z0825/nature-skills47k2 repos~2.1kAutomated safety check: PassApache-2.0
Hypothesis GenerationK-Dense-AI/claude-scientific-writer2.4k2 repos~3.9kAutomated safety check: PassMIT
Good QuestionRimagination/good-question3051 repos~4.3kAutomated safety check: PassMIT
High Stakes Analytics Decision Lablimingrui679-design/high-stakes-analytics-decision-lab1k—~2.2kAutomated safety check: PassMIT

Similar skills

  • Hypothesis Generation

    spacering-net/codeg

    Structured hypothesis formulation from observations. An agent skill from spacering-net/codeg.

    3.9k GitHub starsUsed in 14 repos~3.6k tokens
    Research & ScienceAuto-check: notes
  • Nature Paper Card

    Yuan1z0825/nature-skills

    Builds a structured deep-reading card for one scientific paper, covering methods, how experiments support claims, limitations and research ideas, with a script to prepare the source.

    47k GitHub starsUsed in 2 repos~2.1k tokens
    Research & ScienceAuto-check passed
  • Hypothesis Generation

    K-Dense-AI/claude-scientific-writer

    Formulate evidence-bounded scientific questions, candidate hypotheses, rival explanations, causal or associational claims, discriminating predictions, measurements, and preregistration-ready…

    2.4k GitHub starsUsed in 2 repos~3.9k tokens
    Research & ScienceAuto-check passed
  • Good Question

    Rimagination/good-question

    A skill your agent uses when a researcher is choosing, framing, refining, or stress-testing a research question, hypothesis, thesis topic, project idea, grant direction, paper angle, or stalled…

    305 GitHub starsUsed in 1 repo~4.3k tokens
    Research & ScienceAuto-check passed
  • High Stakes Analytics Decision Lab

    limingrui679-design/high-stakes-analytics-decision-lab

    Build or review source-backed descriptive, diagnostic, predictive, and prescriptive analysis for consequential decisions.

    1k GitHub stars~2.2k tokensUpdated 4 days ago
    Research & ScienceAuto-check passed
  • Claim-Driven Experiment Planner

    zjYao36/Auto-Research-Refine

    Turns a refined research proposal into a claim-to-evidence-to-run-order roadmap instead of a sprawling benchmark wishlist.

    128 GitHub starsUsed in 6 repos~2.3k tokens
    Research & ScienceAuto-check: notes

More from GRIND-Lab-Core/night_owl_research_agent

All 19 skills in this repo
  • Research Review

    GRIND-Lab-Core/night_owl_research_agent

    Get a deep critical review of research idea from GPT via Codex MCP.

    106 GitHub starsUsed in 5 repos~1.1k tokens
    Auto-check: notes
  • Deploy Experiment

    GRIND-Lab-Core/night_owl_research_agent

    Deploy and run experiments for ML/DL training (local, remote, or Modal GPU) AND spatial data science / GIScience experiments (local, data-driven).

    106 GitHub stars~5k tokensUpdated 5 mo ago
    Auto-check: notes
  • Experiment Design

    GRIND-Lab-Core/night_owl_research_agent

    Turn a refined GIScience / remote sensing / spatial data science proposal into a detailed, claim-driven experiment roadmap.

    106 GitHub stars~3.3k tokensUpdated 5 mo ago
    Auto-check: warnings
  • Experiment Design Pipeline

    GRIND-Lab-Core/night_owl_research_agent

    Run an end-to-end workflow that chains the skills refine-research and experiment-design.

    106 GitHub stars~2.1k tokensUpdated 5 mo ago
    Auto-check: warnings
  • Generate Idea

    GRIND-Lab-Core/night_owl_research_agent

    Generate and rank research ideas given a broad direction. An agent skill from GRIND-Lab-Core/night_owl_research_agent.

    106 GitHub stars~3.2k tokensUpdated 5 mo ago
    Auto-check: warnings
  • Paper Writing Pipeline

    GRIND-Lab-Core/night_owl_research_agent

    Full paper writing pipeline. An agent skill from GRIND-Lab-Core/night_owl_research_agent.

    106 GitHub stars~5.5k tokensUpdated 5 mo ago
    Auto-check: warnings

Questions about Data Download

What does Data Download do?

Discover, evaluate, and download publicly available datasets from the internet. Data Download is an agent skill from GRIND-Lab-Core/night_owl_research_agent. Discover, evaluate, and download publicly available datasets from the internet.

When should I use Data Download?

Data Download fits situations like: user says download data; I need boundary files; download census data; needs any external dataset for analysis.

How do I install Data Download in Claude Code?

Run `npx skills add GRIND-Lab-Core/night_owl_research_agent --skill data-download -a claude-code`. Or copy the skill folder (skills/data-download in GRIND-Lab-Core/night_owl_research_agent) into .claude/skills/data-download in your project. Claude Code loads it when a task matches its description.

How do I install Data Download in Codex?

Run `npx skills add GRIND-Lab-Core/night_owl_research_agent --skill data-download -a codex`. Or copy the skill folder (skills/data-download in GRIND-Lab-Core/night_owl_research_agent) into .agents/skills/data-download in your project. Codex loads it when a task matches its description.

Can I use Data Download in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add GRIND-Lab-Core/night_owl_research_agent --skill data-download -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/data-download, .gemini/skills/data-download, .github/skills/data-download and .opencode/skills/data-download in your project.

What does Data Download need to run?

Going by SKILL.md and its folder, Data Download needs the command-line tools its instructions call (curl and wget) and credentials named API_KEY. Our summary lists: Python 3; A credential in API_KEY. Its frontmatter pre-approves these tools: Bash(*), Read, Write, Edit, Grep, Glob, WebSearch, WebFetch, Agent.

Does Data Download access the network?

SKILL.md names 6 domains. In commands or code: api.census.gov, api.open-meteo.com, raw.githubusercontent.com, planetarycomputer.microsoft.com, overpass-api.de and www2.census.gov; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Data Download safe to install?

Our automated static check of SKILL.md flagged 1 warning(s): tells the agent its actions are pre-authorized / not to stop for confirmation. Read the flagged lines before installing; the check is not a guarantee either way.

What licence does Data Download use?

No licence was found for Data Download or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Data Download use?

About 8.1k tokens (SKILL.md is roughly 32k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Data Download?

Skills that share tags, products or a category with Data Download: Hypothesis Generation (spacering-net/codeg, 3.9k stars), Nature Paper Card (Yuan1z0825/nature-skills, 47k stars), Hypothesis Generation (K-Dense-AI/claude-scientific-writer, 2.4k stars) and Good Question (Rimagination/good-question, 305 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Data Download?

GRIND-Lab-Core (a GitHub organization) maintains it in GRIND-Lab-Core/night_owl_research_agent, which has 106 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on May 6, 2026.

Source: GRIND-Lab-Core/night_owl_research_agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.