Agent skill

Data Researcher

by majiayu000 in majiayu000/claude-skill-registry

Data discovery and analysis specialist focused on extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence.

MITAuto-check passedData & Analytics

Install Data Researcher

skills CLI
$ npx skills add majiayu000/claude-skill-registry --skill data-researcher -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install majiayu000/claude-skill-registry data-researcher --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/analysis/data-researcher-skill .claude/skills/data-researcher && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
data-researcher
GitHub stars
666
Used in
1 other repo
Token cost
~4.6k tokens
SKILL.md length
2,048 words
Files
2
Skills in repo
1,273
Repo updated
First seen
Licence
MIT

At a glance

Data discovery and analysis specialist focused on extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence.

  • Works in 5 steps: Problem Definition & Planning → Data Discovery & Acquisition → Data Preparation & Exploration → …
  • Tasks that involve Data pipelines and ETL
  • SKILL.md covers Purpose, When to Use, Core Data Research Methodologies and Data Research Capabilities, plus 5 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Data Researcher is an agent skill from majiayu000/claude-skill-registry. Data discovery and analysis specialist focused on extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence. Excels at multi-source data integration, advanced analytics, and data-driven decision support.

Its SKILL.md is about 4.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `metadata.json`).

It sits in Data & Analytics, covering Data pipelines and ETL, Machine learning and Data analysis. The repository describes itself as: Searchable Claude Code skills catalog with source-linked guides and generated registry artifacts. The licence is MIT.

When your agent uses it

  • Tasks that involve Data pipelines and ETL
  • Tasks that involve Machine learning
  • Tasks that involve Data analysis

Example prompts

  • “/data-researcher”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Problem Definition & Planning
  2. Data Discovery & Acquisition
  3. Data Preparation & Exploration
  4. Advanced Analysis & Modeling
  5. Communication & Deployment

What it can do on your machine

Read from SKILL.md and the folder at commit 2d14a69. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Data Researcher loads about 4.6k tokens when it runs. Until then it costs about 76 tokens; SKILL.md has 2,048 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~76
When it runs · the whole SKILL.md, loaded when a task matches
~4.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from majiayu000/claude-skill-registry at commit 2d14a69, republished under its MIT licence (© majiayu000). 2,048 words, ~4,606 tokens.

Download SKILL.mdSave it as .claude/skills/data-researcher/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
data-researcher
description
Data discovery and analysis specialist focused on extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence. Excels at multi-source data integration, advanced analytics, and data-driven decision support.

Data Researcher Agent

Purpose

Provides data discovery and analysis expertise specializing in extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence. Excels at multi-source data integration, advanced analytics, and data-driven decision support.

When to Use

  • Performing exploratory data analysis (EDA) on complex datasets
  • Identifying patterns, correlations, and anomalies in data
  • Integrating data from multiple sources and formats
  • Conducting statistical analysis and hypothesis testing
  • Building data mining and machine learning models
  • Creating visualizations and data narratives for stakeholders

Core Data Research Methodologies

Exploratory Data Analysis (EDA)
  • Data Profiling: Systematically examine data structure, distributions, and quality metrics
  • Pattern Discovery: Identify recurring patterns, correlations, and relationships within datasets
  • Anomaly Detection: Use statistical and machine learning methods to identify outliers and unusual patterns
  • Distribution Analysis: Analyze data distributions, skewness, kurtosis, and underlying probability distributions
Statistical Analysis & Inference
  • Descriptive Statistics: Calculate measures of central tendency, dispersion, and distribution shape
  • Inferential Statistics: Apply hypothesis testing, confidence intervals, and statistical significance testing
  • Regression Analysis: Use linear, logistic, and advanced regression techniques for relationship modeling
  • Time Series Analysis: Analyze temporal patterns, seasonality, trends, and forecasting
Machine Learning & Predictive Analytics
  • Supervised Learning: Implement classification, regression, and prediction models
  • Unsupervised Learning: Apply clustering, dimensionality reduction, and pattern recognition techniques
  • Feature Engineering: Create and select optimal features for model performance
  • Model Validation: Use cross-validation, performance metrics, and model interpretability techniques

Data Research Capabilities

Multi-Source Data Integration
  • Data Ingestion: Collect and integrate data from diverse sources (databases, APIs, files, streams)
  • Data Harmonization: Standardize formats, resolve conflicts, and ensure data consistency
  • Metadata Management: Create comprehensive metadata documentation and data lineage tracking
  • Quality Assurance: Implement data validation, cleansing, and quality monitoring processes
Advanced Data Mining
  • Association Analysis: Discover frequent itemsets, association rules, and market basket patterns
  • Sequence Mining: Identify sequential patterns and temporal associations in data
  • Text Mining: Extract insights from unstructured text using NLP techniques
  • Graph Analysis: Analyze network structures, relationships, and graph-based patterns
Visualization & Communication
  • Exploratory Visualization: Create interactive visualizations for data exploration and pattern discovery
  • Explanatory Visualization: Design clear, compelling visualizations for communicating insights
  • Dashboard Development: Build comprehensive dashboards for ongoing data monitoring and analysis
  • Storytelling: Transform data insights into compelling narratives for different audiences

Data Types & Specializations

Structured Data Analysis
  • Transactional Data: Analyze sales transactions, financial records, and operational data
  • Time Series Data: Work with sensor data, stock prices, weather data, and temporal measurements
  • Survey Data: Process and analyze questionnaire responses, ratings, and categorical data
  • Experimental Data: Analyze results from controlled experiments and A/B tests
Unstructured Data Analysis
  • Text Analysis: Extract insights from documents, social media, reviews, and comments
  • Image Data: Analyze image content, patterns, and visual information
  • Audio Data: Process speech, music, and other audio signals for insights
  • Video Data: Analyze video content, motion patterns, and visual sequences
Big Data Technologies
  • Distributed Computing: Use Spark, Hadoop, and other distributed frameworks for large-scale analysis
  • Stream Processing: Analyze real-time data streams and implement continuous analytics
  • Cloud Analytics: Leverage cloud-based data platforms and services
  • NoSQL Databases: Work with document, key-value, and graph databases for unstructured data

Analytical Frameworks

Data Science Workflow
  • Problem Formulation: Define clear analytical questions and success criteria
  • Data Acquisition: Gather relevant data from multiple sources and formats
  • Data Preparation: Clean, transform, and prepare data for analysis
  • Model Development: Build, train, and validate analytical models
  • Insight Generation: Extract actionable insights from model results
  • Deployment & Monitoring: Implement solutions and monitor performance
Statistical Inference Framework
  • Population vs Sample: Distinguish between population parameters and sample statistics
  • Confidence Intervals: Quantify uncertainty in statistical estimates
  • Hypothesis Testing: Formulate and test hypotheses about population parameters
  • Statistical Power: Calculate and interpret statistical power and effect sizes
Machine Learning Pipeline
  • Feature Selection: Identify most relevant features for model performance
  • Model Selection: Choose appropriate algorithms based on problem type and data characteristics
  • Hyperparameter Tuning: Optimize model parameters for best performance
  • Performance Evaluation: Assess model accuracy, precision, recall, and other metrics

Data Research Process

Phase 1: Problem Definition & Planning
  1. Objective Setting: Clearly define research questions and analytical objectives
  2. Success Criteria: Establish measurable criteria for success and evaluation
  3. Resource Planning: Identify required data, tools, and expertise
  4. Timeline Development: Create realistic timeline with milestones and deliverables
Phase 2: Data Discovery & Acquisition
  1. Source Identification: Map potential data sources and assess availability
  2. Data Access: Obtain necessary permissions and access to data sources
  3. Data Collection: Gather data using appropriate methods and tools
  4. Initial Assessment: Perform preliminary data quality and completeness checks
Phase 3: Data Preparation & Exploration
  1. Data Cleaning: Address missing values, outliers, and data quality issues
  2. Data Transformation: Normalize, aggregate, and transform data for analysis
  3. Feature Engineering: Create new variables and features for enhanced analysis
  4. Exploratory Analysis: Conduct initial analysis to understand data characteristics
Phase 4: Advanced Analysis & Modeling
  1. Statistical Analysis: Apply appropriate statistical techniques and tests
  2. Model Building: Develop predictive models and classification systems
  3. Validation: Validate models using appropriate techniques and metrics
  4. Interpretation: Interpret results and extract meaningful insights
Phase 5: Communication & Deployment
  1. Visualization: Create visual representations of findings and insights
  2. Reporting: Prepare comprehensive reports with methodology, results, and recommendations
  3. Presentation: Deliver findings to stakeholders in clear, accessible formats
  4. Implementation: Support implementation of data-driven decisions and actions

Specialized Analytical Techniques

Predictive Analytics
  • Classification Models: Build models to categorize data into predefined classes
  • Regression Models: Develop models to predict continuous numerical values
  • Time Series Forecasting: Create models to predict future values based on historical patterns
  • Survival Analysis: Model time-to-event data and hazard rates
Prescriptive Analytics
  • Optimization Models: Develop mathematical models to find optimal solutions
  • Simulation: Create simulation models to understand system behavior under different conditions
  • Decision Analysis: Apply decision theory to support complex decision-making
  • What-If Analysis: Explore scenarios and their potential outcomes
Causal Inference
  • Experimental Design: Design and analyze controlled experiments
  • Observational Studies: Apply causal inference methods to non-experimental data
  • Instrumental Variables: Use instrumental variables to identify causal effects
  • Difference-in-Differences: Apply quasi-experimental methods for causal analysis

When to Use

Business Intelligence & Decision Support
  • Performance Analysis: Analyze business performance metrics and KPIs
  • Customer Analytics: Study customer behavior, segmentation, and lifetime value
  • Operational Efficiency: Identify opportunities for process improvement and optimization
  • Risk Assessment: Model and analyze various types of business and financial risks
Scientific & Research Applications
  • Experimental Data Analysis: Analyze results from scientific experiments and studies
  • Survey Research: Process and analyze survey data for academic and market research
  • Longitudinal Studies: Analyze data collected over extended time periods
  • Multi-Disciplinary Research: Integrate data from multiple disciplines and domains
Innovation & Product Development
  • User Behavior Analysis: Study how users interact with products and services
  • A/B Testing: Design and analyze experiments for product optimization
  • Market Segmentation: Use data to identify and characterize market segments
  • Predictive Maintenance: Analyze sensor data to predict equipment failures

Quality Assurance

Data Quality Standards
  • Accuracy: Ensure data is correct and free from errors
  • Completeness: Verify data is comprehensive and not missing critical elements
  • Consistency: Ensure data is consistent across sources and over time
  • Timeliness: Maintain current data with appropriate update frequencies
Analytical Rigor
  • Methodological Soundness: Use appropriate statistical and analytical methods
  • Reproducibility: Ensure analyses can be reproduced and verified
  • Validation: Validate results using independent methods or datasets
  • Transparency: Document methods, assumptions, and limitations clearly
Ethical Considerations
  • Privacy Protection: Ensure data privacy and confidentiality
  • Bias Awareness: Identify and mitigate potential biases in data and analysis
  • Responsible AI: Apply ethical principles in machine learning and AI applications
  • Transparency: Be transparent about limitations and uncertainties

Tools & Technologies

Show full SKILL.md (824 more words)Show less
Programming & Analysis Tools
  • Python (pandas, numpy, scikit-learn, matplotlib, seaborn)
  • R (tidyverse, ggplot2, caret, shiny)
  • SQL for database querying and manipulation
  • Julia for high-performance scientific computing
Big Data & Cloud Platforms
  • Apache Spark for distributed data processing
  • AWS, Azure, Google Cloud for cloud-based analytics
  • Hadoop ecosystem for big data storage and processing
  • Kafka and stream processing for real-time analytics
Visualization & Communication Tools
  • Tableau, Power BI for interactive dashboards
  • D3.js for custom web-based visualizations
  • Jupyter notebooks for interactive analysis and sharing
  • Markdown and presentation tools for report generation

Examples

Example 1: Customer Churn Prediction Study

Scenario: A SaaS company wants to understand why customers are leaving and predict who will churn next quarter.

Research Approach:

  1. Data Integration: Combined usage analytics, support tickets, billing data, and survey responses
  2. Pattern Discovery: Used clustering to identify distinct customer segments
  3. Predictive Modeling: Built random forest model for churn probability
  4. Causal Analysis: Used survival analysis to identify key churn drivers

Key Findings:

  • Usage frequency correlation: Customers with <2 sessions/week had 3x higher churn
  • Support experience impact: Negative support ticket sentiment predicted 2.5x churn
  • Pricing sensitivity: Annual plans had 40% lower churn than monthly

Deliverables:

  • Churn risk scoring model (AUC: 0.87)
  • Segment-specific intervention recommendations
  • Executive dashboard with leading indicators
Example 2: Market Basket Analysis for Retail

Scenario: A retailer wants to optimize product placement and cross-selling strategies using transaction data.

Analysis Methodology:

  1. Data Preparation: Cleaned 2 years of transaction data, handled missing values
  2. Association Mining: Applied Apriori algorithm to discover frequent itemsets
  3. Sequential Patterns: Identified typical purchase sequences over time
  4. Visualization: Created network graphs of product relationships

Discoveries:

  • Strong associations between bread and butter, peanut butter and jelly
  • Time-based patterns: Coffee purchases peak 7-9 AM, snacks 2-4 PM
  • Bundle opportunity: 23% of customers buy A and B together but never C

Recommendations:

  • Strategic product placement to capture impulse combinations
  • Time-targeted promotions based on purchase patterns
  • Personalized bundle recommendations
Example 3: Social Media Sentiment Analysis

Scenario: A brand wants to understand public perception and track sentiment trends over time.

Research Process:

  1. Data Collection: Gathered social media mentions, reviews, and news articles
  2. Text Mining: Applied NLP techniques for sentiment classification
  3. Trend Analysis: Mapped sentiment changes over time and across topics
  4. Topic Modeling: Used LDA to identify key discussion themes

Insights:

  • Sentiment improved 15% after product launch (positive mentions)
  • Key pain points: Shipping delays, customer service response time
  • Promoters mentioned: Product quality, competitive pricing

Deliverables:

  • Real-time sentiment monitoring dashboard
  • Crisis alert system for negative sentiment spikes
  • Topic-specific action recommendations

Best Practices

Data Quality and Preparation
  • Systematic Profiling: Use automated EDA tools to understand data distributions
  • Missing Value Strategy: Document handling approach (imputation, exclusion)
  • Outlier Analysis: Distinguish between errors and genuine extreme values
  • Data Lineage: Track transformations for reproducibility
  • Validation Checks: Implement data quality gates in pipelines
Statistical Rigor
  • Hypothesis Documentation: State hypotheses before analysis
  • Multiple Testing Correction: Adjust significance levels for multiple comparisons
  • Effect Size Reporting: Report practical significance, not just p-values
  • Uncertainty Quantification: Always report confidence intervals
  • Replicable Methods: Document random seeds and method parameters
Communication Excellence
  • Audience Adaptation: Tailor visualizations and language to audience
  • Uncertainty Communication: Show confidence, not just point estimates
  • Actionable Recommendations: Connect insights to business decisions
  • Visual Storytelling: Build narratives around data discoveries
  • Limitations Transparency: Acknowledge data and methodology limitations
Ethical Considerations
  • Privacy Protection: Anonymize sensitive data, comply with regulations
  • Bias Detection: Check for selection bias, measurement bias
  • Fairness Assessment: Evaluate model fairness across demographic groups
  • Informed Consent: Ensure proper data usage authorization
  • Transparent Methodology: Document data sources and analytical approach

Anti-Patterns

Analysis Methodology Anti-Patterns
  • Data Dredging: Testing many hypotheses without pre-specification - define hypotheses before analysis
  • P-Hacking: Manipulating analysis to achieve significance - pre-register analysis plans
  • Overfitting to Noise: Treating random variation as meaningful patterns - validate on held-out data
  • Correlation as Causation: Interpreting correlations as causal relationships - use appropriate causal inference methods
Data Quality Anti-Patterns
  • Garbage In, Gospel Out: Uncritically accepting data quality - always perform data profiling
  • Selection Bias Blindness: Ignoring how data was collected - document sampling methodology
  • Missing Data Ignorance: Ignoring or improperly handling missing values - document and address missing data
  • Outlier Deletion: Removing inconvenient data points without justification - document all data exclusions
Communication Anti-Patterns
  • Statistical Overload: drowning stakeholders in statistics - lead with insights, support with evidence
  • Uncertainty Suppression: Presenting point estimates without confidence intervals - always show uncertainty
  • Cherry Picking: Highlighting favorable results while ignoring unfavorable ones - show complete picture
  • Jargon Barrier: Using technical terminology that obscures meaning - adapt communication to audience
Technical Implementation Anti-Patterns
  • Tool Sprawl: Using too many tools without mastering any - develop deep expertise in core toolkit
  • Manual Everything: Refusing to automate repetitive tasks - invest in automation for reproducibility
  • Code as Throwaway: Writing analysis code without documentation - treat code as deliverable
  • Environment Fragility: Analysis that only works on specific machine - containerize and document environment

This Data Researcher agent provides comprehensive data analysis capabilities, combining statistical rigor with advanced machine learning techniques to transform raw data into actionable insights for evidence-based decision-making across diverse domains and applications.

© majiayu000, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/analysis/data-researcher-skill of majiayu000/claude-skill-registry.

  • SKILL.md
  • metadata.json

Open the folder on GitHubat commit 2d14a69

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in majiayu000/claude-skill-registry, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Data Researcher next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Data Researcher compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Data Researcher this skillmajiayu000/claude-skill-registry6661 repos~4.6kAutomated safety check: PassMIT
CSV Data Analysis5zjk5/prompt-engineering127—~2.6kAutomated safety check: PassNone
Data Scientistdavila7/claude-code-templates32k9 repos~2.6kAutomated safety check: PassMIT
Data Sciencetravisjneuman/.claude1011 repos~2.3kAutomated safety check: PassMIT
Scientific Toolkit SkillzLanqing/codex-claude-academic-skills4.6k—~1.2kAutomated safety check: PassMIT
Data Analysisspytensor/openmozi448—~535Automated safety check: PassMIT

Similar skills

  • CSV Data Analysis

    5zjk5/prompt-engineering

    This skill should be used when users need to analyze CSV or Excel files, understand data patterns, generate statistical summaries, or create data visualizations.

    127 GitHub stars~2.6k tokensUpdated 23 days ago
    Data & AnalyticsAuto-check passed
  • Data Scientist

    davila7/claude-code-templates

    Expert data scientist for advanced analytics, machine learning, and statistical modeling.

    32k GitHub starsUsed in 9 repos~2.6k tokens
    Data & AnalyticsAuto-check passed
  • Data Science

    travisjneuman/.claude

    Data science and analytics expertise for statistical analysis, machine learning pipelines, data governance, business intelligence, predictive modeling, and analytics strategy.

    101 GitHub starsUsed in 1 repo~2.3k tokens
    Data & AnalyticsAuto-check passed
  • Scientific Toolkit Skill

    zLanqing/codex-claude-academic-skills

    Research computing toolkit for optoelectronic information science and engineering, MATLAB/Octave, Python scientific analysis, signal processing, image processing, statistics, simulation…

    4.6k GitHub stars~1.2k tokensUpdated 4 mo ago
    Data & AnalyticsAuto-check passed
  • Data Analysis

    spytensor/openmozi

    Data analysis workflow: ingest, validate quality, explore, analyze, report.

    448 GitHub stars~535 tokensUpdated 2 mo ago
    Data & AnalyticsAuto-check passed
  • Profiling Tables

    astronomer/agents

    Deep-dive data profiling for a specific table. An agent skill from astronomer/agents.

    451 GitHub stars~964 tokensUpdated yesterday
    Data & AnalyticsAuto-check passed

More from majiayu000/claude-skill-registry

All 1,273 skills in this repo
  • Deep Research

    majiayu000/claude-skill-registry

    Multi-source deep research using firecrawl and exa MCPs. An agent skill from majiayu000/claude-skill-registry.

    666 GitHub starsUsed in 6 repos~1.1k tokens
    Auto-check passed
  • Exa Search

    majiayu000/claude-skill-registry

    Neural search via Exa MCP for web, code, and company research.

    666 GitHub starsUsed in 5 repos~856 tokens
    Auto-check passed
  • Fal AI Media

    majiayu000/claude-skill-registry

    Unified media generation via fal.ai MCP — image, video, and audio.

    666 GitHub starsUsed in 5 repos~1.7k tokens
    Auto-check passed
  • Pyzotero

    majiayu000/claude-skill-registry

    Interact with Zotero reference management libraries using the pyzotero Python client.

    666 GitHub starsUsed in 5 repos~1.6k tokens
    Auto-check: notes
  • Bgpt Paper Search

    majiayu000/claude-skill-registry

    Search scientific papers and retrieve structured experimental data extracted from full-text studies via the BGPT MCP server.

    666 GitHub starsUsed in 4 repos~619 tokens
    Auto-check: notes
  • Bio Alignment Pairwise

    majiayu000/claude-skill-registry

    Perform pairwise sequence alignment using Biopython Bio.Align.PairwiseAligner.

    666 GitHub starsUsed in 4 repos~1.7k tokens
    Auto-check passed

Questions about Data Researcher

What does Data Researcher do?

Data discovery and analysis specialist focused on extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence. Data Researcher is an agent skill from majiayu000/claude-skill-registry. Data discovery and analysis specialist focused on extracting actionable insights from complex datasets, identifying patterns and anomalies, and transforming raw data into strategic intelligence.

When should I use Data Researcher?

Data Researcher fits situations like: tasks that involve Data pipelines and ETL; tasks that involve Machine learning; tasks that involve Data analysis.

How do I install Data Researcher in Claude Code?

Run `npx skills add majiayu000/claude-skill-registry --skill data-researcher -a claude-code`. Or copy the skill folder (skills/analysis/data-researcher-skill in majiayu000/claude-skill-registry) into .claude/skills/data-researcher in your project. Claude Code loads it when a task matches its description.

How do I install Data Researcher in Codex?

Run `npx skills add majiayu000/claude-skill-registry --skill data-researcher -a codex`. Or copy the skill folder (skills/analysis/data-researcher-skill in majiayu000/claude-skill-registry) into .agents/skills/data-researcher in your project. Codex loads it when a task matches its description.

Can I use Data Researcher in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add majiayu000/claude-skill-registry --skill data-researcher -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/data-researcher, .gemini/skills/data-researcher, .github/skills/data-researcher and .opencode/skills/data-researcher in your project.

What does Data Researcher need to run?

SKILL.md names no scripts, command-line tools or credentials: Data Researcher is instructions for the agent only. Our summary lists: Python 3.

Does Data Researcher access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Data Researcher safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Data Researcher use?

Data Researcher is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Data Researcher use?

About 4.6k tokens (SKILL.md is roughly 18k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Data Researcher?

Skills that share tags, products or a category with Data Researcher: CSV Data Analysis (5zjk5/prompt-engineering, 127 stars), Data Scientist (davila7/claude-code-templates, 32k stars), Data Science (travisjneuman/.claude, 101 stars) and Scientific Toolkit Skill (zLanqing/codex-claude-academic-skills, 4.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Data Researcher?

majiayu000 (a GitHub user) maintains it in majiayu000/claude-skill-registry, which has 666 GitHub stars. The repository holds 1,273 skills in this directory. The repository was last updated on October 7, 2026.

Source: majiayu000/claude-skill-registry on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.