Agent skill

Lightgbm Analysis

by aipoch in aipoch/medical-research-skills

A skill your agent uses when training a LightGBM model on tabular data in R and returning model metrics, feature importance ranking tables, and feature importance plots.

MITAuto-check passedData & Analytics

Install Lightgbm Analysis

skills CLI
$ npx skills add aipoch/medical-research-skills --skill lightgbm-analysis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install aipoch/medical-research-skills lightgbm-analysis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/aipoch/medical-research-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/'awesome-med-research-skills/Data Analysis/LightGBM-analysis' .claude/skills/lightgbm-analysis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
lightgbm-analysis
GitHub stars
2k
Token cost
~3.6k tokens
SKILL.md length
1,351 words
Files
12 (incl. scripts, references)
Skills in repo
567
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when training a LightGBM model on tabular data in R and returning model metrics, feature importance ranking tables, and feature importance plots.

  • Works in 7 steps: Confirm the input file exists and the… → Remove identifier or sensitive columns… → Set --drop_cols and optionally… → …
  • Training a LightGBM model on tabular data in R and returning model metrics
  • SKILL.md covers Use This Skill When, Primary Command, Prerequisites and Core Arguments, plus 14 more sections
  • Runs R scripts from its folder; reaches cloud.r-project.org

What it does

Lightgbm Analysis is an agent skill from aipoch/medical-research-skills. Use when training a LightGBM model on tabular data in R and returning model metrics, feature importance ranking tables, and feature importance plots.

Its SKILL.md is about 3.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 15 other files, including scripts and reference files (for example `eval_report_lightgbm-analysis_result.json`, `references/algorithm.md` and `references/cli-guide.md`).

It sits in Data & Analytics, covering Machine learning. The repository describes itself as: Hundreds of agent skills for medical research, including protocol design, data analysis, evidence insights, and academic writing. The licence is MIT.

When your agent uses it

  • Training a LightGBM model on tabular data in R and returning model metrics
  • Feature importance ranking tables
  • Feature importance plots

Example prompts

  • “/lightgbm-analysis”

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. Confirm the input file exists and the target column name is correct.
  2. Remove identifier or sensitive columns such as id, sample_id, patient_id, accession numbers, or the bundled sample identifier column V1…
  3. Set --drop_cols and optionally --feature_cols so the model only sees intended predictors.
  4. If you need overwrite protection, add --fail_if_output_exists or choose a fresh --output_dir.
  5. Run scripts/main.R.
  6. Check table/ for the importance table, model metrics, and remediation guidance.
  7. Check figure/ for the feature importance ranking plot and data/ for the run summary.

What it can do on your machine

Read from SKILL.md and the folder at commit 686e09d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 4 files in scripts/ (R), which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • cloud.r-project.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Lightgbm Analysis loads about 3.6k tokens when it runs, and up to ~7.9k if it reads all its reference files. Until then it costs about 42 tokens; SKILL.md has 1,351 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~42
When it runs · the whole SKILL.md, loaded when a task matches
~3.6k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~7.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from aipoch/medical-research-skills at commit 686e09d, republished under its MIT licence (© aipoch). 1,351 words, ~3,571 tokens.

Download SKILL.mdSave it as .claude/skills/lightgbm-analysis/SKILL.md (or your agent's skills folder). This skill also uses 11 other files; get the full folder from GitHub.
name
lightgbm-analysis
description
Use when training a LightGBM model on tabular data in R and returning model metrics, feature importance ranking tables, and feature importance plots.
license
MIT
author
AIPOCH

Source: https://github.com/aipoch/medical-research-skills

LightGBM Analysis

Use this skill to build a LightGBM model on tabular data and export feature importance ranking results as both a table and a figure.

Use This Skill When

  • You need a command-line LightGBM workflow written in R.
  • You need classification or regression on structured tabular data.
  • You need ranked feature importance outputs for reporting or interpretation.
  • You need standardized outputs under table/, figure/, and data/.

Primary Command

bash
Rscript scripts/main.R \
  --data_file <input_file> \
  --target_var <target_column> \
  --output_dir <output_dir>

Prerequisites

  • Rscript is available in the shell.
  • Required R packages: optparse, data.table, lightgbm.
  • Install basic dependencies with Rscript -e 'install.packages(c("optparse", "data.table"), repos="https://cloud.r-project.org")'.
  • Install the R lightgbm package from the LightGBM project because it is usually not available from CRAN.

Core Arguments

ArgumentRequiredDescription
--data_fileYesInput data file in CSV format or tab-delimited TXT/TSV format
--target_varYesTarget column used for modeling
--output_dirNoOutput directory, default ./LightGBM_Results
--fail_if_output_existsNoStop instead of overwriting when output_dir already contains files
--task_typeNoauto, regression, binary, or multiclass. Default auto
--feature_colsNoComma-separated feature columns. Default uses all columns except target and dropped columns
--drop_colsNoComma-separated columns to exclude before modeling
--importance_typeNogain or split. Default gain
--top_nNoNumber of features to show in the importance plot. Default 20
--output_formatNocsv or txt table export. Default csv

Modeling Arguments

ArgumentDefaultDescription
--metricautoEvaluation metric matched to task type
--test_size0.2Test-set proportion
--valid_size0.2Validation proportion taken from the training partition
--nrounds500Maximum boosting rounds
--learning_rate0.05Shrinkage rate
--num_leaves31Maximum leaf count per tree
--max_depth-1Maximum tree depth, -1 means no explicit limit
--min_data_in_leaf5Minimum samples per leaf
--feature_fraction0.8Column sampling ratio
--bagging_fraction0.8Row sampling ratio
--bagging_freq1Bagging frequency
--lambda_l10L1 regularization
--lambda_l20L2 regularization
--early_stopping_rounds50Early stopping patience
--seed42Random seed

Input Requirements

  • The input file must include the target column.
  • Prefer .csv or .tsv inputs. .txt files must be tab-delimited.
  • The skill expects at least 20 rows after removing missing target values.
  • Features may be numeric, integer, logical, character, or factor-like text.
  • Character features are label-encoded internally for LightGBM.
  • Missing target values are removed before modeling.
  • Missing feature values are left for LightGBM to handle.
  • If task_type=auto, the script infers regression or classification from the target values.

Bundled test data examples:

csv
V1,fustat,CAMK2N2,GGT6,GPR161,RAB26,RIBC2
TCGA-C5-A1M5,1,2.248291938,5.274690305,2.825215762,3.121114894,5.35318565
TCGA-EA-A5O9,0,3.346176843,5.404368414,2.604616977,0.629473197,4.429314674
TCGA-C5-A3HL,0,3.363100974,5.363314779,4.124799581,4.127228806,4.916596068

Minimal Workflow

  1. Confirm the input file exists and the target column name is correct.
  2. Remove identifier or sensitive columns such as id, sample_id, patient_id, accession numbers, or the bundled sample identifier column V1 before training.
  3. Set --drop_cols and optionally --feature_cols so the model only sees intended predictors.
  4. If you need overwrite protection, add --fail_if_output_exists or choose a fresh --output_dir.
  5. Run scripts/main.R.
  6. Check table/ for the importance table, model metrics, and remediation guidance.
  7. Check figure/ for the feature importance ranking plot and data/ for the run summary.

Avoid ambiguous text exports. If a .txt file is parsed as one column, re-export it as tab-delimited text or CSV before rerunning.

For quick validation in small audit environments, prefer the bundled dt_sample3.txt smoke test shown below with reduced --nrounds and --early_stopping_rounds. The full binary example on dt_sample1.csv is still useful as a complete workflow example, but it can exceed short runtime budgets.

If you omit --data_file or --target_var, the script exits with SKILL_MISSING_INPUT.

Outputs

Expected output structure:

text
<output_dir>/
├── table/
├── figure/
└── data/

Primary result files:

  • table/lightgbm_feature_importance.<output_format>
  • table/lightgbm_model_metrics.<output_format>
  • table/lightgbm_remediation.<output_format>
  • figure/lightgbm_feature_importance_<importance_type>.pdf
  • data/lightgbm_run_summary.txt
  • data/lightgbm_categorical_levels.txt when categorical or character predictors were encoded

Feature importance table fields include:

  • feature
  • gain
  • split
  • cover
  • importance_type
  • importance_value
  • rank
  • gain_share
  • split_share

Model metrics include:

  • task_type
  • metric_primary
  • best_iteration
  • train_rows
  • valid_rows
  • test_rows
  • prediction_collapse_flag
  • model_quality_flag
  • interpretation_status
  • primary_issue
  • model_quality_issues
  • rerun_hint
  • model_quality_note
  • task-specific evaluation metrics such as rmse, mae, accuracy, auc, or logloss

Remediation table fields include:

  • task_type
  • model_quality_flag
  • interpretation_status
  • issue_code
  • issue_detail
  • recommended_action
  • suggested_rerun_change

Run summary file includes the task type, best iteration, primary quality fields, top features, and artifact paths for the completed run.

Overwrite Behavior

  • Rerunning into an existing output_dir replaces prior result files with the new metrics, importance table, remediation table, figure, and session metadata.
  • Set --fail_if_output_exists when you want the run to stop instead of replacing prior artifacts.
  • If you need an audit trail, prefer a timestamped or per-run output_dir.
  • The script now warns when output_dir already contains files.

Success And Failure Contract

Success:

  • Console output should end with LightGBM analysis completed successfully.
  • table/lightgbm_model_metrics.<output_format> and table/lightgbm_feature_importance.<output_format> should exist.
  • table/lightgbm_remediation.<output_format> and data/lightgbm_run_summary.txt should exist.
  • figure/lightgbm_feature_importance_<importance_type>.pdf should exist.
  • The importance table should contain at least one non-zero gain or split value.

Failure or caution:

  • If parsing fails, expect a SKILL_* message instead of a raw stack trace.
  • If best_iteration <= 1, predictions collapse to one class, recall is 0, f1 is NA, or the selected importance values are mostly zero, do not treat the ranking as reliable.
  • Review model_quality_flag and model_quality_note in table/lightgbm_model_metrics.csv before interpreting the exported ranking.
  • Use interpretation_status to decide whether the run is report-ready: eligible means interpretation-ready, eligible_with_caveats means the ranking may still be usable with caveats, and caution_only means diagnostic-only.
  • Review table/lightgbm_remediation.csv and rerun_hint for the exact failure mode and recommended rerun changes.
  • Recheck delimiter choice, identifier leakage, and --min_data_in_leaf before trusting the outputs.
Show full SKILL.md (506 more words)Show less

Caution Remediation

  • best_iteration<=1: lower --min_data_in_leaf and verify that the selected predictors have usable signal.
  • single_predicted_class: review class balance and feature selection before using the ranking downstream.
  • recall=0 or no_positive_predictions: revisit --feature_cols and the target balance before treating the run as report-ready.
  • <importance_type>_importance_sparse: compare against the alternate importance type and review whether the retained predictors have enough signal.

Agent Response Contract

When this skill completes, the agent should report:

  • resolved task_type
  • best_iteration
  • primary evaluation metrics from table/lightgbm_model_metrics.<output_format>
  • top ranked features from table/lightgbm_feature_importance.<output_format>
  • model_quality_flag and interpretation_status
  • artifact paths for the metrics table, importance table, remediation table, figure, and run summary file

If model_quality_flag is not ok, the agent must explicitly say the run is diagnostic-only or caveat-limited and include the recommended rerun changes from rerun_hint or table/lightgbm_remediation.<output_format>.

Feature Importance Guidance

  • Use gain when you care about overall contribution to loss reduction.
  • Use split when you care about how often a feature is used in tree splits.
  • Prefer gain for most ranking summaries and reports.
  • Low importance does not imply no business value, especially under correlated features.

Read These Files When Needed

NeedFile
LightGBM method details and importance interpretationreferences/algorithm.md
CLI examplesreferences/cli-guide.md
Error diagnosisreferences/troubleshooting.md
Main entry pointscripts/main.R
Sample test datatests/data/

Quick Examples

Fast smoke test with dt_sample3.txt:

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample3.txt \
  --target_var Group \
  --drop_cols V1 \
  --task_type binary \
  --nrounds 80 \
  --early_stopping_rounds 20 \
  --top_n 15 \
  --output_dir tests/output_smoke_txt

Audit-friendly binary preset for short runtime budgets:

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample1.csv \
  --target_var fustat \
  --drop_cols V1 \
  --task_type binary \
  --nrounds 120 \
  --early_stopping_rounds 20 \
  --output_dir tests/output_binary_fast

Full binary workflow example with dt_sample1.csv:

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample1.csv \
  --target_var fustat \
  --drop_cols V1 \
  --task_type binary \
  --output_dir tests/output_binary

Split-based importance export example with dt_sample2.csv:

Use this to verify split-based ranking output. Review model_quality_flag and interpretation_status before treating the bundled example as report-ready because this path can remain diagnostic-only on small test splits.

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample2.csv \
  --target_var fustat \
  --feature_cols CAMK2N2,GGT6,GPR161,RAB26,RIBC2 \
  --drop_cols V1 \
  --task_type binary \
  --importance_type split \
  --output_dir tests/output_binary_split

Audit-friendly regression preset with dt_sample1.csv and RIBC2 as the target:

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample1.csv \
  --target_var RIBC2 \
  --drop_cols V1 \
  --task_type regression \
  --nrounds 120 \
  --early_stopping_rounds 20 \
  --output_dir tests/output_regression_fast

Full regression workflow with dt_sample1.csv and RIBC2 as the target:

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample1.csv \
  --target_var RIBC2 \
  --drop_cols V1 \
  --task_type regression \
  --output_dir tests/output_regression

Tab-delimited TXT input with automatic binary target encoding from Group:

bash
Rscript scripts/main.R \
  --data_file tests/data/dt_sample3.txt \
  --target_var Group \
  --drop_cols V1 \
  --task_type binary \
  --top_n 15 \
  --output_dir tests/output_group_txt

Validation

bash
Rscript scripts/main.R --help

Use the smoke test under ## Quick Examples for a fast validation pass. After a successful run, verify that these files exist under the selected output_dir:

  • table/lightgbm_feature_importance.csv
  • table/lightgbm_model_metrics.csv
  • table/lightgbm_remediation.csv
  • figure/lightgbm_feature_importance_<importance_type>.pdf
  • data/lightgbm_run_summary.txt
  • data/lightgbm_categorical_levels.txt if categorical or character predictors were encoded

When Not To Use

  • The input file is an unstructured note, JSON blob, or free-text report.
  • The text file delimiter is unknown and you cannot inspect or re-export it.
  • The table still contains sample IDs, patient IDs, accession numbers, or similar identifiers that should not be model features.
  • The input still contains direct identifiers or sensitive fields that you have not reviewed and removed from modeling.

Common Errors

  • SKILL_FILE_NOT_FOUND: Input file path is wrong or inaccessible.
  • SKILL_MISSING_COLUMNS: The target or requested feature columns are missing.
  • SKILL_INVALID_DATA: Data types, target encoding, or row count are unsuitable for LightGBM.
  • SKILL_DEGENERATE_MODEL: Training finished but the exported importance table is all zero and should not be interpreted.
  • SKILL_INVALID_PARAMETER: An argument value is invalid.
  • SKILL_DEPENDENCY_MISSING: Required package such as lightgbm is unavailable.
  • SKILL_TRAINING_FAILED: LightGBM training failed.

Before sharing exported artifacts, verify that identifier-like columns such as V1, sample IDs, or patient IDs were excluded from modeling and from any published tables. If model_quality_flag is not ok, treat the run as a diagnostic result rather than an interpretable ranking.

If the issue is not obvious, read references/troubleshooting.md.

© aipoch, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 11 other files (scripts, references) in awesome-med-research-skills/Data Analysis/LightGBM-analysis of aipoch/medical-research-skills.

  • SKILL.md
  • eval_report_lightgbm-analysis_result.json
  • references/algorithm.md
  • references/cli-guide.md
  • references/troubleshooting.md
  • scripts/functions.R
  • scripts/main.R
  • scripts/run_analysis.R
  • scripts/utils.R
  • tests/data/dt_sample1.csv
  • tests/data/dt_sample2.csv
  • tests/data/dt_sample3.txt

Open the folder on GitHubat commit 686e09d

Compare with similar skills

Lightgbm Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Lightgbm Analysis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Lightgbm Analysis this skillaipoch/medical-research-skills2k—~3.6kAutomated safety check: PassMIT
Scikit LearnzLanqing/codex-claude-academic-skills4.6k17 repos~3.9kAutomated safety check: PassBSD-3-Clause
Senior Data ScientistRaidriar7170/hermes-skilleval1256 repos~1.4kAutomated safety check: PassMIT
Agentic Kaggle WorkflowFrankS-IntelLab/agentic-kaggle-skill188—~4kAutomated safety check: PassMIT
Retention Analysisliangdabiao/claude-data-analysis-ultra-main2901 repos~1.3kAutomated safety check: NotesNone
Geomlitalo-goncalves/geoML109—~4.2kAutomated safety check: PassGPL-3.0

Similar skills

  • Scikit Learn

    zLanqing/codex-claude-academic-skills

    Machine learning in Python with scikit-learn. An agent skill from zLanqing/codex-claude-academic-skills.

    4.6k GitHub starsUsed in 17 repos~3.9k tokens
    Data & AnalyticsAuto-check passed
  • Senior Data Scientist

    Raidriar7170/hermes-skilleval

    World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics.

    125 GitHub starsUsed in 6 repos~1.4k tokens
    Data & AnalyticsAuto-check passed
  • Agentic Kaggle Workflow

    FrankS-IntelLab/agentic-kaggle-skill

    Takes a Kaggle competition from rules and validation design through baselines, ensembling and notebook architecture to a scored submission.

    188 GitHub stars~4k tokensUpdated 3 mo ago
    Data & AnalyticsAuto-check passed
  • Retention Analysis

    liangdabiao/claude-data-analysis-ultra-main

    Analyze user retention and churn using survival analysis, cohort analysis, and machine learning.

    290 GitHub starsUsed in 1 repo~1.3k tokens
    Data & AnalyticsAuto-check: notes
  • Geoml

    italo-goncalves/geoML

    Working knowledge of the geoML Python package (github.com/italo-goncalves/geoML): variational Gaussian processes for spatial data, implicit geological modelling, block models, drillhole data…

    109 GitHub stars~4.2k tokensUpdated 5 days ago
    Data & AnalyticsAuto-check passed
  • Radiomics ML

    Aperivue/medsci-skills

    A skill your agent uses when building or auditing a radiomics or tabular clinical-ML prediction model with a classical learner (LASSO, SVM, random forest, XGBoost and similar).

    329 GitHub starsUsed in 1 repo~2.7k tokens
    Data & AnalyticsAuto-check passed

More from aipoch/medical-research-skills

All 567 skills in this repo
  • Academic Poster Generator

    aipoch/medical-research-skills

    Complete workflow for generating academic research posters from PDF literature; use when you need to extract paper content from PDFs and produce a LaTeX-based poster…

    2k GitHub stars~2.2k tokensUpdated 21 days ago
    Auto-check passed
  • Diagnostic Study Quality Assessment Quadas

    aipoch/medical-research-skills

    Analyzes clinical diagnostic accuracy studies for bias using the QUADAS-2 tool.

    2k GitHub stars~1.4k tokensUpdated 21 days ago
    Auto-check passed
  • Exploratory Data Analysis

    aipoch/medical-research-skills

    Perform comprehensive exploratory data analysis on scientific data files across 200+ file formats.

    2k GitHub stars~3.7k tokensUpdated 21 days ago
    Auto-check passed
  • Iso Certification

    aipoch/medical-research-skills

    A toolkit for preparing ISO 13485:2016 certification documentation for medical device QMS.

    2k GitHub stars~1.8k tokensUpdated 21 days ago
    Auto-check passed
  • Journal Skills

    aipoch/medical-research-skills

    Recommends target journals for manuscript submission by analyzing the paper topic/abstract and the journal distribution of similar PubMed literature; use when users ask for journal…

    2k GitHub stars~1.7k tokensUpdated 21 days ago
    Auto-check passed
  • Latex Posters

    aipoch/medical-research-skills

    Creates academic-poster writing packages for LaTeX using beamerposter, tikzposter, or baposter.

    2k GitHub stars~1.3k tokensUpdated 21 days ago
    Auto-check passed

Questions about Lightgbm Analysis

What does Lightgbm Analysis do?

A skill your agent uses when training a LightGBM model on tabular data in R and returning model metrics, feature importance ranking tables, and feature importance plots. Lightgbm Analysis is an agent skill from aipoch/medical-research-skills. Use when training a LightGBM model on tabular data in R and returning model metrics, feature importance ranking tables, and feature importance plots.

When should I use Lightgbm Analysis?

Lightgbm Analysis fits situations like: training a LightGBM model on tabular data in R and returning model metrics; feature importance ranking tables; feature importance plots.

How do I install Lightgbm Analysis in Claude Code?

Run `npx skills add aipoch/medical-research-skills --skill lightgbm-analysis -a claude-code`. Or copy the skill folder (awesome-med-research-skills/Data Analysis/LightGBM-analysis in aipoch/medical-research-skills) into .claude/skills/lightgbm-analysis in your project. Claude Code loads it when a task matches its description.

How do I install Lightgbm Analysis in Codex?

Run `npx skills add aipoch/medical-research-skills --skill lightgbm-analysis -a codex`. Or copy the skill folder (awesome-med-research-skills/Data Analysis/LightGBM-analysis in aipoch/medical-research-skills) into .agents/skills/lightgbm-analysis in your project. Codex loads it when a task matches its description.

Can I use Lightgbm Analysis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add aipoch/medical-research-skills --skill lightgbm-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/lightgbm-analysis, .gemini/skills/lightgbm-analysis, .github/skills/lightgbm-analysis and .opencode/skills/lightgbm-analysis in your project.

What does Lightgbm Analysis need to run?

Going by SKILL.md and its folder, Lightgbm Analysis needs R for the scripts in its folder.

Does Lightgbm Analysis access the network?

SKILL.md names 1 domain. In commands or code: cloud.r-project.org; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Lightgbm Analysis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Lightgbm Analysis use?

Lightgbm Analysis is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Lightgbm Analysis use?

About 3.6k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 4.4k tokens, read only when the agent opens those files.

What are the alternatives to Lightgbm Analysis?

Skills that share tags, products or a category with Lightgbm Analysis: Scikit Learn (zLanqing/codex-claude-academic-skills, 4.6k stars), Senior Data Scientist (Raidriar7170/hermes-skilleval, 125 stars), Agentic Kaggle Workflow (FrankS-IntelLab/agentic-kaggle-skill, 188 stars) and Retention Analysis (liangdabiao/claude-data-analysis-ultra-main, 290 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Lightgbm Analysis?

aipoch (a GitHub organization) maintains it in aipoch/medical-research-skills, which has 1,974 GitHub stars. The repository holds 567 skills in this directory. The repository was last updated on September 17, 2026.

Source: aipoch/medical-research-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.