Agent skill

Bio Tcr Bcr Analysis Vdjtools Analysis

by GPTomics in GPTomics/bioSkills

Computes immune-repertoire diversity, clonal structure, overlap, and segment usage from TCR/BCR clonotype tables with VDJtools (immunarch as the modern R alternative).

MITAuto-check passedResearch & Science

Install Bio Tcr Bcr Analysis Vdjtools Analysis

skills CLI
$ npx skills add GPTomics/bioSkills --skill bio-tcr-bcr-analysis-vdjtools-analysis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install GPTomics/bioSkills bio-tcr-bcr-analysis-vdjtools-analysis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/GPTomics/bioSkills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/tcr-bcr-analysis/vdjtools-analysis .claude/skills/bio-tcr-bcr-analysis-vdjtools-analysis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
bio-tcr-bcr-analysis-vdjtools-analysis
GitHub stars
1.2k
Used in
1 other repo
Token cost
~4.5k tokens
SKILL.md length
1,848 words
Files
3
Skills in repo
559
Repo updated
First seen
Licence
MIT

At a glance

Computes immune-repertoire diversity, clonal structure, overlap, and segment usage from TCR/BCR clonotype tables with VDJtools (immunarch as the modern R alternative).

  • Works in 3 steps: Depth-dependent through the ln(S)… → It discards richness (it is a rescaled… → It is dominated by the middle of the…
  • Deciding which diversity estimator answers a question (q=0 observed richness/chao1/chaoE
  • SKILL.md covers Version Compatibility, The governing principle:…, Report a Hill profile, not one… and Clonality: the field default…, plus 10 more sections
  • Runs Shell scripts from its folder; calls java

What it does

Bio Tcr Bcr Analysis Vdjtools Analysis is an agent skill from GPTomics/bioSkills. Computes immune-repertoire diversity, clonal structure, overlap, and segment usage from TCR/BCR clonotype tables with VDJtools (immunarch as the modern R alternative). Use when deciding which diversity estimator answers a question (q=0 observed richness/chao1/chaoE, q=1 shannonWienerIndex, q=2 inverseSimpson as a Hill profile); normalizing sequencing depth before any cross-sample claim (DownSample or the resampled CalcDiversityStats table); choosing an overlap metric (depth-robust MorisitaHorn/F2 vs depth-biased…

Its SKILL.md is about 4.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `examples/diversity_analysis.sh` and `usage-guide.md`).

It sits in Research & Science. It works with Java. The repository describes itself as: a set of SKILLS.md for doing bioinformatics with agents like claude code. The licence is MIT.

When your agent uses it

  • Deciding which diversity estimator answers a question (q=0 observed richness/chao1/chaoE
  • Q=1 shannonWienerIndex
  • Q=2 inverseSimpson as a Hill profile)
  • Normalizing sequencing depth before any cross-sample claim (DownSample

Example prompts

  • “Use the bio-tcr-bcr-analysis-vdjtools-analysis skill to compute immune-repertoire diversity, clonal structure, overlap, and segment usage from…”
  • “/bio-tcr-bcr-analysis-vdjtools-analysis”

Requirements

  • A Bash shell

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Depth-dependent through the ln(S) denominator -- compare clonality only on depth-normalized samples.
  2. It discards richness (it is a rescaled evenness): two repertoires with identical clonality can differ 100x in richness.
  3. It is dominated by the middle of the abundance distribution, not the top clones a clinician cares about.

What it can do on your machine

Read from SKILL.md and the folder at commit d91ed3d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships script files (Shell), which the agent can run.

    Shell commands in SKILL.md call:

    • java

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Bio Tcr Bcr Analysis Vdjtools Analysis loads about 4.5k tokens when it runs. Until then it costs about 210 tokens; SKILL.md has 1,848 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~210
When it runs · the whole SKILL.md, loaded when a task matches
~4.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from GPTomics/bioSkills at commit d91ed3d, republished under its MIT licence (© GPTomics). 1,848 words, ~4,527 tokens.

Download SKILL.mdSave it as .claude/skills/bio-tcr-bcr-analysis-vdjtools-analysis/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
bio-tcr-bcr-analysis-vdjtools-analysis
description
Computes immune-repertoire diversity, clonal structure, overlap, and segment usage from TCR/BCR clonotype tables with VDJtools (immunarch as the modern R alternative). Use when deciding which diversity estimator answers a question (q=0 observed richness/chao1/chaoE, q=1 shannonWienerIndex, q=2 inverseSimpson as a Hill profile); normalizing sequencing depth before any cross-sample claim (DownSample or the resampled CalcDiversityStats table); choosing an overlap metric (depth-robust MorisitaHorn/F2 vs depth-biased Jaccard/public counts) and a clonotype match key (-i nt/aa, +/-V/J); summarizing clonality as 1 - normalizedShannonWienerIndex; reading spectratype and V-J usage under primer bias; interpreting public clonotypes; and choosing VDJtools (stable Java CLI) vs immunarch (active tidy R).
tool_type
cli
primary_tool
VDJtools

Version Compatibility

Reference examples tested with: VDJtools 1.2.1+, Java (JRE 8+), R 4.x with ggplot2/reshape2/gridExtra (for Plot* modules), immunarch 0.9+/1.0+

Before using code patterns, verify installed versions match. If versions differ:

  • CLI: java -jar vdjtools.jar prints the current routine list; <routine> with no args prints its flags
  • R: packageVersion('immunarch') then ?repDiversity / ?repOverlap to confirm .method strings

If code throws ImportError, AttributeError, or TypeError, introspect the installed package and adapt the example to match the actual API rather than retrying.

Note: routine names are CamelCase and case-sensitive, and the depth-resampling routine is DownSample (capital S). Run RInstall once so the Plot* modules can call R. VDJtools is post-analysis only: it consumes clonotype tables (from MiXCR etc.), not FASTQ.

VDJtools Analysis

"Compute diversity and compare my TCR/BCR repertoires" -> summarize each repertoire's clonal structure, compare samples at equal depth, and quantify overlap.

  • CLI: java -jar vdjtools.jar CalcDiversityStats | CalcPairwiseDistances | TrackClonotypes
  • R alternative: immunarch repDiversity(), repOverlap(), repClonality()

The governing principle: diversity is sampling-depth-dependent

A repertoire is a sample of an enormous, unevenly expanded clonal population with a long tail of rare clonotypes, so observed richness never saturates: deeper sequencing keeps discovering new clonotypes. Observed richness, Shannon entropy, clonality, Jaccard, and shared-clonotype counts are all functions of read depth. Comparing raw values across libraries of unequal depth measures depth, not biology -- this is the field's single most common and most invalidating error.

The fix is mandatory before any cross-sample claim: bring all samples to a common depth. Two routes:

  • DownSample -x <reads> every sample to a shared depth, then analyze; or
  • read CalcDiversityStats at a common depth from its resampled table. CalcDiversityStats emits two tables, diversity.<i>.txt (original) and diversity.<i>.resampled.txt (downsampled to the smallest sample or -x). Use the resampled/normalized values for between-sample comparison; the original table is for within-sample description only.

Choosing the normalization depth is itself a decision: downsampling every sample to the cohort minimum discards data and can leave everyone underpowered if one library is tiny. Set the common depth near the cohort's lower quartile, and EXCLUDE (do not drag everyone down to) any sample far below it -- a sample whose rarefaction curve is still steeply climbing well below the chosen depth is under-sampled and cannot support a diversity claim at all. Report the chosen depth and any excluded samples. PlotQuantileStats and the rarefaction curves show which samples are safe to include.

Rarefaction makes the problem visible: RarefactionPlot draws interpolated + extrapolated diversity-vs-depth curves. Compare samples at a common x, never at curve endpoints of different depth. Extrapolation is reliable only to ~2-3x observed depth and degrades for q=0 (Chao 2014).

Report a Hill profile, not one number

A single index misleads because indices weight the abundance distribution differently. Report the Hill profile -- effective number of clonotypes at orders q=0, 1, 2 -- whose shape (steep drop from q=0 to q=2 = a few dominant clones over a large rare tail) is the informative object (Greiff 2015 Genome Med 7:49; Chao 2014 Ecol Monogr 84:45-67). Two repertoires can share richness yet have opposite clonality.

CalcDiversityStats emits these columns (each with _mean/_std); Gini is NOT among them (it is an immunarch option, not a VDJtools output):

ColumnHill orderQuestion it answersDepth-robustness
observedDiversityq=0How many distinct clonotypes were seenPoor -- must downsample
chao1q=0Nonparametric richness lower bound (uses singletons f1, doubletons f2)Poor; breaks without count data (f2=0) or if rare clones were pre-filtered; PCR error inflates it
chaoEq=0Chao richness extrapolated, normalized for cross-sample useModerate (VDJtools' preferred richness proxy)
efronThistedq=0Efron-Thisted lower-bound total diversityPoor; a lower bound, not the truth
shannonWienerIndexq=1exp(Shannon), effective number weighting by frequencyModerate
normalizedShannonWienerIndex--Pielou evenness H'/ln(S), range 0-1Depth-dependent through ln(S)
inverseSimpsonq=21/sum(p^2), dominated by abundant clonesBest -- most depth-robust
d50--Fewest top clones covering 50% of readsPoor; coarse descriptor

Default: report q=0 (chaoE or downsampled observedDiversity), q=1 (shannonWienerIndex), and q=2 (inverseSimpson) together. Feed chao1/efronThisted only genuine count data with singletons and doubletons; on non-UMI, non-error-corrected data, PCR/sequencing errors manufacture singletons and inflate them arbitrarily.

Clonality: the field default and its three flaws

Clonality = 1 - normalizedShannonWienerIndex = 1 - H'/ln(S). It runs 0 (even/polyclonal) to 1 (one clone dominates) and is the near-universal one-number summary because it is bounded and intuitive. State its flaws in any report:

  1. Depth-dependent through the ln(S) denominator -- compare clonality only on depth-normalized samples.
  2. It discards richness (it is a rescaled evenness): two repertoires with identical clonality can differ 100x in richness.
  3. It is dominated by the middle of the abundance distribution, not the top clones a clinician cares about.

Fix: report clonality alongside a q=2 Hill number (inverseSimpson) and a rarefaction curve, never alone.

Overlap: pick a depth-robust metric and hold the match key fixed

CalcPairwiseDistances builds an N x N matrix; OverlapPair compares two samples. Any count-of-shared-clonotypes or set index is dominated by the shallower sample's depth: a clone can only be shared if sampled in both, so the shallow sample caps the intersection. Downsample both samples to a common depth first, and prefer abundance-weighted metrics.

MetricBasisBest whenFails when
MorisitaHornAbundance, size-normalizedUnequal depth; the default choice-- (near-invariant to depth; dominated by abundant shared clones)
F2Sum of per-clonotype geometric-mean frequenciesFrequency-weighted overlap robust to a single dominant shared clone-- (preferred VDJtools frequency metric)
FGeometric mean of summed shared frequenciesQuick frequency overlapOne large shared clone dominates it
RPearson of log-frequencies over shared clones onlyConcordance of abundances among shared clonesIgnores private clones entirely
DShared count / geometric-mean diversitiesDescriptiveNumerator (shared count) still depth-biased
JaccardPresence/absenceEqual-depth, denoised samples onlyDominated by the shallower sample's depth

Overlap magnitude also swings by orders of magnitude with the clonotype match key (-i): nt (strict, few coincidental shares) vs aa (convergent recombination inflates sharing), and whether V/J must match (ntV, ntVJ, aaVJ, ...). Fix one key and hold it constant across every comparison in a study; state it in every figure. TrackClonotypes does ordered all-vs-all intersection for time courses -- a clone scoring "absent" at a timepoint is often a sampling zero, so downsample timepoints to common depth before declaring contraction.

An overlap number is only interpretable against a null: some sharing is expected by chance from convergent recombination of high-Pgen clonotypes. To claim overlap EXCEEDS chance, compare the observed statistic to a background of unrelated-donor pairs, or to shuffled/label-permuted repertoires at the same depth, and for public-clonotype claims condition on generation probability (specificity-annotation). Two related individuals or two timepoints from one host will always overlap more than two random donors regardless of biology.

Show full SKILL.md (765 more words)Show less

Public is not antigen-driven

"Public" (a clonotype shared across individuals) is largely an artifact of generation probability, not shared antigen selection. High-Pgen CDR3s -- short, few insertions, near-germline (and fetal-generated) -- are independently produced by many donors, and convergent recombination compounds this at the aa level (Venturi 2006 PNAS 103:18691). So a public/shared count is enriched for stochastic high-Pgen sequences, not evidence of a shared response. To argue antigen association, condition on Pgen (OLGA/IGoR) or intersect with an antigen database (ScanDatabase against VDJdb; immunarch dbAnnotate against VDJdb/McPAS-TCR) -- and even a database hit is a sequence match, not proof of binding. Hand Pgen-aware interpretation off to specificity-annotation.

Segment usage and spectratype

CalcSegmentUsage yields per-sample V/J frequency vectors; CalcSpectratype yields the CDR3-length histogram (PlotFancySpectratype overlays the top-N clones; PlotSpectratypeV stacks by V family; PlotFancyVJUsage is the V-J chord plot).

  • Spectratype shape: a Gaussian/bell length distribution indicates a diverse polyclonal (naive-like) repertoire; skew or spikes at particular lengths indicate clonal expansion(s). Weighting by reads shows expansions; weighting by unique clonotypes shows underlying diversity.
  • Confound: multiplex-PCR primer sets have V-gene-specific amplification bias, so apparent V/J usage differences between platforms or batches are frequently primer artifacts, not biology (Barennes 2021 Nat Biotechnol 39:236). Compare usage only within one protocol, or use 5'-RACE/UMI data. Usage vectors are compositional (sum to 1): CLR-transform before PCA and check that PC1 is not just depth/batch.

VDJtools vs immunarch

Both consume the same clonotype tables; pick by pipeline, not by metric.

VDJtoolsimmunarch
LanguageJava CLI (calls R for plots)R / tidyverse
MaintenanceStable, low activity (~1.2.1)Actively maintained (v1.0 adds airr_*)
IngestionConvert -S <fmt>repLoad() auto-detects MiXCR/Adaptive/10x/AIRR/VDJtools
PlottingFixed Plot* PDFsvis() returns editable ggplot objects
StrengthsReference F/F2/chaoE + resampled tables; legacy reproducibility; pairs with MiXCR/VDJdb10x single-cell, k-mer/motif, publication plots, ML feature matrices

Prefer VDJtools for CLI/legacy MiXCR pipelines and its exact resampled diversity tables; prefer immunarch for R, single-cell, k-mer/motif, or editable figures. Many groups convert with VDJtools and analyze/plot with immunarch.

Core operations in immunarch (verify .method strings on the installed version):

r
library(immunarch)
data <- repLoad('samples_dir/')                       # metadata + tidy clonotype tables

repDiversity(data$data, .method = 'raref')            # rarefaction/extrapolation curves (the depth control)
repDiversity(data$data, .method = 'hill')             # Hill profile across q
repDiversity(data$data, .method = 'inv.simp')         # q=2, depth-robust
repClonality(data$data, .method = 'homeo')            # clonal-space homeostasis (Rare..Hyperexpanded bins)
repOverlap(data$data, .method = 'morisita')           # depth-robust overlap; 'jaccard'/'public' are depth-biased
geneUsage(data$data[[1]])                             # V/J usage vector
trackClonotypes(data$data, list('Sample1', 1:10))     # longitudinal tracking
dbAnnotate(data$data, vdjdb, 'CDR3.aa', 'cdr3')       # antigen-database annotation

Prepare and normalize (CLI)

Convert upstream output, drop nonfunctional clones, and downsample to a shared depth before any comparison.

bash
# Import MiXCR clonotypes to VDJtools format (also: -S migec/immunoseq/imgt/vidjil ...)
java -jar vdjtools.jar Convert -S mixcr mixcr_clones.txt converted/

# Keep only functional (in-frame, no-stop) clonotypes for functional-repertoire analysis
java -jar vdjtools.jar FilterNonFunctional -m metadata.txt filtered/

# Optional: remove cross-sample contamination (barcode switching / chimeras) before cross-sample work
java -jar vdjtools.jar Decontaminate -m metadata.txt decontaminated/

# Downsample every sample to a common read depth (capital S; -x/--size = target reads)
java -jar vdjtools.jar DownSample -x 100000 -m metadata.txt downsampled/

The metadata file is tab-delimited: a #file.name sample.id <covariate...> header row, then one row per sample. Most multi-sample routines consume it via -m.

Diversity and overlap (CLI)

bash
# Diversity: emits diversity.<i>.txt (original) AND diversity.<i>.resampled.txt (depth-normalized)
# Compare across samples using the RESAMPLED table only.
java -jar vdjtools.jar CalcDiversityStats -m metadata.txt diversity/

# Rarefaction curves -- the correct visual for comparing diversity across depths (read at common x)
java -jar vdjtools.jar RarefactionPlot -m metadata.txt rarefaction/

# Pairwise overlap; -i sets the clonotype match key (hold constant study-wide).
# Report MorisitaHorn / F2 columns; treat Jaccard as depth-biased.
java -jar vdjtools.jar CalcPairwiseDistances -i aa -m metadata.txt overlap/
java -jar vdjtools.jar ClusterSamples -e MorisitaHorn overlap/ clustered/

# Longitudinal tracking across an ordered sample set (downsample timepoints first)
java -jar vdjtools.jar TrackClonotypes -m metadata_timecourse.txt tracking/

Parse VDJtools output in Python

python
import pandas as pd

def load_resampled_diversity(prefix):
    '''Load the depth-normalized diversity table for cross-sample comparison.'''
    return pd.read_csv(f'{prefix}.strict.resampled.txt', sep='\t')

def load_overlap_matrix(path, metric='MorisitaHorn'):
    '''Load one depth-robust overlap metric from the pairwise-distance output.'''
    df = pd.read_csv(path, sep='\t')
    return df.pivot(index='1_sample_id', columns='2_sample_id', values=metric)

Common Errors

SymptomCauseFix
Deeper libraries look 'more diverse' every timeComparing raw richness/Shannon/clonality across unequal depth -- measuring depth, not biologyDownSample to a common depth or read the .resampled.txt table; compare rarefaction curves at a common x
Overlap flips when samples are swapped or re-sequencedJaccard / public counts dominated by the shallower sample's depthDownsample both, and report MorisitaHorn or F2
Overlap magnitude differs wildly between studiesDifferent clonotype match key (-i aa vs nt, +/-V/J)Fix one -i value and state it in every figure
chao1/efronThisted are NaN or absurdly largeFed frequency-only or rare-clone-filtered data, or PCR errors created singletonsProvide genuine count data (singletons/doubletons); UMI/error-correct first; or use inverseSimpson
One clonality number reported as 'the diversity'Clonality is a rescaled evenness that discards richness and is depth-dependentReport Hill q=0/1/2 (add inverseSimpson) plus a rarefaction curve
'Public' clones claimed as antigen-specificPublicity is mostly high-Pgen convergent recombinationCondition on Pgen (OLGA) or intersect VDJdb via ScanDatabase; hand off to specificity-annotation
V/J usage differs between cohorts by platformMultiplex-PCR primer bias, not biologyCompare usage only within a protocol; CLR-transform before PCA and check PC1 is not depth/batch
Plot* routine errors on startR plotting dependencies missingRun java -jar vdjtools.jar RInstall once
  • mixcr-analysis - Generate input clonotype tables
  • repertoire-visualization - Rarefaction, spectratype and overlap figures
  • immcantation-analysis - BCR-aware diversity and clonal analysis
  • specificity-annotation - Pgen-aware interpretation of public clonotypes
  • experimental-design/sample-size - Sequencing depth and power planning
  • workflows/tcr-pipeline - End-to-end orchestration

References

  • Shugay M, et al. VDJtools: unifying post-analysis of T cell receptor repertoires. PLoS Comput Biol 2015; 11(11):e1004503.
  • Chao A, et al. Rarefaction and extrapolation with Hill numbers: a framework for sampling and estimation in species diversity studies. Ecol Monogr 2014; 84(1):45-67.
  • Greiff V, et al. A bioinformatic framework for immune repertoire diversity profiling enables detection of immunological status. Genome Med 2015; 7:49.
  • Chao A. Nonparametric estimation of the number of classes in a population. Scand J Stat 1984; 11:265-270.
  • Venturi V, et al. Sharing of T cell receptors in antigen-specific responses is driven by convergent recombination. PNAS 2006; 103(49):18691-18696.
  • Barennes P, et al. Benchmarking of T cell receptor repertoire profiling methods reveals large systematic biases. Nat Biotechnol 2021; 39:236-245.
  • ImmunoMind Team. immunarch: an R package for painless analysis of T-cell and B-cell immune repertoires. CRAN / immunarch.com (v1.x).

© GPTomics, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files in tcr-bcr-analysis/vdjtools-analysis of GPTomics/bioSkills.

  • SKILL.md
  • examples/diversity_analysis.sh
  • usage-guide.md

Open the folder on GitHubat commit d91ed3d

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in GPTomics/bioSkills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Bio Tcr Bcr Analysis Vdjtools Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Bio Tcr Bcr Analysis Vdjtools Analysis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Bio Tcr Bcr Analysis Vdjtools Analysis this skillGPTomics/bioSkills1.2k1 repos~4.5kAutomated safety check: PassMIT
FastreerClawBio/ClawBio1.2k1 repos~3.5kAutomated safety check: NotesGPL-3.0
Pymoojaechang-hits/SciAgent-Skills3741 repos~4.9kAutomated safety check: PassApache-2.0
Snpeff Variant Annotationjaechang-hits/SciAgent-Skills3741 repos~5.4kAutomated safety check: PassMIT
Pyimagej Fiji Bridgejaechang-hits/SciAgent-Skills3741 repos~6.2kAutomated safety check: PassApache-2.0
Hypothesis Generationspacering-net/codeg3.9k14 repos~3.6kAutomated safety check: NotesMIT

Similar skills

  • Fastreer

    ClawBio/ClawBio

    Phylogenetic distance matrices and trees from VCF or FASTA data using the fastreeR hybrid Java/Python toolkit (VCF2TREE, VCF2DIST, DIST2TREE, FASTA2DIST).

    1.2k GitHub starsUsed in 1 repo~3.5k tokens
    Research & ScienceAuto-check: notes
  • Pymoo

    jaechang-hits/SciAgent-Skills

    Python framework for single- and multi-objective optimization with evolutionary algorithms.

    374 GitHub starsUsed in 1 repo~4.9k tokens
    Research & ScienceAuto-check passed
  • Snpeff Variant Annotation

    jaechang-hits/SciAgent-Skills

    Annotate and filter VCF variants with SnpEff and SnpSift. An agent skill from jaechang-hits/SciAgent-Skills.

    374 GitHub starsUsed in 1 repo~5.4k tokens
    Research & ScienceAuto-check passed
  • Pyimagej Fiji Bridge

    jaechang-hits/SciAgent-Skills

    Python bridge to ImageJ2/Fiji for macros, plugins (Bio-Formats, TrackMate, Analyze Particles), NumPy↔ImagePlus/ImgLib2 exchange, and ImageJ Ops.

    374 GitHub starsUsed in 1 repo~6.2k tokens
    Data & AnalyticsAuto-check passed
  • Hypothesis Generation

    spacering-net/codeg

    Structured hypothesis formulation from observations. An agent skill from spacering-net/codeg.

    3.9k GitHub starsUsed in 14 repos~3.6k tokens
    Research & ScienceAuto-check: notes
  • GitHub Deep Research

    bytedance/deer-flow

    Researches a GitHub repository over four rounds using the GitHub API and web search, then writes a structured markdown report with timeline, metrics and Mermaid diagrams.

    84k GitHub starsUsed in 4 repos~1.3k tokens
    Research & ScienceAuto-check passed

More from GPTomics/bioSkills

All 559 skills in this repo
  • Bio Alignment Io

    GPTomics/bioSkills

    Read, write, and convert multiple sequence alignment files using Biopython Bio.AlignIO.

    1.2k GitHub starsUsed in 3 repos~4.9k tokens
    Auto-check passed
  • bioSkills Installer

    GPTomics/bioSkills

    Installs the bioSkills collection of 425 bioinformatics skills in one step, or only chosen categories, so sequencing, RNA-seq, single-cell and variant tasks get specialized help.

    1.2k GitHub starsUsed in 1 repo~789 tokens
    Auto-check passed
  • Bio Write Sequences

    GPTomics/bioSkills

    Write biological sequences to files (FASTA, FASTQ, GenBank, EMBL) using Biopython Bio.SeqIO.

    1.2k GitHub starsUsed in 3 repos~2.1k tokens
    Auto-check passed
  • Amplicon Primer Clipping

    GPTomics/bioSkills

    Soft- or hard-clips PCR primer footprints from aligned amplicon BAMs so primer bases stop masquerading as confirmed reference sequence.

    1.2k GitHub starsUsed in 2 repos~2.2k tokens
    Auto-check passed
  • Filters BAM alignments by FLAG bits, mapping quality and regions with samtools view or pysam, with recipes for common keep and drop cases.

    1.2k GitHub starsUsed in 2 repos~3.6k tokens
    Auto-check passed
  • Bio Alignment Indexing

    GPTomics/bioSkills

    Create and use BAI/CSI indices for BAM/CRAM files using samtools and pysam.

    1.2k GitHub starsUsed in 2 repos~2.4k tokens
    Auto-check passed

Works with

Questions about Bio Tcr Bcr Analysis Vdjtools Analysis

What does Bio Tcr Bcr Analysis Vdjtools Analysis do?

Computes immune-repertoire diversity, clonal structure, overlap, and segment usage from TCR/BCR clonotype tables with VDJtools (immunarch as the modern R alternative). Bio Tcr Bcr Analysis Vdjtools Analysis is an agent skill from GPTomics/bioSkills. Computes immune-repertoire diversity, clonal structure, overlap, and segment usage from TCR/BCR clonotype tables with VDJtools (immunarch as the modern R alternative).

When should I use Bio Tcr Bcr Analysis Vdjtools Analysis?

Bio Tcr Bcr Analysis Vdjtools Analysis fits situations like: deciding which diversity estimator answers a question (q=0 observed richness/chao1/chaoE; Q=1 shannonWienerIndex; Q=2 inverseSimpson as a Hill profile); normalizing sequencing depth before any cross-sample claim (DownSample.

How do I install Bio Tcr Bcr Analysis Vdjtools Analysis in Claude Code?

Run `npx skills add GPTomics/bioSkills --skill bio-tcr-bcr-analysis-vdjtools-analysis -a claude-code`. Or copy the skill folder (tcr-bcr-analysis/vdjtools-analysis in GPTomics/bioSkills) into .claude/skills/bio-tcr-bcr-analysis-vdjtools-analysis in your project. Claude Code loads it when a task matches its description.

How do I install Bio Tcr Bcr Analysis Vdjtools Analysis in Codex?

Run `npx skills add GPTomics/bioSkills --skill bio-tcr-bcr-analysis-vdjtools-analysis -a codex`. Or copy the skill folder (tcr-bcr-analysis/vdjtools-analysis in GPTomics/bioSkills) into .agents/skills/bio-tcr-bcr-analysis-vdjtools-analysis in your project. Codex loads it when a task matches its description.

Can I use Bio Tcr Bcr Analysis Vdjtools Analysis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add GPTomics/bioSkills --skill bio-tcr-bcr-analysis-vdjtools-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/bio-tcr-bcr-analysis-vdjtools-analysis, .gemini/skills/bio-tcr-bcr-analysis-vdjtools-analysis, .github/skills/bio-tcr-bcr-analysis-vdjtools-analysis and .opencode/skills/bio-tcr-bcr-analysis-vdjtools-analysis in your project.

What does Bio Tcr Bcr Analysis Vdjtools Analysis need to run?

Going by SKILL.md and its folder, Bio Tcr Bcr Analysis Vdjtools Analysis needs a shell for the scripts in its folder and the command-line tools its instructions call (java). Our summary lists: A Bash shell.

Does Bio Tcr Bcr Analysis Vdjtools Analysis access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Bio Tcr Bcr Analysis Vdjtools Analysis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Bio Tcr Bcr Analysis Vdjtools Analysis use?

Bio Tcr Bcr Analysis Vdjtools Analysis is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Bio Tcr Bcr Analysis Vdjtools Analysis use?

About 4.5k tokens (SKILL.md is roughly 18k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Bio Tcr Bcr Analysis Vdjtools Analysis?

Skills that share tags, products or a category with Bio Tcr Bcr Analysis Vdjtools Analysis: Fastreer (ClawBio/ClawBio, 1.2k stars), Pymoo (jaechang-hits/SciAgent-Skills, 374 stars), Snpeff Variant Annotation (jaechang-hits/SciAgent-Skills, 374 stars) and Pyimagej Fiji Bridge (jaechang-hits/SciAgent-Skills, 374 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Bio Tcr Bcr Analysis Vdjtools Analysis?

GPTomics (a GitHub organization) maintains it in GPTomics/bioSkills, which has 1,218 GitHub stars. The repository holds 559 skills in this directory. The repository was last updated on August 15, 2026.

Source: GPTomics/bioSkills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.