Agent skill

Bio Copy Number Cnvkit Analysis

by GPTomics in GPTomics/bioSkills

Detect somatic and germline copy number variants from targeted, exome, and whole-genome sequencing with CNVkit, a read-depth caller that combines on-target and off-target (antitarget) coverage.

MITAuto-check passedBusiness, Finance & HR

Install Bio Copy Number Cnvkit Analysis

skills CLI
$ npx skills add GPTomics/bioSkills --skill bio-copy-number-cnvkit-analysis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install GPTomics/bioSkills bio-copy-number-cnvkit-analysis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/GPTomics/bioSkills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/copy-number/cnvkit-analysis .claude/skills/bio-copy-number-cnvkit-analysis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
bio-copy-number-cnvkit-analysis
GitHub stars
1.2k
Used in
2 other repos
Token cost
~4.1k tokens
SKILL.md length
1,648 words
Files
3
Skills in repo
559
Repo updated
First seen
Licence
MIT

At a glance

Detect somatic and germline copy number variants from targeted, exome, and whole-genome sequencing with CNVkit, a read-depth caller that combines on-target and off-target (antitarget) coverage.

  • Calling CNVs from hybrid-capture panels
  • SKILL.md covers Version Compatibility, Where CNVkit Sits — Caller…, Decision Tree by Scenario and Core Pipeline — Tumor-Normal…, plus 11 more sections
  • Runs Shell scripts from its folder; calls pip and python
  • Deciding whether CNVkit (depth-only) is the right tool versus an allele-specific caller

What it does

Bio Copy Number Cnvkit Analysis is an agent skill from GPTomics/bioSkills. Detect somatic and germline copy number variants from targeted, exome, and whole-genome sequencing with CNVkit, a read-depth caller that combines on-target and off-target (antitarget) coverage. Covers panel-of-normals construction, flat-reference tumor-only calling, hybrid/amplicon/WGS modes, CBS vs HMM segmentation selection, purity-aware integer calling, and reconciliation against GATK and allele-specific callers. Use when calling CNVs from hybrid-capture panels or exomes, deciding whether CNVkit (depth-only)…

Its SKILL.md is about 4.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `examples/run_cnvkit.sh` and `usage-guide.md`).

It sits in Business, Finance & HR, covering Bioinformatics and Accounting and bookkeeping. It works with Python. The repository describes itself as: a set of SKILLS.md for doing bioinformatics with agents like claude code. The licence is MIT.

When your agent uses it

  • Calling CNVs from hybrid-capture panels
  • Deciding whether CNVkit (depth-only) is the right tool versus an allele-specific caller
  • Building a panel of normals
  • Diagnosing flat-reference false positives

Example prompts

  • “/bio-copy-number-cnvkit-analysis”

Requirements

  • Python 3
  • A Bash shell

What it can do on your machine

Read from SKILL.md and the folder at commit d91ed3d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships script files (Shell), which the agent can run.

    Shell commands in SKILL.md call:

    • pip
    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Bio Copy Number Cnvkit Analysis loads about 4.1k tokens when it runs. Until then it costs about 181 tokens; SKILL.md has 1,648 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~181
When it runs · the whole SKILL.md, loaded when a task matches
~4.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from GPTomics/bioSkills at commit d91ed3d, republished under its MIT licence (© GPTomics). 1,648 words, ~4,058 tokens.

Download SKILL.mdSave it as .claude/skills/bio-copy-number-cnvkit-analysis/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
bio-copy-number-cnvkit-analysis
description
Detect somatic and germline copy number variants from targeted, exome, and whole-genome sequencing with CNVkit, a read-depth caller that combines on-target and off-target (antitarget) coverage. Covers panel-of-normals construction, flat-reference tumor-only calling, hybrid/amplicon/WGS modes, CBS vs HMM segmentation selection, purity-aware integer calling, and reconciliation against GATK and allele-specific callers. Use when calling CNVs from hybrid-capture panels or exomes, deciding whether CNVkit (depth-only) is the right tool versus an allele-specific caller, building a panel of normals, diagnosing flat-reference false positives, or interpreting log2 ratios into copy-number states.
tool_type
cli
primary_tool
cnvkit

Version Compatibility

Reference examples tested with: CNVkit 0.9.10+, samtools 1.19+, bedtools 2.31+, Python 3.10+, R 4.3+ with DNAcopy 1.76+.

Before using code patterns, verify installed versions match. If versions differ:

  • CLI: cnvkit.py version then cnvkit.py batch --help to confirm flags
  • Python: pip show cnvkit then python -c "import cnvlib; help(cnvlib.read)"
  • R: packageVersion('DNAcopy') (CBS backend)

If a command throws an unrecognized-argument or AttributeError, introspect the installed version and adapt the example rather than retrying. CNVkit segmentation methods (hmm, hmm-tumor, hmm-germline) depend on pomegranate; CBS depends on Bioconductor DNAcopy.

CNVkit Copy Number Analysis

"Detect copy number variants from my exome / panel data" -> Run a read-depth pipeline: normalize on-target and off-target coverage against a reference, segment the log2-ratio profile, and call gains/losses. CNVkit is a depth-only caller — it estimates relative copy number and cannot, on its own, resolve tumor purity, ploidy, or allele-specific state. Choosing CNVkit is a decision that the experiment does not require allelic resolution.

  • CLI: cnvkit.py batch tumor.bam --normal normal.bam --targets panel.bed --fasta ref.fa
  • Python API: cnvlib.read('sample.cnr') for downstream filtering

Where CNVkit Sits — Caller Taxonomy

CallerSignal usedOutputPurity/ploidy awareFails when
CNVkitDepth (on + off-target)Relative log2, threshold-called CNNo (manual --purity)Sample purity < ~40%; hyper-aneuploid genome breaks median centering; balanced events invisible
GATK gCNV / somatic CNVDepth (PCA/tangent denoised)Copy-ratio segments, +/-/0 callNo (somatic); ploidy prior (germline)Recurrent CNV in the PoN normalized away; ModelSegments gives no integer ASCN
ASCAT / Sequenza / FACETSDepth + B-allele frequencyInteger allele-specific CN, purity, ploidyYes (jointly fit)Near-diploid genome cannot anchor purity; low het-SNP density
ExomeDepth / GATK gCNV (cohort)Depth across a cohortGermline CN genotypeGermline ploidy only< ~30-100 technically matched samples; common CNV

CNVkit's niche: a fast, single-sample depth caller for hybrid-capture panels and exomes where antitarget reads recover genome-wide resolution. Its limit: it answers "is this region gained or lost relative to baseline" — not "how many absolute copies, on which haplotype, in what fraction of cells." For tumor integer CN, purity, LOH, or whole-genome doubling, escalate to allele-specific-copy-number.

Decision Tree by Scenario

ScenarioRecommended CNVkit configurationWhy
Hybrid-capture panel or exome, tumor-normalbatch hybrid mode, matched normal as referenceAntitargets recover off-target resolution; matched normal cancels capture bias
Exome cohort, pooled normals availableBuild pooled PoN reference, then batch --referencePooled reference averages out per-normal noise; 5-20+ normals
Amplicon / multiplex-PCR panelbatch --method ampliconNo usable off-target reads; antitarget bins are pure noise — must be dropped
Whole-genome sequencingbatch --method wgsGenome-wide fixed bins; no target/antitarget split
Tumor-only, no normal of any kindbatch with flat reference (omit --normal)Last resort; expect GC/capture-bias false positives — see failure mode below
FFPE / low-input / impure tumorAdd --drop-low-coverage; segment with hmm-tumorFFPE dropout produces zero-coverage bins that CBS reads as deletions
Need absolute CN, LOH, purityDo not use CNVkit aloneEscalate to allele-specific-copy-number (ASCAT/Sequenza/FACETS/PureCN)

Core Pipeline — Tumor-Normal Pair

The batch command wraps target/antitarget generation, coverage, reference building, fix, and segment:

bash
cnvkit.py batch tumor.bam \
    --normal normal.bam \
    --targets panel.bed \
    --annotate refFlat.txt \
    --fasta reference.fa \
    --access access-excludes.bed \
    --output-reference reference.cnn \
    --output-dir results/ \
    --drop-low-coverage \
    --diagram --scatter

--access restricts antitarget bins to mappable, non-gap genome (generate once with cnvkit.py access reference.fa -o access.bed). --drop-low-coverage is effectively mandatory for tumor, FFPE, or any sample with coverage dropout.

Panel of Normals — The Reference Determines Call Quality

A reference built from pooled normals is the single largest quality lever. Process is: build the reference from normals once, then run every tumor against it.

bash
# Build pooled reference from process-matched normals (same capture kit, same lab)
cnvkit.py batch --normal normal*.bam \
    --targets panel.bed --annotate refFlat.txt --fasta reference.fa \
    --access access.bed \
    --output-reference pooled_reference.cnn

# Run each tumor against the pre-built reference
cnvkit.py batch tumor*.bam --reference pooled_reference.cnn \
    --output-dir results/ --drop-low-coverage --scatter --diagram

Step-by-Step Pipeline (Fine-Grained Control)

bash
cnvkit.py target panel.bed --annotate refFlat.txt --split -o targets.bed
cnvkit.py antitarget panel.bed --access access.bed -o antitargets.bed
cnvkit.py coverage tumor.bam targets.bed -o tumor.targetcoverage.cnn
cnvkit.py coverage tumor.bam antitargets.bed -o tumor.antitargetcoverage.cnn
cnvkit.py reference normal*.{target,antitarget}coverage.cnn --fasta reference.fa -o reference.cnn
cnvkit.py fix tumor.targetcoverage.cnn tumor.antitargetcoverage.cnn reference.cnn -o tumor.cnr
cnvkit.py segment tumor.cnr -o tumor.cns --drop-low-coverage
cnvkit.py call tumor.cns -o tumor.call.cns

Segmentation Method Selection

CNVkit's segment step is where the bias-variance trade-off is set. The default CBS is not always correct — see copy-ratio-segmentation for the full algorithm comparison.

bash
cnvkit.py segment tumor.cnr -m cbs -o tumor.cns          # default; precise on focal events
cnvkit.py segment tumor.cnr -m hmm-tumor -o tumor.cns    # heterogeneous tumor, broad states
cnvkit.py segment tumor.cnr -m hmm-germline -o tumor.cns # germline, priors near diploid
cnvkit.py segment tumor.cnr -m haar -o tumor.cns         # fast, low-depth WGS

Rule of thumb: CBS for panels/exomes with adequate depth (precise on small segments); hmm-tumor for impure or heterogeneous tumors where CBS over-fragments; haar for shallow WGS where CBS recall degrades.

Purity-Aware Integer Calling

call converts segmented log2 ratios to copy-number states. The clonal method rescales by tumor purity before rounding to integers — without it, an impure tumor's true CN=4 amplification rounds to CN=3 or CN=2.

bash
# Threshold method (default): fixed log2 cutpoints, no purity correction
cnvkit.py call tumor.cns -o tumor.call.cns

# Clonal method: rescale by purity, then round to integer CN
cnvkit.py call tumor.cns -m clonal --purity 0.65 --ploidy 2 -o tumor.call.cns

# Overlay B-allele frequency from a SNV VCF (for LOH visualization, NOT joint ASCN)
cnvkit.py call tumor.cns -m clonal --purity 0.65 --vcf tumor.vcf.gz -o tumor.call.cns

CNVkit can read BAF from a VCF and report a baf column, but it segments log2 and BAF separately and does not jointly fit purity from them. For a true joint allele-specific model (ASPCF, FACETS joint segmentation), use allele-specific-copy-number.

Failure Modes

Flat reference (tumor-only) — systematic false focal calls

Trigger: No --normal and no pooled PoN; CNVkit builds a flat reference (uniform log2 0) from the FASTA.

Mechanism: A flat reference corrects only GC and (optionally) RepeatMasker content via the FASTA. It cannot correct capture efficiency, which varies 10-100x across probes and is the dominant bias in hybrid capture. Per-probe capture bias is then misread as copy number.

Symptom: Recurrent "CNVs" at the same loci across unrelated tumor-only samples; spiky .cnr profiles; high MAD; calls concentrated at probe boundaries.

Fix: Never rely on a flat reference for clinical or focal calls. Build a pooled PoN from >= 5 process-matched normals. If truly no normal exists, treat tumor-only CNVkit output as hypothesis-generating only and escalate to PureCN (allele-specific-copy-number), which models a normal database explicitly.

Antitarget bins on amplicon panels — pure noise

Trigger: Running default hybrid mode on an amplicon (multiplex-PCR) panel.

Mechanism: Amplicon panels produce essentially no off-target reads. Antitarget bins then contain a handful of stray reads, giving wildly variable log2 that the segmenter chases.

Symptom: Huge antitarget bin spread; nonsensical genome-wide segments between the targeted genes.

Fix: Use --method amplicon, which drops antitargets entirely and calls only from on-target bins. Accept that resolution is limited to the targeted genes.

Low tumor purity — the death zone below ~40%

Trigger: Tumor cellularity below ~40% (common in breast, lung adenocarcinoma, melanoma, low-cellularity biopsies).

Mechanism: Each somatic CN change is diluted by 2-copy normal DNA. A true single-copy loss at 30% purity produces log2 ~ -0.23 — inside the diploid threshold band.

Symptom: Genome looks near-flat; few or no calls; known driver amplifications (e.g. ERBB2, MYC) missed.

Fix: CNVkit cannot rescue this. Confirm purity with an allele-specific caller (BAF gives an orthogonal purity estimate). Below ~20% purity, no depth-based caller is reliable — report as indeterminate.

Show full SKILL.md (670 more words)Show less
Hyper-aneuploid / whole-genome-doubled genome — baseline miscalled

Trigger: Tumor with >50% of the genome altered, or whole-genome doubling.

Mechanism: call --center median (or mode) assumes the commonest log2 state is diploid. In a WGD genome the commonest state is tetraploid; centering on it shifts the whole profile and inverts gain/loss calls.

Symptom: Genome-wide pattern of calls inconsistent with known biology; "deletions" everywhere or "gains" everywhere.

Fix: Do not trust depth-only centering on aneuploid tumors. Anchor the diploid baseline with BAF/SNV data via an allele-specific caller, which estimates absolute ploidy directly.

FFPE / low-input dropout read as homozygous deletions

Trigger: Degraded FFPE DNA or low input; some bins have near-zero coverage.

Mechanism: Zero-coverage bins produce extreme negative log2; CBS joins them into spurious homozygous-deletion segments.

Symptom: Scattered tiny "CN=0" segments, often at hard-to-capture (high-GC) loci.

Fix: Always pass --drop-low-coverage to batch and segment. Inspect MAD; if MAD > 0.5, the sample is too noisy for confident focal calls.

Reconciliation: When CNVkit Disagrees With Another Caller

PatternLikely causeAction
CNVkit calls focal events GATK missesGATK PoN/tangent normalization absorbed the eventTrust CNVkit if the event is rare; suspect GATK if PoN contained tumors
GATK/ASCAT call broad arm events CNVkit flattensCNVkit centered on a non-diploid modeRe-center against an allele-specific ploidy estimate
CNVkit and ASCAT disagree on integer CNCNVkit purity guess wrong, or genome is WGDTrust ASCAT/FACETS — joint BAF+depth fit resolves purity/ploidy
Tumor-only CNVkit calls absent in matched-normal rerunGermline CNV or capture bias misread as somaticRe-run with the matched normal; germline CNVs are not somatic events

Operational rule for high-confidence reporting: Treat a CNVkit call as confident only when (1) the reference was a pooled PoN of process-matched normals, (2) sample MAD < 0.5, (3) the segment spans multiple bins with consistent weight, and (4) for any clinically actionable focal event, it is confirmed by an orthogonal caller or by allele-specific data. Depth-only calls on tumors are screening-grade, not definitive.

Quality Control

bash
cnvkit.py metrics results/*.cnr -s results/*.cns      # MAD, spread, bivar per sample
cnvkit.py sex results/*.cnr                           # detect sex / sample swaps
cnvkit.py segmetrics tumor.cnr -s tumor.cns --ci --pi --bootstrap 100 -o tumor.segmetrics.cns
cnvkit.py genemetrics tumor.cnr -s tumor.cns -t 0.2 --ci --bootstrap 100 -o tumor.genemetrics.tsv

Quantitative Thresholds

ThresholdValueSource / Rationale
Sample MAD (noise)< 0.5 acceptable; < 0.3 goodCNVkit docs; MAD is the median absolute deviation of bin log2
Panel of normals size>= 5; 10-20 preferredTalevich 2016; pooling averages per-normal capture noise
Default call thresholds (log2)-1.1, -0.25, 0.2, 0.7CNVkit call -t defaults: CN 0 / 1 / 2 / 3 / 4+ boundaries
Purity floor for depth calling~40% reliable; ~20% absolute floorBelow ~40% segmentation fails (sCNAphase, Gusnanto 2012)
genemetrics -t gain/loss0.2 (default)2^0.2 ~ 15% copy-ratio change; tune up for impure samples
Antitarget avg bin sizeauto; ~target size x fold-enrichmentCNVkit docs; off-target bins should hold comparable read counts

Common Errors

Error / symptomCauseSolution
Spiky .cnr, recurrent calls across samplesFlat reference; capture bias misreadBuild a pooled PoN; never use flat reference for focal calls
Antitarget spread huge, nonsense segmentsHybrid mode on an amplicon panelUse --method amplicon
Scattered CN=0 micro-segmentsFFPE zero-coverage binsAdd --drop-low-coverage to batch and segment
Integer CN systematically too lowcall without -m clonal --puritySupply purity, or use an allele-specific caller
Whole genome called gain or lossCentering on a non-diploid mode (WGD)Anchor ploidy with BAF; do not depth-center aneuploid tumors
pomegranate ImportError on HMMHMM backend not installedpip install pomegranate; or use -m cbs

References

  • Talevich E et al 2016. CNVkit: genome-wide copy number detection from targeted DNA sequencing. PLoS Comput Biol 12:e1004873
  • Olshen AB et al 2004. Circular binary segmentation for the analysis of array-based DNA copy number data. Biostatistics 5:557 (CBS)
  • Gusnanto A et al 2012. Correcting for cancer genome size and tumour cell content in whole-genome copy number. Bioinformatics 28:40
  • Benjamini Y, Speed TP 2012. Summarizing and correcting the GC content bias in high-throughput sequencing. Nucleic Acids Res 40:e72
  • copy-number/copy-ratio-segmentation - CBS vs HMM choice, depth normalization, bias correction
  • copy-number/allele-specific-copy-number - ASCAT/Sequenza/FACETS/PureCN for purity, ploidy, integer ASCN
  • copy-number/gatk-cnv - GATK depth-based alternative; tangent normalization
  • copy-number/cnv-annotation - Gene and clinical annotation of CNV calls
  • copy-number/cnv-visualization - Profile plots, segmentation views, cohort heatmaps
  • copy-number/recurrent-cnv - GISTIC2 cohort-level recurrent and driver CNV
  • alignment-files/bam-statistics - QC of input BAMs before calling
  • long-read-sequencing/structural-variants - Complementary breakpoint-resolved SV calling

© GPTomics, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files in copy-number/cnvkit-analysis of GPTomics/bioSkills.

  • SKILL.md
  • examples/run_cnvkit.sh
  • usage-guide.md

Open the folder on GitHubat commit d91ed3d

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in GPTomics/bioSkills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Bio Copy Number Cnvkit Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Bio Copy Number Cnvkit Analysis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Bio Copy Number Cnvkit Analysis this skillGPTomics/bioSkills1.2k2 repos~4.1kAutomated safety check: PassMIT
Beancount Importer Authorbex-co/beancount-io297—~2kAutomated safety check: PassMIT
Travel Expense ReimbursementFoxtailsss-Andy/Anna-Agent147—~812Automated safety check: PassMIT
SQL Server Table Reconciliationgithub/awesome-copilot40k1 repos~1.4kAutomated safety check: PassMIT
Financial Data EntrySerein-81/financial_rag148—~1.5kAutomated safety check: NotesNone
CICPA Company Registration Querynigo81/nigo-skills133—~2.8kAutomated safety check: PassMIT

Similar skills

  • Beancount Importer Author

    bex-co/beancount-io

    Write or repair a reusable Beangulp importer from a sample bank export, with reviewed golden files and a passing test harness.

    297 GitHub stars~2k tokensUpdated today
    Business, Finance & HRAuto-check passed
  • Travel Expense Reimbursement

    Foxtailsss-Andy/Anna-Agent

    Create, validate, submit, and verify employee travel and expense reimbursements through Anna Cowork and MCP.

    147 GitHub stars~812 tokensUpdated 2 days ago
    Business, Finance & HRAuto-check passed
  • SQL Server Table Reconciliation

    github/awesome-copilot

    Official

    A skill your agent uses when: comparing SQL Server tables across instances, data migration validation, ETL verification, row mismatch detection, schema drift, reconciliation report, production vs…

    40k GitHub starsUsed in 1 repo~1.4k tokens
    Business, Finance & HRAuto-check passed
  • Financial Data Entry

    Serein-81/financial_rag

    Guides entry of a single financial record by collecting fiscal-year figures, validating them with a script and confirming before submitting to the financial-data API.

    148 GitHub stars~1.5k tokensUpdated 4 mo ago
    Business, Finance & HRAuto-check: notes
  • Looks up Chinese company registration data in the CICPA industry knowledge base, from a quick search to a full 61-dimension export, subsidiary discovery and Excel output.

    133 GitHub stars~2.8k tokensUpdated 1 mo ago
    Business, Finance & HRAuto-check passed
  • Merges bank statements from several banks and formats into one standard Excel layout, then reconciles them against the general ledger with a layered matching engine.

    133 GitHub stars~1.3k tokensUpdated 1 mo ago
    Business, Finance & HRAuto-check passed

More from GPTomics/bioSkills

All 559 skills in this repo
  • Bio Alignment Io

    GPTomics/bioSkills

    Read, write, and convert multiple sequence alignment files using Biopython Bio.AlignIO.

    1.2k GitHub starsUsed in 3 repos~4.9k tokens
    Auto-check passed
  • bioSkills Installer

    GPTomics/bioSkills

    Installs the bioSkills collection of 425 bioinformatics skills in one step, or only chosen categories, so sequencing, RNA-seq, single-cell and variant tasks get specialized help.

    1.2k GitHub starsUsed in 1 repo~789 tokens
    Auto-check passed
  • Bio Write Sequences

    GPTomics/bioSkills

    Write biological sequences to files (FASTA, FASTQ, GenBank, EMBL) using Biopython Bio.SeqIO.

    1.2k GitHub starsUsed in 3 repos~2.1k tokens
    Auto-check passed
  • Amplicon Primer Clipping

    GPTomics/bioSkills

    Soft- or hard-clips PCR primer footprints from aligned amplicon BAMs so primer bases stop masquerading as confirmed reference sequence.

    1.2k GitHub starsUsed in 2 repos~2.2k tokens
    Auto-check passed
  • Filters BAM alignments by FLAG bits, mapping quality and regions with samtools view or pysam, with recipes for common keep and drop cases.

    1.2k GitHub starsUsed in 2 repos~3.6k tokens
    Auto-check passed
  • Bio Alignment Indexing

    GPTomics/bioSkills

    Create and use BAI/CSI indices for BAM/CRAM files using samtools and pysam.

    1.2k GitHub starsUsed in 2 repos~2.4k tokens
    Auto-check passed

Works with

Questions about Bio Copy Number Cnvkit Analysis

What does Bio Copy Number Cnvkit Analysis do?

Detect somatic and germline copy number variants from targeted, exome, and whole-genome sequencing with CNVkit, a read-depth caller that combines on-target and off-target (antitarget) coverage. Bio Copy Number Cnvkit Analysis is an agent skill from GPTomics/bioSkills. Detect somatic and germline copy number variants from targeted, exome, and whole-genome sequencing with CNVkit, a read-depth caller that combines on-target and off-target (antitarget) coverage.

When should I use Bio Copy Number Cnvkit Analysis?

Bio Copy Number Cnvkit Analysis fits situations like: calling CNVs from hybrid-capture panels; deciding whether CNVkit (depth-only) is the right tool versus an allele-specific caller; building a panel of normals; diagnosing flat-reference false positives.

How do I install Bio Copy Number Cnvkit Analysis in Claude Code?

Run `npx skills add GPTomics/bioSkills --skill bio-copy-number-cnvkit-analysis -a claude-code`. Or copy the skill folder (copy-number/cnvkit-analysis in GPTomics/bioSkills) into .claude/skills/bio-copy-number-cnvkit-analysis in your project. Claude Code loads it when a task matches its description.

How do I install Bio Copy Number Cnvkit Analysis in Codex?

Run `npx skills add GPTomics/bioSkills --skill bio-copy-number-cnvkit-analysis -a codex`. Or copy the skill folder (copy-number/cnvkit-analysis in GPTomics/bioSkills) into .agents/skills/bio-copy-number-cnvkit-analysis in your project. Codex loads it when a task matches its description.

Can I use Bio Copy Number Cnvkit Analysis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add GPTomics/bioSkills --skill bio-copy-number-cnvkit-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/bio-copy-number-cnvkit-analysis, .gemini/skills/bio-copy-number-cnvkit-analysis, .github/skills/bio-copy-number-cnvkit-analysis and .opencode/skills/bio-copy-number-cnvkit-analysis in your project.

What does Bio Copy Number Cnvkit Analysis need to run?

Going by SKILL.md and its folder, Bio Copy Number Cnvkit Analysis needs a shell for the scripts in its folder and the command-line tools its instructions call (pip and python). Our summary lists: Python 3; A Bash shell.

Does Bio Copy Number Cnvkit Analysis access the network?

SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Bio Copy Number Cnvkit Analysis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Bio Copy Number Cnvkit Analysis use?

Bio Copy Number Cnvkit Analysis is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Bio Copy Number Cnvkit Analysis use?

About 4.1k tokens (SKILL.md is roughly 16k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Bio Copy Number Cnvkit Analysis?

Skills that share tags, products or a category with Bio Copy Number Cnvkit Analysis: Beancount Importer Author (bex-co/beancount-io, 297 stars), Travel Expense Reimbursement (Foxtailsss-Andy/Anna-Agent, 147 stars), SQL Server Table Reconciliation (github/awesome-copilot, 40k stars) and Financial Data Entry (Serein-81/financial_rag, 148 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Bio Copy Number Cnvkit Analysis?

GPTomics (a GitHub organization) maintains it in GPTomics/bioSkills, which has 1,218 GitHub stars. The repository holds 559 skills in this directory. The repository was last updated on August 15, 2026.

Source: GPTomics/bioSkills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.