Agent skill

Bio Ctdna Mutation Detection

by GPTomics in GPTomics/bioSkills

Detects somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise problem set by error suppression and molecules sampled, not by the choice of caller.

MITAuto-check passedResearch & Science

Install Bio Ctdna Mutation Detection

skills CLI
$ npx skills add GPTomics/bioSkills --skill bio-ctdna-mutation-detection -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install GPTomics/bioSkills bio-ctdna-mutation-detection --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/GPTomics/bioSkills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/liquid-biopsy/ctdna-mutation-detection .claude/skills/bio-ctdna-mutation-detection && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
bio-ctdna-mutation-detection
GitHub stars
1.2k
Used in
1 other repo
Token cost
~5.1k tokens
SKILL.md length
2,441 words
Files
3
Skills in repo
559
Repo updated
First seen
Licence
MIT

At a glance

Detects somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise problem set by error suppression and molecules sampled, not by the choice of caller.

  • Tracking tumor mutations from plasma cfDNA
  • SKILL.md covers Version Compatibility, The Single Most Important…, Methods Landscape and Decision Tree by Scenario, plus 9 more sections
  • Runs Python scripts from its folder; calls pip
  • Setting a VAF threshold

What it does

Bio Ctdna Mutation Detection is an agent skill from GPTomics/bioSkills. Detects somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise problem set by error suppression and molecules sampled, not by the choice of caller. Distinguishes de novo CALLING (scanning a panel for unknown variants, bounded by per-locus error and multiple testing) from tumor-informed DETECTION (tracking a pre-specified variant set, where panel integration reaches single-ppm). Covers VarDict and Mutect2 for de novo calling, UMI-aware callers, and a pysam-based…

Its SKILL.md is about 5.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `examples/detect_ctdna_mutations.py` and `usage-guide.md`).

It sits in Research & Science. It works with pysam. The repository describes itself as: a set of SKILLS.md for doing bioinformatics with agents like claude code. The licence is MIT.

When your agent uses it

  • Tracking tumor mutations from plasma cfDNA
  • Setting a VAF threshold
  • Deciding whether a low-VAF call is tumor versus CHIP

Example prompts

  • “Use the bio-ctdna-mutation-detection skill to detect somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise…”
  • “/bio-ctdna-mutation-detection”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit d91ed3d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships script files (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Bio Ctdna Mutation Detection loads about 5.1k tokens when it runs. Until then it costs about 206 tokens; SKILL.md has 2,441 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~206
When it runs · the whole SKILL.md, loaded when a task matches
~5.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from GPTomics/bioSkills at commit d91ed3d, republished under its MIT licence (© GPTomics). 2,441 words, ~5,104 tokens.

Download SKILL.mdSave it as .claude/skills/bio-ctdna-mutation-detection/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
bio-ctdna-mutation-detection
description
Detects somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise problem set by error suppression and molecules sampled, not by the choice of caller. Distinguishes de novo CALLING (scanning a panel for unknown variants, bounded by per-locus error and multiple testing) from tumor-informed DETECTION (tracking a pre-specified variant set, where panel integration reaches single-ppm). Covers VarDict and Mutect2 for de novo calling, UMI-aware callers, and a pysam-based known-variant VAF tracker, with matched-WBC subtraction as the mandatory defense against clonal hematopoiesis (the dominant false positive). Use when calling or tracking tumor mutations from plasma cfDNA, setting a VAF threshold, or deciding whether a low-VAF call is tumor versus CHIP.
tool_type
mixed
primary_tool
VarDict

Version Compatibility

Reference examples tested with: pysam 0.22+, pandas 2.2+, VarDictJava 1.8+, GATK 4.5+, Ensembl VEP 111+

Before using code patterns, verify installed versions match. If versions differ:

  • Python: pip show <package> then help(module.function) to check signatures
  • CLI: <tool> --version then <tool> --help to confirm flags

If code throws ImportError, AttributeError, or TypeError, introspect the installed package and adapt the example to match the actual API rather than retrying.

Notes specific to this skill: VarDict's -c -S -E -g are 1-based BED COLUMN INDICES, not genomic coordinates; var2vcf_valid.pl's -E suppresses the END tag (opposite meaning to VarDict's -E). VEP gnomAD flags are --af_gnomade (exomes)/--af_gnomadg (genomes); the bare --af_gnomad is a legacy alias that returns only exome AF, so prefer the explicit forms.

ctDNA Mutation Detection

"Detect mutations in my cfDNA sample" -> Either scan a panel for unknown low-VAF somatic variants (de novo calling) or quantify a pre-specified mutation set across samples (tumor-informed tracking) — two different statistical problems.

  • CLI: vardict-java | teststrandbias.R | var2vcf_valid.pl for de novo low-VAF calling on a consensus BAM
  • CLI: gatk Mutect2 with the read-orientation model for de novo calling with artifact filtering
  • Python: pysam pileup of ref/alt counts at fixed loci for known-variant tracking (MRD)

The Single Most Important Modern Insight -- detection is a signal-vs-noise problem, and tracking is not the same problem as calling

Low-VAF ctDNA detection is set by two limits that no caller can overcome: the per-base error floor (raw Illumina ~1e-3 caps naive VAF detection near 0.5-1%) and the number of tumor molecules physically present in the tube (1 ng cfDNA ~= 303 haploid genome-equivalents; at 0.01% VAF in 10 ng the expected mutant count is ~0.3 copies — there is nothing to detect at any depth). The achieved limit of detection is the worse of the two. Error suppression (UMI consensus -> ~1e-5, duplex -> <1e-7) and input mass move the floor; swapping VarDict for Mutect2 does not.

Critically, de novo CALLING and known-variant DETECTION are different statistical problems. De novo calling scans every covered position for an unknown alt and pays a multiple-testing tax across 1e5-1e6 loci, so per-locus thresholds must be stringent (practical LoD ~0.1-0.5% on UMI consensus). Tumor-informed detection tests ONE hypothesis — "is tumor present?" — by integrating signal across a pre-specified set of N patient-specific loci; the multiple-testing penalty collapses and per-locus signal that is individually indistinguishable from noise sums into a confident panel-level call. This is why per-locus LoD is poor while panel-integrated LoD reaches single-ppm. Conflating the two is the most common conceptual error in the field.

Methods Landscape

MethodClassCitationRoleWhen
VarDict / vardict-javade novo callerAstraZeneca-NGSsensitive low-VAF amplicon/capture calling with explicit strand-bias testde novo panel calling on a UMI-consensus BAM
Mutect2 (tumor-only)de novo callerGATKlocal-assembly somatic caller + learned orientation-bias artifact modelde novo calling needing FFPE/OxoG artifact filtering, PoN, germline resource
umi-varcalUMI-aware callerSater 2020 Bioinformatics 36(9):2718own UMI-aware pileup + per-position Poisson test against local backgroundUMI-tagged BAM where a consensus-aware caller is wanted (floor ~0.3%)
CAPP-Seq / iDEStumor-informed integrationNewman 2014 Nat Med 20:548; 2016 Nat Biotechnol 34:547hybrid-capture deep panel + molecular barcoding + in-silico background polishingde novo ctDNA to ~0.02%; iDES stacks ~15x error suppression
INVARtumor-informed integrationWan 2020 Sci Transl Med 12:eaaz8084integrate variant reads across 100s-1000s patient loci, background-weightedMRD/monitoring with tumor WES; quantifies to ~1e-5, best ~2.5 ppm
MRDetecttumor-informed integrationZviran 2020 Nat Med 26:1114shallow WGS vs patient SNV compendium, read-level SVM noise modelMRD trading depth for breadth (~35x WGS, thousands of SNVs); ~1e-5
PhasED-Seqtumor-informed integrationKurtz 2021 Nat Biotechnol 39:1537enrich phased (co-occurring) variants to suppress single-molecule errorsub-ppm MRD where phased variants are available

Per-locus LoD for any of these is error- and sampling-limited (~0.1-0.5%); the tumor-informed methods reach ppm only by integrating across a known, large, patient-specific variant set. Methodology evolves — verify current best practice against each tool's live docs before committing to one.

Decision Tree by Scenario

ScenarioRecommendedWhy
Tumor tissue available, MRD/monitoring of a known cancerTumor-informed tracking (INVAR/MRDetect/Signatera-class), or the pysam tracker below for a fixed listIntegrating across N pre-specified loci is the only route to ppm; CHIP excluded by construction (CHIP variants are not on the tumor list)
No tumor tissue, screening/discoveryde novo panel calling (VarDict or Mutect2) + matched WBCMust scan for unknown variants; CHIP subtraction is mandatory or most calls are not tumor
VAF regime > 1%Any standard caller on a deduplicated BAMAbove the raw error floor; consensus not strictly required
VAF regime 0.1-1%UMI single-strand consensus + VarDict/umi-varcal/Mutect2Below the raw 1e-3 floor; consensus needed to recover real signal from error
VAF regime < 0.1%Duplex consensus + tumor-informed integrationSingle-strand consensus cannot remove one-strand deamination/oxidation; only duplex + panel integration reaches this regime
No matched WBC availableDo NOT report de novo calls as somatic-tumorWithout WBC subtraction, CHIP (the majority of non-germline cfDNA variants) is indistinguishable from tumor

CHIP -- the dominant false positive, not background

Clonal hematopoiesis of indeterminate potential (CHIP) is the single largest source of false-positive somatic calls in plasma, and it is the null hypothesis for any low-VAF cfDNA variant. Razavi 2019 sequenced cfDNA with matched white-blood-cell DNA (508 genes, >60,000x) and found that 53.2% of non-germline cfDNA variants in cancer patients and 81.6% in non-cancer controls had features consistent with clonal hematopoiesis; only ~24.4% of cfDNA somatic variants in patients were also in the matched tumor (the remainder split between white-cell CHIP and variants of uncertain origin). These are bona fide somatic mutations — in cancer genes — that come from lysed leukocytes, not tumor. No error-suppression tier removes them because they are not errors.

The biology: CHIP arises in hematopoietic stem cells and rises steeply with age (Jaiswal 2014: ~10% prevalence over age 70), enriched for PPM1D/TP53/CHEK2 clones after prior chemo/radiation — exactly the monitored population. The canonical genes are DNMT3A, TET2, ASXL1 (the big three), then PPM1D, TP53, JAK2, SF3B1, SRSF2, GNB1, GNAS, CBL, ATM, CHEK2. TP53 and ATM are both CHIP genes and bona fide tumor suppressors, so a low-VAF TP53 cfDNA call is the ambiguous case par excellence.

The only reliable filter is matched buffy-coat/WBC subtraction: sequence the WBC fraction of the same draw at comparable depth and remove any cfDNA variant also present in WBC. gnomAD filtering removes germline only — CHIP variants are somatic and absent from germline databases, so they sail straight through. A canonical-CHIP-gene list (the example's CHIP_GENES) is a heuristic flag for extra scrutiny, NOT a substitute for WBC subtraction. See analytical-validation for the LoB/LoD statistics that quantify how confidently a subtracted call clears background.

De Novo Calling with VarDict

Goal: Scan a target panel for unknown low-VAF somatic variants on a UMI-consensus BAM.

Approach: Run vardict-java with a lowered -f, pipe through the strand-bias test, then convert to VCF — matching -f across both stages so the threshold is not silently re-applied.

bash
AF_THR=0.005   # 0.5% — practical UMI-consensus de novo floor; below this approaches the per-base error floor
vardict-java -G ref.fa -f $AF_THR -N sample -b consensus.bam \
  -c 1 -S 2 -E 3 -g 4 targets.bed | \
  teststrandbias.R | \
  var2vcf_valid.pl -N sample -E -f $AF_THR > sample.vcf

Key flags: -G indexed reference; -f min VAF (VarDict default 0.01); -N sample name; -b BAM. -c 1 -S 2 -E 3 -g 4 are the 1-based BED COLUMN INDICES for chrom/start/end/gene in a standard 4-column BED — they are column positions, not genomic values. On var2vcf_valid.pl, -E means "do NOT print the END tag" (unrelated to VarDict's -E); its -f default is 0.02, so set it to match. For PCR/amplicon data add -P 0 (positional std is expected to be ~0). For paired tumor/normal use the testsomatic.R | var2vcf_paired.pl path instead.

De Novo Calling with Mutect2 and the Orientation-Bias Model

Goal: Call de novo somatic variants while filtering FFPE-deamination (C>T) and OxoG (G>T) strand-biased artifacts that dominate low-VAF false positives.

Approach: Collect F1R2/F2R1 counts during calling, learn the orientation-bias prior, then apply it during filtering alongside a panel of normals and germline resource.

bash
gatk Mutect2 -R ref.fa -I consensus.bam --f1r2-tar-gz f1r2.tar.gz \
  --germline-resource af-only-gnomad.vcf.gz --panel-of-normals pon.vcf.gz \
  -O unfiltered.vcf.gz
gatk LearnReadOrientationModel -I f1r2.tar.gz -O read-orientation-model.tar.gz
gatk FilterMutectCalls -R ref.fa -V unfiltered.vcf.gz \
  --ob-priors read-orientation-model.tar.gz -O filtered.vcf.gz

Mutect2 is run tumor-only here (no normal sample arg); the orientation model is the load-bearing low-VAF filter. At true ctDNA VAFs Mutect2 is underpowered relative to a dedicated UMI/duplex + background-polishing pipeline, and local assembly can miss extremely low-AF alt support — it is a reasonable de novo caller on consensus reads with the orientation model + PoN (and ideally a matched normal, not shown in this tumor-only command), not a substitute for tumor-informed integration at ppm.

Track Known Mutations Across Serial Samples

Goal: Quantify the VAF of a pre-specified mutation set at fixed loci for MRD monitoring — the detection (not calling) problem.

Approach: For each target mutation, pileup reads at the position, count ref/alt/other alleles, and compute VAF with depth; aggregate across loci as the panel-level detection signal. The single-base pileup below tracks SNVs only — indel reporters (e.g. EGFR exon-19 deletions) need read.indel/CIGAR-aware counting; a single-base comparison silently scores every indel read as other and reports the locus as cleared.

python
import pysam

def track_known_variants(bam_file, variants):
    '''Pileup ref/alt counts at fixed (chrom, pos, ref, alt) SNV loci; pos is 1-based.
    SNVs only - indel reporters need read.indel/CIGAR handling, not a single-base compare.'''
    bam = pysam.AlignmentFile(bam_file, 'rb')
    rows = []
    for chrom, pos, ref, alt in variants:
        counts = {'ref': 0, 'alt': 0, 'other': 0}
        for col in bam.pileup(chrom, pos - 1, pos, truncate=True):
            for read in col.pileups:
                if read.is_del or read.is_refskip:
                    continue
                base = read.alignment.query_sequence[read.query_position]
                counts['alt' if base == alt else 'ref' if base == ref else 'other'] += 1
        depth = sum(counts.values())
        rows.append({'chrom': chrom, 'pos': pos, 'ref': ref, 'alt': alt,
                     'depth': depth, 'alt_count': counts['alt'],
                     'vaf': counts['alt'] / depth if depth else 0.0})
    bam.close()
    return rows

Annotate calls for interpretation with Ensembl VEP (--cache --offline --fasta --vcf --everything); the gnomAD allele-frequency flags are --af_gnomade (exomes) and --af_gnomadg (genomes) — the bare --af_gnomad is a legacy alias returning only exome AF. gnomAD presence separates germline; only WBC presence separates CHIP.

Show full SKILL.md (969 more words)Show less

Per-Method Failure Modes

CHIP misclassified as tumor

Trigger: de novo calling without matched WBC. Mechanism: leukocyte-derived clonal somatic variants in cancer genes look identical to tumor signal and pass gnomAD filtering. Symptom: low-VAF calls in DNMT3A/TET2/TP53; "tumor" mutations not in the matched tissue. Fix: subtract matched buffy-coat/WBC genotype; never report somatic-tumor without it.

Strand-biased / deamination artifacts at low VAF

Trigger: FFPE-style C>T or oxidative G>T at VAF near the floor. Mechanism: damage on one template strand is inherited by every PCR copy, so single-strand UMI consensus votes unanimously for the artifact. Symptom: alt support concentrated on one strand. Fix: VarDict strand-bias test or Mutect2 orientation model; for sub-0.1% require duplex consensus.

Calling below the error floor

Trigger: lowering -f to e.g. 0.001 on a non-consensus BAM. Mechanism: the raw ~1e-3 error rate manufactures alt reads at that frequency. Symptom: a flood of low-VAF calls scaling with depth. Fix: do consensus upstream; do not set a VAF threshold below the demonstrated error floor of the input.

Germline-vs-somatic confusion at low coverage

Trigger: classifying by VAF alone when depth is low. Mechanism: a true 50% het reads 3/12 = 0.25 by chance. Symptom: germline hets mislabeled subclonal somatic. Fix: gnomAD + matched-WBC presence (germline ~0.5 in WBC; CHIP at clone VAF; tumor-only absent from WBC).

Per-locus LoD quoted as the assay LoD

Trigger: reporting a single-variant sensitivity for a multi-locus tracking assay (or vice versa). Mechanism: panel-integrated LoD is orders of magnitude below per-locus LoD. Symptom: a "0.1%" claim that does not match observed ppm-level tracking. Fix: state per-locus vs panel-integrated explicitly; see analytical-validation.

Quantitative Thresholds

ThresholdSourceRationale
Raw Illumina error ~1e-3 caps naive VAF near 0.5-1%Schmitt 2012 PNAS 109:14508Per-base miscall rate sets the per-locus VAF floor; alt support below it is mostly error
UMI single-strand consensus -> ~1e-5; duplex -> <1e-7Schmitt 2012; Newman 2016 Nat Biotechnol 34:547Family consensus erases PCR/sequencing error; duplex strand concordance also catches one-strand damage
CAPP-Seq de novo LoD ~0.02% at 96% specificityNewman 2014 Nat Med 20:548Deep hybrid-capture + reporter set; demonstrates the de novo panel floor
iDES ~15x error suppression (UMI ~3x x polishing ~3x)Newman 2016 Nat Biotechnol 34:547Molecular consensus and in-silico background polishing are orthogonal and stack
INVAR quantifies to ~1e-5, detects to ~2.5 ppmWan 2020 Sci Transl Med 12:eaaz8084Integrating variant reads across 100s-1000s of patient loci collapses multiple testing
MRDetect ~1e-5 tumor fraction at ~35x WGS, 95% specZviran 2020 Nat Med 26:1114Breadth (thousands of SNVs) + read-level SVM (~14.4x error reduction) substitutes for depth
CHIP = 53.2% (cancer pts) / 81.6% (controls) of cfDNA variantsRazavi 2019 Nat Med 25:1928Most non-germline cfDNA variants are not tumor; matched WBC is mandatory
Depth >= 1000-5000x unique consensus for panelscommunity / Phallen 2017 Sci Transl Med 9:eaan2415 (~30,000x)Detecting <1% VAF needs enough unique molecules sampled at each locus
~303 genome-equivalents per ng cfDNA (3.3 pg/haploid)standard constantInput mass sets a hard Poisson ceiling on detectable VAF independent of sequencing
LoB / LoD / LoD95 per CLSI EP17CLSI EP17-A2A bare VAF without input mass + replicate detection rate is not a sensitivity spec

Common Errors

Error / symptomCauseSolution
Flood of low-VAF calls scaling with depth-f set below the input's error floor on non-consensus readsDo UMI/duplex consensus first; keep -f >= demonstrated floor
VarDict emits nothing or wrong regions-c -S -E -g read as genomic valuesThey are 1-based BED column indices; use -c 1 -S 2 -E 3 -g 4 for a 4-column BED
var2vcf re-filters away VarDict callsvar2vcf_valid.pl -f default 0.02 mismatchedSet var2vcf -f to match VarDict's -f
Real amplicon calls dropped as positional artifactsvar2vcf -P (filter pstd=0) on by defaultAdd -P 0 for PCR/amplicon data
"Tumor" variants absent from matched tissueCHIP not subtractedSequence and subtract matched WBC; flag CHIP-gene hits
--af_gnomad returns only exome AFbare flag is a legacy exome-only aliasUse --af_gnomade (exomes) / --af_gnomadg (genomes)
ppm "LoD" not reproducibleper-locus LoD quoted for a tracking assayReport panel-integrated LoD with input mass and LoD95

References

  • Schmitt MW, Kennedy SR, Salk JJ, Fox EJ, Hiatt JB, Loeb LA. 2012. Detection of ultra-rare mutations by next-generation sequencing. Proc Natl Acad Sci USA 109(36):14508-14513. — Duplex Sequencing; DCS error <1e-7.
  • Newman AM, Bratman SV, To J, et al. 2014. An ultrasensitive method for quantitating circulating tumor DNA with broad patient coverage. Nat Med 20(5):548-554. — CAPP-Seq; ~0.02% LoD.
  • Newman AM, Lovejoy AF, Klass DM, et al. 2016. Integrated digital error suppression for improved detection of circulating tumor DNA. Nat Biotechnol 34(5):547-555. — iDES; UMI + background polishing.
  • Phallen J, Sausen M, Adleff V, et al. 2017. Direct detection of early-stage cancers using circulating tumor DNA. Sci Transl Med 9(403):eaan2415. — TEC-Seq deep panel.
  • Razavi P, Li BT, Brown DN, et al. 2019. High-intensity sequencing reveals the sources of plasma circulating cell-free DNA variants. Nat Med 25(12):1928-1937. — CHIP is the majority of cfDNA variants; matched WBC.
  • Wan JCM, Heider K, Gale D, et al. 2020. ctDNA monitoring using patient-specific sequencing and integration of variant reads. Sci Transl Med 12(548):eaaz8084. — INVAR; integration to ~2.5 ppm.
  • Zviran A, Schulman RC, Shah M, et al. 2020. Genome-wide cell-free DNA mutational integration enables ultra-sensitive cancer monitoring. Nat Med 26(7):1114-1124. — MRDetect; shallow WGS + read SVM.
  • Kurtz DM, Soo J, Co Ting Keh L, et al. 2021. Enhanced detection of minimal residual disease by targeted sequencing of phased variants in circulating tumor DNA. Nat Biotechnol 39(12):1537-1547. — PhasED-Seq; phased-variant enrichment.
  • Sater V, Viailly P-J, Lecroq T, et al. 2020. UMI-VarCal: a new UMI-based variant caller that efficiently improves low-frequency variant detection in paired-end sequencing NGS libraries. Bioinformatics 36(9):2718-2724. — UMI-aware pileup + per-position Poisson test.
  • cfdna-preprocessing - UMI/duplex consensus input that sets the error floor
  • analytical-validation - LoD/LoB and the panel-integration math behind detection
  • longitudinal-monitoring - track detected variants across serial samples
  • tumor-fraction-estimation - orthogonal burden estimate to cross-check
  • variant-calling/variant-calling - general somatic calling principles
  • clinical-databases/variant-prioritization - clinical annotation and interpretation

© GPTomics, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files in liquid-biopsy/ctdna-mutation-detection of GPTomics/bioSkills.

  • SKILL.md
  • examples/detect_ctdna_mutations.py
  • usage-guide.md

Open the folder on GitHubat commit d91ed3d

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in GPTomics/bioSkills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Bio Ctdna Mutation Detection next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Bio Ctdna Mutation Detection compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Bio Ctdna Mutation Detection this skillGPTomics/bioSkills1.2k1 repos~5.1kAutomated safety check: PassMIT
Pysamdavila7/claude-code-templates32k10 repos~2.5kAutomated safety check: PassMIT
PysamK-Dense-AI/scientific-agent-skills48k1 repos~3.4kAutomated safety check: NotesMIT
Omics ToolsDrugClaw/DrugClaw125—~1.1kAutomated safety check: PassApache-2.0
Tooluniverse Epigenomicswu-yc/LabClaw1.1k2 repos~14kAutomated safety check: PassNone
Biopython Sequence Analysisjaechang-hits/SciAgent-Skills3711 repos~8.5kAutomated safety check: PassBSD-3-Clause

Similar skills

  • Pysam

    davila7/claude-code-templates

    Genomic file toolkit. An agent skill from davila7/claude-code-templates.

    32k GitHub starsUsed in 10 repos~2.5k tokens
    Research & ScienceAuto-check passed
  • Pysam

    K-Dense-AI/scientific-agent-skills

    Provides Python/HTSlib workflows for genomic files. An agent skill from K-Dense-AI/scientific-agent-skills.

    48k GitHub starsUsed in 1 repo~3.4k tokens
    Research & ScienceAuto-check: notes
  • Omics Tools

    DrugClaw/DrugClaw

    Omics and single-cell workflow guide for AnnData, Scanpy-style dataset profiling, PyDESeq2-oriented count checks, pysam alignment inspection, and pyOpenMS mass-spectrometry summaries.

    125 GitHub stars~1.1k tokensUpdated 6 mo ago
    Research & ScienceAuto-check passed
  • Production-ready genomics and epigenomics data processing for BixBench questions.

    1.1k GitHub starsUsed in 2 repos~14k tokens
    Research & ScienceAuto-check passed
  • Biopython Sequence Analysis

    jaechang-hits/SciAgent-Skills

    Biopython sequence analysis: parse FASTA/FASTQ/GenBank/GFF (SeqIO), NCBI Entrez (esearch/efetch/elink), remote/local BLAST, pairwise/MSA alignment (PairwiseAligner, MUSCLE/ClustalW), phylogenetic…

    371 GitHub starsUsed in 1 repo~8.5k tokens
    Research & ScienceAuto-check passed
  • Samtools Bam Processing

    jaechang-hits/SciAgent-Skills

    CLI toolkit for SAM/BAM/CRAM: sort, index, convert, filter, QC alignments.

    371 GitHub starsUsed in 1 repo~4.1k tokens
    Research & ScienceAuto-check passed

More from GPTomics/bioSkills

All 559 skills in this repo
  • Bio Alignment Io

    GPTomics/bioSkills

    Read, write, and convert multiple sequence alignment files using Biopython Bio.AlignIO.

    1.2k GitHub starsUsed in 3 repos~4.9k tokens
    Auto-check passed
  • bioSkills Installer

    GPTomics/bioSkills

    Installs the bioSkills collection of 425 bioinformatics skills in one step, or only chosen categories, so sequencing, RNA-seq, single-cell and variant tasks get specialized help.

    1.2k GitHub starsUsed in 1 repo~789 tokens
    Auto-check passed
  • Bio Write Sequences

    GPTomics/bioSkills

    Write biological sequences to files (FASTA, FASTQ, GenBank, EMBL) using Biopython Bio.SeqIO.

    1.2k GitHub starsUsed in 3 repos~2.1k tokens
    Auto-check passed
  • Amplicon Primer Clipping

    GPTomics/bioSkills

    Soft- or hard-clips PCR primer footprints from aligned amplicon BAMs so primer bases stop masquerading as confirmed reference sequence.

    1.2k GitHub starsUsed in 2 repos~2.2k tokens
    Auto-check passed
  • Filters BAM alignments by FLAG bits, mapping quality and regions with samtools view or pysam, with recipes for common keep and drop cases.

    1.2k GitHub starsUsed in 2 repos~3.6k tokens
    Auto-check passed
  • Bio Alignment Indexing

    GPTomics/bioSkills

    Create and use BAI/CSI indices for BAM/CRAM files using samtools and pysam.

    1.2k GitHub starsUsed in 2 repos~2.4k tokens
    Auto-check passed

Works with

Questions about Bio Ctdna Mutation Detection

What does Bio Ctdna Mutation Detection do?

Detects somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise problem set by error suppression and molecules sampled, not by the choice of caller. Bio Ctdna Mutation Detection is an agent skill from GPTomics/bioSkills. Detects somatic mutations in circulating tumor DNA, treating low-VAF detection as a signal-versus-noise problem set by error suppression and molecules sampled, not by the choice of caller.

When should I use Bio Ctdna Mutation Detection?

Bio Ctdna Mutation Detection fits situations like: tracking tumor mutations from plasma cfDNA; setting a VAF threshold; deciding whether a low-VAF call is tumor versus CHIP.

How do I install Bio Ctdna Mutation Detection in Claude Code?

Run `npx skills add GPTomics/bioSkills --skill bio-ctdna-mutation-detection -a claude-code`. Or copy the skill folder (liquid-biopsy/ctdna-mutation-detection in GPTomics/bioSkills) into .claude/skills/bio-ctdna-mutation-detection in your project. Claude Code loads it when a task matches its description.

How do I install Bio Ctdna Mutation Detection in Codex?

Run `npx skills add GPTomics/bioSkills --skill bio-ctdna-mutation-detection -a codex`. Or copy the skill folder (liquid-biopsy/ctdna-mutation-detection in GPTomics/bioSkills) into .agents/skills/bio-ctdna-mutation-detection in your project. Codex loads it when a task matches its description.

Can I use Bio Ctdna Mutation Detection in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add GPTomics/bioSkills --skill bio-ctdna-mutation-detection -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/bio-ctdna-mutation-detection, .gemini/skills/bio-ctdna-mutation-detection, .github/skills/bio-ctdna-mutation-detection and .opencode/skills/bio-ctdna-mutation-detection in your project.

What does Bio Ctdna Mutation Detection need to run?

Going by SKILL.md and its folder, Bio Ctdna Mutation Detection needs Python for the scripts in its folder and the command-line tools its instructions call (pip). Our summary lists: Python 3.

Does Bio Ctdna Mutation Detection access the network?

SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Bio Ctdna Mutation Detection safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Bio Ctdna Mutation Detection use?

Bio Ctdna Mutation Detection is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Bio Ctdna Mutation Detection use?

About 5.1k tokens (SKILL.md is roughly 20k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Bio Ctdna Mutation Detection?

Skills that share tags, products or a category with Bio Ctdna Mutation Detection: Pysam (davila7/claude-code-templates, 32k stars), Pysam (K-Dense-AI/scientific-agent-skills, 48k stars), Omics Tools (DrugClaw/DrugClaw, 125 stars) and Tooluniverse Epigenomics (wu-yc/LabClaw, 1.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Bio Ctdna Mutation Detection?

GPTomics (a GitHub organization) maintains it in GPTomics/bioSkills, which has 1,217 GitHub stars. The repository holds 559 skills in this directory. The repository was last updated on August 15, 2026.

Source: GPTomics/bioSkills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.