LaminDB Biological Data Management
davila7/claude-code-templates
Manages biological datasets with LaminDB: versioned artifacts, run lineage, ontology-based annotation, schema validation and links to workflow managers and ML tools.
Wrapper skill for running nf-core/rnaseq bulk RNA-seq preprocessing from FASTQ or BAM inputs with strict preflight, reproducibility outputs, and downstream handoff to ClawBio bulk RNA-seq DE skills.
$ npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install ClawBio/ClawBio nfcore-rnaseq-wrapper --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/ClawBio/ClawBio.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/nfcore-rnaseq-wrapper .claude/skills/nfcore-rnaseq-wrapper && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "nfcore-rnaseq-wrapper" agent skill from https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapper into .claude/skills/nfcore-rnaseq-wrapper/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "nfcore-rnaseq-wrapper", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapperType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install ClawBio/ClawBio nfcore-rnaseq-wrapper --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ClawBio/ClawBio.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/nfcore-rnaseq-wrapper .agents/skills/nfcore-rnaseq-wrapper && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "nfcore-rnaseq-wrapper" agent skill from https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapper into .agents/skills/nfcore-rnaseq-wrapper/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "nfcore-rnaseq-wrapper", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install ClawBio/ClawBio nfcore-rnaseq-wrapper --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ClawBio/ClawBio.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/nfcore-rnaseq-wrapper .cursor/skills/nfcore-rnaseq-wrapper && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "nfcore-rnaseq-wrapper" agent skill from https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapper into .cursor/skills/nfcore-rnaseq-wrapper/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "nfcore-rnaseq-wrapper", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/ClawBio/ClawBio.git --path skills/nfcore-rnaseq-wrapper--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install ClawBio/ClawBio nfcore-rnaseq-wrapper --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ClawBio/ClawBio.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/nfcore-rnaseq-wrapper .gemini/skills/nfcore-rnaseq-wrapper && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "nfcore-rnaseq-wrapper" agent skill from https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapper into .gemini/skills/nfcore-rnaseq-wrapper/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "nfcore-rnaseq-wrapper", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install ClawBio/ClawBio nfcore-rnaseq-wrapperInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/ClawBio/ClawBio.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/nfcore-rnaseq-wrapper .github/skills/nfcore-rnaseq-wrapper && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "nfcore-rnaseq-wrapper" agent skill from https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapper into .github/skills/nfcore-rnaseq-wrapper/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "nfcore-rnaseq-wrapper", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install ClawBio/ClawBio nfcore-rnaseq-wrapper --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ClawBio/ClawBio.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/nfcore-rnaseq-wrapper .opencode/skills/nfcore-rnaseq-wrapper && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "nfcore-rnaseq-wrapper" agent skill from https://github.com/ClawBio/ClawBio/tree/main/skills/nfcore-rnaseq-wrapper into .opencode/skills/nfcore-rnaseq-wrapper/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "nfcore-rnaseq-wrapper", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
nfcore-rnaseq-wrapperWrapper skill for running nf-core/rnaseq bulk RNA-seq preprocessing from FASTQ or BAM inputs with strict preflight, reproducibility outputs, and downstream handoff to ClawBio bulk RNA-seq DE skills.
Nfcore Rnaseq Wrapper is an agent skill from ClawBio/ClawBio. Wrapper skill for running nf-core/rnaseq bulk RNA-seq preprocessing from FASTQ or BAM inputs with strict preflight, reproducibility outputs, and downstream handoff to ClawBio bulk RNA-seq DE skills.
Its SKILL.md is about 8.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 39 other files (for example `CHANGELOG.md`, `README.md` and `_isolated_imports.py`).
It sits in Research & Science, covering Bioinformatics and Reproducible research. It works with Nextflow. The repository describes itself as: 🦖 ClawBio - The first bioinformatics-native AI agent skill library. Local-first. Reproducible. Open. Free. The licence is MIT.
5 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 5e045e3. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships script files (Python, from the files we listed), which the agent can run.
Shell commands in SKILL.md call:
pythondockerpython3From the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
nf-co.regithub.comnextflow.iosalmon.readthedocs.iodaehwankimlab.github.iobowtie-bio.sourceforge.netFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Nfcore Rnaseq Wrapper loads about 8.9k tokens when it runs. Until then it costs about 55 tokens; SKILL.md has 3,547 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from ClawBio/ClawBio at commit 5e045e3, republished under its MIT licence (© ClawBio). 3,547 words, ~8,922 tokens.
.claude/skills/nfcore-rnaseq-wrapper/SKILL.md (or your agent's skills folder). This skill also uses 37 other files; get the full folder from GitHub.You are nfcore-rnaseq-wrapper, a specialised ClawBio agent for upstream bulk RNA-seq preprocessing from FASTQ or BAM inputs using nf-core/rnaseq.
Fire when:
nf-core/rnaseqDo NOT fire when:
rnaseq-de.h5ad -> route to nfcore-scrnaseq-wrapperscrna-orchestratorOne skill, one task: run upstream bulk RNA-seq preprocessing through nf-core/rnaseq and produce count-matrix handoff artifacts for downstream ClawBio skills.
This skill does not perform differential expression. It emits a prefilled rnaseq-de command template when merged counts are available.
nf-core/rnaseq v3.26.0 through -params-file with deterministic work/result directories.commands.sh, params.yaml, manifest.json, checksums, environment.yml, and seven provenance JSON files.python clawbio.py run rnaseq --counts ... when a merged count matrix is available.--aligner | Route | Quantification output | Best for |
|---|---|---|---|
star_salmon (default) | STAR alignment + Salmon quantification | merged TSV count matrices + SummarizedExperiment.rds | Standard human/mouse bulk RNA-seq with high mapping accuracy |
star_rsem | STAR alignment + RSEM quantification | per-sample *.genes.results + merged matrix + RDS | Encode-style isoform-level analyses |
hisat2 | HISAT2 alignment only (no quantification) | BAM only — handoff_available=false unless --pseudo-aligner is also set | Alignment-only workflows; add --pseudo-aligner salmon to re-enable downstream DE handoff |
bowtie2_salmon | Bowtie2 alignment + Salmon quantification | merged TSV count matrices + RDS | Prokaryotic transcriptomes (combine with --prokaryotic) |
A pseudo-aligner (--pseudo-aligner salmon or --pseudo-aligner kallisto) runs alongside
--aligner unless paired with --skip-alignment. Each route may use either --genome <iGenomes>
(optionally with additive annotation/transcriptome overrides such as --gtf or --gff,
--additional-fasta, --transcript-fasta, --gene-bed, --splicesites, --salmon-index, or
--kallisto-index) or a fully explicit --fasta/--gtf(/--gff) reference plus optional
pre-built --*-index paths. You may not provide both --genome and your own genome --fasta
or a genome-level index (--star-index/--rsem-index/--hisat2-index/--bowtie2-index).
If both --gtf and --gff are supplied, the wrapper keeps --gtf and drops --gff with a
warning — matching nf-core/rnaseq, which uses the GTF and ignores the GFF when both are given.
For new analyses nf-core/rnaseq recommends supplying explicit --fasta/--gtf directly; the
iGenomes --genome catalogue is supported here for legacy compatibility and convenience.
| Format | Extension | Required Fields | Example |
|---|---|---|---|
| Samplesheet | .csv | sample, fastq_1, strandedness; optional fastq_2 | samplesheet.csv |
| BAM reprocessing samplesheet | .csv | sample, fastq_1, strandedness, plus genome_bam and/or transcriptome_bam; use with --skip-alignment | samplesheet_with_bams.csv |
| Demo mode | n/a | none | python clawbio.py run rnaseq-pipeline --demo |
../rnaseq, or remote nf-core/rnaseq at the pinned version.reproducibility/params.yaml.report.md, result.json, provenance JSON, checksums, and replay commands.rnaseq-de command template using preferred_counts_tsv.# Preflight only; no Nextflow execution
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_check --check \
--genome GRCh38
# Demo mode using upstream test profile
python clawbio.py run rnaseq-pipeline --demo --output ./rnaseq_demo
# STAR + Salmon default route
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_run \
--aligner star_salmon --genome GRCh38
# Explicit FASTA/GTF reference
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_run \
--fasta /refs/genome.fa --gtf /refs/genes.gtf
# RSEM route
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rsem_run \
--aligner star_rsem --genome GRCh38
# Contaminant screening with Kraken2 + Bracken
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_run \
--genome GRCh38 \
--contaminant-screening kraken2_bracken \
--kraken-db /refs/kraken2_db --bracken-precision G
# Auto-handoff to rnaseq-de when all flags are provided
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_run \
--genome GRCh38 --run-downstream \
--metadata metadata.csv --formula "~ batch + condition" \
--contrast "condition,treated,control"
# Prokaryotic transcriptomes via Bowtie2+Salmon
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./prok_run \
--aligner bowtie2_salmon --fasta /refs/genome.fa --gtf /refs/genes.gtf \
--profile docker --prokaryotic
# ARM architecture (Apple M-series, AWS Graviton) — composes -profile docker,arm64
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_arm \
--genome GRCh38 --profile docker --arm
# BAM reprocessing from nf-core samplesheet_with_bams.csv output
python clawbio.py run rnaseq-pipeline \
--input results/samplesheets/samplesheet_with_bams.csv \
--output ./rnaseq_reprocess \
--skip-alignment
# Wrapper runtime controls (parity with scrnaseq/sarek):
# --timeout-hours N wall-clock cap (default 12h; 0 disables for HPC/cloud)
# --work-dir PATH Nextflow work dir (local path or object-store URI; default <output>/upstream/work)
# --nextflow-config / -c / --config extra Nextflow config file(s), repeatable
# --allow-pipeline-version-override run a non-3.26.0 --pipeline-version at your own risk
# --allow-remote-inputs opt in to remote inputs/refs (default local-first)
python clawbio.py run rnaseq-pipeline \
--input samplesheet.csv --output ./rnaseq_run \
--genome GRCh38 --aligner star_salmon \
--timeout-hours 0 --work-dir s3://my-bucket/rnaseq/workpython clawbio.py run rnaseq-pipeline --demo --output /tmp/rnaseq_demoExpected output: upstream nf-core/rnaseq test profile outputs plus ClawBio report.md, result.json, provenance/, and reproducibility/.
The wrapper uses a gated 7-step flow. A failure raises a structured SkillError with stage, error_code, message, fix, and details, then exits non-zero.
Key methods:
s3://, https://, ... — only accepted with --allow-remote-inputs) are passed through unchanged.params.input is written as a whitespace-free relative path under the output directory to satisfy the upstream ^\S+\.csv$ schema.--genome, --fasta --gtf, or --fasta --gff.--genome accepts additive annotation/transcriptome overrides (--gtf or --gff, --gene-bed, --transcript-fasta, --additional-fasta, --splicesites, --salmon-index, --kallisto-index) — matching nf-core/rnaseq — but is mutually exclusive with a genome --fasta or a genome-level index (--star-index/--rsem-index/--hisat2-index/--bowtie2-index).--gtf and --gff together are not rejected: nf-core/rnaseq uses the GTF and ignores the GFF when both are given, so the wrapper drops --gff (with a warning) and proceeds with --gtf, matching upstream in every reference mode (--genome, explicit --fasta, prebuilt indices).handoff_available=false.rnaseq-de.# nf-core/rnaseq Wrapper Report
## Summary
- Aligner: `star_salmon`
- Samples: `5`
## Outputs
- Preferred counts TSV: `/run/upstream/results/star_salmon/salmon.merged.gene_counts_length_scaled.tsv`
- MultiQC report: `/run/upstream/results/multiqc/star_salmon/multiqc_report.html`
## Next Steps
python clawbio.py run rnaseq --counts <preferred_counts_tsv> --metadata <your_metadata.csv> ...output/
├── report.md
├── result.json
├── logs/
├── upstream/
│ ├── results/
│ │ ├── samplesheets/
│ │ │ └── samplesheet_with_bams.csv # only when --save-align-intermeds; use with --skip-alignment for BAM reprocessing
│ │ ├── star_salmon/ # star_salmon aligner outputs
│ │ │ ├── *.markdup.sorted.bam # sorted, deduplicated BAMs (one per sample)
│ │ │ ├── log/ # STAR alignment logs (*.Log.final.out, *.SJ.out.tab)
│ │ │ ├── salmon.merged.*.tsv # merged gene/transcript count matrices
│ │ │ └── salmon.merged.*.rds # SummarizedExperiment objects
│ │ └── ...
│ └── work/
├── provenance/
└── reproducibility/
├── samplesheet.valid.csv # demo run → samplesheet.demo.csv; test profile → samplesheet.noinput.csv
├── params.yaml
├── commands.sh
├── remap_paths.py
├── manifest.json
├── environment.yml
└── checksums.sha256Required
strandedness is required per row and must be auto, forward, reverse, or unstranded..fq, .fastq, .fq.gz, or .fastq.gz (all four are accepted by the nf-core/rnaseq schema). Only the basename must be whitespace-free; parent directory paths may contain spaces.s3://.../https://.... Local paths are normalized and existence-checked; remote URIs are preserved unchanged and left for Nextflow to stage.--genome may be combined with additive annotation/transcriptome overrides (--gtf or --gff, --gene-bed, --transcript-fasta, --additional-fasta, --splicesites, --salmon-index, --kallisto-index) — this matches nf-core/rnaseq and supports common cases such as ERCC spike-ins (--genome GRCh38 --additional-fasta ercc.fa) or overriding the dated iGenomes annotation (--genome GRCh38 --gtf custom.gtf). It is rejected only with a second genome sequence source (--fasta) or a genome-level index (--star-index/--rsem-index/--hisat2-index/--bowtie2-index), which would be ambiguous. If both --gtf and --gff are supplied, --gff is dropped with a warning and --gtf is used (matching nf-core/rnaseq). Names not in the built-in iGenomes catalogue emit a preflight warning but do not block execution — this is expected when using a user-defined genome catalogue (pass it via --nextflow-config my_genomes.config). If you intended an iGenomes entry, check the exact spelling and case (e.g. GRCh38, GRCm38).gencode: true from gene_type/havana_gene markers in the GTF) only inspects local --gtf files; for remote (s3:///https://) GTFs it is skipped silently — pass --gencode explicitly in that case. Autodetection scans only the first 10 feature records of the GTF (gzip is detected case-insensitively, e.g. .gtf.gz and .gtf.GZ); if your GENCODE markers appear later in the file, pass --gencode explicitly.--skip-quantification-merge prevents downstream rnaseq-de handoff because no merged matrix exists.--aligner hisat2 is alignment-only for this handoff contract.--with-umi requires a barcode pattern unless --skip-umi-extract is set. Conversely, UMI options (--umitools-bc-pattern, --umi-dedup-tool, etc.) set without --with-umi are inert — preflight warns so a run is not mistaken for UMI-deduplicated when it is not.--output must be outside the ClawBio source tree. An output directory inside the repository is rejected at preflight with OUTPUT_DIR_INSIDE_REPO, so multi-gigabyte pipeline artifacts never pollute (or get committed to) the checkout — choose a path under your analysis workspace. This matches the nfcore-sarek and nfcore-scrnaseq wrappers./tmp. The wrapper writes a macOS Docker compatibility config whose per-process memory ceiling is derived from host RAM (75% share, floored at 8 GB, capped at 15 GB) and then capped to 90% of the actual Docker VM memory (docker info, when available) so a container process is never OOM-killed by requesting more than the VM has. Its per-process time ceiling tracks --timeout-hours (default 12, floored at 1 h) so raising the wrapper timeout does not leave processes capped at 12 h.--timeout-hours (default 12). Raise it for large cohorts (e.g. --timeout-hours 48) so a long but healthy run is not terminated, or pass --timeout-hours 0 to disable the cap entirely for long HPC/cloud runs whose walltime is enforced by the scheduler (negative values are rejected). On a timeout the wrapper terminates Nextflow's process group, but containers started by the Docker/Singularity daemon are not in that group and may keep running — the timeout error reminds you to check for and remove leftover containers (e.g. docker ps).--fasta/--gtf/--gff/--transcript-fasta/--additional-fasta/--gene-bed) must resolve to a path without whitespace — the nf-core schema pattern ^\S+ rejects spaces. Preflight catches a whitespace-containing resolved path early with a precise REFERENCE_PATH_HAS_WHITESPACE error (mirroring the samplesheet input guard) instead of letting Nextflow abort late. Move or symlink the reference into a space-free directory.--check validates that Nextflow is present but defers the >=25.04.3 version gate to the real run; it emits a warning so a passing check is not mistaken for confirmation of a compatible Nextflow version.upstream/results because the wrapper launches Nextflow with cwd=<output>; the relative path keeps the nf-core ^\S+$ outdir schema valid even when --output contains spaces (common on macOS). This is a deliberate local-first design. Running against cloud executors that require an absolute publish path (e.g. outdir on s3:///gs://) is outside the wrapper's audited surface.--plaintext_email, --max_multiqc_email_size, --monochrome_logs, --trace_report_suffix, --custom_config_*) are intentionally not exposed. Non-parametric runtime settings (executor, resource limits, institutional config) are supplied through --nextflow-config.../rnaseq checkout is auto-detected and used, but its manifest.version must be 3.26.0 (the version this wrapper's validations are pinned to). A different version is rejected unless --allow-pipeline-version-override is passed; an unparseable manifest version is warned, not blocked.--rseqc-modules is validated against the eight nf-core/rnaseq 3.26.0 module names; a typo is rejected at preflight instead of failing later inside Nextflow.--contaminant-screening kraken2/kraken2_bracken requires --kraken-db, and --contaminant-screening sylph requires --sylph-db; local database paths are existence-checked before Nextflow starts, while URI schemes such as s3:// and https:// are passed through for Nextflow to stage. --bracken-precision only applies to kraken2_bracken and is warned (no effect) otherwise.--skip-alignment + --pseudo-aligner salmon/kallisto + --transcript-fasta or a prebuilt --salmon-index/--kallisto-index + --gtf/--gff) is accepted without a genome --fasta. A pseudo-aligner running alongside a genome aligner still requires the genome reference.--fasta: a genome index matching the aligner (--star-index/--hisat2-index/--bowtie2-index, or --rsem-index for star_rsem) plus --gtf/--gff and, for the Salmon routes, a transcript source (--transcript-fasta or --salmon-index) is accepted. A bare genome index without a transcript source (Salmon routes) or without --rsem-index/--fasta (RSEM) is rejected because quantification cannot run.--pseudo-aligner-kmer-size must be an odd integer in 1..31 (Salmon and Kallisto both encode the index k-mer in a 64-bit word, so 31 is their shared hard cap; pipeline default 31). Preflight rejects an even or out-of-range value with INVALID_PRESET_CONFIGURATION instead of letting the pseudo-aligner indexing step crash. Lower it for short reads (<50 bp).--prokaryotic, --rapid-quant, and --arm are profile-modifier flags. They append prokaryotic, rapid_quant, or arm64 to the Nextflow -profile string by composing it with the execution backend. Use --profile docker --prokaryotic (composes -profile docker,prokaryotic). --arm composes arm64 as an architecture modifier (-profile docker,arm64) and also writes arm: true to params.yaml — arm is a real hidden boolean parameter in the nf-core/rnaseq 3.26.0 schema ("Use ARM architecture containers.").sample, fastq_1, strandedness, plus at least one of genome_bam or transcriptome_bam. Use the nf-core-generated samplesheet_with_bams.csv with --skip-alignment. Rows with BAMs and an empty fastq_1 are rejected because they no longer match the audited nf-core/rnaseq 3.26.0 samplesheet contract. Reprocess with the same --aligner used to generate the BAMs: nf-core/rnaseq cannot mix quantifier types between BAM generation and reprocessing (BAMs from star_salmon must be reprocessed with star_salmon, star_rsem with star_rsem). The wrapper defaults to star_salmon, so pass --aligner star_rsem explicitly when reprocessing RSEM BAMs; preflight emits a reminder warning whenever BAM reprocessing is detected. The samplesheet_with_bams.csv you reprocess from is only produced when the original alignment run used --save-align-intermeds — nf-core/rnaseq creates it solely in that case, so add --save-align-intermeds to the run whose BAMs you intend to reprocess later.--ribo-database-manifest is preflight-checked when it is a local path; missing files or directories are rejected before Nextflow starts. URI schemes are preserved unchanged in params.yaml.--use-parabricks-star requires --aligner star_salmon; --use-sentieon-star requires a STAR-based aligner (star_salmon or star_rsem); --use-gpu-ribodetector requires --remove-ribo-rna --ribo-removal-tool ribodetector.rnaseq-de handoff is opt-in via --run-downstream. It launches rnaseq-de only when --run-downstream is set and --metadata, --formula, and --contrast are all provided. With --run-downstream but any of those three missing, only a copy-paste template reproducibility/rnaseq_de_handoff.sh is written. Without --run-downstream (the default, including --demo), no handoff is launched and no template file is written — the report.md "Next Steps" section still shows the suggested rnaseq-de command. --skip-downstream suppresses the template even when --run-downstream is set.--rseqc-modules runs a default set of 7 modules. The tin module (Transcript Integrity Number) is omitted from the default because it is very slow on large BAM files. Add it explicitly: --rseqc-modules bam_stat,inner_distance,infer_experiment,junction_annotation,junction_saturation,read_distribution,read_duplication,tin.--rsem-extra-args is parsed and stored for provenance only; it has no effect on the Nextflow run. nf-core/rnaseq ≥3.14 removed extra_rsem_quant_args from the schema. Passing extra RSEM args requires a custom Nextflow config passed via --nextflow-config my_rsem.config.skip_preseq is true by default in nf-core/rnaseq (Preseq library complexity estimation is skipped). Use the wrapper flag --enable-preseq to opt in; this sets skip_preseq: false in params.yaml. Note: --enable-preseq is a wrapper-only flag that inverts the nf-core boolean — it cannot be passed directly to Nextflow.--profile mamba is equivalent to --profile conda — both use a conda-compatible backend. The wrapper accepts either spelling.--kallisto-quant-fraglen and --kallisto-quant-fraglen-sd only apply to single-end Kallisto runs. Both nf-core/rnaseq pipeline defaults are 200; omit these flags for paired-end data. Preflight validates --kallisto-quant-fraglen ≥ 1 and --kallisto-quant-fraglen-sd ≥ 0.--min-trimmed-reads must be ≥ 0 (pipeline default: 10000). Preflight rejects negative values. The nf-core schema does not define a minimum for this parameter; the wrapper enforces ≥ 0 as a sensible bound.params.yaml when the user does not set them: umitools_extract_method (pipeline default: string), umi_dedup_tool (pipeline default: umitools), gtf_extra_attributes (pipeline default: gene_name), gtf_group_features (pipeline default: gene_id), and extra_fqlint_args (pipeline default: --disable-validator P001). Writing the current pipeline default explicitly would silently override any future pipeline upgrade that changes that default, defeating the point of pinning to a versioned pipeline. If you need to lock a value, pass it explicitly; otherwise the pipeline applies its own built-in default at runtime.test, test_full, test_prokaryotic, test_full_aws, test_full_gcp, test_full_azure, test_gpu) ship with params.input in their profile config and do not require --input. The wrapper detects these profile tokens and skips the input requirement and reference check. test_full* profiles use genome='GRCh37' via iGenomes — the wrapper does not set igenomes_ignore: true (nor aligner, unless you pass --aligner explicitly) for these, letting the profile config own them. --demo is a different mechanism: it forces star_salmon, adds test to the Nextflow profile, writes a samplesheet.demo.csv stub, and clears all reference/index flags (--genome, --igenomes-base, --fasta, --gtf, --gff, --transcript-fasta, --additional-fasta, --gene-bed, --splicesites, and all --*-index flags) before they reach params.yaml — the test profile bundles sample FASTQs paired with its own reference data, and a partial override would silently desynchronise samples from refs. Self-contained test profile runs produce samplesheet.noinput.csv instead so provenance audits can distinguish them. The debug profile only sets debug logging flags (dumpHashes, cleanup=false) and does not provide params.input — it still requires --input.--demo requires network access. It runs the upstream nf-core -profile test, whose sample FASTQs and reference FASTA/GTF are fetched from remote GitHub URLs (nf-core's design — the wrapper does not bundle local test data). On an offline/sandboxed host set NXF_OFFLINE, and the wrapper fails fast at preflight with DEMO_REQUIRES_NETWORK and a clear message, instead of a cryptic Nextflow does not exist abort during schema validation. This does not violate the local-first guarantee, which governs your genetic data (never uploaded); --demo only downloads nf-core's public test data. For a fully offline run, use a real analysis with your own local --input samplesheet and references.--gene_bed, --transcript_fasta, …): clawbio.py run rnaseq-pipeline treats _ and - as equivalent when matching the flag allowlist and forwards the wrapper's hyphenated spelling. No manual underscore-to-hyphen conversion is needed.process.resourceLimits config scaled to this host (on macOS: host-RAM share capped to the Docker VM; on Linux/other: physical RAM minus headroom) so a real run does not abort with Process requirement exceeds available memory when an nf-core default request — e.g. MAKE_TRANSCRIPTS_FASTA — is larger than your machine (--demo is exempt: -profile test carries its own limits). If it still aborts (a non-docker backend, or one process that genuinely needs more RAM than the host has), override with your own -c config, e.g. process { resourceLimits = [ memory: '12.GB', cpus: 4 ] }; do not delete resource labels to force it through. On an IPv6-only / NAT64 host the JVM prefers IPv4 and downloads fail with Network is unreachable; export NXF_OPTS='-Djava.net.preferIPv6Addresses=true' and re-run. The wrapper inherits your environment and never overrides NXF_OPTS.--demo bundle. Do not "fix" a replay by deleting the output directory. Unlike the sarek/scrnaseq bundles (which replay Nextflow directly, and Nextflow tolerates a populated output dir), the rnaseq commands.sh re-invokes the wrapper, whose preflight rejects a non-empty --output with OUTPUT_DIR_NOT_EMPTY. So commands.sh carries a guard that adds --resume when the target output dir already holds a completed run of this bundle (reproducibility/manifest.json present); a fresh or remap_paths.py --output-dir-relocated directory has no manifest and runs clean. --demo bundles get the same guard: Nextflow's -resume is orthogonal to -profile test (nf-core documents no incompatibility), the demo samplesheet stub is content-stable so its checksum matches on replay, and the run's work tree (upstream/work) and Nextflow session cache (.nextflow/) both live under the output dir. Resuming across the demo/real boundary is still blocked — demo is compared against the manifest like aligner/profile/arm.remap_paths.py --output-dir <new-path>. The rnaseq bundle bakes the --output directory into commands.sh (its replay re-invokes the wrapper), so after moving the output tree run python3 reproducibility/remap_paths.py --output-dir <new-path> to rewrite it (it keeps the replay guard's manifest path in sync). Use --old/--new for relocated FASTQs and --refs-old/--refs-new for relocated references in params.yaml. The sarek bundle exposes the same --output-dir; the scrnaseq bundle self-relocates (its commands.sh self-anchors) and accepts --output-dir only for parity, as a no-op that confirms no rewrite is needed.REMOTE_INPUT_NOT_ALLOWED) unless --allow-remote-inputs is explicitly passed, which also logs a runtime warning naming every path fetched over the network. The object-store --work-dir is not gated. --allow-remote-inputs relaxes only the wrapper's own preflight check: remote FASTQ/reference URIs are then written into the normalized samplesheet/params.yaml verbatim and staged natively by Nextflow at run time. The wrapper does not download them itself, so remote inputs require outbound network access and are incompatible with NXF_OFFLINE — under offline mode Nextflow's own file-existence validation (nf-schema) still runs and will fail on the remote paths.--params-file: only the audited CLI surface is translated to params.yaml. --nextflow-config forwards user-supplied -c config file(s) for trusted runtime settings such as process, executor, profiles, labels, institutional module tuning, and params.genomes custom genome catalogues. Configs that define params in any form — block (params { … }), property (params.x), assignment (params = …), subscript (params['x']), or map-merge (params << …) — are rejected so they cannot bypass the audited parameter surface (the documented params.genomes catalogue is the sole exception). Every locally-resolvable includeConfig target is audited recursively under the same rule; includes the wrapper cannot read (remote URIs, ${…}-interpolated paths, or missing files) are surfaced as preflight warnings rather than silently trusted, so unaudited surface is always visible.--resume is rejected when the pipeline source/version, profile, aligner, pseudo-aligner, --demo/--prokaryotic/--arm modifiers, params checksum, or samplesheet checksum drift. --demo is part of that contract because it composes the upstream test profile, which supplies both its own samplesheet and its own bundled references — resuming across the demo/real boundary would swap both underneath the run.ClawBio is a research and educational tool. It is not a medical device and does not provide clinical diagnoses. Consult a healthcare professional before making any medical decisions.
Use this skill to produce upstream bulk RNA-seq preprocessing outputs. Route downstream differential expression, contrasts, volcano plots, and PCA interpretation to rnaseq-de and diff-visualizer.
rnaseq-de: bulk/pseudo-bulk differential expression from preferred_counts_tsvdiff-visualizer: plots from downstream DE resultsmultiqc-reporter: optional QC aggregation/reporting follow-upbio-orchestrator: routes inbound bulk RNA-seq preprocessing requests to this wrapperPinned upstream: nf-core/rnaseq v3.26.0. Before changing the default version, audit nextflow.config, assets/schema_input.json, nextflow_schema.json, docs/output.md, and changed module configs, then update tests and reproducibility/pinned_versions.json.
© ClawBio, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 37 other files in skills/nfcore-rnaseq-wrapper of ClawBio/ClawBio.
Open the folder on GitHubat commit 5e045e3
We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in ClawBio/ClawBio, which our catalogue first saw on October 7, 2026.
Nfcore Rnaseq Wrapper next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Nfcore Rnaseq Wrapper this skillClawBio/ClawBio | 1.2k | 1 repos | ~8.9k | Automated safety check: Pass | MIT | |
| LaminDB Biological Data Managementdavila7/claude-code-templates | 32k | 12 repos | ~3.6k | Automated safety check: Pass | MIT | |
| Latchbio Integrationdavila7/claude-code-templates | 32k | 11 repos | ~2.4k | Automated safety check: Pass | MIT | |
| Latchbio IntegrationK-Dense-AI/scientific-agent-skills | 48k | 1 repos | ~2.5k | Automated safety check: Notes | MIT | |
| PacsomaticK-Dense-AI/scientific-agent-skills | 48k | 1 repos | ~1.6k | Automated safety check: Pass | MIT | |
| Dnanexus IntegrationK-Dense-AI/scientific-agent-skills | 48k | 1 repos | ~3.1k | Automated safety check: Pass | MIT |
davila7/claude-code-templates
Manages biological datasets with LaminDB: versioned artifacts, run lineage, ontology-based annotation, schema validation and links to workflow managers and ML tools.
davila7/claude-code-templates
Latch platform for bioinformatics workflows. An agent skill from davila7/claude-code-templates.
K-Dense-AI/scientific-agent-skills
Builds, registers, debugs, and operates bioinformatics workflows on Latch using the Python SDK, CLI, Latch Data and Registry, Nextflow, Snakemake, programmatic execution, and Latch MCP.
K-Dense-AI/scientific-agent-skills
Prepares and launches nf-core/pacsomatic matched tumor-normal PacBio HiFi genomics workflows from unaligned BAM inputs.
K-Dense-AI/scientific-agent-skills
Builds and operates reproducible genomics workloads on DNAnexus with the dx CLI, dxpy, apps/applets, native workflows, dxCompiler, and Nextflow.
GPTomics/bioSkills
Authors portable, strongly-typed bioinformatics pipelines in the Common Workflow Language (CWL v1.2) as CommandLineTool/Workflow/ExpressionTool documents, validated with cwltool and run at scale on…
ClawBio/ClawBio
Fetch a region of cis-eQTL summary statistics from EBI eQTL Catalogue v7+ via tabix-on-FTP.
ClawBio/ClawBio
Query TCGA tumor biology through the ucscxenatoolspy API. An agent skill from ClawBio/ClawBio.
ClawBio/ClawBio
Fetch a region of GWAS summary statistics from the NHGRI-EBI GWAS Catalog harmonised collection via tabix-on-FTP.
ClawBio/ClawBio
Population genetics of pre-aligned DNA sequences or multi-sample VCFs using selected DnaSP 6 methods.
ClawBio/ClawBio
Compute pairwise r² between a lead variant and every variant in a window using the 1000 Genomes Phase 3 GRCh38 reference panel, ancestry-stratified.
ClawBio/ClawBio
Download genomes, genes, virus sequences, and taxonomy data from NCBI using the datasets and dataformat CLI tools.
Works with
Categories
Wrapper skill for running nf-core/rnaseq bulk RNA-seq preprocessing from FASTQ or BAM inputs with strict preflight, reproducibility outputs, and downstream handoff to ClawBio bulk RNA-seq DE skills. Nfcore Rnaseq Wrapper is an agent skill from ClawBio/ClawBio. Wrapper skill for running nf-core/rnaseq bulk RNA-seq preprocessing from FASTQ or BAM inputs with strict preflight, reproducibility outputs, and downstream handoff to ClawBio bulk RNA-seq DE skills.
Nfcore Rnaseq Wrapper fits situations like: tasks that involve Bioinformatics; tasks that involve Reproducible research.
Run `npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a claude-code`. Or copy the skill folder (skills/nfcore-rnaseq-wrapper in ClawBio/ClawBio) into .claude/skills/nfcore-rnaseq-wrapper in your project. Claude Code loads it when a task matches its description.
Run `npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a codex`. Or copy the skill folder (skills/nfcore-rnaseq-wrapper in ClawBio/ClawBio) into .agents/skills/nfcore-rnaseq-wrapper in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ClawBio/ClawBio --skill nfcore-rnaseq-wrapper -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/nfcore-rnaseq-wrapper, .gemini/skills/nfcore-rnaseq-wrapper, .github/skills/nfcore-rnaseq-wrapper and .opencode/skills/nfcore-rnaseq-wrapper in your project.
Going by SKILL.md and its folder, Nfcore Rnaseq Wrapper needs Python for the scripts in its folder and the command-line tools its instructions call (python, docker and python3). Our summary lists: Python 3; Docker.
SKILL.md names 6 domains. As links in the text: nf-co.re, github.com, nextflow.io, salmon.readthedocs.io, daehwankimlab.github.io and bowtie-bio.sourceforge.net. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Nfcore Rnaseq Wrapper is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 8.9k tokens (SKILL.md is roughly 36k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Nfcore Rnaseq Wrapper: LaminDB Biological Data Management (davila7/claude-code-templates, 32k stars), Latchbio Integration (davila7/claude-code-templates, 32k stars), Latchbio Integration (K-Dense-AI/scientific-agent-skills, 48k stars) and Pacsomatic (K-Dense-AI/scientific-agent-skills, 48k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
ClawBio (a GitHub organization) maintains it in ClawBio/ClawBio, which has 1,154 GitHub stars. The repository holds 104 skills in this directory. The repository was last updated on October 7, 2026.
Source: ClawBio/ClawBio on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.