← Life Sciences NGS AnalysisCONTENT HISTORY

Update to Life Sciences NGS Analysis

Snapshot Sep 30, 2026 · 22:50 UTC · version 1.0.3

Collection source: not recorded for this historical snapshot.

WHAT CHANGED · RULE-BASED ANALYSIS

First saved snapshot

No earlier snapshot is available to establish a change.

Compare saved observations

Download comparison JSON
Full technical diff · 0 changed fields
Full snapshot data
{
  "name": "ngs-amplicon-microbiome",
  "description": "Kick off public 16S, 18S, ITS, COI, or other marker-gene amplicon microbiome workflows using nf-core/ampliseq, QIIME2, DADA2, and Cutadapt.",
  "included_files": [
    {
      "relative_path": "agents/openai.yaml",
      "size_in_bytes": 267
    }
  ],
  "skill_md_contents": "---\nname: ngs-amplicon-microbiome\ndescription: Kick off public 16S, 18S, ITS, COI, or other marker-gene amplicon microbiome workflows using nf-core/ampliseq, QIIME2, DADA2, and Cutadapt.\n---\n\n# Amplicon Microbiome\n\nUse this skill for marker-gene microbiome analysis from amplicon FASTQs.\n\n## Essential Inputs\n\nConfirm:\n\n- marker region: 16S, 18S, ITS, COI, or custom\n- primer sequences and orientation\n- paired-end or single-end reads\n- whether reads should be merged\n- taxonomy database and version\n- sample metadata\n- endpoint: ASV table, taxonomy, diversity, differential abundance, or plots\n\n## Public Defaults\n\nPrefer `nf-core/ampliseq` for reproducible end-to-end runs. Use QIIME2 or DADA2 directly when the user wants notebook-level control or an existing lab protocol requires it.\n\n## Preflight\n\n```bash\npython plugins/ngs-analysis/scripts/ngs_preflight.py --pipeline amplicon_microbiome --emit-install-plan\n```\n\n## Local Execution Package\n\nFor FASTQ intake/QC before primer, ASV, and taxonomy decisions, use:\n\n```bash\npython plugins/ngs-analysis/scripts/run_fastq_assay_package.py \\\n  --lane amplicon_microbiome \\\n  --sample-sheet amplicon_samples.tsv \\\n  --execute\n```\n\nThis validates read paths and structure, runs seqkit stats and FastQC/MultiQC when available, and writes `amplicon_analysis_status.json`. The runner now also emits `methods/amplicon_methods.json` plus a concrete backend handoff bundle under `workflow/` so primer, denoiser, truncation, normalization, and taxonomy choices are machine-readable even before a full backend is run.\n\nIf the user asks for a full amplicon analysis rather than QC/readiness, do not treat FASTQs alone as sufficient. Require primer sequences, primer orientation, taxonomy database plus version, and sample metadata before presenting the run as analysis-ready. Without that context, run the local execution package and describe the result as a read-QC/readiness bundle only.\n\nFor backend ASV/taxonomy/diversity execution when primers, metadata, and taxonomy resources are available, use:\n\n```bash\npython plugins/ngs-analysis/scripts/run_amplicon_microbiome.py \\\n  --sample-sheet amplicon_samples.tsv \\\n  --backend qiime2 \\\n  --primer-forward GTGYCAGCMGCCGCGGTAA \\\n  --primer-reverse GGACTACNVGGGTWTCTAAT \\\n  --taxonomy-classifier silva-138-classifier.qza \\\n  --metadata sample_metadata.tsv \\\n  --execute\n```\n\nUse `--backend dada2` for a direct R/Bioconductor ASV path. The plugin includes `workflows/amplicon_microbiome/run_dada2_backend.R`; the runner checks for `Rscript` and the `dada2` R package before execution, then writes normalized ASV, representative-sequence, read-retention, and optional taxonomy tables under `tables/`.\n\nFor nf-core execution, use `plugins/ngs-analysis/scripts/run_nfcore_pipeline.py --pipeline ampliseq`.\n\nThe direct backend runner also emits `resources/resource_plan.json`, `resource_manifest.tsv`, `resource_env.sh`, and `resource_readiness.md`. The resource check is advisory by default when a QIIME classifier is supplied directly; add `--bundle-root silva_138_amplicon=<path>`, `--include-optional-resources`, and `--require-resource-plan` when missing registered taxonomy databases should block readiness.\n\nThe backend runner writes native normalized tables when QIIME2/DADA2/nf-core outputs are present:\n\n- `tables/asv_table.tsv`\n- `tables/representative_sequences.fasta` for direct DADA2 runs\n- `tables/taxonomy.tsv`\n- `tables/read_retention.tsv`\n- `tables/amplicon_backend_summary.json`\n- `tables/alpha_diversity.tsv`, `tables/bray_curtis_distance.tsv`, and `tables/top_taxa_or_features.tsv` when a normalized ASV/feature table is available\n\nQIIME2 BIOM-only feature-table exports are recorded as requiring conversion, with a `biom convert` command in the backend summary. Do not claim diversity or taxonomy interpretation unless these normalized tables or equivalent supplied inputs exist.\n\n## Kickoff Pattern\n\nnf-core preflight run:\n\n```bash\nnextflow run nf-core/ampliseq \\\n  -profile test,docker \\\n  --outdir results/ampliseq_test\n```\n\nBefore a real run, verify primer trimming and truncation choices from read-quality profiles.\n\n## Visualization Outputs\n\nThe local FASTQ package always writes `visualizations/index.html` and `visualizations/visualization_manifest.json`. With only FASTQs, this is a read-QC/readiness bundle. If an ASV/feature table is available, pass it to the runner with `--asv-table` to generate alpha diversity, Bray-Curtis PCoA, and rarefaction artifacts. If a feature taxonomy table is available, pass `--taxonomy-table` to generate taxa barplots. When downstream tables are labeled synthetic or contain sample columns that are not present in the real sample sheet, the runner marks the run review-only and blocks beta-diversity/PCoA unless `--allow-synthetic-diversity` is set explicitly.\n\nThe run also emits `qc_verdict.json` and, for amplicon runs, `qc_interpretation.json` with machine-readable reason codes, a readiness verdict, and follow-on command templates for generating ASV/taxonomy tables and re-rendering plugin-native plots. Backend runs additionally write `tables/amplicon_backend_summary.json` so exported ASV, taxonomy, read-retention, and BIOM-conversion status are auditable. When a normalized ASV/feature table is available, the backend runner also writes `tables/amplicon_diversity_summary.json`, `visualizations/amplicon_backend_dashboard.html`, and SVG plots for sample depth, Shannon diversity, and top taxa/features. If the ASV table is absent, these outputs remain explicitly unavailable rather than inferred from FASTQ QC.\n\n## Guardrails\n\n- Do not choose truncation lengths before looking at quality distributions.\n- Do not mix taxonomy database versions without recording them.\n- Preserve negative controls and extraction blanks in metadata.\n"
}

SHA-256: 6103a0fe553fb3d9e6d0008f48043a07ae00c4490ce6c9575e8e643cf5344eb3