Core pipeline

version %E2%89%A524.04.2 green?style=flat&logo=nextflow&logoColor=white&color=%230DC09D&link=https%3A%2F%2Fnextflow
nf  core template 3.3.1 green?style=flat&logo=nfcore&logoColor=white&color=%2324B064&link=https%3A%2F%2Fnf co
conda

Steps to generate data for the G-nom core analyses are bundled in the core pipeline. It currently includes the following tools:

  1. RepeatMasker

  2. BUSCO

fCat and taXaminer are not yet included pending a bioconda release.

The core pipeline is not executed by G-nom itself but designed to run in a separate HPC environment (it is generally advisable to isolate compute clusters from web-facing applications). The pipeline uses nextflow and is built on top of nf-core. Nextflow supports all major HPC schedulers, including e.g. SLURM and LSF and can be set up in a few minutes. Nf-core provides a common style for the development of nextflow pipelines. The environments required to run the various tools are set up using conda.

Further usage instructions can be found in the repository. Please report issues in the pipeline repository rather then the main G-nom repository.

Automatic import

Future versions of G-nom will include HTTP endpoints to automate the import of pipeline outputs. For progress, check the issue on Github.