snakemake-workflow-template/config/README.md at main · MPUSP/snakemake-workflow-template

Workflow overview

This workflow is a best-practice workflow for <detailed description>. The workflow is built using snakemake and consists of the following steps:

Download genome reference from NCBI
Validate downloaded genome (python script)
Simulate short read sequencing data on the fly (dwgsim)
Check quality of input read data (FastQC)
Collect statistics from tool output (MultiQC)

Running the workflow

Input data

This template workflow creates artificial sequencing data in *.fastq.gz format. It does not contain actual input data. The simulated input files are nevertheless created based on a mandatory table linked in the config.yaml file (default: .test/samples.tsv). The sample sheet has the following layout:

sample	condition	replicate	read1	read2
sample1	wild_type	1	sample1.bwa.read1.fastq.gz	sample1.bwa.read2.fastq.gz
sample2	wild_type	2	sample2.bwa.read1.fastq.gz	sample2.bwa.read2.fastq.gz

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Workflow overview

Running the workflow

Input data

FilesExpand file tree

README.md

Latest commit

History

README.md

File metadata and controls

Workflow overview

Running the workflow

Input data