damidseq_pipeline
damidseq_pipeline automates processing and normalization of DNA adenine methyltransferase identification (DamID) FASTQ datasets to reduce background noise and identify genomic regions bound by DNA-binding or DNA-associated proteins.
Key Features:
- Automatic normalization: Implements normalization algorithms for DamID-seq data to reduce high background signals that can obscure protein-bound regions.
- Background reduction: Focuses on minimizing background noise to enhance accuracy and reliability of detected binding sites.
- Comprehensive processing steps: Performs sequence alignment, read extension, binned counting, normalization, pseudocount addition, and generation of final ratio files for downstream analysis.
- Batch handling: Capable of processing multiple DamID-seq datasets to support large-scale sequencing experiments.
- Implementation and compatibility: Implemented in Perl and compatible with Unix-based operating systems such as Linux and Mac OS X.
Scientific Applications:
- Mapping protein-DNA interactions: Identifies genomic regions bound by DNA-binding or DNA-associated proteins from DamID-seq data.
- Gene regulation studies: Supports analyses that infer regulatory relationships by locating protein binding sites relative to genes.
- Chromatin organization research: Enables investigation of chromatin-associated protein binding patterns across the genome.
- Comparative binding analyses: Facilitates comparison of binding profiles across conditions or samples using normalized ratio outputs.
Methodology:
Operates on DamID FASTQ datasets and automates sequence alignment, read extension, binned counts, normalization, pseudocount addition, and generation of final ratio files.
Topics
Details
- License:
- GPL-2.0
- Tool Type:
- command-line tool, workflow
- Operating Systems:
- Linux, Mac
- Programming Languages:
- Perl
- Added:
- 5/28/2018
- Last Updated:
- 12/10/2018
Operations
Data Inputs & Outputs
Peak calling
Outputs
Publications
Marshall OJ, Brand AH. damidseq_pipeline: an automated pipeline for processing DamID sequencing datasets. Bioinformatics. 2015;31(20):3371-3373. doi:10.1093/bioinformatics/btv386. PMID:26112292. PMCID:PMC4595905.