damidseq_pipeline

damidseq_pipeline automates processing and normalization of DNA adenine methyltransferase identification (DamID) FASTQ datasets to reduce background noise and identify genomic regions bound by DNA-binding or DNA-associated proteins.


Key Features:

  • Automatic normalization: Implements normalization algorithms for DamID-seq data to reduce high background signals that can obscure protein-bound regions.
  • Background reduction: Focuses on minimizing background noise to enhance accuracy and reliability of detected binding sites.
  • Comprehensive processing steps: Performs sequence alignment, read extension, binned counting, normalization, pseudocount addition, and generation of final ratio files for downstream analysis.
  • Batch handling: Capable of processing multiple DamID-seq datasets to support large-scale sequencing experiments.
  • Implementation and compatibility: Implemented in Perl and compatible with Unix-based operating systems such as Linux and Mac OS X.

Scientific Applications:

  • Mapping protein-DNA interactions: Identifies genomic regions bound by DNA-binding or DNA-associated proteins from DamID-seq data.
  • Gene regulation studies: Supports analyses that infer regulatory relationships by locating protein binding sites relative to genes.
  • Chromatin organization research: Enables investigation of chromatin-associated protein binding patterns across the genome.
  • Comparative binding analyses: Facilitates comparison of binding profiles across conditions or samples using normalized ratio outputs.

Methodology:

Operates on DamID FASTQ datasets and automates sequence alignment, read extension, binned counts, normalization, pseudocount addition, and generation of final ratio files.

Topics

Details

License:
GPL-2.0
Tool Type:
command-line tool, workflow
Operating Systems:
Linux, Mac
Programming Languages:
Perl
Added:
5/28/2018
Last Updated:
12/10/2018

Operations

Data Inputs & Outputs

Peak calling

Publications

Marshall OJ, Brand AH. damidseq_pipeline: an automated pipeline for processing DamID sequencing datasets. Bioinformatics. 2015;31(20):3371-3373. doi:10.1093/bioinformatics/btv386. PMID:26112292. PMCID:PMC4595905.

Documentation

Links