SAMSA2

SAMSA2 processes metatranscriptomic RNA-seq datasets to provide rapid, efficient transcript-level analysis and customizable reference-based annotation in supercomputing cluster environments.


Key Features:

  • Speed and Efficiency: Improved speed relative to the original SAMSA pipeline for faster processing of large RNA-seq datasets.
  • Lightweight Design: Optimized to reduce computational resource consumption for high-throughput analyses on clusters.
  • Enhanced Options: Provides expanded configuration options to tailor analysis parameters to specific experimental needs.
  • Simplified Output: Generates streamlined outputs that can be directly examined or further processed for downstream analysis.
  • Customizable Reference Databases: Supports upgrading, altering, or customizing reference databases for annotation.

Scientific Applications:

  • Metatranscriptomics: Analysis of community-level RNA expression from metatranscriptomic RNA-seq datasets.
  • Large-scale RNA-seq processing in HPC: Processing large-scale RNA-seq datasets within supercomputing cluster or high-performance computing environments.

Methodology:

Pipeline optimizations for rapid and efficient analysis of large RNA-seq datasets on supercomputing clusters, reduced computational load, and support for customizable reference databases.

Topics

Details

License:
GPL-3.0
Tool Type:
workflow
Operating Systems:
Linux, Windows, Mac
Programming Languages:
Python
Added:
7/31/2018
Last Updated:
12/10/2018

Operations

Publications

Westreich ST, Treiber ML, Mills DA, Korf I, Lemay DG. SAMSA2: a standalone metatranscriptome analysis pipeline. BMC Bioinformatics. 2018;19(1). doi:10.1186/s12859-018-2189-z. PMID:29783945. PMCID:PMC5963165.

Funding: - National Institutes of Health: R01AT008759, T32-GM008799 - Agricultural Research Service: 2032-53000-001-00-D

Documentation

Links