SAMSA2
SAMSA2 processes metatranscriptomic RNA-seq datasets to provide rapid, efficient transcript-level analysis and customizable reference-based annotation in supercomputing cluster environments.
Key Features:
- Speed and Efficiency: Improved speed relative to the original SAMSA pipeline for faster processing of large RNA-seq datasets.
- Lightweight Design: Optimized to reduce computational resource consumption for high-throughput analyses on clusters.
- Enhanced Options: Provides expanded configuration options to tailor analysis parameters to specific experimental needs.
- Simplified Output: Generates streamlined outputs that can be directly examined or further processed for downstream analysis.
- Customizable Reference Databases: Supports upgrading, altering, or customizing reference databases for annotation.
Scientific Applications:
- Metatranscriptomics: Analysis of community-level RNA expression from metatranscriptomic RNA-seq datasets.
- Large-scale RNA-seq processing in HPC: Processing large-scale RNA-seq datasets within supercomputing cluster or high-performance computing environments.
Methodology:
Pipeline optimizations for rapid and efficient analysis of large RNA-seq datasets on supercomputing clusters, reduced computational load, and support for customizable reference databases.
Topics
Details
- License:
- GPL-3.0
- Tool Type:
- workflow
- Operating Systems:
- Linux, Windows, Mac
- Programming Languages:
- Python
- Added:
- 7/31/2018
- Last Updated:
- 12/10/2018
Operations
Publications
Westreich ST, Treiber ML, Mills DA, Korf I, Lemay DG. SAMSA2: a standalone metatranscriptome analysis pipeline. BMC Bioinformatics. 2018;19(1). doi:10.1186/s12859-018-2189-z. PMID:29783945. PMCID:PMC5963165.
Funding: - National Institutes of Health: R01AT008759, T32-GM008799
- Agricultural Research Service: 2032-53000-001-00-D