RNA-Bloom2
RNA-Bloom2 assembles long-read transcriptome sequencing data de novo to reconstruct transcripts without a reference genome for comparative transcriptomics when high-quality draft assemblies are unavailable.
Key Features:
- Reference-free de novo assembly: Performs reference-free (de novo) assembly of long-read transcriptome sequencing data.
- Long-read optimization: Tailored for long-read sequencing technologies that capture full-length transcripts to improve transcript reconstruction.
- Benchmark evaluation: Evaluated using simulated datasets and spike-in control data for accuracy and reliability assessment.
- Competitive assembly quality: Produces assembly quality comparable to traditional reference-based methods.
- Resource efficiency: Demonstrates reduced resource usage, requiring 27.0%–80.6% of peak memory and 3.6%–10.8% of total wall-clock runtime compared to other reference-free methods.
Scientific Applications:
- De novo transcriptome reconstruction: Enables transcriptome assembly for organisms lacking high-quality genome assemblies.
- Comparative transcriptomics: Facilitates large-scale comparative transcriptomic analyses across species and conditions without relying on references.
- Real-world assembly: Has been applied to assemble a transcriptome from Picea sitchensis (Sitka spruce).
Methodology:
Performs reference-free de novo assembly of long-read transcriptome sequencing data and is evaluated using simulated datasets and spike-in control data with benchmarking of peak memory and wall-clock runtime.
Topics
Details
- License:
- GPL-3.0
- Cost:
- Free of charge
- Tool Type:
- command-line tool
- Operating Systems:
- Mac, Linux, Windows
- Programming Languages:
- Java, Shell
- Added:
- 1/2/2024
- Last Updated:
- 11/24/2024
Operations
Publications
Nip KM, Hafezqorani S, Gagalova KK, Chiu R, Yang C, Warren RL, Birol I. Reference-free assembly of long-read transcriptome sequencing data with RNA-Bloom2. Nature Communications. 2023;14(1). doi:10.1038/s41467-023-38553-y. PMID:37217540. PMCID:PMC10202958.
PMID: 37217540
PMCID: PMC10202958
Funding: - Genome British Columbia: (243FOR
- Genome Canada: (243FOR
- U.S. Department of Health & Human Services | National Institutes of Health: 2R01HG007182-04A1