GTO
GTO provides a modular command-line toolkit for processing, analyzing, simulating, compressing, transforming, and visualizing genomic and proteomic sequence data from next-generation sequencing, supporting FASTQ, FASTA, and SEQ formats.
Key Features:
- Modular Architecture: Modular design that composes pipelines from subprograms to build tailored analyses.
- Comprehensive Toolset: Functions for data analysis, simulation, compression, development, transformation, and visualization of sequence data.
- Format Support: Native support for FASTQ, FASTA, and SEQ file formats.
- Unix Compatibility and Performance: Targets ultra-fast computations on Unix-based systems for large datasets.
- Integration via Pipes: Supports Unix pipes to chain internal subprograms and external tools for interoperability.
Scientific Applications:
- Next-Generation Sequencing Analysis: Facilitates processing and analysis of large NGS datasets.
- Genomic and Proteomic Data Processing: Enables combined analyses and transformations of genomic and proteomic sequence data.
- Translational and Evolutionary Research: Supports analytical workflows relevant to personalized medicine, evolutionary biology, and molecular diagnostics.
Methodology:
Implemented in C; employs a modular set of subprograms and Unix pipes to compose pipelines and performs simulation, compression, transformation, visualization, and sequence analysis on FASTQ, FASTA, and SEQ files targeting ultra-fast computations on Unix-based systems.
Topics
Details
- License:
- MIT
- Tool Type:
- web application
- Programming Languages:
- C
- Added:
- 1/18/2021
- Last Updated:
- 1/25/2021
Operations
Publications
Almeida JR, Pinho AJ, Oliveira JL, Fajarda O, Pratas D. GTO: a toolkit to unify pipelines in genomic and proteomic research. Unknown Journal. 2020. doi:10.1101/2020.01.07.882845.