GTO

GTO provides a modular command-line toolkit for processing, analyzing, simulating, compressing, transforming, and visualizing genomic and proteomic sequence data from next-generation sequencing, supporting FASTQ, FASTA, and SEQ formats.


Key Features:

  • Modular Architecture: Modular design that composes pipelines from subprograms to build tailored analyses.
  • Comprehensive Toolset: Functions for data analysis, simulation, compression, development, transformation, and visualization of sequence data.
  • Format Support: Native support for FASTQ, FASTA, and SEQ file formats.
  • Unix Compatibility and Performance: Targets ultra-fast computations on Unix-based systems for large datasets.
  • Integration via Pipes: Supports Unix pipes to chain internal subprograms and external tools for interoperability.

Scientific Applications:

  • Next-Generation Sequencing Analysis: Facilitates processing and analysis of large NGS datasets.
  • Genomic and Proteomic Data Processing: Enables combined analyses and transformations of genomic and proteomic sequence data.
  • Translational and Evolutionary Research: Supports analytical workflows relevant to personalized medicine, evolutionary biology, and molecular diagnostics.

Methodology:

Implemented in C; employs a modular set of subprograms and Unix pipes to compose pipelines and performs simulation, compression, transformation, visualization, and sequence analysis on FASTQ, FASTA, and SEQ files targeting ultra-fast computations on Unix-based systems.

Topics

Details

License:
MIT
Tool Type:
web application
Programming Languages:
C
Added:
1/18/2021
Last Updated:
1/25/2021

Operations

Publications

Almeida JR, Pinho AJ, Oliveira JL, Fajarda O, Pratas D. GTO: a toolkit to unify pipelines in genomic and proteomic research. Unknown Journal. 2020. doi:10.1101/2020.01.07.882845.