SISEQ
SISEQ extracts and converts DNA sequences from large database entries for molecular biology and phylogenetic analyses.
Key Features:
- Sequence extraction: Extracts DNA sequences corresponding to coding sequences (CDS) or RNA fields from substantial database files.
- Multi-sequence conversion: Converts extracted sequences into multi-sequence formats suitable for downstream phylogenetic and molecular biological analyses.
- Large-scale dataset handling: Processes substantial database entries to support analyses of extensive genomic or transcriptomic datasets.
- Workflow automation: Supports script-driven automation to facilitate repetitive processing and integration into larger computational workflows.
- Data preparation for downstream analysis: Prepares sequence data specifically for phylogenetic and molecular biology applications.
Scientific Applications:
- Evolutionary biology and phylogenetics: Enables extraction and formatting of sequences for phylogenetic inference and comparative evolutionary studies.
- Genomics and comparative genomics: Facilitates retrieval of coding and RNA sequences from large databases for genome-scale comparisons.
- Molecular genetics: Provides targeted sequence extraction of CDS and RNA fields for gene-focused analyses.
- Transcriptomics: Supports handling and conversion of transcript-related sequences for large-scale transcriptomic investigations.
Methodology:
Extracts DNA sequences corresponding to CDS or RNA fields from large database files and performs multi-sequence conversion for phylogenetic and molecular biological analyses, with support for script-driven automation.
Topics
Details
- Tool Type:
- command-line tool
- Operating Systems:
- Linux, Windows, Mac
- Programming Languages:
- C
- Added:
- 12/18/2017
- Last Updated:
- 12/14/2018
Operations
Data Inputs & Outputs
Formatting
Other operations do not define inputs or outputs.
Publications
Sato N. SISEQ: manipulation of multiple sequence and large database files for common platforms. Bioinformatics. 2000;16(2):180-181. doi:10.1093/bioinformatics/16.2.180. PMID:10842742.
PMID: 10842742