FISH - sequence search

FISH - sequence search identifies the family membership of protein domains within query sequences by integrating sequence and structural information using structure-anchored Hidden Markov Models (saHMMs).


Key Features:

  • High Accuracy: Achieves a 99.3% top-hit family assignment success rate on SCOP sequences at an E-value cutoff of 0.1.
  • Structure-Anchored Hidden Markov Models (saHMMs): Employs saHMMs that integrate sequence and structural information to detect domains with low sequence similarity to known proteins.
  • Functional Annotation: Provides functional annotations for identified domain families to suggest potential roles within proteins.
  • Structural Predictions: Produces probable 2D and 3D structural predictions for identified domains to inform protein architecture and function analyses.
  • Sequence Alignments: Generates pairwise and multiple sequence alignments with low-identity homologues to support comparative and evolutionary analyses.
  • Custom Searches: Supports analysis of user-supplied protein sequences and searches of public sequence databases using individual saHMMs.

Scientific Applications:

  • Protein Function Prediction: Identifies domain families and associated annotations to aid prediction of functions for uncharacterized proteins.
  • Structural Biology: Provides structural predictions to support studies of protein folding, stability, and interaction networks.
  • Comparative Genomics: Produces low-identity sequence alignments enabling evolutionary comparisons and analysis of conserved domain architectures across species.

Methodology:

Leverages structure-anchored Hidden Markov Models to integrate sequence and structural information for robust identification of protein domain families.

Topics

Details

Tool Type:
web application
Operating Systems:
Linux, Windows, Mac
Added:
12/6/2015
Last Updated:
11/25/2024

Operations

Data Inputs & Outputs

Prediction and recognition

Publications

Tangrot J, Wang L, Kagstrom B, Sauer UH. FISH--family identification of sequence homologues using structure anchored hidden Markov models. Nucleic Acids Research. 2006;34(Web Server):W10-W14. doi:10.1093/nar/gkl330. PMID:16844969. PMCID:PMC1538871.

Documentation