pyHCA

pyHCA applies Hydrophobic Cluster Analysis (HCA) to protein sequences to detect hydrophobic clusters and infer structural domains, including regions lacking similarity to annotated domain database entries.


Key Features:

  • Hydrophobic Cluster Analysis (HCA): Performs detection and analysis of hydrophobic clusters in protein sequences to infer structural domain organization.
  • Protein Sequence Processing API: Provides Python classes and functions for automated integration of HCA into computational workflows and analysis pipelines.
  • Executable HCA Toolkit: Executes HCA through standalone programs (HCAtk) for large-scale protein sequence analysis.
  • Enhanced Domain Detection Methodologies: Incorporates improved HCA-based approaches to detect domains in unannotated proteomes, fast-diverging proteins, recently emerged proteins, and proteins containing intrinsically disordered regions.

Scientific Applications:

  • Novel Domain Identification: Detects protein domains without sequence similarity to known domain database entries, enabling structural and evolutionary analysis of unannotated and rapidly evolving proteins.
  • Emerging Domain Characterization: Analyzes hydrophobic cluster patterns to investigate structural features of newly emerged protein domains.

Methodology:

Applies hydrophobic cluster analysis to protein sequences to identify regions enriched in hydrophobic residues. Cluster distribution patterns are used to infer potential structural domains and functional properties independently of sequence similarity-based annotation methods.

Collections

Details

License:
CECILL-C
Cost:
Free of charge
Tool Type:
command-line tool
Operating Systems:
Linux, Mac
Programming Languages:
Python
Added:
8/29/2022
Last Updated:
11/24/2024

Operations

Publications

Bitard-Feildel T, Callebaut I. HCAtk and pyHCA: A Toolkit and Python API for the Hydrophobic Cluster Analysis of Protein Sequences. Unknown Journal. 2018. doi:10.1101/249995.