eProbalign
eProbalign computes maximal expected accuracy multiple sequence alignments using partition function posterior probabilities to generate accurate residue-pairing for protein sequence analysis.
Key Features:
- Maximal Expected Accuracy (MEA): Constructs alignments that maximize expected accuracy using the maximum expected accuracy optimization criterion.
- Partition function posterior probabilities: Uses partition function-derived posterior probabilities and pairwise residue posterior probabilities to guide alignment decisions within the Probalign framework.
- Benchmark performance: Outperforms Probcons, MAFFT, and MUSCLE on benchmarks including BAliBASE 3.0, HOMSTRAD, and OXBENCH with reported statistical significance (P-value < 0.005).
- Enhanced performance on complex datasets: Shows large accuracy improvements on datasets with N/C-terminal extensions, long or heterogeneous-length proteins, and protein repeats, with at least 10% and 15% higher accuracy when the standard deviation of length exceeds 300 and 400, respectively.
Scientific Applications:
- Complex protein dataset analysis: Alignment of proteins with N/C-terminal extensions, long sequences, heterogeneous lengths, and repeats.
- Evolutionary biology: Precise sequence comparisons for phylogenetic and comparative analyses.
- Structural genomics: Improved residue alignments to support structural inference and modeling.
- Functional annotation of proteins: Accurate alignments to inform functional residue and domain annotation.
Methodology:
Computes partition function posterior probabilities, derives pairwise residue posterior probabilities, and applies maximum expected accuracy optimization within the Probalign framework.
Topics
Details
- Tool Type:
- web application
- Added:
- 2/14/2017
- Last Updated:
- 11/25/2024
Operations
Publications
Roshan U, Livesay DR. Probalign: multiple sequence alignment using partition function posterior probabilities. Bioinformatics. 2006;22(22):2715-2721. doi:10.1093/bioinformatics/btl472. PMID:16954142.
Chikkagoudar S, Roshan U, Livesay D. eProbalign: generation and manipulation of multiple sequence alignments using partition function posterior probabilities. Nucleic Acids Research. 2007;35(Web Server):W675-W677. doi:10.1093/nar/gkm267. PMID:17485479. PMCID:PMC1933135.