PDBselect

PDBselect provides a curated, non-redundant set of representative protein chains from the Protein Data Bank (PDB) for unbiased statistical and comparative analyses by enforcing low mutual sequence identity.


Key Features:

  • Curated representative chains: A static curated list of representative protein chains sampled from the Protein Data Bank (PDB).
  • Low mutual sequence identity: Selected chains are chosen to have low mutual sequence identity to minimize redundancy.
  • Non-redundant dataset: Produces a dataset intended to reflect the structural diversity present in the PDB.
  • PDBfilter-select: Provides the PDBfilter-select service to generate custom selections from the PDB based on specified criteria.
  • Bias reduction: Designed to reduce bias in statistical analyses by preventing overrepresentation of similar sequences.

Scientific Applications:

  • Comparative structural analysis: Supports comparative studies by supplying diverse, representative chains for structure-based comparisons.
  • Evolutionary analysis: Supports evolutionary and phylogenetic analyses by providing non-redundant sequence samples.
  • Unbiased statistics: Enables unbiased statistical analysis of sequence and structural properties by minimizing redundancy.

Methodology:

Selection enforces low mutual sequence identity among chosen protein chains drawn from the Protein Data Bank (PDB).

Topics

Details

Tool Type:
web application
Added:
3/27/2017
Last Updated:
11/25/2024

Operations

Publications

Griep S, Hobohm U. PDBselect 1992–2009 and PDBfilter-select. Nucleic Acids Research. 2009;38(suppl_1):D318-D319. doi:10.1093/nar/gkp786. PMID:19783827. PMCID:PMC2808879.