Harvester

Harvester aggregates bioinformatic data from multiple public databases and prediction servers to compile, compare, and enable proteome-wide searches of human protein information.


Key Features:

  • Bulk Data Collection: Collects data by interfacing with Uniprot/SWISSprot, ensEMBL, BLAST (NCBI), SOURCE, SMART, STRING, PSORT2, CDART, UniGene, and SOSUI.
  • Integrated Data Presentation: Compiles database-derived information on individual proteins into a single HTML page combining screenshots and plain text outputs.
  • Full-Text Meta Search Engine: Implements a full-text meta search engine for rapid screening and retrieval across the human proteome.
  • Comparative Analysis: Presents outputs from multiple databases and prediction algorithms on one page to enable direct comparison and assessment of database entries and predictions.

Scientific Applications:

  • Protein Function Prediction: Aggregates predictions from multiple sources to support generation and cross-validation of hypotheses about protein function.
  • Data Quality Assessment: Enables identification of discrepancies across Uniprot/SWISSprot, ensEMBL, SOURCE, and other resources to assess entry reliability.
  • Genome-Wide Proteomic Studies: Supports proteome-scale querying and retrieval for large-scale genomics and proteomics analyses.

Methodology:

Collects data from specified bioinformatic resources (Uniprot/SWISSprot, ensEMBL, BLAST (NCBI), SOURCE, SMART, STRING, PSORT2, CDART, UniGene, SOSUI), assembles the information into a single HTML page per protein, and implements a full-text meta search engine for efficient querying across the human proteome.

Topics

Collections

Details

Tool Type:
web application
Operating Systems:
Linux, Windows, Mac
Added:
5/2/2017
Last Updated:
11/24/2024

Operations

Publications

Liebel U, Kindler B, Pepperkok R. ‘Harvester’: a fast meta search engine of human protein resources. Bioinformatics. 2004;20(12):1962-1963. doi:10.1093/bioinformatics/bth146. PMID:14988114.