Harvester
Harvester aggregates bioinformatic data from multiple public databases and prediction servers to compile, compare, and enable proteome-wide searches of human protein information.
Key Features:
- Bulk Data Collection: Collects data by interfacing with Uniprot/SWISSprot, ensEMBL, BLAST (NCBI), SOURCE, SMART, STRING, PSORT2, CDART, UniGene, and SOSUI.
- Integrated Data Presentation: Compiles database-derived information on individual proteins into a single HTML page combining screenshots and plain text outputs.
- Full-Text Meta Search Engine: Implements a full-text meta search engine for rapid screening and retrieval across the human proteome.
- Comparative Analysis: Presents outputs from multiple databases and prediction algorithms on one page to enable direct comparison and assessment of database entries and predictions.
Scientific Applications:
- Protein Function Prediction: Aggregates predictions from multiple sources to support generation and cross-validation of hypotheses about protein function.
- Data Quality Assessment: Enables identification of discrepancies across Uniprot/SWISSprot, ensEMBL, SOURCE, and other resources to assess entry reliability.
- Genome-Wide Proteomic Studies: Supports proteome-scale querying and retrieval for large-scale genomics and proteomics analyses.
Methodology:
Collects data from specified bioinformatic resources (Uniprot/SWISSprot, ensEMBL, BLAST (NCBI), SOURCE, SMART, STRING, PSORT2, CDART, UniGene, SOSUI), assembles the information into a single HTML page per protein, and implements a full-text meta search engine for efficient querying across the human proteome.
Topics
Collections
Details
- Tool Type:
- web application
- Operating Systems:
- Linux, Windows, Mac
- Added:
- 5/2/2017
- Last Updated:
- 11/24/2024
Operations
Publications
Liebel U, Kindler B, Pepperkok R. ‘Harvester’: a fast meta search engine of human protein resources. Bioinformatics. 2004;20(12):1962-1963. doi:10.1093/bioinformatics/bth146. PMID:14988114.
PMID: 14988114