Biobtree
Biobtree performs search and mapping of large-scale bioinformatics datasets using identifiers and specialized keywords to support genomic data integration and analysis.
Key Features:
- MapReduce-based framework: MapReduce-based computational framework that optimizes processing power and storage resources for efficient handling of large-scale genomic data.
- B+ tree database: B+ tree-based database structure that provides uniform search results and accelerates query speed and accuracy.
- Chain mapping queries: Support for chain mapping queries across different datasets to establish complex relationships between biological entities.
- Identifier and keyword search: Search and mapping via identifiers or specialized keywords such as species names.
Scientific Applications:
- Identifier-driven mapping: Identifier- and keyword-driven search and mapping of large-scale genomic and bioinformatics datasets.
- Cross-dataset relationship discovery: Establishing complex relationships between biological entities across multiple datasets via chain mapping queries.
- High-throughput querying: Efficient processing and querying of extensive genomic data volumes using MapReduce and B+ tree indexing.
Methodology:
Biobtree uses a MapReduce-based framework and a B+ tree-based database to index datasets and perform identifier- or keyword-driven chain mapping queries.
Topics
Details
- License:
- BSD-3-Clause
- Added:
- 11/14/2019
- Last Updated:
- 12/5/2020
Operations
Publications
Gur T. Biobtree: A tool to search and map bioinformatics identifiers and special keywords. F1000Research. 2019;8:145. doi:10.12688/f1000research.17927.2.