Biobtree

Biobtree performs search and mapping of large-scale bioinformatics datasets using identifiers and specialized keywords to support genomic data integration and analysis.


Key Features:

  • MapReduce-based framework: MapReduce-based computational framework that optimizes processing power and storage resources for efficient handling of large-scale genomic data.
  • B+ tree database: B+ tree-based database structure that provides uniform search results and accelerates query speed and accuracy.
  • Chain mapping queries: Support for chain mapping queries across different datasets to establish complex relationships between biological entities.
  • Identifier and keyword search: Search and mapping via identifiers or specialized keywords such as species names.

Scientific Applications:

  • Identifier-driven mapping: Identifier- and keyword-driven search and mapping of large-scale genomic and bioinformatics datasets.
  • Cross-dataset relationship discovery: Establishing complex relationships between biological entities across multiple datasets via chain mapping queries.
  • High-throughput querying: Efficient processing and querying of extensive genomic data volumes using MapReduce and B+ tree indexing.

Methodology:

Biobtree uses a MapReduce-based framework and a B+ tree-based database to index datasets and perform identifier- or keyword-driven chain mapping queries.

Topics

Details

License:
BSD-3-Clause
Added:
11/14/2019
Last Updated:
12/5/2020

Operations

Publications

Gur T. Biobtree: A tool to search and map bioinformatics identifiers and special keywords. F1000Research. 2019;8:145. doi:10.12688/f1000research.17927.2.