Related Experiment Video
Updated: Feb 2, 2026

The Participant-Reported Implementation Update and Score PRIUS: A Novel Method for Capturing Implementation-Related Data Over Time
Published on: February 19, 2021
DREAM-Yara: an exact read mapper for very large databases with short update time
Temesgen Hailemariam Dadi1, Enrico Siragusa2, Vitor C Piro3,4
1Algorithmic Bioinformatics, Institute for Bioinformatics, FU Berlin, Berlin, Germany.
The DREAM framework introduces a novel Bloom filter approach for efficient indexing of large biological sequence databases. This enables faster approximate matching for next-generation sequencing (NGS) reads, overcoming FM-index limitations.
Area of Science:
- Bioinformatics
- Computational Biology
- Genomics
Background:
- Mapping-based approaches using FM-indexes are computationally intensive for large reference databases (>10 GB).
- Index construction for massive datasets, crucial for next-generation sequencing (NGS) read analysis, can take over a day on high-memory machines.
- The difficulty in updating FM-indexes hinders the incorporation of new genomic data and analysis of dynamic biological systems.
Purpose of the Study:
- To develop an alternative indexing strategy that overcomes the limitations of FM-indexes for large-scale biological data.
- To enable faster and more efficient approximate matching of NGS reads to extensive reference databases.
- To facilitate easier index updates and maintain rapid search performance.
Main Methods:
- Introduction of the DREAM (Dynamic seaRchablE pArallel coMpressed index) framework.
- Development of an approximate search distributor utilizing a novel interleaved Bloom filter data structure.
- Implementation of DREAM-Yara, a distributed, fully sensitive read mapper within the DREAM framework.
Main Results:
- The interleaved Bloom filter effectively excludes reads from irrelevant database portions, significantly speeding up search times.
- The DREAM framework allows for parallel processing and dynamic index rebuilding, addressing the bottleneck of large index construction.
- DREAM-Yara demonstrates efficient and sensitive read mapping capabilities within the proposed framework.
Conclusions:
- The DREAM framework provides a scalable and efficient solution for indexing and searching large biological sequence databases.
- The novel Bloom filter application enhances the speed and manageability of NGS read analysis.
- DREAM-Yara offers a robust tool for metagenomics and other applications requiring rapid and accurate read mapping.
More Related Videos
07:38Mass Spectrometry-Based Proteomics Analyses Using the OpenProt Database to Unveil Novel Proteins Translated from Non-Canonical Open Reading Frames
Published on: April 11, 2019
11:18Generation of Comprehensive Thoracic Oncology Database - Tool for Translational Research
Published on: January 22, 2011
Related Concept Videos
Uncertainty in Measurement: Reading Instruments
Fisher's Exact Test
Dreaming
Lucid Dreaming
Studies have shown...
Short-distance Transport of Resources
Power System Three-Phase Short Circuits