Related Experiment Video
Updated: Jun 22, 2025

Facet-to-facet Linking of Shape-anisotropic Colloidal Cadmium Chalcogenide Nanostructures
Published on: August 10, 2017
Label-guided seed-chain-extend alignment on annotated De Bruijn graphs
Harun Mustafa1,2,3, Mikhail Karasikov1,2,3, Nika Mansouri Ghiasi4
1Department of Computer Science, ETH Zurich, Zurich, 8092, Switzerland.
We developed multi-label alignment (MLA), a new scoring model for De Bruijn graphs, to improve long-read alignment accuracy in fragmented sequencing data. MLA enhances taxonomic classification and reduces errors compared to existing methods.
Area of Science:
- Bioinformatics
- Computational Biology
- Genomics
Background:
- Scalable De Bruijn graph-based (DBG) indexing is crucial for searching large sequencing databases.
- Low-depth sequencing creates fragmented subgraphs, hindering accurate long-read alignment.
- Existing aligners struggle with fragmentation, leading to low recall or reduced accuracy due to irrelevant sample combinations.
Purpose of the Study:
- Introduce a novel scoring model, multi-label alignment (MLA), for annotated DBGs.
- Address challenges in long-read alignment caused by fragmented sequencing data.
- Improve the biological relevance and accuracy of alignments in large sequencing datasets.
Main Methods:
- Developed MLA with 'Label Change' and 'Node Length Change' operations for biologically relevant sample combinations and improved connectivity.
- Implemented MLA using a two-step approach: single-label seed-chain-extend aligner (SCA) and multi-label chainer (MLC).
- SCA provides initial alignments, while MLC refines them using MLA scoring for multi-label chains and final alignments.
Main Results:
- MLA significantly improves taxonomic classification accuracy, reducing average weighted UniFrac errors by 63.1%-66.8%.
- MLA covers 45.5%-47.4% more long-read query characters compared to state-of-the-art aligners.
- MLA achieves competitive runtimes, faster than single-label alignment and comparable to label-combining methods.
Conclusions:
- MLA effectively produces biologically relevant alignments, overcoming fragmentation issues in DBGs.
- The method enhances the accuracy and recall of long-read alignment in large sequencing databases.
- MLA offers a scalable and accurate solution for analyzing fragmented sequencing data.
Related Concept Videos
Radical Chain-Growth Polymerization: Chain Branching
Maxam-Gilbert Sequencing
Challenges of the Maxam-Gilbert Method
The...
Structure of Conjugated Dienes
Conjugated dienes are compounds characterized by the presence of alternating double and single bonds. In a conjugated system like 1,3-butadiene, the unhybridized 2p orbital on each carbon overlaps continuously, allowing the π electrons to be delocalized across the entire molecule. In contrast, this type of overlap does not occur in cumulated and isolated dienes, such as 2,3-pentadiene and 1,4-pentadiene, respectively. Instead, the π electrons remain localized between the double...
Ziegler–Natta Chain-Growth Polymerization: Overview
Ligand Binding and Linkage
Labeling DNA Probes
Radioisotopes, fluorophores, or small molecule binding partners like biotin or digoxigenin, are the most widely used reporter tags for labeling DNA probes. These labels can be attached to the probe DNA molecule via...

