Related Experiment Video
Updated: Aug 6, 2025

14:06
Detection of Rare Genomic Variants from Pooled Sequencing Using SPLINTER
Published on: June 23, 2012
15.3K
nPoRe: n-polymer realigner for improved pileup-based variant calling
Tim Dunn1, David Blaauw2, Reetuparna Das2
1University of Michigan, Ann Arbor, USA. timdunn@umich.edu.
BMC Bioinformatics
|March 17, 2023
Summary
Improving germline variant calling for insertions and deletions (INDELs) using nanopore sequencing is crucial. New alignment methods enhance INDEL recall by accurately handling homopolymers and tandem repeats, boosting accuracy in genetic analysis.
Area of Science:
- Genomics
- Bioinformatics
Background:
- Nanopore sequencing offers rapid genomic data generation.
- Current nanopore basecalling has high accuracy for single nucleotide polymorphisms (SNPs) but struggles with insertions and deletions (INDELs).
- INDEL recall remains below 80% for standard R9.4.1 flow cells, limiting comprehensive germline variant analysis.
Purpose of the Study:
- To improve the accuracy of germline INDEL variant calling from nanopore sequencing data.
- To develop enhanced alignment algorithms for better detection of INDELs, particularly in repetitive regions.
Main Methods:
- Extension of the Needleman-Wunsch affine gap alignment algorithm.
- Introduction of novel gap penalties specifically designed for homopolymers and tandem repeats.
- Application of haplotype phasing and realignment techniques to nanopore sequencing reads.
Main Results:
- Haplotype phasing improved INDEL recall from 63.76% to [Formula: see text] at consistent precision.
- Further improvement in INDEL recall to [Formula: see text] was achieved using nPoRe realignment.
- The enhanced alignment demonstrated superior accuracy in regions with repetitive sequences.
Conclusions:
- Read phasing and realignment significantly enhance INDEL recall in nanopore sequencing data.
- Modified gap penalties in alignment are effective for accurately calling INDELs in homopolymers and tandem repeats.
- These advancements address a key limitation in nanopore-based germline variant calling, improving genomic analysis.
Related Concept Videos
Long-patch Base Excision Repair
7.1K
Since the discovery of the two BER pathways, there has been a debate about how a cell chooses one pathway over the other and the factors determining this selection. Numerous in vitro experiments have pointed out multiple determinants for the sub-pathway selection. These are:
7.1K
Single Nucleotide Polymorphisms-SNPs
15.4K
A single nucleotide polymorphism or SNP is a single nucleotide variation at a specific genomic position in a large population. It is the most prevalent type of sequence variation found in the human genome. Point mutations that occur in more than 1% of the population qualify as SNPs. These are present once every 1000 nucleotides on an average in the human genome. Replacement of a purine with another purine (A/G) or a pyrimidine with another pyrimidine (C/T) is known as a transition. In contrast,...
15.4K
Comparing Copy Number Variations and SNPs
17.8K
Sequencing of the human genome has opened up several best-kept secrets of the genome. Scientists have identified thousands of genome variations that exist within a population. These variations can be a single nucleotide or a larger chromosomal variation.
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
17.8K
Sanger Sequencing
755.5K
DNA sequencing is a fundamental technique that is routinely used in the biological sciences. This method can be applied to a range of questions at different scales - from the sequencing of a cloned DNA fragment or the study of a mutation in a gene up to whole-genome sequencing. However, despite the widespread use of sequencing today, it was not until 1977 that Fredrick Sanger and his collaborators developed the chain-termination method to decode DNA sequences. It relies on the separation of a...
755.5K
Polymer Classification: Stereospecificity
2.5K
Polymerization generates chiral centers along the entire backbone of a polymer chain. Accordingly, the stereochemistry of the substituent group has a significant effect on polymer properties. Polymers formed from monosubstituted alkene monomers feature chiral carbons at every alternate position in the polymer backbone. Relative to the predominant orientation of substituents at the adjacent chiral carbons, the polymer can exist in three different configurations: isotactic, syndiotactic, and...
2.5K
Fixing Double-strand Breaks
12.7K
The double-stranded structure of DNA has two major advantages. First, it serves as a safe repository of genetic information where one strand serves as the back-up in case the other strand is damaged. Second, the double-helical structure can be wrapped around proteins called histones to form nucleosomes, which can then be tightly wound to form chromosomes. This way, DNA chains up to 2 inches long can be contained within microscopic structures in a cell. A double-stranded break not only damages...
12.7K

