Related Experiment Video
Updated: Feb 6, 2026

Targeted Next-generation Sequencing and Bioinformatics Pipeline to Evaluate Genetic Determinants of Constitutional Disease
Published on: April 4, 2018
Bioinformatic removal of NUMT-associated variants in mitotiling next-generation sequencing data from whole blood
Joseph David Ring1,2, Kimberly Sturk-Andreaggi1,2, Michelle Alyse Peck1,2
1Armed Forces Medical Examiner System's Armed Forces DNA Identification Laboratory (AFMES-AFDIL), DE, United States.
Abstract:
Nuclear mitochondrial DNA segments (NUMTs) have arisen because of the transposition of segments of the mitochondrial DNA genome (mitogenome) into the nuclear genome. When using a "mitotiling" strategy, NUMTs may be more readily amplified when targeting the entire mitogenome compared to the control region, as hundreds of primers are required for complete sequencing coverage. In samples with a high percentage of nuclear DNA copies per cell, such as whole blood, NUMT coenrichment may be exacerbated. The present study examined bioinformatic approaches for removing NUMTs and NUMT-associated variants (NAVs) from next-generation sequence data generated using two mitotiling kits (Precision ID and QIAseq). Across 16 samples with low mtDNA copy number, NUMT coenrichment produced 890 NAVs with >5% variant frequency. The use of the consensus sequence to eliminate NUMT reads proved to be effective for QIAseq data, and resulted in >85% NAV removal in Precision ID data. This method was bolstered by NAV filtering in Precision ID analysis. Alternative high stringency mapping to the revised Cambridge Reference Sequence (rCRS) and the human genome reference GRCh38 for the QIAseq data caused a reduction in mitogenome coverage without complete NUMT removal. These bioinformatic solutions facilitate mitotiling sequence data analysis for low-level variant detection.
Related Concept Videos
Next-generation Sequencing
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Histone Variants at the Centromere
Cis-regulatory Sequences
How Data are Classified: Categorical Data
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...
How Data are Classified: Numerical Data
Quantitative data may be either discrete or continuous. All quantitative data that take on only specific numerical...
Sequences

