Related Experiment Video
Updated: Dec 8, 2025

08:38
Targeted DNA Methylation Analysis by Next-generation Sequencing
Published on: February 24, 2015
37.8K
Removing reference bias and improving indel calling in ancient DNA data analysis by mapping to a sequence variation
Rui Martiniano1, Erik Garrison2,3, Eppie R Jones4
1Department of Genetics, University of Cambridge, Cambridge, CB3 0DH, UK.
Genome Biology
|September 18, 2020
Summary
Ancient DNA (aDNA) analysis is improved using variation graphs to overcome reference bias. This method enhances variant detection, particularly for insertions and deletions (indels), in degraded DNA sequences.
Area of Science:
- Genomics
- Bioinformatics
- Paleogenetics
Background:
- Ancient DNA (aDNA) analysis is crucial for studying past human populations.
- Degraded aDNA exhibits short sequences and post-mortem mutations, reducing mapping accuracy and causing reference bias.
- Reference bias favors mapping reads with reference alleles over non-reference ones.
Purpose of the Study:
- To evaluate the effectiveness of variation graph software (vg) in mitigating reference bias for aDNA analysis.
- To compare vg performance against existing methods like BWA for aDNA sequence alignment.
Main Methods:
- Alignment of simulated and real aDNA samples using vg against a variation graph.
- Comparison of vg alignments with BWA alignments to the human linear reference genome.
- Evaluation of variant detection sensitivity, especially for insertions and deletions (indels).
Main Results:
- Variation graph alignment with vg effectively removed reference bias, achieving balanced allelic representation.
- vg demonstrated more sensitive variant detection compared to BWA, particularly for indels.
- Alternative methods like relaxed BWA parameters or filtering reduced bias but were less sensitive than vg for indels.
Conclusions:
- Aligning aDNA sequences to variation graphs is an effective strategy to mitigate reference bias.
- Variation graph analysis retains mapping sensitivity and improves the detection of genetic variations, including previously missed indels.
Related Concept Videos
Evolutionary Relationships through Genome Comparisons
6.7K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
6.7K
Sanger Sequencing
770.1K
DNA sequencing is a fundamental technique that is routinely used in the biological sciences. This method can be applied to a range of questions at different scales - from the sequencing of a cloned DNA fragment or the study of a mutation in a gene up to whole-genome sequencing. However, despite the widespread use of sequencing today, it was not until 1977 that Fredrick Sanger and his collaborators developed the chain-termination method to decode DNA sequences. It relies on the separation of a...
770.1K
Next-generation Sequencing
96.8K
The first human genome sequencing project cost $2.7 billion and was declared complete in 2003, after 15 years of international cooperation and collaboration between several research teams and funding agencies. Today, with the advent of next-generation sequencing technologies, the cost and time of sequencing a human genome have dropped over 100 fold.
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
96.8K
Comparing Copy Number Variations and SNPs
18.4K
Sequencing of the human genome has opened up several best-kept secrets of the genome. Scientists have identified thousands of genome variations that exist within a population. These variations can be a single nucleotide or a larger chromosomal variation.
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
18.4K

