Related Experiment Video
Updated: May 6, 2026

10:36
Rare Event Detection Using Error-corrected DNA and RNA Sequencing
Published on: August 3, 2018
14.6K
Detecting Inter-Individual Contamination and Mismatches in Multiomics Next-Generation Sequencing Data
Zachary Whitfield1, Przemyslaw Kiljan2, Monika Krzyzanowska2
1Genentech Research and Early Development; Rancho Biosciences, LLC.
Journal of Visualized Experiments : Jove
|May 4, 2026
Summary
A new bioinformatics workflow accurately matches patient biosamples using genome-wide single-nucleotide polymorphism (SNP) comparisons. This method enhances the reliability of next-generation sequencing (NGS) data for clinical trials and biomarker discovery.
Area of Science:
- Genomics
- Bioinformatics
- Clinical Trials
Background:
- High-throughput sequencing of patient biosamples generates vast molecular data.
- Accurate sample tracking and matching are crucial for interpreting clinical trial results.
- Bioinformatics solutions can verify sample origin and improve data integrity.
Purpose of the Study:
- To present a bioinformatics workflow for identifying matched samples from the same individual.
- To enable robust verification of patient sample origin in next-generation sequencing (NGS) datasets.
- To support precise sample tracking throughout the biospecimen chain of custody.
Main Methods:
- Development of a genome-wide comparison scoring algorithm for sample verification.
- Utilizing single-nucleotide polymorphisms (SNPs) within linkage disequilibrium blocks for sample identification.
- Establishing threshold combinations for stringent and permissive matching of samples.
Main Results:
- Demonstrated utility in quality control and validation of clinical tumor and blood samples.
- Successfully applied to multi-omics data from over 2,000 patients.
- Identified specific SNP-based criteria for accurate sample matching.
Conclusions:
- The presented bioinformatics workflow reliably confirms sample origin, enhancing data accuracy.
- This method is applicable to various NGS datasets and clinical sample types.
- It provides a critical tool for robust interpretation of biomarker trial results and quality control.
Related Concept Videos
Next-generation Sequencing
88.0K
The first human genome sequencing project cost $2.7 billion and was declared complete in 2003, after 15 years of international cooperation and collaboration between several research teams and funding agencies. Today, with the advent of next-generation sequencing technologies, the cost and time of sequencing a human genome have dropped over 100 fold.
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
88.0K
Mismatch Repair
5.4K
Organisms are capable of detecting and fixing nucleotide mismatches that occur during DNA replication. This sophisticated process requires identifying the new strand and replacing the erroneous bases with correct nucleotides. Mismatch repair is coordinated by many proteins in both prokaryotes and eukaryotes.
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...
5.4K
Mismatch Repair
38.2K
Overview
38.2K
Comparing Copy Number Variations and SNPs
11.6K
Sequencing of the human genome has opened up several best-kept secrets of the genome. Scientists have identified thousands of genome variations that exist within a population. These variations can be a single nucleotide or a larger chromosomal variation.
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
11.6K

