Related Experiment Video
Updated: Jan 28, 2026

Amplicon Sequencing using the Long-Read Sequencing Technologies
Published on: August 29, 2025
vi-HMM: a novel HMM-based method for sequence variant identification in short-read data.
Man Tang1, Mohammad Shabbir Hasan2, Hongxiao Zhu1
1Department of Statistics, Virginia Tech, 250 Drillfield Drive, Blacksburg, 24061, VA, USA.
This study introduces vi-HMM, a novel hidden Markov model for accurate variant calling in next-generation sequencing data. It improves the identification of single nucleotide polymorphisms (SNPs) and insertion-deletion polymorphisms (INDELs) by considering linkage disequilibrium.
Area of Science:
- Genomics
- Bioinformatics
- Computational Biology
Background:
- Accurate identification of sequence variants like single nucleotide polymorphisms (SNPs) and insertion-deletion polymorphisms (INDELs) is crucial for next-generation sequencing (NGS).
- Current variant calling methods often assume positional independence and do not utilize the linkage disequilibrium (LD) between nearby loci.
Purpose of the Study:
- To develop a novel hidden Markov model (HMM)-based method for accurate SNP and INDEL calling in mapped short-read sequencing data.
- To leverage the dependence between genotypes at nearby loci caused by LD for improved variant identification.
Main Methods:
- Proposed vi-HMM, a hidden Markov model (HMM) that allows transitions between hidden states (SNP, Ins, Del, Match) of adjacent genomic bases.
- Utilized the Viterbi algorithm to determine the optimal hidden state path for identifying SNPs and INDELs.
Main Results:
- Simulation studies demonstrated that vi-HMM outperforms existing variant calling methods in sensitivity and F1 score across various sequencing depths.
- Application to real data confirmed vi-HMM's higher accuracy in calling both SNPs and INDELs.
Conclusions:
- vi-HMM provides a more accurate and reliable approach for variant calling in NGS data compared to existing methods.
- The method effectively accounts for linkage disequilibrium, leading to improved SNP and INDEL identification.
More Related Videos
08:23De novo Identification of Actively Translated Open Reading Frames with Ribosome Profiling Data
Published on: February 18, 2022
12:08Hybrid De Novo Genome Assembly for the Generation of Complete Genomes of Urinary Bacteria using Short- and Long-read Sequencing Technologies
Published on: August 20, 2021
Related Concept Videos
Histone Variants at the Centromere
Methods of Classification and Identification
Uncertainty in Measurement: Reading Instruments
Cis-regulatory Sequences
Statistical Methods for Analyzing Epidemiological Data
Statistical Methods to Analyze Parametric Data: ANOVA
One-way ANOVA is applied when a single independent variable or factor is scrutinized. It compares...