Related Experiment Video
Updated: Feb 3, 2026

DNA Sequence Recognition by DNA Primase Using High-Throughput Primase Profiling
Published on: October 8, 2019
A benchmark study of k-mer counting methods for high-throughput sequencing
Swati C Manekar1, Shailesh R Sathe1
1Department of Computer Science and Engineering, Visvesvaraya National Institute of Technology, Nagpur 440 010, India.
Abstract:
The rapid development of high-throughput sequencing technologies means that hundreds of gigabytes of sequencing data can be produced in a single study. Many bioinformatics tools require counts of substrings of length k in DNA/RNA sequencing reads obtained for applications such as genome and transcriptome assembly, error correction, multiple sequence alignment, and repeat detection. Recently, several techniques have been developed to count k-mers in large sequencing datasets, with a trade-off between the time and memory required to perform this function. We assessed several k-mer counting programs and evaluated their relative performance, primarily on the basis of runtime and memory usage. We also considered additional parameters such as disk usage, accuracy, parallelism, the impact of compressed input, performance in terms of counting large k values and the scalability of the application to larger datasets.We make specific recommendations for the setup of a current state-of-the-art program and suggestions for further development.
Related Concept Videos
Cis-regulatory Sequences
Sequences
Sanger Sequencing
Arithmetic Sequences
Methods for Studying Drug Absorption: In vitro
The diffusion cell method uses a two-compartment cell, including a donor compartment with the drug solution, which simulates the environment where the drug is applied, and a receptor compartment with a buffer solution, which simulates the environment...
Next-generation Sequencing
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....

