Related Experiment Video
Updated: Mar 18, 2026

G2-seq: A High Throughput Sequencing-based Technique for Identifying Late Replicating Regions of the Genome
Published on: March 22, 2018
TagDigger: user-friendly extraction of read counts from GBS and RAD-seq data
Lindsay V Clark1, Erik J Sacks1
1Department of Crop Sciences, University of Illinois at Urbana-Champaign, 1201 W. Gregory Drive, Urbana, IL 61802 USA.
TagDigger software improves genotyping-by-sequencing (GBS) and RAD-seq analysis by providing accurate read counts and genotypes. This open-source tool facilitates SNP analysis across projects and simplifies data preparation for public archives.
Area of Science:
- Genomics
- Bioinformatics
- Population Genetics
Background:
- Read depth is crucial for genotype quality and allele dosage estimation in genotyping-by-sequencing (GBS) and restriction site-associated DNA sequencing (RAD-seq).
- Existing GBS and RAD-seq pipelines lack accurate and accessible read count formats.
- Current pipelines offer limited flexibility in selecting specific loci for analysis and struggle with standardized SNP naming for cross-project meta-analysis.
Purpose of the Study:
- To develop user-friendly software for accurate read count extraction and genotype calling in GBS and RAD-seq data.
- To enable flexible analysis by allowing users to specify subsets of loci for examination.
- To facilitate cross-project meta-analysis by consolidating marker names and sequences.
Main Methods:
- Developed TagDigger, a suite of three Python 3 scripts: tagdigger_interactive.py, tag_manager.py, and barcode_splitter.py.
- tagdigger_interactive.py extracts read counts and genotypes from FASTQ files, supporting user-defined barcodes and tags, with CSV input/output.
- tag_manager.py consolidates marker information across projects, while barcode_splitter.py prepares FASTQ files for data archival.
Main Results:
- TagDigger provides accurate read counts and genotypes in an accessible CSV format.
- The software supports importing tag sequences from popular pipelines (Stacks, TASSEL-GBSv2, TASSEL-UNEAK, pyRAD) and allows manual selection of loci.
- TagDigger processes over 100 million FASTQ reads per hour using a scalable, rapid search algorithm.
Conclusions:
- TagDigger is an open-source, freely available software package that enhances GBS and RAD-seq data analysis.
- It offers a user-friendly interface requiring no programming skills and runs on standard operating systems.
- The software efficiently processes large datasets without generating excessive intermediate files, making it suitable for laptops.
More Related Videos
05:07Rup (RNA-seq Usability Assessment Pipeline) - Quality Control for Bulk RNA-seq Experiments in Eukaryotes
Published on: November 7, 2025
04:58Author Spotlight: Investigating the Role of Repetitive DNA Misregulation in Cancer Initiation and Immunotherapy Resistance
Published on: December 13, 2024