Related Experiment Video
Updated: Jun 27, 2025

08:05
Guided Protocol for Fecal Microbial Characterization by 16S rRNA-Amplicon Sequencing
Published on: March 19, 2018
19.7K
Guidelines for the Analysis of DNA Barcoding/Metabarcoding Sequencing Data and Interpretation of Publicly Available
Natalie Damaso1, Kyleen E Elwick2, James M Robertson3
1Counter WMD Systems, Massachusetts Institute of Technology, Lincoln Laboratory, Lexington, MA, USA.
Methods in Molecular Biology (Clifton, N.J.)
|April 29, 2024
Summary
This guide explains how to use DNA sequencing data for taxonomic identification with GenBank and Barcode of Life Data System (BOLD) databases. It covers data preparation, querying, and interpreting results for reliable species identification.
Area of Science:
- Molecular Biology
- Bioinformatics
- Taxonomy
Background:
- Accurate taxonomic identification is crucial for biological research and conservation.
- Public DNA sequence databases like GenBank and BOLD are essential resources for species identification.
- Standardized procedures are needed for effectively utilizing these databases.
Purpose of the Study:
- To provide a comprehensive guide on using DNA sequence data for taxonomic identification.
- To detail the process of preparing and uploading sequences to GenBank and BOLD.
- To explain how to query these databases and interpret the results for reliable identification.
Main Methods:
- Quality control and preparation of DNA sequences for database submission.
- Utilizing BLAST (Basic Local Alignment Search Tool) for GenBank queries.
- Employing the BOLD identification engine for Barcode of Life Data System queries.
- Developing guidelines for interpreting taxonomic assignments from database searches.
- Establishing protocols for assessing the accuracy and reliability of retrieved sequence data.
Main Results:
- Demonstrated procedures for preparing high-quality DNA sequences for public databases.
- Outlined effective methods for querying GenBank and BOLD using their respective identification engines.
- Provided clear guidelines for interpreting taxonomic identifications derived from sequence similarity.
- Presented a framework for evaluating the reliability of sequence data obtained from public repositories.
Conclusions:
- Standardized DNA sequencing and database querying protocols enhance taxonomic identification accuracy.
- Effective use of GenBank and BOLD facilitates reliable species identification in diverse biological studies.
- Critical evaluation of sequence data is essential for robust taxonomic conclusions.
Related Concept Videos
Evolutionary Relationships through Genome Comparisons
5.7K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
5.7K
RNA-seq
9.9K
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
9.9K
Sanger Sequencing
754.1K
DNA sequencing is a fundamental technique that is routinely used in the biological sciences. This method can be applied to a range of questions at different scales - from the sequencing of a cloned DNA fragment or the study of a mutation in a gene up to whole-genome sequencing. However, despite the widespread use of sequencing today, it was not until 1977 that Fredrick Sanger and his collaborators developed the chain-termination method to decode DNA sequences. It relies on the separation of a...
754.1K

