Related Experiment Video
Updated: Mar 11, 2026

Author Spotlight: Investigating the Role of Repetitive DNA Misregulation in Cancer Initiation and Immunotherapy Resistance
Published on: December 13, 2024
Managing Sequence Data
Christopher O'Sullivan1, Benjamin Busby1, Ilene Karsch Mizrachi2
1National Center for Biotechnology Information, National Library of Medicine, National Institutes of Health, Bethesda, MD, USA.
Abstract:
Nucleotide and protein sequences are the foundation for all bioinformatics tools and resources. Researchers can analyze these sequences to discover genes or predict the function of their products. The INSDC (International Nucleotide Sequence Database-DDBJ/ENA/GenBank + SRA) is an international, centralized primary sequence resource that is freely available on the Internet. This database contains all publicly available nucleotide and derived protein sequences. This chapter discusses the structure and history of the nucleotide sequence database resources built at NCBI, provides information on how to submit sequences to the databases, and explains how to access the sequence data.
Related Concept Videos
RNA-seq
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
Maxam-Gilbert Sequencing
Challenges of the Maxam-Gilbert Method
The...
Next-generation Sequencing
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Multi-species Conserved Sequences
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
Sanger Sequencing
Evolutionary Relationships through Genome Comparisons

