DDBJ in the stream of various biological data
S Miyazaki1, H Sugawara, K Ikeo
1Center for Information Biology and DNA Data Bank of Japan, National Institute of Genetics, Yata, Mishima 411-8540, Japan.
Nucleic Acids Research
|December 19, 2003
Summary
The DNA Data Bank of Japan (DDBJ) has expanded its biological data resources, increasing submissions of human, ascidian, and rice genome data. New tools and services facilitate data access and integration worldwide.
Area of Science:
- Bioinformatics
- Genomics
- Molecular Biology
Background:
- The DNA Data Bank of Japan (DDBJ) is a crucial repository for biological data.
- Continuous growth in biological data necessitates efficient management and accessibility.
Purpose of the Study:
- To report on the DDBJ's activities and advancements over the past year.
- To highlight the expansion of data submissions and the development of new data retrieval and access tools.
Main Methods:
- Analysis of data submission statistics, including increments in bases and entries.
- Development and implementation of new bioinformatics tools and web services.
- Expansion of public database offerings, such as the CIBEX gene expression database.
Main Results:
- A 50.6% increase in the number of bases and a 46.5% increase in the number of entries submitted to DDBJ.
- Genome data from human, ascidian, and rice represent the top three submissions.
- Introduction of a regular expression-based sequence retrieval tool, SOAP server, and web services.
- Launch of the public gene expression database, CIBEX.
Conclusions:
- DDBJ has significantly expanded its data holdings and improved data accessibility.
- The new tools and services enhance the utility of DDBJ resources for the global research community.
- DDBJ continues to play a vital role in managing and disseminating large-scale biological data.
Related Concept Videos
Genomic DNA in Eukaryotes
Eukaryotes have large genomes compared to prokaryotes. To fit their genomes into a cell, eukaryotic DNA is packaged extraordinarily tightly inside the nucleus. To achieve this, DNA is tightly wound around proteins called histones, which are packaged into nucleosomes that are joined by linker DNA and coil into chromatin fibers. Additional fibrous proteins further compact the chromatin, which is recognizable as chromosomes during certain phases of cell division.
Complementary DNA
Overview
Genomics
Genomics is the science of genomes: it is the study of all the genetic material of an organism. In humans, the genome consists of information carried in 23 pairs of chromosomes in the nucleus, as well as mitochondrial DNA. In genomics, both coding and non-coding DNA is sequenced and analyzed. Genomics allows a better understanding of all living things, their evolution, and their diversity. It has a myriad of uses: for example, to build phylogenetic trees, to improve productivity and...
Exon Recombination
The evolution of new genes is critical for speciation. Exon recombination, also known as exon shuffling or domain shuffling, is an important means of new gene formation. It is observed across vertebrates, invertebrates, and in some plants such as potatoes and sunflowers. During exon recombination, exons from the same or different genes recombine and produce new exon-intron combinations, which might evolve into new genes.
Exon shuffling follows “splice frame rules.” Each exon has three reading...
Exon shuffling follows “splice frame rules.” Each exon has three reading...
DNA Microarrays
Microarrays are high-throughput and relatively inexpensive assays that can be automated to analyze large quantities of data at a time. They are used in genome-wide studies to compare gene or protein expression under two varied conditions, such as healthy and diseased states. Microarrays consist of glass or silica slides on which probe molecules are covalently attached through surface functionalization. Most commonly, the slides are prepared through the chemisorption of silanes to silica...
Genome Annotation and Assembly
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.


