PDB NextGen 档案:通过全球蛋白质数据库集中访问集成注释和丰富的结构信息
Preeti Choudhary1, Zukang Feng2, John Berrisford1
1Protein Data Bank in Europe, European Molecular Biology Laboratory, European Bioinformatics Institute Wellcome Genome Campus, Hinxton, Cambridgeshire, CB10 1SD, UK.
概括
蛋白质数据库 (PDB) 下一代档案集中了结构注释,改善了对生物分子数据的访问. 该资源通过提供来自可信来源的综合,最新信息来增强研究.
科学领域:
- 结构生物学 结构生物学
- 生物信息学是一种生物信息学.
- 数据归档数据归档
背景情况:
- 蛋白质数据库 (PDB) 是3D生物分子结构的全球存储库.
- 将外部注释集成到 PDB 中是具有挑战性的,因为它具有档案性质和分布式数据中心.
研究的目的:
- 建立一个集中式系统,以丰富结构注释.
- 简化访问PDB合作伙伴和外部生物数据资源的最新信息.
主要方法:
- 开发了PDB下一代 (NextGen) 档案的开发.
- 在PDB结构,UniProt序列和域注释 (Pfam,SCOP2,CATH) 之间集中映射.
- 包括分子内连接信息.
主要成果:
- 下一代档案提供集成数据,包括蛋白质结构,UniProt序列和域注释.
- 自推出以来,用户参与度大幅度,数据文件下载量超过350万.
- 确保研究人员能够获得准确,最新和易于访问的结构注释.
结论:
- 据了解,PDB NextGen 档案成功地集中并简化了对丰富结构注释的访问.
- 该档案通过向研究人员提供可靠和综合的生物分子数据来增强研究.
- 高用户参与度表明了NextGen档案在科学界的价值和实用性.
相关概念视频
Genome Annotation and Assembly
18.8K
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
18.8K
Next-generation Sequencing
88.7K
The first human genome sequencing project cost $2.7 billion and was declared complete in 2003, after 15 years of international cooperation and collaboration between several research teams and funding agencies. Today, with the advent of next-generation sequencing technologies, the cost and time of sequencing a human genome have dropped over 100 fold.
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
88.7K
Nucleic Acid Structure
6.1K
The pentose sugar in DNA is deoxyribose, while in RNA the pentose sugar is ribose. The difference between the sugars is the presence of the hydroxyl group on the ribose's second carbon and a hydrogen on the deoxyribose's second carbon. The phosphate residue attaches to the hydroxyl group of the 5′ carbon of one sugar and the hydroxyl group of the 3′ carbon of the sugar of the next nucleotide, which forms a 5′ to 3′ phosphodiester linkage.
DNA Structure
DNA...
DNA Structure
DNA...
6.1K
Multi-species Conserved Sequences
3.9K
Next-generation sequencing technologies have created large genomic databases of a variety of animals and plants. Ever since the human genome project was completed, scientists studied the genome of primates, mammals, and other phylogenetically distant living beings. Such large-scale studies have provided new insights into the evolutionary relationship between organisms.
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
3.9K
RNA-seq
9.9K
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
9.9K
Gene Families
8.8K
Gene families consist of groups of genes proposed to have originated from a common ancestor. Typically these arise through events in which a gene or genes are mistakenly duplicated during cell division. Unlike their parent genes (which are subject to selection pressure to maintain function), these gene copies do not need to preserve their sequences and may evolve at a relatively faster rate.
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
8.8K


