用Domain2GO对蛋白质域功能进行基于相互注释的预测
Erva Ulusoy1,2, Tunca Doğan1,2
1Biological Data Science Lab, Department of Computer Engineering, Hacettepe University, Ankara, Turkey.
概括
Domain2GO通过将蛋白质域与基因本体学 (GO) 术语联系起来来预测蛋白质功能. 这种计算方法有效地识别出未知的蛋白质作用,有助于生物学研究和疾病理解.
科学领域:
- 计算生物学 计算生物学
- 生物信息学是一种生物信息学.
- 蛋白质功能预测的预测
背景情况:
- 了解蛋白质功能对于健康和疾病研究至关重要.
- 蛋白质域是关键的功能单位,但实验性确定是昂贵和缓慢的.
- 计算方法是有效的蛋白质功能预测所需的.
研究的目的:
- 介绍Domain2GO,一种用于推断蛋白质功能的新计算方法.
- 将蛋白质功能预测重新定义为一个域函数预测问题.
- 在蛋白质域和基因本体学 (GO) 术语之间建立可靠的关联.
主要方法:
- 域2GO使用蛋白质水平的GO注释和域注释.
- 统计重新抽样识别了域名和GO术语之间的共同注释模式.
- 方法使用文献审查和功能评论评价注释3 (CAFA3) 数据集进行验证.
主要成果:
- 域2GO有效地预测分子功能和生物过程术语.
- 该方法以非常低的计算成本证明了可解释的结果.
- 通过域-GO术语传播成功应用在预测未知的蛋白质功能.
结论:
- Domain2GO提供了一种强大,高效和可解释的方法来预测蛋白质功能.
- 该方法可以扩展到其他生物本体和实体.
- 通过开源代码和用户友好的在线工具来访问,用于广泛的科学用途.
更多相关视频
12:04Interactome-Seq: A Protocol for Domainome Library Construction, Validation and Selection by Phage Display and Next Generation Sequencing
Published on: October 3, 2018
8.9K
06:50Author Spotlight: A Computational Approach to Decipher Amino Acid Preferences in Multispecific Protein-Protein Interactions
Published on: January 26, 2024
1.8K
相关概念视频
Conservation of Protein Domains Over Different Proteins
10.8K
Protein domains are small structurally independent units that are part of a single amino acid chain. Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
10.8K
Conservation of Protein Domains
3.1K
3.1K
Genome Annotation and Assembly
18.8K
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
18.8K
Conserved Binding Sites
4.2K
Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
4.2K
Protein Families
15.3K
Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism. Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members. If these new proteins contain similar amino acids in key...
15.3K
Protein Networks
3.9K
An organism can have thousands of different proteins, and these proteins must cooperate to ensure the health of an organism. Proteins bind to other proteins and form complexes to carry out their functions. Many proteins interact with multiple other proteins creating a complex network of protein interactions.
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
3.9K
