从重复到融合:扩展代霍夫的蛋白质进化模型
Yusran Abdillah Muthahari1, Lilian Magnus1, Paola Laurino1,2
1Protein Engineering and Evolution Unit, Okinawa Institute of Science and Technology, Okinawa, Japan.
Protein science : a publication of the Protein Society
|February 19, 2025
概括
蛋白质进化涉及基因重复和融合,从更简单的蛋白质中创建复杂的蛋白质. 融合在确定重复蛋白质的进化路径方面发挥着关键作用,从而导致新的功能.
科学领域:
- 进化生物学是进化的生物学.
- 分子生物学分子生物学
- 生物化学 生物化学
背景情况:
- 戴霍夫的假设认为,复杂的蛋白质来自更简单的,重复的和融合的或域.
- 基因重复在蛋白质进化中的作用得到了充分的研究,但蛋白质融合的影响不太清楚.
- 蛋白质进化扩大了生物的表现,使其功能多样化.
研究的目的:
- 突出蛋白质融合在进化过程中的关键作用.
- 强调融合是补充其他进化事件的关键机制,如基因重复.
- 探索融合如何影响蛋白质的功能多样化.
主要方法:
- 这是一个视角的作品,不是基于实验数据.
- 它综合了有关蛋白质进化,基因复制和蛋白质融合的现有知识.
- 重点是对进化机制的理论和概念分析.
主要成果:
- 蛋白质融合是蛋白质进化的重要驱动因素,与基因重复一起.
- 融合决定了重复的蛋白质单元 (原体) 的进化轨迹.
- 融合扩大了新型蛋白质功能和结构探索的潜力.
结论:
- 蛋白质融合对于蛋白质的进化能力至关重要,使得新的结构和功能空间的探索成为可能.
- 融合补充了突变,插入和删除在塑造蛋白质进化的过程中.
- 了解融合是全面了解蛋白质复杂性和多样性的关键.
相关概念视频
Gene Duplication and Divergence
6.0K
The seminal work of Ohno in 1970 popularized the idea of gene duplication and divergence. DNA sequence comparison studies reveal that a large portion of the genes in bacteria, archaebacteria, and eukaryotes was generated by gene duplication and divergence, indicating its critical role in evolution.
The duplicated copies of the gene are called Paralogs. Paralogs with similar sequences and functions form a gene family. Across several species, a large number of gene families are...
The duplicated copies of the gene are called Paralogs. Paralogs with similar sequences and functions form a gene family. Across several species, a large number of gene families are...
6.0K
Conservation of Protein Domains Over Different Proteins
10.7K
Protein domains are small structurally independent units that are part of a single amino acid chain. Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
10.7K
Gene Families
8.7K
Gene families consist of groups of genes proposed to have originated from a common ancestor. Typically these arise through events in which a gene or genes are mistakenly duplicated during cell division. Unlike their parent genes (which are subject to selection pressure to maintain function), these gene copies do not need to preserve their sequences and may evolve at a relatively faster rate.
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
8.7K
Conservation of Protein Domains
3.1K
3.1K
Protein Families
15.2K
Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism. Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members. If these new proteins contain similar amino acids in key...
15.2K
Protein Complexes with Interchangeable Parts
2.5K
Groups of proteins may form a complex where each protein in this complex has a different role in the overall execution of the complex’s function. Often some of the proteins in the complex can be replaced by a closely related variant to give a complex that contains many of the same components yet is functionally distinct.
The SCF ubiquitin ligase is a protein complex of five individual proteins. This complex attaches ubiquitin to other target proteins to mark them for degradation. In order...
The SCF ubiquitin ligase is a protein complex of five individual proteins. This complex attaches ubiquitin to other target proteins to mark them for degradation. In order...
2.5K


