为可靠的全基因组合成检测提供无偏见的
Karl K Käther1,2, Andreas Remmel3,4, Steffen Lemke3,4
1Bioinformatics Group, Department of Computer Science, and Interdisciplinary Center for Bioinformatics, Leipzig University, Härtelstrasse 16-18, D-04017, Leipzig, Germany. karl@bioinf.uni-leipzig.de.
Algorithms for molecular biology : AMB
|April 5, 2025
概括
这项研究引入了一种新的无注释方法,用于 ортология推断,改善基因识别在比较基因组学. 新方法提高了准确性,特别是在密切相关的基因组中,为现有的基于注释的技术提供了有价值的替代方案.
科学领域:
- 基因组学就是基因组学.
- 生物信息学是一种生物信息学.
- 计算生物学 计算生物学
背景情况:
- 正理学推断对于比较基因组学至关重要,它可以识别来自共同祖先的基因.
- ортология推断中的挑战包括序列分歧,基因重复和基因组重组.
- 目前的方法通常依赖于基因组注释,限制了它们的适用性和准确性.
研究的目的:
- 开发和评估一个无注释的方法,用于 ортология推断.
- 将无注释方法的性能与基于注释的合成分析进行比较.
- 评估新算法在不同程度的基因组相关性的实用性.
主要方法:
- 开发一种新的算法用于 ортология推断,不需要预先存在的基因组注释.
- 对无注释方法与使用基因组注释的传统合成分析进行比较分析.
- 对近距离相关基因组的性能评估.
主要成果:
- 没有注释的方法在密切相关的基因组中表现出卓越的性能.
- 基于注释的合成分析显示,更远的相关基因组的性能更好.
- 开发的算法为依赖注释的方法提供了强大的替代方案.
结论:
- 新的无注释的正统学推断方法是对比较基因组学工具的宝贵补充.
- 这种方法可以在特定场景中优于现有的基于注释的方法,特别是在密切相关的物种中.
- 该算法提供了灵活性和更高的准确性,推进了基因组研究领域.
相关概念视频
Synteny and Evolution
3.2K
John H. Renwick first coined the term “synteny” in 1971, which refers to the genes present on the same chromosomes, even if they are not genetically linked. The species with common ancestry tend to show conserved syntenic regions. Therefore, the concept of synteny is nowadays used to describe the evolutionary relationship between species.
Around 80 million years ago, the human and mice lineages diverged from the common ancestor. During the course of evolution, the ancestral...
Around 80 million years ago, the human and mice lineages diverged from the common ancestor. During the course of evolution, the ancestral...
3.2K
Multi-species Conserved Sequences
3.9K
Next-generation sequencing technologies have created large genomic databases of a variety of animals and plants. Ever since the human genome project was completed, scientists studied the genome of primates, mammals, and other phylogenetically distant living beings. Such large-scale studies have provided new insights into the evolutionary relationship between organisms.
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
3.9K
Genome Annotation and Assembly
18.7K
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
18.7K
Condensins
3.2K
Condensins are large protein complexes that use ATP to fuel the assembly of chromosomes during mitosis. They transform the tangled, shapeless mass of post-interphase DNA into individualized chromosomes by compacting, organizing, and segregating chromosomal DNA.
The plant and animal cells contain two types of condensin complexes—condensin I and condensin II. Both complexes have five subunits: two SMC (Structural Maintenance of Chromosomes) subunits, a kleisin subunit, and two HEAT-repeat...
The plant and animal cells contain two types of condensin complexes—condensin I and condensin II. Both complexes have five subunits: two SMC (Structural Maintenance of Chromosomes) subunits, a kleisin subunit, and two HEAT-repeat...
3.2K
Conserved Binding Sites
4.1K
Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
4.1K


