ReAlign-N:用于多个核酸序列对齐的综合调整方法,结合全球和本地调整
Yixiao Zhai1,2, Tong Zhou1,2, Yanming Wei2,3
1Institute of Fundamental and Frontier Sciences, University of Electronic Science and Technology of China, No.2006, Xiyuan Avenue, Pidu Zone, Chengdu 610054, China.
NAR genomics and bioinformatics
|December 20, 2024
概括
ReAlign-N通过整合全球和本地策略来提高核酸序列对齐的准确性. 这种新方法比现有的大规模分析工具更快,使用更少的内存.
科学领域:
- 生物信息学是一种生物信息学.
- 计算生物学 计算生物学
- 基因组学就是基因组学.
背景情况:
- 精确的多重序列对齐 (MSA) 对于生物序列分析至关重要.
- 现有的工具在核酸序列的进化变化中扎,需要重新调整.
- 对长核酸序列缺乏专门的调整方法.
研究的目的:
- 介绍ReAlign-N,一种用于多重核酸序列对齐的新型调整方法.
- 为了提高核酸序列对齐的准确性和效率.
- 解决当前工具在处理复杂的进化关系方面的局限性.
主要方法:
- ReAlign-N结合了全球和地方的调整策略.
- 全球调整使用K频段和节省内存的动态编程来提高效率.
- 局部调整采用完全匹配和得分,使用MAFFT进行改进.
主要成果:
- ReAlign-N在模拟和真实数据集上的初始对齐方面表现出优越的性能.
- 该方法在多个核酸序列对齐中实现了更高的准确性.
- 与ReformAlign相比,ReAlign-N显著减少了运行时间和内存使用量.
结论:
- ReAlign-N为准确的多重核酸序列调整提供了一个有效的解决方案.
- 该工具高效且节省内存,适合大规模数据集.
- ReAlign-N通过提供用于核酸序列分析的专用工具来推进生物信息学领域.
相关概念视频
Evolutionary Relationships through Genome Comparisons
5.7K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
5.7K
Nucleic Acid Structure
5.9K
The pentose sugar in DNA is deoxyribose, while in RNA the pentose sugar is ribose. The difference between the sugars is the presence of the hydroxyl group on the ribose's second carbon and a hydrogen on the deoxyribose's second carbon. The phosphate residue attaches to the hydroxyl group of the 5′ carbon of one sugar and the hydroxyl group of the 3′ carbon of the sugar of the next nucleotide, which forms a 5′ to 3′ phosphodiester linkage.
DNA Structure
DNA...
DNA Structure
DNA...
5.9K
RNA-seq
9.8K
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while...
9.8K
Next-generation Sequencing
87.3K
The first human genome sequencing project cost $2.7 billion and was declared complete in 2003, after 15 years of international cooperation and collaboration between several research teams and funding agencies. Today, with the advent of next-generation sequencing technologies, the cost and time of sequencing a human genome have dropped over 100 fold.
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
Next-Generation Sequencing Methods
Although all next-generation methods use different technologies, they all share a set of standard features....
87.3K
Multi-species Conserved Sequences
3.9K
Next-generation sequencing technologies have created large genomic databases of a variety of animals and plants. Ever since the human genome project was completed, scientists studied the genome of primates, mammals, and other phylogenetically distant living beings. Such large-scale studies have provided new insights into the evolutionary relationship between organisms.
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
Although the genome of each species varies greatly from each other, a few sequences are highly conserved. Such conserved...
3.9K
Genome Annotation and Assembly
18.8K
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
18.8K


