DAVID Ortholog:一个整合性工具,通过ortologs来增强功能分析
Brad T Sherman1, Ganesh Panzade1, Tomozumi Imamichi1
1Laboratory of Human Retrovirology and Immunoinformatics, Frederick National Laboratory for Cancer Research, Frederick, MD 21702, United States.
Bioinformatics (Oxford, England)
|October 16, 2024
概括
DAVID Ortholog使物种之间的基因列表转换成为可能,通过利用模型生物数据来增强功能解释. 这种工具有助于跨物种分析,以获得更深入的生物学见解.
科学领域:
- 生物信息学是一种生物信息学.
- 基因组学就是基因组学.
- 计算生物学 计算生物学
背景情况:
- 标注,可视化和综合发现数据库 (DAVID) 是一种广泛使用的生物信息学工具,用于解释高通量实验中的基因列表.
- 目前的DAVID分析通常仅限于研究物种,当基因数据不完整或不可用时,这可能是一个限制.
- 在广泛研究的模型生物的背景下分析基因列表可以为生物学主题和基因功能提供宝贵的见解.
研究的目的:
- 开发一种方法来转换不同物种之间的基因列表,以克服数据限制.
- 通过在具有良好的特征模型生物的背景下进行分析来增强基因列表的功能解释.
- 将跨物种分析无整合到现有的DAVID工作流程中.
主要方法:
- 开发了DAVID Ortholog,这是一个用于物种之间转换基因列表的工具.
- 使用来自Orthologous MAtrix (OMA) 和Ensembl Compara的正统数据进行转换.
- 集成的正统识别器ID配对信息进入DAVID知识库.
主要成果:
- DAVID Ortholog成功地将用户提供的基因列表转换为所需物种的ortolog列表.
- 该工具有助于在目标物种的背景下无地继续下游的DAVID分析.
- 允许用户通过功能注释更好地了解他们的基因列表的生物含义.
结论:
- DAVID Ortholog 扩大了 DAVID 的实用性,通过实现跨物种基因列表分析.
- 该工具通过利用模型生物体的功能注释来增强生物解释.
- 促进对不同物种的基因功能和生物通路的更深入的理解.
相关概念视频
Genome Annotation and Assembly
18.8K
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
18.8K
Evolutionary Relationships through Genome Comparisons
5.7K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
5.7K
Protein Networks
3.9K
An organism can have thousands of different proteins, and these proteins must cooperate to ensure the health of an organism. Proteins bind to other proteins and form complexes to carry out their functions. Many proteins interact with multiple other proteins creating a complex network of protein interactions.
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
3.9K
Epistasis Analysis
4.9K
Although Mendel chose seven unrelated traits in peas to study gene segregation, most traits involve multiple gene interactions that create a spectrum of phenotypes. When the interaction of various genes or alleles at different locations influences a phenotype, this is called epistasis. Epistasis often involves one gene masking or interfering with the expression of another (antagonistic epistasis). Epistasis often occurs when different genes are part of the same biochemical pathway. The...
4.9K
Gene Families
8.8K
Gene families consist of groups of genes proposed to have originated from a common ancestor. Typically these arise through events in which a gene or genes are mistakenly duplicated during cell division. Unlike their parent genes (which are subject to selection pressure to maintain function), these gene copies do not need to preserve their sequences and may evolve at a relatively faster rate.
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
Occasionally these regions can be adapted to take on new roles within the organism, becoming novel genes...
8.8K
Gene Evolution - Fast or Slow?
7.0K
The genomes of eukaryotes are punctuated by long stretches of sequence which do not code for proteins or RNAs. Although some of these regions do contain crucial regulatory sequences, the vast majority of this DNA serves no known function. Typically, these regions of the genome are the ones in which the fastest change, in evolutionary terms, is observed, because there is typically little to no selection pressure acting on these regions to preserve their sequences.
In contrast, regions which code...
In contrast, regions which code...
7.0K


