通过多源域识别嵌入和特征投影网络预测circRNA-疾病关联
Si-Zhe Liang1, Lei Wang2,3, Zhu-Hong You4
1School of Information Engineering, Xijing Univerity, Xi'an 710123, China.
Journal of chemical information and modeling
|January 20, 2025
概括
这项研究介绍了MNDCDA,一种用于预测循环RNA (circRNA) -疾病关联的计算方法. 它有效地识别出新的联系,帮助疾病机制研究,降低实验成本.
科学领域:
- 生物化学 生物化学
- 基因组学就是基因组学.
- 计算生物学 计算生物学
背景情况:
- 循环RNAs (circRNAs) 在各种疾病中发挥着重要作用.
- 准确预测circRNA与疾病的关联对于理解疾病机制至关重要.
- 由于已知的关联有限和高的实验成本,现有的方法面临挑战.
研究的目的:
- 开发一种计算方法 (MNDCDA) 来预测circRNA与疾病的关联.
- 整合多个生物数据源,以提高预测准确度.
- 为识别新型circRNA与疾病的联系提供一个具有成本效益的工具.
主要方法:
- 使用全面的生物识别数据构建了四个相似性网络.
- 采用了社区意识嵌入模型来捕获结构信息.
- 利用深度特征投影网络进行高阶特征交互学习.
- 应用了双线解码器来识别新的circRNA疾病关联.
主要成果:
- 在一个基准数据集上,MNDCDA模型实现了0.9070的曲线下的面积 (AUC).
- 案例研究通过实验和文献验证了30个预测的circRNA-疾病对中的25个.
- 在预测circRNA疾病关联方面表现出强度和有效性.
结论:
- MNDCDA是一种强大的计算工具,用于预测circRNA与疾病的关联.
- 该方法为疾病机制提供了宝贵的见解.
- 它有效地降低了生物实验的成本和精力.
相关概念视频
Genome-wide Association Studies-GWAS
12.4K
Genome-wide association studies or GWAS are used to identify whether common SNPs are associated with certain diseases. Suppose specific SNPs are more frequently observed in individuals with a particular disease than those without the disease. In that case, those SNPs are said to be associated with the disease. Chi-square analysis is performed to check the probability of the allele likely to be associated with the disease.
GWAS does not require the identification of the target gene involved in...
GWAS does not require the identification of the target gene involved in...
12.4K
lncRNA - Long Non-coding RNAs
8.5K
In humans, more than 80% of the genome gets transcribed. However, only around 2% of the genome codes for proteins. The remaining part produces non-coding RNAs which includes ribosomal RNAs, transfer RNAs, telomerase RNAs, and regulatory RNAs, among other types. A large number of regulatory non-coding RNAs have been classified into two groups depending upon their length – small non-coding RNAs, such as microRNA, which are less than 200 nucleotides in length, and long non-coding RNA...
8.5K
Single Nucleotide Polymorphisms-SNPs
13.9K
A single nucleotide polymorphism or SNP is a single nucleotide variation at a specific genomic position in a large population. It is the most prevalent type of sequence variation found in the human genome. Point mutations that occur in more than 1% of the population qualify as SNPs. These are present once every 1000 nucleotides on an average in the human genome. Replacement of a purine with another purine (A/G) or a pyrimidine with another pyrimidine (C/T) is known as a transition. In contrast,...
13.9K


