使用DNA动机预测区域体质突变率.
Cong Liu1, Zengmiao Wang2, Jun Wang1
1Department of Chemistry and Biochemistry, University of California San Diego, La Jolla, California, United States of America.
PLoS computational biology
|October 2, 2023
概括
科学家们使用DNA图案来预测13种癌症体内突变率. 他们确定了影响突变率的关键动机,并发现特定的基因组区域可以预测癌症类型,揭示了癌症发展的洞察力.
科学领域:
- 基因组学和表观遗传学
- 癌症研究 癌症研究
- 计算生物学 计算生物学
背景情况:
- 在表观遗传修饰中,局部特异性的调节尚未得到充分理解.
- 表观遗传酶通过识别序列动机 (epi-motifs) 的DNA结合因子被招募到特定的DNA位点.
- 通过预测身体突变率等生物输出,可以证实这些表观动机的功能性.
研究的目的:
- 为了研究DNA基因和癌症体质突变率之间的关系.
- 识别影响区域突变率的特定DNA基因.
- 为了确定具有高突变率的独特基因组区域是否可以预测癌症类型.
主要方法:
- 使用DNA图案,包括转录因子 (TF) 图案和表观图案,作为表观遗传信号的替代品.
- 应用了一个可解释的神经网络模型 (上下文回归) 来预测13种癌症类型的体质突变率,分辨率为23kbp.
- 分析了突变特征对癌症相关和癌症独立基因组区域的差异性贡献.
主要成果:
- 成功学习了DNA动机与区域突变率之间的通用关系.
- 确定了对突变速率产生影响的动机,如TP53和H3K9me3相关的表观动机.
- 发现具有显著更高突变率的特定基因组区域可以准确预测癌症类型,并确定了导致突变特征的动机.
结论:
- DNA 基因是体质突变率和表观遗传状态的强有力的预测因素.
- 该研究揭示了调节区域突变率的新型表观动机和TF动机.
- 识别与癌症相关的基因组区域及其相关动机为癌症类型的预测和理解突变过程提供了新的途径.
更多相关视频
相关概念视频
Gene Evolution - Fast or Slow?
7.2K
The genomes of eukaryotes are punctuated by long stretches of sequence which do not code for proteins or RNAs. Although some of these regions do contain crucial regulatory sequences, the vast majority of this DNA serves no known function. Typically, these regions of the genome are the ones in which the fastest change, in evolutionary terms, is observed, because there is typically little to no selection pressure acting on these regions to preserve their sequences.
In contrast, regions which code...
In contrast, regions which code...
7.2K
Spontaneous and Induced Mutations
36
Spontaneous mutations arise infrequently during DNA replication due to errors in the process. A key factor behind these errors is tautomeric shifts in nitrogenous bases, where bases transition from keto to enol forms or amino to imino forms. This shift can alter base-pairing rules, leading to mutations. Additionally, reactive oxygen species (ROS) arising from aerobic metabolism can damage DNA, resulting in depurination (loss of a purine base) or depyrimidination (loss of a pyrimidine base).
36
Mismatch Repair
4.9K
Organisms are capable of detecting and fixing nucleotide mismatches that occur during DNA replication. This sophisticated process requires identifying the new strand and replacing the erroneous bases with correct nucleotides. Mismatch repair is coordinated by many proteins in both prokaryotes and eukaryotes.
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...
4.9K
Nucleotide Excision Repair
3.5K
DNA Distortion and Damage
Cells are regularly exposed to mutagens—factors in the environment that can damage DNA and generate mutations. UV radiation is one of the most common mutagens and is estimated to introduce a significant number of changes in DNA. These include bends or kinks in the structure, which can block DNA replication or transcription. If these errors are not fixed, the damage can cause mutations, which in turn can result in cancer or disease depending on which sequences are...
Cells are regularly exposed to mutagens—factors in the environment that can damage DNA and generate mutations. UV radiation is one of the most common mutagens and is estimated to introduce a significant number of changes in DNA. These include bends or kinks in the structure, which can block DNA replication or transcription. If these errors are not fixed, the damage can cause mutations, which in turn can result in cancer or disease depending on which sequences are...
3.5K
Cancers Originate from Somatic Mutations in a Single Cell
12.0K
Cancer arises from mutations in the critical genes that allow healthy cells to escape cell cycle regulation and acquire the ability to proliferate indefinitely. Though originating from a single mutation event in one of the originator cells, cancer progresses when the mutant cell lines continue to gain more and more mutations, and finally, become malignant. For example, chronic myelogenous leukemia (CML) develops initially as a non-lethal increase in white blood cells, which progressively...
12.0K
Conserved Binding Sites
4.2K
Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
4.2K


