Related Experiment Video
Updated: Sep 8, 2025

Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers
Published on: March 1, 2024
Multiview Deep Learning Framework for Precise Prediction of Transcription Factor Binding Sites
Yiben Lin1,2, Huiliang Luo2, Liang Yan3
1Key Laboratory of Micro-nano Sensing and IoT of Wenzhou, Wenzhou Institute of Hangzhou Dianzi University, Wenzhou 325038, China.
Abstract:
Transcription factors (TFs) are essential proteins that regulate gene expression by specifically binding to transcription factor binding sites (TFBSs) within DNA sequences. Their ability to precisely control the transcription process is crucial for understanding gene regulatory networks, uncovering disease mechanisms, and designing synthetic biology tools. Accurate TFBS prediction, therefore, holds significant importance in advancing these areas of research. While machine learning methods, particularly deep learning approaches, have achieved notable progress in TFBS prediction in recent years, several challenges persist. These include modeling the intricate structural features of the DNA double helix, capturing long-range dependencies within sequences and integrating diverse biological data sources. To address these issues, we propose an innovative solution known as multiview deep learning for Transcription Factor Binding Prediction (MDNet-TFP), which leverages multiple views of DNA sequences─including different representational forms and diverse processing strategies─to enhance prediction capabilities. Specifically, our framework introduces a bidirectional reverse complement module (BiRC-Mamba) that effectively accounts for the bidirectional and reverse complement properties characteristic of DNA sequences. Furthermore, we developed a multiscale convolutional recurrent attention network (MCRAN) that extracts both structural and functional DNA features across multiple dimensions while integrating information from various biological data sets. These advancements allow our model to outperform existing methods across 165 ChIP-seq data sets, achieving an average ACC of 88.13% (±0.47), an ROC-AUC of 93.72% (±0.15), and a PR-AUC of 93.40% (±0.21). The model not only excels with this specific data set but also maintains its high performance across a wider array of 690 ChIP-seq data sets. To further validate the model's effectiveness, we employ motif visualization techniques. This approach reveals that the regions receiving high attention from our model align with known transcription factor binding motifs, offering valuable biological insights. Additionally, this correspondence substantiates the model's ability to generalize and interpret complex genomic data effectively. By addressing critical limitations in the field, MDNet-TFP offers a promising new avenue for advancing research in transcriptional regulation and biomedical applications.
More Related Videos
Related Concept Videos
Cooperative Binding of Transcription Regulators
General Transcription Factors
Transcription Factors
Conserved Binding Sites
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
Master Transcription Regulators
Combinatorial Gene Control
The expression of more than 30,000 genes is controlled by approximately 2000-3000 transcription factors. This is possible because a single transcription factor can recognize more than one regulatory sequence. The specificity in gene...

