Related Experiment Video
Updated: Apr 5, 2026

Chromatin Immunoprecipitation ChIP of Histone Modifications from Saccharomyces cerevisiae
Published on: December 29, 2017
Contribution of Sequence Motif, Chromatin State, and DNA Structure Features to Predictive Models of Transcription
Zing Tsung-Yeh Tsai1, Shin-Han Shiu2, Huai-Kuang Tsai3
1Institute of Information Science, Academia Sinica, Taipei, Taiwan; Department of Plant Biology, Michigan State University, East Lansing, Michigan, United States of America.
Abstract:
Transcription factor (TF) binding is determined by the presence of specific sequence motifs (SM) and chromatin accessibility, where the latter is influenced by both chromatin state (CS) and DNA structure (DS) properties. Although SM, CS, and DS have been used to predict TF binding sites, a predictive model that jointly considers CS and DS has not been developed to predict either TF-specific binding or general binding properties of TFs. Using budding yeast as model, we found that machine learning classifiers trained with either CS or DS features alone perform better in predicting TF-specific binding compared to SM-based classifiers. In addition, simultaneously considering CS and DS further improves the accuracy of the TF binding predictions, indicating the highly complementary nature of these two properties. The contributions of SM, CS, and DS features to binding site predictions differ greatly between TFs, allowing TF-specific predictions and potentially reflecting different TF binding mechanisms. In addition, a "TF-agnostic" predictive model based on three DNA "intrinsic properties" (in silico predicted nucleosome occupancy, major groove geometry, and dinucleotide free energy) that can be calculated from genomic sequences alone has performance that rivals the model incorporating experiment-derived data. This intrinsic property model allows prediction of binding regions not only across TFs, but also across DNA-binding domain families with distinct structural folds. Furthermore, these predicted binding regions can help identify TF binding sites that have a significant impact on target gene expression. Because the intrinsic property model allows prediction of binding regions across DNA-binding domain families, it is TF agnostic and likely describes general binding potential of TFs. Thus, our findings suggest that it is feasible to establish a TF agnostic model for identifying functional regulatory regions in potentially any sequenced genome.
Related Concept Videos
Cis-regulatory Sequences
Cooperative Binding of Transcription Regulators
Chromatin Modification in iPS Cells
Compact chromatin makes reprogramming difficult. Enzymes, such as histone demethylases and acetyltransferases, are often added during reprogramming to loosen the chromatin, making the DNA more accessible to transcription factors. Molecules that inhibit histone...
Histone Modification
Acetylation
The enzyme histone acetyltransferase adds acetyl group to the histones. Another enzyme, histone...
Chromatin Structure Regulates pre-mRNA Processing
The chromatin structure, especially...
Co-activators and Co-repressors

