Related Experiment Video
Updated: Feb 4, 2026

A Protocol for Computer-Based Protein Structure and Function Prediction
Published on: November 3, 2011
Coevolutionary Signals and Structure-Based Models for the Prediction of Protein Native Conformations
Ricardo Nascimento Dos Santos1, Xianli Jiang2, Leandro Martínez1
1Institute of Chemistry, University of Campinas (UNICAMP), Campinas, SP, Brazil.
Abstract:
The analysis of coevolutionary signals from families of evolutionarily related sequences is a recent conceptual framework that provides valuable information about unique intramolecular interactions and, therefore, can assist in the elucidation of biomolecular conformations. It is based on the idea that compensatory mutations at specific residue positions in a sequence help preserve stability of protein architecture and function and leave a statistical signature related to residue-residue interactions in the 3D structure of the protein. Consequently, statistical analysis of these correlated mutations in subsets of protein sequence alignments can be used to predict which residue pairs should be in spatial proximity in the native functional protein fold. These predicted signals can be then used to guide molecular dynamics (MD) simulations to predict the three-dimensional coordinates of a functional amino acid chain. In this chapter, we introduce a general and efficient methodology to perform coevolutionary analysis on protein sequences and to use this information in combination with computational physical models to predict the native 3D conformation of functional polypeptides. We present a step-by-step methodology that includes the description and application of software tools and databases required to infer tertiary structures of a protein fold. The general pipeline includes instructions on (1) how to obtain direct amino acid couplings from protein sequences using direct coupling analysis (DCA), (2) how to incorporate such signals as interaction potentials in Cα structure-based models (SBMs) to drive protein-folding MD simulations, (3) a procedure to estimate secondary structure and how to include such estimates in the topology files required in the MD simulations, and (4) how to build full atomic models based on the top Cα candidates selected in the pipeline. The information presented in this chapter is self-contained and sufficient to allow a computational scientist to predict structures of proteins using publicly available algorithms and databases.
Related Concept Videos
Protein and Protein Structure
A protein's shape is critical to its function. For example, an enzyme...
Conformity
Structural Protein Function
Collagen, the most abundant protein in mammals, is found throughout the body. In connective tissue, such as skin, ligaments, and tendons, it provides tensile strength and elasticity. In bones and teeth, it mineralizes to...
Structural Protein Function
Protein and Protein Structures
Predicting Molecular Geometry

