Hypervariable loop profiling decodes sequence determinants of antibody stability

Yue Wan1, Jiahao Liang1, Yile Dai1

  • 1Department of Pharmaceutical Chemistry, University of California, San Francisco, CA, USA.

Antibody folding and aggregation are major challenges in the development of relevant reagents and therapeutics. Antibodies face a biophysical trade-off; the immense diversity in complementarity-determining regions (CDRs), which is crucial for broad antigen recognition, comes at the cost of folding stability. How CDR sequences influence antibody folding remains poorly understood because of their sequence diversity and lack of large-scale data. Here we develop a high-throughput 'deep loop profiling' approach to quantify folding fitness across millions of diverse CDRs. Machine learning models trained on this dataset predict folding propensity directly from sequence and identify interpretable residue-level rules that reveal CDR1 and CDR2 as key folding determinants. Using these insights, we rescue two unstable nanobodies, including an aggregation-prone SARS-CoV-2 binder and a G-protein-coupled receptor-targeting intrabody, and build next-generation synthetic libraries enriched for biophysically optimized nanobodies. This approach provides a scalable framework for understanding and engineering folding competence in antibody-based scaffolds.

Related Concept Videos

Antibody Structure01:10

Antibody Structure

12.0K
Antibody Structure01:10

Antibody Structure

Overview
Antibodies, also known as immunoglobulins (Ig), are essential players of the adaptive immune system. These antigen-binding proteins are produced by B cells and make up 20 percent of the total blood plasma by weight. In mammals, antibodies fall into five different classes, which each elicits a different biological response upon antigen binding.
The Y-Shaped Structure of Antibodies Consists of Four Polypeptide Chains
Antibodies consist of four polypeptide chains: two identical heavy...
51.8K
Conserved Binding Sites01:49

Conserved Binding Sites

Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
4.1K
Antibody Structure and Classes01:25

Antibody Structure and Classes

Antibodies, also known as immunoglobulins, are produced by B cells in response to foreign substances, such as bacteria and viruses. These proteins are critical for recognizing and neutralizing these substances, protecting the body from potential harm.
The basic structure of an antibody consists of four protein chains: two identical heavy chains and two identical light chains. These chains are held together by disulfide bonds and other non-covalent interactions, forming a Y-shaped structure.
7.7K
Conservation of Protein Domains Over Different Proteins02:26

Conservation of Protein Domains Over Different Proteins

Protein domains are small structurally independent units that are part of a single amino acid chain.  Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
11.8K