AbTune: layer-wise selective fine-tuning of protein language models for antibodies

Xiaotong Xu1, Alexandre M J J Bonvin1

  • 1Bijvoet Centre for Biomolecular Research, Faculty of Science, Chemistry, Utrecht University, Heidelberglaan 8, 3584 CS Utrecht, Netherlands.

Antibodies play central roles in immune defense and are widely used as therapeutic agents. However, the high structural and sequence diversity of antigen-binding loops, combined with limited experimental data and weak co-evolutionary signals, makes it difficult to develop generalizable predictive models. In this work, we investigate test-time fine-tuning strategies to improve protein language model (pLM) performance in low-data settings, with a focus on antibody-related tasks. Systematic evaluations across tasks show that carefully constrained fine-tuning greatly enhances performance while preserving generalization. In particular, depth-selective fine-tuning consistently outperforms full-depth fine-tuning, with optimal performance achieved when tuning 50%-75% of model layers for medium- to small-sized pLMs. We introduce AbTune, a test-time fine-tuning framework that leverages this depth-controlled adaptation strategy. Across antibody structure prediction, mutation effect prediction, and binding affinity prediction, AbTune outperforms both standard pLM baselines and task-specific predictors, achieving the best performance among the evaluated baselines on two of the three tasks. To gain insight into the adaptation process and identify optimal AbTune protocols, we analyzed representation shifts, examined how sequence properties influence fine-tuning dynamics, and evaluated metrics that capture potential overfitting. Our results show that fine-tuning depth, duration, and perplexity jointly influence performance and must be carefully controlled to achieve optimal results.

Related Concept Videos

Antibody Structure01:10

Antibody Structure

Overview
Antibodies, also known as immunoglobulins (Ig), are essential players of the adaptive immune system. These antigen-binding proteins are produced by B cells and make up 20 percent of the total blood plasma by weight. In mammals, antibodies fall into five different classes, which each elicits a different biological response upon antigen binding.
The Y-Shaped Structure of Antibodies Consists of Four Polypeptide Chains
Antibodies consist of four polypeptide chains: two identical heavy...
Antibody Structure01:10

Antibody Structure

Overview
Antibodies, also known as immunoglobulins (Ig), are essential players of the adaptive immune system. These antigen-binding proteins are produced by B cells and make up 20 percent of the total blood plasma by weight. In mammals, antibodies fall into five different classes, which each elicits a different biological response upon antigen binding.
The Y-Shaped Structure of Antibodies Consists of Four Polypeptide Chains
Antibodies consist of four polypeptide chains: two identical heavy...
Affinity and Avidity01:41

Affinity and Avidity

Overview
Antibody Structure and Classes01:25

Antibody Structure and Classes

Antibodies, also known as immunoglobulins, are produced by B cells in response to foreign substances, such as bacteria and viruses. These proteins are critical for recognizing and neutralizing these substances, protecting the body from potential harm.
The basic structure of an antibody consists of four protein chains: two identical heavy chains and two identical light chains. These chains are held together by disulfide bonds and other non-covalent interactions, forming a Y-shaped structure.
Improving Translational Accuracy02:07

Improving Translational Accuracy

Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
Leaky Scanning02:28

Leaky Scanning

During most eukaryotic translation processes, the small 40S ribosome subunit scans an mRNA from its 5' end until it encounters the first start AUG codon. The large 60S ribosomal subunit then joins the smaller one to initiate protein synthesis. The location of the translation initiation is largely determined by the nucleotides near the start codon as there may be multiple translation initiation sites present on the mRNA.  Marilyn Kozak discovered that the sequence RCCAUGG (where R stands for...