Related Experiment Video
Updated: Jan 16, 2026

A Protocol for Computer-Based Protein Structure and Function Prediction
Published on: November 3, 2011
Efficient inference, training, and fine-tuning of protein language models
Muhammed Hasan Çelik1,2, Xiaohui Xie1
1Department of Computer Science, University of California, Irvine, Irvine, CA, USA.
None:
Protein language models (PLMs) have shown great promise in protein structure and function predictions, but their adoption is limited by computational cost. We address this challenge by enhancing the efficiency of evolutionary scale modeling (ESM). Using FlashAttention and sequence packing, we achieve 4-9× faster inference and 3-14× lower memory usage. Four-bit quantization of billion-parameter models further reduces memory by 2-3× while preserving accuracy for missense variant effect prediction. Training is also optimized, cutting runtime 6-fold with methods, such as activation checkpointing and DeepSpeed zero-offload. Parameter-efficient fine-tuning of a few adapter weights yields state-of-the-art performance at protein property and function predictions, resulting in 70% Spearman's correlation for melting point and 87% AU-PRC for transcription factor identification. Our efficient ESM (ESME) implementation significantly lowers the barrier to using these powerful models, making them accessible to academic laboratories with limited computational resources. The code is available on GitHub.
More Related Videos
Related Concept Videos
Improving Translational Accuracy
Improving Translational Accuracy
Leaky Scanning
Conservation of Protein Domains Over Different Proteins
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
Termination of Translation
Proteins: From Genes to Degradation
Transcription is the synthesis of RNA...

