Related Experiment Video
Updated: Jan 15, 2026

Author Spotlight: Real-Time Imaging of Bonding in 3D-Printed Layers
Published on: September 1, 2023
Benchmarking Large Language Models for Polymer Property Predictions
Sonakshi Gupta1, Akhlak Mahmood2, Shivank Shukla2
1School of Computational Science and Engineering, Georgia Institute of Technology, Atlanta, GA, 30332, USA.
None:
Machine learning and artificial intelligence have revolutionized polymer science by enhancing the ability to rapidly predict key polymer properties and enabling generative design. The utilization of large language models (LLMs) in polymer informatics may offers additional opportunities for advancement. Unlike traditional methods that depend on large labeled datasets, hand-crafted representations of the materials, and complex feature engineering, LLM-based methods utilize natural language inputs via a transfer learning process and eliminate the need for complex representation and fingerprinting, thus significantly simplifying the training process. In this study, we fine-tune general-purpose LLMs-open-source Llama-3-8B and commercial GPT-3.5-on a curated dataset of 11,740 entries to predict key thermal properties: glass transition ( ), melting ( ), and decomposition ( ) temperatures. Using parameter-efficient fine-tuning and hyperparameter optimization, we benchmark these models against traditional fingerprinting-based approaches including Polymer Genome, polyGNN, and polyBERT, under both single-task (ST) and multi-task (MT) learning frameworks. We find that while LLM informatics techniques can come close to traditional methods, they generally underperform in terms of predictive accuracy and computational efficiency. The fine-tuned Llama-3 model consistently outperforms GPT-3.5, likely due to the flexibility and tunability of the open-source architecture. Additionally, ST learning proves more effective than MT for LLMs, which struggle to exploit cross-property correlations-a significant and known advantage of traditional methods. The analysis of molecular embeddings learned by the models provides insight into the inner workings of the LLMs, revealing fundamental limitations of general-purpose LLMs in capturing nuanced chemo-structural information compared to the handcrafted features and domain-specific embeddings utilized in the traditional methods. These findings offer insights into the interplay between molecular embeddings and natural language processing, and provide guidance for LLM model selection within the context of polymer informatics.
Related Concept Videos
Polymers: Molecular Weight Distribution
Polymers: Defining Molecular Weight
The number average molecular weight (Mn) is the summation of the number...
Polymer Classification: Architecture
Molecular Weight of Step-Growth Polymers
As the step-growth polymerization involves step-wise condensation of monomers, the molecular weight also builds up eventually. Consequently, high molecular weight polymers are obtained at the late stages of the polymerization, where 99% of monomers have been consumed.
The extent of the...
Polymer Classification: Stereospecificity
Step-Growth Polymerization: Overview
Many natural and synthetic polymers are produced by...

