SearcharxivSearch

arXiv subjects

Shivank Shukla

Publications and source records attributed to Shivank Shukla.

4 recordsLinked to original sources

On-the-Fly Machine-Learned Force Fields for High-Fidelity Polymer Glass Transition Simulations

Predicting polymer glass transition temperatures (Tg) with first-principles fidelity has long remained out of reach, as cooling multi-thousand-atom systems over a broad temperature range at acceptable rates exceeds the computational limits of ab initio molecular dynamics (AIMD). Here we employ a hybrid scheme that merges AIMD with accelerated on-the-fly (OTF) machine-learned force-field (MLFF) construction, enabling Tg prediction at quantum-mechanical accuracy with near-classical computational cost. The OTF protocol to construct MLFFs adaptively triggers first-principles calculations only when newly encountered configurations lie outside the current model's domain of confidence, allowing robust, parameter-free MLFFs to be built from merely 1000 AIMD-sampled configurations per polymer. These MLFFs are then utilized to perform long-time cooling simulations on amorphous supercells containing several thousand atoms. Applied across twelve polymers spanning aromatic, aliphatic, heteroatomic, and branched chemistries, the method yields predictions in excellent accord with experiment while reducing computational cost by approximately six orders of magnitude relative to AIMD. This work establishes a new paradigm for predictive polymer modeling, demonstrating that OTF-MLFFs provide a generalizable, accurate, and scalable route to simulating the thermophysical behavior of complex disordered materials at near quantum-mechanical fidelity.

cond-mat.mtrl-sci

Benchmarking Large Language Models for Polymer Property Predictions

Machine learning has revolutionized polymer science by enabling rapid property prediction and generative design. Large language models (LLMs) offer further opportunities in polymer informatics by simplifying workflows that traditionally rely on large labeled datasets, handcrafted representations, and complex feature engineering. LLMs leverage natural language inputs through transfer learning, eliminating the need for explicit fingerprinting and streamlining training. In this study, we finetune general purpose LLMs -- open-source LLaMA-3-8B and commercial GPT-3.5 -- on a curated dataset of 11,740 entries to predict key thermal properties: glass transition, melting, and decomposition temperatures. Using parameter-efficient fine-tuning and hyperparameter optimization, we benchmark these models against traditional fingerprinting-based approaches -- Polymer Genome, polyGNN, and polyBERT -- under single-task (ST) and multi-task (MT) learning. We find that while LLM-based methods approach traditional models in performance, they generally underperform in predictive accuracy and efficiency. LLaMA-3 consistently outperforms GPT-3.5, likely due to its tunable open-source architecture. Additionally, ST learning proves more effective than MT, as LLMs struggle to capture cross-property correlations, a key strength of traditional methods. Analysis of molecular embeddings reveals limitations of general purpose LLMs in representing nuanced chemo-structural information compared to handcrafted features and domain-specific embeddings. These findings provide insight into the interplay between molecular embeddings and natural language processing, guiding LLM selection for polymer informatics.

cs.CE

polyBART: A Chemical Linguist for Polymer Property Prediction and Generative Design

Designing polymers for targeted applications and accurately predicting their properties is a key challenge in materials science owing to the vast and complex polymer chemical space. While molecular language models have proven effective in solving analogous problems for molecular discovery, similar advancements for polymers are limited. To address this gap, we propose polyBART, a language model-driven polymer discovery capability that enables rapid and accurate exploration of the polymer design space. Central to our approach is Pseudo-polymer SELFIES (PSELFIES), a novel representation that allows for the transfer of molecular language models to the polymer space. polyBART is, to the best of our knowledge, the first language model capable of bidirectional translation between polymer structures and properties, achieving state-of-the-art results in property prediction and design of novel polymers for electrostatic energy storage. Further, polyBART is validated through a combination of both computational and laboratory experiments. We report what we believe is the first successful synthesis and validation of a polymer designed by a language model, predicted to exhibit high thermal degradation temperature and confirmed by our laboratory measurements. Our work presents a generalizable strategy for adapting molecular language models to the polymer space and introduces a polymer foundation model, advancing generative polymer design that may be adapted for a variety of applications.

cond-mat.soft

Informatics-Driven Selection of Polymers for Fuel-Cell Applications

Modern fuel cell technologies use Nafion as the material of choice for the proton exchange membrane (PEM) and as the binding material (ionomer), used to assemble the catalyst layers of the anode and cathode. These applications demand high proton conductivity as well as other requirements. For example, PEM is expected to block electrons, oxygen, and hydrogen from penetrating and diffusing while the anode/cathode ionomer should allow hydrogen/oxygen to move easily, so that they can reach the catalyst nanoparticles. Given some of the well-known limits of Nafion, such as low glass-transition temperature, the community is in the midst of an active search for Nafion replacements. In this work, we present an informatics-based scheme to search large polymer chemical spaces, which includes establishing a list of properties needed for the targeted applications, developing predictive machine-learning models for these properties, defining a search space, and using the developed models to screen the search space. Using the scheme, we have identified 60 new polymer candidates for PEM, anode ionomer, and cathode ionomer that we hope will be advanced to the next step, i.e., validating the designs through synthesis and testing. The proposed informatics scheme is generic, and can be used to select polymers for multiple applications in the future.

physics.app-ph