SearcharxivSearch

arXiv subjects

Eric Beyerle

Publications and source records attributed to Eric Beyerle.

5 recordsLinked to original sources

An Information Bottleneck Approach for Markov Model Construction

Markov state models (MSMs) are valuable for studying dynamics of protein conformational changes via statistical analysis of molecular dynamics (MD) simulations. In MSMs, the complex configuration space is coarse-grained into conformational states, with the dynamics modeled by a series of Markovian transitions among these states at discrete lag times. Constructing the Markovian model at a specific lag time requires state defined without significant internal energy barriers, enabling internal dynamics relaxation within the lag time. This process coarse grains time and space, integrating out rapid motions within metastable states. This work introduces a continuous embedding approach for molecular conformations using the state predictive information bottleneck (SPIB), which unifies dimensionality reduction and state space partitioning via a continuous, machine learned basis set. Without explicit optimization of VAMP-based scores, SPIB demonstrates state-of-the-art performance in identifying slow dynamical processes and constructing predictive multi-resolution Markovian models. When applied to mini-proteins trajectories, SPIB showcases unique advantages compared to competing methods. It automatically adjusts the number of metastable states based on a specified minimal time resolution, eliminating the need for manual tuning. While maintaining efficacy in dynamical properties, SPIB excels in accurately distinguishing metastable states and capturing numerous well-populated macrostates. Furthermore, SPIB's ability to learn a low-dimensional continuous embedding of the underlying MSMs enhances the interpretation of dynamic pathways. Accordingly, we propose SPIB as an easy-to-implement methodology for end-to-end MSM construction.

physics.bio-ph

Thermodynamically Optimized Machine-learned Reaction Coordinates for Hydrophobic Ligand Dissociation

Ligand unbinding is mediated by the free energy change, which has intertwined contributions from both energy and entropy. It is important but not easy to quantify their individual contributions. We model hydrophobic ligand unbinding for two systems, a methane particle and a C60 fullerene, both unbinding from hydrophobic pockets in all-atom water. By using a modified deep learning framework, we learn a thermodynamically optimized reaction coordinate to describe hydrophobic ligand dissociation for both systems. Interpretation of these reaction coordinates reveals the roles of entropic and enthalpic forces as ligand and pocket sizes change. Irrespective of the contrasting roles of energy and entropy, we also find that for both the systems the transition from the bound to unbound states is driven primarily by solvation of the pocket and ligand, independent of ligand size. Our framework thus gives useful thermodynamic insight into hydrophobic ligand dissociation problems that are otherwise difficult to glean.

physics.chem-ph

JARVIS-Leaderboard: A Large Scale Benchmark of Materials Design Methods

Lack of rigorous reproducibility and validation are major hurdles for scientific development across many fields. Materials science in particular encompasses a variety of experimental and theoretical approaches that require careful benchmarking. Leaderboard efforts have been developed previously to mitigate these issues. However, a comprehensive comparison and benchmarking on an integrated platform with multiple data modalities with both perfect and defect materials data is still lacking. This work introduces JARVIS-Leaderboard, an open-source and community-driven platform that facilitates benchmarking and enhances reproducibility. The platform allows users to set up benchmarks with custom tasks and enables contributions in the form of dataset, code, and meta-data submissions. We cover the following materials design categories: Artificial Intelligence (AI), Electronic Structure (ES), Force-fields (FF), Quantum Computation (QC) and Experiments (EXP). For AI, we cover several types of input data, including atomic structures, atomistic images, spectra, and text. For ES, we consider multiple ES approaches, software packages, pseudopotentials, materials, and properties, comparing results to experiment. For FF, we compare multiple approaches for material property predictions. For QC, we benchmark Hamiltonian simulations using various quantum algorithms and circuits. Finally, for experiments, we use the inter-laboratory approach to establish benchmarks. There are 1281 contributions to 274 benchmarks using 152 methods with more than 8 million data-points, and the leaderboard is continuously expanding. The JARVIS-Leaderboard is available at the website: https://pages.nist.gov/jarvis_leaderboard

cond-mat.mtrl-sci

Driving and characterizing nucleation of urea and glycine polymorphs in water

Crystal nucleation is relevant across the domains of fundamental and applied sciences. However, in many cases its mechanism remains unclear due to a lack of temporal or spatial resolution. To gain insights to the molecular details of nucleation, some form of molecular dynamics simulations is typically performed; these simulations, in turn, are limited by their ability to run long enough to sample the nucleation event thoroughly. To overcome the timescale limits in typical molecular dynamics simulations in a manner free of prior human bias, here we employ the machine learning augmented molecular dynamics framework ``Reweighted Autoencoded Variational Bayes for enhanced sampling (RAVE)". We study two molecular systems, urea and glycine in explicit all-atom water, due to their enrichment in polymorphic structures and common utility in commercial applications. From our simulations, we observe multiple back-and-forth liquid-solid transitions of different polymorphs and from these trajectories calculate the polymorph stability relative to the dissolved liquid state. We further observe that the obtained reaction coordinates and transitions are highly non-classical.

cond-mat.soft

Identifying the leading dynamics of ubiquitin: a comparison between the tICA and the LE4PD slow fluctuations in amino acids' position

Molecular Dynamics (MD) simulations of proteins implicitly contain the information connecting the atomistic molecular structure and proteins' biologically relevant motion, where large-scale fluctuations are deemed to guide folding and function. In the complex multiscale processes described by MD trajectories it is difficult to identify, separate, and study those large-scale fluctuations. This problem can be formulated as the need to identify a small number of collective variables that guide the slow kinetic processes. Among the methods used to study the slow, leading processes in proteins' dynamics, the time-lagged independent component analysis, or tICA, has been extensively used. Recently, we developed a Langevin coarse-grained approach for the dynamics of proteins, called the Langevin Equation for Protein Dynamics or LE4PD. This approach partitions the protein's MD dynamics into uncorrelated, wavelength-dependent, diffusive modes, and it associates to each mode a free-energy map. In the free energy maps, we measure the spatial extension and the time evolution of the mode-dependent, slow dynamical fluctuations, using the string method and a Markov state model. The theory identifies the slow collective variables in the rescaled LE4PD normal modes. Here, we compare the tICA modes' predictions with the collective LE4PD modes. We observe that the two methods consistently identify the nature and extension of the slowest fluctuation processes. The tICA separates the slow, leading processes in a smaller number of modes than the LE4PD does. However, LE4PD provides time-dependent information and a formal connection to the physics of the kinetic processes missing in the pure statistical analysis of tICA.

physics.bio-ph