Two machine-learned representations are compared, DeepTICA and SPIB-VAE, with linear TICA for recovering steady-state observables from flawed distributions and a kinetic score computed from a coarse MSM at a resolution comparable to that used for RiteWeight random clustering is provided.
Abstract
The increasing use of generative models has made ensembles of short molecular dynamics trajectories increasingly common, creating a growing need for methods that can recover physically meaningful steady-state populations and kinetics from improperly weighted conformational ensembles. Randomized Iterative Trajectory Reweighting (RiteWeight) addresses this problem through repeated random clustering and iterative reweighting, without requiring the fixed Markovian discretization used in conventional Markov state models (MSM). However, the choice of reduced feature space in which RiteWeight performs random clustering has not been systematically investigated. Here, we compare two machine-learned representations, DeepTICA and SPIB-VAE, with linear TICA for recovering steady-state observables from flawed distributions. DeepTICA learns nonlinear coordinates by targeting slow transfer-operator eigenmodes, whereas SPIB-VAE compresses configurations into a low-dimensional latent space while retaining information predictive of future metastable states. DeepTICA provided a comparatively robust RiteWeight representation under limited hyperparameter exploration, whereas SPIB-VAE benefited more strongly from broader optimization. Moreover, a kinetic score computed from a coarse MSM at a resolution comparable to that used for RiteWeight random clustering provided a useful criterion for efficiently selecting reduced representations and their hyperparameters for RiteWeight.
BICePs-reweighted Reversible DeepMSMs are introduced, combining Bayesian Inference of Conformational Populations (BICePs), variational learning of Markov processes, and maximum entropy (MaxEnt)/maximum caliber (MaxCal) principles to infer consistent thermodynamics and minimally perturbed kinetics.
Many scientific and machine learning systems, from molecular dynamics to diffusion models and beyond, are governed by stochastic dynamics with low-dimensional structure, evolving on slow timescales. However, target trajectories, used to identify and interpret such dynamics, are often inaccessible: only biased or static...
Vladimir R. Kostic, Karim Lounici, Hélène Halconruy et al.· 0 citations
Deep learning surrogates have become powerful tools for simulating and forecasting complex dynamical systems, yet their utility remains limited by catastrophic error accumulation during long-term autoregressive rollouts. This behavior is partly tied to the nature of the underlying systems: chaotic spatiotemporal system...
Molecular systems have many degrees of freedom, but their metastable behavior can often be described by a few collective variables. Identifying these variables and estimating free energies along them from limited simulation data remains a challenging, important problem. Separate short trajectories may sample different...
The arc of CV discovery is traced from intuition-driven heuristics to modern data-driven and generative frameworks, critically assess the strengths and limitations of each class of methods, and outline how the convergence of machine-learned potentials, automated CV learning, generative sampling, and causal interpretabi...
R. Talmazan, Cheng Giuseppe Chen, Chen-Yu Tang et al.· Digital Discovery· 0 citations
A transition kernel defined over the RBM sequence used in DT is proposed, enabling nonlocal moves within a single transition while leaving the RBM sequence invariant and mitigates the training failures observed with BGS- and DT-based learning.
Kaiji Sekimoto, Muneki Yasuda· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.