Source author record

Eric Vanden-Eijnden

Eric Vanden-Eijnden appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

27works
23topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

27 published item(s)

preprint2026arXiv

FEAT: Free energy Estimators with Adaptive Transport

We present Free energy Estimators with Adaptive Transport (FEAT), a novel framework for free energy estimation -- a critical challenge across scientific domains. FEAT leverages learned transports implemented via stochastic interpolants and provides consistent, minimum-variance estimators based on escorted Jarzynski equality and controlled Crooks theorem, alongside variational upper and lower bounds on free energy differences. Unifying equilibrium and non-equilibrium methods under a single theoretical framework, FEAT establishes a principled foundation for neural free energy calculations. Experimental validation on toy examples, molecular simulations, and quantum field theory demonstrates improvements over existing learning-based methods. Our PyTorch implementation is available at https://github.com/jiajunhe98/FEAT.

preprint2022arXiv

A Dynamical Central Limit Theorem for Shallow Neural Networks

Recent theoretical works have characterized the dynamics of wide shallow neural networks trained via gradient descent in an asymptotic mean-field limit when the width tends towards infinity. At initialization, the random sampling of the parameters leads to deviations from the mean-field limit dictated by the classical Central Limit Theorem (CLT). However, since gradient descent induces correlations among the parameters, it is of interest to analyze how these fluctuations evolve. Here, we use a dynamical CLT to prove that the asymptotic fluctuations around the mean limit remain bounded in mean square throughout training. The upper bound is given by a Monte-Carlo resampling error, with a variance that that depends on the 2-norm of the underlying measure, which also controls the generalization error. This motivates the use of this 2-norm as a regularization term during training. Furthermore, if the mean-field dynamics converges to a measure that interpolates the training data, we prove that the asymptotic deviation eventually vanishes in the CLT scaling. We also complement these results with numerical experiments.

preprint2022arXiv

Adaptive Monte Carlo augmented with normalizing flows

Many problems in the physical sciences, machine learning, and statistical inference necessitate sampling from a high-dimensional, multi-modal probability distribution. Markov Chain Monte Carlo (MCMC) algorithms, the ubiquitous tool for this task, typically rely on random local updates to propagate configurations of a given system in a way that ensures that generated configurations will be distributed according to a target probability distribution asymptotically. In high-dimensional settings with multiple relevant metastable basins, local approaches require either immense computational effort or intricately designed importance sampling strategies to capture information about, for example, the relative populations of such basins. Here we analyze an adaptive MCMC which augments MCMC sampling with nonlocal transition kernels parameterized with generative models known as normalizing flows. We focus on a setting where there is no preexisting data, as is commonly the case for problems in which MCMC is used. Our method uses: (i) a MCMC strategy that blends local moves obtained from any standard transition kernel with those from a generative model to accelerate the sampling and (ii) the data generated this way to adapt the generative model and improve its efficacy in the MCMC algorithm. We provide a theoretical analysis of the convergence properties of this algorithm, and investigate numerically its efficiency, in particular in terms of its propensity to equilibrate fast between metastable modes whose rough location is known \textit{a~priori} but respective probability weight is not. We show that our algorithm can sample effectively across large free energy barriers, providing dramatic accelerations relative to traditional MCMC algorithms.

preprint2022arXiv

Averaged equation for energy diffusion on a graph reveals bifurcation diagram and thermally assisted reversal times in spin-torque driven nanomagnets

Driving nanomagnets by spin-polarized currents offers exciting prospects in magnetoelectronics, but the response of the magnets to such currents remains poorly understood. We show that an averaged equation describing the diffusion of energy on a graph captures the low-damping dynamics of these systems. From this equation we obtain the bifurcation diagram of the magnets, including the critical currents to induce stable precessional states and magnetization switching, as well as the mean times of thermally assisted magnetization reversal in situations where the standard reaction rate theory of Kramers is no longer valid. These results match experimental observations and give a theoretical basis for a Néel-Brown-type formula with an effective energy barrier for the reversal times.

preprint2022arXiv

Dual Training of Energy-Based Models with Overparametrized Shallow Neural Networks

Energy-based models (EBMs) are generative models that are usually trained via maximum likelihood estimation. This approach becomes challenging in generic situations where the trained energy is non-convex, due to the need to sample the Gibbs distribution associated with this energy. Using general Fenchel duality results, we derive variational principles dual to maximum likelihood EBMs with shallow overparametrized neural network energies, both in the feature-learning and lazy linearized regimes. In the feature-learning regime, this dual formulation justifies using a two time-scale gradient ascent-descent (GDA) training algorithm in which one updates concurrently the particles in the sample space and the neurons in the parameter space of the energy. We also consider a variant of this algorithm in which the particles are sometimes restarted at random samples drawn from the data set, and show that performing these restarts at every iteration step corresponds to score matching training. These results are illustrated in simple numerical experiments, which indicates that GDA performs best when features and particles are updated using similar time scales.

preprint2022arXiv

Minimum Action Method for Nonequilibrium Phase Transitions

First-order nonequilibrium phase transitions observed in active matter, fluid dynamics, biology, climate science, and other systems with irreversible dynamics are challenging to analyze because they cannot be inferred from a simple free energy minimization principle. Rather the mechanism of these transitions depends crucially on the system's dynamics, which requires us to analyze them in trajectory space rather than in phase space. Here we consider situations where the path of these transitions can be characterized as the minimizer of an action, whose minimum value can be used in a nonequilibrium generalization of the Arrhenius law to calculate the system's phase diagram. We also develop efficient numerical tools for the minimization of this action. These tools are general enough to be transportable to many situations of interest, in particular when the fluctuations present in the microscopic system are non-Gaussian and its dynamics is not governed by the standard Langevin equation. As an illustration, first-order phase transitions in two spatially-extended nonequilibrium systems are analyzed: a modified Ginzburg-Landau equation with a chemical potential which is non-gradient, and a reaction-diffusion network based on the Schlögl model. The phase diagrams of both systems are calculated as a function of their control parameters, and the paths of the transitions, including their critical nuclei, are identified. These results clearly demonstrate the nonequilibrium nature of the transitions, with differing forward and backward paths.

preprint2022arXiv

On Feature Learning in Neural Networks with Global Convergence Guarantees

We study the optimization of wide neural networks (NNs) via gradient flow (GF) in setups that allow feature learning while admitting non-asymptotic global convergence guarantees. First, for wide shallow NNs under the mean-field scaling and with a general class of activation functions, we prove that when the input dimension is no less than the size of the training set, the training loss converges to zero at a linear rate under GF. Building upon this analysis, we study a model of wide multi-layer NNs whose second-to-last layer is trained via GF, for which we also prove a linear-rate convergence of the training loss to zero, but regardless of the input dimension. We also show empirically that, unlike in the Neural Tangent Kernel (NTK) regime, our multi-layer model exhibits feature learning and can achieve better generalization performance than its NTK counterpart.

preprint2019arXiv

Infinite Switch Simulated Tempering in Force (FISST)

Many proteins in cells are capable of sensing and responding to piconewton scale forces, a regime in which conformational changes are small but significant for biological processes. In order to efficiently and effectively sample the response of these proteins to small forces, enhanced sampling techniques will be required. In this work, we derive, implement, and evaluate an efficient method to simultaneously sample the result of applying any constant pulling force within a specified range to a molecular system of interest. We start from Simulated Tempering in Force, whereby force is applied as a linear bias on a collective variable to the system's Hamiltonian, and the coefficient is taken as a continuous auxiliary degree of freedom. We derive a formula for an average collective-variable-dependent force, which depends on a set of weights, learned on-the-fly throughout a simulation, that reflect the limit where force varies infinitely quickly. These weights can then be used to retroactively compute averages of any observable at any force within the specified range. This technique is based on recent work deriving similar equations for Infinite Switch Simulated Tempering in Temperature, that showed the infinite switch limit is the most efficient for sampling. Here, we demonstrate that our method accurately and simultaneously samples molecular systems at all forces within a user defined force range, and show how it can serve as an enhanced sampling tool for cases where the pulling direction destabilizes states of low free-energy at zero-force. This method is implemented in, and will be freely-distributed with, the PLUMED open-source sampling library, and hence can be readily applied to problems using a wide range of molecular dynamics software packages.

preprint2016arXiv

Fluctuations in the heterogeneous multiscale methods for fast-slow systems

How heterogeneous multiscale methods (HMM) handle fluctuations acting on the slow variables in fast-slow systems is investigated. In particular, it is shown via analysis of central limit theorems (CLT) and large deviation principles (LDP) that the standard version of HMM artificially amplifies these fluctuations. A simple modification of HMM, termed parallel HMM, is introduced and is shown to remedy this problem, capturing fluctuations correctly both at the level of the CLT and the LDP. Similar type of arguments can also be used to justify that the tau-leaping method used in the context of Gillespie's stochastic simulation algorithm for Markov jump processes also captures the right CLT and LDP for these processes.

preprint2016arXiv

Optimized Markov State Models for Metastable Systems

A method is proposed to identify target states that optimize a metastability index amongst a set of trial states and use these target states as milestones (or core sets) to build Markov State Models (MSMs). If the optimized metastability index is small, this automatically guarantees the accuracy of the MSM, in the sense that the transitions between the target milestones is indeed approximately Markovian. The method is simple to implement and use, it does not require that the dynamics on the trial milestones be Markovian, and it also offers the possibility to partition the system's state-space by assigning every trial milestone to the target milestones it is most likely to visit next and to identify transition state regions. Here the method is tested on the Gly-Ala-Gly peptide, where it shown to correctly identify the expected metastable states in the dihedral angle space of the molecule without \textit{a~priori} information about these states. It is also applied to analyze the folding landscape of the Beta3s mini-protein, where it is shown to identify the folded basin as a connecting hub between an helix-rich region, which is entropically stabilized, and a beta-rich region, which is energetically stabilized and acts as a kinetic trap.

preprint2015arXiv

Continuous-time Random Walks for the Numerical Solution of Stochastic Differential Equations

This paper introduces time-continuous numerical schemes to simulate stochastic differential equations (SDEs) arising in mathematical finance, population dynamics, chemical kinetics, epidemiology, biophysics, and polymeric fluids. These schemes are obtained by spatially discretizing the Kolmogorov equation associated with the SDE in such a way that the resulting semi-discrete equation generates a Markov jump process that can be realized exactly using a Monte Carlo method. In this construction the spatial increment of the approximation can be bounded uniformly in space, which guarantees that the schemes are numerically stable for both finite and long time simulation of SDEs. By directly analyzing the generator of the approximation, we prove that the approximation has a sharp stochastic Lyapunov function when applied to an SDE with a drift field that is locally Lipschitz continuous and weakly dissipative. We use this stochastic Lyapunov function to extend a local semimartingale representation of the approximation. This extension permits to analyze the complexity of the approximation. Using the theory of semigroups of linear operators on Banach spaces, we show that the approximation is (weakly) accurate in representing finite and infinite-time statistics, with an order of accuracy identical to that of its generator. The proofs are carried out in the context of both fixed and variable spatial step sizes. Theoretical and numerical studies confirm these statements, and provide evidence that these schemes have several advantages over standard methods based on time-discretization. In particular, they are accurate, eliminate nonphysical moves in simulating SDEs with boundaries (or confined domains), prevent exploding trajectories from occurring when simulating stiff SDEs, and solve first exit problems without time-interpolation errors.

preprint2015arXiv

Large Deviations in Fast-Slow Systems

The incidence of rare events in fast-slow systems is investigated via analysis of the large deviation principle (LDP) that characterizes the likelihood and pathway of large fluctuations of the slow variables away from their mean behavior -- such fluctuations are rare on short timescales but become ubiquitous eventually. This LDP involves an Hamilton-Jacobi equation whose Hamiltonian is related to the leading eigenvalue of the generator of the fast process, and is typically non-quadratic in the momenta -- in other words, the LDP for the slow variables in fast-slow systems is different in general from that of any stochastic differential equation (SDE) one would write for the slow variables alone. It is shown here that the eigenvalue problem for the Hamiltonian can be reduced to a simpler algebraic equation for this Hamiltonian for a specific class of systems in which the fast variables satisfy a linear equation whose coefficients depend nonlinearly on the slow variables, and the fast variables enter quadratically the equation for the slow variables. These results are illustrated via examples, inspired by kinetic theories of turbulent flows and plasma, in which the quasipotential characterizing the long time behavior of the system is calculated and shown again to be different from that of an SDE.

preprint2014arXiv

Exact dynamical coarse-graining without time-scale separation

A family of collective variables is proposed to perform exact dynamical coarse-graining even in systems without time scale separation. More precisely, it is shown that these variables are not slow in general but they satisfy an overdamped Langevin equation that statistically preserves the sequence in which any regions in collective variable space are visited and permits to calculate exactly the mean first passage times from any such region to another. The role of the free energy and diffusion coefficient in this overdamped Langevin equation is discussed, along with the way they transform under any change of variable in collective variable space. These results apply both for systems with and without inertia, and they can be generalized to using several collective variables simultaneously. The view they offer on what makes collective variables and reaction coordinates optimal breaks from the standard notion that good collective variable must be slow variable, and it suggests new ways to interpret data from molecular dynamic simulations and experiments.

preprint2014arXiv

Flows in Complex Networks: Theory, Algorithms, and Application to Lennard-Jones Cluster Rearrangement

A set of analytical and computational tools based on transition path theory (TPT) is proposed to analyze flows in complex networks. Specifically, TPT is used to study the statistical properties of the reactive trajectories by which transitions occur between specific groups of nodes on the network. Sampling tools are built upon the outputs of TPT that allow to generate these reactive trajectories directly, or even transition paths that travel from one group of nodes to the other without making any detour and carry the same probability current as the reactive trajectories. These objects permit to characterize the mechanism of the transitions, for example by quantifying the width of the tubes by which these transitions occur, the location and distribution of their dynamical bottlenecks, etc. These tools are applied to a network modeling the dynamics of the Lennard-Jones cluster with 38 atoms (LJ38) and used to understand the mechanism by which this cluster rearranges itself between its two most likely states at various temperatures.

preprint2014arXiv

Metropolis Integration Schemes for Self-Adjoint Diffusions

We present explicit methods for simulating diffusions whose generator is self-adjoint with respect to a known (but possibly not normalizable) density. These methods exploit this property and combine an optimized Runge-Kutta algorithm with a Metropolis-Hastings Monte-Carlo scheme. The resulting numerical integration scheme is shown to be weakly accurate at finite noise and to gain higher order accuracy in the small noise limit. It also permits to avoid computing explicitly certain terms in the equation, such as the divergence of the mobility tensor, which can be tedious to calculate. Finally, the scheme is shown to be ergodic with respect to the exact equilibrium probability distribution of the diffusion when it exists. These results are illustrated on several examples including a Brownian dynamics simulation of DNA in a solvent. In this example, the proposed scheme is able to accurately compute dynamics at time step sizes that are an order of magnitude (or more) larger than those permitted with commonly used explicit predictor-corrector schemes.

preprint2014arXiv

Relevance of instantons in Burgers turbulence

Instanton calculations are performed in the context of stationary Burgers turbulence to estimate the tails of the probability density function (PDF) of velocity gradients. These results are then compared to those obtained from massive direct numerical simulations (DNS) of the randomly forced Burgers equation. The instanton predictions are shown to agree with the DNS in a wide range of regimes, including those that are far from the limiting cases previously considered in the literature. These results settle the controversy of the relevance of the instanton approach for the prediction of the velocity gradient PDF tail exponents. They also demonstrate the usefulness of the instanton formalism in Burgers turbulence, and suggest that this approach may be applicable in other contexts, such as 2D and 3D turbulence in compressible and incompressible flows.

preprint2014arXiv

Stochastic Mode-Reduction in Models with Conservative Fast Sub-Systems

A stochastic mode reduction strategy is applied to multiscale models with a deterministic energy-conserving fast sub-system. Specifically, we consider situations where the slow variables are driven stochastically and interact with the fast sub-system in an energy-conserving fashion. Since the stochastic terms only affect the slow variables, the fast-subsystem evolves deterministically on a sphere of constant energy. However, in the full model the radius of the sphere slowly changes due to the coupling between the slow and fast dynamics. Therefore, the energy of the fast sub-system becomes an additional hidden slow variable that must be accounted for in order to apply the stochastic mode reduction technique to systems of this type.

preprint2012arXiv

A granocentric model captures the statistical properties of monodisperse random packings

We present a generalization of the granocentric model proposed in [Clusel et al., Nature, 2009, 460, 611615] that is capable of describing the local fluctuations inside not only polydisperse but also monodisperse packings of spheres. This minimal model does not take into account the relative particle positions, yet it captures positional disorder through local stochastic processes sampled by efficient Monte Carlo methods. The disorder is characterized by the distributions of local parameters, such as the number of neighbors and contacts, filled solid angle around a central particle and the cell volumes. The model predictions are in good agreement with our experimental data on monodisperse random close packings of PMMA particles. Moreover, the model can be used to predict the distributions of local fluctuations in any packing, as long as the average number of neighbors, contacts and the packing fraction are known. These distributions give a microscopic foundation to the statistical mechanics framework for jammed matter and allow us to calculate thermodynamic quantities such as the compactivity in the phase space of possible jammed configurations.

preprint2012arXiv

An observable for vacancy characterization and diffusion in crystals

To locate the position and characterize the dynamics of a vacancy in a crystal, we propose to represent it by the ground state density of a quantum probe quasi-particle for the Hamiltonian associated to the potential energy field generated by the atoms in the sample. In this description, the h^2/2mu coefficient of the kinetic energy term is a tunable parameter controlling the density localization in the regions of relevant minima of the potential energy field. Based on this description, we derive a set of collective variables that we use in rare event simulations to identify some of the vacancy diffusion paths in a 2D crystal. Our simulations reveal, in addition to the simple and expected nearest neighbor hopping path, a collective migration mechanism of the vacancy. This mechanism involves several lattice sites and produces a long range migration of the vacancy. Finally, we also observed a vacancy induced crystal reorientation process.

preprint2012arXiv

Force-clamp analysis techniques reveal stretched exponential unfolding kinetics in ubiquitin

Force-clamp spectroscopy reveals the unfolding and disulfide bond rupture times of single protein molecules as a function of the stretching force, point mutations and solvent conditions. The statistics of these times reveal whether the protein domains are independent of one another, the mechanical hierarchy in the polyprotein chain, and the functional form of the probability distribution from which they originate. It is therefore important to use robust statistical tests to decipher the correct theoretical model underlying the process. Here we develop multiple techniques to compare the well-established experimental data set on ubiquitin with existing theoretical models as a case study. We show that robustness against filtering, agreement with a maximum likelihood function that takes into account experimental artifacts, the Kuiper statistic test and alignment with synthetic data all identify the Weibull or stretched exponential distribution as the best fitting model. Our results are inconsistent with recently proposed models of Gaussian disorder in the energy landscape or noise in the applied force as explanations for the observed non-exponential kinetics. Since the physical model in the fit affects the characteristic unfolding time, these results have important implications on our understanding of the biological function of proteins.

preprint2012arXiv

Thermal Stability of the Magnetization in Perpendicularly Magnetized Thin Film Nanomagnets

Understanding the stability of thin film nanomagnets with perpendicular magnetic anisotropy (PMA) against thermally induced magnetization reversal is important when designing perpendicularly magnetized patterned media and magnetic random access memories. The leading-order dependence of magnetization reversal rates are governed by the energy barrier the system needs to surmount in order for reversal to proceed. In this paper we study the reversal dynamics of these systems and compute the relevant barriers using the string method of E, Vanden-Eijnden, and Ren. We find the reversal to be often spatially incoherent; that is, rather than the magnetization flipping as a rigid unit, reversal proceeds instead through a soliton-like domain wall sweeping through the system. We show that for square nanomagnetic elements the energy barrier increases with element size up to a critical length scale, beyond which the energy barrier is constant. For circular elements the energy barrier continues to increase indefinitely, albeit more slowly beyond a critical size. In both cases the energy barriers are smaller than those expected for coherent magnetization reversal.

preprint2010arXiv

A patch that imparts unconditional stability to certain explicit integrators for SDEs

This paper proposes a simple strategy to simulate stochastic differential equations (SDE) arising in constant temperature molecular dynamics. The main idea is to patch an explicit integrator with Metropolis accept or reject steps. The resulting `Metropolized integrator' preserves the SDE's equilibrium distribution and is pathwise accurate on finite time intervals. As a corollary the integrator can be used to estimate finite-time dynamical properties along an infinitely long solution. The paper explains how to implement the patch (even in the presence of multiple-time-stepsizes and holonomic constraints), how it scales with system size, and how much overhead it requires. We test the integrator on a Lennard-Jones cluster of particles and `dumbbells' at constant temperature.

preprint2010arXiv

Non-asymptotic mixing of the MALA algorithm

The Metropolis-Adjusted Langevin Algorithm (MALA), originally introduced to sample exactly the invariant measure of certain stochastic differential equations (SDE) on infinitely long time intervals, can also be used to approximate pathwise the solution of these SDEs on finite time intervals. However, when applied to an SDE with a nonglobally Lipschitz drift coefficient, the algorithm may not have a spectral gap even when the SDE does. This paper reconciles MALA's lack of a spectral gap with its ergodicity to the invariant measure of the SDE and finite time accuracy. In particular, the paper shows that its convergence to equilibrium happens at exponential rate up to terms exponentially small in time-stepsize. This quantification relies on MALA's ability to exactly preserve the SDE's invariant measure and accurately represent the SDE's transition probability on finite time intervals.

preprint2010arXiv

Stability of 2pi domain walls in ferromagnetic nanorings

The stability of 2pi domain walls in ferromagnetic nanorings is investigated via calculation of the minimum energy path that separates a 2pi domain wall from the vortex state of a ferromagnetic nanoring. Trapped domains are stable when they exist between certain types of transverse domain walls, i.e., walls in which the edge defects on the same side of the magnetic strip have equal sign and thus repel. Here the energy barriers between these configurations and vortex magnetization states are obtained using the string method. Due to the geometry of a ring, two types of 2pi walls must be distinguished that differ by their overall topological index and exchange energy. The minimum energy path corresponds to the expulsion of a vortex. The energy barrier for annihilation of a 2pi wall is compared to the activation energy for transitions between the two ring vortex states.

preprint2009arXiv

Pathwise Accuracy and Ergodicity of Metropolized Integrators for SDEs

Metropolized integrators for ergodic stochastic differential equations (SDE) are proposed which (i) are ergodic with respect to the (known) equilibrium distribution of the SDE and (ii) approximate pathwise the solutions of the SDE on finite time intervals. Both these properties are demonstrated in the paper and precise strong error estimates are obtained. It is also shown that the Metropolized integrator retains these properties even in situations where the drift in the SDE is nonglobally Lipschitz, and vanilla explicit integrators for SDEs typically become unstable and fail to be ergodic.

preprint2007arXiv

Solvent coarse-graining and the string method applied to the hydrophobic collapse of a hydrated chain

Using computer simulations of over 100,000 atoms, the mechanism for the hydrophobic collapse of an idealized hydrated chain is obtained. This is done by coarse-graining the atomistic water molecule positions over 129,000 collective variables that represent the water density field and then using the string method in these variables to compute the minimum free energy pathway (MFEP) for the collapsing chain. The dynamical relevance of the MFEP (i.e. its coincidence with the mechanism of collapse) is validated a posteriori using conventional molecular dynamics trajectories. Analysis of the MFEP provides atomistic confirmation for the mechanism of hydrophobic collapse proposed by ten Wolde and Chandler. In particular, it is shown that lengthscale-dependent hydrophobic dewetting is the rate-limiting step in the hydrophobic collapse of the considered chain.

preprint2005arXiv

Non-meanfield deterministic limits in chemical reaction kinetics far from equilibrium

A general mechanism is proposed by which small intrinsic fluctuations in a system far from equilibrium can result in nearly deterministic dynamical behaviors which are markedly distinct from those realized in the meanfield limit. The mechanism is demonstrated for the kinetic Monte-Carlo version of the Schnakenberg reaction where we identified a scaling limit in which the global deterministic bifurcation picture is fundamentally altered by fluctuations. Numerical simulations of the model are found to be in quantitative agreement with theoretical predictions.