Source author record

Xiantao Li

Xiantao Li appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

21works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

21 published item(s)

preprint2022arXiv

A Local Convergence Theory for the Stochastic Gradient Descent Method in Non-Convex Optimization With Non-isolated Local Minima

Loss functions with non-isolated minima have emerged in several machine learning problems, creating a gap between theory and practice. In this paper, we formulate a new type of local convexity condition that is suitable to describe the behavior of loss functions near non-isolated minima. We show that such condition is general enough to encompass many existing conditions. In addition we study the local convergence of the SGD under this mild condition by adopting the notion of stochastic stability. The corresponding concentration inequalities from the convergence analysis help to interpret the empirical observation from some practical training results.

preprint2022arXiv

On quantum algorithms for the Schrödinger equation in the semi-classical regime

Solving the time-dependent Schrödinger equation is an important application area for quantum algorithms. We consider Schrödinger's equation in the semi-classical regime. Here the solutions exhibit strong multiple-scale behavior due to a small parameter $\hbar$, in the sense that the dynamics of the quantum states and the induced observables can occur on different spatial and temporal scales. Such a Schrödinger equation finds many applications, including in Born-Oppenheimer molecular dynamics and Ehrenfest dynamics. This paper considers quantum analogues of pseudo-spectral (PS) methods on classical computers. Estimates on the gate counts in terms of $\hbar$ and the precision $\varepsilon$ are obtained. It is found that the number of required qubits, $m$, scales only logarithmically with respect to $\hbar$. When the solution has bounded derivatives up to order $\ell$, the symmetric Trotting method has gate complexity $\mathcal{O}\Big({ (\varepsilon \hbar)^{-\frac12} \mathrm{polylog}(\varepsilon^{-\frac{3}{2\ell}} \hbar^{-1-\frac{1}{2\ell}})}\Big),$ provided that the diagonal unitary operators in the pseudo-spectral methods can be implemented with $\mathrm{poly}(m)$ operations. When physical observables are the desired outcomes, however, the step size in the time integration can be chosen independently of $\hbar$. The gate complexity in this case is reduced to $\mathcal{O}\Big({\varepsilon^{-\frac12} \mathrm{polylog}( \varepsilon^{-\frac3{2\ell}} \hbar^{-1} )}\Big),$ with $\ell$ again indicating the smoothness of the solution.

preprint2022arXiv

Some Error Analysis for the Quantum Phase Estimation Algorithms

This paper is concerned with the phase estimation algorithm in quantum computing algorithms, especially the scenarios where (1) the input vector is not an eigenvector; (2) the unitary operator is not exactly implemented; (3) random approximations are used for the unitary operator, e.g., the QDRIFT method. We characterize the probability of computing the phase values in terms of the consistency error, including the residual error, Trotter splitting error, or statistical mean-square error. In the first two cases, we show that in order to obtain the phase value with {error less or equal to $2^{-n}$ } and probability at least $1-ε$, the required number of qubits is $ t \geq n + \log \big(2 + \frac{δ^2 }{2 εΔ\!E^2 } \big).$ The parameter $δ$ quantifies the error associated with the inexact eigenvector and/or the unitary operator, and $Δ\! E$ characterizes the spectral gap, i.e., the separation from the rest of the phase values. For the third case, we found a similar estimate, but the number of random steps has to be sufficiently large.

preprint2020arXiv

Data-driven molecular modeling with the generalized Langevin equation

The complexity of molecular dynamics simulations necessitates dimension reduction and coarse-graining techniques to enable tractable computation. The generalized Langevin equation (GLE) describes coarse-grained dynamics in reduced dimensions. In spite of playing a crucial role in non-equilibrium dynamics, the memory kernel of the GLE is often ignored because it is difficult to characterize and expensive to solve. To address these issues, we construct a data-driven rational approximation to the GLE. Building upon previous work leveraging the GLE to simulate simple systems, we extend these results to more complex molecules, whose many degrees of freedom and complicated dynamics require approximation methods. We demonstrate the effectiveness of our approximation by testing it against exact methods and comparing observables such as autocorrelation and transition rates.

preprint2020arXiv

Markovian Embedding Procedures for Non-Markovian Stochastic Schrödinger Equations

We present embedding procedures for the non-Markovian stochastic Schrödinger equations, arising from studies of quantum systems coupled with bath environments. By introducing auxiliary wave functions, it is demonstrated that the non-Markovian dynamics can be embedded in extended, but Markovian, stochastic models. Two embedding procedures are presented. The first method leads to nonlinear stochastic equations, the implementation of which is much more efficient than the non-Markovian stochastic Schrödinger equations. The stochastic Schrödinger equations obtained from the second procedure involve more auxiliary wave functions, but the equations are linear, and we derive the corresponding generalized quantum master equation for the density-matrix. The accuracy of the embedded models is ensured by fitting to the power spectrum. The stochastic force is represented using a linear superposition of Ornstein-Uhlenbeck processes, which are incorporated as multiplicative noise in the auxiliary Schrödinger equations. The asymptotic behavior of the spectral density in the low frequency regime is preserved by using correlated stochastic processes. The approximations are verified by using a spin-boson system as a test example.

preprint2020arXiv

Random Batch Algorithms for Quantum Monte Carlo simulations

Random batch algorithms are constructed for quantum Monte Carlo simulations. The main objective is to alleviate the computational cost associated with the calculations of two-body interactions, including the pairwise interactions in the potential energy, and the two-body terms in the Jastrow factor. In the framework of variational Monte Carlo methods, the random batch algorithm is constructed based on the over-damped Langevin dynamics, so that updating the position of each particle in an $N$-particle system only requires $\mathcal{O}(1)$ operations, thus for each time step the computational cost for $N$ particles is reduced from $\mathcal{O}(N^2)$ to $\mathcal{O}(N)$. For diffusion Monte Carlo methods, the random batch algorithm uses an energy decomposition to avoid the computation of the total energy in the branching step. The effectiveness of the random batch method is demonstrated using a system of liquid ${}^4$He atoms interacting with a graphite surface.

preprint2020arXiv

The strong convergence of operator-splitting methods for the Langevin dynamics model

We study the strong convergence of some operator-splitting methods for the Langevin dynamics model with additive noise. It will be shown that a direct splitting of deterministic and random terms, including the symmetric splitting methods, only offers strong convergence of order 1. To improve the order of strong convergence, a new class of operator-splitting methods based on Kunita's solution representation are proposed. We present stochastic algorithms with strong orders up to 3. Both mathematical analysis and numerical evidence are provided to verify the desired order of accuracy.

preprint2019arXiv

Exponential Integrators for Stochastic Schrödinger Equation

We present a class of exponential integrators to compute solutions of the stochastic Schrödinger equation arising from the modeling of open quantum systems. In order to be able to implement the methods within the same framework as the deterministic counterpart, we express the solution using the Kunita's representation. With appropriate truncations, the solution operator can be written as matrix exponentials, which can be efficiently implemented by the Krylov subspace projection. The accuracy is examined in terms of the strong convergence, by comparing trajectories, and the weak convergence, by comparing the density-matrix operator. We show that the local accuracy can be further improved by introducing a third-order commutator in the exponential. The effectiveness of the proposed methods is tested using the example from Di Ventra et al. [Journal of Physics: Condensed Matter, 2004].

preprint2016arXiv

Data-driven parameterization of the generalized Langevin equation

We present a data-driven approach to determine the memory kernel and random noise in generalized Langevin equations. To facilitate practical implementations, we parameterize the kernel function in the Laplace domain by a rational function, with coefficients directly linked to the equilibrium statistics of the coarse-grain variables. We show that such an approximation can be constructed to arbitrarily high order and the resulting generalized Langevin dynamics can be embedded in an extended stochastic model without explicit memory. We demonstrate how to introduce the stochastic noise so that the second fluctuation-dissipation theorem is exactly satisfied. Results from several numerical tests are presented to demonstrate the effectiveness of the proposed method.

preprint2016arXiv

PEXSI-$Σ$: A Green's function embedding method for Kohn-Sham density functional theory

In this paper, we propose a new Green's function embedding method called PEXSI-$Σ$ for describing complex systems within the Kohn-Sham density functional theory (KSDFT) framework, after revisiting the physics literature of Green's function embedding methods from a numerical linear algebra perspective. The PEXSI-$Σ$ method approximates the density matrix using a set of nearly optimally chosen Green's functions evaluated at complex frequencies. For each Green's function, the complex boundary conditions are described by a self energy matrix $Σ$ constructed from a physical reference Green's function, which can be computed relatively easily. In the linear regime, such treatment of the boundary condition can be numerically exact. The support of the $Σ$ matrix is restricted to degrees of freedom near the boundary of computational domain, and can be interpreted as a frequency dependent surface potential. This makes it possible to perform KSDFT calculations with $\mathcal{O}(N^2)$ computational complexity, where $N$ is the number of atoms within the computational domain. Green's function embedding methods are also naturally compatible with atomistic Green's function methods for relaxing the atomic configuration outside the computational domain. As a proof of concept, we demonstrate the accuracy of the PEXSI-$Σ$ method for graphene with divacancy and dislocation dipole type of defects using the DFTB+ software package.

preprint2015arXiv

An atomistic/continuum coupling method using enriched bases

A common observation from an atomistic to continuum coupling method is that the error is often generated and concentrated near the interface, where the two models are combined. In this paper, a new method is proposed to suppress the error at the interface, and as a consequence, the overall accuracy is improved. The method is motivated by formulating the molecular mechanics model as a two-stage minimization problem. In particular, it is demonstrated that the error at the interface can be considerably reduced when new basis functions are introduced in a Galerkin projection formalism. The improvement of the accuracy is illustrated by two examples. Further, the comparison to some quasicontinuum-type methods is provided.

preprint2015arXiv

Parametric Reduced Models for the Nonlinear Schrödinger Equation

Reduced models for the (defocusing) nonlinear Schrödinger equation are developed. In particular, we develop reduced models that only involve the low-frequency modes given noisy observations of these modes. The ansatz of the reduced parametric models are obtained by employing a rational approximation and a colored noise approximation, respectively, on the memory terms and the random noise of a generalized Langevin equation that is derived from the standard Mori-Zwanzig formalism. The parameters in the resulting reduced models are inferred from noisy observations with a recently developed ensemble Kalman filter-based parameterization method. The forecasting skill across different temperature regimes are verified by comparing the moments up to order four, a two-time correlation function statistics, and marginal densities of the coarse-grained variables.

preprint2015arXiv

Some New Symplectic Multiple Timestepping Methods for Multiscale Molecular Dynamics Models

We derived a number of numerical methods to treat biomolecular systems with multiple time scales. Based on the splitting of the operators associated with the slow-varying and fast-varying forces, new multiple time-stepping (MTS) methods are obtained by eliminating the dominant terms in the error. These new methods can be viewed as a generalization of the impulse method. In the implementation of these methods, the long-range forces only need to be computed on the slow time scale, which reduces the computational cost considerably. Preliminary analysis for the energy conservation property is provided.

preprint2015arXiv

Traction Boundary Conditions for Molecular Static Simulations

This paper presents a consistent approach to prescribe traction boundary conditions in atomistic models. Due to the typical multiple-neighbor interactions, finding an appropriate boundary condition that models a desired traction is a non-trivial task. We first present a one-dimensional example, which demonstrates how such boundary conditions can be formulated. We further analyze the stability, and derive its continuum limit. We also show how the boundary conditions can be extended to higher dimensions with an application to a dislocation dipole problem under shear stress.

preprint2014arXiv

Computation of the Memory Functions in the Generalized Langevin Models for Collective Dynamics of Macromolecules

We present a numerical method to compute the approximation of the memory functions in the generalized Langevin models for collective dynamics of macromolecules. We first derive the exact expressions of the memory functions, obtained from projection to subspaces that correspond to the selection of coarse-grain variables. In particular, the memory functions are expressed in the forms of matrix functions, which will then be approximated by Krylov-subspace methods. It will also be demonstrated that the random noise can be approximated under the same framework, and the fluctuation-dissipation theorem is automatically satisfied. The accuracy of the method is examined through several numerical examples.

preprint2014arXiv

On Consistent Definitions of Momentum and Energy Fluxes for Molecular Dynamics Models with Multi-body Interatomic Potentials

In this paper, we propose a two-level criteria to check the consistency of the definitions of continuum quantities in Molecular Dynamics. As examples, we follow the control- volume approach, derive the definitions of the tractions and energy fluxes for EAM potential and Tersoff potential, and provide the pseudo code the computing. Then, we verify the consistency of the definitions by analytical and numerical methods.

preprint2013arXiv

A study on the quasiconinuum approximations of a one-dimensional fracture model

We study three quasicontinuum approximations of a lattice model for crack propagation. The influence of the approximation on the bifurcation patterns is investigated. The estimate of the modeling error is applicable to near and beyond bifurcation points, which enables us to evaluate the approximation over a finite range of loading and multiple mechanical equilibria.

preprint2013arXiv

Accurate Evaluations of Strain and Stress in Atomistic Simulations of Crystalline Solids

In this paper, we study the accuracy of Irving-Kirkwood type of formulas for the approximation of continuum quantities from atomistic simulations. Such formulas are derived by expressing the displacement, deformation gradient and stress in terms of certain kernel functions. We propose two criteria for choosing the kernel functions to significantly improve the sampling accuracy. We present a simple procedure to construct kernel functions that meet these criteria. Further, numerical tests on homogeneous and non-homogeneous systems provide validations for our analysis.

preprint2013arXiv

On the Cauchy-Born Approximation at Finite Temperature

We address several issues regarding the derivation and implementation of the Cauchy-Born approximation of the stress at finite temperature. In particular, an asymptotic expansion is employed to derive a closed form expression for the first Piola-Kirchhoff stress. For systems under periodic boundary conditions, a derivation is presented, which takes into account the translational invariance and clarifies the removal of the zero phonon modes. Also revealed by the asymptotic approach is the role of the smoothness of the interatomic potential. Several numerical examples are provided to validate this approach.

preprint2012arXiv

Coarse-graining molecular dynamics models using an extended Galerkin projection

We present a new framework for coarse-graining molecular dynamics models for crystalline solids. The reduction method is based on a Galerkin projection to a subspace, whose dimension is much smaller than that of the full atomistic model. The subspace is expanded by adding more coarse-grain variables near the interface between lattice defects and the surrounding regions. This effectively minimizes reflection of phonons at the interface. In this approach, there is no need to pre-compute the memory function in the generalized Langevin equations, a typical model of interface conditions. Moreover, the variational formulation preserves the stability of mechanical equilibria.