Source author record

Hailiang Liu

Hailiang Liu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

28works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

28 published item(s)

preprint2022arXiv

A global convergence theory for deep ReLU implicit networks via over-parameterization

Implicit deep learning has received increasing attention recently due to the fact that it generalizes the recursive prediction rules of many commonly used neural network architectures. Its prediction rule is provided implicitly based on the solution of an equilibrium equation. Although a line of recent empirical studies has demonstrated its superior performances, the theoretical understanding of implicit neural networks is limited. In general, the equilibrium equation may not be well-posed during the training. As a result, there is no guarantee that a vanilla (stochastic) gradient descent (SGD) training nonlinear implicit neural networks can converge. This paper fills the gap by analyzing the gradient flow of Rectified Linear Unit (ReLU) activated implicit neural networks. For an $m$-width implicit neural network with ReLU activation and $n$ training samples, we show that a randomly initialized gradient descent converges to a global minimum at a linear rate for the square loss function if the implicit neural network is \textit{over-parameterized}. It is worth noting that, unlike existing works on the convergence of (S)GD on finite-layer over-parameterized neural networks, our convergence results hold for implicit neural networks, where the number of layers is \textit{infinite}.

preprint2022arXiv

Global Dynamics and Photon Loss in the Kompaneets Equation

The Kompaneets equation governs dynamics of the photon energy spectrum in certain high temperature (or low density) plasmas. We prove several results concerning the long-time convergence of solutions to Bose--Einstein equilibria and the failure of photon conservation. In particular, we show the total photon number can decrease with time via an outflux of photons at the zero-energy boundary. The ensuing accumulation of photons at zero energy is analogous to Bose--Einstein condensation. We provide two conditions that guarantee that photon loss occurs, and show that once loss is initiated then it persists forever. We prove that as $t\to \infty$, solutions necessarily converge to equilibrium and we characterize the limit in terms of the total photon loss. Additionally, we provide a few results concerning the behavior of the solution near the zero-energy boundary, an Oleinik inequality, a comparison principle, and show that the solution operator is a contraction in $L^1$. None of these results impose a boundary condition at the zero-energy boundary.

preprint2022arXiv

SGEM: stochastic gradient with energy and momentum

In this paper, we propose SGEM, Stochastic Gradient with Energy and Momentum, to solve a large class of general non-convex stochastic optimization problems, based on the AEGD method that originated in the work [AEGD: Adaptive Gradient Descent with Energy. arXiv: 2010.05109]. SGEM incorporates both energy and momentum at the same time so as to inherit their dual advantages. We show that SGEM features an unconditional energy stability property, and derive energy-dependent convergence rates in the general nonconvex stochastic setting, as well as a regret bound in the online convex setting. A lower threshold for the energy variable is also provided. Our experimental results show that SGEM converges faster than AEGD and generalizes better or at least as well as SGDM in training some deep neural networks.

preprint2021arXiv

Energy stable Runge-Kutta discontinuous Galerkin schemes for fourth order gradient flows

We present unconditionally energy stable Runge-Kutta (RK) discontinuous Galerkin (DG) schemes for solving a class of fourth order gradient flows. Our algorithm is geared toward arbitrarily high order approximations in both space and time, while energy dissipation remains preserved without imposing any restriction on time steps and meshes. We achieve this in two steps. First, taking advantage of the penalty free DG method introduced by Liu and Yin [J Sci. Comput. 77:467--501, 2018] for spatial discretization, we reformulate an extended linearized ODE system by the energy quadratization (EQ) approach. Second, we apply an s-stage algebraically stable RK method for temporal discretization. The resulting fully discrete DG schemes are linear and unconditionally energy stable. In addition, we introduce a prediction-correction procedure to improve both the accuracy and stability of the scheme. We illustrate the effectiveness of the proposed schemes by numerical tests with benchmark problems.

preprint2021arXiv

Necessary conditions for blow-up solutions to the restricted Euler--Poisson equations

In this work, we study the behavior of blow-up solutions to the multidimensional restricted Euler--Poisson equations which are the localized version of the full Euler--Poisson system. We provide necessary conditions for the existence of finite-time blow-up solutions in terms of the initial data, and describe the asymptotic behavior of the solutions near blow up times. We also identify a rich set of the initial data which yields global bounded solutions.

preprint2021arXiv

Positivity-preserving third order DG schemes for Poisson--Nernst--Planck equations

In this paper, we design and analyze third order positivity-preserving discontinuous Galerkin (DG) schemes for solving the time-dependent system of Poisson--Nernst--Planck (PNP) equations, which has found much use in diverse applications. Our DG method with Euler forward time discretization is shown to preserve the positivity of cell averages at all time steps. The positivity of numerical solutions is then restored by a scaling limiter in reference to positive weighted cell averages. The method is also shown to preserve steady states. Numerical examples are presented to demonstrate the third order accuracy and illustrate the positivity-preserving property in both one and two dimensions.

preprint2021arXiv

Unconditionally energy stable discontinuous Galerkin schemes for the Cahn-Hilliard equation

In this paper, we introduce novel discontinuous Galerkin (DG) schemes for the Cahn-Hilliard equation, which arises in many applications. The method is designed by integrating the mixed DG method for the spatial discretization with the \emph{Invariant Energy Quadratization} (IEQ) approach for the time discretization. Coupled with a spatial projection, the resulting IEQ-DG schemes are shown to be unconditionally energy dissipative, and can be efficiently solved without resorting to any iteration method. Both one and two dimensional numerical examples are provided to verify the theoretical results, and demonstrate the good performance of IEQ-DG in terms of efficiency, accuracy, and preservation of the desired solution properties.

preprint2020arXiv

A positivity-preserving and energy stable scheme for a quantum diffusion equation

We propose a new fully-discretized finite difference scheme for a quantum diffusion equation, in both one and two dimensions. This is the first fully-discretized scheme with proven positivity-preserving and energy stable properties using only standard finite difference discretization. The difficulty in proving the positivity-preserving property lies in the lack of a maximum principle for fourth order PDEs. To overcome this difficulty, we reformulate the scheme as an optimization problem based on variational structure and use the singular nature of the energy functional near the boundary values to exclude the possibility of non-positive solutions. The scheme is also shown to be mass conservative and consistent.

preprint2020arXiv

An energy stable and positivity-preserving scheme for the Maxwell-Stefan diffusion system

We develop a new finite difference scheme for the Maxwell-Stefan diffusion system. The scheme is conservative, energy stable and positivity-preserving. These nice properties stem from a variational structure and are proved by reformulating the finite difference scheme into an equivalent optimization problem. The solution to the scheme emerges as the minimizer of the optimization problem, and as a consequence energy stability and positivity-preserving properties are obtained.

preprint2020arXiv

Dynamics of many species through competition for resources

This paper is concerned with a mathematical model of competition for resource where species consume noninteracting resources. This system of differential equations is formally obtained by renormalizing the MacArthur's competition model at equilibrium, and agrees with the trait-continuous model studied by Mirrahimi S, Perthame B, Wakano JY [J. Math. Biol. 64(7): 1189-1223, 2012]. As a dynamical system, self-organized generation of distinct species occurs. The necessary conditions for survival are given. We prove the existence of the evolutionary stable distribution (ESD) through an optimization problem and present an independent algorithm to compute the ESD directly. Under certain structural conditions, solutions of the system are shown to approach the discrete ESD as time evolves. The time discretization of the system is proven to satisfy two desired properties: positivity and energy dissipation. Numerical examples are given to illustrate certain interesting biological phenomena.

preprint2020arXiv

Efficient, positive, and energy stable schemes for multi-D Poisson-Nernst-Planck systems

In this paper, we design, analyze, and numerically validate positive and energy-dissipating schemes for solving the time-dependent multi-dimensional system of Poisson-Nernst-Planck (PNP) equations, which has found much use in the modeling of biological membrane channels and semiconductor devices. The semi-implicit time discretization based on a reformulation of the system gives a well-posed elliptic system, which is shown to preserve solution positivity for arbitrary time steps. The first order (in time) fully-discrete scheme is shown to preserve solution positivity and mass conservation unconditionally, and energy dissipation with only a mild $O(1)$ time step restriction. The scheme is also shown to preserve the steady-state. For the fully second order (in both time and space) scheme with large time steps, solution positivity is restored by a local scaling limiter, which is shown to maintain the spatial accuracy. These schemes are easy to implement. Several three-dimensional numerical examples verify our theoretical findings and demonstrate the accuracy, efficiency, and robustness of the proposed schemes, as well as the fast approach to steady states.

preprint2020arXiv

On the SAV-DG method for a class of fourth order gradient flows

For a class of fourth order gradient flow problems, integration of the scalar auxiliary variable (SAV) time discretization with the penalty-free discontinuous Galerkin (DG) spatial discretization leads to SAV-DG schemes. These schemes are linear and shown unconditionally energy stable. But the reduced linear systems are rather expensive to solve due to the dense coefficient matrices. In this paper, we provide a procedure to pre-evaluate the auxiliary variable in the piecewise polynomial space. As a result, the computational complexity of $O(\mathcal{N}^2)$ reduces to $O(\mathcal{N})$ when exploiting the conjugate gradient (CG) solver. This hybrid SAV-DG method is more efficient and able to deliver satisfactory results of high accuracy. This was also compared with solving the full augmented system of the SAV-DG schemes.

preprint2020arXiv

Radially symmetric solutions of the ultra-relativistic Euler equations

The ultra-relativistic Euler equations for an ideal gas are described in terms of the pressure $p$, the spatial part $\underline{u} \in \R^3$ of the dimensionless four-velocity and the particle density $n$. Radially symmetric solutions of these equations are studied. Analytical solutions are presented for the linearized system. For the original nonlinear equations we design and analyze a numerical scheme for simulating radially symmetric solutions in three space dimensions. The good performance of the scheme is demonstrated by numerical examples. In particular, it was observed that the method has the capability to capture accurately the pressure singularity formation caused by shock wave reflections at the origin.

preprint2020arXiv

Selection dynamics for deep neural networks

This paper presents a partial differential equation framework for deep residual neural networks and for the associated learning problem. This is done by carrying out the continuum limits of neural networks with respect to width and depth. We study the wellposedness, the large time solution behavior, and the characterization of the steady states of the forward problem. Several useful time-uniform estimates and stability/instability conditions are presented. We state and prove optimality conditions for the inverse deep learning problem, using standard variational calculus, the Hamilton-Jacobi-Bellmann equation and the Pontryagin maximum principle. This serves to establish a mathematical foundation for investigating the algorithmic and theoretical connections between neural networks, PDE theory, variational analysis, optimal control, and deep learning.

preprint2019arXiv

Positive and free energy satisfying schemes for diffusion with interaction potentials

In this paper, we design and analyze second order positive and free energy satisfying schemes for solving diffusion equations with interaction potentials. The semi-discrete scheme is shown to conserve mass, preserve solution positivity, and satisfy a discrete free energy dissipation law for nonuniform meshes. These properties for the fully-discrete scheme (first order in time) remain preserved without a strict restriction on time steps. For the fully second order (in both time and space) scheme, we use a local scaling limiter to restore solution positivity when necessary. It is proved that such limiter does not destroy the second order accuracy. In addition, these schemes are easy to implement, and efficient in simulations over long time. Both one and two dimensional numerical examples are presented to demonstrate the performance of these schemes.

preprint2016arXiv

A free energy satisfying discontinuous Galerkin method for one-dimensional Poisson--Nernst--Planck systems

We design an arbitrary-order free energy satisfying discontinuous Galerkin (DG) method for solving time-dependent Poisson-Nernst-Planck systems. Both the semi-discrete and fully discrete DG methods are shown to satisfy the corresponding discrete free energy dissipation law for positive numerical solutions. Positivities of numerical solutions are enforced by an accuracy-preserving limiter in reference to positive cell averages. Numerical examples are presented to demonstrate the high resolution of the numerical algorithm and to illustrate the proven properties of mass conservation, free energy dissipation, as well as the preservation of steady states.

preprint2016arXiv

An entropy satisfying discontinuous Galerkin method for nonlinear Fokker-Planck equations

We propose a high order discontinuous Galerkin (DG) method for solving nonlinear Fokker-Planck equations with a gradient flow structure. For some of these models it is known that the transient solutions converge to steady-states when time tends to infinity. The scheme is shown to satisfy a discrete version of the entropy dissipation law and preserve steady-states, therefore providing numerical solutions with satisfying long-time behavior. The positivity of numerical solutions is enforced through a reconstruction algorithm, based on positive cell averages. For the model with trivial potential, a parameter range sufficient for positivity preservation is rigorously established. For other cases, cell averages can be made positive at each time step by tuning the numerical flux parameters. A selected set of numerical examples is presented to confirm both the high-order accuracy and the efficiency to capture the large-time asymptotic.

preprint2015arXiv

Sobolev and Max Norm Error Estimates for Gaussian Beam Superpositions

This work is concerned with the accuracy of Gaussian beam superpositions, which are asymptotically valid high frequency solutions to linear hyperbolic partial differential equations and the Schrödinger equation. We derive Sobolev and max norms estimates for the difference between an exact solution and the corresponding Gaussian beam approximation, in terms of the short wavelength $\varepsilon$. The estimates are performed for the scalar wave equation and the Schrödinger equation. Our result demonstrates that a Gaussian beam superposition with $k$-th order beams converges to the exact solution as $O(\varepsilon^{k/2-s})$ in order $s$ Sobolev norms. This result is valid in any number of spatial dimensions and it is unaffected by the presence of caustics in the solution. In max norm, we show that away from caustics the convergence rate is $O(\varepsilon^{\lceil k/2\rceil})$ and away from the essential support of the solution, the convergence is spectral in $\varepsilon$. However, in the neighborhood of a caustic point we are only able to show the slower, and dimensional dependent, rate $O(\varepsilon^{(k-n)/2})$ in $n$ spatial dimensions.

preprint2014arXiv

A free energy satisfying finite difference method for Poisson--Nernst--Planck equations

In this work we design and analyze a free energy satisfying finite difference method for solving Poisson-Nernst-Planck equations in a bounded domain. The algorithm is of second order in space, with numerical solutions satisfying all three desired properties: i) mass conservation, ii) positivity preserving, and iii) free energy satisfying in the sense that these schemes satisfy a discrete free energy dissipation inequality. These ensure that the computed solution is a probability density, and the schemes are energy stable and preserve the equilibrium solutions. Both one and two-dimensional numerical results are provided to demonstrate the good qualities of the algorithm, as well as effects of relative size of the data given.

preprint2014arXiv

Error Estimates of the Bloch Band-Based Gaussian Beam Superposition for the Schrödinger Equation

This work is concerned with asymptotic approximations of the semi-classical Schrödinger equation in periodic media using Gaussian beams. For the underlying equation, subject to a highly oscillatory initial data, a hybrid of the Gaussian beam approximation and homogenization leads to the Bloch eigenvalue problem and associated evolution equations for Gaussian beam components in each Bloch band. We formulate a superposition of Bloch-band based Gaussian beams to generate high frequency approximate solutions to the original wave field. For initial data of a sum of finite number of band eigen-functions, we prove that the first-order Gaussian beam superposition converges to the original wave field at a rate of $ε^{1/2}$, with $ε$ the semiclassically scaled constant, as long as the initial data for Gaussian beam components in each band are prepared with same order of error or smaller. For a natural choice of initial approximation, a rate of $ε^{1/2}$ of initial error is verified.

preprint2013arXiv

Thresholds for shock formation in traffic flow models with Arrhenius look-ahead dynamics

We investigate a class of nonlocal conservation laws with the nonlinear advection coupling both local and nonlocal mechanism, which arises in several applications such as the collective motion of cells and traffic flows. It is proved that the $C^1$ solution regularity of this class of conservation laws will persist at least for a short time. This persistency may continue as long as the solution gradient remains bounded. Based on this result, we further identify sub-thresholds for finite time shock formation in traffic flow models with Arrhenius look-ahead dynamics.

preprint2011arXiv

Error Estimates for Gaussian Beam Superpositions

Gaussian beams are asymptotically valid high frequency solutions to hyperbolic partial differential equations, concentrated on a single curve through the physical domain. They can also be extended to some dispersive wave equations, such as the Schrödinger equation. Superpositions of Gaussian beams provide a powerful tool to generate more general high frequency solutions that are not necessarily concentrated on a single curve. This work is concerned with the accuracy of Gaussian beam superpositions in terms of the wavelength $ε$. We present a systematic construction of Gaussian beam superpositions for all strictly hyperbolic and Schrödinger equations subject to highly oscillatory initial data of the form $Ae^{iΦ/ε}$. Through a careful estimate of an oscillatory integral operator, we prove that the $k$-th order Gaussian beam superposition converges to the original wave field at a rate proportional to $ε^{k/2}$ in the appropriate norm dictated by the well-posedness estimate. In particular, we prove that the Gaussian beam superposition converges at this rate for the acoustic wave equation in the standard, $ε$-scaled, energy norm and for the Schrödinger equation in the $L^2$ norm. The obtained results are valid for any number of spatial dimensions and are unaffected by the presence of caustics. We present a numerical study of convergence for the constant coefficient acoustic wave equation in $\Real^2$ to analyze the sharpness of the theoretical results.

preprint2010arXiv

Global Well-Posedness for the Microscopic FENE Model with a Sharp Boundary Condition

We prove global well-posedness for the microscopic FENE model under a sharp boundary requirement. The well-posedness of the FENE model that consists of the incompressible Navier-Stokes equation and the Fokker-Planck equation has been studied intensively, mostly with the zero flux boundary condition. Recently it was illustrated by C. Liu and H. Liu [2008, SIAM J. Appl. Math., 68(5):1304--1315] that any preassigned boundary value of a weighted distribution will become redundant once the non-dimensional parameter $b>2$. In this article, we show that for the well-posedness of the microscopic FENE model ($b>2$) the least boundary requirement is that the distribution near boundary needs to approach zero faster than the distance function. Under this condition, it is shown that there exists a unique weak solution in a weighted Sobolev space. Moreover, such a condition still ensures that the distribution is a probability density. The sharpness of this boundary requirement is shown by a construction of infinitely many solutions when the distribution approaches zero as fast as the distance function.

preprint2010arXiv

Numerical approximation of the Euler-Poisson-Boltzmann model in the quasineutral limit

This paper analyzes various schemes for the Euler-Poisson-Boltzmann (EPB) model of plasma physics. This model consists of the pressureless gas dynamics equations coupled with the Poisson equation and where the Boltzmann relation relates the potential to the electron density. If the quasi-neutral assumption is made, the Poisson equation is replaced by the constraint of zero local charge and the model reduces to the Isothermal Compressible Euler (ICE) model. We compare a numerical strategy based on the EPB model to a strategy using a reformulation (called REPB formulation). The REPB scheme captures the quasi-neutral limit more accurately.

preprint2010arXiv

The Cauchy-Dirichlet problem for the FENE dumbbell model of polymeric fluids

The FENE dumbbell model consists of the incompressible Navier-Stokes equation and the Fokker-Planck equation for the polymer distribution. In such a model, the polymer elongation cannot exceed a limit $\sqrt{b}$, yielding all interesting features near the boundary. In this paper we establish the local well-posedness for the FENE dumbbell model under a class of Dirichlet-type boundary conditions dictated by the parameter $b$. As a result, for each $b>0$ we identify a sharp boundary requirement for the underlying density distribution, while the sharpness follows from the existence result for each specification of the boundary behavior. It is shown that the probability density governed by the Fokker-Planck equation approaches zero near boundary, necessarily faster than the distance function $d$ for $b>2$, faster than $d|ln d|$ for $b=2$, and as fast as $d^{b/2}$ for $0<b<2$. Moreover, the sharp boundary requirement for $b\geq 2$ is also sufficient for the distribution to remain a probability density.

preprint2009arXiv

Kinetic models for polymers with inertial effects

Novel kinetic models for both Dumbbell-like and rigid-rod like polymers are derived, based on the probability distribution function $f(t, x, n, \dot n)$ for a polymer molecule positioned at $x$ to be oriented along direction $n$ while embedded in a $\dot n$ environment created by inertial effects. It is shown that the probability distribution function of the extended model, when converging, will lead to well accepted kinetic models when inertial effects are ignored such as the Doi models for rod like polymers, and the Finitely Extensible Non-linear Elastic (FENE) models for Dumbbell like polymers.

preprint2007arXiv

Rigorous derivation of the hydrodynamical equations for rotating superfluids

Using a modified WKB approach, we present a rigorous semi-classical analysis for solutions of nonlinear Schroedinger equations with rotational forcing. This yields a rigorous justification for the hydrodynamical system of rotating superfluids. In particular it is shown that global-in-time semi-classical convergence holds whenever the limiting hydrodynamical system has global smooth solutions and we also discuss the semi-classical dynamics of several physical quantities describing rotating superfluids.

preprint2003arXiv

Rotation Prevents Finite-Time Breakdown

We consider a two-dimensional convection model augmented with the rotational Coriolis forcing, $U_t + U\cdot\nabla_x U = 2k U^\perp$, with a fixed $2k$ being the inverse Rossby number. We ask whether the action of dispersive rotational forcing alone, $U^\perp$, prevents the generic finite time breakdown of the free nonlinear convection. The answer provided in this work is a conditional yes. Namely, we show that the rotating Euler equations admit global smooth solutions for a subset of generic initial configurations. With other configurations, however, finite time breakdown of solutions may and actually does occur. Thus, global regularity depends on whether the initial configuration crosses an intrinsic, ${\mathcal O}(1)$ critical threshold, which is quantified in terms of the initial vorticity, $ω_0=\nabla \times U_0$, and the initial spectral gap associated with the $2\times 2$ initial velocity gradient, $η_0:=λ_2(0)-λ_1(0), λ_j(0)= λ_j(\nabla U_0)$. Specifically, global regularity of the rotational Euler equation is ensured if and only if $4k ω_0(α) +η^2_0(α) <4k^2, \forall α\in \R^2$ . We also prove that the velocity field remains smooth if and only if it is periodic. We observe yet another remarkable periodic behavior exhibited by the {\em gradient} of the velocity field. The spectral dynamics of the Eulerian formulation reveals that the vorticity and the eigenvalues (and hence the divergence) of the flow evolve with their own path-dependent period. We conclude with a kinetic formulation of the rotating Euler equation.