Source author record

Zhiqiang Cai

Zhiqiang Cai appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

15works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

15 published item(s)

preprint2022arXiv

Adaptive Two-Layer ReLU Neural Network: I. Best Least-squares Approximation

In this paper, we introduce adaptive neuron enhancement (ANE) method for the best least-squares approximation using two-layer ReLU neural networks (NNs). For a given function f(x), the ANE method generates a two-layer ReLU NN and a numerical integration mesh such that the approximation accuracy is within the prescribed tolerance. The ANE method provides a natural process for obtaining a good initialization which is crucial for training nonlinear optimization problems. Numerical results of the ANE method are presented for functions of two variables exhibiting either intersecting interface singularities or sharp interior layers.

preprint2022arXiv

Learn Quasi-stationary Distributions of Finite State Markov Chain

We propose a reinforcement learning (RL) approach to compute the expression of quasi-stationary distribution. Based on the fixed-point formulation of quasi-stationary distribution, we minimize the KL-divergence of two Markovian path distributions induced by the candidate distribution and the true target distribution. To solve this challenging minimization problem by gradient descent, we apply the reinforcement learning technique by introducing the reward and value functions. We derive the corresponding policy gradient theorem and design an actor-critic algorithm to learn the optimal solution and the value function. The numerical examples of finite state Markov chain are tested to demonstrate the new method.

preprint2022arXiv

Least-Squares ReLU Neural Network (LSNN) Method For Scalar Nonlinear Hyperbolic Conservation Law

We introduced the least-squares ReLU neural network (LSNN) method for solving the linear advection-reaction problem with discontinuous solution and showed that the method outperforms mesh-based numerical methods in terms of the number of degrees of freedom. This paper studies the LSNN method for scalar nonlinear hyperbolic conservation law. The method is a discretization of an equivalent least-squares (LS) formulation in the set of neural network functions with the ReLU activation function. Evaluation of the LS functional is done by using numerical integration and conservative finite volume scheme. Numerical results of some test problems show that the method is capable of approximating the discontinuous interface of the underlying problem automatically through the free breaking lines of the ReLU neural network. Moreover, the method does not exhibit the common Gibbs phenomena along the discontinuous interface.

preprint2022arXiv

Self-adaptive deep neural network: Numerical approximation to functions and PDEs

Designing an optimal deep neural network for a given task is important and challenging in many machine learning applications. To address this issue, we introduce a self-adaptive algorithm: the adaptive network enhancement (ANE) method, written as loops of the form train, estimate and enhance. Starting with a small two-layer neural network (NN), the step train is to solve the optimization problem at the current NN; the step estimate is to compute a posteriori estimator/indicators using the solution at the current NN; the step enhance is to add new neurons to the current NN. Novel network enhancement strategies based on the computed estimator/indicators are developed in this paper to determine how many new neurons and when a new layer should be added to the current NN. The ANE method provides a natural process for obtaining a good initialization in training the current NN; in addition, we introduce an advanced procedure on how to initialize newly added neurons for a better approximation. We demonstrate that the ANE method can automatically design a nearly minimal NN for learning functions exhibiting sharp transitional layers as well as discontinuous solutions of hyperbolic partial differential equations.

preprint2020arXiv

Deep least-squares methods: an unsupervised learning-based numerical method for solving elliptic PDEs

This paper studies an unsupervised deep learning-based numerical approach for solving partial differential equations (PDEs). The approach makes use of the deep neural network to approximate solutions of PDEs through the compositional construction and employs least-squares functionals as loss functions to determine parameters of the deep neural network. There are various least-squares functionals for a partial differential equation. This paper focuses on the so-called first-order system least-squares (FOSLS) functional studied in [3], which is based on a first-order system of scalar second-order elliptic PDEs. Numerical results for second-order elliptic PDEs in one dimension are presented.

preprint2020arXiv

Generalized Prager-Synge Inequality and Equilibrated Error Estimators for Discontinuous Elements

The well-known Prager-Synge identity is valid in $H^1(Ω)$ and serves as a foundation for developing equilibrated a posteriori error estimators for continuous elements. In this paper, we introduce a new inequality, that may be regarded as a generalization of the Prager-Synge identity, to be valid for piecewise $H^1(Ω)$ functions for diffusion problems. The inequality is proved to be identity in two dimensions. For nonconforming finite element approximation of arbitrary odd order, we propose a fully explicit approach that recovers an equilibrated flux in $H(div; Ω)$ through a local element-wise scheme and that recovers a gradient in $H(curl;Ω)$ through a simple averaging technique over edges. The resulting error estimator is then proved to be globally reliable and locally efficient. Moreover, the reliability and efficiency constants are independent of the jump of the diffusion coefficient regardless of its distribution.

preprint2016arXiv

An Empirical Study on Academic Commentary and Its Implications on Reading and Writing

The relationship between reading and writing (RRW) is one of the major themes in learning science. One of its obstacles is that it is difficult to define or measure the latent background knowledge of the individual. However, in an academic research setting, scholars are required to explicitly list their background knowledge in the citation sections of their manuscripts. This unique opportunity was taken advantage of to observe RRW, especially in the published academic commentary scenario. RRW was visualized under a proposed topic process model by using a state of the art version of latent Dirichlet allocation (LDA). The empirical study showed that the academic commentary is modulated both by its target paper and the author's background knowledge. Although this conclusion was obtained in a unique environment, we suggest its implications can also shed light on other similar interesting areas, such as dialog and conversation, group discussion, and social media.

preprint2016arXiv

Finite Element Methods for Interface Problems: Robust Residual-Based A Posteriori Error Estimates

For elliptic interface problems, this paper studies residual-based a posteriori error estimations for various finite element approximations. For the conforming and the Raviart-Thomas mixed elements in two-dimension and for the Crouzeix-Raviart nonconforming and the discontinuous Galerkin elements in both two- and three-dimensions, the global reliability bounds are established with constants independent of the jump of the diffusion coefficient. Moreover, we obtain these estimates with no assumption on the distribution of the diffusion coefficient.

preprint2016arXiv

Improved ZZ A Posteriori Error Estimators for Diffusion Problems: Conforming Linear Elements

In \cite{CaZh:09}, we introduced and analyzed an improved Zienkiewicz-Zhu (ZZ) estimator for the conforming linear finite element approximation to elliptic interface problems. The estimator is based on the piecewise "constant" flux recovery in the $H(div;Ω)$ conforming finite element space. This paper extends the results of \cite{CaZh:09} to diffusion problems with full diffusion tensor and to the flux recovery both in piecewise constant and piecewise linear $H(div)$ space.

preprint2016arXiv

Residual-based a Posteriori Error Estimate for Interface Problems: Nonconforming Linear Elements

In this paper, we study a modified residual-based a posteriori error estimator for the nonconforming linear finite element approximation to the interface problem. The reliability of the estimator is analyzed by a new and direct approach without using the Helmholtz decomposition. It is proved that the estimator is reliable with constant independent of the jump of diffusion coefficients across the interfaces, without the assumption that the diffusion coefficient is quasi-monotone. Numerical results for one test problem with intersecting interfaces are also presented.

preprint2016arXiv

Robust A Posteriori Error Estimation for Finite Element Approximation to H(curl) Problem

In this paper, we introduce a novel a posteriori error estimator for the conforming finite element approximation to the H(curl) problem with inhomogeneous media and with the right-hand side only in L^2. The estimator is of the recovery type. Independent with the current approximation to the primary variable (the electric field), an auxiliary variable (the magnetizing field) is recovered in parallel by solving a similar H(curl) problem. An alternate way of recovery is presented as well by localizing the error flux. The estimator is then defined as the sum of the modified element residual and the residual of the constitutive equation defining the auxiliary variable. It is proved that the estimator is approximately equal to the true error in the energy norm without the quasi-monotonicity assumption. Finally, we present numerical results for two H(curl) interface problems.

preprint2015arXiv

A Recovery-Based A Posteriori Error Estimator for H(curl) Interface Problems

This paper introduces a new recovery-based a posteriori error estimator for the lowest order Nedelec finite element approximation to the H(curl) interface problem. The error estimator is analyzed by establishing both the reliability and the efficiency bounds and is supported by numerical results. Under certain assumptions, it is proved that the reliability and efficiency constants are independent of the jumps of the coefficients.

preprint2015arXiv

Finite Element Methods for Interface Problems: Robust and Local Optimal A Priori Error Estimates

For elliptic interface problems in two- and three-dimensions, this paper establishes a priori error estimates for Crouzeix-Raviart nonconforming, Raviart-Thomas mixed, and discontinuous Galerkin finite element approximations. These estimates are robust with respect to the diffusion coefficient and optimal with respect to local regularity of the solution. Moreover, we obtain these estimates with no assumption on the distribution of the diffusion coefficient.

preprint2014arXiv

Div First-Order System LL* (FOSLL*) for Second-Order Elliptic Partial Differential Equations

The first-order system LL* (FOSLL*) approach for general second-order elliptic partial differential equations was proposed and analyzed in [10], in order to retain the full efficiency of the L2 norm first-order system least-squares (FOSLS) ap- proach while exhibiting the generality of the inverse-norm FOSLS approach. The FOSLL* approach in [10] was applied to the div-curl system with added slack vari- ables, and hence it is quite complicated. In this paper, we apply the FOSLL* approach to the div system and establish its well-posedness. For the corresponding finite ele- ment approximation, we obtain a quasi-optimal a priori error bound under the same regularity assumption as the standard Galerkin method, but without the restriction to sufficiently small mesh size. Unlike the FOSLS approach, the FOSLL* approach does not have a free a posteriori error estimator, we then propose an explicit residual error estimator and establish its reliability and efficiency bounds

preprint2014arXiv

Recovery-Based Error Estimators for Diffusion Problems: Explicit Formulas

We introduced and analyzed robust recovery-based a posteriori error estimators for various lower order finite element approximations to interface problems in [9, 10], where the recoveries of the flux and/or gradient are implicit (i.e., requiring solutions of global problems with mass matrices). In this paper, we develop fully explicit recovery-based error estimators for lower order conforming, mixed, and non- conforming finite element approximations to diffusion problems with full coefficient tensor. When the diffusion coefficient is piecewise constant scalar and its distribution is local quasi-monotone, it is shown theoretically that the estimators developed in this paper are robust with respect to the size of jumps. Numerical experiments are also performed to support the theoretical results.