Source author record

Shi Shu

Shi Shu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2022arXiv

Error Analysis of Virtual Element Methods for the Time-dependent Poisson-Nernst-Planck Equations

We discuss and analyze the virtual element method on general polygonal meshes for the time-dependent Poisson-Nernst-Planck equations, which are a nonlinear coupled system widely used in semiconductors and ion channels. The spatial discretization is based on the elliptic projection and the $L^2$ projection operator, and for the temporal discretization, the backward Euler scheme is employed. After presenting the semi and fully discrete schemes, we derive the a priori error estimates in the $L^2$ and $H^1$ norms. Finally, a numerical experiment verifies the theoretical convergence results.

preprint2022arXiv

Parallel Multi-Stage Preconditioners with Adaptive Setup for the Black Oil Model

The black oil model is widely used to describe multiphase porous media flow in the petroleum industry. The fully implicit method features strong stability and weak constraints on time step-sizes; hence, commonly used in the current mainstream commercial reservoir simulators. In this paper, a CPR-type preconditioner with an adaptive "setup phase" is developed to improve parallel efficiency of petroleum reservoir simulation. Furthermore, we propose a multi-color Gauss-Seidel (GS) algorithm for algebraic multigrid method based on the coefficient matrix of strong connections. Numerical experiments show that the proposed preconditioner can improve the parallel performance for both OpenMP and CUDA implements. Moreover, the proposed algorithm yields good parallel speedup as well as same convergence behavior as the corresponding single-threaded algorithm. In particular, for a three-phase benchmark problem, the parallel speedup of the OpenMP version is over 6.5 with 16 threads and the CUDA version reaches more than 9.5.

preprint2021arXiv

A multigrid-reduction-in-time solver with a new two-level convergence for unsteady fractional Laplacian problems

The multigrid-reduction-in-time (MGRIT) technique has proven to be successful in achieving higher run-time speedup by exploiting parallelism in time. The goal of this article is to develop and analyze a MGRIT algorithm, using FCF-relaxation with time-dependent time-grid propagators, to seek the finite element approximations of unsteady fractional Laplacian problems. The multigrid with line smoother proposed in [L. Chen, R. H. Nochetto, E. Ot{á}rola, A. J. Salgado, Math. Comp. 85 (2016) 2583--2607] is chosen to be the spatial solver. Motivated by [B. S. Southworth, SIAM J. Matrix Anal. Appl. 40 (2019) 564--608], we provide a new temporal eigenvalue approximation property and then deduce a generalized two-level convergence theory which removes the previous unitary diagonalization assumption on the fine and coarse time-grid propagators required in [X. Q. Yue, S. Shu, X. W. Xu, W. P. Bu, K. J. Pan, Comput. Math. Appl. 78 (2019) 3471--3484]. Numerical computations are included to confirm the theoretical predictions and demonstrate the sharpness of the derived convergence upper bound.

preprint2020arXiv

A Dynamic Subspace Based BFGS Method for Large Scale Optimization Problem

Large-scale unconstrained optimization is a fundamental and important class of, yet not well-solved problems in numerical optimization. The main challenge in designing an algorithm is to require a few storage locations or very inexpensive computations while preserving global convergence. In this work, we propose a novel approach solving large-scale unconstrained optimization problem by combining the dynamic subspace technique and the BFGS update algorithm. It is clearly demonstrated that our approach has the same rate of convergence in the dynamic subspace as the BFGS and less memory than L-BFGS. Further, we give the convergence analysis by constructing the mapping of low-dimensional Euclidean space to the adaptive subspace. We compare our hybrid algorithm with the BFGS and L-BFGS approaches. Experimental results show that our hybrid algorithm offers several significant advantages such as parallel computing, convergence efficiency, and robustness.

preprint2020arXiv

Adaptive-Multilevel BDDC algorithm for three-dimensional plane wave Helmholtz systems

In this paper, we are concerned with the weighted plane wave least-squares (PWLS) method for three-dimensional Helmholtz equations, and develop the multi-level adaptive BDDC algorithms for solving the resulting discrete system. In order to form the adaptive coarse components, the local generalized eigenvalue problems for each common face and each common edge are carefully designed. The condition number of the two-level adaptive BDDC preconditioned system is proved to be bounded above by a user-defined tolerance and a constant which is dependent on the maximum number of faces and edges per subdomain and the number of subdomains sharing a common edge. The efficiency of these algorithms is illustrated on a benchmark problem. The numerical results show the robustness of our two-level adaptive BDDC algorithms with respect to the wave number, the number of subdomains and the mesh size, and illustrate that our multi-level adaptive BDDC algorithm can reduce the scale of the coarse problem and can be used to solve large wave number problems efficiently.

preprint2020arXiv

Algebraic multigrid block preconditioning for multi-group radiation diffusion equations

The paper focuses on developing and studying efficient block preconditioners based on classical algebraic multigrid for the large-scale sparse linear systems arising from the fully coupled and implicitly cell-centered finite volume discretization of multi-group radiation diffusion equations, whose coefficient matrices can be rearranged into the $(G+2)\times(G+2)$ block form, where $G$ is the number of energy groups. The preconditioning techniques are based on the monolithic classical algebraic multigrid method, physical-variable based coarsening two-level algorithm and two types of block Schur complement preconditioners. The classical algebraic multigrid is applied to solve the subsystems that arise in the last three block preconditioners. The coupling strength and diagonal dominance are further explored to improve performance. We use representative one-group and twenty-group linear systems from capsule implosion simulations to test the robustness, efficiency, strong and weak parallel scaling properties of the proposed methods. Numerical results demonstrate that block preconditioners lead to mesh- and problem-independent convergence, and scale well both algorithmically and in parallel.

preprint2020arXiv

Local Averaging Type a Posteriori Error Estimates for the Nonlinear Steady-state Poisson-Nernst-Planck Equations

The a posteriori error estimates are studied for a class of nonlinear stead-state Poisson-Nernst-Planck equations, which are a coupled system consisting of the Nernst-Planck equation and the Poisson equation. Both the global upper bounds and the local lower bounds of the error estimators are obtained by using a local averaging operator. Numerical experiments are given to confirm the reliability and efficiency of the error estimators.

preprint2016arXiv

Error estimates on a finite volume method for diffusion problems with interface on Eulerian grids

The finite volume methods are frequently employed in the discretization of diffusion problems with interface. In this paper, we firstly present a vertex-centered MACH-like finite volume method for solving stationary diffusion problems with strong discontinuity and multiple material cells on the Eulerian grids. This method is motivated by Frese [No. AMRC-R-874, Mission Research Corp., Albuquerque, NM, 1987]. Then, the local truncation error and global error estimates of the degenerate five-point MACH-like scheme are derived by introducing some new techniques. Especially under some assumptions, we prove that this scheme can reach the asymptotic optimal error estimate $O(h^2 |\ln h|)$ in the maximum norm. Finally, numerical experiments verify theoretical results.

preprint2014arXiv

PIBM: Particulate immersed boundary method for fluid-particle interaction problems

It is well known that the number of particles should be scaled up to enable industrial scale simulation. The calculations are more computationally intensive when the motion of the surrounding fluid is considered. Besides the advances in computer hardware and numerical algorithms, the coupling scheme also plays an important role on the computational efficiency. In this study, a particle immersed boundary method (PIBM) for simulating the fluid-particle multiphase flow was presented and assessed in both two- and three-dimensional applications. The idea behind PIBM derives from the conventional momentum exchange-based immersed boundary method (IBM) by treating each Lagrangian point as a solid particle. This treatment enables LBM to be coupled with fine particles residing within a particular grid cell. Compared with the conventional IBM, dozens of times speedup in two-dimensional simulation and hundreds of times in three-dimensional simulation can be expected under the same particle and mesh number. Numerical simulations of particle sedimentation in the Newtonian flows were conducted based on a combined lattice Boltzmann method - particle immersed boundary method - discrete element method scheme, showing that the PIBM can capture the feature of particulate flows in fluid and is indeed a promising scheme for the solution of the fluid-particle interaction problems.

preprint2013arXiv

Numerical Study of Geometric Multigrid Methods on CPU--GPU Heterogeneous Computers

The geometric multigrid method (GMG) is one of the most efficient solving techniques for discrete algebraic systems arising from elliptic partial differential equations. GMG utilizes a hierarchy of grids or discretizations and reduces the error at a number of frequencies simultaneously. Graphics processing units (GPUs) have recently burst onto the scientific computing scene as a technology that has yielded substantial performance and energy-efficiency improvements. A central challenge in implementing GMG on GPUs, though, is that computational work on coarse levels cannot fully utilize the capacity of a GPU. In this work, we perform numerical studies of GMG on CPU--GPU heterogeneous computers. Furthermore, we compare our implementation with an efficient CPU implementation of GMG and with the most popular fast Poisson solver, Fast Fourier Transform, in the cuFFT library developed by NVIDIA.

preprint2011arXiv

Breaking a chaotic image encryption algorithm based on perceptron model

Recently, a chaotic image encryption algorithm based on perceptron model was proposed. The present paper analyzes security of the algorithm and finds that the equivalent secret key can be reconstructed with only one pair of known-plaintext/ciphertext, which is supported by both mathematical proof and experiment results. In addition, some other security defects are also reported.

preprint2011arXiv

Cryptanalyzing a chaos-based image encryption algorithm using alternate structure

Recently, a chaos-based image encryption algorithm using alternate structure (IEAS) was proposed. This paper focuses on differential cryptanalysis of the algorithm and finds that some properties of IEAS can support a differential attack to recover equivalent secret key with a little small number of known plain-images. Detailed approaches of the cryptanalysis for cryptanalyzing IEAS of the lower round number are presented and the breaking method can be extended to the case of higher round number. Both theoretical analysis and experiment results are provided to support vulnerability of IEAS against differential attack. In addition, some other security defects of IEAS, including insensitivity with respect to changes of plain-images and insufficient size of key space, are also reported.