Source author record

Wei Kang

Wei Kang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

27works
17topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

27 published item(s)

preprint2026arXiv

Movable Antenna Assisted Dual-Polarized Multi-Cell Cooperative AirComp: An Alternating Optimization Approach

Over-the-air computation (AirComp) is a key enabler for distributed optimization, since it leverages analog waveform superposition to perform aggregation and thereby mitigates the communication bottleneck caused by iterative information exchange. However, AirComp is sensitive to wireless environment and conventional systems with fixed single-polarized base-station arrays cannot fully exploit spatial degrees of freedom while also suffering from polarization mismatch. To overcome these limitations, this paper proposes a multi-cell cooperative air-computation framework assisted by dual-polarized movable antennas (D-PMA), and formulates a mean squared error (MSE) minimization problem by jointly optimizing the combining matrix, polarization vectors, antenna positions, and user transmit coefficients. The resulting problem is highly nonconvex, so an alternating algorithm is developed in which closed-form updates are obtained for the combining matrix and transmit coefficients. Then a method based on successive convex approximation (SCA) and semidefinite relaxation (SDR) is proposed to refine polarization vectors, and the antenna positions are updated using a gradient-based method. In addition, we develop a statistical-channel-based scheme for optimizing the antenna locations, and we further present the corresponding algorithm to efficiently obtain the solution. Numerical results show that the proposed movable dual-polarized scheme consistently outperforms movable single-polarized and fixed-antenna baselines under both instantaneous and statistical channels.

preprint2022arXiv

An Actor Critic Method for Free Terminal Time Optimal Control

Optimal control problems with free terminal time present many challenges including nonsmooth and discontinuous control laws, irregular value functions, many local optima, and the curse of dimensionality. To overcome these issues, we propose an adaptation of the model-based actor-critic paradigm from the field of Reinforcement Learning via an exponential transformation to learn an approximate feedback control and value function pair. We demonstrate the algorithm's effectiveness on prototypical examples featuring each of the main pathological issues present in problems of this type.

preprint2022arXiv

Machine Learning based Optimal Feedback Control for Microgrid Stabilization

Microgrids have more operational flexibilities as well as uncertainties than conventional power grids, especially when renewable energy resources are utilized. An energy storage based feedback controller can compensate undesired dynamics of a microgrid to improve its stability. However, the optimal feedback control of a microgrid subject to a large disturbance needs to solve a Hamilton-Jacobi-Bellman problem. This paper proposes a machine learning-based optimal feedback control scheme. Its training dataset is generated from a linear-quadratic regulator and a brute-force method respectively addressing small and large disturbances. Then, a three-layer neural network is constructed from the data for the purpose of optimal feedback control. A case study is carried out for a microgrid model based on a modified Kundur two-area system to test the real-time performance of the proposed control scheme.

preprint2022arXiv

Pruned RNN-T for fast, memory-efficient ASR training

The RNN-Transducer (RNN-T) framework for speech recognition has been growing in popularity, particularly for deployed real-time ASR systems, because it combines high accuracy with naturally streaming recognition. One of the drawbacks of RNN-T is that its loss function is relatively slow to compute, and can use a lot of memory. Excessive GPU memory usage can make it impractical to use RNN-T loss in cases where the vocabulary size is large: for example, for Chinese character-based ASR. We introduce a method for faster and more memory-efficient RNN-T loss computation. We first obtain pruning bounds for the RNN-T recursion using a simple joiner network that is linear in the encoder and decoder embeddings; we can evaluate this without using much memory. We then use those pruning bounds to evaluate the full, non-linear joiner network.

preprint2022arXiv

The Observability in Unobservable Systems

In this paper, we introduce the concept of observability of targeted state variables for systems that may not be fully observable. For their estimation, we introduce and exemplify a deep filter, which is a neural network specifically designed for the estimation of targeted state variables without computing the trajectory of the entire system. The observability definition is quantitative rather than a yes or no answer so that one can compare the level of observability between different sensor locations.

preprint2022arXiv

Towards A Critical Evaluation of Robustness for Deep Learning Backdoor Countermeasures

Since Deep Learning (DL) backdoor attacks have been revealed as one of the most insidious adversarial attacks, a number of countermeasures have been developed with certain assumptions defined in their respective threat models. However, the robustness of these countermeasures is inadvertently ignored, which can introduce severe consequences, e.g., a countermeasure can be misused and result in a false implication of backdoor detection. For the first time, we critically examine the robustness of existing backdoor countermeasures with an initial focus on three influential model-inspection ones that are Neural Cleanse (S&P'19), ABS (CCS'19), and MNTD (S&P'21). Although the three countermeasures claim that they work well under their respective threat models, they have inherent unexplored non-robust cases depending on factors such as given tasks, model architectures, datasets, and defense hyper-parameter, which are \textit{not even rooted from delicate adaptive attacks}. We demonstrate how to trivially bypass them aligned with their respective threat models by simply varying aforementioned factors. Particularly, for each defense, formal proofs or empirical studies are used to reveal its two non-robust cases where it is not as robust as it claims or expects, especially the recent MNTD. This work highlights the necessity of thoroughly evaluating the robustness of backdoor countermeasures to avoid their misleading security implications in unknown non-robust cases.

preprint2020arXiv

The Capacity of Private Information Retrieval Under Arbitrary Collusion Patterns

We study the private information retrieval (PIR) problem under arbitrary collusion pattern for replicated databases. We find its capacity, which is the same as the capacity of the original PIR problem with the number of databases $N$ replaced by a number $S^*$. The number $S^*$ is the optimal solution to a linear programming problem that is a function of the collusion pattern. Hence, the collusion pattern affects the capacity of the PIR problem only through the number $S^*$.

preprint2016arXiv

Adaptive Optimal PMU Placement Based on Empirical Observability Gramian

In this paper, we compare four measures of the empirical observability gramian, including the determinant, the trace, the minimum eigenvalue, and the condition number, which can be used to quantify the observability of system states and to obtain the optimal PMU placement for power system dynamic state estimation. An adaptive optimal PMU placement method is proposed by automatically choosing proper measures as the objective function. It is shown that when the number of PMUs is small and thus the observability is very weak, the minimum eigenvalue and the condition number are better measures of the observability and are preferred to be chosen as the objective function. The effectiveness of the proposed method is validated by performing dynamic state estimation on an Northeast Power Coordinating Council (NPCC) 48-machine 140-bus system with the square-root unscented Kalman filter.

preprint2016arXiv

An Upper Bound on the Sum Capacity of the Downlink Multicell Processing with Finite Backhaul Capacity

In this paper, we study upper bounds on the sum capacity of the downlink multicell processing model with finite backhaul capacity for the simple case of 2 base stations and 2 mobile users. It is modelled as a two-user multiple access diamond channel. It consists of a first hop from the central processor to the base stations via orthogonal links of finite capacity, and the second hop from the base stations to the mobile users via a Gaussian interference channel. The converse is derived using the converse tools of the multiple access diamond channel and that of the Gaussian MIMO broadcast channel. Through numerical results, it is shown that our upper bound improves upon the existing upper bound greatly in the medium backhaul capacity range, and as a result, the gap between the upper bounds and the sum rate of the time-sharing of the known achievable schemes is significantly reduced.

preprint2016arXiv

Extended First-Principles Molecular Dynamics Method From Cold Materials to Hot Dense Plasmas

An extended first-principles molecular dynamics (FPMD) method based on Kohn-Sham scheme is proposed to elevate the temperature limit of the FPMD method in the calculation of dense plasmas. The extended method treats the wave functions of high energy electrons as plane waves analytically, and thus expands the application of the FPMD method to the region of hot dense plasmas without suffering from the formidable computational costs. In addition, the extended method inherits the high accuracy of the Kohn-Sham scheme and keeps the information of elec- tronic structures. This gives an edge to the extended method in the calculation of the lowering of ionization potential, X-ray absorption/emission spectra, opacity, and high-Z dense plasmas, which are of particular interest to astrophysics, inertial confinement fusion engineering, and laboratory astrophysics.

preprint2016arXiv

Mitigating the Curse of Dimensionality: Sparse Grid Characteristics Method for Optimal Feedback Control and HJB Equations

We address finding the semi-global solutions to optimal feedback control and the Hamilton--Jacobi--Bellman (HJB) equation. Using the solution of an HJB equation, a feedback optimal control law can be implemented in real-time with minimum computational load. However, except for systems with two or three state variables, using traditional techniques for numerically finding a semi-global solution to an HJB equation for general nonlinear systems is infeasible due to the curse of dimensionality. Here we present a new computational method for finding feedback optimal control and solving HJB equations which is able to mitigate the curse of dimensionality. We do not discretize the HJB equation directly, instead we introduce a sparse grid in the state space and use the Pontryagin's maximum principle to derive a set of necessary conditions in the form of a boundary value problem, also known as the characteristic equations, for each grid point. Using this approach, the method is spatially causality free, which enjoys the advantage of perfect parallelism on a sparse grid. Compared with dense grids, a sparse grid has a significantly reduced size which is feasible for systems with relatively high dimensions, such as the $6$-D system shown in the examples. Once the solution obtained at each grid point, high-order accurate polynomial interpolation is used to approximate the feedback control at arbitrary points. We prove an upper bound for the approximation error and approximate it numerically. This sparse grid characteristics method is demonstrated with two examples of rigid body attitude control using momentum wheels.

preprint2016arXiv

Nonlinear Modal Decoupling of Multi-Oscillator Systems with Applications to Power Systems

Many natural and manmade dynamical systems that are modeled as large nonlinear multi-oscillator systems like power systems are hard to analyze. For such a system, we propose a nonlinear modal decoupling (NMD) approach inversely constructing as many decoupled nonlinear oscillators as the system oscillation modes so that individual decoupled oscillators can easily be analyzed to infer dynamics and stability of the original system. The NMD follows a similar idea to the normal form except that we eliminate inter-modal terms but allow intra-modal terms of desired nonlinearities in decoupled systems, so decoupled systems can flexibly be shaped into desired forms of nonlinear oscillators. The NMD is then applied to power systems towards two types of nonlinear oscillators, i.e. the single-machine-infinite-bus (SMIB) systems and a proposed non-SMIB oscillator. Numerical studies on a 3-machine 9-bus system and New England 10-machine 39-bus system show that (i) decoupled oscillators keep a majority of the original system modal nonlinearities and the NMD provides a bigger validity region than the normal form, and (ii) decoupled non-SMIB oscillators may keep more authentic dynamics of the original system than decoupled SMIB systems.

preprint2016arXiv

Optimal Placement of Dynamic Var Sources by Using Empirical Controllability Covariance

In this paper, the empirical controllability covariance (ECC), which is calculated around the considered operating condition of a power system, is applied to quantify the degree of controllability of system voltages under specific dynamic var source locations. An optimal dynamic var source placement method addressing fault-induced delayed voltage recovery (FIDVR) issues is further formulated as an optimization problem that maximizes the determinant of ECC. The optimization problem is effectively solved by the NOMAD solver, which implements the Mesh Adaptive Direct Search algorithm. The proposed method is tested on an NPCC 140-bus system and the results show that the proposed method with fault specified ECC can solve the FIDVR issue caused by the most severe contingency with fewer dynamic var sources than the Voltage Sensitivity Index (VSI) based method. The proposed method with fault unspecified ECC does not depend on the settings of the contingency and can address more FIDVR issues than VSI method when placing the same number of SVCs under different fault durations. It is also shown that the proposed method can help mitigate voltage collapse.

preprint2016arXiv

Robust Ferroelectricity in Monolayer Group-IV Monochalcogenides

Ferroelectricity usually fades away when materials are thinned down below a critical value. Employing the first-principles density functional theory and modern theory of polarization, we show that the unique ionic-potential anharmonicity can induce spontaneous in-plane electrical polarizations and ferroelectricity in monolayer group-IV monochalcogenides MX (M=Ge, Sn; X=S, Se). Using Monte Carlo simulations with an effective Hamiltonian extracted from the parameterized energy space, we show these materials exhibit a two-dimensional ferroelectric phase transition that is described by fourth-order Landau theory. We also show the ferroelectricity in these materials is robust and the corresponding Curie temperature is higher than room temperature, making these materials promising for realizing ultra-thin ferroelectric devices of broad interest.

preprint2015arXiv

First-Principles Calculation of Principal Hugoniot and K-Shell X-ray Absorption Spectra for Warm Dense KCl

Principal Hugoniot and K-shell X-ray absorption spectra of warm dense KCl are calculated using the first-principles molecular dynamics method. Evolution of electronic structures as well as the influence of the approximate description of ionization on pressure (caused by the underestimation of the energy gap between conduction bands and valence bands) in the first-principles method are illustrated by the calculation. Pressure ionization and thermal smearing are shown as the major factors to prevent the deviation of pressure from global accumulation along the Hugoniot. In addition, cancellation between electronic kinetic pressure and virial pressure further reduces the deviation. The calculation of X-ray absorption spectra shows that the band gap of KCl persists after the pressure ionization of the $3p$ electrons of Cl and K taking place at lower energy, which provides a detailed understanding to the evolution of electronic structures of warm dense matter.

preprint2015arXiv

Link between K-absorption edges and thermodynamic properties of warm-dense plasmas established by improved first-principles method

A precise calculation that translates shifts of X-ray K-absorption edges to variations of thermodynamic properties allows quantitative characterization of interior thermodynamic properties of warm dense plasmas by X-ray absorption techniques, which provides essential information for inertial confinement fusion and other astrophysical applications. We show that this interpretation can be achieved through an improved first-principles method. Our calculation shows that the shift of K-edges exhibits selective sensitivity to thermal parameters and thus would be a suitable temperature index to warm dense plasmas. We also show with a simple model that the shift of K-edges can be used to detect inhomogeneity inside warm dense plasmas when combined with other experimental tools.

preprint2015arXiv

The Gaussian Multiple Access Diamond Channel

In this paper, we study the capacity of the diamond channel. We focus on the special case where the channel between the source node and the two relay nodes are two separate links with finite capacities and the link from the two relay nodes to the destination node is a Gaussian multiple access channel. We call this model the Gaussian multiple access diamond channel. We first propose an upper bound on the capacity. This upper bound is a single-letterization of an $n$-letter upper bound proposed by Traskov and Kramer, and is tighter than the cut-set bound. As for the lower bound, we propose an achievability scheme based on sending correlated codes through the multiple access channel with superposition structure. We then specialize this achievable rate to the Gaussian multiple access diamond channel. Noting the similarity between the upper and lower bounds, we provide sufficient and necessary conditions that a Gaussian multiple access diamond channel has to satisfy such that the proposed upper and lower bounds meet. Thus, for a Gaussian multiple access diamond channel that satisfies these conditions, we have found its capacity.

preprint2014arXiv

Deception with Side Information in Biometric Authentication Systems

In this paper, we study the probability of successful deception of an uncompressed biometric authentication system with side information at the adversary. It represents the scenario where the adversary may have correlated side information, e.g.,~a partial finger print or a DNA sequence of a relative of the legitimate user. We find the optimal exponent of the deception probability by proving both the achievability and the converse. Our proofs are based on the connection between the problem of deception with side information and the rate distortion problem with side information at both the encoder and decoder.

preprint2014arXiv

Manipulation of electronic and magnetic properties of M$_2$C (M=Hf, Nb, Sc, Ta, Ti, V, Zr) monolayer by applying mechanical strains

Tuning the electronic and magnetic properties of a material through strain engineering is an effective strategy to enhance the performance of electronic and spintronic devices. Recently synthesized two-dimensional transition metal carbides M$_2$C (M=Hf, Nb, Sc, Ta, Ti, V, Zr), known as MXenes, has aroused increasingly attentions in nanoelectronic technology due to their unusual properties. In this paper, first-principles calculations based on density functional theory are carried out to investigate the electronic and magnetic properties of M$_2$C subjected to biaxial symmetric mechanical strains. At the strain-free state, all these MXenes exhibit no spontaneous magnetism except for Ti$_2$C and Zr$_2$C which show a magnetic moment of 1.92 and 1.25 $μ_B$/unit, respectively. As the tensile strain increases, the magnetic moments of MXenes are greatly enhanced and a transition from nonmagnetism to ferromagnetism is observed for those nonmagnetic MXenes at zero strains. The most distinct transition is found in Hf$_2$C, in which the magnetic moment is elevated to 1.5 $μ_B$/unit at a strain of 15%. We further show that the magnetic properties of Hf$_2$C are attributed to the band shift mainly composed of Hf(5$d$) states. This strain-tunable magnetism can be utilized to design future spintronics based on MXenes.

preprint2014arXiv

Optimal PMU Placement for Power System Dynamic State Estimation by Using Empirical Observability Gramian

In this paper the empirical observability Gramian calculated around the operating region of a power system is used to quantify the degree of observability of the system states under specific phasor measurement unit (PMU) placement. An optimal PMU placement method for power system dynamic state estimation is further formulated as an optimization problem which maximizes the determinant of the empirical observability Gramian and is efficiently solved by the NOMAD solver, which implements the Mesh Adaptive Direct Search (MADS) algorithm. The implementation, validation, and also the robustness to load fluctuations and contingencies of the proposed method are carefully discussed. The proposed method is tested on WSCC 3-machine 9-bus system and NPCC 48-machine 140-bus system by performing dynamic state estimation with square-root unscented Kalman filter. The simulation results show that the determined optimal PMU placements by the proposed method can guarantee good observability of the system states, which further leads to smaller estimation errors and larger number of convergent states for dynamic state estimation compared with random PMU placements. Under optimal PMU placements an obvious observability transition can be observed. The proposed method is also validated to be very robust to both load fluctuations and contingencies.

preprint2014arXiv

Partial Observability and its Consistency for PDEs

In this paper, a quantitative measure of partial observability is defined for PDEs. The quantity is proved to be consistent if the PDE is approximated using well-posed approximation schemes. A first order approximation of an unobservability index using an empirical Gramian is introduced. Several examples are presented to illustrate the concept of partial observability, including Burgers' equation and a one-dimensional nonlinear shallow water equation.

preprint2014arXiv

The potential applications of phosphorene as anode materials in Li-ion batteries

The capacity and stability of constituent electrodes determine the performance of Li-ion batteries. In this study, density functional theory is employed to explore the potential application of recently synthesized two dimensional phosphorene as electrode materials. Our results show that Li atoms can bind strongly with phosphorene monolayer and double layer with significant electron transfer. Besides, the structure of phosphorene is not much influenced by lithiation and the volume change is only 0.2\%. A semiconducting to metallic transition is observed after lithiation. The diffusion barrier is calculated to 0.76 and 0.72 eV on monolayer and double layer phosphorene. The theoretical specific capacity of phosphorene monolayer is 432.79 mAh/g, which is larger than other commercial anodes materials. Our findings show that the high capacity, low open circuit voltage, small volume change and electrical conductivity of phosphorene make it a good candidate as electrode material.

preprint2013arXiv

Antifferomagnetic FeSe monolayer on SiTiO$_{3}$: The charge doping and electric field effects

We present theoretically the electronic structure of antiferromagnetic (AFM) FeSe monolayer on TiO$_{2}$ terminated SrTiO$_{3}$(001) surface. It is revealed that the striking disappearance of the Fermi surface around the Brillouin zone (BZ) center can be well explained by the antiferromatnetic (AFM) phase. We show that the system has a considerable charge transfer from SrTiO$_{3}$(001) substrate to FeSe monolayer, and so has a self-constructed electric field. The FeSe monolayer band structure near the BZ center is sensitive to charge doping, and the spin-resolved energy bands at BZ corner are distorted to be flattened by the perpendicular electric field. We propose a tight-binding model Hamiltonian to take these key factors into account. We also show that this composite structure is an ideal electron-hole bilayer system, with electrons and holes respectively formed in FeSe monolayer and TiO$_{2}$ surface layer.

preprint2013arXiv

Gas adsorption on MoS2 monolayer from first-principles calculations

First-principles calculations within density functional theory (DFT) have been carried out to investigate the adsorption of various gas molecules including CO, CO2, NH3, NO and NO2 on MoS2 monolayer in order to fully exploit the gas sensing capabilities of MoS2. By including van der Waals (vdW) interactions between gas molecules and MoS2, we find that only NO and NO2 can bind strongly to MoS2 sheet with large adsorption energies, which is in line with experimental observations. The charge transfer and the variation of electronic structures are discussed in view of the density of states and molecular orbitals of the gas molecules. Our results thus provide a theoretical basis for the potential applications of MoS2 monolayer in gas sensing and give an explanation for recent experimental findings.

preprint2011arXiv

The Consistency of Partial Observability for PDEs

In this paper, a new definition of observability is introduced for PDEs. It is a quantitative measure of partial observability. The quantity is proved to be consistent if approximated using well posed approximation schemes. A first order approximation of an unobservability index using empirical gramian is introduced. For linear systems with full state observability, the empirical gramian is equivalent to the observability gramian in control theory. The consistency of the defined observability is exemplified using a Burgers' equation.

preprint2010arXiv

Enhanced Static Approximation to the Electron Self-Energy Operator for Efficient Calculation of Quasiparticle Energies

An enhanced static approximation for the electron self energy operator is proposed for efficient calculation of quasiparticle energies. Analysis of the static COHSEX approximation originally proposed by Hedin shows that most of the error derives from the short wavelength contributions of the assumed adiabatic accumulation of the Coulomb-hole. A wavevector dependent correction factor can be incorporated as the basis for a new static approximation. This factor can be approximated by a single scaling function, determined from the homogeneous electron gas model. The local field effect in real materials is captured by a simple ansatz based on symmetry consideration. As inherited from the COHSEX approximation, the new approximation presents a Hermitian self-energy operator and the summation over empty states is eliminated from the evaluation of the self energy operator. Tests were conducted comparing the new approximation to GW calculations for diverse materials ranging from crystals and nanotubes. The accuracy for the minimum gap is about 10% or better. Like in the COHSEX approximation, the occupied bandwidth is overestimated.

preprint2010arXiv

Quasiparticle and Optical Properties of Rutile and Anatase TiO$_{2}$

Quasiparticle excitation energies and optical properties of TiO$_{2}$ in the rutile and anatase structures are calculated using many-body perturbation theory methods. Calculations are performed for a frozen crystal lattice; electron-phonon coupling is not explicitly considered. In the GW method, several approximations are compared and it is found that inclusion of the full frequency dependence as well as explicit treatment of the Ti semicore states are essential for accurate calculation of the quasiparticle energy band gap. The calculated quasiparticle energies are in good agreement with available photoemission and inverse photoemission experiments. The results of the GW calculations, together with the calculated static screened Coulomb interaction, are utilized in the Bethe-Salpeter equation to calculate the dielectric function $ε_{2}(ω)$ for both the rutile and anatase structures. The results are in good agreement with experimental observations, particularly the onset of the main absorption features around 4 eV. For comparison to low temperature optical absorption measurements that resolve individual excitonic transitions in rutile, the low-lying discrete excitonic energy levels are calculated with electronic screening only. The lowest energy exciton found in the energy gap of rutile has a binding energy of 0.13 eV. In agreement with experiment, it is not dipole allowed, but the calculated exciton energy exceeds that measured in absorption experiments by about 0.22 eV and the scale of the exciton binding energy is also too large. The quasiparticle energy alignment of rutile is calculated for non-polar (110) surfaces. In the GW approximation, the valence band maximum is 7.8 eV below the vacuum level, showing a small shift from density functional theory results.