Source author record

Xu He

Xu He appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

22works
16topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

22 published item(s)

preprint2025arXiv

Exploring the Potential of Spiking Neural Networks in UWB Channel Estimation

Although existing deep learning-based Ultra-Wide Band (UWB) channel estimation methods achieve high accuracy, their computational intensity clashes sharply with the resource constraints of low-cost edge devices. Motivated by this, this letter explores the potential of Spiking Neural Networks (SNNs) for this task and develops a fully unsupervised SNN solution. To enable a comprehensive performance analysis, we devise an extensive set of comparative strategies and evaluate them on a compelling public benchmark. Experimental results show that our unsupervised approach still attains 80% test accuracy, on par with several supervised deep learning-based strategies. Moreover, compared with complex deep learning methods, our SNN implementation is inherently suited to neuromorphic deployment and offers a drastic reduction in model complexity, bringing significant advantages for future neuromorphic practice.

preprint2022arXiv

DeepScalper: A Risk-Aware Reinforcement Learning Framework to Capture Fleeting Intraday Trading Opportunities

Reinforcement learning (RL) techniques have shown great success in many challenging quantitative trading tasks, such as portfolio management and algorithmic trading. Especially, intraday trading is one of the most profitable and risky tasks because of the intraday behaviors of the financial market that reflect billions of rapidly fluctuating capitals. However, a vast majority of existing RL methods focus on the relatively low frequency trading scenarios (e.g., day-level) and fail to capture the fleeting intraday investment opportunities due to two major challenges: 1) how to effectively train profitable RL agents for intraday investment decision-making, which involves high-dimensional fine-grained action space; 2) how to learn meaningful multi-modality market representation to understand the intraday behaviors of the financial market at tick-level. Motivated by the efficient workflow of professional human intraday traders, we propose DeepScalper, a deep reinforcement learning framework for intraday trading to tackle the above challenges. Specifically, DeepScalper includes four components: 1) a dueling Q-network with action branching to deal with the large action space of intraday trading for efficient RL optimization; 2) a novel reward function with a hindsight bonus to encourage RL agents making trading decisions with a long-term horizon of the entire trading day; 3) an encoder-decoder architecture to learn multi-modality temporal market embedding, which incorporates both macro-level and micro-level market information; 4) a risk-aware auxiliary task to maintain a striking balance between maximizing profit and minimizing risk. Through extensive experiments on real-world market data spanning over three years on six financial futures, we demonstrate that DeepScalper significantly outperforms many state-of-the-art baselines in terms of four financial criteria.

preprint2022arXiv

Dynamic control of octahedral rotation in perovskites by defect engineering

Engineering oxygen octahedra rotation patterns in $ABO_3$ perovskites is a powerful route to design functional materials. Here we propose a strategy that exploits point defects that create local electric dipoles and couple to the oxygen sublattice, enabling direct actuation on the rotational degrees of freedom. This approach, which relies on substituting an $A$ site with a smaller ion, paves a way to couple dynamically octahedra rotations to external electric fields. A common antisite defect, $\mathrm{Al_{La}}$ in rhombohedral LaAlO$_3$ is taken as a prototype to validate the idea, with atomistic density functional theory calculations supported with an effective lattice model to simulate the dynamics of switching of the local rotational degrees of freedom to long distances. Our simulations provide an insight of the main parameters that govern the operation of the proposed mechanism, and allow to define guidelines for screening other systems where this approach could be used for tuning the properties of the host material.

preprint2022arXiv

Phase Diagram of Infinite-layer Nickelate Compounds from First- and Second-principles Calculations

The fundamental properties of infinite-layer rare-earth nickelates (RNiO2) are carefully revisited and compared with those of CaCuO2 and RNiO3 perovskites. Combining first-principles and finite-temperature second-principles calculations, we highlight that bulk NdNiO2 compound are far from equivalent to CaCuO2, together at the structural, electronic, and magnetic levels. Structurally, it is shown to be prone to spin-phonon coupling induced oxygen square rotation motion, which might be responsible for the intriguing upturn of the resistivity. At the electronic and magnetic levels, we point out orbital-selective Mott localization with strong out-of-plane band dispersion, which should result in the isotropic upper critical fields and weakly three-dimensional magnetic interactions with in-plane local moment and out-of-plane itinerant moment. We further demonstrate that as in RNiO3 perovskites, oxygen rotation motion and rare-earth ion controlled electronic and magnetic properties can give rise in RNiO2 compounds to a rich phase diagram and high tunability of various appealing properties. In line with that, we reveal that key ingredients of high-Tc superconductor such as orbital polarization, Fermi surface, and antiferromagnetic interactions can be deliberately controlled in NdNiO2 through epitaxial strain. Exploiting strain-orbital engineering, a crossover from three- to two-dimensional magnetic transition can be established, making then NdNiO2 thin film a true analog of high-Tc cuprates.

preprint2020arXiv

Contextual User Browsing Bandits for Large-Scale Online Mobile Recommendation

Online recommendation services recommend multiple commodities to users. Nowadays, a considerable proportion of users visit e-commerce platforms by mobile devices. Due to the limited screen size of mobile devices, positions of items have a significant influence on clicks: 1) Higher positions lead to more clicks for one commodity. 2) The 'pseudo-exposure' issue: Only a few recommended items are shown at first glance and users need to slide the screen to browse other items. Therefore, some recommended items ranked behind are not viewed by users and it is not proper to treat this kind of items as negative samples. While many works model the online recommendation as contextual bandit problems, they rarely take the influence of positions into consideration and thus the estimation of the reward function may be biased. In this paper, we aim at addressing these two issues to improve the performance of online mobile recommendation. Our contributions are four-fold. First, since we concern the reward of a set of recommended items, we model the online recommendation as a contextual combinatorial bandit problem and define the reward of a recommended set. Second, we propose a novel contextual combinatorial bandit method called UBM-LinUCB to address two issues related to positions by adopting the User Browsing Model (UBM), a click model for web search. Third, we provide a formal regret analysis and prove that our algorithm achieves sublinear regret independent of the number of items. Finally, we evaluate our algorithm on two real-world datasets by a novel unbiased estimator. An online experiment is also implemented in Taobao, one of the most popular e-commerce platforms in the world. Results on two CTR metrics show that our algorithm outperforms the other contextual bandit algorithms.

preprint2020arXiv

Continual Learning from the Perspective of Compression

Connectionist models such as neural networks suffer from catastrophic forgetting. In this work, we study this problem from the perspective of information theory and define forgetting as the increase of description lengths of previous data when they are compressed with a sequentially learned model. In addition, we show that continual learning approaches based on variational posterior approximation and generative replay can be considered as approximations to two prequential coding methods in compression, namely, the Bayesian mixture code and maximum likelihood (ML) plug-in code. We compare these approaches in terms of both compression and forgetting and empirically study the reasons that limit the performance of continual learning methods based on variational posterior approximation. To address these limitations, we propose a new continual learning method that combines ML plug-in and Bayesian mixture codes.

preprint2020arXiv

Lattice-based designs possessing quasi-optimal separation distance on all projections

Experimental designs that spread out points apart from each other on projections are important for computer experiments when not necessarily all factors have substantial influence on the response. We provide a theoretical framework to generate designs that possess quasi-optimal separation distance on all of the projections and quasi-optimal fill distance on univariate margins. The key is to use special techniques to rotate certain lattices. One such type of design is densest packing-based maximum projection designs, which outperform existing types of space-filling designs in many scenarios. Computer code to generate these designs is provided in R package LatticeDesign.

preprint2020arXiv

Learning Behaviors with Uncertain Human Feedback

Human feedback is widely used to train agents in many domains. However, previous works rarely consider the uncertainty when humans provide feedback, especially in cases that the optimal actions are not obvious to the trainers. For example, the reward of a sub-optimal action can be stochastic and sometimes exceeds that of the optimal action, which is common in games or real-world. Trainers are likely to provide positive feedback to sub-optimal actions, negative feedback to the optimal actions and even do not provide feedback in some confusing situations. Existing works, which utilize the Expectation Maximization (EM) algorithm and treat the feedback model as hidden parameters, do not consider uncertainties in the learning environment and human feedback. To address this challenge, we introduce a novel feedback model that considers the uncertainty of human feedback. However, this incurs intractable calculus in the EM algorithm. To this end, we propose a novel approximate EM algorithm, in which we approximate the expectation step with the Gradient Descent method. Experimental results in both synthetic scenarios and two real-world scenarios with human participants demonstrate the superior performance of our proposed approach.

preprint2020arXiv

Learning Efficient Multi-agent Communication: An Information Bottleneck Approach

We consider the problem of the limited-bandwidth communication for multi-agent reinforcement learning, where agents cooperate with the assistance of a communication protocol and a scheduler. The protocol and scheduler jointly determine which agent is communicating what message and to whom. Under the limited bandwidth constraint, a communication protocol is required to generate informative messages. Meanwhile, an unnecessary communication connection should not be established because it occupies limited resources in vain. In this paper, we develop an Informative Multi-Agent Communication (IMAC) method to learn efficient communication protocols as well as scheduling. First, from the perspective of communication theory, we prove that the limited bandwidth constraint requires low-entropy messages throughout the transmission. Then inspired by the information bottleneck principle, we learn a valuable and compact communication protocol and a weight-based scheduler. To demonstrate the efficiency of our method, we conduct extensive experiments in various cooperative and competitive multi-agent tasks with different numbers of agents and different bandwidths. We show that IMAC converges faster and leads to efficient communication among agents under the limited bandwidth as compared to many baseline methods.

preprint2020arXiv

Learning from Naturalistic Driving Data for Human-like Autonomous Highway Driving

Driving in a human-like manner is important for an autonomous vehicle to be a smart and predictable traffic participant. To achieve this goal, parameters of the motion planning module should be carefully tuned, which needs great effort and expert knowledge. In this study, a method of learning cost parameters of a motion planner from naturalistic driving data is proposed. The learning is achieved by encouraging the selected trajectory to approximate the human driving trajectory under the same traffic situation. The employed motion planner follows a widely accepted methodology that first samples candidate trajectories in the trajectory space, then select the one with minimal cost as the planned trajectory. Moreover, in addition to traditional factors such as comfort, efficiency and safety, the cost function is proposed to incorporate incentive of behavior decision like a human driver, so that both lane change decision and motion planning are coupled into one framework. Two types of lane incentive cost -- heuristic and learning based -- are proposed and implemented. To verify the validity of the proposed method, a data set is developed by using the naturalistic trajectory data of human drivers collected on the motorways in Beijing, containing samples of lane changes to the left and right lanes, and car followings. Experiments are conducted with respect to both lane change decision and motion planning, and promising results are achieved.

preprint2020arXiv

Learning to Collaborate in Multi-Module Recommendation via Multi-Agent Reinforcement Learning without Communication

With the rise of online e-commerce platforms, more and more customers prefer to shop online. To sell more products, online platforms introduce various modules to recommend items with different properties such as huge discounts. A web page often consists of different independent modules. The ranking policies of these modules are decided by different teams and optimized individually without cooperation, which might result in competition between modules. Thus, the global policy of the whole page could be sub-optimal. In this paper, we propose a novel multi-agent cooperative reinforcement learning approach with the restriction that different modules cannot communicate. Our contributions are three-fold. Firstly, inspired by a solution concept in game theory named correlated equilibrium, we design a signal network to promote cooperation of all modules by generating signals (vectors) for different modules. Secondly, an entropy-regularized version of the signal network is proposed to coordinate agents' exploration of the optimal global policy. Furthermore, experiments based on real-world e-commerce data demonstrate that our algorithm obtains superior performance over baselines.

preprint2020arXiv

When the Differences in Frequency Domain are Compensated: Understanding and Defeating Modulated Replay Attacks on Automatic Speech Recognition

Automatic speech recognition (ASR) systems have been widely deployed in modern smart devices to provide convenient and diverse voice-controlled services. Since ASR systems are vulnerable to audio replay attacks that can spoof and mislead ASR systems, a number of defense systems have been proposed to identify replayed audio signals based on the speakers' unique acoustic features in the frequency domain. In this paper, we uncover a new type of replay attack called modulated replay attack, which can bypass the existing frequency domain based defense systems. The basic idea is to compensate for the frequency distortion of a given electronic speaker using an inverse filter that is customized to the speaker's transform characteristics. Our experiments on real smart devices confirm the modulated replay attacks can successfully escape the existing detection mechanisms that rely on identifying suspicious features in the frequency domain. To defeat modulated replay attacks, we design and implement a countermeasure named DualGuard. We discover and formally prove that no matter how the replay audio signals could be modulated, the replay attacks will either leave ringing artifacts in the time domain or cause spectrum distortion in the frequency domain. Therefore, by jointly checking suspicious features in both frequency and time domains, DualGuard can successfully detect various replay attacks including the modulated replay attacks. We implement a prototype of DualGuard on a popular voice interactive platform, ReSpeaker Core v2. The experimental results show DualGuard can achieve 98% accuracy on detecting modulated replay attacks.

preprint2016arXiv

Engineering charge ordering into multiferroicity

Multiferroic materials have attracted great interests but are rare in nature. In many transitional metal oxides, charge ordering and magnetic ordering coexist, so that a method of engineering charge-ordered materials into ferroelectric materials would lead to a large class of multiferroic materials. We propose a strategy for designing new ferroelectric or even multiferroic materials by inserting a spacing layer into each two layers of charge-ordered materials and artificially making a superlattice. One example of the model demonstrated here is the perovskite (LaFeO$_3$)$_2$/LaTiO$_3$ (111) superlattice, in which the LaTiO$_3$ layer acts as the donor and the spacing layer, and the LaFeO$_3$ layer is half doped and performs charge ordering. The collaboration of the charge ordering and the spacing layer breaks the space inversion symmetry, resulting in a large ferroelectric polarization. As the charge ordering also leads to a ferrimagnetic structure, the (LaFeO$_3$)$_2$/LaTiO$_3$ is multiferroic. It is expected that this work can encourage the designing and experimentally implementation of a large class of multiferroic structures with novel properties.

preprint2016arXiv

Evolution of the electronic and lattice structure with carrier injection in BiFeO$_3$

We report a density functional study on the evolution of the electronic and lattice structure in BiFeO$_3$ with injected electrons and holes. First, the self-trapping of electrons and holes were investigated. We found that the injected electrons tend to be localized on Fe sites due to the local lattice expansion, the on-site Coulomb interaction of Fe $3d$ electrons, and the antiferromagnetic order in BiFeO$_3$. The injected holes tend to be delocalized if the on-site Coulomb interaction of O $2p$ is weak (in other words, $U_\mathrm{O}$ is small). Single center polarons and multi-center polarons are formed with large and intermediate $U_\mathrm{O}$, respectively. With intermediate $U_\mathrm{O}$, multi-center polarons can be formed. We also studied the lattice distortion with the injection of carriers by assuming the delocalization of these carriers. We found that the ferroelectric off-centering of BiFeO$_3$ increases with the concentration of the electrons injected and decreases with that of the holes injected. It was also found that a structural phase transition from $R3c$ to the non-ferroelectric $Pbnm$ occurs, with the hole concentration over 8.7$\times10^{19} cm^{-3}$. The change of the off-centering is mainly due to the change of the lattice volume. The understanding of the carrier localization mechanism can help to optimize the functionality of ferroelectric diodes and the ferroelectric photovoltage devices, while the understanding of the evolution of the lattice with carriers can help tuning the ferroelectric properties by the carriers in BiFeO$_3$.

preprint2016arXiv

Insulating phase at low temperature in ultrathin La0.8Sr0.2MnO3 films

Metal-insulator transition is observed in the La0.8Sr0.2MnO3 thin films with thickness larger than 5 unit cells. Insulating phase at lower temperature appeared in the ultrathin films with thickness ranging from 6 unit cells to 10 unit cells and it is found that the Mott variable range hopping conduction dominates in this insulating phase at low temperature with a decrease of localization length in thinner films. A deficiency of oxygen content and a resulted decrease of the Mn valence have been observed in the ultrathin films with thickness smaller than or equal to 10 unit cells by studying the aberration-corrected scanning transmission electron microscopy and electron energy loss spectroscopy of the films. These results suggest that the existence of the oxygen vacancies in thinner films suppresses the double-exchange mechanism and contributes to the enhancement of disorder, leading to a decrease of the Curie temperature and the low temperature insulating phase in the ultrathin films. In addition, the suppression of the magnetic properties in thinner films indicates stronger disorder of magnetic moments, which is considered to be the reason for this decrease of the localization length.

preprint2016arXiv

Oxygen vacancy induced room temperature metal-insulator transition in nickelates films and its potential application in photovoltaics

Oxygen vacancy is intrinsically coupled with magnetic, electronic and transport properties of transition-metal oxide materials and directly determines their multifunctionality. Here, we demonstrate reversible control of oxygen content by post-annealing at temperature lower than 300 degree centigrade and realize the reversible metal-insulator transition in epitaxial NdNiO3 films. Importantly, over six orders of magnitude in the resistance modulation and a large change in optical band gap are demonstrated at room temperature without destroying the parent framework and changing the p-type conductive mechanism. Further study revealed that oxygen vacancies stabilized the insulating phase at room temperature is universal for perovskite nickelates films. Acting as electron donors, oxygen vacancies not only stabilize the insulating phase at room temperature, but also induce a large magnetization of ~50 emu/cm3 due to the formation of strongly correlated Ni2+ t2g6eg2 states. The band gap opening is an order of magnitude larger than that of the thermally driven metal-insulator transition and continuously tunable. Potential application of the newly found insulating phase in photovoltaics has been demonstrated in the nickelates-based heterojunctions. Our discovery opens up new possibilities for strongly correlated perovskite nickelates.

preprint2016arXiv

Persisting of Polar Distortion with Electron Doping in Lone-Pair Driven Ferroelectrics

Free electrons can screen out long-range Coulomb interaction and destroy the polar distortion in some ferroelectric materials, whereas the coexistence of polar distortion and metallicity were found in several non-central-symmetric metals (NCSMs). Therefore, the mechanisms and designing of NCSMs have attracted great interests. In this work, by first-principles calculation, we found the polar distortion in the lone-pair driven ferroelectric material PbTiO$_3$ can not only persist, but also increase with electron doping. We further analyzed the mechanisms of the persisting of the polar distortion. We found that the Ti site polar instability is suppressed but the Pb site polar instability is intact with the electron doping. The Pb-site instability is due to the lone-pair mechanism which can be viewed as a pseudo-Jahn-Teller effect, a mix of the ground state and the excited state by ion displacement from the central symmetric position. The lone-pair mechanism is not strongly affected by the electron doping because neither the ground state or the excited state involved is at the Fermi energy. The enhancement of the polar distortion is related to the increasing of the Ti ion size by doping. These results show the lone-pair stereoactive ions can be used in designing NCSMs.

preprint2016arXiv

Rotated sphere packing designs

We propose a new class of space-filling designs called rotated sphere packing designs for computer experiments. The approach starts from the asymptotically optimal positioning of identical balls that covers the unit cube. Properly scaled, rotated, translated and extracted, such designs are excellent in maximin distance criterion, low in discrepancy, good in projective uniformity and thus useful in both prediction and numerical integration purposes. We provide a fast algorithm to construct such designs for any numbers of dimensions and points with R codes available online. Theoretical and numerical results are also provided.

preprint2015arXiv

Ferroelectric Control of Metal-Insulator Transition

We propose a method of controlling the metal-insulator transition of one perovskite material at its interface with a another ferroelectric material based on first principle calculations. The operating principle is that the rotation of oxygen octahedra tuned by the ferroelectric polarization can modulate the superexchange interaction in this perovskite. We designed a tri-color superlattice of (BiFeO$_3$)$_N$/LaNiO$_3$/LaTiO$_3$, in which the BiFeO$_3$ layers are ferroelectric, the LaNiO$_3$ layer is the layer of which the electronic structure is to be tuned, and LaTiO$_3$ layer is inserted to enhance the inversion asymmetry. By reversing the ferroelectric polarization in this structure, there is a metal-insulator transition of the LaNiO$_3$ layer because of the changes of crystal field splitting of the Ni $e_g$ orbitals and the bandwidth of the Ni in-plane $e_g$ orbital. It is highly expected that a metal-transition can be realized by designing the structures at the interfaces for more materials.

preprint2014arXiv

A central limit theorem for general orthogonal array based space-filling designs

Orthogonal array based space-filling designs (Owen [Statist. Sinica 2 (1992a) 439-452]; Tang [J. Amer. Statist. Assoc. 88 (1993) 1392-1397]) have become popular in computer experiments, numerical integration, stochastic optimization and uncertainty quantification. As improvements of ordinary Latin hypercube designs, these designs achieve stratification in multi-dimensions. If the underlying orthogonal array has strength $t$, such designs achieve uniformity up to $t$ dimensions. Existing central limit theorems are limited to these designs with only two-dimensional stratification based on strength two orthogonal arrays. We develop a new central limit theorem for these designs that possess stratification in arbitrary multi-dimensions associated with orthogonal arrays of general strength. This result is useful for building confidence statements for such designs in various statistical applications.

preprint2014arXiv

A regression tree approach to identifying subgroups with differential treatment effects

In the fight against hard-to-treat diseases such as cancer, it is often difficult to discover new treatments that benefit all subjects. For regulatory agency approval, it is more practical to identify subgroups of subjects for whom the treatment has an enhanced effect. Regression trees are natural for this task because they partition the data space. We briefly review existing regression tree algorithms. Then we introduce three new ones that are practically free of selection bias and are applicable to two or more treatments, censored response variables, and missing values in the predictor variables. The algorithms extend the GUIDE approach by using three key ideas: (i) treatment as a linear predictor, (ii) chi-squared tests to detect residual patterns and lack of fit, and (iii) proportional hazards modeling via Poisson regression. Importance scores with thresholds for identifying influential variables are obtained as by-products. A bootstrap technique is used to construct confidence intervals for the treatment effects in each node. Real and simulated data are used to compare the methods.

preprint2012arXiv

A Comparative Study of Discretization Approaches for Granular Association Rule Mining

Granular association rule mining is a new relational data mining approach to reveal patterns hidden in multiple tables. The current research of granular association rule mining considers only nominal data. In this paper, we study the impact of discretization approaches on mining semantically richer and stronger rules from numeric data. Specifically, the Equal Width approach and the Equal Frequency approach are adopted and compared. The setting of interval numbers is a key issue in discretization approaches, so we compare different settings through experiments on a well-known real life data set. Experimental results show that: 1) discretization is an effective preprocessing technique in mining stronger rules; 2) the Equal Frequency approach helps generating more rules than the Equal Width approach; 3) with certain settings of interval numbers, we can obtain much more rules than others.