Source author record

Hao Yuan

Hao Yuan appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

16works
12topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

16 published item(s)

preprint2023arXiv

FlowX: Towards Explainable Graph Neural Networks via Message Flows

We investigate the explainability of graph neural networks (GNNs) as a step toward elucidating their working mechanisms. While most current methods focus on explaining graph nodes, edges, or features, we argue that, as the inherent functional mechanism of GNNs, message flows are more natural for performing explainability. To this end, we propose a novel method here, known as FlowX, to explain GNNs by identifying important message flows. To quantify the importance of flows, we propose to follow the philosophy of Shapley values from cooperative game theory. To tackle the complexity of computing all coalitions' marginal contributions, we propose a flow sampling scheme to compute Shapley value approximations as initial assessments of further training. We then propose an information-controlled learning algorithm to train flow scores toward diverse explanation targets: necessary or sufficient explanations. Experimental studies on both synthetic and real-world datasets demonstrate that our proposed FlowX and its variants lead to improved explainability of GNNs. The code is available at https://github.com/divelab/DIG.

preprint2022arXiv

Explainability in Graph Neural Networks: A Taxonomic Survey

Deep learning methods are achieving ever-increasing performance on many artificial intelligence tasks. A major limitation of deep models is that they are not amenable to interpretability. This limitation can be circumvented by developing post hoc techniques to explain the predictions, giving rise to the area of explainability. Recently, explainability of deep models on images and texts has achieved significant progress. In the area of graph data, graph neural networks (GNNs) and their explainability are experiencing rapid developments. However, there is neither a unified treatment of GNN explainability methods, nor a standard benchmark and testbed for evaluations. In this survey, we provide a unified and taxonomic view of current GNN explainability methods. Our unified and taxonomic treatments of this subject shed lights on the commonalities and differences of existing methods and set the stage for further methodological developments. To facilitate evaluations, we generate a set of benchmark graph datasets specifically for GNN explainability. We summarize current datasets and metrics for evaluating GNN explainability. Altogether, this work provides a unified methodological treatment of GNN explainability and a standardized testbed for evaluations.

preprint2022arXiv

Observation of novel topological states in hyperbolic lattices

The discovery of novel topological states has served as a major branch in physics and material science. However, to date, most of the established topological states of matter have been employed in Euclidean systems, where the interplay between unique geometrical characteristics of curved spaces and exotic topological phases is less explored, especially on the experimental perspective. Recently, the experimental realization of the hyperbolic lattice, which is the regular tessellation in non-Euclidean spaces with a constant negative curvature, has attracted much attention in the field of simulating exotic phenomena from quantum physics in curved spaces to the general relativity. The question is whether there are novel topological states in such a non-Euclidean system without analogues in Euclidean spaces. Here, we demonstrate both in theory and experiment that novel topological states possessing unique properties compared with their Euclidean counterparts can exist in engineered hyperbolic lattices. Specially, based on the extended Haldane model, the boundary-dominated first-order Chern edge state with a nontrivial real-space Chern number is achieved, and the associated one-way propagation is proven. Furthermore, we show that fractal-like midgap higher-order zero modes appear in deformed hyperbolic lattices, where the number of zero modes increases exponentially with the increase of lattice size. These novel topological states are observed in designed hyperbolic circuit networks by measuring site-resolved impendence responses and dynamics of voltage packets. Our findings suggest a novel platform to study topological phases beyond Euclidean space and may have potential applications in the field of designing high-efficient topological devices, such as topological lasers, with extremely fewer trivial regions.

preprint2022arXiv

Restricted mean survival time regression model with time-dependent covariates

In clinical or epidemiological follow-up studies, methods based on time scale indicators such as the restricted mean survival time (RMST) have been developed to some extent. Compared with traditional hazard rate indicator system methods, the RMST is easier to interpret and does not require the proportional hazard assumption. To date, regression models based on the RMST are indirect or direct models of the RMST and baseline covariates. However, time-dependent covariates are becoming increasingly common in follow-up studies. Based on the inverse probability of censoring weighting (IPCW) method, we developed a regression model of the RMST and time-dependent covariates. Through Monte Carlo simulation, we verified the estimation performance of the regression parameters of the proposed model. Compared with the time-dependent Cox model and the fixed (baseline) covariate RMST model, the time-dependent RMST model has a better prediction ability. Finally, an example of heart transplantation was used to verify the above conclusions.

preprint2021arXiv

Node2Seq: Towards Trainable Convolutions in Graph Neural Networks

Investigating graph feature learning becomes essentially important with the emergence of graph data in many real-world applications. Several graph neural network approaches are proposed for node feature learning and they generally follow a neighboring information aggregation scheme to learn node features. While great performance has been achieved, the weights learning for different neighboring nodes is still less explored. In this work, we propose a novel graph network layer, known as Node2Seq, to learn node embeddings with explicitly trainable weights for different neighboring nodes. For a target node, our method sorts its neighboring nodes via attention mechanism and then employs 1D convolutional neural networks (CNNs) to enable explicit weights for information aggregation. In addition, we propose to incorporate non-local information for feature learning in an adaptive manner based on the attention scores. Experimental results demonstrate the effectiveness of our proposed Node2Seq layer and show that the proposed adaptively non-local information learning can improve the performance of feature learning.

preprint2020arXiv

Deep Learning of High-Order Interactions for Protein Interface Prediction

Protein interactions are important in a broad range of biological processes. Traditionally, computational methods have been developed to automatically predict protein interface from hand-crafted features. Recent approaches employ deep neural networks and predict the interaction of each amino acid pair independently. However, these methods do not incorporate the important sequential information from amino acid chains and the high-order pairwise interactions. Intuitively, the prediction of an amino acid pair should depend on both their features and the information of other amino acid pairs. In this work, we propose to formulate the protein interface prediction as a 2D dense prediction problem. In addition, we propose a novel deep model to incorporate the sequential information and high-order pairwise interactions to perform interface predictions. We represent proteins as graphs and employ graph neural networks to learn node features. Then we propose the sequential modeling method to incorporate the sequential information and reorder the feature matrix. Next, we incorporate high-order pairwise interactions to generate a 3D tensor containing different pairwise interactions. Finally, we employ convolutional neural networks to perform 2D dense predictions. Experimental results on multiple benchmarks demonstrate that our proposed method can consistently improve the protein interface prediction performance.

preprint2020arXiv

Experimental demonstration of complementarity relations between quantum steering criteria

The ability that one system immediately affects another one by using local measurements is regarded as quantum steering, which can be detected by various steering criteria. Recently, Mondal et al. [Phys. Rev. A 98, 052330 (2018)] derived the complementarity relations of coherence steering criteria, and revealed that the quantum steering of system can be observed through the average coherence of subsystem. Here, we experimentally verify the complementarity relations between quantum steering criteria by employing two-photon Bell-like states and three Pauli operators. The results demonstrate that if prepared quantum states can violate two setting coherence steering criteria and turn out to be steerable states, then it cannot violate the complementary settings criteria. Three measurement settings inequality, which establish a complementarity relation between these two coherence steering criteria, always holds in experiment. Besides, we experimentally certify that the strengths of coherence steering criteria dependent on the choice of coherence measure. In comparison with two setting coherence steering criteria based on l1 norm of coherence and relative entropy of coherence, our experimental results show that the steering criterion based on skew information of coherence is more stronger in detecting the steerability of quantum states. Thus, our experimental demonstrations can deepen the understanding of the relation between the quantum steering and quantum coherence.

preprint2020arXiv

Large-scale Hybrid Approach for Predicting User Satisfaction with Conversational Agents

Measuring user satisfaction level is a challenging task, and a critical component in developing large-scale conversational agent systems serving the needs of real users. An widely used approach to tackle this is to collect human annotation data and use them for evaluation or modeling. Human annotation based approaches are easier to control, but hard to scale. A novel alternative approach is to collect user's direct feedback via a feedback elicitation system embedded to the conversational agent system, and use the collected user feedback to train a machine-learned model for generalization. User feedback is the best proxy for user satisfaction, but is not available for some ineligible intents and certain situations. Thus, these two types of approaches are complementary to each other. In this work, we tackle the user satisfaction assessment problem with a hybrid approach that fuses explicit user feedback, user satisfaction predictions inferred by two machine-learned models, one trained on user feedback data and the other human annotation data. The hybrid approach is based on a waterfall policy, and the experimental results with Amazon Alexa's large-scale datasets show significant improvements in inferring user satisfaction. A detailed hybrid architecture, an in-depth analysis on user feedback data, and an algorithm that generates data sets to properly simulate the live traffic are presented in this paper.

preprint2020arXiv

XGNN: Towards Model-Level Explanations of Graph Neural Networks

Graphs neural networks (GNNs) learn node features by aggregating and combining neighbor information, which have achieved promising performance on many graph tasks. However, GNNs are mostly treated as black-boxes and lack human intelligible explanations. Thus, they cannot be fully trusted and used in certain application domains if GNN models cannot be explained. In this work, we propose a novel approach, known as XGNN, to interpret GNNs at the model-level. Our approach can provide high-level insights and generic understanding of how GNNs work. In particular, we propose to explain GNNs by training a graph generator so that the generated graph patterns maximize a certain prediction of the model.We formulate the graph generation as a reinforcement learning task, where for each step, the graph generator predicts how to add an edge into the current graph. The graph generator is trained via a policy gradient method based on information from the trained GNNs. In addition, we incorporate several graph rules to encourage the generated graphs to be valid. Experimental results on both synthetic and real-world datasets show that our proposed methods help understand and verify the trained GNNs. Furthermore, our experimental results indicate that the generated graphs can provide guidance on how to improve the trained GNNs.

preprint2019arXiv

Experimental certification of steering criterion based on general entropic uncertainty relation

Quantum steering describes the phenomenon that one system can be immediately influenced by another with local measurements. It can be detected by the violation of a powerful and useful steering criterion from general entropic uncertainty relation. This criterion, in principle, can be evaluated straightforwardly and achieved by only probability distributions from a finite set of measurement settings. Herein, we experimentally verify the steering criterion by means of the two-photon Werner-like states and three Pauli measurements. The results indicate that quantum steering can be verified by the criterion in a convenient way. In particular, it is no need to perform the usual quantum state tomography in experiment, which reduces the required experimental resources greatly. Moreover, we demonstrate that the criterion is stronger than the linear one for the detecting quantum steering of the Werner-like states.

preprint2019arXiv

Experimental investigation of entropic uncertainty relations and coherence uncertainty relations

Uncertainty relation usually is one of the most important features in quantum mechanics, and is the backbone of quantum theory, which distinguishes from the rule in classical counterpart. Specifically, entropy-based uncertainty relations are of fundamental importance in the region of quantum information theory, offering one nontrivial bound of key rate towards quantum key distribution. In this work, we experimentally demonstrate the entropic uncertainty relations and coherence-based uncertainty relations in an all-optics platform. By means of preparing two kinds of bipartite initial states with high fidelity, i.e., Bell-like states and Bell-like diagonal states, we carry on local projective measurements over a complete set of mutually unbiased bases on the measured subsystem. In terms of quantum tomography, the density matrices of the initial states and the post-measurement states are reconstructed. It shows that our experimental results coincide with the theoretical predictions very well. Additionally, we also verify that the lower bounds of both the entropy-based and coherence-based uncertainty can be tightened by imposing the Holevo quantity and mutual information, and the entropic uncertainty is inversely correlated with the coherence. Our demonstrations might offer an insight into their uncertainty relations and their connection to quantum coherence in quantum information science, which might be applicable to the security analysis of quantum key distributions.

preprint2019arXiv

Experimental observation the Einstein-Podolsky-Rosen Steering based on the detection of entanglement

The Einstein-Podolsky-Rosen (EPR) steering is an intermediate quantum nonlocality between entanglement and Bell nonlocality, which plays an important role in quantum information processing tasks. In the past few years, the investigations concerning EPR steering have been demonstrated in a series of experiments. However, these studies rely on the relevant steering inequalities and the choices of measurement settings. Here, we experimentally verify the EPR steering via entanglement detection without using any steering inequality and measurement setting. By constructing two new states from a two-qubit target state, we observe the EPR steering by detecting the entanglement of these new states. The results show that the entanglement of the newly constructed states can be regarded as a new kind of steering witness for target states. Compared to the results of Xiao et al. [Phys. Rev. Lett. 118, 140404 (2017)], we find that the ability of detecting EPR steering in our scenario is stronger than two-setting projective measurements, which can observe more steerable states. Hence, our demonstrations can deepen the understanding of the connection between the EPR steering and entanglement.

preprint2015arXiv

Polarization-controlled anisotropic coding metamaterials at terahertz frequencies

Metamaterials based on effective media have achieved a lot of unusual physics (e.g. negative refraction and invisibility cloaking) owing to their abilities to tailor the effective medium parameters that do not exist in nature. Recently, coding metamaterials have been suggested to control electromagnetic waves by designing the coding sequences of digital elements '0' and '1', which possess opposite phase responses. Here, we propose the concept of anisotropic coding metamaterial at terahertz frequencies, in which coding behaviors in different directions are dependent on the polarization status of terahertz waves. We experimentally demonstrate an ultrathin and flexible polarization-controlled anisotropic coding metasurface functioning in the terahertz regime using specially- designed coding elements. By encoding the elements with elaborately-designed digital sequences (in both 1 bit and 2 bits), the x- and y-polarized reflected waves can be deflected or diffused independently in three dimensions. The simulated far-field scattering patterns as well as near-electric-field distributions are given to illustrate the bifunctional performance of the encoded metasurface, which show good agreement to the measurement results. We further demonstrate the abilities of anisotropic coding metasurface to generate beam splitter and realize anomalous reflection and polarization conversion simultaneously, providing powerful controls of differently-polarized terahertz waves. The proposed method enables versatile beam behaviors under orthogonal polarizations using a single metasurface, and hence will promise interesting terahertz devices.

preprint2013arXiv

On the Complexity of $t$-Closeness Anonymization and Related Problems

An important issue in releasing individual data is to protect the sensitive information from being leaked and maliciously utilized. Famous privacy preserving principles that aim to ensure both data privacy and data integrity, such as $k$-anonymity and $l$-diversity, have been extensively studied both theoretically and empirically. Nonetheless, these widely-adopted principles are still insufficient to prevent attribute disclosure if the attacker has partial knowledge about the overall sensitive data distribution. The $t$-closeness principle has been proposed to fix this, which also has the benefit of supporting numerical sensitive attributes. However, in contrast to $k$-anonymity and $l$-diversity, the theoretical aspect of $t$-closeness has not been well investigated. We initiate the first systematic theoretical study on the $t$-closeness principle under the commonly-used attribute suppression model. We prove that for every constant $t$ such that $0\leq t<1$, it is NP-hard to find an optimal $t$-closeness generalization of a given table. The proof consists of several reductions each of which works for different values of $t$, which together cover the full range. To complement this negative result, we also provide exact and fixed-parameter algorithms. Finally, we answer some open questions regarding the complexity of $k$-anonymity and $l$-diversity left in the literature.

preprint2012arXiv

Supercongruences and Complex Multiplication

We study congruences involving truncated hypergeometric series of the form_rF_{r-1}(1/2,...,1/2;1,...,1;λ)_{(mp^s-1)/2} = \sum_{k=0}^{(mp^s-1)/2} ((1/2)_k/k!)^r λ^k where p is a prime and m, s, r are positive integers. These truncated hypergeometric series are related to the arithmetic of a family of algebraic varieties and exhibit Atkin and Swinnerton-Dyer type congruences. In particular, when r=3, they are related to K3 surfaces. For special values of λ, with s=1 and r=3, our congruences are stronger than what can be predicted by the theory of formal groups because of the presence of elliptic curves with complex multiplications. They generalize a conjecture made by Rodriguez-Villegas for the λ=1 case and confirm some other supercongruence conjectures at special values of λ.

preprint2011arXiv

Testing Bell inequalities with circuit QEDs by joint spectral measurements

We propose a feasible approach to test Bell's inequality with the experimentally-demonstrated circuit QED system, consisting of two well-separated superconducting charge qubits (SCQs) dispersively coupled to a common one-dimensional transmission line resonator (TLR). Our proposal is based on the joint spectral measurements of the two SCQs, i.e., their quantum states in the computational basis $\{|kl>,\,k,l=0,1\}$ can be measured by detecting the transmission spectra of the driven TLR: each peak marks one of the computational basis and its relative height corresponds to the probability superposed. With these joint spectral measurements, the generated Bell states of the two SCQs can be robustly confirmed without the standard tomographic technique. Furthermore, the statistical nonlocal-correlations between these two distant qubits can be directly read out by the joint spectral measurements, and consequently the Bell's inequality can be tested by sequentially measuring the relevant correlations related to the suitably-selected sets of the classical local variables $\{θ_j,θ_j', j=1,2\}$. The experimental challenges of our proposal are also analyzed.