Source author record

Fang Chen

Fang Chen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

28works
20topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

28 published item(s)

preprint2025arXiv

E-LENS: User Requirements-Oriented AI Ethics Assurance

Despite the much proliferation of AI ethical principles in recent years, there is a challenge of assuring AI ethics with current AI ethics frameworks in real-world applications. While system safety has emerged as a distinct discipline for a long time, originated from safety concerns in early aircraft manufacturing. The safety assurance is now an indispensable component in safety critical domains. Motivated by the assurance approaches for safety-critical systems such as aviation, this paper introduces the concept of AI ethics assurance cases into the AI ethics assurance. Three pillars of user requirements, evidence, and validation are proposed as key components and integrated into AI ethics assurance cases for a new approach of user requirements-oriented AI ethics assurance. The user requirements-oriented AI ethics assurance case is set up based on three pillars and hazard analysis methods used in the safety assurance of safety-critical systems. This paper also proposes a platform named Ethical-Lens (E-LENS) to implement the user requirements-oriented AI ethics assurance approach. The proposed user requirements-based E-LENS platform is then applied to assure AI ethics of an AI-driven human resource shortlisting system as a case study to show the effectiveness of the proposed approach.

preprint2023arXiv

Explaining Imitation Learning through Frames

As one of the prevalent methods to achieve automation systems, Imitation Learning (IL) presents a promising performance in a wide range of domains. However, despite the considerable improvement in policy performance, the corresponding research on the explainability of IL models is still limited. Inspired by the recent approaches in explainable artificial intelligence methods, we proposed a model-agnostic explaining framework for IL models called R2RISE. R2RISE aims to explain the overall policy performance with respect to the frames in demonstrations. It iteratively retrains the black-box IL model from the randomized masked demonstrations and uses the conventional evaluation outcome environment returns as the coefficient to build an importance map. We also conducted experiments to investigate three major questions concerning frames' importance equality, the effectiveness of the importance map, and connections between importance maps from different IL models. The result shows that R2RISE successfully distinguishes important frames from the demonstrations.

preprint2023arXiv

Genetic Imitation Learning by Reward Extrapolation

Imitation learning demonstrates remarkable performance in various domains. However, imitation learning is also constrained by many prerequisites. The research community has done intensive research to alleviate these constraints, such as adding the stochastic policy to avoid unseen states, eliminating the need for action labels, and learning from the suboptimal demonstrations. Inspired by the natural reproduction process, we proposed a method called GenIL that integrates the Genetic Algorithm with imitation learning. The involvement of the Genetic Algorithm improves the data efficiency by reproducing trajectories with various returns and assists the model in estimating more accurate and compact reward function parameters. We tested GenIL in both Atari and Mujoco domains, and the result shows that it successfully outperforms the previous extrapolation methods over extrapolation accuracy, robustness, and overall policy performance when input data is limited.

preprint2022arXiv

A Review on Constraint Handling Techniques for Population-based Algorithms: from single-objective to multi-objective optimization

This presented study provides a novel analysis of scholarly literature on constraint handling techniques for single-objective and multi-objective population-based algorithms according to the most relevant journals, keywords, authors, and articles. The paper reviews the main ideas of the most state-of-the-art constraint handling techniques in multi-objective population-based optimization, and then the study addresses the bibliometric analysis in the field. The extracted papers include research articles, reviews, book/book chapters, and conference papers published between 2000 and 2020 for the analysis. The results indicate that the constraint handling techniques for multi-objective optimization have received much less attention compared with single-objective optimization. The most promising algorithms for such optimization were determined to be genetic algorithms, differential evolutionary algorithms, and particle swarm intelligence.

preprint2022arXiv

De-biased Representation Learning for Fairness with Unreliable Labels

Removing bias while keeping all task-relevant information is challenging for fair representation learning methods since they would yield random or degenerate representations w.r.t. labels when the sensitive attributes correlate with labels. Existing works proposed to inject the label information into the learning procedure to overcome such issues. However, the assumption that the observed labels are clean is not always met. In fact, label bias is acknowledged as the primary source inducing discrimination. In other words, the fair pre-processing methods ignore the discrimination encoded in the labels either during the learning procedure or the evaluation stage. This contradiction puts a question mark on the fairness of the learned representations. To circumvent this issue, we explore the following question: \emph{Can we learn fair representations predictable to latent ideal fair labels given only access to unreliable labels?} In this work, we propose a \textbf{D}e-\textbf{B}iased \textbf{R}epresentation Learning for \textbf{F}airness (DBRF) framework which disentangles the sensitive information from non-sensitive attributes whilst keeping the learned representations predictable to ideal fair labels rather than observed biased ones. We formulate the de-biased learning framework through information-theoretic concepts such as mutual information and information bottleneck. The core concept is that DBRF advocates not to use unreliable labels for supervision when sensitive information benefits the prediction of unreliable labels. Experiment results over both synthetic and real-world data demonstrate that DBRF effectively learns de-biased representations towards ideal labels.

preprint2022arXiv

GANExplainer: GAN-based Graph Neural Networks Explainer

With the rapid deployment of graph neural networks (GNNs) based techniques into a wide range of applications such as link prediction, node classification, and graph classification the explainability of GNNs has become an indispensable component for predictive and trustworthy decision-making. Thus, it is critical to explain why graph neural network (GNN) makes particular predictions for them to be believed in many applications. Some GNNs explainers have been proposed recently. However, they lack to generate accurate and real explanations. To mitigate these limitations, we propose GANExplainer, based on Generative Adversarial Network (GAN) architecture. GANExplainer is composed of a generator to create explanations and a discriminator to assist with the Generator development. We investigate the explanation accuracy of our models by comparing the performance of GANExplainer with other state-of-the-art methods. Our empirical results on synthetic datasets indicate that GANExplainer improves explanation accuracy by up to 35\% compared to its alternatives.

preprint2022arXiv

Incident duration prediction using a bi-level machine learning framework with outlier removal and intra-extra joint optimisation

Predicting the duration of traffic incidents is a challenging task due to the stochastic nature of events. The ability to accurately predict how long accidents will last can provide significant benefits to both end-users in their route choice and traffic operation managers in handling of non-recurrent traffic congestion. This paper presents a novel bi-level machine learning framework enhanced with outlier removal and intra-extra joint optimisation for predicting the incident duration on three heterogeneous data sets collected for both arterial roads and motorways from Sydney, Australia and San-Francisco, U.S.A. Firstly, we use incident data logs to develop a binary classification prediction approach, which allows us to classify traffic incidents as short-term or long-term. We find the optimal threshold between short-term versus long-term traffic incident duration, targeting both class balance and prediction performance while also comparing the binary versus multi-class classification approaches. Secondly, for more granularity of the incident duration prediction to the minute level, we propose a new Intra-Extra Joint Optimisation algorithm (IEO-ML) which extends multiple baseline ML models tested against several regression scenarios across the data sets. Final results indicate that: a) 40-45 min is the best split threshold for identifying short versus long-term incidents and that these incidents should be modelled separately, b) our proposed IEO-ML approach significantly outperforms baseline ML models in $66\%$ of all cases showcasing its great potential for accurate incident duration prediction. Lastly, we evaluate the feature importance and show that time, location, incident type, incident reporting source and weather at among the top 10 critical factors which influence how long incidents will last.

preprint2022arXiv

Instance Image Retrieval by Learning Purely From Within the Dataset

Quality feature representation is key to instance image retrieval. To attain it, existing methods usually resort to a deep model pre-trained on benchmark datasets or even fine-tune the model with a task-dependent labelled auxiliary dataset. Although achieving promising results, this approach is restricted by two issues: 1) the domain gap between benchmark datasets and the dataset of a given retrieval task; 2) the required auxiliary dataset cannot be readily obtained. In light of this situation, this work looks into a different approach which has not been well investigated for instance image retrieval previously: {can we learn feature representation \textit{specific to} a given retrieval task in order to achieve excellent retrieval?} Our finding is encouraging. By adding an object proposal generator to generate image regions for self-supervised learning, the investigated approach can successfully learn feature representation specific to a given dataset for retrieval. This representation can be made even more effective by boosting it with image similarity information mined from the dataset. As experimentally validated, such a simple ``self-supervised learning + self-boosting'' approach can well compete with the relevant state-of-the-art retrieval methods. Ablation study is conducted to show the appealing properties of this approach and its limitation on generalisation across datasets.

preprint2022arXiv

Self-Attentive Pooling for Efficient Deep Learning

Efficient custom pooling techniques that can aggressively trim the dimensions of a feature map and thereby reduce inference compute and memory footprint for resource-constrained computer vision applications have recently gained significant traction. However, prior pooling works extract only the local context of the activation maps, limiting their effectiveness. In contrast, we propose a novel non-local self-attentive pooling method that can be used as a drop-in replacement to the standard pooling layers, such as max/average pooling or strided convolution. The proposed self-attention module uses patch embedding, multi-head self-attention, and spatial-channel restoration, followed by sigmoid activation and exponential soft-max. This self-attention mechanism efficiently aggregates dependencies between non-local activation patches during down-sampling. Extensive experiments on standard object classification and detection tasks with various convolutional neural network (CNN) architectures demonstrate the superiority of our proposed mechanism over the state-of-the-art (SOTA) pooling techniques. In particular, we surpass the test accuracy of existing pooling techniques on different variants of MobileNet-V2 on ImageNet by an average of 1.2%. With the aggressive down-sampling of the activation maps in the initial layers (providing up to 22x reduction in memory consumption), our approach achieves 1.43% higher test accuracy compared to SOTA techniques with iso-memory footprints. This enables the deployment of our models in memory-constrained devices, such as micro-controllers (without losing significant accuracy), because the initial activation maps consume a significant amount of on-chip memory for high-resolution images required for complex vision tasks. Our proposed pooling method also leverages the idea of channel pruning to further reduce memory footprints.

preprint2021arXiv

Magnetohydrodynamic shock refraction at an inclined density interface

Shock wave refraction at a sharp density interface is a classical problem in hydrodynamics. Presently, we investigate the strongly planar refraction of a magnetohydrodynamic (MHD) shock wave at an inclined density interface. A magnetic field is applied that is initially oriented either perpendicular and parallel to the motion of incident shock. We explore flow structure by varying the magnitude of the magnetic field governed by the non-dimensional parameter $β\in (0.5, 10^6)$ and the inclination angle of density interface $α\in (0.30, 1.52)$. The regular MHD shock refraction process results in a pair of outer fast shocks (reflected and transmitted) and a set of inner nonlinear magneto-sonic waves. By varying magnetic field (strength and direction) and inclination interface angle, the latter waves can be slow shocks, slow expansion fans, intermediate shocks or slow-mode compound waves. For a chosen incident shock strength, and density ratio, the MHD shock refraction transitions from regular (all nonlinear waves meeting at a single point) into irregular when the inclined density interface angle is less than a critical value. Since the MHD shock refraction is self-similar, we further explore by converting the initial value problem (IVP) into a boundary value problem (BVP) by a self-similar coordinate transformation. The self-similar solution to the BVP is numerically solved using an iterative method and implemented using the p4est adaptive mesh framework. The simulation shows that a Mach stem occurs in irregular MHD shock refraction, the flow structure can be an MHD equivalent to a single Mach reflexion irregular refraction $MRR$ and convex-forwards irregular refraction $CFR$ that occur in the hydrodynamic case.

preprint2021arXiv

Optimal stopping time on discounted semi-Markov processes

This paper attempts to study the optimal stopping time for semi-Markov processes (SMPs) under the discount optimization criteria with unbounded cost rates. In our work, we introduce an explicit construction of the equivalent semi-Markov decision processes (SMDPs). The equivalence is embodied in the value functions of SMPs and SMDPs, that is, every stopping time of SMPs can induce a policy of SMDPs such that the value functions are equal, and vice versa. The existence of the optimal stopping time of SMPs is proved by this equivalence relation. Next, we give the optimality equation of the value function and develop an effective iterative algorithm for computing it. Moreover, we show that the optimal and ε-optimal stopping time can be characterized by the hitting time of the special sets. Finally, to illustrate the validity of our results, an example of a maintenance system is presented in the end.

preprint2020arXiv

An improved online learning algorithm for general fuzzy min-max neural network

This paper proposes an improved version of the current online learning algorithm for a general fuzzy min-max neural network (GFMM) to tackle existing issues concerning expansion and contraction steps as well as the way of dealing with unseen data located on decision boundaries. These drawbacks lower its classification performance, so an improved algorithm is proposed in this study to address the above limitations. The proposed approach does not use the contraction process for overlapping hyperboxes, which is more likely to increase the error rate as shown in the literature. The empirical results indicated the improvement in the classification accuracy and stability of the proposed method compared to the original version and other fuzzy min-max classifiers. In order to reduce the sensitivity to the training samples presentation order of this new on-line learning algorithm, a simple ensemble method is also proposed.

preprint2020arXiv

Detecting Community Depression Dynamics Due to COVID-19 Pandemic in Australia

The recent COVID-19 pandemic has caused unprecedented impact across the globe. We have also witnessed millions of people with increased mental health issues, such as depression, stress, worry, fear, disgust, sadness, and anxiety, which have become one of the major public health concerns during this severe health crisis. For instance, depression is one of the most common mental health issues according to the findings made by the World Health Organisation (WHO). Depression can cause serious emotional, behavioural and physical health problems with significant consequences, both personal and social costs included. This paper studies community depression dynamics due to COVID-19 pandemic through user-generated content on Twitter. A new approach based on multi-modal features from tweets and Term Frequency-Inverse Document Frequency (TF-IDF) is proposed to build depression classification models. Multi-modal features capture depression cues from emotion, topic and domain-specific perspectives. We study the problem using recently scraped tweets from Twitter users emanating from the state of New South Wales in Australia. Our novel classification model is capable of extracting depression polarities which may be affected by COVID-19 and related events during the COVID-19 period. The results found that people became more depressed after the outbreak of COVID-19. The measures implemented by the government such as the state lockdown also increased depression levels. Further analysis in the Local Government Area (LGA) level found that the community depression level was different across different LGAs. Such granular level analysis of depression dynamics not only can help authorities such as governmental departments to take corresponding actions more objectively in specific regions if necessary but also allows users to perceive the dynamics of depression over the time.

preprint2020arXiv

Utilizing machine learning to prevent water main breaks by understanding pipeline failure drivers

Data61 and Western Water worked collaboratively to apply engineering expertise and Machine Learning tools to find a cost-effective solution to the pipe failure problem in the region west of Melbourne, where on average 400 water main failures occur per year. To achieve this objective, we constructed a detailed picture and understanding of the behaviour of the water pipe network by 1) discovering the underlying drivers of water main breaks, and 2) developing a Machine Learning system to assess and predict the failure likelihood of water main breaking using historical failure records, descriptors of pipes, and other environmental factors. The ensuing results open up an avenue for Western Water to identify the priority of pipe renewals

preprint2016arXiv

On Improving Informativity and Grammaticality for Multi-Sentence Compression

Multi Sentence Compression (MSC) is of great value to many real world applications, such as guided microblog summarization, opinion summarization and newswire summarization. Recently, word graph-based approaches have been proposed and become popular in MSC. Their key assumption is that redundancy among a set of related sentences provides a reliable way to generate informative and grammatical sentences. In this paper, we propose an effective approach to enhance the word graph-based MSC and tackle the issue that most of the state-of-the-art MSC approaches are confronted with: i.e., improving both informativity and grammaticality at the same time. Our approach consists of three main components: (1) a merging method based on Multiword Expressions (MWE); (2) a mapping strategy based on synonymy between words; (3) a re-ranking step to identify the best compression candidates generated using a POS-based language model (POS-LM). We demonstrate the effectiveness of this novel approach using a dataset made of clusters of English newswire sentences. The observed improvements on informativity and grammaticality of the generated compressions show that our approach is superior to state-of-the-art MSC methods.

preprint2015arXiv

Dynamic Sleep Control in Green Relay-Assisted Networks for Energy Saving and QoS Improving

We study the relay station (RS) sleep control mechanism targeting on reducing energy consumption while improving users' quality of service (QoS) in green relay-assisted cellular networks, where the base station (BS) is powered by grid power and the RSs are powered by renewable energy. By adopting green RSs, the grid power consumption of the BS is greatly reduced. But due to the uncertainty and stochastic characteristics of the renewable energy, power supply for RSs is not always sufficient. Thus the harvested energy needs to be scheduled appropriately to cater to the dynamic traffic so as to minimize the energy saving in the long term. An optimization problem is formulated to find the optimal sleep ratio of RSs to match the time variation of energy harvesting and traffic arrival. To fully use the renewable energy, green-RS-first principle is adopted in the user association process. The optimal RS sleeping policy is obtained through dynamic programming (DP) approach, which divides the original optimization problem into per-stage subproblems. A reduced DP algorithm and a greedy algorithm are further proposed to greatly reduce the computation complexity. By simulations, the reduced DP algorithm outperforms the greedy algorithm in achieving satisfactory energy saving and QoS performance.

preprint2015arXiv

Simultaneous Acquisition of Multi-nuclei Enhanced NMR/MRI by Solution State Dynamic Nuclear Polarization

Dynamic nuclear polarization (DNP) has become a very important hyperpolarization method because it can dramatically increase the sensitivity of nuclear magnetic resonance (NMR) of various molecules. Liquid-state DNP based on Overhauser effect is capable of directly enhancing polarizations of all kinds of nuclei in the system. The combination of simultaneous Overhauser multi-nuclei enhancements with the multi-nuclei parallel acquisitions provides a variety of important applications in both MR spectroscopy (MRS) and image (MRI). Here we present two simple illustrative examples for simultaneously enhanced multi-nuclear spectra and images to demonstrate the principle and superiority. We have observed very large simultaneous DNP enhancements for different nuclei, such as 1H and 23Na, 1H and 31P, 19F and 31P, especially for the first time to report sodium ion enhancement in liquid. We have also obtained the simultaneous imaging of 19H and 31P at low field by solution-state DNP for the first time. This method can obtain considerably complementary structure-determination information of miscellaneous biomolecules from a single measurement. It can also be used in combination with the fast acquisition schemes and quantitative analysis with reduced scan time.

preprint2014arXiv

Journey to the Center of the Fuzzball

We study two-charge fuzzball geometries, with attention to the use of the proper duality frame. For zero angular momentum there is an onion-like structure, and the smooth D1-D5 geometries are not valid for typical states. Rather, they are best approximated by geometries with stringy sources, or by a free CFT. For non-zero angular momentum we find a regime where smooth fuzzball solutions are the correct description. Our analysis rests on the comparison of three radii: the typical fuzzball radius, the entropy radius determined by the microscopic theory, and the breakdown radius where the curvature becomes large. We attempt to draw more general lessons.

preprint2013arXiv

Gauge/Gravity Duality in Heterotic String Theory

Gravity duals for little string theories --- which give rise to four-dimensional theories that undergo permanent confinement in the infrared --- have not been studied in great detail. We address this question in the framework of heterotic SO(32) and E_8 x E_8 string theory, constructing these backgrounds by wrapping heterotic five-branes on calibrated two-cycles of non-Kahler resolved conifolds. Related to deformations of the underlying little string theories, we find numerous analytic solutions preserving N = 1 supersymmetry in four-dimensions. These theories all have non-abelian global symmetries that generally arise from both the heterotic vector bundle and from certain orbifold states. In the decoupling limit, we argue that the gravity duals are given by non-Kahler manifolds that have both blown-up two-cycles and three-cycles at the origin. We argue this following certain duality sequences that include M-theory torsional manifolds at an intermediate step, which help us to construct new type I' gauge/gravity duality pairs. In the M-theory duality frame, we also elucidate new sequences of flips and flops.

preprint2013arXiv

Non extremal geometries and holographic phase transitions

Using the low energy limit of type IIB superstring theory, we obtain the non-extremal limit of deformed conifold geometry which is dual to the IR limit of large N thermal QCD.At low temperatures, the extremal geometry without black hole is favored while at high temperatures, the field theory is described by non-extremal black hole geometry. We compute the ten dimensional on shell action for extremal and non-extremal geometries and demonstrate that at a critical temperature $T_c$ there is a first order confinement to deconfinement phase transition. We compute $T_c$ as a function of 'tHooft coupling and study the thermodynamics of the dual gauge theory by evaluating the free energy and entropy of the ten dimensional geometry. We find agreement with the conformal limit while thermodynamics of non-conformal strongly coupled gauge theories is explored using the black hole geometries in non-AdS space.

preprint2012arXiv

A UV complete model of Large N Thermal QCD

Many recent works on large N holographic QCD in the planar limit have not considered UV completions, restricting exclusively towards analyzing the IR physics. Due to this, the UV problems like Landau poles and divergences of Wilson loops including instabilities at high temperatures have not been addressed. In some of our recent papers, we have discussed a possible UV completion, which is conformal in the UV and confining in the far IR, that avoids the Landau poles and the Wilson loop divergences. In this paper we give a field theory realization of this including the complete RG flow. We extend our UV complete model to study scenarios both above and below the deconfinement temperature and argue how phase transition in our model should be understood. Interestingly, because of the UV completion, subtle issues like instability due to negative specific heat do not appear. We also briefly elucidate the advantages that our model may have over other models studying large N thermal QCD.

preprint2012arXiv

Non-Extremality, Chemical Potential and the Infrared limit of Large N Thermal QCD

Non-extremal solution with warped resolved-deformed conifold background is important to study the infrared limit of large N thermal QCD. Earlier works in this direction have not taken into account all the back-reactions on the geometry, namely from the branes, fluxes, and black-hole carefully. In the present work we make some progress in this direction by solving explicitly the supergravity equations of motions in the presence of the backreaction from the black-hole. The backreactions from the branes and the fluxes on the other hand and to the order that we study, are comparatively suppressed. Our analysis reveal, among other things, how the resolution parameter would depend on the horizon radius and how the RG flows of the coupling constants should be understood in these scenarios, including their effects on the background three-form fluxes. We also study the effect of switching on a chemical potential in the background and, in a particularly simplified scenario, compute the actual value of the chemical potential for our case.

preprint2012arXiv

On the Scalar Spectrum of the Y^{p,q} Manifolds

The spectra of supergravity modes in anti de Sitter (AdS) space on a five-sphere endowed with the round metric (which is the simplest 5d Sasaki-Einstein space) has been studied in detail in the past. However for the more general class of cohomogeneity one Sasaki-Einstein metrics on S^2 x S^3, given by the Y^{p, q} class, a complete study of the spectra has not been attempted. Earlier studies on scalar spectrum were restricted to only the first few eigenstates. In this paper we take a step in this direction by analysing the full scalar spectrum on these spaces. However it turns out that finding the exact solution of the corresponding eigenvalue problem in closed form is not feasible since the computation of the eigenvalues of the Laplacian boils down to the analysis of a one-dimensional operator of Heun type, whose spectrum cannot be computed in closed form. However, despite this analytical obstacle, we manage to get both lower and upper bounds on the eigenvalues of the scalar spectrum by comparing the eigenvalue problem with a simpler, solvable system. We also briefly touch upon various other new avenues such as non-commutative and dipole deformations as well as possible non-conformal extensions of these models.

preprint2011arXiv

Metastable dark matter mechanisms for INTEGRAL 511 keV $γ$ rays and DAMA/CoGeNT events

We explore dark matter mechanisms that can simultaneously explain the galactic 511 keV gamma rays observed by INTEGRAL/SPI, the DAMA/LIBRA annual modulation, and the excess of low-recoil dark matter candidates observed by CoGeNT. It requires three nearly degenerate states of dark matter in the 4-7 GeV mass range, with splittings respectively of order an MeV and a few keV. The top two states have the small mass gap and transitions between them, either exothermic or endothermic, can account for direct detections. Decays from one of the top states to the ground state produce low-energy positrons in the galaxy whose associated 511 keV gamma rays are seen by INTEGRAL. This decay can happen spontaneously, if the excited state is metastable (longer-lived than the age of the universe), or it can be triggered by inelastic scattering of the metastable states into the shorter-lived ones. We focus on a simple model where the DM is a triplet of an SU(2) hidden sector gauge symmetry, broken at the scale of a few GeV, giving masses of order \lsim 1 GeV to the dark gauge bosons, which mix kinetically with the standard model hypercharge. The purely decaying scenario can give the observed angular dependence of the 511 keV signal with no positron diffusion, while the inelastic scattering mechanism requires transport of the positrons over distances \sim 1 kpc before annihilating. We note that an x-ray line of several keV in energy, due to single-photon decays involving the top DM states, could provide an additional component to the diffuse x-ray background. The model is testable by proposed low-energy fixed target experiments.

preprint2011arXiv

Supersymmetric Configurations, Geometric Transitions and New Non-Kahler Manifolds

We give a detailed derivation of a supersymmetric configuration of wrapped D5-branes on a two-cycle of a warped resolved conifold. Our analysis reveals that the resolved conifold should support a non-Kahler metric with an SU(3) structure. We use this as a starting point of the geometric transition in type IIB theory. A mirror, and a subsequent flop transition using an intermediate M-theory configuration with a G_2 structure, gives rise to the complete IR geometric transition in type IIA theory. A further mirror transformation gives the type IIB gravity dual of the IR gauge theory on the wrapped D5-branes. Expectedly non-Kahler deformations of the resolved and the deformed conifolds appear as the gravity duals of the confining gauge theories in type IIA and type IIB theories respectively, although in more generic cases these manifolds could also be non-geometric. In the local limit we reproduce precisely the scenarios presented in our earlier works. Our present work should therefore be viewed as providing a supergravity proof of geometric transitions in the full global scenarios in type II theories.

preprint2011arXiv

Toward the Gravity Dual of Heterotic Small Instantons

The question of what happens when the heterotic SO(32) instanton becomes small was answered sometime back by Witten. The heterotic theory develops an enhanced Sp(2k) gauge symmetry for k small instantons, besides the allowed SO(32) gauge symmetry. An interesting question now is to ask what happens when we take the large k limit. In this paper we argue that in some special cases, where Gauss' law allows the large k limit, the dynamics of the large k small instantons can be captured by a dual gravitational description. For the cases that we elaborate in this paper, the gravity duals are non-Kahler manifolds although in general they could be non-geometric. These small instantons are heterotic five-branes and the duality allows us to study the strongly coupled field theories on these five-branes. We review and elaborate on some of the recent observations pointing towards this duality, and argue that in certain cases the gauge/gravity duality may be understood as small instanton transitions under which the instantons smoothen out and consequently lose the Sp(2k) gauge symmetry. This may explain how branes disappear on the dual side and are replaced by fluxes. We analyse the torsion classes before and after the transitions, and discuss briefly how the ADHM sigma model and related vector bundles could be studied for these scenarios.

preprint2010arXiv

Exciting dark matter in the galactic center

We reconsider the proposal of excited dark matter (DM) as an explanation for excess 511 keV gamma rays from positrons in the galactic center. We quantitatively compute the cross section for DM annihilation to nearby excited states, mediated by exchange of a new light gauge boson with off-diagonal couplings to the DM states. In models where both excited states must be heavy enough to decay into e^+ e^- and the ground state, the predicted rate of positron production is never large enough to agree with observations, unless one makes extreme assumptions about the local circular velocity in the Milky Way, or alternatively if there exists a metastable population of DM states which can be excited through a mass gap of less than 650 keV, before decaying into electrons and positrons.

preprint2009arXiv

A new twist on excited dark matter: implications for INTEGRAL, PAMELA/ATIC/PPB-BETS, DAMA

We show that the 511 keV gamma ray excess observed by INTEGRAL/SPI can be more robustly explained by exciting dark matter (DM) at the center of the galaxy, if there is a peculiar spectrum of DM states chi_0, chi_1 and chi_2, with masses M_0 ~ 500 GeV, M_1 <~ M_0 + 2 m_e, and M_2 = M_1 + delta M >~ M_0 + 2 m_e. The small mass splitting delta M should be <~ 100 keV. In addition, we require at least two new gauge bosons (preferably three), with masses ~100 MeV. With this spectrum, chi_1 is stable, but can be excited to chi_2 by low-velocity DM scatterings near the galactic center, which are Sommerfeld-enhanced by two of the 100 MeV gauge boson exchanges. The excited state chi_2 decays to chi_0 and nonrelativistic e+e-, mediated by the third gauge boson, which mixes with the photon and Z. Although such a small 100 keV splitting has been independently proposed for explaining the DAMA annual modulation through the inelastic DM mechanism, the need for stability of chi_1 (and hence seqestering it from the Standard Model) implies that our scenario cannot account for the DAMA signal. It can however address the PAMELA/ATIC positron excess via DM annihilation in the galaxy, and it offers the possibility of a sharper feature in the ATIC spectrum relative to previously proposed models. The data are consistent with three new gauge bosons, whose couplings fit naturally into a broken SU(2) gauge theory where the DM is a triplet of the SU(2). We propose a simple model in which the SU(2) is broken by new Higgs triplet and 5-plet VEV's, giving rise to the right spectrum of DM, and mixing of one of the new gauge bosons with the photon and Z boson. A coupling of the DM to a heavy Z' may also be necessary to get the right relic density and PAMELA/ATIC signals.