Source author record

Alex P. Leung

Alex P. Leung appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2022arXiv

On Improving the Performance of Glitch Classification for Gravitational Wave Detection by using Generative Adversarial Networks

Spectrogram classification plays an important role in analyzing gravitational wave data. In this paper, we propose a framework to improve the classification performance by using Generative Adversarial Networks (GANs). As substantial efforts and expertise are required to annotate spectrograms, the number of training examples is very limited. However, it is well known that deep networks can perform well only when the sample size of the training set is sufficiently large. Furthermore, the imbalanced sample sizes in different classes can also hamper the performance. In order to tackle these problems, we propose a GAN-based data augmentation framework. While standard data augmentation methods for conventional images cannot be applied on spectrograms, we found that a variant of GANs, ProGAN, is capable of generating high-resolution spectrograms which are consistent with the quality of the high-resolution original images and provide a desirable diversity. We have validated our framework by classifying glitches in the {\it Gravity Spy} dataset with the GAN-generated spectrograms for training. We show that the proposed method can provide an alternative to transfer learning for the classification of spectrograms using deep networks, i.e. using a high-resolution GAN for data augmentation instead. Furthermore, fluctuations in classification performance with small sample sizes for training and evaluation can be greatly reduced. Using the trained network in our framework, we have also examined the spectrograms with label anomalies in {\it Gravity Spy}.

preprint2020arXiv

An investigation on the factors affecting machine learning classifications in $γ$-ray astronomy

We have investigated a number of factors that can have significant impacts on the classification performance of $γ$-ray sources detected by Fermi Large Area Telescope (LAT) with machine learning techniques. We show that a framework of automatic feature selection can construct a simple model with a small set of features which yields better performance over previous results. Secondly, because of the small sample size of the training/test sets of certain classes in $γ$-ray, nested re-sampling and cross-validations are suggested for quantifying the statistical fluctuations of the quoted accuracy. We have also constructed a test set by cross-matching the identified active galactic nuclei (AGNs) and the pulsars (PSRs) in the Fermi LAT eight-year point source catalog (4FGL) with those unidentified sources in the previous 3$^{\rm rd}$ Fermi LAT Source Catalog (3FGL). Using this cross-matched set, we show that some features used for building classification model with the identified source can suffer from the problem of covariate shift, which can be a result of various observational effects. This can possibly hamper the actual performance when one applies such model in classifying unidentified sources. Using our framework, both AGN/PSR and young pulsar (YNG)/millisecond pulsar (MSP) classifiers are automatically updated with the new features and the enlarged training samples in 4FGL catalog incorporated. Using a two-layer model with these updated classifiers, we have selected 20 promising MSP candidates with confidence scores $>98\%$ from the unidentified sources in 4FGL catalog which can provide inputs for a multi-wavelength identification campaign.

preprint2020arXiv

Searches for Pulsar-like Candidates from Unidentified Objects in the Third Catalog of Hard Fermi-LAT (3FHL) sources with Machine Learning Techniques

We report the results of searching pulsar-like candidates from the unidentified objects in the $3^{\rm rd}$ Catalog of Hard Fermi-LAT sources (3FHL). Using a machine-learning based classification scheme with a nominal accuracy of $\sim98\%$, we have selected 27 pulsar-like objects from 200 unidentified 3FHL sources for an identification campaign. Using archival data, X-ray sources are found within the $γ-$ray error ellipses of 10 3FHL pulsar-like candidates. Within the error circles of the much better constrained X-ray positions, we have also searched for the optical/infrared counterparts and examined their spectral energy distributions. Among our short-listed candidates, the most secure identification is the association of 3FHL J1823.3-1339 and its X-ray counterpart with the globular cluster Mercer 5. The $γ-$rays from the source can be contributed by a population of millisecond pulsars residing in the cluster. This makes Mercer 5 as one of the slowly growing hard $γ-$ray population of globular clusters with emission $>10$ GeV. Very recently, another candidate picked by our classification scheme, 3FHL J1405.1-6118, has been identified as a new $γ-$ray binary with an orbital period of $13.7$ days. Our X-ray analysis with a short Chandra observation has found a possible periodic signal candidate of $\sim1.4$ hrs and a putative extended X-ray tail of $\sim20$ arcsec long. Spectral energy distribution of its optical/infrared counterpart conforms with a blackbody of $T_{\rm bb}\sim40000$ K and $R_{\rm bb}\sim12R_{\odot}$ at a distance of 7.7 kpc. This is consistent with its identification as an early O star as found by infrared spectroscopy.

preprint2014arXiv

PinView: Implicit Feedback in Content-Based Image Retrieval

This paper describes PinView, a content-based image retrieval system that exploits implicit relevance feedback collected during a search session. PinView contains several novel methods to infer the intent of the user. From relevance feedback, such as eye movements or pointer clicks, and visual features of images, PinView learns a similarity metric between images which depends on the current interests of the user. It then retrieves images with a specialized online learning algorithm that balances the tradeoff between exploring new images and exploiting the already inferred interests of the user. We have integrated PinView to the content-based image retrieval system PicSOM, which enables applying PinView to real-world image databases. With the new algorithms PinView outperforms the original PicSOM, and in online experiments with real users the combination of implicit and explicit feedback gives the best results.