Source author record

Bruce A. Bassett

Bruce A. Bassett appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

24works
14topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

24 published item(s)

preprint2022arXiv

A Hitchhiker's Guide to Anomaly Detection with Astronomaly

The next generation of telescopes such as the SKA and the Rubin Observatory will produce enormous data sets, requiring automated anomaly detection to enable scientific discovery. Here, we present an overview and friendly user guide to the Astronomaly framework for active anomaly detection in astronomical data. Astronomaly uses active learning to combine the raw processing power of machine learning with the intuition and experience of a human user, enabling personalised recommendations of interesting anomalies. It makes use of a Python backend to perform data processing, feature extraction and machine learning to detect anomalous objects; and a JavaScript frontend to allow interaction with the data, labelling of interesting anomalous and active learning. Astronomaly is designed to be modular, extendable and run on almost any type of astronomical data. In this paper, we detail the structure of the Astronomaly code and provide guidelines for basic usage.

preprint2022arXiv

The Hydrogen Intensity and Real-time Analysis eXperiment: 256-Element Array Status and Overview

The Hydrogen Intensity and Real-time Analysis eXperiment (HIRAX) is a radio interferometer array currently in development, with an initial 256-element array to be deployed at the South African Radio Astronomy Observatory (SARAO) Square Kilometer Array (SKA) site in South Africa. Each of the 6m, $f/0.23$ dishes will be instrumented with dual-polarisation feeds operating over a frequency range of 400-800 MHz. Through intensity mapping of the 21 cm emission line of neutral hydrogen, HIRAX will provide a cosmological survey of the distribution of large-scale structure over the redshift range of $0.775 < z < 2.55$ over $\sim$15,000 square degrees of the southern sky. The statistical power of such a survey is sufficient to produce $\sim$7 percent constraints on the dark energy equation of state parameter when combined with measurements from the Planck satellite. Additionally, HIRAX will provide a highly competitive platform for radio transient and HI absorber science while enabling a multitude of cross-correlation studies. In this paper, we describe the science goals of the experiment, overview of the design and status of the sub-components of the telescope system, and describe the expected performance of the initial 256-element array as well as the planned future expansion to the final, 1024-element array.

preprint2021arXiv

Classification of Multiwavelength Transients with Machine Learning

With the advent of powerful telescopes such as the Square Kilometer Array and the Vera C. Rubin Observatory, we are entering an era of multiwavelength transient astronomy that will lead to a dramatic increase in data volume. Machine learning techniques are well suited to address this data challenge and rapidly classify newly detected transients. We present a multiwavelength classification algorithm consisting of three steps: (1) interpolation and augmentation of the data using Gaussian processes; (2) feature extraction using wavelets; and (3) classification with random forests. Augmentation provides improved performance at test time by balancing the classes and adding diversity into the training set. In the first application of machine learning to the classification of real radio transient data, we apply our technique to the Green Bank Interferometer and other radio light curves. We find we are able to accurately classify most of the 11 classes of radio variables and transients after just eight hours of observations, achieving an overall test accuracy of 78 percent. We fully investigate the impact of the small sample size of 82 publicly available light curves and use data augmentation techniques to mitigate the effect. We also show that on a significantly larger simulated representative training set that the algorithm achieves an overall accuracy of 97 percent, illustrating that the method is likely to provide excellent performance on future surveys. Finally, we demonstrate the effectiveness of simultaneous multiwavelength observations by showing how incorporating just one optical data point into the analysis improves the accuracy of the worst performing class by 19 percent.

preprint2020arXiv

Climate & BCG: Effects on COVID-19 Death Growth Rates

Multiple studies have suggested the spread of COVID-19 is affected by factors such as climate, BCG vaccinations, pollution and blood type. We perform a joint study of these factors using the death growth rates of 40 regions worldwide with both machine learning and Bayesian methods. We find weak, non-significant (< 3$σ$) evidence for temperature and relative humidity as factors in the spread of COVID-19 but little or no evidence for BCG vaccination prevalence or $\text{PM}_{2.5}$ pollution. The only variable detected at a statistically significant level (>3$σ$) is the rate of positive COVID-19 tests, with higher positive rates correlating with higher daily growth of deaths.

preprint2019arXiv

A Flexible Framework for Anomaly Detection via Dimensionality Reduction

Anomaly detection is challenging, especially for large datasets in high dimensions. Here we explore a general anomaly detection framework based on dimensionality reduction and unsupervised clustering. We release DRAMA, a general python package that implements the general framework with a wide range of built-in options. We test DRAMA on a wide variety of simulated and real datasets, in up to 3000 dimensions, and find it robust and highly competitive with commonly-used anomaly detection algorithms, especially in high dimensions. The flexibility of the DRAMA framework allows for significant optimization once some examples of anomalies are available, making it ideal for online anomaly detection, active learning and highly unbalanced datasets.

preprint2016arXiv

Application of Bayesian graphs to SN Ia data analysis and compression

Bayesian graphical models are an efficient tool for modelling complex data and derive self-consistent expressions of the posterior distribution of model parameters. We apply Bayesian graphs to perform statistical analyses of Type Ia supernova (SN Ia) luminosity distance measurements from the joint light-curve analysis (JLA) data set. In contrast to the $χ^2$ approach used in previous studies, the Bayesian inference allows us to fully account for the standard-candle parameter dependence of the data covariance matrix. Comparing with $χ^2$ analysis results, we find a systematic offset of the marginal model parameter bounds. We demonstrate that the bias is statistically significant in the case of the SN Ia standardization parameters with a maximal 6 $σ$ shift of the SN light-curve colour correction. In addition, we find that the evidence for a host galaxy correction is now only 2.4 $σ$. Systematic offsets on the cosmological parameters remain small, but may increase by combining constraints from complementary cosmological probes. The bias of the $χ^2$ analysis is due to neglecting the parameter-dependent log-determinant of the data covariance, which gives more statistical weight to larger values of the standardization parameters. We find a similar effect on compressed distance modulus data. To this end, we implement a fully consistent compression method of the JLA data set that uses a Gaussian approximation of the posterior distribution for fast generation of compressed data. Overall, the results of our analysis emphasize the need for a fully consistent Bayesian statistical approach in the analysis of future large SN Ia data sets.

preprint2016arXiv

Bayes Factors via Savage-Dickey Supermodels

We outline a new method to compute the Bayes Factor for model selection which bypasses the Bayesian Evidence. Our method combines multiple models into a single, nested, Supermodel using one or more hyperparameters. Since the models are now nested the Bayes Factors between the models can be efficiently computed using the Savage-Dickey Density Ratio (SDDR). In this way model selection becomes a problem of parameter estimation. We consider two ways of constructing the supermodel in detail: one based on combined models, and a second based on combined likelihoods. We report on these two approaches for a Gaussian linear model for which the Bayesian evidence can be calculated analytically and a toy nonlinear problem. Unlike the combined model approach, where a standard Monte Carlo Markov Chain (MCMC) struggles, the combined-likelihood approach fares much better in providing a reliable estimate of the log-Bayes Factor. This scheme potentially opens the way to computationally efficient ways to compute Bayes Factors in high dimensions that exploit the good scaling properties of MCMC, as compared to methods such as nested sampling that fail for high dimensions.

preprint2015arXiv

Bayesian Inference for Radio Observations

New telescopes like the Square Kilometre Array (SKA) will push into a new sensitivity regime and expose systematics, such as direction-dependent effects, that could previously be ignored. Current methods for handling such systematics rely on alternating best estimates of instrumental calibration and models of the underlying sky, which can lead to inadequate uncertainty estimates and biased results because any correlations between parameters are ignored. These deconvolution algorithms produce a single image that is assumed to be a true representation of the sky, when in fact it is just one realization of an infinite ensemble of images compatible with the noise in the data. In contrast, here we report a Bayesian formalism that simultaneously infers both systematics and science. Our technique, Bayesian Inference for Radio Observations (BIRO), determines all parameters directly from the raw data, bypassing image-making entirely, by sampling from the joint posterior probability distribution. This enables it to derive both correlations and accurate uncertainties, making use of the flexible software MEQTREES to model the sky and telescope simultaneously. We demonstrate BIRO with two simulated sets of Westerbork Synthesis Radio Telescope data sets. In the first, we perform joint estimates of 103 scientific (flux densities of sources) and instrumental (pointing errors, beamwidth and noise) parameters. In the second example, we perform source separation with BIRO. Using the Bayesian evidence, we can accurately select between a single point source, two point sources and an extended Gaussian source, allowing for 'super-resolution' on scales much smaller than the synthesized beam.

preprint2015arXiv

Nonparametric Transient Classification using Adaptive Wavelets

Classifying transients based on multi band light curves is a challenging but crucial problem in the era of GAIA and LSST since the sheer volume of transients will make spectroscopic classification unfeasible. Here we present a nonparametric classifier that uses the transient's light curve measurements to predict its class given training data. It implements two novel components: the first is the use of the BAGIDIS wavelet methodology - a characterization of functional data using hierarchical wavelet coefficients. The second novelty is the introduction of a ranked probability classifier on the wavelet coefficients that handles both the heteroscedasticity of the data in addition to the potential non-representativity of the training set. The ranked classifier is simple and quick to implement while a major advantage of the BAGIDIS wavelets is that they are translation invariant, hence they do not need the light curves to be aligned to extract features. Further, BAGIDIS is nonparametric so it can be used for blind searches for new objects. We demonstrate the effectiveness of our ranked wavelet classifier against the well-tested Supernova Photometric Classification Challenge dataset in which the challenge is to correctly classify light curves as Type Ia or non-Ia supernovae. We train our ranked probability classifier on the spectroscopically-confirmed subsample (which is not representative) and show that it gives good results for all supernova with observed light curve timespans greater than 100 days (roughly 55% of the dataset). For such data, we obtain a Ia efficiency of 80.5% and a purity of 82.4% yielding a highly competitive score of 0.49 whilst implementing a truly "model-blind" approach to supernova classification. Consequently this approach may be particularly suitable for the classification of astronomical transients in the era of large synoptic sky surveys.

preprint2015arXiv

Observational Constraints on Redshift Remapping

There are two redshifts in cosmology: $z_{obs}$, the observed redshift computed via spectral lines, and the model redshift, $z$, defined by the effective FLRW scale factor. In general these do not coincide. We place observational constraints on the allowed distortions of $z$ away from $z_{obs}$ - a possibility we dub redshift remapping. Remapping is degenerate with cosmic dynamics for either $d_L(z)$ or $H(z)$ observations alone: for example, the simple remapping $z = α_1 z_{obs} +α_2 z_{obs}^2$ allows a decelerating Einstein de Sitter universe to fit the observed supernova Hubble diagram as successfully as $Λ$CDM, highlighting that supernova data alone cannot prove that the universe is accelerating. We show however, that redshift remapping leads to apparent violations of cosmic distance duality that can be used to detect its presence even when neither a specific theory of gravity nor the Copernican Principle are assumed. Combining current data sets favours acceleration but does not yet rule out redshift remapping as an alternative to dark energy. Future surveys, however, will provide exquisite constraints on remapping and any models -- such as backreaction -- that predict it.

preprint2014arXiv

Extending BEAMS to incorporate correlated systematic uncertainties

New supernova surveys such as the Dark Energy Survey, Pan-STARRS and the LSST will produce an unprecedented number of photometric supernova candidates, most with no spectroscopic data. Avoiding biases in cosmological parameters due to the resulting inevitable contamination from non-Ia supernovae can be achieved with the BEAMS formalism, allowing for fully photometric supernova cosmology studies. Here we extend BEAMS to deal with the case in which the supernovae are correlated by systematic uncertainties. The analytical form of the full BEAMS posterior requires evaluating 2^N terms, where N is the number of supernova candidates. This `exponential catastrophe' is computationally unfeasible even for N of order 100. We circumvent the exponential catastrophe by marginalising numerically instead of analytically over the possible supernova types: we augment the cosmological parameters with nuisance parameters describing the covariance matrix and the types of all the supernovae, τ_i, that we include in our MCMC analysis. We show that this method deals well even with large, unknown systematic uncertainties without a major increase in computational time, whereas ignoring the correlations can lead to significant biases and incorrect credible contours. We then compare the numerical marginalisation technique with a perturbative expansion of the posterior based on the insight that future surveys will have exquisite light curves and hence the probability that a given candidate is a Type Ia will be close to unity or zero, for most objects. Although this perturbative approach changes computation of the posterior from a 2^N problem into an N^2 or N^3 one, we show that it leads to biases in general through a small number of misclassifications, implying that numerical marginalisation is superior.

preprint2014arXiv

Towards the Future of Supernova Cosmology

For future surveys, spectroscopic follow-up for all supernovae will be extremely difficult. However, one can use light curve fitters, to obtain the probability that an object is a Type Ia. One may consider applying a probability cut to the data, but we show that the resulting non-Ia contamination can lead to biases in the estimation of cosmological parameters. A different method, which allows the use of the full dataset and results in unbiased cosmological parameter estimation, is Bayesian Estimation Applied to Multiple Species (BEAMS). BEAMS is a Bayesian approach to the problem which includes the uncertainty in the types in the evaluation of the posterior. Here we outline the theory of BEAMS and demonstrate its effectiveness using both simulated datasets and SDSS-II data. We also show that it is possible to use BEAMS if the data are correlated, by introducing a numerical marginalisation over the types of the objects. This is largely a pedagogical introduction to BEAMS with references to the main BEAMS papers.

preprint2013arXiv

The Effect of Weak Lensing on Distance Estimates from Supernovae

Using a sample of 608 Type Ia supernovae from the SDSS-II and BOSS surveys, combined with a sample of foreground galaxies from SDSS-II, we estimate the weak lensing convergence for each supernova line-of-sight. We find that the correlation between this measurement and the Hubble residuals is consistent with the prediction from lensing (at a significance of 1.7sigma. Strong correlations are also found between the residuals and supernova nuisance parameters after a linear correction is applied. When these other correlations are taken into account, the lensing signal is detected at 1.4sigma. We show for the first time that distance estimates from supernovae can be improved when lensing is incorporated by including a new parameter in the SALT2 methodology for determining distance moduli. The recovered value of the new parameter is consistent with the lensing prediction. Using CMB data from WMAP7, H0 data from HST and SDSS BAO measurements, we find the best-fit value of the new lensing parameter and show that the central values and uncertainties on Omega_m and w are unaffected. The lensing of supernovae, while only seen at marginal significance in this low redshift sample, will be of vital importance for the next generation of surveys, such as DES and LSST, which will be systematics dominated.

preprint2012arXiv

BEAMS: separating the wheat from the chaff in supernova analysis

We introduce Bayesian Estimation Applied to Multiple Species (BEAMS), an algorithm designed to deal with parameter estimation when using contaminated data. We present the algorithm and demonstrate how it works with the help of a Gaussian simulation. We then apply it to supernova data from the Sloan Digital Sky Survey (SDSS), showing how the resulting confidence contours of the cosmological parameters shrink significantly.

preprint2012arXiv

Fisher Matrix Preloaded -- Fisher4Cast

The Fisher Matrix is the backbone of modern cosmological forecasting. We describe the Fisher4Cast software: a general-purpose, easy-to-use, Fisher Matrix framework. It is open source, rigorously designed and tested and includes a Graphical User Interface (GUI) with automated LATEX file creation capability and point-and-click Fisher ellipse generation. Fisher4Cast was designed for ease of extension and, although written in Matlab, is easily portable to open-source alternatives such as Octave and Scilab. Here we use Fisher4Cast to present new 3-D and 4-D visualisations of the forecasting landscape and to investigate the effects of growth and curvature on future cosmological surveys. Early releases have been available at http://www.cosmology.org.za since May 2008 with 750 downloads in the first year. Version 2.2 is made public with this paper and includes a Quick Start guide and the code used to produce the figures in this paper, in the hope that it will be useful to the cosmology and wider scientific communities.

preprint2012arXiv

Non-Gaussian Posteriors arising from Marginal Detections

We show that in cases of marginal detections (~ 3σ), such as that of Baryonic Acoustic Oscillations (BAO) in cosmology, the often-used Gaussian approximation to the full likelihood is very poor, especially beyond ~3σ. This can radically alter confidence intervals on parameters and implies that one cannot naively extrapolate 1σ-errorbars to 3σ, and beyond. We propose a simple fitting formula which corrects for this effect in posterior probabilities arising from marginal detections. Alternatively the full likelihood should be used for parameter estimation rather than the Gaussian approximation of a just mean and an error.

preprint2010arXiv

Statistical Classification Techniques for Photometric Supernova Typing

Future photometric supernova surveys will produce vastly more candidates than can be followed up spectroscopically, highlighting the need for effective classification methods based on lightcurves alone. Here we introduce boosting and kernel density estimation techniques which have minimal astrophysical input, and compare their performance on 20,000 simulated Dark Energy Survey lightcurves. We demonstrate that these methods are comparable to the best template fitting methods currently used, and in particular do not require the redshift of the host galaxy or candidate. However both methods require a training sample that is representative of the full population, so typical spectroscopic supernova subsamples will lead to poor performance. To enable the full potential of such blind methods, we recommend that representative training samples should be used and so specific attention should be given to their creation in the design phase of future photometric surveys.

preprint2009arXiv

Optimizing baryon acoustic oscillation surveys II: curvature, redshifts, and external datasets

We extend our study of the optimization of large baryon acoustic oscillation (BAO) surveys to return the best constraints on the dark energy, building on Paper I of this series (Parkinson et al. 2007). The survey galaxies are assumed to be pre-selected active, star-forming galaxies observed by their line emission with a constant number density across the redshift bin. Star-forming galaxies have a redshift desert in the region 1.6 < z < 2, and so this redshift range was excluded from the analysis. We use the Seo & Eisenstein (2007) fitting formula for the accuracies of the BAO measurements, using only the information for the oscillatory part of the power spectrum as distance and expansion rate rulers. We go beyond our earlier analysis by examining the effect of including curvature on the optimal survey configuration and updating the expected `prior' constraints from Planck and SDSS. We once again find that the optimal survey strategy involves minimizing the exposure time and maximizing the survey area (within the instrumental constraints), and that all time should be spent observing in the low-redshift range (z<1.6) rather than beyond the redshift desert, z>2. We find that when assuming a flat universe the optimal survey makes measurements in the redshift range 0.1 < z <0.7, but that including curvature as a nuisance parameter requires us to push the maximum redshift to 1.35, to remove the degeneracy between curvature and evolving dark energy. The inclusion of expected other data sets (such as WiggleZ, BOSS and a stage III SN-Ia survey) removes the necessity of measurements below redshift 0.9, and pushes the maximum redshift up to 1.5. We discuss considerations in determining the best survey strategy in light of uncertainty in the true underlying cosmological model.

preprint2005arXiv

A Measurement of the Quadrupole Power Spectrum in the Clustering of the 2dF QSO Survey

We report a measurement of the quadrupole power spectrum in the two degree field (2dF) QSO redshift (2QZ) survey. The analysis uses an algorithm parallel to that for the estimation of the standard monopole power spectrum without first requiring computation of the correlation function or the anisotropic power spectrum. The error on the quadrupole spectrum is rather large but the best fit value of the bias parameter from the quadrupole spectrum is consistent with that from previous investigations of the 2dF data.

preprint2005arXiv

WFMOS - Sounding the Dark Cosmos

Vast sound waves traveling through the relativistic plasma during the first million years of the universe imprint a preferred scale in the density of matter. We now have the ability to detect this characteristic fingerprint in the clustering of galaxies at various redshifts and use it to measure the acceleration of the expansion of the Universe. The Wide-Field Multi-Object Spectrograph (WFMOS) would use this test to shed significant light on the true nature of dark energy, the mysterious source of this cosmic acceleration. WFMOS would also revolutionise studies of the kinematics of the Milky Way and provide deep insights into the clustering of galaxies at redshifts up to z~4. In this article we discuss the recent progress in large galaxy redshift surveys and detail how WFMOS will help unravel the mystery of dark energy.

preprint2004arXiv

Testing for double inflation with WMAP

With the WMAP data we can now begin to test realistic models of inflation involving multiple scalar fields. These naturally lead to correlated adiabatic and isocurvature (entropy) perturbations with a running spectral index. We present the first full (9 parameter) likelihood analysis of double inflation with WMAP data and find that despite the extra freedom, supersymmetric hybrid potentials are strongly constrained with less than 7% correlated isocurvature component allowed when standard priors are imposed on the cosomological parameters. As a result we also find that Akaike & Bayesian model selection criteria rather strongly prefer single-field inflation, just as equivalent analysis prefers a cosmological constant over dynamical dark energy in the late universe. It appears that simplicity is the best guide to our universe.

preprint2003arXiv

Mapping the Dark Energy with Varying Alpha

Cosmological dark energy is a natural source of variation of the fine structure constant. Using a model-independent approach we show that once general assumptions about the alpha-varying interactions are made, astronomical probes of its variation constrain the dark energy equation of state today to satisfy -1 < w_f < -0.96 at 3-sigma and significantly disfavour late-time changes in the equation of state. We show how dark-energy-induced spatial perturbations of alpha are linked to violations of the Equivalence Principle and are thus negligible at low-redshift, in stark contrast to the BSBM theories. This provides a new test of dark energy as the source of alpha variation.

preprint2000arXiv

Geometrodynamics of Variable-Speed-of-Light Cosmologies

This paper is dedicated to the memory of Dennis Sciama. Variable-Speed-of-Light (VSL) cosmologies are currently attracting interest as an alternative to inflation. We investigate the fundamental geometrodynamic aspects of VSL cosmologies and provide several implementations which do not explicitly break Lorentz invariance (no "hard" breaking). These "soft" implementations of Lorentz symmetry breaking provide particularly clean answers to the question "VSL with respect to what?". The class of VSL cosmologies we consider are compatible with both classical Einstein gravity and low-energy particle physics. These models solve the "kinematic" puzzles of cosmology as well as inflation does, but cannot by themselves solve the flatness problem, since in their purest form no violation of the strong energy condition occurs. We also consider a heterotic model (VSL plus inflation) which provides a number of observational implications for the low-redshift universe if chi contributes to the "dark energy" either as CDM or quintessence. These implications include modified gravitational lensing, birefringence, variation of fundamental constants and rotation of the plane of polarization of light from distant sources.

preprint1998arXiv

Geometric Reheating after Inflation

Inflationary reheating via resonant production of non-minimally coupled scalar particles with only gravitational coupling is shown to be extremely strong, exhibiting a negative coupling instability for $ξ< 0$ and a wide resonance decay for $ξ\gg 1$. Since non-minimal fields are generic after renormalisation in curved spacetime, this offers a new paradigm in reheating - one which naturally allows for efficient production of the massive bosons needed for GUT baryogenesis. We also show that both vector and tensor fields are produced resonantly during reheating, extending the previously known correspondences between bosonic fields of different spins during preheating.