Source author record

Emille E. O. Ishida

Emille E. O. Ishida appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2022arXiv

Fink: early supernovae Ia classification using active learning

We describe how the Fink broker early supernova Ia classifier optimizes its ML classifications by employing an active learning (AL) strategy. We demonstrate the feasibility of implementation of such strategies in the current Zwicky Transient Facility (ZTF) public alert data stream. We compare the performance of two AL strategies: uncertainty sampling and random sampling. Our pipeline consists of 3 stages: feature extraction, classification and learning strategy. Starting from an initial sample of 10 alerts (5 SN Ia and 5 non-Ia), we let the algorithm identify which alert should be added to the training sample. The system is allowed to evolve through 300 iterations. Our data set consists of 23 840 alerts from the ZTF with confirmed classification via cross-match with SIMBAD database and the Transient name server (TNS), 1 600 of which were SNe Ia (1 021 unique objects). The data configuration, after the learning cycle was completed, consists of 310 alerts for training and 23 530 for testing. Averaging over 100 realizations, the classifier achieved 89% purity and 54% efficiency. From 01/November/2020 to 31/October/2021 Fink has applied its early supernova Ia module to the ZTF stream and communicated promising SN Ia candidates to the TNS. From the 535 spectroscopically classified Fink candidates, 459 (86%) were proven to be SNe Ia. Our results confirm the effectiveness of active learning strategies for guiding the construction of optimal training samples for astronomical classifiers. It demonstrates in real data that the performance of learning algorithms can be highly improved without the need of extra computational resources or overwhelmingly large training samples. This is, to our knowledge, the first application of AL to real alerts data.

preprint2022arXiv

From Data to Software to Science with the Rubin Observatory LSST

The Vera C. Rubin Observatory Legacy Survey of Space and Time (LSST) dataset will dramatically alter our understanding of the Universe, from the origins of the Solar System to the nature of dark matter and dark energy. Much of this research will depend on the existence of robust, tested, and scalable algorithms, software, and services. Identifying and developing such tools ahead of time has the potential to significantly accelerate the delivery of early science from LSST. Developing these collaboratively, and making them broadly available, can enable more inclusive and equitable collaboration on LSST science. To facilitate such opportunities, a community workshop entitled "From Data to Software to Science with the Rubin Observatory LSST" was organized by the LSST Interdisciplinary Network for Collaboration and Computing (LINCC) and partners, and held at the Flatiron Institute in New York, March 28-30th 2022. The workshop included over 50 in-person attendees invited from over 300 applications. It identified seven key software areas of need: (i) scalable cross-matching and distributed joining of catalogs, (ii) robust photometric redshift determination, (iii) software for determination of selection functions, (iv) frameworks for scalable time-series analyses, (v) services for image access and reprocessing at scale, (vi) object image access (cutouts) and analysis at scale, and (vii) scalable job execution systems. This white paper summarizes the discussions of this workshop. It considers the motivating science use cases, identified cross-cutting algorithms, software, and services, their high-level technical specifications, and the principles of inclusive collaborations needed to develop them. We provide it as a useful roadmap of needs, as well as to spur action and collaboration between groups and individuals looking to develop reusable software for early LSST science.

preprint2022arXiv

How have astronomers cited other fields in the last decade?

We present a citation pattern analysis between astronomical papers and 13 other disciplines, based on the arXiv database over the past decade ($2010 - 2020$). We analyze 12,600 astronomical papers citing over 14,531 unique publications outside astronomy. Two striking patterns are unraveled. First, general relativity recently became the most cited field by astronomers, a trend highly correlated with the discovery of gravitational waves. Secondly, the fast growth of referenced papers in computer science and statistics, the first with a notable 15-fold increase since 2015. Such findings confirm the critical role of interdisciplinary efforts involving astronomy, statistics, and computer science in recent astronomical research.

preprint2019arXiv

Photometry of high-redshift blended galaxies using deep learning

The new generation of deep photometric surveys requires unprecedentedly precise shape and photometry measurements of billions of galaxies to achieve their main science goals. At such depths, one major limiting factor is the blending of galaxies due to line-of-sight projection, with an expected fraction of blended galaxies of up to 50%. Current deblending approaches are in most cases either too slow or not accurate enough to reach the level of requirements. This work explores the use of deep neural networks to estimate the photometry of blended pairs of galaxies in monochrome space images, similar to the ones that will be delivered by the Euclid space telescope. Using a clean sample of isolated galaxies from the CANDELS survey, we artificially blend them and train two different network models to recover the photometry of the two galaxies. We show that our approach can recover the original photometry of the galaxies before being blended with $\sim$7% accuracy without any human intervention and without any assumption on the galaxy shape. This represents an improvement of at least a factor of 4 compared to the classical SExtractor approach. We also show that forcing the network to simultaneously estimate a binary segmentation map results in a slightly improved photometry. All data products and codes will be made public to ease the comparison with other approaches on a common data set.

preprint2016arXiv

Large Magellanic Cloud Near-Infrared Synoptic Survey. III. A Statistical Study of Non-Linearity in the Leavitt Laws

We present a detailed statistical analysis of possible non-linearities in the Period-Luminosity (P-L), Period-Wesenheit (P-W) and Period-Color (P-C) relations for Cepheid variables in the LMC at optical ($VI$) and near-infrared ($JHK_{s}$) wavelengths. We test for the presence of possible non-linearities and determine their statistical significance by applying a variety of robust statistical tests ($F$-test, Random-Walk, Testimator and the Davies test) to optical data from OGLE III and near-infrared data from LMCNISS. For fundamental-mode Cepheids, we find that the optical P-L, P-W and P-C relations are non-linear at 10 days. The near-infrared P-L and the $W^H_{V,I}$ relations are non-linear around 18 days; this break is attributed to a distinct variation in mean Fourier amplitude parameters near this period for longer wavelengths as compared to optical bands. The near-infrared P-W relations are also non-linear except for the $W_{H,K_s}$ relation. For first-overtone mode Cepheids, a significant change in the slope of P-L, P-W and P-C relations is found around 2.5 days only at optical wavelengths. We determine a global slope of $\textrm{-}3.212\pm0.013$ for the $W^H_{V,I}$ relation by combining our LMC data with observations of Cepheids in Supernovae host galaxies \citep{riess11}. We find this slope to be consistent with the corresponding LMC relation at short periods, and significantly different to the long-period value. We do not find any significant difference in the slope of the global-fit solution using a linear or non-linear LMC P-L relation as calibrator, but the linear version provides a $2\times$ better constraint on the slope and metallicity coefficient.

preprint2012arXiv

Kernel PCA for type Ia supernovae photometric classification

In this work, we propose the use of Kernel Principal Component Analysis (KPCA) combined with k = 1 nearest neighbour algorithm (1NN) as a framework for supernovae (SNe) photometric classification. The classification is entirely based on information within the spectroscopic confirmed sample and each new light curve is classified one at a time. This allows us to update the principal component (PC) parameter space if a new spectroscopic light curve is available while also avoids the need of re-determining it for each individual new classification. We applied the method to different instances of the \textit{Supernova Photometric Classification Challenge} (SNPCC) data set. Our method provide good purity results in all data sample analysed, when SNR$\geq$5. As a consequence, we can state that if a sample as the post-SNPCC was available today, we would be able to classify $\approx 15%$ of the initial data set with purity $\gtrsim$ 90% (D$_{7}$+SNR3). Results from the original SNPCC sample, reported as a function of redshift, show that our method provides high purity (up to $\approx 97%$), specially in the range of $0.2\leq z < 0.4$, when compared to results from the SNPCC, while maintaining a moderate figure of merit ($\approx 0.25$). We also present results for SNe photometric classification using only pre-maximum epochs, obtaining 63% purity and 77% successful classification rates (SNR$\geq$5). Results are sensitive to the information contained in each light curve, as a consequence, higher quality data points lead to higher successful classification rates. The method is flexible enough to be applied to other astrophysical transients, as long as a training and a test sample are provided.

preprint2012arXiv

Searching for the first stars with the Gaia mission

We construct a theoretical model to predict the number of orphan afterglows (OA) from gamma-ray bursts (GRBs) triggered by primordial metal-free (Pop III) stars expected to be observed by the Gaia mission. In particular, we consider primordial metal-free stars that were affected by radiation from other stars (Pop III.2) as a possible target. We use a semi-analytical approach that includes all relevant feedback effects to construct cosmic star formation history and its connection with the cumulative number of GRBs. The OA events are generated using the Monte Carlo method, and realistic simulations of Gaia's scanning law are performed to derive the observation probability expectation. We show that Gaia can observe up to 2.28 $\pm$ 0.88 off-axis afterglows and 2.78 $\pm$ 1.41 on-axis during the five-year nominal mission. This implies that a nonnegligible percentage of afterglows that may be observed by Gaia ($\sim 10%$) could have Pop III stars as progenitors.

preprint2011arXiv

Hubble parameter reconstruction from a principal component analysis: minimizing the bias

A model-independent reconstruction of the cosmic expansion rate is essential to a robust analysis of cosmological observations. Our goal is to demonstrate that current data are able to provide reasonable constraints on the behavior of the Hubble parameter with redshift, independently of any cosmological model or underlying gravity theory. Using type Ia supernova data, we show that it is possible to analytically calculate the Fisher matrix components in a Hubble parameter analysis without assumptions about the energy content of the Universe. We used a principal component analysis to reconstruct the Hubble parameter as a linear combination of the Fisher matrix eigenvectors (principal components). To suppress the bias introduced by the high redshift behavior of the components, we considered the value of the Hubble parameter at high redshift as a free parameter. We first tested our procedure using a mock sample of type Ia supernova observations, we then applied it to the real data compiled by the Sloan Digital Sky Survey (SDSS) group. In the mock sample analysis, we demonstrate that it is possible to drastically suppress the bias introduced by the high redshift behavior of the principal components. Applying our procedure to the real data, we show that it allows us to determine the behavior of the Hubble parameter with reasonable uncertainty, without introducing any ad-hoc parameterizations. Beyond that, our reconstruction agrees with completely independent measurements of the Hubble parameter obtained from red-envelope galaxies.

preprint2011arXiv

Probing cosmic star formation up to z = 9.4 with GRBs

We propose a novel approach, based on Principal Components Analysis, to the use of Gamma-Ray Bursts (GRBs) as probes of cosmic star formation history (SFH) up to very high redshifts. The main advantage of such approach is to avoid the necessity of assuming an \textit{ad hoc} parameterization of the SFH. We first validate the method by reconstructing a known SFH from Monte Carlo-generated mock data. We then apply the method to the most recent \textit{Swift} data of GRBs with known redshift and compare it against the SFH obtained by independent methods. The main conclusion is that the level of star formation activity at $z \approx 9.4$ could have been already as high as the present-day one ($\approx 0.01 M_\odot$ yr$^{-1}$ Mpc$^{-3}$). This is a factor 3-5 times higher than deduced from high-$z$ galaxy searches through drop-out techniques. If true, this might alleviate the long-standing problem of a photon-starving reionization; it might also indicate that galaxies accounting for most of the star formation activity at high redshift go undetected by even the most deep searches.

preprint2011arXiv

The Effect of a Single Supernova Explosion on the Cuspy Density Profile of a Small-Mass Dark Matter Halo

Some observations of galaxies, and in particular dwarf galaxies, indicate a presence of cored density profiles in apparent contradiction with cusp profiles predicted by dark matter N-body simulations. We constructed an analytical model, using particle distribution functions (DFs), to show how a supernova (SN) explosion can transform a cusp density profile in a small-mass dark matter halo into a cored one. Considering the fact that a SN efficiently removes matter from the centre of the first haloes, we study the effect of mass removal through a SN perturbation in the DFs. We found that the transformation from a cusp into a cored profile is present even for changes as small as 0.5% of the total energy of the halo, that can be produced by the expulsion of matter caused by a single SN explosion.

preprint2010arXiv

An analytical approach to the dwarf galaxies cusp problem

An analytical solution for the discrepancy between observed core-like profiles and predicted cusp profiles in dark matter halos is studied. We calculate the distribution function for Navarro-Frenk-White halos and extract energy from the distribution, taking into account the effects of baryonic physics processes. We show with a simple argument that we can reproduce the evolution of a cusp to a flat density profile by a decrease of the initial potential energy.

preprint2006arXiv

Statefinder Revisited

The quality of supernova data will dramatically increase in the next few years by new experiments that will add high-redshift supernova to the currently known ones. In order to use this new data to discriminate between different dark energy models, the statefinder diagnostic was suggested and investigated by Alam et al. in the light of the proposed SuperNova Acceleration Probe (SNAP) satellite. By making use of the same procedure presented by these authors, we compare their analyzes with ours, which shows a more realistic supernovae redshift distribution and do not assume that the intercept is known. We also analyzed the behavior of the statefinder pair {r,s} and the alternative pair {s,q} in the presence of offset errors.