Catalog footprint

What is connected

40works
20topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

40 published item(s)

preprint2022arXiv

European historical evidence of the supernova of AD 1054 coins of Constantine IX and SN 1054

We investigate a possible depiction of the famous SN 1054 event in specially minted coins produced in the Eastern Roman Empire in 1054 A.D. On these coins, we investigate if the head of the Emperor, Constantine IX, might represent the Sun with a bright 'star' on either side - Venus in the east and SN 1054 in the west, perhaps also representing the newly split Christian churches. We explore the idea that the eastern star represents the stable and well-known Venus and the Eastern Orthodox Church, while the western star represents the short-lived 'new star' and the 'fading' Western Catholic church. We examined 36 coins of this rare Constantine IX Class IV batch. While no exact date could be associated to any of these coins, they most likely were minted during the last six months of Constantine IX's rule in 1054. We hypothesise that the stance of the church concerning the order of the Universe, as well as the chaos surrounding the Great Schism, played a crucial role in stopping the official reporting of an obvious event in the sky, yet a dangerous omen. A temporal coincidence of all these events could be a reasonable explanation as well.

preprint2020arXiv

Building the largest spectroscopic sample of ultra-compact massive galaxies with the Kilo Degree Survey

Ultra-compact massive galaxies UCMGs, i.e. galaxies with stellar masses $M_{*} > 8 \times 10^{10} M_{\odot}$ and effective radii $R_{e} < 1.5$ kpc, are very rare systems, in particular at low and intermediate redshifts. Their origin as well as their number density across cosmic time are still under scrutiny, especially because of the paucity of spectroscopically confirmed samples. We have started a systematic census of UCMG candidates within the ESO Kilo Degree Survey, together with a large spectroscopic follow-up campaign to build the largest possible sample of confirmed UCMGs. This is the third paper of the series and the second based on the spectroscopic follow-up program. Here, we present photometrical and structural parameters of 33 new candidates at redshifts $0.15 \lesssim z \lesssim 0.5$ and confirm 19 of them as UCMGs, based on their nominal spectroscopically inferred $M_{*}$ and $R_{e}$. This corresponds to a success rate of $\sim 58\%$, nicely consistent with our previous findings. The addition of these 19 newly confirmed objects, allows us to fully assess the systematics on the system selection, and finally reduce the number density uncertainties. Moreover, putting together the results from our current and past observational campaigns and some literature data, we build the largest sample of UCMGs ever collected, comprising 92 spectroscopically confirmed objects at $0.1 \lesssim z \lesssim 0.5$. This number raises to 116, allowing for a $3σ$ tolerance on the $M_{*}$ and $R_{e}$ thresholds for the UCMG definition. For all these galaxies we have estimated the velocity dispersion values at the effective radii which have been used to derive a preliminary mass-velocity dispersion correlation.

preprint2016arXiv

A cooperative approach among methods for photometric redshifts estimation: an application to KiDS data

Photometric redshifts (photo-z's) are fundamental in galaxy surveys to address different topics, from gravitational lensing and dark matter distribution to galaxy evolution. The Kilo Degree Survey (KiDS), i.e. the ESO public survey on the VLT Survey Telescope (VST), provides the unprecedented opportunity to exploit a large galaxy dataset with an exceptional image quality and depth in the optical wavebands. Using a KiDS subset of about 25,000 galaxies with measured spectroscopic redshifts, we have derived photo-z's using i) three different empirical methods based on supervised machine learning, ii) the Bayesian Photometric Redshift model (or BPZ), and iii) a classical SED template fitting procedure (Le Phare). We confirm that, in the regions of the photometric parameter space properly sampled by the spectroscopic templates, machine learning methods provide better redshift estimates, with a lower scatter and a smaller fraction of outliers. SED fitting techniques, however, provide useful information on the galaxy spectral type which can be effectively used to constrain systematic errors and to better characterize potential catastrophic outliers. Such classification is then used to specialize the training of regression machine learning models, by demonstrating that a hybrid approach, involving SED fitting and machine learning in a single collaborative framework, can be effectively used to improve the accuracy of photo-z estimates.

preprint2016arXiv

An analysis of feature relevance in the classification of astronomical transients with machine learning methods

The exploitation of present and future synoptic (multi-band and multi-epoch) surveys requires an extensive use of automatic methods for data processing and data interpretation. In this work, using data extracted from the Catalina Real Time Transient Survey (CRTS), we investigate the classification performance of some well tested methods: Random Forest, MLPQNA (Multi Layer Perceptron with Quasi Newton Algorithm) and K-Nearest Neighbors, paying special attention to the feature selection phase. In order to do so, several classification experiments were performed. Namely: identification of cataclysmic variables, separation between galactic and extra-galactic objects and identification of supernovae.

preprint2016arXiv

DAMEWARE - Data Mining & Exploration Web Application Resource

Astronomy is undergoing through a methodological revolution triggered by an unprecedented wealth of complex and accurate data. DAMEWARE (DAta Mining & Exploration Web Application and REsource) is a general purpose, Web-based, Virtual Observatory compliant, distributed data mining framework specialized in massive data sets exploration with machine learning methods. We present the DAMEWARE (DAta Mining & Exploration Web Application REsource) which allows the scientific community to perform data mining and exploratory experiments on massive data sets, by using a simple web browser. DAMEWARE offers several tools which can be seen as working environments where to choose data analysis functionalities such as clustering, classification, regression, feature extraction etc., together with models and algorithms.

preprint2016arXiv

METAPHOR: A machine learning based method for the probability density estimation of photometric redshifts

A variety of fundamental astrophysical science topics require the determination of very accurate photometric redshifts (photo-z's). A wide plethora of methods have been developed, based either on template models fitting or on empirical explorations of the photometric parameter space. Machine learning based techniques are not explicitly dependent on the physical priors and able to produce accurate photo-z estimations within the photometric ranges derived from the spectroscopic training set. These estimates, however, are not easy to characterize in terms of a photo-z Probability Density Function (PDF), due to the fact that the analytical relation mapping the photometric parameters onto the redshift space is virtually unknown. We present METAPHOR (Machine-learning Estimation Tool for Accurate PHOtometric Redshifts), a method designed to provide a reliable PDF of the error distribution for empirical techniques. The method is implemented as a modular workflow, whose internal engine for photo-z estimation makes use of the MLPQNA neural network (Multi Layer Perceptron with Quasi Newton learning rule), with the possibility to easily replace the specific machine learning model chosen to predict photo-z's. We present a summary of results on SDSS-DR9 galaxy data, used also to perform a direct comparison with PDF's obtained by the Le Phare SED template fitting. We show that METAPHOR is capable to estimate the precision and reliability of photometric redshifts obtained with three different self-adaptive techniques, i.e. MLPQNA, Random Forest and the standard K-Nearest Neighbors models.

preprint2016arXiv

PhotoRaptor - Photometric Research Application To Redshifts

Due to the necessity to evaluate photo-z for a variety of huge sky survey data sets, it seemed important to provide the astronomical community with an instrument able to fill this gap. Besides the problem of moving massive data sets over the network, another critical point is that a great part of astronomical data is stored in private archives that are not fully accessible on line. So, in order to evaluate photo-z it is needed a desktop application that can be downloaded and used by everyone locally, i.e. on his own personal computer or more in general within the local intranet hosted by a data center. The name chosen for the application is PhotoRApToR, i.e. Photometric Research Application To Redshift (Cavuoti et al. 2015, 2014; Brescia 2014b). It embeds a machine learning algorithm and special tools dedicated to preand post-processing data. The ML model is the MLPQNA (Multi Layer Perceptron trained by the Quasi Newton Algorithm), which has been revealed particularly powerful for the photo-z calculation on the base of a spectroscopic sample (Cavuoti et al. 2012; Brescia et al. 2013, 2014a; Biviano et al. 2013). The PhotoRApToR program package is available, for different platforms, at the official website (http://dame.dsf.unina.it/dame_photoz.html#photoraptor).

preprint2016arXiv

The design strategy of scientific data quality control software for Euclid mission

The most valuable asset of a space mission like Euclid are the data. Due to their huge volume, the automatic quality control becomes a crucial aspect over the entire lifetime of the experiment. Here we focus on the design strategy for the Science Ground Segment (SGS) Data Quality Common Tools (DQCT), which has the main role to provide software solutions to gather, evaluate, and record quality information about the raw and derived data products from a primarily scientific perspective. The SGS DQCT will provide a quantitative basis for evaluating the application of reduction and calibration reference data, as well as diagnostic tools for quality parameters, flags, trend analysis diagrams and any other metadata parameter produced by the pipeline. In a large programme like Euclid, it is prohibitively expensive to process large amount of data at the pixel level just for the purpose of quality evaluation. Thus, all measures of quality at the pixel level are implemented in the individual pipeline stages, and passed along as metadata in the production. In this sense most of the tasks related to science data quality are delegated to the pipeline stages, even though the responsibility for science data quality is managed at a higher level. The DQCT subsystem of the SGS is currently under development, but its path to full realization will likely be different than that of other subsystems. Primarily because, due to a high level of parallelism and to the wide pipeline processing redundancy, for instance the mechanism of double Science Data Center for each processing function, the data quality tools have not only to be widely spread over all pipeline segments and data levels, but also to minimize the occurrences of potential diversity of solutions implemented for similar functions, ensuring the maximum of coherency and standardization for quality evaluation and reporting in the SGS.

preprint2015arXiv

Automated physical classification in the SDSS DR10. A catalogue of candidate Quasars

We discuss whether modern machine learning methods can be used to characterize the physical nature of the large number of objects sampled by the modern multi-band digital surveys. In particular, we applied the MLPQNA (Multi Layer Perceptron with Quasi Newton Algorithm) method to the optical data of the Sloan Digital Sky Survey - Data Release 10, investigating whether photometric data alone suffice to disentangle different classes of objects as they are defined in the SDSS spectroscopic classification. We discuss three groups of classification problems: (i) the simultaneous classification of galaxies, quasars and stars; (ii) the separation of stars from quasars; (iii) the separation of galaxies with normal spectral energy distribution from those with peculiar spectra, such as starburst or starforming galaxies and AGN. While confirming the difficulty of disentangling AGN from normal galaxies on a photometric basis only, MLPQNA proved to be quite effective in the three-class separation. In disentangling quasars from stars and galaxies, our method achieved an overall efficiency of 91.31% and a QSO class purity of ~95%. The resulting catalogue of candidate quasars/AGNs consists of ~3.6 million objects, of which about half a million are also flagged as robust candidates, and will be made available on CDS VizieR facility.

preprint2015arXiv

Feature Selection based on Machine Learning in MRIs for Hippocampal Segmentation

Neurodegenerative diseases are frequently associated with structural changes in the brain. Magnetic Resonance Imaging (MRI) scans can show these variations and therefore be used as a supportive feature for a number of neurodegenerative diseases. The hippocampus has been known to be a biomarker for Alzheimer disease and other neurological and psychiatric diseases. However, it requires accurate, robust and reproducible delineation of hippocampal structures. Fully automatic methods are usually the voxel based approach, for each voxel a number of local features were calculated. In this paper we compared four different techniques for feature selection from a set of 315 features extracted for each voxel: (i) filter method based on the Kolmogorov-Smirnov test; two wrapper methods, respectively, (ii) Sequential Forward Selection and (iii) Sequential Backward Elimination; and (iv) embedded method based on the Random Forest Classifier on a set of 10 T1-weighted brain MRIs and tested on an independent set of 25 subjects. The resulting segmentations were compared with manual reference labelling. By using only 23 features for each voxel (sequential backward elimination) we obtained comparable state of-the-art performances with respect to the standard tool FreeSurfer.

preprint2015arXiv

Machine Learning based photometric redshifts for the KiDS ESO DR2 galaxies

We estimated photometric redshifts (zphot) for more than 1.1 million galaxies of the ESO Public Kilo-Degree Survey (KiDS) Data Release 2. KiDS is an optical wide-field imaging survey carried out with the VLT Survey Telescope (VST) and the OmegaCAM camera, which aims at tackling open questions in cosmology and galaxy evolution, such as the origin of dark energy and the channel of galaxy mass growth. We present a catalogue of photometric redshifts obtained using the Multi Layer Perceptron with Quasi Newton Algorithm (MLPQNA) model, provided within the framework of the DAta Mining and Exploration Web Application REsource (DAMEWARE). These photometric redshifts are based on a spectroscopic knowledge base which was obtained by merging spectroscopic datasets from GAMA (Galaxy And Mass Assembly) data release 2 and SDSS-III data release 9. The overall 1 sigma uncertainty on Delta z = (zspec - zphot) / (1+ zspec) is ~ 0.03, with a very small average bias of ~ 0.001, a NMAD of ~ 0.02 and a fraction of catastrophic outliers (| Delta z | > 0.15) of ~0.4%.

preprint2015arXiv

Mapping the Galaxy Color-Redshift Relation: Optimal Photometric Redshift Calibration Strategies for Cosmology Surveys

Calibrating the photometric redshifts of >10^9 galaxies for upcoming weak lensing cosmology experiments is a major challenge for the astrophysics community. The path to obtaining the required spectroscopic redshifts for training and calibration is daunting, given the anticipated depths of the surveys and the difficulty in obtaining secure redshifts for some faint galaxy populations. Here we present an analysis of the problem based on the self-organizing map, a method of mapping the distribution of data in a high-dimensional space and projecting it onto a lower-dimensional representation. We apply this method to existing photometric data from the COSMOS survey selected to approximate the anticipated Euclid weak lensing sample, enabling us to robustly map the empirical distribution of galaxies in the multidimensional color space defined by the expected Euclid filters. Mapping this multicolor distribution lets us determine where - in galaxy color space - redshifts from current spectroscopic surveys exist and where they are systematically missing. Crucially, the method lets us determine whether a spectroscopic training sample is representative of the full photometric space occupied by the galaxies in a survey. We explore optimal sampling techniques and estimate the additional spectroscopy needed to map out the color-redshift relation, finding that sampling the galaxy distribution in color space in a systematic way can efficiently meet the calibration requirements. While the analysis presented here focuses on the Euclid survey, similar analysis can be applied to other surveys facing the same calibration challenge, such as DES, LSST, and WFIRST.

preprint2015arXiv

Photometric redshift estimation based on data mining with PhotoRApToR

Photometric redshifts (photo-z) are crucial to the scientific exploitation of modern panchromatic digital surveys. In this paper we present PhotoRApToR (Photometric Research Application To Redshift): a Java/C++ based desktop application capable to solve non-linear regression and multi-variate classification problems, in particular specialized for photo-z estimation. It embeds a machine learning algorithm, namely a multilayer neural network trained by the Quasi Newton learning rule, and special tools dedicated to pre- and postprocessing data. PhotoRApToR has been successfully tested on several scientific cases. The application is available for free download from the DAME Program web site.

preprint2015arXiv

The first and second data releases of the Kilo-Degree Survey

The Kilo-Degree Survey (KiDS) is an optical wide-field imaging survey carried out with the VLT Survey Telescope and the OmegaCAM camera. KiDS will image 1500 square degrees in four filters (ugri), and together with its near-infrared counterpart VIKING will produce deep photometry in nine bands. Designed for weak lensing shape and photometric redshift measurements, the core science driver of the survey is mapping the large-scale matter distribution in the Universe back to a redshift of ~0.5. Secondary science cases are manifold, covering topics such as galaxy evolution, Milky Way structure, and the detection of high-redshift clusters and quasars. KiDS is an ESO Public Survey and dedicated to serving the astronomical community with high-quality data products derived from the survey data, as well as with calibration data. Public data releases will be made on a yearly basis, the first two of which are presented here. For a total of 148 survey tiles (~160 sq.deg.) astrometrically and photometrically calibrated, coadded ugri images have been released, accompanied by weight maps, masks, source lists, and a multi-band source catalog. A dedicated pipeline and data management system based on the Astro-WISE software system, combined with newly developed masking and source classification software, is used for the data production of the data products described here. The achieved data quality and early science projects based on the data products in the first two data releases are reviewed in order to validate the survey data. Early scientific results include the detection of nine high-z QSOs, fifteen candidate strong gravitational lenses, high-quality photometric redshifts and galaxy structural parameters for hundreds of thousands of galaxies. (Abridged)

preprint2014arXiv

DAMEWARE: A web cyberinfrastructure for astrophysical data mining

Astronomy is undergoing through a methodological revolution triggered by an unprecedented wealth of complex and accurate data. The new panchromatic, synoptic sky surveys require advanced tools for discovering patterns and trends hidden behind data which are both complex and of high dimensionality. We present DAMEWARE (DAta Mining & Exploration Web Application REsource): a general purpose, web-based, distributed data mining environment developed for the exploration of large datasets, and finely tuned for astronomical applications. By means of graphical user interfaces, it allows the user to perform classification, regression or clustering tasks with machine learning methods. Salient features of DAMEWARE include its capability to work on large datasets with minimal human intervention, and to deal with a wide variety of real problems such as the classification of globular clusters in the galaxy NGC1399, the evaluation of photometric redshifts and, finally, the identification of candidate Active Galactic Nuclei in multiband photometric surveys. In all these applications, DAMEWARE allowed to achieve better results than those attained with more traditional methods. With the aim of providing potential users with all needed information, in this paper we briefly describe the technological background of DAMEWARE, give a short introduction to some relevant aspects of data mining, followed by a summary of some science cases and, finally, we provide a detailed description of a template use case.

preprint2014arXiv

Data-Rich Astronomy: Mining Sky Surveys with PhotoRApToR

In the last decade a new generation of telescopes and sensors has allowed the production of a very large amount of data and astronomy has become a data-rich science. New automatic methods largely based on machine learning are needed to cope with such data tsunami. We present some results in the fields of photometric redshifts and galaxy classification, obtained using the MLPQNA algorithm available in the DAMEWARE (Data Mining and Web Application Resource) for the SDSS galaxies (DR9 and DR10). We present PhotoRApToR (Photometric Research Application To Redshift): a Java based desktop application capable to solve regression and classification problems and specialized for photo-z estimation.

preprint2014arXiv

Immersive and Collaborative Data Visualization Using Virtual Reality Platforms

Effective data visualization is a key part of the discovery process in the era of big data. It is the bridge between the quantitative content of the data and human intuition, and thus an essential component of the scientific path from data into knowledge and understanding. Visualization is also essential in the data mining process, directing the choice of the applicable algorithms, and in helping to identify and remove bad data from the analysis. However, a high complexity or a high dimensionality of modern data sets represents a critical obstacle. How do we visualize interesting structures and patterns that may exist in hyper-dimensional data spaces? A better understanding of how we can perceive and interact with multi dimensional information poses some deep questions in the field of cognition technology and human computer interaction. To this effect, we are exploring the use of immersive virtual reality platforms for scientific data visualization, both as software and inexpensive commodity hardware. These potentially powerful and innovative tools for multi dimensional data visualization can also provide an easy and natural path to a collaborative data visualization and exploration, where scientists can interact with their data and their colleagues in the same visual space. Immersion provides benefits beyond the traditional desktop visualization tools: it leads to a demonstrably better perception of a datascape geometry, more intuitive data understanding, and a better retention of the perceived relationships in the data.

preprint2013arXiv

Astrophysical data mining with GPU. A case study: genetic classification of globular clusters

We present a multi-purpose genetic algorithm, designed and implemented with GPGPU / CUDA parallel computing technology. The model was derived from our CPU serial implementation, named GAME (Genetic Algorithm Model Experiment). It was successfully tested and validated on the detection of candidate Globular Clusters in deep, wide-field, single band HST images. The GPU version of GAME will be made available to the community by integrating it into the web application DAMEWARE (DAta Mining Web Application REsource (http://dame.dsf.unina.it/beta_info.html), a public data mining service specialized on massive astrophysical data. Since genetic algorithms are inherently parallel, the GPGPU computing paradigm leads to a speedup of a factor of 200x in the training phase with respect to the CPU based version.

preprint2013arXiv

Feature Selection Strategies for Classifying High Dimensional Astronomical Data Sets

The amount of collected data in many scientific fields is increasing, all of them requiring a common task: extract knowledge from massive, multi parametric data sets, as rapidly and efficiently possible. This is especially true in astronomy where synoptic sky surveys are enabling new research frontiers in the time domain astronomy and posing several new object classification challenges in multi dimensional spaces; given the high number of parameters available for each object, feature selection is quickly becoming a crucial task in analyzing astronomical data sets. Using data sets extracted from the ongoing Catalina Real-Time Transient Surveys (CRTS) and the Kepler Mission we illustrate a variety of feature selection strategies used to identify the subsets that give the most information and the results achieved applying these techniques to three major astronomical problems.

preprint2013arXiv

Photometric classification of emission line galaxies with Machine Learning methods

In this paper we discuss an application of machine learning based methods to the identification of candidate AGN from optical survey data and to the automatic classification of AGNs in broad classes. We applied four different machine learning algorithms, namely the Multi Layer Perceptron (MLP), trained respectively with the Conjugate Gradient, Scaled Conjugate Gradient and Quasi Newton learning rules, and the Support Vector Machines (SVM), to tackle the problem of the classification of emission line galaxies in different classes, mainly AGNs vs non-AGNs, obtained using optical photometry in place of the diagnostics based on line intensity ratios which are classically used in the literature. Using the same photometric features we discuss also the behavior of the classifiers on finer AGN classification tasks, namely Seyfert I vs Seyfert II and Seyfert vs LINER. Furthermore we describe the algorithms employed, the samples of spectroscopically classified galaxies used to train the algorithms, the procedure followed to select the photometric parameters and the performances of our methods in terms of multiple statistical indicators. The results of the experiments show that the application of self adaptive data mining algorithms trained on spectroscopic data sets and applied to carefully chosen photometric parameters represents a viable alternative to the classical methods that employ time-consuming spectroscopic observations.

preprint2013arXiv

The MICA Experiment: Astrophysics in Virtual Worlds

We describe the work of the Meta-Institute for Computational Astrophysics (MICA), the first professional scientific organization based in virtual worlds. MICA was an experiment in the use of this technology for science and scholarship, lasting from the early 2008 to June 2012, mainly using the Second Life and OpenSimulator as platforms. We describe its goals and activities, and our future plans. We conducted scientific collaboration meetings, professional seminars, a workshop, classroom instruction, public lectures, informal discussions and gatherings, and experiments in immersive, interactive visualization of high-dimensional scientific data. Perhaps the most successful of these was our program of popular science lectures, illustrating yet again the great potential of immersive VR as an educational and outreach platform. While the members of our research groups and some collaborators found the use of immersive VR as a professional telepresence tool to be very effective, we did not convince a broader astrophysics community to adopt it at this time, despite some efforts; we discuss some possible reasons for this non-uptake. On the whole, we conclude that immersive VR has a great potential as a scientific and educational platform, as the technology matures and becomes more broadly available and accepted.

preprint2012arXiv

A 2-dimensional Geometry for Biological Time

This paper proposes an abstract mathematical frame for describing some features of biological time. The key point is that usual physical (linear) representation of time is insufficient, in our view, for the understanding key phenomena of life, such as rhythms, both physical (circadian, seasonal ...) and properly biological (heart beating, respiration, metabolic ...). In particular, the role of biological rhythms do not seem to have any counterpart in mathematical formalization of physical clocks, which are based on frequencies along the usual (possibly thermodynamical, thus oriented) time. We then suggest a functional representation of biological time by a 2-dimensional manifold as a mathematical frame for accommodating autonomous biological rhythms. The "visual" representation of rhythms so obtained, in particular heart beatings, will provide, by a few examples, hints towards possible applications of our approach to the understanding of interspecific differences or intraspecific pathologies. The 3-dimensional embedding space, needed for purely mathematical reasons, allows to introduce a suitable extra-dimension for "representation time", with a cognitive significance.

preprint2012arXiv

Astroinformatics, data mining and the future of astronomical research

Astronomy, as many other scientific disciplines, is facing a true data deluge which is bound to change both the praxis and the methodology of every day research work. The emerging field of astroinformatics, while on the one end appears crucial to face the technological challenges, on the other is opening new exciting perspectives for new astronomical discoveries through the implementation of advanced data mining procedures. The complexity of astronomical data and the variety of scientific problems, however, call for innovative algorithms and methods as well as for an extreme usage of ICT technologies.

preprint2012arXiv

Characterising galaxy groups: spectroscopic observations of the Shakhbazyan sample

Groups of galaxies are the most common cosmic structures. However, due to the poor statistics, projection effects and the lack of accurate distances, our understanding of their dynamical and evolutionary status is still limited. This is particularly true for the so called Shakhbazyan groups (SHK) which are still largely unexplored due to the lack of systematic spectroscopic studies of both their member galaxies and the surrounding environment. In our previous paper, we investigated the statistical properties of a large sample of SHK groups using Sloan Digital Sky Survey data and photometric redshifts. Here we present the follow-up of 5 SHK groups (SHK 10, 71, 75, 80, 259) observed within our spectroscopic campaign with the Telescopio Nazionale Galileo, aimed at confirming their physical reality and strengthening our photometric results. For each of the selected groups we were able to identify between 6 and 13 spectroscopic members, thus confirming the robustness of the photometric redshift approach in identifying real galaxy over-densities. Consistently with the finding of our previous paper, the structures studied here have properties spanning from those of compact and isolated groups to those of loose groups. For what the global physical properties are concerned (total mass, mass-to-light ratios, etc.), we find systematic differences with those reported in the literature by previous studies. Our analysis suggests that previous results should be revisited; we show in fact that, if the literature data are re-analysed in a consistent and homogeneous way, the properties obtained are in agreement with those estimated for our sample.

preprint2012arXiv

Constraints on sterile neutrino dark matter from XMM--Newton observation of M33

Using archival XMM-Newton observations of the diffuse and unresolved components emission in the inner disc of M33 we exclude the possible contribution from narrow line emission in the energy range 0.5-5 keV more intense than 10^{-6}-10^{-5} erg/s. Under the hypothesis that sterile neutrinos constitute the majority of the dark matter in M33, we use this result in order to put constraints on their parameter space in the 1-10 keV mass range.

preprint2012arXiv

Data challenges of time domain astronomy

Astronomy has been at the forefront of the development of the techniques and methodologies of data intensive science for over a decade with large sky surveys and distributed efforts such as the Virtual Observatory. However, it faces a new data deluge with the next generation of synoptic sky surveys which are opening up the time domain for discovery and exploration. This brings both new scientific opportunities and fresh challenges, in terms of data rates from robotic telescopes and exponential complexity in linked data, but also for data mining algorithms used in classification and decision making. In this paper, we describe how an informatics-based approach-part of the so-called "fourth paradigm" of scientific discovery-is emerging to deal with these. We review our experiences with the Palomar-Quest and Catalina Real-Time Transient Sky Surveys; in particular, addressing the issue of the heterogeneity of data associated with transient astronomical events (and other sensor networks) and how to manage and analyze it.

preprint2012arXiv

From Physics to Biology by Extending Criticality and Symmetry Breakings

Symmetries play a major role in physics, in particular since the work by E. Noether and H. Weyl in the first half of last century. Herein, we briefly review their role by recalling how symmetry changes allow to conceptually move from classical to relativistic and quantum physics. We then introduce our ongoing theoretical analysis in biology and show that symmetries play a radically different role in this discipline, when compared to those in current physics. By this comparison, we stress that symmetries must be understood in relation to conservation and stability properties, as represented in the theories. We posit that the dynamics of biological organisms, in their various levels of organization, are not "just" processes, but permanent (extended, in our terminology) critical transitions and, thus, symmetry changes. Within the limits of a relative structural stability (or interval of viability), variability is at the core of these transitions.

preprint2012arXiv

Genetic Algorithm Modeling with GPU Parallel Computing Technology

We present a multi-purpose genetic algorithm, designed and implemented with GPGPU / CUDA parallel computing technology. The model was derived from a multi-core CPU serial implementation, named GAME, already scientifically successfully tested and validated on astrophysical massive data classification problems, through a web application resource (DAMEWARE), specialized in data mining based on Machine Learning paradigms. Since genetic algorithms are inherently parallel, the GPGPU computing paradigm has provided an exploit of the internal training features of the model, permitting a strong optimization in terms of processing performances and scalability.

preprint2012arXiv

No entailing laws, but enablement in the evolution of the biosphere

Biological evolution is a complex blend of ever changing structural stability, variability and emergence of new phenotypes, niches, ecosystems. We wish to argue that the evolution of life marks the end of a physics world view of law entailed dynamics. Our considerations depend upon discussing the variability of the very "contexts of life": the interactions between organisms, biological niches and ecosystems. These are ever changing, intrinsically indeterminate and even unprestatable: we do not know ahead of time the "niches" which constitute the boundary conditions on selection. More generally, by the mathematical unprestatability of the "phase space" (space of possibilities), no laws of motion can be formulated for evolution. We call this radical emergence, from life to life. The purpose of this paper is the integration of variation and diversity in a sound conceptual frame and situate unpredictability at a novel theoretical level, that of the very phase space. Our argument will be carried on in close comparisons with physics and the mathematical constructions of phase spaces in that discipline. The role of (theoretical) symmetries as invariant preserving transformations will allow us to understand the nature of physical phase spaces and to stress the differences required for a sound biological theoretizing. In this frame, we discuss the novel notion of "enablement". This will restrict causal analyses to differential cases (a difference that causes a difference). Mutations or other causal differences will allow us to stress that "non conservation principles" are at the core of evolution, in contrast to physical dynamics, largely based on conservation principles as symmetries. Critical transitions, the main locus of symmetry changes in physics, will be discussed, and lead to "extended criticality" as a conceptual frame for a better understanding of the living state of matter.

preprint2012arXiv

Photometric redshifts with Quasi Newton Algorithm (MLPQNA). Results in the PHAT1 contest

Context. Since the advent of modern multiband digital sky surveys, photometric redshifts (photo-z's) have become relevant if not crucial to many fields of observational cosmology, from the characterization of cosmic structures, to weak and strong lensing. Aims. We describe an application to an astrophysical context, namely the evaluation of photometric redshifts, of MLPQNA, a machine learning method based on Quasi Newton Algorithm. Methods. Theoretical methods for photo-z's evaluation are based on the interpolation of a priori knowledge (spectroscopic redshifts or SED templates) and represent an ideal comparison ground for neural networks based methods. The MultiLayer Perceptron with Quasi Newton learning rule (MLPQNA) described here is a computing effective implementation of Neural Networks for the first time exploited to solve regression problems in the astrophysical context and is offered to the community through the DAMEWARE (DAta Mining & ExplorationWeb Application REsource) infrastructure. Results. The PHAT contest (Hildebrandt et al. 2010) provides a standard dataset to test old and new methods for photometric redshift evaluation and with a set of statistical indicators which allow a straightforward comparison among different methods. The MLPQNA model has been applied on the whole PHAT1 dataset of 1984 objects after an optimization of the model performed by using as training set the 515 available spectroscopic redshifts. When applied to the PHAT1 dataset, MLPQNA obtains the best bias accuracy (0.0006) and very competitive accuracies in terms of scatter (0.056) and outlier percentage (16.3%), scoring as the second most effective empirical method among those which have so far participated to the contest. MLPQNA shows better generalization capabilities than most other empirical methods especially in presence of underpopulated regions of the Knowledge Base.

preprint2012arXiv

Protention and retention in biological systems

This paper proposes an abstract mathematical frame for describing some features of cognitive and biological time. We focus here on the so called "extended present" as a result of protentional and retentional activities (memory and anticipation). Memory, as retention, is treated in some physical theories (relaxation phenomena, which will inspire our approach), while protention (or anticipation) seems outside the scope of physics. We then suggest a simple functional representation of biological protention. This allows us to introduce the abstract notion of "biological inertia".

preprint2011arXiv

Astroinformatics of galaxies and quasars: a new general method for photometric redshifts estimation

With the availability of the huge amounts of data produced by current and future large multi-band photometric surveys, photometric redshifts have become a crucial tool for extragalactic astronomy and cosmology. In this paper we present a novel method, called Weak Gated Experts (WGE), which allows to derive photometric redshifts through a combination of data mining techniques. \noindent The WGE, like many other machine learning techniques, is based on the exploitation of a spectroscopic knowledge base composed by sources for which a spectroscopic value of the redshift is available. This method achieves a variance σ^2(Δz)=2.3x10^{-4} (σ^2(Δz) =0.08), where Δz = z_{phot} - z_{spec}) for the reconstruction of the photometric redshifts for the optical galaxies from the SDSS and for the optical quasars respectively, while the Root Mean Square (RMS) of the Δz variable distributions for the two experiments is respectively equal to 0.021 and 0.35. The WGE provides also a mechanism for the estimation of the accuracy of each photometric redshift. We also present and discuss the catalogs obtained for the optical SDSS galaxies, for the optical candidate quasars extracted from the DR7 SDSS photometric dataset {The sample of SDSS sources on which the accuracy of the reconstruction has been assessed is composed of bright sources, for a subset of which spectroscopic redshifts have been measured.}, and for optical SDSS candidate quasars observed by GALEX in the UV range. The WGE method exploits the new technological paradigm provided by the Virtual Observatory and the emerging field of Astroinformatics.

preprint2011arXiv

Extracting Knowledge From Massive Astronomical Data Sets

The exponential growth of astronomical data collected by both ground based and space borne instruments has fostered the growth of Astroinformatics: a new discipline laying at the intersection between astronomy, applied computer science, and information and computation (ICT) technologies. At the very heart of Astroinformatics is a complex set of methodologies usually called Data Mining (DM) or Knowledge Discovery in Data Bases (KDD). In the astronomical domain, DM/KDD are still in a very early usage stage, even though new methods and tools are being continuously deployed in order to cope with the Massive Data Sets (MDS) that can only grow in the future. In this paper, we briefly outline some general problems encountered when applying DM/KDD methods to astrophysical problems, and describe the DAME (DAta Mining & Exploration) web application. While specifically tailored to work on MDS, DAME can be effectively applied also to smaller data sets. As an illustration, we describe two application of DAME to two different problems: the identification of candidate globular clusters in external galaxies, and the classification of active galactic nuclei (AGN). We believe that tools and services of this nature will become increasingly necessary for the data-intensive astronomy (and indeed all sciences) in the 21st century.

preprint2011arXiv

Randomness and Multi-level Interactions in Biology

The dynamic instability of the living systems and the "superposition" of different forms of randomness are viewed as a component of the contingently increasing organization of life along evolution. We briefly survey how classical and quantum physics define randomness differently. We then discuss why this requires, in our view, an enrichment of the understanding of the effects of their concurrent presence in biological systems' dynamics. Biological randomness is then presented as an essential component of the heterogeneous determination and intrinsic unpredictability proper to life phenomena, due to the nesting and interaction of many levels of organization. Even increasing organization itself induces growing disorder, by energy dispersal effects of course, but also by variability and differentiation. Co-operation between diverse components in networks implies at the same time the presence of constraints due to the peculiar forms of (bio-)resonance and (bio-)entanglement we discuss.

preprint2010arXiv

DAME: A Web Oriented Infrastructure for Scientific Data Mining & Exploration

Nowadays, many scientific areas share the same need of being able to deal with massive and distributed datasets and to perform on them complex knowledge extraction tasks. This simple consideration is behind the international efforts to build virtual organizations such as, for instance, the Virtual Observatory (VObs). DAME (DAta Mining & Exploration) is an innovative, general purpose, Web-based, VObs compliant, distributed data mining infrastructure specialized in Massive Data Sets exploration with machine learning methods. Initially fine tuned to deal with astronomical data only, DAME has evolved in a general purpose platform which has found applications also in other domains of human endeavor. We present the products and a short outline of a science case, together with a detailed description of main features available in the beta release of the web application now released.

preprint2010arXiv

Radio emission from dark matter annihilation in the Large Magellanic Cloud

The Large Magellanic Cloud, at only 50 kpc away from us and known to be dark matter dominated, is clearly an interesting place where to search for dark matter annihilation signals. In this paper, we estimate the synchrotron emission due to WIMP annihilation in the halo of the LMC at two radio frequencies, 1.4 and 4.8 GHz, and compare it to the observed emission, in order to impose constraints in the WIMP mass vs. annihilation cross section plane. We use available Faraday rotation data from background sources to estimate the magnitude of the magnetic field in different regions of the LMC's disc, where we calculate the radio signal due to dark matter annihilation. We account for the e+ e- energy losses due to synchrotron, Inverse Compton Scattering and bremsstrahlung, using the observed hydrogen and dust temperature distribution on the LMC to estimate their efficiency. The extensive use of observations, allied with conservative choices adopted in all the steps of the calculation, allow us to obtain very realistic constraints.

preprint2010arXiv

The Evolution and Development of the Universe

This document is the Special Issue of the First International Conference on the Evolution and Development of the Universe (EDU 2008). Please refer to the preface and introduction for more details on the contributions. Keywords: acceleration, artificial cosmogenesis, artificial life, Big Bang, Big History, biological evolution, biological universe, biology, causality, classical vacuum energy, complex systems, complexity, computational universe, conscious evolution, cosmological artificial selection, cosmological natural selection, cosmology, critique, cultural evolution, dark energy, dark matter, development of the universe, development, emergence, evolution of the universe evolution, exobiology, extinction, fine-tuning, fractal space-time, fractal, information, initial conditions, intentional evolution, linear expansion of the universe, log-periodic laws, macroevolution, materialism, meduso-anthropic principle, multiple worlds, natural sciences, Nature, ontology, order, origin of the universe, particle hierarchy, philosophy, physical constants, quantum darwinism, reduction, role of intelligent life, scale relativity, scientific evolution, self-organization, speciation, specification hierarchy, thermodynamics, time, universe, vagueness.

preprint2009arXiv

Searching for Dark Matter in Messier 33

Among various approaches for indirect detection of dark matter, synchrotron emission due to secondary electrons/positrons produced in galactic WIMPs annihilation is raising an increasing interest. In this paper we propose a new method to derive bounds in the mchi - <sigmaA*v> plane by using radio continuum observations of Messier 33, paying particular attention to a low emitting Radio Cavity. The comparison of the expected radio emission due to the galactic dark matter distribution with the observed one provides bounds which are comparable to those obtained from a similar analysis of the Milky Way. Remarkably, the present results are simply based on archival data and thus largely improvable by means of specifically tailored observations. The potentiality of the method compared with more standard searches is discussed by considering the optimistic situation of a vanishing flux (within the experimental sensitivity) measured inside the cavity by a high resolution radio telescope like ALMA. Under the best conditions our technique is able to produce bounds which are comparable to the ones expected after five years of Fermi LAT data taking for an hadronic annihilation channel. Furthermore, it allows to test the hypothesis that space telescopes like Pamela and Fermi LAT are actually observing electrons and positrons due to galactic dark matter annihilation into leptons.

preprint2002arXiv

Neural Networks and Photometric Redshifts

We present a neural network based approach to the determination of photometric redshift. The method was tested on the Sloan Digital Sky Survey Early Data Release (SDSS-EDR) reaching an accuracy comparable and, in some cases, better than SED template fitting techniques. Different neural networks architecture have been tested and the combination of a Multi Layer Perceptron with 1 hidden layer (22 neurons) operated in a Bayesian framework, with a Self Organizing Map used to estimate the accuracy of the results, turned out to be the most effective. In the best experiment, the implemented network reached an accuracy of 0.020 (interquartile error) in the range 0<zphot<0.3, and of 0.022 in the range 0<zphot<0.5.