Source author record

Massimo Caccia

Massimo Caccia appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Machine Learning Artificial Intelligence Computer Vision physics.ins-det Applications astro-ph.IM astro-ph.SR Computation and Language math.OC physics.flu-dyn quant-ph Robotics

Catalog footprint

What is connected

10works

12topics

4close collaborators

Actions

Connect this record

Open graph Browse works

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

preprint2022arXiv

Multi-fidelity hydrodynamic analysis of an autonomous surface vehicle at surveying speed in deep water subject to variable payload

Autonomous surface vehicles (ASV) allow the investigation of coastal areas, ports and harbors as well as harsh and dangerous environments such as the arctic regions. Despite receiving increasing attention, the hydrodynamic analysis of ASV performance subject to variable operational parameters is little investigated. In this context, this paper presents a multi-fidelity (MF) hydrodynamic analysis of an ASV, namely the Shallow Water Autonomous Multipurpose Platform (SWAMP), at surveying speed in calm water and subject to variable payload and location of the center of mass, accounting for the variety of equipment that the vehicle can carry. The analysis is conducted in deep water, which is the condition mostly encountered by the ASV during surveys of coastal and harbors areas. Quantities of interest are the resistance, the vehicle attitude, and the wave generated in the region between the catamaran hulls. These are assessed using a Reynolds Averaged Navier Stokes Equation (RANSE) code and a linear potential flow (PF) solver. The objective is to accurately assess the quantities of interest, along with identifying the limitation of PF analysis in the current context. Finally, a multi-fidelity Gaussian Process (MF-GP) model is obtained combining RANSE and PF solutions. The latter also include variable grid refinement and coupling between hydrodynamic loads and rigid body equations of motion. The surrogate model is iteratively refined using an active learning approach. Numerical results show that the MF-GP is effective in producing response surfaces of the SWAMP performance with a limited computational cost. It is highlighted how the SWAMP performance is significantly affected not only by the payload, but also by the location of the center of mass. The latter can be therefore properly calibrated to minimize the resistance and allow for longer-range operations.

preprint2022arXiv

Understanding Continual Learning Settings with Data Distribution Drift Analysis

Classical machine learning algorithms often assume that the data are drawn i.i.d. from a stationary probability distribution. Recently, continual learning emerged as a rapidly growing area of machine learning where this assumption is relaxed, i.e. where the data distribution is non-stationary and changes over time. This paper represents the state of data distribution by a context variable $c$. A drift in $c$ leads to a data distribution drift. A context drift may change the target distribution, the input distribution, or both. Moreover, distribution drifts might be abrupt or gradual. In continual learning, context drifts may interfere with the learning process and erase previously learned knowledge; thus, continual learning algorithms must include specialized mechanisms to deal with such drifts. In this paper, we aim to identify and categorize different types of context drifts and potential assumptions about them, to better characterize various continual-learning scenarios. Moreover, we propose to use the distribution drift framework to provide more precise definitions of several terms commonly used in the continual learning field.

preprint2021arXiv

Online Fast Adaptation and Knowledge Accumulation: a New Approach to Continual Learning

Continual learning studies agents that learn from streams of tasks without forgetting previous ones while adapting to new ones. Two recent continual-learning scenarios have opened new avenues of research. In meta-continual learning, the model is pre-trained to minimize catastrophic forgetting of previous tasks. In continual-meta learning, the aim is to train agents for faster remembering of previous tasks through adaptation. In their original formulations, both methods have limitations. We stand on their shoulders to propose a more general scenario, OSAKA, where an agent must quickly solve new (out-of-distribution) tasks, while also requiring fast remembering. We show that current continual learning, meta-learning, meta-continual learning, and continual-meta learning techniques fail in this new scenario. We propose Continual-MAML, an online extension of the popular MAML algorithm as a strong baseline for this scenario. We empirically show that Continual-MAML is better suited to the new scenario than the aforementioned methodologies, as well as standard continual learning and meta-learning approaches.

preprint2020arXiv

CVPR 2020 Continual Learning in Computer Vision Competition: Approaches, Results, Current Challenges and Future Directions

In the last few years, we have witnessed a renewed and fast-growing interest in continual learning with deep neural networks with the shared objective of making current AI systems more adaptive, efficient and autonomous. However, despite the significant and undoubted progress of the field in addressing the issue of catastrophic forgetting, benchmarking different continual learning approaches is a difficult task by itself. In fact, given the proliferation of different settings, training and evaluation protocols, metrics and nomenclature, it is often tricky to properly characterize a continual learning algorithm, relate it to other solutions and gauge its real-world applicability. The first Continual Learning in Computer Vision challenge held at CVPR in 2020 has been one of the first opportunities to evaluate different continual learning algorithms on a common hardware with a large set of shared evaluation metrics and 3 different settings based on the realistic CORe50 video benchmark. In this paper, we report the main results of the competition, which counted more than 79 teams registered, 11 finalists and 2300$ in prizes. We also summarize the winning approaches, current challenges and future research directions.

preprint2020arXiv

Language GANs Falling Short

Generating high-quality text with sufficient diversity is essential for a wide range of Natural Language Generation (NLG) tasks. Maximum-Likelihood (MLE) models trained with teacher forcing have consistently been reported as weak baselines, where poor performance is attributed to exposure bias (Bengio et al., 2015; Ranzato et al., 2015); at inference time, the model is fed its own prediction instead of a ground-truth token, which can lead to accumulating errors and poor samples. This line of reasoning has led to an outbreak of adversarial based approaches for NLG, on the account that GANs do not suffer from exposure bias. In this work, we make several surprising observations which contradict common beliefs. First, we revisit the canonical evaluation framework for NLG, and point out fundamental flaws with quality-only evaluation: we show that one can outperform such metrics using a simple, well-known temperature parameter to artificially reduce the entropy of the model's conditional distributions. Second, we leverage the control over the quality / diversity trade-off given by this parameter to evaluate models over the whole quality-diversity spectrum and find MLE models constantly outperform the proposed GAN variants over the whole quality-diversity space. Our results have several implications: 1) The impact of exposure bias on sample quality is less severe than previously thought, 2) temperature tuning provides a better quality / diversity trade-off than adversarial training while being easier to train, easier to cross-validate, and less computationally expensive. Code to reproduce the experiments is available at github.com/pclucas14/GansFallingShort

preprint2020arXiv

Online Learned Continual Compression with Adaptive Quantization Modules

We introduce and study the problem of Online Continual Compression, where one attempts to simultaneously learn to compress and store a representative dataset from a non i.i.d data stream, while only observing each sample once. A naive application of auto-encoders in this setting encounters a major challenge: representations derived from earlier encoder states must be usable by later decoder states. We show how to use discrete auto-encoders to effectively address this challenge and introduce Adaptive Quantization Modules (AQM) to control variation in the compression ability of the module at any given stage of learning. This enables selecting an appropriate compression for incoming samples, while taking into account overall memory constraints and current progress of the learned compression. Unlike previous methods, our approach does not require any pretraining, even on challenging datasets. We show that using AQM to replace standard episodic memory in continual learning settings leads to significant gains on continual learning benchmarks. Furthermore we demonstrate this approach with larger images, LiDAR, and reinforcement learning environments.

preprint2016arXiv

Adaptive Experimental Design for Path-following Performance Assessment of Unmanned Vehicles

The definition of Good Experimental Methodologies (GEMs) in robotics is a topic of widespread interest due also to the increasing employment of robots in everyday civilian life. The present work contributes to the ongoing discussion on GEMs for Unmanned Surface Vehicles (USVs). It focuses on the definition of GEMs and provides specific guidelines for path-following experiments. Statistically designed experiments (DoE) offer a valid basis for developing an empirical model of the system being investigated. A two-step adaptive experimental procedure for evaluating path-following performance and based on DoE, is tested on the simulator of the Charlie USV. The paper argues the necessity of performing extensive simulations prior to the execution of field trials.

preprint2014arXiv

A simple and robust method to study after-pulses in Silicon Photomultipliers

The after-pulsing probability in Silicon Photomulti- pliers and its time constant are obtained measuring the mean number of photo-electrons in a variable time window following a light pulse. The method, experimentally simple and statistically robust due to the use of the Central Limit Theorem, has been applied to an HAMAMATSU MPPC S10362-11-100C.

preprint2011arXiv

Atmospheric fluctuations below 0.1 Hz during drift-scan solar diameter measurements

Measurements of the power spectrum of the seeing in the range 0.001-1 Hz have been performed in order to understand the criticity of the transits' method for solar diameter monitoring.

preprint2010arXiv

Photon-number statistics with Silicon photomultipliers

We present a description of the operation of a multi-pixel detector in the presence of non-negligible dark-count and cross-talk effects. We apply the model to devise self-consistent calibration strategies to be performed on the very light under investigation.

Massimo Caccia

What is connected

Connect this record

See the researcher in context

Building this map preview

10 published item(s)

Multi-fidelity hydrodynamic analysis of an autonomous surface vehicle at surveying speed in deep water subject to variable payload

Understanding Continual Learning Settings with Data Distribution Drift Analysis

Online Fast Adaptation and Knowledge Accumulation: a New Approach to Continual Learning

CVPR 2020 Continual Learning in Computer Vision Competition: Approaches, Results, Current Challenges and Future Directions

Language GANs Falling Short

Online Learned Continual Compression with Adaptive Quantization Modules

Adaptive Experimental Design for Path-following Performance Assessment of Unmanned Vehicles

A simple and robust method to study after-pulses in Silicon Photomultipliers

Atmospheric fluctuations below 0.1 Hz during drift-scan solar diameter measurements

Photon-number statistics with Silicon photomultipliers