Source author record

A. Lonardo

A. Lonardo appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

22works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

22 published item(s)

preprint2022arXiv

Architectural improvements and technological enhancements for the APEnet+ interconnect system

The APEnet+ board delivers a point-to-point, low-latency, 3D torus network interface card. In this paper we describe the latest generation of APEnet NIC, APEnet v5, integrated in a PCIe Gen3 board based on a state-of-the-art, 28 nm Altera Stratix V FPGA. The NIC features a network architecture designed following the Remote DMA paradigm and tailored to tightly bind the computing power of modern GPUs to the communication fabric. For the APEnet v5 board we show characterizing figures as achieved bandwidth and BER obtained by exploiting new high performance ALTERA transceivers and PCIe Gen3 compliancy.

preprint2022arXiv

Progress report on the online processing upgrade at the NA62 experiment

A new FPGA-based low-level trigger processor has been installed at the NA62 experiment. It is intended to extend the features of its predecessor due to a faster interconnection technology and additional logic resources available on the new platform. With the aim of improving trigger selectivity and exploring new architectures for complex trigger computation, a GPU system has been developed and a neural network on FPGA is in progress. They both process data streams from the Ring Imaging Cherenkov detector of the experiment to extract in real time high level features for the trigger logic. Description of the systems, latest developments and design flows are reported in this paper.

preprint2020arXiv

Dependence of atmospheric muon flux on seawater depth measured with the first KM3NeT detection units

KM3NeT is a research infrastructure located in the Mediterranean Sea, that will consist of two deep-sea Cherenkov neutrino detectors. With one detector (ARCA), the KM3NeT Collaboration aims at identifying and studying TeV-PeV astrophysical neutrino sources. With the other detector (ORCA), the neutrino mass ordering will be determined by studying GeV-scale atmospheric neutrino oscillations. The first KM3NeT detection units were deployed at the Italian and French sites between 2015 and 2017. In this paper, a description of the detector is presented, together with a summary of the procedures used to calibrate the detector in-situ. Finally, the measurement of the atmospheric muon flux between 2232-3386 m seawater depth is obtained.

preprint2019arXiv

Sensitivity of the KM3NeT/ARCA neutrino telescope to point-like neutrino sources

KM3NeT will be a network of deep-sea neutrino telescopes in the Mediterranean Sea. The KM3NeT/ARCA detector, to be installed at the Capo Passero site (Italy), is optimised for the detection of high-energy neutrinos of cosmic origin. Thanks to its geographical location on the Northern hemisphere, KM3NeT/ARCA can observe upgoing neutrinos from most of the Galactic Plane, including the Galactic Centre. Given its effective area and excellent pointing resolution, KM3NeT/ARCA will measure or significantly constrain the neutrino flux from potential astrophysical neutrino sources. At the same time, it will test flux predictions based on gamma-ray measurements and the assumption that the gamma-ray flux is of hadronic origin. Assuming this scenario, discovery potentials and sensitivities for a selected list of Galactic sources and to generic point sources with an $E^{-2}$ spectrum are presented. These spectra are assumed to be time independent. The results indicate that an observation with $3σ$ significance is possible in about six years of operation for the most intense sources, such as Supernovae Remnants RX\,J1713.7-3946 and Vela Jr. If no signal will be found during this time, the fraction of the gamma-ray flux coming from hadronic processes can be constrained to be below 50\% for these two objects.

preprint2019arXiv

The Control Unit of the KM3NeT Data Acquisition System

The KM3NeT Collaboration runs a multi-site neutrino observatory in the Mediterranean Sea. Water Cherenkov particle detectors, deep in the sea and far off the coasts of France and Italy, are already taking data while incremental construction progresses. Data Acquisition Control software is operating off-shore detectors as well as testing and qualification stations for their components. The software, named Control Unit, is highly modular. It can undergo upgrades and reconfiguration with the acquisition running. Interplay with the central database of the Collaboration is obtained in a way that allows for data taking even if Internet links fail. In order to simplify the management of computing resources in the long term, and to cope with possible hardware failures of one or more computers, the KM3NeT Control Unit software features a custom dynamic resource provisioning and failover technology, which is especially important for ensuring continuity in case of rare transient events in multi-messenger astronomy. The software architecture relies on ubiquitous tools and broadly adopted technologies and has been successfully tested on several operating systems.

preprint2016arXiv

GPU-based Real-time Triggering in the NA62 Experiment

Over the last few years the GPGPU (General-Purpose computing on Graphics Processing Units) paradigm represented a remarkable development in the world of computing. Computing for High-Energy Physics is no exception: several works have demonstrated the effectiveness of the integration of GPU-based systems in high level trigger of different experiments. On the other hand the use of GPUs in the low level trigger systems, characterized by stringent real-time constraints, such as tight time budget and high throughput, poses several challenges. In this paper we focus on the low level trigger in the CERN NA62 experiment, investigating the use of real-time computing on GPUs in this synchronous system. Our approach aimed at harvesting the GPU computing power to build in real-time refined physics-related trigger primitives for the RICH detector, as the the knowledge of Cerenkov rings parameters allows to build stringent conditions for data selection at trigger level. Latencies of all components of the trigger chain have been analyzed, pointing out that networking is the most critical one. To keep the latency of data transfer task under control, we devised NaNet, an FPGA-based PCIe Network Interface Card (NIC) with GPUDirect capabilities. For the processing task, we developed specific multiple ring trigger algorithms to leverage the parallel architecture of GPUs and increase the processing throughput to keep up with the high event rate. Results obtained during the first months of 2016 NA62 run are presented and discussed.

preprint2016arXiv

Letter of Intent for KM3NeT 2.0

The main objectives of the KM3NeT Collaboration are i) the discovery and subsequent observation of high-energy neutrino sources in the Universe and ii) the determination of the mass hierarchy of neutrinos. These objectives are strongly motivated by two recent important discoveries, namely: 1) The high-energy astrophysical neutrino signal reported by IceCube and 2) the sizable contribution of electron neutrinos to the third neutrino mass eigenstate as reported by Daya Bay, Reno and others. To meet these objectives, the KM3NeT Collaboration plans to build a new Research Infrastructure consisting of a network of deep-sea neutrino telescopes in the Mediterranean Sea. A phased and distributed implementation is pursued which maximises the access to regional funds, the availability of human resources and the synergetic opportunities for the earth and sea sciences community. Three suitable deep-sea sites are identified, namely off-shore Toulon (France), Capo Passero (Italy) and Pylos (Greece). The infrastructure will consist of three so-called building blocks. A building block comprises 115 strings, each string comprises 18 optical modules and each optical module comprises 31 photo-multiplier tubes. Each building block thus constitutes a 3-dimensional array of photo sensors that can be used to detect the Cherenkov light produced by relativistic particles emerging from neutrino interactions. Two building blocks will be configured to fully explore the IceCube signal with different methodology, improved resolution and complementary field of view, including the Galactic plane. One building block will be configured to precisely measure atmospheric neutrino oscillations.

preprint2016arXiv

Long term monitoring of the optical background in the Capo Passero deep-sea site with the NEMO tower prototype

The NEMO Phase-2 tower is the first detector which was operated underwater for more than one year at the "record" depth of 3500 m. It was designed and built within the framework of the NEMO (NEutrino Mediterranean Observatory) project. The 380 m high tower was successfully installed in March 2013 80 km offshore Capo Passero (Italy). This is the first prototype operated on the site where the italian node of the KM3NeT neutrino telescope will be built. The installation and operation of the NEMO Phase-2 tower has proven the functionality of the infrastructure and the operability at 3500 m depth. A more than one year long monitoring of the deep water characteristics of the site has been also provided. In this paper the infrastructure and the tower structure and instrumentation are described. The results of long term optical background measurements are presented. The rates show stable and low baseline values, compatible with the contribution of 40K light emission, with a small percentage of light bursts due to bioluminescence. All these features confirm the stability and good optical properties of the site.

preprint2015arXiv

The prototype detection unit of the KM3NeT detector

A prototype detection unit of the KM3NeT deep-sea neutrino telescope has been installed at 3500m depth 80km offshore the Italian coast. KM3NeT in its final configuration will contain several hundreds of detection units. Each detection unit is a mechanical structure anchored to the sea floor, held vertical by a submerged buoy and supporting optical modules for the detection of Cherenkov light emitted by charged secondary particles emerging from neutrino interactions. This prototype string implements three optical modules with 31 photomultiplier tubes each. These optical modules were developed by the KM3NeT Collaboration to enhance the detection capability of neutrino interactions. The prototype detection unit was operated since its deployment in May 2014 until its decommissioning in July 2015. Reconstruction of the particle trajectories from the data requires a nanosecond accuracy in the time calibration. A procedure for relative time calibration of the photomultiplier tubes contained in each optical module is described. This procedure is based on the measured coincidences produced in the sea by the 40K background light and can easily be expanded to a detector with several thousands of optical modules. The time offsets between the different optical modules are obtained using LED nanobeacons mounted inside them. A set of data corresponding to 600 hours of livetime was analysed. The results show good agreement with Monte Carlo simulations of the expected optical background and the signal from atmospheric muons. An almost background-free sample of muons was selected by filtering the time correlated signals on all the three optical modules. The zenith angle of the selected muons was reconstructed with a precision of about 3°.

preprint2014arXiv

Measurement of the atmospheric muon depth intensity relation with the NEMO Phase-2 tower

The results of the analysis of the data collected with the NEMO Phase-2 tower, deployed at 3500 m depth about 80 km off-shore Capo Passero (Italy), are presented. Cherenkov photons detected with the photomultipliers tubes were used to reconstruct the tracks of atmospheric muons. Their zenith-angle distribution was measured and the results compared with Monte Carlo simulations. An evaluation of the systematic effects due to uncertainties on environmental and detector parameters is also included. The associated depth intensity relation was evaluated and compared with previous measurements and theoretical predictions. With the present analysis, the muon depth intensity relation has been measured up to 13 km of water equivalent.

preprint2014arXiv

NaNet: a flexible and configurable low-latency NIC for real-time trigger systems based on GPUs

NaNet is an FPGA-based PCIe X8 Gen2 NIC supporting 1/10 GbE links and the custom 34 Gbps APElink channel. The design has GPUDirect RDMA capabilities and features a network stack protocol offloading module, making it suitable for building low-latency, real-time GPU-based computing systems. We provide a detailed description of the NaNet hardware modular architecture. Benchmarks for latency and bandwidth for GbE and APElink channels are presented, followed by a performance analysis on the case study of the GPU-based low level trigger for the RICH detector in the NA62 CERN experiment, using either the NaNet GbE and APElink channels. Finally, we give an outline of project future activities.

preprint2014arXiv

NaNet: a Low-Latency, Real-Time, Multi-Standard Network Interface Card with GPUDirect Features

While the GPGPU paradigm is widely recognized as an effective approach to high performance computing, its adoption in low-latency, real-time systems is still in its early stages. Although GPUs typically show deterministic behaviour in terms of latency in executing computational kernels as soon as data is available in their internal memories, assessment of real-time features of a standard GPGPU system needs careful characterization of all subsystems along data stream path. The networking subsystem results in being the most critical one in terms of absolute value and fluctuations of its response latency. Our envisioned solution to this issue is NaNet, a FPGA-based PCIe Network Interface Card (NIC) design featuring a configurable and extensible set of network channels with direct access through GPUDirect to NVIDIA Fermi/Kepler GPU memories. NaNet design currently supports both standard - GbE (1000BASE-T) and 10GbE (10Base-R) - and custom - 34~Gbps APElink and 2.5~Gbps deterministic latency KM3link - channels, but its modularity allows for a straightforward inclusion of other link technologies. To avoid host OS intervention on data stream and remove a possible source of jitter, the design includes a network/transport layer offload module with cycle-accurate, upper-bound latency, supporting UDP, KM3link Time Division Multiplexing and APElink protocols. After NaNet architecture description and its latency/bandwidth characterization for all supported links, two real world use cases will be presented: the GPU-based low level trigger for the RICH detector in the NA62 experiment at CERN and the on-/off-shore data link for KM3 underwater neutrino telescope.

preprint2013arXiv

Applications of Many-Core Technologies to On-line Event Reconstruction in High Energy Physics Experiments

Interest in many-core architectures applied to real time selections is growing in High Energy Physics (HEP) experiments. In this paper we describe performance measurements of many-core devices when applied to a typical HEP online task: the selection of events based on the trajectories of charged particles. We use as benchmark a scaled-up version of the algorithm used at CDF experiment at Tevatron for online track reconstruction - the SVT algorithm - as a realistic test-case for low-latency trigger systems using new computing architectures for LHC experiment. We examine the complexity/performance trade-off in porting existing serial algorithms to many-core devices. We measure performance of different architectures (Intel Xeon Phi and AMD GPUs, in addition to NVidia GPUs) and different software environments (OpenCL, in addition to NVidia CUDA). Measurements of both data processing and data transfer latency are shown, considering different I/O strategies to/from the many-core devices.

preprint2013arXiv

Many-core applications to online track reconstruction in HEP experiments

Interest in parallel architectures applied to real time selections is growing in High Energy Physics (HEP) experiments. In this paper we describe performance measurements of Graphic Processing Units (GPUs) and Intel Many Integrated Core architecture (MIC) when applied to a typical HEP online task: the selection of events based on the trajectories of charged particles. We use as benchmark a scaled-up version of the algorithm used at CDF experiment at Tevatron for online track reconstruction - the SVT algorithm - as a realistic test-case for low-latency trigger systems using new computing architectures for LHC experiment. We examine the complexity/performance trade-off in porting existing serial algorithms to many-core devices. Measurements of both data processing and data transfer latency are shown, considering different I/O strategies to/from the parallel devices.

preprint2011arXiv

High-speed data transfer with FPGAs and QSFP+ modules

We present test results and characterization of a data transmission system based on a last generation FPGA and a commercial QSFP+ (Quad Small Form Pluggable +) module. QSFP+ standard defines a hot-pluggable transceiver available in copper or optical cable assemblies for an aggregated bandwidth of up to 40 Gbps. We implemented a complete testbench based on a commercial development card mounting an Altera Stratix IV FPGA with 24 serial transceivers at 8.5 Gbps, together with a custom mezzanine hosting three QSFP+ modules. We present test results and signal integrity measurements up to an aggregated bandwidth of 12 Gbps.

preprint2010arXiv

Measurement of the atmospheric muon flux with the NEMO Phase-1 detector

The NEMO Collaboration installed and operated an underwater detector including prototypes of the critical elements of a possible underwater km3 neutrino telescope: a four-floor tower (called Mini-Tower) and a Junction Box. The detector was developed to test some of the main systems of the km3 detector, including the data transmission, the power distribution, the timing calibration and the acoustic positioning systems as well as to verify the capabilities of a single tridimensional detection structure to reconstruct muon tracks. We present results of the analysis of the data collected with the NEMO Mini-Tower. The position of photomultiplier tubes (PMTs) is determined through the acoustic position system. Signals detected with PMTs are used to reconstruct the tracks of atmospheric muons. The angular distribution of atmospheric muons was measured and results compared with Monte Carlo simulations.

preprint2003arXiv

apeNEXT: A multi-TFlops Computer for Simulations in Lattice Gauge Theory

We present the APE (Array Processor Experiment) project for the development of dedicated parallel computers for numerical simulations in lattice gauge theories. While APEmille is a production machine in today's physics simulations at various sites in Europe, a new machine, apeNEXT, is currently being developed to provide multi-Tflops computing performance. Like previous APE machines, the new supercomputer is largely custom designed and specifically optimized for simulations of Lattice QCD.

preprint2003arXiv

Status of the apeNEXT project

We present the current status of the apeNEXT project. Aim of this project is the development of the next generation of APE machines which will provide multi-teraflop computing power. Like previous machines, apeNEXT is based on a custom designed processor, which is specifically optimized for simulating QCD. We discuss the machine design, report on benchmarks, and give an overview on the status of the software development.

preprint2003arXiv

The apeNEXT project (Status report)

We present the current status of the apeNEXT project. Aim of this project is the development of the next generation of APE machines which will provide multi-teraflop computing power. Like previous machines, apeNEXT is based on a custom designed processor, which is specifically optimized for simulating QCD. We discuss the machine design, report on benchmarks, and give an overview on the status of the software development.

preprint2001arXiv

Status of APEmille

This paper presents the status of the APEmille project, which is essentially completed, as far as machine development and construction is concerned. Several large installations of APEmille are in use for physics production runs leading to many new results presented at this conference. This paper briefly summarizes the APEmille architecture, reviews the status of the installations and presents some performance figures for physics codes.