Source author record

David Schultz

David Schultz appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2022arXiv

The anachronism of whole-GPU accounting

NVIDIA has been making steady progress in increasing the compute performance of its GPUs, resulting in order of magnitude compute throughput improvements over the years. With several models of GPUs coexisting in many deployments, the traditional accounting method of treating all GPUs as being equal is not reflecting compute output anymore. Moreover, for applications that require significant CPU-based compute to complement the GPU-based compute, it is becoming harder and harder to make full use of the newer GPUs, requiring sharing of those GPUs between multiple applications in order to maximize the achievable science output. This further reduces the value of whole-GPU accounting, especially when the sharing is done at the infrastructure level. We thus argue that GPU accounting for throughput-oriented infrastructures should be expressed in GPU core hours, much like it is normally done for the CPUs. While GPU core compute throughput does change between GPU generations, the variability is similar to what we expect to see among CPU cores. To validate our position, we present an extensive set of run time measurements of two IceCube photon propagation workflows on 14 GPU models, using both on-prem and Cloud resources. The measurements also outline the influence of GPU sharing at both HTCondor and Kubernetes infrastructure level.

preprint2020arXiv

An Average-Compress Algorithm for the Sample Mean Problem under Dynamic Time Warping

Computing a sample mean of time series under dynamic time warping (DTW) is NP-hard. Consequently, there is an ongoing research effort to devise efficient heuristics. The majority of heuristics have been developed for the constrained sample mean problem that assumes a solution of predefined length. In contrast, research on the unconstrained sample mean problem is underdeveloped. In this article, we propose a generic average-compress (AC) algorithm for solving the unconstrained problem. The algorithm alternates between averaging (A-step) and compression (C-step). The A-step takes an initial guess as input and returns an approximation of a sample mean. Then the C-step reduces the length of the approximate solution. The compressed approximation serves as initial guess of the A-step in the next iteration. The purpose of the C-step is to direct the algorithm to more promising solutions of shorter length. The proposed algorithm is generic in the sense that any averaging and any compression method can be used. Experimental results show that the AC algorithm substantially outperforms current state-of-the-art algorithms for time series averaging.

preprint2020arXiv

Running a Pre-Exascale, Geographically Distributed, Multi-Cloud Scientific Simulation

As we approach the Exascale era, it is important to verify that the existing frameworks and tools will still work at that scale. Moreover, public Cloud computing has been emerging as a viable solution for both prototyping and urgent computing. Using the elasticity of the Cloud, we have thus put in place a pre-exascale HTCondor setup for running a scientific simulation in the Cloud, with the chosen application being IceCube's photon propagation simulation. I.e. this was not a purely demonstration run, but it was also used to produce valuable and much needed scientific results for the IceCube collaboration. In order to reach the desired scale, we aggregated GPU resources across 8 GPU models from many geographic regions across Amazon Web Services, Microsoft Azure, and the Google Cloud Platform. Using this setup, we reached a peak of over 51k GPUs corresponding to almost 380 PFLOP32s, for a total integrated compute of about 100k GPU hours. In this paper we provide the description of the setup, the problems that were discovered and overcome, as well as a short description of the actual science output of the exercise.

preprint2015arXiv

Detecting particles with cell phones: the Distributed Electronic Cosmic-ray Observatory

In 2014 the number of active cell phones worldwide for the first time surpassed the number of humans. Cell phone camera quality and onboard processing power (both CPU and GPU) continue to improve rapidly. In addition to their primary purpose of detecting photons, camera image sensors on cell phones and other ubiquitous devices such as tablets, laptops and digital cameras can detect ionizing radiation produced by cosmic rays and radioactive decays. While cosmic rays have long been understood and characterized as a nuisance in astronomical cameras, they can also be identified as a signal in idle camera image sensors. We present the Distributed Electronic Cosmic-ray Observatory (DECO), a platform for outreach and education as well as for citizen science. Consisting of an app and associated database and web site, DECO harnesses the power of distributed camera image sensors for cosmic-ray detection.

preprint2011arXiv

Laboratory Astrophysics White Paper (based on the 2010 NASA Laboratory Astrophysics Workshop in Gatlinberg, Tennessee, 25-28 October 2010)

The purpose of the 2010 NASA Laboratory Astrophysics Workshop (LAW) was, as given in the Charter from NASA, "to provide a forum within which the scientific community can review the current state of knowledge in the field of Laboratory Astrophysics, assess the critical data needs of NASA's current and future Space Astrophysics missions, and identify the challenges and opportunities facing the field as we begin a new decade". LAW 2010 was the fourth in a roughly quadrennial series of such workshops sponsored by the Astrophysics Division of the NASA Science Mission Directorate. In this White Paper, we report the findings of the workshop.