Source author record

Maram Akila

Maram Akila appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2022arXiv

Tailored Uncertainty Estimation for Deep Learning Systems

Uncertainty estimation bears the potential to make deep learning (DL) systems more reliable. Standard techniques for uncertainty estimation, however, come along with specific combinations of strengths and weaknesses, e.g., with respect to estimation quality, generalization abilities and computational complexity. To actually harness the potential of uncertainty quantification, estimators are required whose properties closely match the requirements of a given use case. In this work, we propose a framework that, firstly, structures and shapes these requirements, secondly, guides the selection of a suitable uncertainty estimation method and, thirdly, provides strategies to validate this choice and to uncover structural weaknesses. By contributing tailored uncertainty estimation in this sense, our framework helps to foster trustworthy DL systems. Moreover, it anticipates prospective machine learning regulations that require, e.g., in the EU, evidences for the technical appropriateness of machine learning systems. Our framework provides such evidences for system components modeling uncertainty.

preprint2021arXiv

A Novel Regression Loss for Non-Parametric Uncertainty Optimization

Quantification of uncertainty is one of the most promising approaches to establish safe machine learning. Despite its importance, it is far from being generally solved, especially for neural networks. One of the most commonly used approaches so far is Monte Carlo dropout, which is computationally cheap and easy to apply in practice. However, it can underestimate the uncertainty. We propose a new objective, referred to as second-moment loss (SML), to address this issue. While the full network is encouraged to model the mean, the dropout networks are explicitly used to optimize the model variance. We intensively study the performance of the new objective on various UCI regression datasets. Comparing to the state-of-the-art of deep ensembles, SML leads to comparable prediction accuracies and uncertainty estimates while only requiring a single model. Under distribution shift, we observe moderate improvements. As a side result, we introduce an intuitive Wasserstein distance-based uncertainty measure that is non-saturating and thus allows to resolve quality differences between any two uncertainty estimates.

preprint2021arXiv

Inspect, Understand, Overcome: A Survey of Practical Methods for AI Safety

The use of deep neural networks (DNNs) in safety-critical applications like mobile health and autonomous driving is challenging due to numerous model-inherent shortcomings. These shortcomings are diverse and range from a lack of generalization over insufficient interpretability to problems with malicious inputs. Cyber-physical systems employing DNNs are therefore likely to suffer from safety concerns. In recent years, a zoo of state-of-the-art techniques aiming to address these safety concerns has emerged. This work provides a structured and broad overview of them. We first identify categories of insufficiencies to then describe research activities aiming at their detection, quantification, or mitigation. Our paper addresses both machine learning experts and safety engineers: The former ones might profit from the broad range of machine learning topics covered and discussions on limitations of recent methods. The latter ones might gain insights into the specifics of modern ML methods. We moreover hope that our contribution fuels discussions on desiderata for ML systems and strategies on how to propel existing approaches accordingly.

preprint2020arXiv

Characteristics of Monte Carlo Dropout in Wide Neural Networks

Monte Carlo (MC) dropout is one of the state-of-the-art approaches for uncertainty estimation in neural networks (NNs). It has been interpreted as approximately performing Bayesian inference. Based on previous work on the approximation of Gaussian processes by wide and deep neural networks with random weights, we study the limiting distribution of wide untrained NNs under dropout more rigorously and prove that they as well converge to Gaussian processes for fixed sets of weights and biases. We sketch an argument that this property might also hold for infinitely wide feed-forward networks that are trained with (full-batch) gradient descent. The theory is contrasted by an empirical analysis in which we find correlations and non-Gaussian behaviour for the pre-activations of finite width NNs. We therefore investigate how (strongly) correlated pre-activations can induce non-Gaussian behavior in NNs with strongly correlated weights.

preprint2020arXiv

Local correlations in dual-unitary kicked chains

We show that for dual-unitary kicked chains, built upon a pair of complex Hadamard matrices, correlators of strictly local, traceless operators vanish identically for sufficiently long chains. On the other hand, operators supported at pairs of adjacent chain sites, generically, exhibit nontrivial correlations along the light cone edges. In agreement with Bertini et. al. [Phys. Rev. Lett. 123, 210601 (2019)], they can be expressed through the expectation values of a transfer matrix $T$. Furthermore, we identify a remarkable family of dual-unitary models where an explicit information on the spectrum of $T$ is available. For this class of models we provide a closed analytical formula for the corresponding two-point correlators. This result, in turn, allows an evaluation of local correlators in the vicinity of the dual-unitary regime which is exemplified on the kicked Ising spin chain.

preprint2020arXiv

Transition from Quantum Chaos to Localization in Spin Chains

Recent years have seen an increasing interest in quantum chaos and related aspects of spatially extended systems, such as spin chains. However, the results are strongly system dependent, generic approaches suggest the presence of many-body localization while analytical calculations for certain system classes, here referred to as the ``self-dual case'', prove adherence to universal (chaotic) spectral behavior. We address these issues studying the level statistics in the vicinity of the latter case, thereby revealing transitions to many-body localization as well as the appearance of several non-standard random-matrix universality classes.

preprint2016arXiv

Trace formula for spin chains

While detailed information about the semiclassics for single-particle systems is available, much less is known about the connection between quantum and classical dynamics for many-body systems. As an example, we focus on spin chains which are of considerable conceptual and practical importance. We derive a trace formula for coupled spin $j$ particles which relates the quantum energy levels to the classical dynamics. Our derivation is valid in the limit $j\rightarrow\infty$ with $j\hbar={\rm const.}$ and applies to time-continuous as well as to periodically driven dynamics. We provide a simple explanation why the Solari-Kochetov phase can be omitted if the correct classical Hamiltonian is chosen.

preprint2015arXiv

Spectral statistics of nearly unidirectional quantum graphs

The energy levels of a quantum graph with time reversal symmetry and unidirectional classical dynamics are doubly degenerate and obey the spectral statistics of the Gaussian Unitary Ensemble. These degeneracies, however, are lifted when the unidirectionality is broken in one of the graph's vertices by a singular perturbation. Based on a Random Matrix model we derive an analytic expression for the nearest neighbour distribution between energy levels of such systems. As we demonstrate the result agrees excellently with the actual statistics for graphs with a uniform distribution of eigenfunctions. Yet, it exhibits quite substantial deviations for classes of graphs which show strong scarring.