Source author record

Alexander Pritzel

Alexander Pritzel appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

9works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

9 published item(s)

preprint2022arXiv

Normalizing flows for atomic solids

We present a machine-learning approach, based on normalizing flows, for modelling atomic solids. Our model transforms an analytically tractable base distribution into the target solid without requiring ground-truth samples for training. We report Helmholtz free energy estimates for cubic and hexagonal ice modelled as monatomic water as well as for a truncated and shifted Lennard-Jones system, and find them to be in excellent agreement with literature values and with estimates from established baseline methods. We further investigate structural properties and show that the model samples are nearly indistinguishable from the ones obtained with molecular dynamics. Our results thus demonstrate that normalizing flows can provide high-quality samples and free energy estimates without the need for multi-staging.

preprint2020arXiv

Never Give Up: Learning Directed Exploration Strategies

We propose a reinforcement learning agent to solve hard exploration games by learning a range of directed exploratory policies. We construct an episodic memory-based intrinsic reward using k-nearest neighbors over the agent's recent experience to train the directed exploratory policies, thereby encouraging the agent to repeatedly revisit all states in its environment. A self-supervised inverse dynamics model is used to train the embeddings of the nearest neighbour lookup, biasing the novelty signal towards what the agent can control. We employ the framework of Universal Value Function Approximators (UVFA) to simultaneously learn many directed exploration policies with the same neural network, with different trade-offs between exploration and exploitation. By using the same neural network for different degrees of exploration/exploitation, transfer is demonstrated from predominantly exploratory policies yielding effective exploitative policies. The proposed method can be incorporated to run with modern distributed RL agents that collect large amounts of experience from many actors running in parallel on separate environment instances. Our method doubles the performance of the base agent in all hard exploration in the Atari-57 suite while maintaining a very high score across the remaining games, obtaining a median human normalised score of 1344.0%. Notably, the proposed method is the first algorithm to achieve non-zero rewards (with a mean score of 8,400) in the game of Pitfall! without using demonstrations or hand-crafted features.

preprint2016arXiv

Deep Exploration via Bootstrapped DQN

Efficient exploration in complex environments remains a major challenge for reinforcement learning. We propose bootstrapped DQN, a simple algorithm that explores in a computationally and statistically efficient manner through use of randomized value functions. Unlike dithering strategies such as epsilon-greedy exploration, bootstrapped DQN carries out temporally-extended (or deep) exploration; this can lead to exponentially faster learning. We demonstrate these benefits in complex stochastic MDPs and in the large-scale Arcade Learning Environment. Bootstrapped DQN substantially improves learning times and performance across most Atari games.

preprint2016arXiv

Model-Free Episodic Control

State of the art deep reinforcement learning algorithms take many millions of interactions to attain human-level performance. Humans, on the other hand, can very quickly exploit highly rewarding nuances of an environment upon first discovery. In the brain, such rapid learning is thought to depend on the hippocampus and its capacity for episodic memory. Here we investigate whether a simple model of hippocampal episodic control can learn to solve difficult sequential decision-making tasks. We demonstrate that it not only attains a highly rewarding strategy significantly faster than state-of-the-art deep reinforcement learning algorithms, but also achieves a higher overall reward on some of the more challenging domains.

preprint2015arXiv

Large-N ground state of the Lieb-Liniger model and Yang-Mills theory on a two-sphere

We derive the large particle number limit of the Bethe equations for the ground state of the attractive one-dimensional Bose gas (Lieb-Liniger model) on a ring and solve it for arbitrary coupling. We show that the ground state of this system can be mapped to the large-N saddle point of Euclidean Yang-Mills theory on a two-sphere with a U(N) gauge group, and the phase transition that interpolates between the homogeneous and solitonic regime is dual to the Douglas-Kazakov confimenent-deconfinement phase transition.

preprint2014arXiv

Topological Model for Domain Walls in (Super-)Yang-Mills Theories

We derive a topological action that describes the confining phase of (Super-)Yang-Mills theories with gauge group $SU(N)$, similar to the work recently carried out by Seiberg and collaborators. It encodes all the Aharonov-Bohm phases of the possible non-local operators and phases generated by the intersection of flux tubes. Within this topological framework we show that the worldvolume theory of domain walls contains a Chern-Simons term at level $N$ also seen in string theory constructions. The discussion can also illuminate dynamical differences of domain walls in the supersymmetric and non-supersymmetric framework. Two further analogies, to string theory and the fractional quantum Hall effect might lead to additional possibilities to investigate the dynamics.

preprint2013arXiv

Scrambling in the Black Hole Portrait

Recently a quantum portrait of black holes was suggested according to which a macroscopic black hole is a Bose-Einstein condensate of soft gravitons stuck at the critical point of a quantum phase transition. We explain why quantum criticality and instability are the key for efficient generation of entanglement and consequently of the scrambling of information. By studying a simple Bose-Einstein prototype, we show that the scrambling time, which is set by the quantum break time of the system, goes as $\log N \,$ for $N$ the number of quantum constituents or equivalently the black hole entropy.

preprint2011arXiv

Localization of gauge fields and Maxwell-Chern-Simons theory

We propose an explicit model, where an axionic domain wall dynamically localizes a U(1)-component of a nonabelian gauge theory living in a 3+1 dimensional bulk. The effective theory on the wall is 2+1d Maxwell-Chern-Simons theory with a compact U(1) gauge group. This setup allows us to understand all key properties of MCS theory in terms of the dynamics of the underlying 3+1 dimensional gauge theory. Our findings can also shed some light on branes in supersymmetric gluodynamics.

preprint2011arXiv

On ghosts in theories of self-interacting massive spin-2 particles

We consider general theories of a massive spin-2 particle $h_{μν}$ on a Minkowski background. A decomposition of $h_{μν}$ in terms of helicity eigenstates allows us to directly test whether any given theory possesses a consistent description as a massive spin-2 representation of the Poincaré group. We demonstrate (i) that any nonlinear theory with an Einsteinian derivative structure either contains ghosts or does not describe a weakly coupled spin-2 and (ii) that there exists a two-parameter family of non-Einsteinian cubic self-interactions which constitute a ghost-free massive spin-2 theory.