Source author record

Heinrich Küttler

Heinrich Küttler appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2021arXiv

Learning with AMIGo: Adversarially Motivated Intrinsic Goals

A key challenge for reinforcement learning (RL) consists of learning in environments with sparse extrinsic rewards. In contrast to current RL methods, humans are able to learn new skills with little or no reward by using various forms of intrinsic motivation. We propose AMIGo, a novel agent incorporating -- as form of meta-learning -- a goal-generating teacher that proposes Adversarially Motivated Intrinsic Goals to train a goal-conditioned "student" policy in the absence of (or alongside) environment reward. Specifically, through a simple but effective "constructively adversarial" objective, the teacher learns to propose increasingly challenging -- yet achievable -- goals that allow the student to learn general skills for acting in a new environment, independent of the task to be solved. We show that our method generates a natural curriculum of self-proposed goals which ultimately allows the agent to solve challenging procedurally-generated tasks where other forms of intrinsic motivation and state-of-the-art RL methods fail.

preprint2021arXiv

PAQ: 65 Million Probably-Asked Questions and What You Can Do With Them

Open-domain Question Answering models which directly leverage question-answer (QA) pairs, such as closed-book QA (CBQA) models and QA-pair retrievers, show promise in terms of speed and memory compared to conventional models which retrieve and read from text corpora. QA-pair retrievers also offer interpretable answers, a high degree of control, and are trivial to update at test time with new knowledge. However, these models lack the accuracy of retrieve-and-read systems, as substantially less knowledge is covered by the available QA-pairs relative to text corpora like Wikipedia. To facilitate improved QA-pair models, we introduce Probably Asked Questions (PAQ), a very large resource of 65M automatically-generated QA-pairs. We introduce a new QA-pair retriever, RePAQ, to complement PAQ. We find that PAQ preempts and caches test questions, enabling RePAQ to match the accuracy of recent retrieve-and-read models, whilst being significantly faster. Using PAQ, we train CBQA models which outperform comparable baselines by 5%, but trail RePAQ by over 15%, indicating the effectiveness of explicit retrieval. RePAQ can be configured for size (under 500MB) or speed (over 1K questions per second) whilst retaining high accuracy. Lastly, we demonstrate RePAQ's strength at selective QA, abstaining from answering when it is likely to be incorrect. This enables RePAQ to ``back-off" to a more expensive state-of-the-art model, leading to a combined system which is both more accurate and 2x faster than the state-of-the-art model alone.

preprint2016arXiv

DeepMind Lab

DeepMind Lab is a first-person 3D game platform designed for research and development of general artificial intelligence and machine learning systems. DeepMind Lab can be used to study how autonomous artificial agents may learn complex tasks in large, partially observed, and visually diverse worlds. DeepMind Lab has a simple and flexible API enabling creative task-designs and novel AI-designs to be explored and quickly iterated upon. It is powered by a fast and widely recognised game engine, and tailored for effective use by the research community.

preprint2014arXiv

Anderson's orthogonality catastrophe

We give an upper bound on the modulus of the ground-state overlap of two non-interacting fermionic quantum systems with $N$ particles in a large but finite volume $L^d$ of $d$-dimensional Euclidean space. The underlying one-particle Hamiltonians of the two systems are standard Schrödinger operators that differ by a non-negative compactly supported scalar potential. In the thermodynamic limit, the bound exhibits an asymptotic power-law decay in the system size $L$, showing that the ground-state overlap vanishes for macroscopic systems. The decay exponent can be interpreted in terms of the total scattering cross section averaged over all incident directions. The result confirms and generalises P. W. Anderson's informal computation [Phys. Rev. Lett. 18, 1049--1051 (1967)].

preprint2013arXiv

Anderson's Orthogonality Catastrophe for One-dimensional Systems

We derive rigorously the leading asymptotics of the so-called Anderson integral in the thermodynamic limit for one-dimensional, non-relativistic, spin-less Fermi systems. The coefficient, $γ$, of the leading term is computed in terms of the S-matrix. This implies a lower and an upper bound on the exponent in Anderson's orthogonality catastrophe, $\tilde CN^{-\tildeγ}\leq \mathcal{D}_N\leq CN^{-γ}$ pertaining to the overlap, $\mathcal{D}_N$, of ground states of non-interacting fermions.