Source author record

Zachary Kenton

Zachary Kenton appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

6works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

6 published item(s)

preprint2022arXiv

Discovering Agents

Causal models of agents have been used to analyse the safety aspects of machine learning systems. But identifying agents is non-trivial -- often the causal model is just assumed by the modeler without much justification -- and modelling failures can lead to mistakes in the safety analysis. This paper proposes the first formal causal definition of agents -- roughly that agents are systems that would adapt their policy if their actions influenced the world in a different way. From this we derive the first causal discovery algorithm for discovering agents from empirical data, and give algorithms for translating between causal models and game-theoretic influence diagrams. We demonstrate our approach by resolving some previous confusions caused by incorrect causal modelling of agents.

preprint2022arXiv

Safe Deep RL in 3D Environments using Human Feedback

Agents should avoid unsafe behaviour during both training and deployment. This typically requires a simulator and a procedural specification of unsafe behaviour. Unfortunately, a simulator is not always available, and procedurally specifying constraints can be difficult or impossible for many real-world tasks. A recently introduced technique, ReQueST, aims to solve this problem by learning a neural simulator of the environment from safe human trajectories, then using the learned simulator to efficiently learn a reward model from human feedback. However, it is yet unknown whether this approach is feasible in complex 3D environments with feedback obtained from real humans - whether sufficient pixel-based neural simulator quality can be achieved, and whether the human data requirements are viable in terms of both quantity and quality. In this paper we answer this question in the affirmative, using ReQueST to train an agent to perform a 3D first-person object collection task using data entirely from human contractors. We show that the resulting agent exhibits an order of magnitude reduction in unsafe behaviour compared to standard reinforcement learning.

preprint2021arXiv

Imitating Interactive Intelligence

A common vision from science fiction is that robots will one day inhabit our physical spaces, sense the world as we do, assist our physical labours, and communicate with us through natural language. Here we study how to design artificial agents that can interact naturally with humans using the simplification of a virtual environment. This setting nevertheless integrates a number of the central challenges of artificial intelligence (AI) research: complex visual perception and goal-directed physical control, grounded language comprehension and production, and multi-agent social interaction. To build agents that can robustly interact with humans, we would ideally train them while they interact with humans. However, this is presently impractical. Therefore, we approximate the role of the human with another learned agent, and use ideas from inverse reinforcement learning to reduce the disparities between human-human and agent-agent interactive behaviour. Rigorously evaluating our agents poses a great challenge, so we develop a variety of behavioural tests, including evaluation by humans who watch videos of agents or interact directly with them. These evaluations convincingly demonstrate that interactive training and auxiliary losses improve agent behaviour beyond what is achieved by supervised learning of actions alone. Further, we demonstrate that agent capabilities generalise beyond literal experiences in the dataset. Finally, we train evaluation models whose ratings of agents agree well with human judgement, thus permitting the evaluation of new agent models without additional effort. Taken together, our results in this virtual environment provide evidence that large-scale human behavioural imitation is a promising tool to create intelligent, interactive agents, and the challenge of reliably evaluating such agents is possible to surmount.

preprint2016arXiv

The squeezed limit of the bispectrum in multi-field inflation

We calculate the squeezed limit of the bispectrum produced by inflation with multiple light fields. To achieve this we allow for different horizon exit times for each mode and calculate the intrinsic field-space three-point function in the squeezed limit using soft-limit techniques. We then use the $δN$ formalism from the time the last mode exits the horizon to calculate the bispectrum of the primordial curvature perturbation. We apply our results to calculate the spectral index of the halo bias, $n_{δb}$, an important observational probe of the squeezed limit of the primordial bispectrum and compare our results with previous formulae. We give an example of a curvaton model with $n_{δb} \sim {\cal O}(n_s-1)$ for which we find a 20% correction to observable parameters for squeezings relevant to future experiments. For completeness, we also calculate the squeezed limit of three-point correlation functions involving gravitons for multiple field models.

preprint2015arXiv

D-brane Potentials in the Warped Resolved Conifold and Natural Inflation

In this paper we obtain a model of Natural Inflation from string theory with a Planckian decay constant. We investigate D-brane dynamics in the background of the warped resolved conifold (WRC) throat approximation of Type IIB string compactifications on Calabi-Yau manifolds. When we glue the throat to a compact bulk Calabi-Yau, we generate a D-brane potential which is a solution to the Laplace equation on the resolved conifold. We can exactly solve this equation, including dependence on the angular coordinates. The solutions are valid down to the tip of the resolved conifold, which is not the case for the more commonly used deformed conifold. This allows us to exploit the effect of the warping, which is strongest at the tip. We inflate near the tip using an angular coordinate of a D5-brane in the WRC which has a discrete shift symmetry, and feels a cosine potential, giving us a model of Natural Inflation, from which it is possible to get a Planckian decay constant whilst maintaining control over the backreaction. This is because the decay constant for a wrapped brane contains powers of the warp factor, and so can be made large, while the wrapping parameter can be kept small enough so that backreaction is under control.

preprint2015arXiv

Generating the cosmic microwave background power asymmetry with $g_{NL}$

We consider a higher order term in the $δN$ expansion for the CMB power asymmetry generated by a superhorizon isocurvature field fluctuation. The term can generate the asymmetry without requiring a large value of $f_{NL}$. Instead it produces a non-zero value of $g_{NL}$. A combination of constraints lead to an allowed region in $f_{NL}-g_{NL}$ space. To produce the asymmetry with this term without a large value of $f_{NL}$ we find that the isocurvature field needs to contribute less than the inflaton towards the power spectrum of the curvature perturbation.