Source author record

David Howard

David Howard appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

14works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

14 published item(s)

preprint2022arXiv

Assessing Evolutionary Terrain Generation Methods for Curriculum Reinforcement Learning

Curriculum learning allows complex tasks to be mastered via incremental progression over `stepping stone' goals towards a final desired behaviour. Typical implementations learn locomotion policies for challenging environments through gradual complexification of a terrain mesh generated through a parameterised noise function. To date, researchers have predominantly generated terrains from a limited range of noise functions, and the effect of the generator on the learning process is underrepresented in the literature. We compare popular noise-based terrain generators to two indirect encodings, CPPN and GAN. To allow direct comparison between both direct and indirect representations, we assess the impact of a range of representation-agnostic MAP-Elites feature descriptors that compute metrics directly from the generated terrain meshes. Next, performance and coverage are assessed when training a humanoid robot in a physics simulator using the PPO algorithm. Results describe key differences between the generators that inform their use in curriculum learning, and present a range of useful feature descriptors for uptake by the community.

preprint2022arXiv

EvoRobogami: Co-designing with Humans in Evolutionary Robotics Experiments

We study the effects of injecting human-generated designs into the initial population of an evolutionary robotics experiment, where subsequent population of robots are optimised via a Genetic Algorithm and MAP-Elites. First, human participants interact via a graphical front-end to explore a directly-parameterised legged robot design space and attempt to produce robots via a combination of intuition and trial-and-error that perform well in a range of environments. Environments are generated whose corresponding high-performance robot designs range from intuitive to complex and hard to grasp. Once the human designs have been collected, their impact on the evolutionary process is assessed by replacing a varying number of designs in the initial population with human designs and subsequently running the evolutionary algorithm. Our results suggest that a balance of random and hand-designed initial solutions provides the best performance for the problems considered, and that human designs are most valuable when the problem is intuitive. The influence of human design in an evolutionary algorithm is a highly understudied area, and the insights in this paper may be valuable to the area of AI-based design more generally.

preprint2020arXiv

Diversity-based Design Assist for Large Legged Robots

We combine MAP-Elites and highly parallelisable simulation to explore the design space of a class of large legged robots, which stand at around 2m tall and whose design and construction is not well-studied. The simulation is modified to account for factors such as motor torque and weight, and presents a reasonable fidelity search space. A novel robot encoding allows for bio-inspired features such as legs scaling along the length of the body. The impact of three possible control generation schemes are assessed in the context of body-brain co-evolution, showing that even constrained problems benefit strongly from coupling-promoting mechanisms. A two stage process in implemented. In the first stage, a library of possible robots is generated, treating user requirements as constraints. In the second stage, the most promising robot niches are analysed and a suite of human-understandable design rules generated related to the values of their feature variables. These rules, together with the library, are then ready to be used by a (human) robot designer as a Design Assist tool.

preprint2020arXiv

Path Towards Multilevel Evolution of Robots

Multi-level evolution is a bottom-up robotic design paradigm which decomposes the design problem into layered sub-tasks that involve concurrent search for appropriate materials, component geometry and overall morphology. Each of the three layers operate with the goal of building a library of diverse candidate solutions which will be used either as building blocks for the layer above or provided to the decision maker for final use. In this paper we provide a theoretical discussion on the concepts and technologies that could potentially be used as building blocks for this framework.

preprint2020arXiv

Real World Morphological Evolution is Feasible

Evolutionary algorithms offer great promise for the automatic design of robot bodies, tailoring them to specific environments or tasks. Most research is done on simplified models or virtual robots in physics simulators, which do not capture the natural noise and richness of the real world. Very few of these virtual robots are built as physical robots, and the few that are will rarely be further improved in the actual environment they operate in, limiting the effectiveness of the automatic design process. We utilize our shape-shifting quadruped robot, which allows us to optimize the design in its real-world environment. The robot is able to change the length of its legs during operation, and is robust enough for complex experiments and tasks. We have co-evolved control and morphology in several different scenarios, and have seen that the algorithm is able to exploit the dynamic morphology solely through real-world experiments.

preprint2020arXiv

Towards Crossing the Reality Gap with Evolved Plastic Neurocontrollers

A critical issue in evolutionary robotics is the transfer of controllers learned in simulation to reality. This is especially the case for small Unmanned Aerial Vehicles (UAVs), as the platforms are highly dynamic and susceptible to breakage. Previous approaches often require simulation models with a high level of accuracy, otherwise significant errors may arise when the well-designed controller is being deployed onto the targeted platform. Here we try to overcome the transfer problem from a different perspective, by designing a spiking neurocontroller which uses synaptic plasticity to cross the reality gap via online adaptation. Through a set of experiments we show that the evolved plastic spiking controller can maintain its functionality by self-adapting to model changes that take place after evolutionary training, and consequently exhibit better performance than its non-plastic counterpart.

preprint2020arXiv

Traversing the Reality Gap via Simulator Tuning

The large demand for simulated data has made the reality gap a problem on the forefront of robotics. We propose a method to traverse the gap by tuning available simulation parameters. Through the optimisation of physics engine parameters, we show that we are able to narrow the gap between simulated solutions and a real world dataset, and thus allow more ready transfer of leaned behaviours between the two. We subsequently gain understanding as to the importance of specific simulator parameters, which is of broad interest to the robotic machine learning community. We find that even optimised for different tasks that different physics engine perform better in certain scenarios and that friction and maximum actuator velocity are tightly bounded parameters that greatly impact the transference of simulated solutions.

preprint2017arXiv

Differential Evolution and Bayesian Optimisation for Hyper-Parameter Selection in Mixed-Signal Neuromorphic Circuits Applied to UAV Obstacle Avoidance

The Lobula Giant Movement Detector (LGMD) is a an identified neuron of the locust that detects looming objects and triggers its escape responses. Understanding the neural principles and networks that lead to these fast and robust responses can lead to the design of efficient facilitate obstacle avoidance strategies in robotic applications. Here we present a neuromorphic spiking neural network model of the LGMD driven by the output of a neuromorphic Dynamic Vision Sensor (DVS), which has been optimised to produce robust and reliable responses in the face of the constraints and variability of its mixed signal analogue-digital circuits. As this LGMD model has many parameters, we use the Differential Evolution (DE) algorithm to optimise its parameter space. We also investigate the use of Self-Adaptive Differential Evolution (SADE) which has been shown to ameliorate the difficulties of finding appropriate input parameters for DE. We explore the use of two biological mechanisms: synaptic plasticity and membrane adaptivity in the LGMD. We apply DE and SADE to find parameters best suited for an obstacle avoidance system on an unmanned aerial vehicle (UAV), and show how it outperforms state-of-the-art Bayesian optimisation used for comparison.

preprint2016arXiv

A rainbow $r$-partite version of the Erdős-Ko-Rado theorem

Let $f(n,r,k)$ be the minimal number such that every hypergraph larger than $f(n,r,k)$ contained in $\binom{[n]}{r}$ contains a matching of size $k$, and let $g(n,r,k)$ be the minimal number such that every hypergraph larger than $g(n,r,k)$ contained in the $r$-partite $r$-graph $[n]^{r}$ contains a matching of size $k$. The Erdős-Ko-Rado theorem states that $f(n,r,2)=\binom{n-1}{r-1}$~~($r \le \frac{n}{2}$) and it is easy to show that $g(n,r,k)=(k-1)n^{r-1}$. The conjecture inspiring this paper is that if $F_1,F_2,\ldots,F_k\subseteq \binom{[n]}{r}$ are of size larger than $f(n,r,k)$ or $F_1,F_2,\ldots,F_k\subseteq [n]^{r}$ are of size larger than $g(n,r,k)$ then there exists a rainbow matching, i.e. a choice of disjoint edges $f_i \in F_i$. In this paper we deal mainly with the second part of the conjecture, and prove it for $r\le 3$. \vspace{.1cm} We also prove that for every $r$ and $k$ there exists $n_0=n_0(r,k)$ such that the $r$-partite version of the conjecture is true for $n>n_0$.

preprint2016arXiv

Cross-intersecting pairs of hypergraphs

Two hypergraphs $H_1,\ H_2$ are called {\em cross-intersecting} if $e_1 \cap e_2 \neq \emptyset$ for every pair of edges $e_1 \in H_1,~e_2 \in H_2$. Each of the hypergraphs is then said to {\em block} the other. Given parameters $n,r,m$ we determine the maximal size of a sub-hypergraph of $[n]^r$ (meaning that it is $r$-partite, with all sides of size $n$) for which there exists a blocking sub-hypergraph of $[n]^r$ of size $m$. The answer involves a fractal-like (that is, self-similar) sequence, first studied by Knuth. We also study the same question with $\binom{n}{r}$ replacing $[n]^r$.

preprint2015arXiv

A Cognitive Architecture Based on a Learning Classifier System with Spiking Classifiers

Learning Classifier Systems (LCS) are population-based reinforcement learners that were originally designed to model various cognitive phenomena. This paper presents an explicitly cognitive LCS by using spiking neural networks as classifiers, providing each classifier with a measure of temporal dynamism. We employ a constructivist model of growth of both neurons and synaptic connections, which permits a Genetic Algorithm (GA) to automatically evolve sufficiently-complex neural structures. The spiking classifiers are coupled with a temporally-sensitive reinforcement learning algorithm, which allows the system to perform temporal state decomposition by appropriately rewarding "macro-actions," created by chaining together multiple atomic actions. The combination of temporal reinforcement learning and neural information processing is shown to outperform benchmark neural classifier systems, and successfully solve a robotic navigation task.

preprint2015arXiv

Evolving Unipolar Memristor Spiking Neural Networks

Neuromorphic computing --- brainlike computing in hardware --- typically requires myriad CMOS spiking neurons interconnected by a dense mesh of nanoscale plastic synapses. Memristors are frequently citepd as strong synapse candidates due to their statefulness and potential for low-power implementations. To date, plentiful research has focused on the bipolar memristor synapse, which is capable of incremental weight alterations and can provide adaptive self-organisation under a Hebbian learning scheme. In this paper we consider the Unipolar memristor synapse --- a device capable of non-Hebbian switching between only two states (conductive and resistive) through application of a suitable input voltage --- and discuss its suitability for neuromorphic systems. A self-adaptive evolutionary process is used to autonomously find highly fit network configurations. Experimentation on a two robotics tasks shows that unipolar memristor networks evolve task-solving controllers faster than both bipolar memristor networks and networks containing constant nonplastic connections whilst performing at least comparably.

preprint2013arXiv

On a Generalization of the Ryser-Brualdi-Stein Conjecture

A rainbow matching for (not necessarily distinct) sets F_1,...,F_k of hypergraph edges is a matching consisting of k edges, one from each F_i. The aim of the paper is twofold - to put order in the multitude of conjectures that relate to this concept (some of them first presented here), and to present some partial results on one of these conjectures, that seems central among them.

preprint2012arXiv

Revolutionaries and Spies

Let $G = (V,E)$ be a graph and let $r,s,k$ be positive integers. "Revolutionaries and Spies", denoted $\cG(G,r,s,k)$, is the following two-player game. The sets of positions for player 1 and player 2 are $V^r$ and $V^s$ respectively. Each coordinate in $p \in V^r$ gives the location of a "revolutionary" in $G$. Similarly player 2 controls $s$ "spies". We say $u, u' \in V(G)^n$ are adjacent, $u \sim u'$, if for all $1 \leq i \leq n$, $u_i = u'_i$ or ${u_i,u'_i} \in E(G)$. In round 0 player 1 picks $p_0 \in V^r$ and then player 2 picks $q_0 \in V^s$. In each round $i \geq 1$ player 1 moves to $p_i \sim p_{i-1}$ and then player 2 moves to $q_i \sim q_{i-1}$. Player 1 wins the game if he can place $k$ revolutionaries on a vertex $v$ in such a way that player 1 cannot place a spy on $v$ in his following move. Player 2 wins the game if he can prevent this outcome. Let $s(G,r,k)$ be the minimum $s$ such that player 2 can win $\cG(G,r,s,k)$. We show that for $d \geq 2$, $s(\Z^d,r,2)\geq 6 \lfloor \frac{r}{8} \rfloor$. Here $a,b \in \Z^{d}$ with $a \neq b$ are connected by an edge if and only if $|a_i - b_i| \leq 1$ for all $i$ with $1 \leq i \leq d$.