Source author record

J. Michael Herrmann

J. Michael Herrmann appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2020arXiv

Evolutionary Selective Imitation: Interpretable Agents by Imitation Learning Without a Demonstrator

We propose a new method for training an agent via an evolutionary strategy (ES), in which we iteratively improve a set of samples to imitate: Starting with a random set, in every iteration we replace a subset of the samples with samples from the best trajectories discovered so far. The evaluation procedure for this set is to train, via supervised learning, a randomly initialised neural network (NN) to imitate the set and then execute the acquired policy against the environment. Our method is thus an ES based on a fitness function that expresses the effectiveness of imitating an evolving data subset. This is in contrast to other ES techniques that iterate over the weights of the policy directly. By observing the samples that the agent selects for learning, it is possible to interpret and evaluate the evolving strategy of the agent more explicitly than in NN learning. In our experiments, we trained an agent to solve the OpenAI Gym environment Bipedalwalker-v3 by imitating an evolutionarily selected set of only 25 samples with a NN with only a few thousand parameters. We further test our method on the Procgen game Plunder and show here as well that the proposed method is an interpretable, small, robust and effective alternative to other ES or policy gradient methods.

preprint2015arXiv

Critical Parameters in Particle Swarm Optimisation

Particle swarm optimisation is a metaheuristic algorithm which finds reasonable solutions in a wide range of applied problems if suitable parameters are used. We study the properties of the algorithm in the framework of random dynamical systems which, due to the quasi-linear swarm dynamics, yields analytical results for the stability properties of the particles. Such considerations predict a relationship between the parameters of the algorithm that marks the edge between convergent and divergent behaviours. Comparison with simulations indicates that the algorithm performs best near this margin of instability.

preprint2015arXiv

Small-world structure induced by spike-timing-dependent plasticity in networks with critical dynamics

The small-world property in the context of complex networks implies structural benefits to the processes taking place within a network, such as optimal information transmission and robustness. In this paper, we study a model network of integrate-and-fire neurons that are subject to activity-dependent synaptic plasticity. We find the learning rule that gives rise to a small-world structure when the collective dynamics of the system reaches a critical state which is characterised by power-law distributions of activity clusters. Moreover, by analysing the motif profile of the networks, we observe that bidirectional connectivity is impaired by the effects of this type of plasticity.

preprint2015arXiv

The success of complex networks at criticality

In spiking neural networks an action potential could in principle trigger subsequent spikes in the neighbourhood of the initial neuron. A successful spike is that which trigger subsequent spikes giving rise to cascading behaviour within the system. In this study we introduce a metric to assess the success of spikes emitted by integrate-and-fire neurons arranged in complex topologies and whose collective behaviour is undergoing a phase transition that is identified by neuronal avalanches that become clusters of activation whose distribution of sizes can be approximated by a power-law. In numerical simulations we report that scale-free networks with the small-world property is the structure in which neurons possess more successful spikes. As well, we conclude both analytically and in numerical simulations that fully-connected networks are structures in which neurons perform worse. Additionally, we study how the small-world property affects spiking behaviour and its success in scale-free networks.