Source author record

Thomas House

Thomas House appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

25works
15topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

25 published item(s)

preprint2024arXiv

Modelling and classifying joint trajectories of self-reported mood and pain in a large cohort study

It is well-known that mood and pain interact with each other, however individual-level variability in this relationship has been less well quantified than overall associations between low mood and pain. Here, we leverage the possibilities presented by mobile health data, in particular the "Cloudy with a Chance of Pain" study, which collected longitudinal data from the residents of the UK with chronic pain conditions. Participants used an App to record self-reported measures of factors including mood, pain and sleep quality. The richness of these data allows us to perform model-based clustering of the data as a mixture of Markov processes. Through this analysis we discover four endotypes with distinct patterns of co-evolution of mood and pain over time. The differences between endotypes are sufficiently large to play a role in clinical hypothesis generation for personalised treatments of comorbid pain and low mood.

preprint2022arXiv

NuZZ: numerical Zig-Zag sampling for general models

Markov chain Monte Carlo (MCMC) is a key algorithm in computational statistics, and as datasets grow larger and models grow more complex, many popular MCMC algorithms become too computationally expensive to be practical. Recent progress has been made on this problem through development of MCMC algorithms based on Piecewise Deterministic Markov Processes (PDMPs), irreversible processes that can be engineered to converge at a rate which is independent of the size of data. While there has understandably been a surge of theoretical studies following these results, PDMPs have so far only been implemented for models where certain gradients can be bounded, which is not possible in many statistical contexts. Focusing on the Zig-Zag process, we present the Numerical Zig-Zag (NuZZ) algorithm, which is applicable to general statistical models without the need for bounds on the gradient of the log posterior. This allows us to perform numerical experiments on: (i) how the Zig-Zag dynamics behaves on some test problems with common challenging features; and (ii) how the error between the target and sampled distributions evolves as a function of computational effort for different MCMC algorithms including NuZZ. Moreover, due to the specifics of the NuZZ algorithms, we are able to give an explicit bound on the Wasserstein distance between the exact posterior and its numerically perturbed counterpart in terms of the user-specified numerical tolerances of NuZZ.

preprint2022arXiv

The role of regular asymptomatic testing in reducing the impact of a COVID-19 wave

Testing for infection with SARS-CoV-2 is an important intervention in reducing onwards transmission of COVID-19, particularly when combined with the isolation and contact-tracing of positive cases. Many countries with the capacity to do so have made use of lab-processed Polymerase Chain Reaction (PCR) testing targeted at individuals with symptoms and the contacts of confirmed cases. Alternatively, Lateral Flow Tests (LFTs) are able to deliver a result quickly, without lab-processing and at a relatively low cost. Their adoption can support regular mass asymptomatic testing, allowing earlier detection of infection and isolation of infectious individuals. In this paper we extend and apply the agent-based epidemic modelling framework Covasim to explore the impact of regular asymptomatic testing on the peak and total number of infections in an emerging COVID-19 wave. We explore testing with LFTs at different frequency levels within a population with high levels of immunity and with background symptomatic PCR testing, case isolation and contact tracing for testing. The effectiveness of regular asymptomatic testing was compared with `lockdown' interventions seeking to reduce the number of non-household contacts across the whole population through measures such as mandating working from home and restrictions on gatherings. Since regular asymptomatic testing requires only those with a positive result to reduce contact, while lockdown measures require the whole population to reduce contact, any policy decision that seeks to trade off harms from infection against other harms will not automatically favour one over the other. Our results demonstrate that, where such a trade off is being made, at moderate rates of early exponential growth regular asymptomatic testing has the potential to achieve significant infection control without the wider harms associated with additional lockdown measures.

preprint2020arXiv

Challenges in control of Covid-19: short doubling time and long delay to effect of interventions

Early assessments of the spreading rate of COVID-19 were subject to significant uncertainty, as expected with limited data and difficulties in case ascertainment, but more reliable inferences can now be made. Here, we estimate from European data that COVID-19 cases are expected to double initially every three days, until social distancing interventions slow this growth, and that the impact of such measures is typically only seen nine days - i.e. three doubling times - after their implementation. We argue that such temporal patterns are more critical than precise estimates of the basic reproduction number for initiating interventions. This observation has particular implications for the low- and middle-income countries currently in the early stages of their local epidemics.

preprint2020arXiv

Fast Approximate Bayesian Contextual Cold Start Learning (FAB-COST)

Cold-start is a notoriously difficult problem which can occur in recommendation systems, and arises when there is insufficient information to draw inferences for users or items. To address this challenge, a contextual bandit algorithm -- the Fast Approximate Bayesian Contextual Cold Start Learning algorithm (FAB-COST) -- is proposed, which is designed to provide improved accuracy compared to the traditionally used Laplace approximation in the logistic contextual bandit, while controlling both algorithmic complexity and computational cost. To this end, FAB-COST uses a combination of two moment projection variational methods: Expectation Propagation (EP), which performs well at the cold start, but becomes slow as the amount of data increases; and Assumed Density Filtering (ADF), which has slower growth of computational cost with data size but requires more data to obtain an acceptable level of accuracy. By switching from EP to ADF when the dataset becomes large, it is able to exploit their complementary strengths. The empirical justification for FAB-COST is presented, and systematically compared to other approaches on simulated data. In a benchmark against the Laplace approximation on real data consisting of over $670,000$ impressions from autotrader.co.uk, FAB-COST demonstrates at one point increase of over $16\%$ in user clicks. On the basis of these results, it is argued that FAB-COST is likely to be an attractive approach to cold-start recommendation systems in a variety of contexts.

preprint2020arXiv

Key Questions for Modelling COVID-19 Exit Strategies

Combinations of intense non-pharmaceutical interventions ('lockdowns') were introduced in countries worldwide to reduce SARS-CoV-2 transmission. Many governments have begun to implement lockdown exit strategies that allow restrictions to be relaxed while attempting to control the risk of a surge in cases. Mathematical modelling has played a central role in guiding interventions, but the challenge of designing optimal exit strategies in the face of ongoing transmission is unprecedented. Here, we report discussions from the Isaac Newton Institute 'Models for an exit strategy' workshop (11-15 May 2020). A diverse community of modellers who are providing evidence to governments worldwide were asked to identify the main questions that, if answered, will allow for more accurate predictions of the effects of different exit strategies. Based on these questions, we propose a roadmap to facilitate the development of reliable models to guide exit strategies. The roadmap requires a global collaborative effort from the scientific community and policy-makers, and is made up of three parts: i) improve estimation of key epidemiological parameters; ii) understand sources of heterogeneity in populations; iii) focus on requirements for data collection, particularly in Low-to-Middle-Income countries. This will provide important information for planning exit strategies that balance socio-economic benefits with public health.

preprint2020arXiv

Using statistics and mathematical modelling to understand infectious disease outbreaks: COVID-19 as an example

During an infectious disease outbreak, biases in the data and complexities of the underlying dynamics pose significant challenges in mathematically modelling the outbreak and designing policy. Motivated by the ongoing response to COVID-19, we provide a toolkit of statistical and mathematical models beyond the simple SIR-type differential equation models for analysing the early stages of an outbreak and assessing interventions. In particular, we focus on parameter estimation in the presence of known biases in the data, and the effect of non-pharmaceutical interventions in enclosed subpopulations, such as households and care homes. We illustrate these methods by applying them to the COVID-19 pandemic.

preprint2016arXiv

Stochastic epidemic dynamics on extremely heterogeneous networks

Networks of contacts capable of spreading infectious diseases are often observed to be highly heterogeneous, with the majority of individuals having fewer contacts than the mean, and a significant minority having relatively very many contacts. We derive a two-dimensional diffusion model for the full temporal behavior of the stochastic susceptible-infectious-recovered (SIR) model on such a network, by making use of a time-scale separation in the deterministic limit of the dynamics. This low-dimensional process is an accurate approximation to the full model in the limit of large populations, even for cases when the time-scale separation is not too pronounced, provided the maximum degree is not of the order of the population size.

preprint2015arXiv

Exact and approximate moment closures for non-Markovian network epidemics

Moment-closure techniques are commonly used to generate low-dimensional deterministic models to approximate the average dynamics of stochastic systems on networks. The quality of such closures is usually difficult to asses and the relationship between model assumptions and closure accuracy are often difficult, if not impossible, to quantify. Here we carefully examine some commonly used moment closures, in particular a new one based on the concept of maximum entropy, for approximating the spread of epidemics on networks by reconstructing the probability distributions over triplets based on those over pairs. We consider various models (SI, SIR, SEIR and Reed-Frost-type) under Markovian and non-Markovian assumption characterising the latent and infectious periods. We initially study two special networks, namely the open triplet and closed triangle, for which we can obtain analytical results. We then explore numerically the exactness of moment closures for a wide range of larger motifs, thus gaining understanding of the factors that introduce errors in the approximations, in particular the presence of a random duration of the infectious period and the presence of overlapping triangles in a network. We also derive a simpler and more intuitive proof than previously available concerning the known result that pair-based moment closure is exact for the Markovian SIR model on tree-like networks under pure initial conditions. We also extend such a result to all infectious models, Markovian and non-Markovian, in which susceptibles escape infection independently from each infected neighbour and for which infectives cannot regain susceptible status, provided the network is tree-like and initial conditions are pure. This works represent a valuable step in deepening understanding of the assumptions behind moment closure approximations and for putting them on a more rigorous mathematical footing.

preprint2015arXiv

Near-critical SIR epidemic on a random graph with given degrees

Emergence of new diseases and elimination of existing diseases is a key public health issue. In mathematical models of epidemics, such phenomena involve the process of infections and recoveries passing through a critical threshold where the basic reproductive ratio is 1. In this paper, we study near-critical behaviour in the context of a susceptible-infective-recovered (SIR) epidemic on a random (multi)graph on $n$ vertices with a given degree sequence. We concentrate on the regime just above the threshold for the emergence of a large epidemic, where the basic reproductive ratio is $1 + ω(n) n^{-1/3}$, with $ω(n)$ tending to infinity slowly as the population size, $n$, tends to infinity. We determine the probability that a large epidemic occurs, and the size of a large epidemic. Our results require basic regularity conditions on the degree sequences, and the assumption that the third moment of the degree of a random susceptible vertex stays uniformly bounded as $n \to \infty$. As a corollary, we determine the probability and size of a large near-critical epidemic on a standard binomial random graph in the `sparse' regime, where the average degree is constant. As a further consequence of our method, we obtain an improved result on the size of the giant component in a random graph with given degrees just above the critical window, proving a conjecture by Janson and Luczak.

preprint2015arXiv

Real-time growth rate for general stochastic SIR epidemics on unclustered networks

Networks have become an important tool for infectious disease epidemiology. Most previous theoretical studies of transmission network models have either considered simple Markovian dynamics at the individual level, or have focused on the invasion threshold and final outcome of the epidemic. Here, we provide a general theory for early real-time behaviour of epidemics on large configuration model networks (i.e. static and locally unclustered), in particular focusing on the computation of the Malthusian parameter that describes the early exponential epidemic growth. Analytical, numerical and Monte-Carlo methods under a wide variety of Markovian and non-Markovian assumptions about the infectivity profile are presented. Numerous examples provide explicit quantification of the impact of the network structure on the temporal dynamics of the spread of infection and provide a benchmark for validating results of large scale simulations.

preprint2014arXiv

Algebraic moment closure for population dynamics on discrete structures

Moment closure on general discrete structures often requires one of the following: (i) an absence of short closed loops (zero clustering); (ii) existence of a spatial scale; (iii) ad hoc assumptions. Algebraic methods are presented to avoid the use of such assumptions for populations based on clumps, and are applied to both SIR and macroparasite disease dynamics. One approach involves a series of approximations that can be derived systematically, and another is exact and based on Lie algebraic methods.

preprint2014arXiv

For principled model fitting in mathematical biology

The mathematical models used to capture features of complex, biological systems are typically non-linear, meaning that there are no generally valid simple relationships between their outputs and the data that might be used to validate them. This invalidates the assumptions behind standard statistical methods such as linear regression, and often the methods used to parameterise biological models from data are ad hoc. In this perspective, I will argue for an approach to model fitting in mathematical biology that incorporates modern statistical methodology without losing the insights gained through non-linear dynamic models, and will call such an approach principled model fitting. Principled model fitting therefore involves defining likelihoods of observing real data on the basis of models that capture key biological mechanisms.

preprint2014arXiv

Non-Markovian stochastic epidemics in extremely heterogeneous populations

A feature often observed in epidemiological networks is significant heterogeneity in degree. A popular modelling approach to this has been to consider large populations with highly heterogeneous discrete contact rates. This paper defines an individual-level non-Markovian stochastic process that converges on standard ODE models of such populations in the appropriate asymptotic limit. A generalised Sellke construction is derived for this model, and this is then used to consider final outcomes in the case where heterogeneity follows a truncated Zipf distribution.

preprint2013arXiv

Dynamics of Stochastic Epidemics on Heterogeneous Networks

Epidemic models currently play a central role in our attempts to understand and control infectious diseases. Here, we derive a model for the diffusion limit of stochastic susceptible-infectious-removed (SIR) epidemic dynamics on a heterogeneous network. Using this, we consider analytically the early asymptotic exponential growth phase of such epidemics, showing how the higher order moments of the network degree distribution enter into the stochastic behaviour of the epidemic. We find that the first three moments of the network degree distribution are needed to specify the variance in disease prevalence fully, meaning that the skewness of the degree distribution affects the variance of the prevalence of infection. We compare these asymptotic results to simulation and find a close agreement for city-sized populations.

preprint2013arXiv

Endemic infections are always possible on regular networks

We study the dependence of the largest component in regular networks on the clustering coefficient, showing that its size changes smoothly without undergoing a phase transition. We explain this behaviour via an analytical approach based on the network structure, and provide an exact equation describing the numerical results. Our work indicates that intrinsic structural properties always allow the spread of epidemics on regular networks.

preprint2013arXiv

Higher-order structure and epidemic dynamics in clustered networks

Clustering is typically measured by the ratio of triangles to all triples, open or closed. Generating clustered networks, and how clustering affects dynamics on networks, is reasonably well understood for certain classes of networks \cite{vmclust, karrerclust2010}, e.g., networks composed of lines and non-overlapping triangles. In this paper we show that it is possible to generate networks which, despite having the same degree distribution and equal clustering, exhibit different higher-order structure, specifically, overlapping triangles and other order-four (a closed network motif composed of four nodes) structures. To distinguish and quantify these additional structural features, we develop a new network metric capable of measuring order-four structure which, when used alongside traditional network metrics, allows us to more accurately describe a network's topology. Three network generation algorithms are considered: a modified configuration model and two rewiring algorithms. By generating homogeneous networks with equal clustering we study and quantify their structural differences, and using SIS (Susceptible-Infected-Susceptible) and SIR (Susceptible-Infected-Recovered) dynamics we investigate computationally how differences in higher-order structure impact on epidemic threshold, final epidemic or prevalence levels and time evolution of epidemics. Our results suggest that characterising and measuring higher-order network structure is needed to advance our understanding of the impact of network topology on dynamics unfolding on the networks.

preprint2013arXiv

The rate of convergence to early asymptotic behaviour in age-structured epidemic models

Age structure is incorporated in many types of epidemic model. Often it is convenient to assume that such models converge to early asymptotic behaviour quickly, before the susceptible population has been appreciably depleted. We make use of dynamical systems theory to show that for some reasonable parameter values, this convergence can be slow. Such a possibility should therefore be considered when parameterising age-structured epidemic models.

preprint2012arXiv

Exact epidemic dynamics for generally clustered, complex networks

The last few years have seen remarkably fast progress in the understanding of statistics and epidemic dynamics of various clustered networks. This paper considers a class of networks based around a concept (the locale) that allows asymptotically exact results to be derived for epidemic dynamics. While there is no restriction on the motifs that can be found in such graphs, each node must be uniquely assigned to a generally clustered subgraph to obtain analytic traction.

preprint2011arXiv

Lie algebra solution of population models based on time-inhomogeneous Markov chains

Many natural populations are well modelled through time-inhomogeneous stochastic processes. Such processes have been analysed in the physical sciences using a method based on Lie algebras, but this methodology is not widely used for models with ecological, medical and social applications. This paper presents the Lie algebraic method, and applies it to three biologically well motivated examples. The result of this is a solution form that is often highly computationally advantageous.

preprint2011arXiv

Modelling Epidemics on Networks

Infectious disease remains, despite centuries of work to control and mitigate its effects, a major problem facing humanity. This paper reviews the mathematical modelling of infectious disease epidemics on networks, starting from the simplest Erdos-Renyi random graphs, and building up structure in the form of correlations, heterogeneity and preference, paying particular attention to the links between random graph theory, percolation and dynamical systems representing transmission. Finally, the problems posed by networks with a large number of short closed looks are discussed.

preprint2010arXiv

Epidemic prediction and control in clustered populations

There has been much recent interest in modelling epidemics on networks, particularly in the presence of substantial clustering. Here, we develop pairwise methods to answer questions that are often addressed using epidemic models, in particular: on the basis of potential observations early in an outbreak, what can be predicted about the epidemic outcomes and the levels of intervention necessary to control the epidemic? We find that while some results are independent of the level of clustering (early growth predicts the level of `leaky' vaccine needed for control and peak time, while the basic reproductive ratio predicts the random vaccination threshold) the relationship between other quantities is very sensitive to clustering.

preprint2010arXiv

Generalised network clustering and its dynamical implications

A parameterisation of generalised network clustering, in the form of four-motif prevalences, is presented. This involves three real parameters that are conditional on one- two- and three-motif prevalences. Interpretations of these real parameters are presented that motivate a set of rewiring schemes to create appropriately clustered networks. Finally, the dynamical implications of higher order structure, as parameterised, for a contact process are considered.

preprint2010arXiv

Networks and the Epidemiology of Infectious Disease

The science of networks has revolutionised research into the dynamics of interacting elements. It could be argued that epidemiology in particular has embraced the potential of network theory more than any other discipline. Here we review the growing body of research concerning the spread of infectious diseases on networks, focusing on the interplay between network theory and epidemiology. The review is split into four main sections, which examine: the types of network relevant to epidemiology; the multitude of ways these networks can be characterised; the statistical methods that can be applied to infer the epidemiological parameters on a realised network; and finally simulation and analytical methods to determine epidemic dynamics on a given network. Given the breadth of areas covered and the ever-expanding number of publications, a comprehensive review of all work is impossible. Instead, we provide a personalised overview into the areas of network epidemiology that have seen the greatest progress in recent years or have the greatest potential to provide novel insights. As such, considerable importance is placed on analytical approaches and statistical methods which are both rapidly expanding fields. Throughout this review we restrict our attention to epidemiological issues.