Source author record

Terry Lyons

Terry Lyons appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

33works
18topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

33 published item(s)

preprint2022arXiv

An asymptotic radius of convergence for the Loewner equation and simulation of $SLE_k$ traces via splitting

In this paper, we shall study the convergence of Taylor approximations for the backward Loewner differential equation (driven by Brownian motion) near the origin. More concretely, whenever the initial condition of the backward Loewner equation (which lies in the upper half plane) is small and has the form $Z_{0} = \varepsilon i$, we show these approximations exhibit an $O(\varepsilon)$ error provided the time horizon is $\varepsilon^{2+δ}$ for $δ> 0$. Statements of this theorem will be given using both rough path and $L^{2}(\mathbb{P})$ estimates. Furthermore, over the time horizon of $\varepsilon^{2-δ}$, we shall see that "higher degree" terms within the Taylor expansion become larger than "lower degree" terms for small $\varepsilon$. In this sense, the time horizon on which approximations are accurate scales like $\varepsilon^{2}$. This scaling comes naturally from the Loewner equation when growing vector field derivatives are balanced against decaying iterated integrals of the Brownian motion. As well as being of theoretical interest, this scaling may be used as a guiding principle for developing adaptive step size strategies which perform efficiently near the origin. In addition, this result highlights the limitations of using stochastic Taylor methods (such as the Euler-Maruyama and Milstein methods) for approximating $SLE_κ$ traces. Due to the analytically tractable vector fields of the Loewner equation, we will show Ninomiya-Victoir (or Strang) splitting is particularly well suited for SLE simulation. As the singularity at the origin can lead to large numerical errors, we shall employ the adaptive step size proposed by Tom Kennedy to discretize $SLE_κ$ traces using this splitting. We believe that the Ninomiya-Victoir scheme is the first high order numerical method that has been successfully applied to $SLE_κ$ traces.

preprint2022arXiv

ImageSig: A signature transform for ultra-lightweight image recognition

This paper introduces a new lightweight method for image recognition. ImageSig is based on computing signatures and does not require a convolutional structure or an attention-based encoder. It is striking to the authors that it achieves: a) an accuracy for 64 X 64 RGB images that exceeds many of the state-of-the-art methods and simultaneously b) requires orders of magnitude less FLOPS, power and memory footprint. The pretrained model can be as small as 44.2 KB in size. ImageSig shows unprecedented performance on hardware such as Raspberry Pi and Jetson-nano. ImageSig treats images as streams with multiple channels. These streams are parameterized by spatial directions. We contribute to the functionality of signature and rough path theory to stream-like data and vision tasks on static images beyond temporal streams. With very few parameters and small size models, the key advantage is that one could have many of these "detectors" assembled on the same chip; moreover, the feature acquisition can be performed once and shared between different models of different tasks - further accelerating the process. This contributes to energy efficiency and the advancements of embedded AI at the edge.

preprint2021arXiv

A Generalised Signature Method for Multivariate Time Series Feature Extraction

The 'signature method' refers to a collection of feature extraction techniques for multivariate time series, derived from the theory of controlled differential equations. There is a great deal of flexibility as to how this method can be applied. On the one hand, this flexibility allows the method to be tailored to specific problems, but on the other hand, can make precise application challenging. This paper makes two contributions. First, the variations on the signature method are unified into a general approach, the \emph{generalised signature method}, of which previous variations are special cases. A primary aim of this unifying framework is to make the signature method more accessible to any machine learning practitioner, whereas it is now mostly used by specialists. Second, and within this framework, we derive a canonical collection of choices that provide a domain-agnostic starting point. We derive these choices as a result of an extensive empirical study on 26 datasets and go on to show competitive performance against current benchmarks for multivariate time series classification. Finally, to ease practical application, we make our techniques available as part of the open-source [redacted] project.

preprint2021arXiv

Estimating the probability that a given vector is in the convex hull of a random sample

For a $d$-dimensional random vector $X$, let $p_{n, X}(θ)$ be the probability that the convex hull of $n$ independent copies of $X$ contains a given point $θ$. We provide several sharp inequalities regarding $p_{n, X}(θ)$ and $N_X(θ)$ denoting the smallest $n$ for which $p_{n, X}(θ)\ge1/2$. As a main result, we derive the totally general inequality $1/2 \le α_X(θ)N_X(θ)\le 3d + 1$, where $α_X(θ)$ (a.k.a. the Tukey depth) is the minimum probability that $X$ is in a fixed closed halfspace containing the point $θ$. We also show several applications of our general results: one is a moment-based bound on $N_X(\mathbb{E}[X])$, which is an important quantity in randomized approaches to cubature construction or measure reduction problem. Another application is the determination of the canonical convex body included in a random convex polytope given by independent copies of $X$, where our combinatorial approach allows us to generalize existing results in random matrix community significantly.

preprint2021arXiv

Modelling Paralinguistic Properties in Conversational Speech to Detect Bipolar Disorder and Borderline Personality Disorder

Bipolar disorder (BD) and borderline personality disorder (BPD) are two chronic mental health conditions that clinicians find challenging to distinguish based on clinical interviews, due to their overlapping symptoms. In this work, we investigate the automatic detection of these two conditions by modelling both verbal and non-verbal cues in a set of interviews. We propose a new approach of modelling short-term features with visibility-signature transform, and compare it with widely used high-level statistical functions. We demonstrate the superior performance of our proposed signature-based model. Furthermore, we show the role of different sets of features in characterising BD and BPD.

preprint2021arXiv

Signatory: differentiable computations of the signature and logsignature transforms, on both CPU and GPU

Signatory is a library for calculating and performing functionality related to the signature and logsignature transforms. The focus is on machine learning, and as such includes features such as CPU parallelism, GPU support, and backpropagation. To our knowledge it is the first GPU-capable library for these operations. Signatory implements new features not available in previous libraries, such as efficient precomputation strategies. Furthermore, several novel algorithmic improvements are introduced, producing substantial real-world speedups even on the CPU without parallelism. The library operates as a Python wrapper around C++, and is compatible with the PyTorch ecosystem. It may be installed directly via \texttt{pip}. Source code, documentation, examples, benchmarks and tests may be found at \texttt{\url{https://github.com/patrick-kidger/signatory}}. The license is Apache-2.0.

preprint2021arXiv

The shifted ODE method for underdamped Langevin MCMC

In this paper, we consider the underdamped Langevin diffusion (ULD) and propose a numerical approximation using its associated ordinary differential equation (ODE). When used as a Markov Chain Monte Carlo (MCMC) algorithm, we show that the ODE approximation achieves a $2$-Wasserstein error of $\varepsilon$ in $\mathcal{O}\big(d^{\frac{1}{3}}/\varepsilon^{\frac{2}{3}}\big)$ steps under the standard smoothness and strong convexity assumptions on the target distribution. This matches the complexity of the randomized midpoint method proposed by Shen and Lee [NeurIPS 2019] which was shown to be order optimal by Cao, Lu and Wang. However, the main feature of the proposed numerical method is that it can utilize additional smoothness of the target log-density $f$. More concretely, we show that the ODE approximation achieves a $2$-Wasserstein error of $\varepsilon$ in $\mathcal{O}\big(d^{\frac{2}{5}}/\varepsilon^{\frac{2}{5}}\big)$ and $\mathcal{O}\big(\sqrt{d}/\varepsilon^{\frac{1}{3}}\big)$ steps when Lipschitz continuity is assumed for the Hessian and third derivative of $f$. By discretizing this ODE using a third order Runge-Kutta method, we can obtain a practical MCMC method that uses just two additional gradient evaluations per step. In our experiment, where the target comes from a logistic regression, this method shows faster convergence compared to other unadjusted Langevin MCMC algorithms.

preprint2020arXiv

A Data-driven Market Simulator for Small Data Environments

Neural network based data-driven market simulation unveils a new and flexible way of modelling financial time series without imposing assumptions on the underlying stochastic dynamics. Though in this sense generative market simulation is model-free, the concrete modelling choices are nevertheless decisive for the features of the simulated paths. We give a brief overview of currently used generative modelling approaches and performance evaluation metrics for financial time series, and address some of the challenges to achieve good results in the latter. We also contrast some classical approaches of market simulation with simulation based on generative modelling and highlight some advantages and pitfalls of the new approach. While most generative models tend to rely on large amounts of training data, we present here a generative model that works reliably in environments where the amount of available training data is notoriously small. Furthermore, we show how a rough paths perspective combined with a parsimonious Variational Autoencoder framework provides a powerful way for encoding and evaluating financial time series in such environments where available training data is scarce. Finally, we also propose a suitable performance evaluation metric for financial time series and discuss some connections of our Market Generator to deep hedging.

preprint2020arXiv

An optimal polynomial approximation of Brownian motion

In this paper, we will present a strong (or pathwise) approximation of standard Brownian motion by a class of orthogonal polynomials. The coefficients that are obtained from the expansion of Brownian motion in this polynomial basis are independent Gaussian random variables. Therefore it is practical (requires $N$ independent Gaussian coefficients) to generate an approximate sample path of Brownian motion that respects integration of polynomials with degree less than $N$. Moreover, since these orthogonal polynomials appear naturally as eigenfunctions of an integral operator defined by the Brownian bridge covariance function, the proposed approximation is optimal in a certain weighted $L^{2}(\mathbb{P})$ sense. In addition, discretizing Brownian paths as piecewise parabolas gives a locally higher order numerical method for stochastic differential equations (SDEs) when compared to the standard piecewise linear approach. We shall demonstrate these ideas by simulating Inhomogeneous Geometric Brownian Motion (IGBM). This numerical example will also illustrate the deficiencies of the piecewise parabola approximation when compared to a new version of the asymptotically efficient log-ODE (or Castell-Gaines) method.

preprint2020arXiv

Generalised Interpretable Shapelets for Irregular Time Series

The shapelet transform is a form of feature extraction for time series, in which a time series is described by its similarity to each of a collection of `shapelets'. However it has previously suffered from a number of limitations, such as being limited to regularly-spaced fully-observed time series, and having to choose between efficient training and interpretability. Here, we extend the method to continuous time, and in doing so handle the general case of irregularly-sampled partially-observed multivariate time series. Furthermore, we show that a simple regularisation penalty may be used to train efficiently without sacrificing interpretability. The continuous-time formulation additionally allows for learning the length of each shapelet (previously a discrete object) in a differentiable manner. Finally, we demonstrate that the measure of similarity between time series may be generalised to a learnt pseudometric. We validate our method by demonstrating its performance and interpretability on several datasets; for example we discover (purely from data) that the digits 5 and 6 may be distinguished by the chirality of their bottom loop, and that a kind of spectral gap exists in spoken audio classification.

preprint2020arXiv

Numerical method for model-free pricing of exotic derivatives using rough path signatures

We estimate prices of exotic options in a discrete-time model-free setting when the trader has access to market prices of a rich enough class of exotic and vanilla options. This is achieved by estimating an unobservable quantity called "implied expected signature" from such market prices, which are used to price other exotic derivatives. The implied expected signature is an object that characterises the market dynamics.

preprint2020arXiv

Universal Approximation with Deep Narrow Networks

The classical Universal Approximation Theorem holds for neural networks of arbitrary width and bounded depth. Here we consider the natural `dual' scenario for networks of bounded width and arbitrary depth. Precisely, let $n$ be the number of inputs neurons, $m$ be the number of output neurons, and let $ρ$ be any nonaffine continuous function, with a continuous nonzero derivative at some point. Then we show that the class of neural networks of arbitrary depth, width $n + m + 2$, and activation function $ρ$, is dense in $C(K; \mathbb{R}^m)$ for $K \subseteq \mathbb{R}^n$ with $K$ compact. This covers every activation function possible to use in practice, and also includes polynomial activation functions, which is unlike the classical version of the theorem, and provides a qualitative difference between deep narrow networks and shallow wide networks. We then consider several extensions of this result. In particular we consider nowhere differentiable activation functions, density in noncompact domains with respect to the $L^p$-norm, and how the width may be reduced to just $n + m + 1$ for `most' activation functions.

preprint2016arXiv

Discretely sampled signals and the rough Hoff process

We introduce a canonical method for transforming a discrete sequential data set into an associated rough path made up of lead-lag increments. In particular, by sampling a $d$-dimensional continuous semimartingale $X:[0,1] \rightarrow \mathbb{R}^d$ at a set of times $D=(t_i)$, we construct a piecewise linear, axis-directed process $X^D: [0,1] \rightarrow\mathbb{R}^{2d}$ comprised of a past and future component. We call such an object the Hoff process associated with the discrete data $\{X_{t}\}_{t_i\in D}$. The Hoff process can be lifted to its natural rough path enhancement and we consider the question of convergence as the sampling frequency increases. We prove that the Itô integral can be recovered from a sequence of random ODEs driven by the components of $X^D$. This is in contrast to the usual Stratonovich integral limit suggested by the classical Wong-Zakai Theorem. Such random ODEs have a natural interpretation in the context of mathematical finance.

preprint2016arXiv

Learning from the past, predicting the statistics for the future, learning an evolving system

We bring the theory of rough paths to the study of non-parametric statistics on streamed data. We discuss the problem of regression where the input variable is a stream of information, and the dependent response is also (potentially) a stream. A certain graded feature set of a stream, known in the rough path literature as the signature, has a universality that allows formally, linear regression to be used to characterise the functional relationship between independent explanatory variables and the conditional distribution of the dependent response. This approach, via linear regression on the signature of the stream, is almost totally general, and yet it still allows explicit computation. The grading allows truncation of the feature set and so leads to an efficient local description for streams (rough paths). In the statistical context this method offers potentially significant, even transformational dimension reduction. By way of illustration, our approach is applied to stationary time series including the familiar AR model and ARCH model. In the numerical examples we examined, our predictions achieve similar accuracy to the Gaussian Process (GP) approach with much lower computational cost especially when the sample size is large.

preprint2015arXiv

Expected signature of Brownian motion up to the first exit time from a bounded domain

The signature of a path provides a top down description of the path in terms of its effects as a control [Differential Equations Driven by Rough Paths (2007) Springer]. The signature transforms a path into a group-like element in the tensor algebra and is an essential object in rough path theory. The expected signature of a stochastic process plays a similar role to that played by the characteristic function of a random variable. In [Chevyrev (2013)], it is proved that under certain boundedness conditions, the expected value of a random signature already determines the law of this random signature. It becomes of great interest to be able to compute examples of expected signatures and obtain the upper bounds for the decay rates of expected signatures. For instance, the computation for Brownian motion on $[0,1]$ leads to the ``cubature on Wiener space'' methodology [Lyons and Victoir, Proc. R. Soc. Lond. Ser. A Math. Phys. Eng. Sci. 460 (2004) 169-198]. In this paper we fix a bounded domain $Γ$ in a Euclidean space $E$ and study the expected signature of a Brownian path starting at $z\inΓ$ and stopped at the first exit time from $Γ$. We denote this tensor series valued function by $Φ_Γ(z)$ and focus on the case $E=\mathbb{R}^d$. We show that $Φ_Γ(z)$ satisfies an elliptic PDE system and a boundary condition. The equations determining $Φ_Γ$ can be recursively solved; by an iterative application of Sobolev estimates we are able, under certain smoothness and boundedness condition of the domain $Γ$, to prove geometric bounds for the terms in $Φ_Γ(z)$. However, there is still a gap and we have not shown that $Φ_Γ(z)$ determines the law of the signature of this stopped Brownian motion even if $Γ$ is a unit ball.

preprint2015arXiv

Pathwise approximation of SDEs by coupling piecewise abelian rough paths

We present a new pathwise approximation scheme for stochastic differential equations driven by multidimensional Brownian motion which does not require the simulation of Lévy area and has a Wasserstein convergence rate better than the Euler scheme's strong error rate of $O(\sqrt{h})$, where $h$ is the step-size. By using rough path theory we avoid imposing any non-degenerate Hörmander or ellipticity assumptions on the vector fields of the SDE, in contrast to the similar papers of Alfonsi, Davie, and Malliavin et al. The scheme is based on the log-ODE method with the Lévy area increments replaced by Gaussian approximations with the same covariance structure. The Wasserstein coupling is achieved by making small changes to the argument of Davie, the latter being an extension of the Komlós-Major-Tusnády Theorem. We prove that the convergence of the scheme in the Wasserstein metric is of the order $O(h^{1-2/γ-\varepsilon})$ when the vector fields are $γ$-Lipschitz in the sense of Stein.

preprint2015arXiv

The adaptive patched cubature filter and its implementation

There are numerous contexts where one wishes to describe the state of a randomly evolving system. Effective solutions combine models that quantify the underlying uncertainty with available observational data to form scientifically reasonable estimates for the uncertainty in the system state. Stochastic differential equations are often used to mathematically model the underlying system. The Kusuoka-Lyons-Victoir (KLV) approach is a higher order particle method for approximating the weak solution of a stochastic differential equation that uses a weighted set of scenarios to approximate the evolving probability distribution to a high order of accuracy. The algorithm can be performed by integrating along a number of carefully selected bounded variation paths. The iterated application of the KLV method has a tendency for the number of particles to increase. This can be addressed and, together with local dynamic recombination, which simplifies the support of discrete measure without harming the accuracy of the approximation, the KLV method becomes eligible to solve the filtering problem in contexts where one desires to maintain an accurate description of the ever-evolving conditioned measure. In addition to the alternate application of the KLV method and recombination, we make use of the smooth nature of the likelihood function and high order accuracy of the approximations to lead some of the particles immediately to the next observation time and to build into the algorithm a form of automatic high order adaptive importance sampling.

preprint2015arXiv

The Signature of a Rough Path: Uniqueness

In the context of controlled differential equations, the signature is the exponential function on paths. B. Hambly and T. Lyons proved that the signature of a bounded variation path is trivial if and only if the path is tree-like. We extend Hambly-Lyons' result and their notion of tree-like paths to the setting of weakly geometric rough paths in a Banach space. At the heart of our approach is a new definition for reduced path and a lemma identifying the reduced path group with the space of signatures.

preprint2015arXiv

The theory of rough paths via one-forms and the extension of an argument of Schwartz to rough differential equations

We give an overview of the recent approach to the integration of rough paths that reduces the problem to classical Young integration. As an application, we extend an argument of Schwartz to rough differential equations, and prove the existence, uniqueness and continuity of the solution, which is applicable when the driving path takes values in nilpotent Lie group or Butcher group.

preprint2014arXiv

Extracting information from the signature of a financial data stream

Market events such as order placement and order cancellation are examples of the complex and substantial flow of data that surrounds a modern financial engineer. New mathematical techniques, developed to describe the interactions of complex oscillatory systems (known as the theory of rough paths) provides new tools for analysing and describing these data streams and extracting the vital information. In this paper we illustrate how a very small number of coefficients obtained from the signature of financial data can be sufficient to classify this data for subtle underlying features and make useful predictions. This paper presents financial examples in which we learn from data and then proceed to classify fresh streams. The classification is based on features of streams that are specified through the coordinates of the signature of the path. At a mathematical level the signature is a faithful transform of a multidimensional time series. (Ben Hambly and Terry Lyons \cite{uniqueSig}), Hao Ni and Terry Lyons \cite{NiLyons} introduced the possibility of its use to understand financial data and pointed to the potential this approach has for machine learning and prediction. We evaluate and refine these theoretical suggestions against practical examples of interest and present a few motivating experiments which demonstrate information the signature can easily capture in a non-parametric way avoiding traditional statistical modelling of the data. In the first experiment we identify atypical market behaviour across standard 30-minute time buckets sampled from the WTI crude oil future market (NYMEX). The second and third experiments aim to characterise the market "impact" of and distinguish between parent orders generated by two different trade execution algorithms on the FTSE 100 Index futures market listed on NYSE Liffe.

preprint2014arXiv

Rough paths, Signatures and the modelling of functions on streams

Rough path theory is focused on capturing and making precise the interactions between highly oscillatory and non-linear systems. It draws on the analysis of LC Young and the geometric algebra of KT Chen. The concepts and the uniform estimates, have widespread application and have simplified proofs of basic questions from the large deviation theory and extended Ito's theory of SDEs; the recent applications contribute to (Graham) automated recognition of Chinese handwriting and (Hairer) formulation of appropriate SPDEs to model randomly evolving interfaces. At the heart of the mathematics is the challenge of describing a smooth but potentially highly oscillatory and vector valued path $x_{t}$ parsimoniously so as to effectively predict the response of a nonlinear system such as $dy_{t}=f(y_{t})dx_{t}$, $y_{0}=a$. The Signature is a homomorphism from the monoid of paths into the grouplike elements of a closed tensor algebra. It provides a graduated summary of the path $x$. Hambly and Lyons have shown that this non-commutative transform is faithful for paths of bounded variation up to appropriate null modifications. Among paths of bounded variation with given Signature there is always a unique shortest representative. These graduated summaries or features of a path are at the heart of the definition of a rough path; locally they remove the need to look at the fine structure of the path. Taylor's theorem explains how any smooth function can, locally, be expressed as a linear combination of certain special functions (monomials based at that point). Coordinate iterated integrals form a more subtle algebra of features that can describe a stream or path in an analogous way; they allow a definition of rough path and a natural linear "basis" for functions on streams that can be used for machine learning.

preprint2013arXiv

Integrability and tail estimates for Gaussian rough differential equations

We derive explicit tail-estimates for the Jacobian of the solution flow for stochastic differential equations driven by Gaussian rough paths. In particular, we deduce that the Jacobian has finite moments of all order for a wide class of Gaussian process including fractional Brownian motion with Hurst parameter H>1/4. We remark on the relevance of such estimates to a number of significant open problems.

preprint2013arXiv

Kusuoka-Stroock gradient bounds for the solution of the filtering equation

We obtain sharp gradient bounds for perturbed diffusion semigroups. In contrast with existing results, the perturbation is here random and the bounds obtained are pathwise. Our approach builds on the classical work of Kusuoka and Stroock [7],[9],[10],[11], and extends their program developed for the heat semi-group to solutions of stochastic partial differential equations. The work is motivated by and applied to nonlinear filtering. The analysis allows us to derive pathwise gradient bounds for the un-normalised conditional distribution of a partially observed signal. It uses a pathwise representation of the perturbed semigroup in the spirit of classical work by Ocone [14]. The estimates we derive have sharp small time asymptotics.

preprint2013arXiv

Physical Brownian motion in magnetic field as rough path

The indefinite integral of the homogenized Ornstein-Uhlenbeck process is a well-known model for physical Brownian motion, modelling the behaviour of an object subject to random impulses [L. S. Ornstein, G. E. Uhlenbeck: On the theory of Brownian Motion. In: Physical Review. 36, 1930, 823-841]. One can scale these models by changing the mass of the particle and in the small mass limit one has almost sure uniform convergence in distribution to the standard idealized model of mathematical Brownian motion. This provides one well known way of realising the Wiener process. However, this result is less robust than it would appear and important generic functionals of the trajectories of the physical Brownian motion do not necessarily converge to the same functionals of Brownian motion when one takes the small mass limit. In presence of a magnetic field the area process associated to the physical process converges - but not to Lévy's stochastic area. As this area is felt generically in settings where the particle interacts through force fields in a nonlinear way, the remark is physically significant and indicates that classical Brownian motion, with its usual stochastic calculus, is not an appropriate model for the limiting behaviour. We compute explicitly the area correction term and establish convergence, in the small mass limit, of the physical Brownian motion in the rough path sense. The small mass limit for the motion of a charged particle in the presence of a magnetic field is, in distribution, an easily calculable, but "non-canonical" rough path lift of Brownian motion. Viewing the trajectory of a charged Brownian particle with small mass as a rough path is informative and allows one to retain information that would be lost if one only considered it as a classical trajectory. We comment on the importance of this point of view.

preprint2013arXiv

The adaptive patched particle filter and its implementation

There are numerous contexts where one wishes to describe the state of a randomly evolving system. Effective solutions combine models that quantify the underlying uncertainty with available observational data to form relatively optimal estimates for the uncertainty in the system state. Stochastic differential equations are often used to mathematically model the underlying system. The Kusuoka-Lyons-Victoir (KLV) approach is a higher order particle method for approximating the weak solution of a stochastic differential equation that uses a weighted set of scenarios to approximate the evolving probability distribution to a high order of accuracy. The algorithm can be performed by integrating along a number of carefully selected bounded variation paths and the iterated application of the KLV method has a tendency for the number of particles to increase. Together with local dynamic recombination that simplifies the support of discrete measure without harming the accuracy of the approximation, the KLV method becomes eligible to solve the filtering problem for which one has to maintain an accurate description of the ever-evolving conditioned measure. Besides the alternate application of the KLV method and recombination for the entire family of particles, we make use of the smooth nature of likelihood to lead some of the particles immediately to the next observation time and to build an algorithm that is a form of automatic high order adaptive importance sampling.

preprint2012arXiv

Backward stochastic dynamics on a filtered probability space

We demonstrate that backward stochastic differential equations (BSDE) may be reformulated as ordinary functional differential equations on certain path spaces. In this framework, neither Itô's integrals nor martingale representation formulate are needed. This approach provides new tools for the study of BSDE, and is particularly useful for the study of BSDE with partial information. The approach allows us to study the following type of backward stochastic differential equations: \[dY_t^j=-f_0^j(t,Y_t,L(M)_t) dt-\sum_{i=1}^df_i^j(t,Y_t), dB_t^i+dM_t^j\] with $Y_T=ξ$, on a general filtered probability space $(Ω,\mathcal{F},\mathcal{F}_t,P)$, where $B$ is a $d$-dimensional Brownian motion, $L$ is a prescribed (nonlinear) mapping which sends a square-integrable $M$ to an adapted process $L(M)$ and $M$, a correction term, is a square-integrable martingale to be determined. Under certain technical conditions, we prove that the system admits a unique solution $(Y,M)$. In general, the associated partial differential equations are not only nonlinear, but also may be nonlocal and involve integral operators.

preprint2011arXiv

Inversion of signature for paths of bounded variation

We develop two methods to reconstruct a path of bounded variation from its signature. The first method gives a simple and explicit expression of any axis path in terms of its signature, but it does not apply directlty to more general ones. The second method, based on an approximation scheme, recovers any tree-reduced path from its signature as the limit of a uniformly convergent sequence of lattice paths.

preprint2006arXiv

Uniqueness for the signature of a path of bounded variation and the reduced path group

We introduce the notions of tree-like path and tree-like equivalence between paths and prove that the latter is an equivalence relation for paths of finite length. We show that the equivalence classes form a group with some similarity to a free group, and that in each class there is one special tree reduced path. The set of these paths is the Reduced Path Group. It is a continuous analogue to the group of reduced words. The signature of the path is a power series whose coefficients are definite iterated integrals of the path. We identify the paths with trivial signature as the tree-like paths, and prove that two paths are in tree-like equivalence if and only if they have the same signature. In this way, we extend Chen's theorems on the uniqueness of the sequence of iterated integrals associated with a piecewise regular path to finite length paths and identify the appropriate extended meaning for reparameterisation in the general setting. It is suggestive to think of this result as a non-commutative analogue of the result that integrable functions on the circle are determined, up to Lebesgue null sets, by their Fourier coefficients. As a second theme we give quantitative versions of Chen's theorem in the case of lattice paths and paths with continuous derivative, and as a corollary derive results on the triviality of exponential products in the tensor algebra.