Source author record

Mathieu Rosenbaum

Mathieu Rosenbaum appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

33works
15topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

33 published item(s)

preprint2022arXiv

Deep calibration of the quadratic rough Heston model

The quadratic rough Heston model provides a natural way to encode Zumbach effect in the rough volatility paradigm. We apply multi-factor approximation and use deep learning methods to build an efficient calibration procedure for this model. We show that the model is able to reproduce very well both SPX and VIX implied volatilities. We typically obtain VIX option prices within the bid-ask spread and an excellent fit of the SPX at-the-money skew. Moreover, we also explain how to use the trained neural networks for hedging with instantaneous computation of hedging quantities.

preprint2022arXiv

Docent: A content-based recommendation system to discover contemporary art

Recommendation systems have been widely used in various domains such as music, films, e-shopping etc. After mostly avoiding digitization, the art world has recently reached a technological turning point due to the pandemic, making online sales grow significantly as well as providing quantitative online data about artists and artworks. In this work, we present a content-based recommendation system on contemporary art relying on images of artworks and contextual metadata of artists. We gathered and annotated artworks with advanced and art-specific information to create a completely unique database that was used to train our models. With this information, we built a proximity graph between artworks. Similarly, we used NLP techniques to characterize the practices of the artists and we extracted information from exhibitions and other event history to create a proximity graph between artists. The power of graph analysis enables us to provide an artwork recommendation system based on a combination of visual and contextual information from artworks and artists. After an assessment by a team of art specialists, we get an average final rating of 75% of meaningful artworks when compared to their professional evaluations.

preprint2022arXiv

On the universality of the volatility formation process: when machine learning and rough volatility agree

We train an LSTM network based on a pooled dataset made of hundreds of liquid stocks aiming to forecast the next daily realized volatility for all stocks. Showing the consistent outperformance of this universal LSTM relative to other asset-specific parametric models, we uncover nonparametric evidences of a universal volatility formation mechanism across assets relating past market realizations, including daily returns and volatilities, to current volatilities. A parsimonious parametric forecasting device combining the rough fractional stochastic volatility and quadratic rough Heston models with fixed parameters results in the same level of performance as the universal LSTM, which confirms the universality of the volatility formation process from a parametric perspective.

preprint2021arXiv

From quadratic Hawkes processes to super-Heston rough volatility models with Zumbach effect

Using microscopic price models based on Hawkes processes, it has been shown that under some no-arbitrage condition, the high degree of endogeneity of markets together with the phenomenon of metaorders splitting generate rough Heston-type volatility at the macroscopic scale. One additional important feature of financial dynamics, at the heart of several influential works in econophysics, is the so-called feedback or Zumbach effect. This essentially means that past trends in returns convey significant information on future volatility. A natural way to reproduce this property in microstructure modeling is to use quadratic versions of Hawkes processes. We show that after suitable rescaling, the long term limits of these processes are refined versions of rough Heston models where the volatility coefficient is enhanced compared to the square root characterizing Heston-type dynamics. Furthermore the Zumbach effect remains explicit in these limiting rough volatility models.

preprint2021arXiv

Optimal make take fees in a multi market maker environment

Following the recent literature on make take fees policies, we consider an exchange wishing to set a suitable contract with several market makers in order to improve trading quality on its platform. To do so, we use a principal-agent approach, where the agents (the market makers) optimise their quotes in a Nash equilibrium fashion, providing best response to the contract proposed by the principal (the exchange). This contract aims at attracting liquidity on the platform. This is because the wealth of the exchange depends on the arrival of market orders, which is driven by the spread of market makers. We compute the optimal contract in quasi explicit form and also derive the optimal spread policies for the market makers. Several new phenomena appears in this multi market maker setting. In particular we show that it is not necessarily optimal to have a large number of market makers in the presence of a contracting scheme.

preprint2020arXiv

On bid and ask side-specific tick sizes

The tick size, which is the smallest increment between two consecutive prices for a given asset, is a key parameter of market microstructure. In particular, the behavior of high frequency market makers is highly related to its value. We take the point of view of an exchange and investigate the relevance of having different tick sizes on the bid and ask sides of the order book. Using an approach based on the model with uncertainty zones, we show that when side-specific tick sizes are suitably chosen, it enables the exchange to improve the quality of liquidity provision.

preprint2020arXiv

Optimal auction duration: A price formation viewpoint

We consider an auction market in which market makers fill the order book during a given time period while some other investors send market orders. We define the clearing price of the auction as the price maximizing the exchanged volume at the clearing time according to the supply and demand of each market participants. Then we derive in a semi-explicit form the error made between this clearing price and the efficient price as a function of the auction duration. We study the impact of the behavior of market takers on this error. To do so we consider the case of naive market takers and that of rational market takers playing a Nash equilibrium to minimize their transaction costs. We compute the optimal duration of the auctions for 77 stocks traded on Euronext and compare the quality of price formation process under this optimal value to the case of a continuous limit order book. Continuous limit order books are found to be usually sub-optimal. However, in term of our metric, they only moderately impair the quality of price formation process. Order of magnitude of optimal auction durations is from 2 to 10 minutes.

preprint2020arXiv

The quadratic rough Heston model and the joint S&P 500/VIX smile calibration problem

Fitting simultaneously SPX and VIX smiles is known to be one of the most challenging problems in volatility modeling. A long-standing conjecture due to Julien Guyon is that it may not be possible to calibrate jointly these two quantities with a model with continuous sample-paths. We present the quadratic rough Heston model as a counterexample to this conjecture. The key idea is the combination of rough volatility together with a price-feedback (Zumbach) effect.

preprint2016arXiv

Asymptotic Optimal Tracking: Feedback Strategies

This is a companion paper to (Cai, Rosenbaum and Tankov, Asymptotic lower bounds for optimal tracking: a linear programming approach, arXiv:1510.04295). We consider a class of strategies of feedback form for the problem of tracking and study their performance under the asymptotic framework of the above reference. The strategies depend only on the current state of the system and keep the deviation from the target inside a time-varying domain. Although the dynamics of the target is non-Markovian, it turns out that such strategies are asympototically optimal for a large list of examples.

preprint2016arXiv

Linear and Conic Programming Estimators in High-Dimensional Errors-in-variables Models

We consider the linear regression model with observation error in the design. In this setting, we allow the number of covariates to be much larger than the sample size. Several new estimation methods have been recently introduced for this model. Indeed, the standard Lasso estimator or Dantzig selector turn out to become unreliable when only noisy regressors are available, which is quite common in practice. We show in this work that under suitable sparsity assumptions, the procedure introduced in Rosenbaum and Tsybakov (2013) is almost optimal in a minimax sense and, despite non-convexities, can be efficiently computed by a single linear programming problem. Furthermore, we provide an estimator attaining the minimax efficiency bound. This estimator is written as a second order cone programming minimisation problem which can be solved numerically in polynomial time.

preprint2016arXiv

The characteristic function of rough Heston models

It has been recently shown that rough volatility models, where the volatility is driven by a fractional Brownian motion with small Hurst parameter, provide very relevant dynamics in order to reproduce the behavior of both historical and implied volatilities. However, due to the non-Markovian nature of the fractional Brownian motion, they raise new issues when it comes to derivatives pricing. Using an original link between nearly unstable Hawkes processes and fractional volatility models, we compute the characteristic function of the log-price in rough Heston models. In the classical Heston model, the characteristic function is expressed in terms of the solution of a Riccati equation. Here we show that rough Heston models exhibit quite a similar structure, the Riccati equation being replaced by a fractional Riccati equation.

preprint2015arXiv

Asymptotic Lower Bounds for Optimal Tracking: a Linear Programming Approach

We consider the problem of tracking a target whose dynamics is modeled by a continuous Itō semi-martingale. The aim is to minimize both deviation from the target and tracking efforts. We establish the existence of asymptotic lower bounds for this problem, depending on the cost structure. These lower bounds can be related to the time-average control of Brownian motion, which is characterized as a deterministic linear programming problem. A comprehensive list of examples with explicit expressions for the lower bounds is provided.

preprint2015arXiv

Ergodicity and diffusivity of Markovian order book models: a general framework

We present a general Markovian framework for order book modeling. Through our approach, we aim at providing a tool enabling to get a better understanding of the price formation process and of the link between microscopic and macroscopic features of financial assets. To do so, we propose a new method of order book representation, and decompose the problem of order book modeling into two sub-problems: dynamics of a continuous-time double auction system with a fixed reference price; interactions between the double auction system and the reference price movements. State dependency is included in our framework by allowing the order flow intensities to depend on the order book state. Furthermore, contrary to most existing models, the impact of the order book updates on the reference price dynamics is not assumed to be instantaneous. We first prove that under general assumptions, our system is ergodic. Then we deduce the convergence towards a Brownian motion of the rescaled price process.

preprint2015arXiv

How to predict the consequences of a tick value change? Evidence from the Tokyo Stock Exchange pilot program

The tick value is a crucial component of market design and is often considered the most suitable tool to mitigate the effects of high frequency trading. The goal of this paper is to demonstrate that the approach introduced in Dayri and Rosenbaum (2015) allows for an ex ante assessment of the consequences of a tick value change on the microstructure of an asset. To that purpose, we analyze the pilot program on tick value modifications started in 2014 by the Tokyo Stock Exchange in light of this methodology. We focus on forecasting the future cost of market and limit orders after a tick value change and show that our predictions are very accurate. Furthermore, for each asset involved in the pilot program, we are able to define (ex ante) an optimal tick value. This enables us to classify the stocks according to the relevance of their tick value, before and after its modification.

preprint2015arXiv

Limit theorems for nearly unstable Hawkes processes

Because of their tractability and their natural interpretations in term of market quantities, Hawkes processes are nowadays widely used in high-frequency finance. However, in practice, the statistical estimation results seem to show that very often, only nearly unstable Hawkes processes are able to fit the data properly. By nearly unstable, we mean that the $L^1$ norm of their kernel is close to unity. We study in this work such processes for which the stability condition is almost violated. Our main result states that after suitable rescaling, they asymptotically behave like integrated Cox-Ingersoll-Ross models. Thus, modeling financial order flows as nearly unstable Hawkes processes may be a good way to reproduce both their high and low frequency stylized facts. We then extend this result to the Hawkes-based price model introduced by Bacry et al. [Quant. Finance 13 (2013) 65-77]. We show that under a similar criticality condition, this process converges to a Heston model. Again, we recover well-known stylized facts of prices, both at the microstructure level and at the macroscopic scale.

preprint2015arXiv

Rough fractional diffusions as scaling limits of nearly unstable heavy tailed Hawkes processes

We investigate the asymptotic behavior as time goes to infinity of Hawkes processes whose regression kernel has $L^1$ norm close to one and power law tail of the form $x^{-(1+α)}$, with $α\in(0,1)$. We in particular prove that when $α\in(1/2,1)$, after suitable rescaling, their law converges to that of a kind of integrated fractional Cox-Ingersoll-Ross process, with associated Hurst parameter $H=α-1/2$. This result is in contrast to the case of a regression kernel with light tail, where a classical Brownian CIR process is obtained at the limit. Interestingly, it shows that persistence properties in the point process can lead to an irregular behavior of the limiting process. This theoretical result enables us to give an agent-based foundation to some recent findings about the rough nature of volatility in financial markets.

preprint2015arXiv

The different asymptotic regimes of nearly unstable autoregressive processes

We extend classical results about the convergence of nearly unstable AR(p) processes to the infinite order case. To do so, we proceed as in recent works about Hawkes processes by using limit theorems for some well chosen geometric sums. We prove that when the coefficients sequence has a light tail, infinite order nearly unstable autoregressive processes behave as Ornstein-Uhlenbeck models. However, in the heavy tail case, we show that fractional diffusions arise as limiting laws for such processes.

preprint2014arXiv

An $\{l_1,l_2,l_{\infty}\}$-Regularization Approach to High-Dimensional Errors-in-variables Models

Several new estimation methods have been recently proposed for the linear regression model with observation error in the design. Different assumptions on the data generating process have motivated different estimators and analysis. In particular, the literature considered (1) observation errors in the design uniformly bounded by some $\bar δ$, and (2) zero mean independent observation errors. Under the first assumption, the rates of convergence of the proposed estimators depend explicitly on $\bar δ$, while the second assumption has been applied when an estimator for the second moment of the observational error is available. This work proposes and studies two new estimators which, compared to other procedures for regression models with errors in the design, exploit an additional $l_{\infty}$-norm regularization. The first estimator is applicable when both (1) and (2) hold but does not require an estimator for the second moment of the observational error. The second estimator is applicable under (2) and requires an estimator for the second moment of the observation error. Importantly, we impose no assumption on the accuracy of this pilot estimator, in contrast to the previously known procedures. As the recent proposals, we allow the number of covariates to be much larger than the sample size. We establish the rates of convergence of the estimators and compare them with the bounds obtained for related estimators in the literature. These comparisons show interesting insights on the interplay of the assumptions and the achievable rates of convergence.

preprint2014arXiv

Asymptotically optimal discretization of hedging strategies with jumps

In this work, we consider the hedging error due to discrete trading in models with jumps. Extending an approach developed by Fukasawa [In Stochastic Analysis with Financial Applications (2011) 331-346 Birkhäuser/Springer Basel AG] for continuous processes, we propose a framework enabling us to (asymptotically) optimize the discretization times. More precisely, a discretization rule is said to be optimal if for a given cost function, no strategy has (asymptotically, for large cost) a lower mean square discretization error for a smaller cost. We focus on discretization rules based on hitting times and give explicit expressions for the optimal rules within this class.

preprint2014arXiv

Optimal discretization of hedging strategies with directional views

We consider the hedging error of a derivative due to discrete trading in the presence of a drift in the dynamics of the underlying asset. We suppose that the trader wishes to find rebalancing times for the hedging portfolio which enable him to keep the discretization error small while taking advantage of market trends. Assuming that the portfolio is readjusted at high frequency, we introduce an asymptotic framework in order to derive optimal discretization strategies. More precisely, we formulate the optimization problem in terms of an asymptotic expectation-error criterion. In this setting, the optimal rebalancing times are given by the hitting times of two barriers whose values can be obtained by solving a linear-quadratic optimal control problem. In specific contexts such as in the Black-Scholes model, explicit expressions for the optimal rebalancing times can be derived.

preprint2014arXiv

Simulating and analyzing order book data: The queue-reactive model

Through the analysis of a dataset of ultra high frequency order book updates, we introduce a model which accommodates the empirical properties of the full order book together with the stylized facts of lower frequency financial data. To do so, we split the time interval of interest into periods in which a well chosen reference price, typically the mid price, remains constant. Within these periods, we view the limit order book as a Markov queuing system. Indeed, we assume that the intensities of the order flows only depend on the current state of the order book. We establish the limiting behavior of this model and estimate its parameters from market data. Then, in order to design a relevant model for the whole period of interest, we use a stochastic mechanism that allows for switches from one period of constant reference price to another. Beyond enabling to reproduce accurately the behavior of market data, we show that our framework can be very useful for practitioners, notably as a market simulator or as a tool for the transaction cost analysis of complex trading algorithms.

preprint2014arXiv

Volatility is rough

Estimating volatility from recent high frequency data, we revisit the question of the smoothness of the volatility process. Our main result is that log-volatility behaves essentially as a fractional Brownian motion with Hurst exponent H of order 0.1, at any reasonable time scale. This leads us to adopt the fractional stochastic volatility (FSV) model of Comte and Renault. We call our model Rough FSV (RFSV) to underline that, in contrast to FSV, H<1/2. We demonstrate that our RFSV model is remarkably consistent with financial time series data; one application is that it enables us to obtain improved forecasts of realized volatility. Furthermore, we find that although volatility is not long memory in the RFSV model, classical statistical procedures aiming at detecting volatility persistence tend to conclude the presence of long memory in data generated from it. This sheds light on why long memory of volatility has been widely accepted as a stylized fact. Finally, we provide a quantitative market microstructure-based foundation for our findings, relating the roughness of volatility to high frequency trading and order splitting.

preprint2013arXiv

Estimating the efficient price from the order flow: a Brownian Cox process approach

At the ultra high frequency level, the notion of price of an asset is very ambiguous. Indeed, many different prices can be defined (last traded price, best bid price, mid price,...). Thus, in practice, market participants face the problem of choosing a price when implementing their strategies. In this work, we propose a notion of efficient price which seems relevant in practice. Furthermore, we provide a statistical methodology enabling to estimate this price form the order flow.

preprint2013arXiv

Large tick assets: implicit spread and optimal tick size

In this work, we provide a framework linking microstructural properties of an asset to the tick value of the exchange. In particular, we bring to light a quantity, referred to as implicit spread, playing the role of spread for large tick assets, for which the effective spread is almost always equal to one tick. The relevance of this new parameter is shown both empirically and theoretically. This implicit spread allows us to quantify the tick sizes of large tick assets and to define a notion of optimal tick size. Moreover, our results open the possibility of forecasting the behavior of relevant market quantities after a change in the tick value and to give a way to modify it in order to reach an optimal tick size.

preprint2013arXiv

Quarticity and other functionals of volatility: Efficient estimation

We consider a multidimensional Ito semimartingale regularly sampled on [0,t] at high frequency $1/Δ_n$, with $Δ_n$ going to zero. The goal of this paper is to provide an estimator for the integral over [0,t] of a given function of the volatility matrix. To approximate the integral, we simply use a Riemann sum based on local estimators of the pointwise volatility. We show that although the accuracy of the pointwise estimation is at most $Δ_n^{1/4}$, this procedure reaches the parametric rate $Δ_n^{1/2}$, as it is usually the case in integrated functionals estimation. After a suitable bias correction, we obtain an unbiased central limit theorem for our estimator and show that it is asymptotically efficient within some classes of sub models.

preprint2013arXiv

Some explicit formulas for the Brownian bridge, Brownian meander and Bessel process under uniform sampling

We show that simple explicit formulas can be obtained for several relevant quantities related to the laws of the uniformly sampled Brownian bridge, Brownian meander and three dimensional Bessel process. To prove such results, we use the distribution of a triplet of random variables associated to the pseudo-Brownian bridge together with various relationships between the laws of these four processes.

preprint2012arXiv

Estimation of volatility functionals: the case of a square root n window

We consider a multidimensional Ito semimartingale regularly sampled on [0,t] at high frequency 1/Δ_n, with Δ_n going to zero. The goal of this paper is to provide an estimator for the integral over [0,t] of a given function of the volatility matrix, with the optimal rate 1/\sqrt{Δ_n} and minimal asymptotic variance. To achieve this we use spot volatility estimators based on observations within time intervals of length k_nΔ_n. In [5] this was done with k_n tending to infinity and k_n\sqrt{Δ_n} tending to 0, and a central limit theorem was given after suitable de-biasing. Here we do the same with k_n of order 1/\sqrt{Δ_n}. This results in a smaller bias, although more difficult to eliminate.

preprint2012arXiv

Testing the finiteness of the support of a distribution: a statistical look at Tsirelson's equation

We consider the following statistical problem: based on an i.i.d.sample of size n of integer valued random variables with common law m, is it possible to test whether or not the support of m is finite as n goes to infinity? This question is in particular connected to a simple case of Tsirelson's equation, for which it is natural to distinguish between two main configurations, the first one leading only to laws with finite support, and the second one including laws with infinite support. We show that it is in fact not possible to discriminate between the two situations, even using a very weak notion of statistical test.

preprint2011arXiv

Improved Matrix Uncertainty Selector

We consider the regression model with observation error in the design: y=Xθ* + e, Z=X+N. Here the random vector y in R^n and the random n*p matrix Z are observed, the n*p matrix X is unknown, N is an n*p random noise matrix, e in R^n is a random noise vector, and θ* is a vector of unknown parameters to be estimated. We consider the setting where the dimension p can be much larger than the sample size n and θ* is sparse. Because of the presence of the noise matrix N, the commonly used Lasso and Dantzig selector are unstable. An alternative procedure called the Matrix Uncertainty (MU) selector has been proposed in Rosenbaum and Tsybakov (2010) in order to account for the noise. The properties of the MU selector have been studied in Rosenbaum and Tsybakov (2010) for sparse θ* under the assumption that the noise matrix N is deterministic and its values are small. In this paper, we propose a modification of the MU selector when N is a random matrix with zero-mean entries having the variances that can be estimated. This is, for example, the case in the model where the entries of X are missing at random. We show both theoretically and numerically that, under these conditions, the new estimator called the Compensated MU selector achieves better accuracy of estimation than the original MU selector.

preprint2010arXiv

Asymptotic results and statistical procedures for time-changed Lévy processes sampled at hitting times

We provide asymptotic results and develop high frequency statistical procedures for time-changed Lévy processes sampled at random instants. The sampling times are given by first hitting times of symmetric barriers whose distance with respect to the starting point is equal to $\varepsilon$. This setting can be seen as a first step towards a model for tick-by-tick financial data allowing for large jumps. For a wide class of Lévy processes, we introduce a renormalization depending on $\varepsilon$, under which the Lévy process converges in law to an $α$-stable process as $\varepsilon$ goes to $0$. The convergence is extended to moments of hitting times and overshoots. In particular, these results allow us to construct consistent estimators of the time change and of the Blumenthal-Getoor index of the underlying Lévy process. Convergence rates and a central limit theorem are established under additional assumptions.

preprint2010arXiv

Sparse recovery under matrix uncertainty

We consider the model {eqnarray*}y=Xθ^*+ξ, Z=X+Ξ,{eqnarray*} where the random vector $y\in\mathbb{R}^n$ and the random $n\times p$ matrix $Z$ are observed, the $n\times p$ matrix $X$ is unknown, $Ξ$ is an $n\times p$ random noise matrix, $ξ\in\mathbb{R}^n$ is a noise independent of $Ξ$, and $θ^*$ is a vector of unknown parameters to be estimated. The matrix uncertainty is in the fact that $X$ is observed with additive error. For dimensions $p$ that can be much larger than the sample size $n$, we consider the estimation of sparse vectors $θ^*$. Under matrix uncertainty, the Lasso and Dantzig selector turn out to be extremely unstable in recovering the sparsity pattern (i.e., of the set of nonzero components of $θ^*$), even if the noise level is very small. We suggest new estimators called matrix uncertainty selectors (or, shortly, the MU-selectors) which are close to $θ^*$ in different norms and in the prediction risk if the restricted eigenvalue assumption on $X$ is satisfied. We also show that under somewhat stronger assumptions, these estimators recover correctly the sparsity pattern.