Source author record

Olivier Wintenberger

Olivier Wintenberger appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

28works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

28 published item(s)

preprint2025arXiv

Tail asymptotics and precise large deviations for some Poisson cluster processes

We study the tail asymptotics of two functionals (the maximum and the sum of the marks) of a generic cluster in two sub-models of the marked Poisson cluster process, namely the renewal Poisson cluster process and the Hawkes process. Under the hypothesis that the governing components of the processes are regularly varying, we extend results due to [18] and [5] notably, relying on Karamata's Tauberian Theorem to do so. We use these asymptotics to derive precise large deviation results in the fashion of [30] for the above-mentioned processes.

preprint2023arXiv

Multivariate sparse clustering for extremes

Identifying directions where extreme events occur is a major challenge in multivariate extreme value analysis. In this paper, we use the concept of sparse regular variation introduced by Meyer and Wintenberger (2021)} to infer the tail dependence of a random vector X. This approach relies on the Euclidean projection onto the simplex which better exhibits the sparsity structure of the tail of X than the standard methods. Our procedure based on a rigorous methodology aims at capturing clusters of extremal coordinates of X. It also includes the identification of the threshold above which the values taken by X are considered as extreme. We provide an efficient and scalable algorithm called MUSCLE and apply it on numerical examples to highlight the relevance of our findings. Finally we illustrate our approach with financial return data.

preprint2022arXiv

Large deviations of lp-blocks of regularly varying time series and applications to cluster inference

In the regularly varying time series setting, a cluster of exceedances is a short period for which the supremum norm exceeds a high threshold. We propose to study a generalization of this notion considering short periods, or blocks, with lp-norm above a high threshold. Our main result derives new large deviation principles of extremal lp-blocks, which guide us to define and characterize spectral cluster processes in lp. We then study cluster inference in lp to motivate our results. We design consistent disjoint block estimators to infer features of cluster processes. Our estimators promote the use of large empirical quantiles from the lp-norm of blocks as threshold levels which eases implementation and also facilitates comparison for different p>0. Our approach highlights the advantages of cluster inference based on extremal l$α$--blocks, where $α$>0 is the index of regular variation of the series. We focus on inferring important indices in extreme value theory, e.g., the extremal index.

preprint2021arXiv

AdaVol: An Adaptive Recursive Volatility Prediction Method

Quasi-Maximum Likelihood (QML) procedures are theoretically appealing and widely used for statistical inference. While there are extensive references on QML estimation in batch settings, it has attracted little attention in streaming settings until recently. An investigation of the convergence properties of the QML procedure in a general conditionally heteroscedastic time series model is conducted, and the classical batch optimization routines extended to the framework of streaming and large-scale problems. An adaptive recursive estimation routine for GARCH models named AdaVol is presented. The AdaVol procedure relies on stochastic approximations combined with the technique of Variance Targeting Estimation (VTE). This recursive method has computationally efficient properties, while VTE alleviates some convergence difficulties encountered by the usual QML estimation due to a lack of convexity. Empirical results demonstrate a favorable trade-off between AdaVol's stability and the ability to adapt to time-varying estimates for real-life data.

preprint2021arXiv

Consistent Regression using Data-Dependent Coverings

In this paper, we introduce a novel method to generate interpretable regression function estimators. The idea is based on called data-dependent coverings. The aim is to extract from the data a covering of the feature space instead of a partition. The estimator predicts the empirical conditional expectation over the cells of the partitions generated from the coverings. Thus, such estimator has the same form as those issued from data-dependent partitioning algorithms. We give sufficient conditions to ensure the consistency, avoiding the sufficient condition of shrinkage of the cells that appears in the former literature. Doing so, we reduce the number of covering elements. We show that such coverings are interpretable and each element of the covering is tagged as significant or insignificant. The proof of the consistency is based on a control of the error of the empirical estimation of conditional expectations which is interesting on its own.

preprint2021arXiv

Hidden regular variation for point processes and the single/multiple large point heuristic

We consider regular variation for marked point processes with independent heavy-tailed marks and prove a single large point heuristic: the limit measure is concentrated on the cone of point measures with one single point. We then investigate successive hidden regular variation removing the cone of point measures with at most $k$ points, $k\geq 1$, and prove a multiple large point phenomenon: the limit measure is concentrated on the cone of point measures with $k+1$ points. We show how these results imply hidden regular variation in Skorokhod space of the associated risk process. Finally, we provide an application to risk theory in a reinsurance model where the $k$ largest claims are covered and we study the asymptotic behavior of the residual risk.

preprint2021arXiv

Sparse regular variation

Regular variation provides a convenient theoretical framework to study large events. In the multivariate setting, the dependence structure of the positive extremes is characterized by a measure - the spectral measure - defined on the positive orthant of the unit sphere. This measure gathers information on the localization of extreme events and has often a sparse support since severe events do not simultaneously occur in all directions. However, it is defined through weak convergence which does not provide a natural way to capture this sparsity structure.In this paper, we introduce the notion of sparse regular variation which allows to better learn the dependence structure of extreme events. This concept is based on the Euclidean projection onto the simplex for which efficient algorithms are known. We prove that under mild assumptions sparse regular variation and regular variation are two equivalent notions and we establish several results for sparsely regularly varying random vectors. Finally, we illustrate on numerical examples how this new concept allows one to detect extremal directions.

preprint2020arXiv

Kalman Recursions Aggregated Online

In this article, we aim at improving the prediction of expert aggregation by using the underlying properties of the models that provide expert predictions. We restrict ourselves to the case where expert predictions come from Kalman recursions, fitting state-space models. By using exponential weights, we construct different algorithms of Kalman recursions Aggregated Online (KAO) that compete with the best expert or the best convex combination of experts in a more or less adaptive way. We improve the existing results on expert aggregation literature when the experts are Kalman recursions by taking advantage of the second-order properties of the Kalman recursions. We apply our approach to Kalman recursions and extend it to the general adversarial expert setting by state-space modeling the errors of the experts. We apply these new algorithms to a real dataset of electricity consumption and show how it can improve forecast performances comparing to other exponentially weighted average procedures.

preprint2020arXiv

Stochastic Online Optimization using Kalman Recursion

We study the Extended Kalman Filter in constant dynamics, offering a bayesian perspective of stochastic optimization. We obtain high probability bounds on the cumulative excess risk in an unconstrained setting. In order to avoid any projection step we propose a two-phase analysis. First, for linear and logistic regressions, we prove that the algorithm enters a local phase where the estimate stays in a small region around the optimum. We provide explicit bounds with high probability on this convergence time. Second, for generalized linear regressions, we provide a martingale analysis of the excess risk in the local phase, improving existing ones in bounded stochastic optimization. The EKF appears as a parameter-free online algorithm with O(d^2) cost per iteration that optimally solves some unconstrained optimization problems.

preprint2016arXiv

Exponential inequalities for unbounded functions of geometrically ergodic Markov chains. Applications to quantitative error bounds for regenerative Metropolis algorithms

The aim of this note is to investigate the concentration properties of unbounded functions of geometrically ergodic Markov chains. We derive concentration properties of centered functions with respect to the square of the Lyapunov's function in the drift condition satisfied by the Markov chain. We apply the new exponential inequalities to derive confidence intervals for MCMC algorithms. Quantitative error bounds are providing for the regenerative Metropolis algorithm of [5].

preprint2016arXiv

Goodness-of-fit tests for extended Log-GARCH models

This paper studies goodness of fit tests and specification tests for an extension of the log-GARCH model which is stable by scaling. A Lagrange-Multiplier test is derived for testing the null assumption of extended log-GARCH against more general formulations including the Exponential GARCH (EGARCH). The null assumption of an EGARCH is also tested. Portmanteau goodness-of-fit tests are developed for the extended log-GARCH. Simulations illustrating the theoretical results and an application to real financial data are proposed.

preprint2016arXiv

Optimal learning with Bernstein Online Aggregation

We introduce a new recursive aggregation procedure called Bernstein Online Aggregation (BOA). The exponential weights include an accuracy term and a second order term that is a proxy of the quadratic variation as in Hazan and Kale (2010). This second term stabilizes the procedure that is optimal in different senses. We first obtain optimal regret bounds in the deterministic context. Then, an adaptive version is the first exponential weights algorithm that exhibits a second order bound with excess losses that appears first in Gaillard et al. (2014). The second order bounds in the deterministic context are extended to a general stochastic context using the cumulative predictive risk. Such conversion provides the main result of the paper, an inequality of a novel type comparing the procedure with any deterministic aggregation procedure for an integrated criteria. Then we obtain an observable estimate of the excess of risk of the BOA procedure. To assert the optimality, we consider finally the iid case for strongly convex and Lipschitz continuous losses and we prove that the optimal rate of aggregation of Tsybakov (2003) is achieved. The batch version of the BOA procedure is then the first adaptive explicit algorithm that satisfies an optimal oracle inequality with high probability.

preprint2016arXiv

Regular variation of a random length sequence of random variables and application to risk assessment

When assessing risks on a finite-time horizon, the problem can often be reduced to the study of a random sequence $C(N)=(C_1,\ldots,C_N)$ of random length $N$, where $C(N)$ comes from the product of a matrix $A(N)$ of random size $N \times N$ and a random sequence $X(N)$ of random length $N$. Our aim is to build a regular variation framework for such random sequences of random length, to study their spectral properties and, subsequently, to develop risk measures. In several applications, many risk indicators can be expressed from the asymptotic behavior of $\vert \vert C(N)\vert\vert$, for some norm $\Vert \cdot \Vert$. We propose a generalization of Breiman Lemma that gives way to an asymptotic equivalent to $\Vert C(N) \Vert$ and provides risk indicators such as the ruin probability and the tail index for Shot Noise Processes on a finite-time horizon. Lastly, we apply our final result to a model used in dietary risk assessment and in non-life insurance mathematics to illustrate the applicability of our method.

preprint2016arXiv

Sparse Accelerated Exponential Weights

We consider the stochastic optimization problem where a convex function is minimized observing recursively the gradients. We introduce SAEW, a new procedure that accelerates exponential weights procedures with the slow rate $1/\sqrt{T}$ to procedures achieving the fast rate $1/T$. Under the strong convexity of the risk, we achieve the optimal rate of convergence for approximating sparse parameters in $\mathbb{R}^d$. The acceleration is achieved by using successive averaging steps in an online fashion. The procedure also produces sparse estimators thanks to additional hard threshold steps.

preprint2014arXiv

Weak transport inequalities and applications to exponential inequalities and oracle inequalities

We extend the dimension free Talagrand inequalities for convex distance \cite{talagrand:1995} using an extension of Marton's weak transport \cite{marton:1996a} to other metrics than the Hamming distance. We study the dual form of these weak transport inequalities for the euclidian norm and prove that it implies sub-gaussianity and convex Poincaré inequality \cite{bobkov:gotze:1999a}. We obtain new weak transport inequalities for non products measures extending the results of Samson in \cite{samson:2000}. Many examples are provided to show that the euclidian norm is an appropriate metric for classical time series. Our approach, based on trajectories coupling, is more efficient to obtain dimension free concentration than existing contractive assumptions \cite{djellout:guillin:wu:2004,marton:2004}. Expressing the concentration properties of the ordinary least square estimator as a conditional weak transport problem, we derive new oracle inequalities with fast rates of convergence in dependent settings.

preprint2013arXiv

Continuous invertibility and stable QML estimation of the EGARCH(1,1) model

We introduce the notion of continuous invertibility on a compact set for volatility models driven by a Stochastic Recurrence Equation (SRE). We prove the strong consistency of the Quasi Maximum Likelihood Estimator (QMLE) when the optimization procedure is done on a continuously invertible domain. This approach gives for the first time the strong consistency of the QMLE used by Nelson in \cite{nelson:1991} for the EGARCH(1,1) model under explicit but non observable conditions. In practice, we propose to stabilize the QMLE by constraining the optimization procedure to an empirical continuously invertible domain. The new method, called Stable QMLE (SQMLE), is strongly consistent when the observations follow an invertible EGARCH(1,1) model. We also give the asymptotic normality of the SQMLE under additional minimal assumptions.

preprint2013arXiv

GARCH models without positivity constraints: Exponential or Log GARCH?

This paper provides a probabilistic and statistical comparison of the log-GARCH and EGARCH models, which both rely on multiplicative volatility dynamics without positivity constraints. We compare the main probabilistic properties (strict stationarity, existence of moments, tails) of the EGARCH model, which are already known, with those of an asymmetric version of the log-GARCH. The quasi-maximum likelihood estimation of the log-GARCH parameters is shown to be strongly consistent and asymptotically normal. Similar estimation results are only available for particular EGARCH models, and under much stronger assumptions. The comparison is pursued via simulation experiments and estimation on real data.

preprint2013arXiv

The cluster index of regularly varying sequences with applications to limit theory for functions of multivariate Markov chains

We introduce the cluster index of a multivariate regularly varying stationary sequence and characterize the index in terms of the spectral tail process. This index plays a major role in limit theory for partial sums of regularly varying sequences. We illustrate the use of the cluster index by characterizing infinite variance stable limit distributions and precise large deviation results for sums of multivariate functions acting on a stationary Markov chain under a drift condition.

preprint2012arXiv

Fast rates in learning with dependent observations

In this paper we tackle the problem of fast rates in time series forecasting from a statistical learning perspective. In a serie of papers (e.g. Meir 2000, Modha and Masry 1998, Alquier and Wintenberger 2012) it is shown that the main tools used in learning theory with iid observations can be extended to the prediction of time series. The main message of these papers is that, given a family of predictors, we are able to build a new predictor that predicts the series as well as the best predictor in the family, up to a remainder of order $1/\sqrt{n}$. It is known that this rate cannot be improved in general. In this paper, we show that in the particular case of the least square loss, and under a strong assumption on the time series (phi-mixing) the remainder is actually of order $1/n$. Thus, the optimal rate for iid variables, see e.g. Tsybakov 2003, and individual sequences, see \cite{lugosi} is, for the first time, achieved for uniformly mixing processes. We also show that our method is optimal for aggregating sparse linear combinations of predictors.

preprint2012arXiv

Model selection for weakly dependent time series forecasting

Observing a stationary time series, we propose a two-step procedure for the prediction of the next value of the time series. The first step follows machine learning theory paradigm and consists in determining a set of possible predictors as randomized estimators in (possibly numerous) different predictive models. The second step follows the model selection paradigm and consists in choosing one predictor with good properties among all the predictors of the first steps. We study our procedure for two different types of bservations: causal Bernoulli shifts and bounded weakly dependent processes. In both cases, we give oracle inequalities: the risk of the chosen predictor is close to the best prediction risk in all predictive models that we consider. We apply our procedure for predictive models such as linear predictors, neural networks predictors and non-parametric autoregressive.

preprint2012arXiv

Precise large deviations for dependent regularly varying sequences

We study a precise large deviation principle for a stationary regularly varying sequence of random variables. This principle extends the classical results of A.V. Nagaev (1969) and S.V. Nagaev (1979) for iid regularly varying sequences. The proof uses an idea of Jakubowski (1993,1997) in the context of centra limit theorems with infinite variance stable limits. We illustrate the principle for \sv\ models, functions of a Markov chain satisfying a polynomial drift condition and solutions of linear and non-linear stochastic recurrence equations.

preprint2012arXiv

Prediction of time series by statistical learning: general losses and fast rates

We establish rates of convergences in time series forecasting using the statistical learning approach based on oracle inequalities. A series of papers extends the oracle inequalities obtained for iid observations to time series under weak dependence conditions. Given a family of predictors and $n$ observations, oracle inequalities state that a predictor forecasts the series as well as the best predictor in the family up to a remainder term $Δ_n$. Using the PAC-Bayesian approach, we establish under weak dependence conditions oracle inequalities with optimal rates of convergence. We extend previous results for the absolute loss function to any Lipschitz loss function with rates $Δ_n\sim\sqrt{c(Θ)/ n}$ where $c(Θ)$ measures the complexity of the model. We apply the method for quantile loss functions to forecast the french GDP. Under additional conditions on the loss functions (satisfied by the quadratic loss function) and on the time series, we refine the rates of convergence to $Δ_n \sim c(Θ)/n$. We achieve for the first time these fast rates for uniformly mixing processes. These rates are known to be optimal in the iid case and for individual sequences. In particular, we generalize the results of Dalalyan and Tsybakov on sparse regression estimation to the case of autoregression.

preprint2011arXiv

Parametric inference and forecasting in continuously invertible volatility models

We introduce the notion of continuously invertible volatility models that relies on some Lyapunov condition and some regularity condition. We show that it is almost equivalent to the ability of the volatilities forecasting using the parametric inference approach based on the SRE given in [16]. Under very weak assumptions, we prove the strong consistency and the asymptotic normality of the parametric inference. Based on this parametric estimation, a natural strongly consistent forecast of the volatility is given. We apply successfully this approach to recover known results on univariate and multivariate GARCH type models and to the EGARCH(1,1) model. We prove the strong consistency of the forecasting as soon as the model is invertible and the asymptotic normality of the parametric inference as soon as the limiting variance exists. Finally, we give some encouraging empirical results of our approach on simulations and real data.

preprint2010arXiv

Detecting multiple change-points in general causal time series using penalized quasi-likelihood

This paper is devoted to the off-line multiple change-point detection in a semiparametric framework. The time series is supposed to belong to a large class of models including AR($\infty$), ARCH($\infty$), TARCH($\infty$),... models where the coefficients change at each instant of breaks. The different unknown parameters (number of changes, change dates and parameters of successive models) are estimated using a penalized contrast built on conditional quasi-likelihood. Under Lipshitzian conditions on the model, the consistency of the estimator is proved when the moment order $r$ of the process satisfies $r\geq 2$. If $r\geq 4$, the same convergence rates for the estimators than in the case of independent random variables are obtained. The particular cases of AR($\infty$), ARCH($\infty$) and TARCH($\infty$) show that our method notably improves the existing results.

preprint2010arXiv

Stable limits for sums of dependent infinite variance random variables

The aim of this paper is to provide conditions which ensure that the affinely transformed partial sums of a strictly stationary process converge in distribution to an infinite variance stable distribution. Conditions for this convergence to hold are known in the literature. However, most of these results are qualitative in the sense that the parameters of the limit distribution are expressed in terms of some limiting point process. In this paper we will be able to determine the parameters of the limiting stable distribution in terms of some tail characteristics of the underlying stationary sequence. We will apply our results to some standard time series models, including the GARCH(1, 1) process and its squares, the stochastic volatility models and solutions to stochastic recurrence equations.

preprint2009arXiv

Deviation inequalities for sums of weakly dependent time series

In this paper we give new deviation inequalities of Bernstein's type for the partial sums of weakly dependent time series. The loss from the independent case is studied carefully. We give non mixing examples such that dynamical systems and Bernoulli shifts for whom our deviation inequalities hold. The proofs are based on the blocks technique and different coupling arguments.

preprint2008arXiv

Adaptive density estimation under dependence

Assume that $(X_t)_{t\in\Z}$ is a real valued time series admitting a common marginal density $f$ with respect to Lebesgue's measure. Donoho {\it et al.} (1996) propose a near-minimax method based on thresholding wavelets to estimate $f$ on a compact set in an independent and identically distributed setting. The aim of the present work is to extend these results to general weak dependent contexts. Weak dependence assumptions are expressed as decreasing bounds of covariance terms and are detailed for different examples. The threshold levels in estimators $\widehat f_n$ depend on weak dependence properties of the sequence $(X_t)_{t\in\Z}$ through the constant. If these properties are unknown, we propose cross-validation procedures to get new estimators. These procedures are illustrated via simulations of dynamical systems and non causal infinite moving averages. We also discuss the efficiency of our estimators with respect to the decrease of covariances bounds.