Source author record

Holger Drees

Holger Drees appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2020arXiv

Asymptotics for sliding blocks estimators of rare events

Drees and Rootzén (2010) have established limit theorems for a general class of empirical processes of statistics that are useful for the extreme value analysis of time series, but do not apply to statistics of sliding blocks, including so-called runs estimators. We generalize these results to empirical processes which cover both the class considered by Drees and Rootzén (2010) and processes of sliding blocks statistics. Using this approach, one can analyze different types of statistics in a unified framework. We show that statistics based on sliding blocks are asymptotically normal with an asymptotic variance which, under rather mild conditions, is smaller than or equal to the asymptotic variance of the corresponding estimator based on disjoint blocks. Finally, the general theory is applied to three well-known estimators of the extremal index. It turns out that they all have the same limit distribution, a fact which has so far been overlooked in the literature.

preprint2020arXiv

Limit theorems for empirical processes of cluster functionals

Let $(X_{n,i})_{1\le i\le n,n\in\mathbb{N}}$ be a triangular array of row-wise stationary $\mathbb{R}^d$-valued random variables. We use a "blocks method" to define clusters of extreme values: the rows of $(X_{n,i})$ are divided into $m_n$ blocks $(Y_{n,j})$, and if a block contains at least one extreme value, the block is considered to contain a cluster. The cluster starts at the first extreme value in the block and ends at the last one. The main results are uniform central limit theorems for empirical processes $Z_n(f):=\frac{1}{\sqrt {nv_n}}\sum_{j=1}^{m_n}(f(Y_{n,j})-Ef(Y_{n,j})),$ for $v_n=P\{X_{n,i}\neq0\}$ and $f$ belonging to classes of cluster functionals, that is, functions of the blocks $Y_{n,j}$ which only depend on the cluster values and which are equal to 0 if $Y_{n,j}$ does not contain a cluster. Conditions for finite-dimensional convergence include $β$-mixing, suitable Lindeberg conditions and convergence of covariances. To obtain full uniform convergence, we use either "bracketing entropy" or bounds on covering numbers with respect to a random semi-metric. The latter makes it possible to bring the powerful Vapnik--Červonenkis theory to bear. Applications include multivariate tail empirical processes and empirical processes of cluster values and of order statistics in clusters. Although our main field of applications is the analysis of extreme values, the theory can be applied more generally to rare events occurring, for example, in nonparametric curve estimation.

preprint2020arXiv

On a minimum distance procedure for threshold selection in tail analysis

Power-law distributions have been widely observed in different areas of scientific research. Practical estimation issues include how to select a threshold above which observations follow a power-law distribution and then how to estimate the power-law tail index. A minimum distance selection procedure (MDSP) is proposed in Clauset et al. (2009) and has been widely adopted in practice, especially in the analyses of social networks. However, theoretical justifications for this selection procedure remain scant. In this paper, we study the asymptotic behavior of the selected threshold and the corresponding power-law index given by the MDSP. We find that the MDSP tends to choose too high a threshold level and leads to Hill estimates with large variances and root mean squared errors for simulated data with Pareto-like tails.

preprint2016arXiv

A stochastic volatility model with flexible extremal dependence structure

Stochastic volatility processes with heavy-tailed innovations are a well-known model for financial time series. In these models, the extremes of the log returns are mainly driven by the extremes of the i.i.d. innovation sequence which leads to a very strong form of asymptotic independence, that is, the coefficient of tail dependence is equal to $1/2$ for all positive lags. We propose an alternative class of stochastic volatility models with heavy-tailed volatilities and examine their extreme value behavior. In particular, it is shown that, while lagged extreme observations are typically asymptotically independent, their coefficient of tail dependence can take on any value between $1/2$ (corresponding to exact independence) and 1 (related to asymptotic dependence). Hence, this class allows for a much more flexible extremal dependence between consecutive observations than classical SV models and can thus describe the observed clustering of financial returns more realistically. The extremal dependence structure of lagged observations is analyzed in the framework of regular variation on the cone $(0,\infty)^d$. As two auxiliary results which are of interest on their own we derive a new Breiman-type theorem about regular variation on $(0,\infty)^d$ for products of a random matrix and a regularly varying random vector and a statement about the joint extremal behavior of products of i.i.d. regularly varying random variables.

preprint2016arXiv

Hypotheses tests in boundary regression models

Consider a nonparametric regression model with one-sided errors and regression function in a general Hölder class. We estimate the regression function via minimization of the local integral of a polynomial approximation. We show uniform rates of convergence for the simple regression estimator as well as for a smooth version. These rates carry over to mean regression models with a symmetric and bounded error distribution. In such a setting, one obtains faster rates for irregular error distributions concentrating sufficient mass near the endpoints than for the usual regular distributions. The results are applied to prove asymptotic $\sqrt{n}$-equivalence of a residual-based (sequential) empirical distribution function to the (sequential) empirical distribution function of unobserved errors in the case of irregular error distributions. This result is remarkably different from corresponding results in mean regression with regular errors. It can readily be applied to develop goodness-of-fit tests for the error distribution. We present some examples and investigate the small sample performance in a simulation study. We further discuss asymptotically distribution-free hypotheses tests for independence of the error distribution from the points of measurement and for monotonicity of the boundary function as well.

preprint2016arXiv

Joint exceedances of random products

We analyze the joint extremal behavior of $n$ random products of the form $\prod_{j=1}^m X_j^{a_{ij}}, 1 \leq i \leq n,$ for non-negative, independent regularly varying random variables $X_1, \ldots, X_m$ and general coefficients $a_{ij} \in \mathbb{R}$. Products of this form appear for example if one observes a linear time series with gamma type innovations at $n$ points in time. We combine arguments of linear optimization and a generalized concept of regular variation on cones to show that the asymptotic behavior of joint exceedance probabilities of these products is determined by the solution of a linear program related to the matrix $\mathbf{A}=(a_{ij})$.

preprint2015arXiv

Bootstrapping Empirical Processes of Cluster Functionals with Application to Extremograms

In the extreme value analysis of time series, not only the tail behavior is of interest, but also the serial dependence plays a crucial role. Drees and Rootzén (2010) established limit theorems for a general class of empirical processes of so-called cluster functionals which can be used to analyse various aspects of the extreme value behavior of mixing time series. However, usually the limit distribution is too complex to enable a direct construction of confidence regions. Therefore, we suggest a multiplier block bootstrap analog to the empirical processes of cluster functionals. It is shown that under virtually the same conditions as used by Drees and Rootzén (2010), conditionally on the data, the bootstrap processes converge to the same limit distribution. These general results are applied to construct confidence regions for the empirical extremogram introduced by Davis and Mikosch (2009). In a simulation study, the confidence intervals constructed by our multiplier block bootstrap approach compare favorably to the stationary bootstrap proposed by Davis et al.\ (2012).

preprint2015arXiv

Estimating failure probabilities

In risk management, often the probability must be estimated that a random vector falls into an extreme failure set. In the framework of bivariate extreme value theory, we construct an estimator for such failure probabilities and analyze its asymptotic properties under natural conditions. It turns out that the estimation error is mainly determined by the accuracy of the statistical analysis of the marginal distributions if the extreme value approximation to the dependence structure is at least as accurate as the generalized Pareto approximation to the marginal distributions. Moreover, we establish confidence intervals and briefly discuss generalizations to higher dimensions and issues arising in practical applications as well.

preprint2014arXiv

Statistics for Tail Processes of Markov Chains

At high levels, the asymptotic distribution of a stationary, regularly varying Markov chain is conveniently given by its tail process. The latter takes the form of a geometric random walk, the increment distribution depending on the sign of the process at the current state and on the flow of time, either forward or backward. Estimation of the tail process provides a nonparametric approach to analyze extreme values. A duality between the distributions of the forward and backward increments provides additional information that can be exploited in the construction of more efficient estimators. The large-sample distribution of such estimators is derived via empirical process theory for cluster functionals. Their finite-sample performance is evaluated via Monte Carlo simulations involving copula-based Markov models and solutions to stochastic recurrence equations. The estimators are applied to stock price data to study the absence or presence of symmetries in the succession of large gains and losses.

preprint2011arXiv

Bias correction for estimators of the extremal index

We investigate the joint asymptotic behavior of so-called blocks estimator of the extremal index, that determines the mean length of clusters of extremes, based on the exceedances over different thresholds. Due to the large bias of these estimators, the resulting estimates are usually very sensitive to the choice of the threshold and thus difficult to interpret. We propose and examine a bias correction that asymptotically removes the leading bias term while the rate of convergence of the random error is preserved.

preprint2011arXiv

Extreme value analysis of actuarial risks: estimation and model validation

We give an overview of several aspects arising in the statistical analysis of extreme risks with actuarial applications in view. In particular it is demonstrated that empirical process theory is a very powerful tool, both for the asymptotic analysis of extreme value estimators and to devise tools for the validation of the underlying model assumptions. While the focus of the paper is on univariate tail risk analysis, the basic ideas of the analysis of the extremal dependence between different risks are also outlined. Here we emphasize some of the limitation of classical multivariate extreme value theory and sketch how a different model proposed by Ledford and Tawn can help to avoid pitfalls. Finally, these theoretical results are used to analyze a data set of large claim sizes from health insurance.