Source author record

Johan Segers

Johan Segers appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

38works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

38 published item(s)

preprint2026arXiv

Simulation of Multivariate Extremes: a Wasserstein-Aitchison GAN approach

Economically responsible mitigation of multivariate extreme risks-such as extreme rainfall over large areas, large simultaneous variations in many stock prices, or widespread breakdowns in transportation systems-requires assessing the resilience of the systems under plausible stress scenarios. This paper uses Extreme Value Theory (EVT) to develop a new approach to simulating such multivariate extreme events. Specifically, we assume that after transformation to a standard scale the distribution of the random phenomenon of interest is multivariate regular varying and use this to provide a sampling procedure for extremes on the original scale. Our procedure combines a Wasserstein-Aitchison Generative Adversarial Network (WA-GAN) to simulate the tail dependence structure on the standard scale with joint modeling of the univariate marginal tails on the original scale. The WA-GAN procedure relies on the angular measure-encoding the distribution on the unit simplex of the angles of extreme observations-after transformation to Aitchison coordinates, which allows the Wasserstein-GAN algorithm to be run in a linear space. Our method is applied both to simulated data under various tail dependence scenarios and to a financial data set from the Kenneth French Data Library. The proposed algorithm demonstrates strong performance compared to existing alternatives in the literature, both in capturing tail dependence structures and in generating accurate new extreme observations.

preprint2023arXiv

Max-linear graphical models with heavy-tailed factors on trees of transitive tournaments

Graphical models with heavy-tailed factors can be used to model extremal dependence or causality between extreme events. In a Bayesian network, variables are recursively defined in terms of their parents according to a directed acyclic graph (DAG). We focus on max-linear graphical models with respect to a special type of graphs, which we call a tree of transitive tournaments. The latter are block graphs combining in a tree-like structure a finite number of transitive tournaments, each of which is a DAG in which every two nodes are connected. We study the limit of the joint tails of the max-linear model conditionally on the event that a given variable exceeds a high threshold. Under a suitable condition, the limiting distribution involves the factorization into independent increments along the shortest trail between two variables, thereby imitating the behavior of a Markov random field. We are also interested in the identifiability of the model parameters in case some variables are latent and only a subvector is observed. It turns out that the parameters are identifiable under a criterion on the nodes carrying the latent variables which is easy and quick to check.

preprint2023arXiv

Tail inference using extreme U-statistics

Extreme U-statistics arise when the kernel of a U-statistic has a high degree but depends only on its arguments through a small number of top order statistics. As the kernel degree of the U-statistic grows to infinity with the sample size, estimators built out of such statistics form an intermediate family in between those constructed in the block maxima and peaks-over-threshold frameworks in extreme value analysis. The asymptotic normality of extreme U-statistics based on location-scale invariant kernels is established. Although the asymptotic variance coincides with the one of the Hájek projection, the proof goes beyond considering the first term in Hoeffding's variance decomposition. We propose a kernel depending on the three highest order statistics leading to a location-scale invariant estimator of the extreme value index resembling the Pickands estimator. This extreme Pickands U-estimator is asymptotically normal and its finite-sample performance is competitive with that of the pseudo-maximum likelihood estimator.

preprint2022arXiv

Graphical and uniform consistency of estimated optimal transport plans

A general theory is provided delivering convergence of maximal cyclically monotone mappings containing the supports of coupling measures of sequences of pairs of possibly random probability measures on Euclidean space. The theory is based on the identification of such a mapping with a closed subset of a Cartesian product of Euclidean spaces and leveraging tools from random set theory. Weak convergence in the appropriate Fell space together with the maximal cyclical monotonicity then automatically yields local uniform convergence of the associated mappings. Viewing such mappings as optimal transport plans between probability measures with respect to the squared Euclidean distance as cost function yields consistency results for notions of multivariate ranks and quantiles based on optimal transport, notably the empirical center-outward distribution and quantile functions.

preprint2021arXiv

Inference on extremal dependence in the domain of attraction of a structured Hüsler-Reiss distribution motivated by a Markov tree with latent variables

A Markov tree is a probabilistic graphical model for a random vector indexed by the nodes of an undirected tree encoding conditional independence relations between variables. One possible limit distribution of partial maxima of samples from such a Markov tree is a max-stable Hüsler-Reiss distribution whose parameter matrix inherits its structure from the tree, each edge contributing one free dependence parameter. Our central assumption is that, upon marginal standardization, the data-generating distribution is in the max-domain of attraction of the said Hüsler-Reiss distribution, an assumption much weaker than the one that data are generated according to a graphical model. Even if some of the variables are unobservable (latent), we show that the underlying model parameters are still identifiable if and only if every node corresponding to a latent variable has degree at least three. Three estimation procedures, based on the method of moments, maximum composite likelihood, and pairwise extremal coefficients, are proposed for usage on multivariate peaks over thresholds data when some variables are latent. A typical application is a river network in the form of a tree where, on some locations, no data are available. We illustrate the model and the identifiability criterion on a data set of high water levels on the Seine, France, with two latent variables. The structured Hüsler-Reiss distribution is found to fit the observed extremal dependence patterns well. The parameters being identifiable we are able to quantify tail dependence between locations for which there are no data.

preprint2021arXiv

Maxima and near-maxima of a Gaussian random assignment field

The assumption that the elements of the cost matrix in the classical assignment problem are drawn independently from a standard Gaussian distribution motivates the study of a particular Gaussian field indexed by the symmetric permutation group. The correlation structure of the field is determined by the Hamming distance between two permutations. The expectation of the maximum of the field is shown to go to infinity in the same way as if all variables of the field were independent. However, the variance of the maximum is shown to converge to zero at a rate which is slower than under independence, as the variance cannot be smaller than the one of the cost of the average assignment. Still, the convergence to zero of the variance means that the maximum possesses a property known as superconcentration. Finally, the dimension of the set of near-optimal assignments is shown to converge to zero.

preprint2021arXiv

Multivariate goodness-of-Fit tests based on Wasserstein distance

Goodness-of-fit tests based on the empirical Wasserstein distance are proposed for simple and composite null hypotheses involving general multivariate distributions. For group families, the procedure is to be implemented after preliminary reduction of the data via invariance.This property allows for calculation of exact critical values and p-values at finite sample sizes. Applications include testing for location--scale families and testing for families arising from affine transformations, such as elliptical distributions with given standard radial density and unspecified location vector and scatter matrix. A novel test for multivariate normality with unspecified mean vector and covariance matrix arises as a special case. For more general parametric families, we propose a parametric bootstrap procedure to calculate critical values. The lack of asymptotic distribution theory for the empirical Wasserstein distance means that the validity of the parametric bootstrap under the null hypothesis remains a conjecture. Nevertheless, we show that the test is consistent against fixed alternatives. To this end, we prove a uniform law of large numbers for the empirical distribution in Wasserstein distance, where the uniformity is over any class of underlying distributions satisfying a uniform integrability condition but no additional moment assumptions. The calculation of test statistics boils down to solving the well-studied semi-discrete optimal transport problem. Extensive numerical experiments demonstrate the practical feasibility and the excellent performance of the proposed tests for the Wasserstein distance of order p = 1 and p = 2 and for dimensions at least up to d = 5. The simulations also lend support to the conjecture of the asymptotic validity of the parametric bootstrap.

preprint2020arXiv

Resampling Procedures with Empirical Beta Copulas

The empirical beta copula is a simple but effective smoother of the empirical copula. Because it is a genuine copula, from which, moreover, it is particularly easy to sample, it is reasonable to expect that resampling procedures based on the empirical beta copula are expedient and accurate. In this paper, after reviewing the literature on some bootstrap approximations for the empirical copula process, we first show the asymptotic equivalence of several bootstrapped processes related to the empirical copula and empirical beta copula. Then we investigate the finite-sample properties of resampling schemes based on the empirical (beta) copula by Monte Carlo simulation. More specifically, we consider interval estimation for some functionals such as rank correlation coefficients and dependence parameters of several well-known families of copulas, constructing confidence intervals by several methods and comparing their accuracy and efficiency. We also compute the actual size and power of symmetry tests based on several resampling schemes for the empirical copula and empirical beta copula.

preprint2016arXiv

A continuous updating weighted least squares estimator of tail dependence in high dimensions

Likelihood-based procedures are a common way to estimate tail dependence parameters. They are not applicable, however, in non-differentiable models such as those arising from recent max-linear structural equation models. Moreover, they can be hard to compute in higher dimensions. An adaptive weighted least-squares procedure matching nonparametric estimates of the stable tail dependence function with the corresponding values of a parametrically specified proposal yields a novel minimum-distance estimator. The estimator is easy to calculate and applies to a wide range of sampling schemes and tail dependence models. In large samples, it is asymptotically normal with an explicit and estimable covariance matrix. The minimum distance obtained forms the basis of a goodness-of-fit statistic whose asymptotic distribution is chi-square. Extensive Monte Carlo simulations confirm the excellent finite-sample performance of the estimator and demonstrate that it is a strong competitor to currently available methods. The estimator is then applied to disentangle sources of tail dependence in European stock markets.

preprint2016arXiv

Marginal standardization of upper semicontinuous processes. with application to max-stable processes

Extreme-value theory for random vectors and stochastic processes with continuous trajectories is usually formulated for random objects all of whose univariate marginal distributions are identical. In the spirit of Sklar's theorem from copula theory, such marginal standardization is carried out by the pointwise probability integral transform. Certain situations, however, call for stochastic models whose trajectories are not continuous but merely upper semicontinuous (usc). Unfortunately, the pointwise application of the probability integral transform to a usc process does in general not preserve the upper semicontinuity of the trajectories. In the present work, we give sufficient conditions for marginal standardization of usc processes to be possible, and we state a partial extension of Sklar's theorem for usc processes. We specialize the results to max-stable processes whose marginal distributions and normalizing sequences are allowed to vary with the coordinate.

preprint2016arXiv

Maximum likelihood estimation for the Fréchet distribution based on block maxima extracted from a time series

The block maxima method in extreme-value analysis proceeds by fitting an extreme-value distribution to a sample of block maxima extracted from an observed stretch of a time series. The method is usually validated under two simplifying assumptions: the block maxima should be distributed according to an extreme-value distribution and the sample of block maxima should be independent. Both assumptions are only approximately true. For general triangular arrays of block maxima attracted to the Fréchet distribution, consistency and asymptotic normality is established for the maximum likelihood estimator of the parameters of the limiting Fréchet distribution. The results are specialized to the setting of block maxima extracted from a strictly stationary time series. The case where the underlying random variables are independent and identically distributed is further worked out in detail. The results are illustrated by theoretical examples and Monte Carlo simulations.

preprint2016arXiv

The Empirical Beta Copula

Given a sample from a multivariate distribution $F$, the uniform random variates generated independently and rearranged in the order specified by the componentwise ranks of the original sample look like a sample from the copula of $F$. This idea can be regarded as a variant on Baker's [J. Multivariate Anal. 99 (2008) 2312--2327] copula construction and leads to the definition of the empirical beta copula. The latter turns out to be a particular case of the empirical Bernstein copula, the degrees of all Bernstein polynomials being equal to the sample size. Necessary and sufficient conditions are given for a Bernstein polynomial to be a copula. These imply that the empirical beta copula is a genuine copula. Furthermore, the empirical process based on the empirical Bernstein copula is shown to be asymptotically the same as the ordinary empirical copula process under assumptions which are significantly weaker than those given in Janssen, Swanepoel and Veraverbeke [J. Stat. Plan. Infer. 142 (2012) 1189--1197]. A Monte Carlo simulation study shows that the empirical beta copula outperforms the empirical copula and the empirical checkerboard copula in terms of both bias and variance. Compared with the empirical Bernstein copula with the smoothing rate suggested by Janssen et al., its finite-sample performance is still significantly better in several cases, especially in terms of bias.

preprint2015arXiv

An M-estimator of spatial tail dependence

Tail dependence models for distributions attracted to a max-stable law are fitted using observations above a high threshold. To cope with spatial, high-dimensional data, a rank-based M-estimator is proposed relying on bivariate margins only. A data-driven weight matrix is used to minimize the asymptotic variance. Empirical process arguments show that the estimator is consistent and asymptotically normal. Its finite-sample performance is assessed in simulation experiments involving popular max-stable processes perturbed with additive noise. An analysis of wind speed data from the Netherlands illustrates the method.

preprint2014arXiv

Detecting changes in cross-sectional dependence in multivariate time series

Classical and more recent tests for detecting distributional changes in multivariate time series often lack power against alternatives that involve changes in the cross-sectional dependence structure. To be able to detect such changes better, a test is introduced based on a recently studied variant of the sequential empirical copula process. In contrast to earlier attempts, ranks are computed with respect to relevant subsamples, with beneficial consequences for the sensitivity of the test. For the computation of p-values we propose a multiplier resampling scheme that takes the serial dependence into account. The large-sample theory for the test statistic and the resampling scheme is developed. The finite-sample performance of the procedure is assessed by Monte Carlo simulations. Two case studies involving time series of financial returns are presented as well.

preprint2014arXiv

Extreme value copula estimation based on block maxima of a multivariate stationary time series

The core of the classical block maxima method consists of fitting an extreme value distribution to a sample of maxima over blocks extracted from an underlying series. In asymptotic theory, it is usually postulated that the block maxima are an independent random sample of an extreme value distribution. In practice however, block sizes are finite, so that the extreme value postulate will only hold approximately. A more accurate asymptotic framework is that of a triangular array of block maxima, the block size depending on the size of the underlying sample in such a way that both the block size and the number of blocks within that sample tend to infinity. The copula of the vector of componentwise maxima in a block is assumed to converge to a limit, which, under mild conditions, is then necessarily an extreme value copula. Under this setting and for absolutely regular stationary sequences, the empirical copula of the sample of vectors of block maxima is shown to be a consistent and asymptotically normal estimator for the limiting extreme value copula. Moreover, the empirical copula serves as a basis for rank-based, nonparametric estimation of the Pickands dependence function of the extreme value copula. The results are illustrated by theoretical examples and a Monte Carlo simulation study.

preprint2014arXiv

Hybrid Copula Estimators

An extension of the empirical copula is considered by combining an estimator of a multivariate cumulative distribution function with estimators of the marginal cumulative distribution functions for marginal estimators that are not necessarily equal to the margins of the joint estimator. Such a hybrid estimator may be reasonable when there is additional information available for some margins in the form of additional data or stronger modelling assumptions. A functional central limit theorem is established and some examples are developed.

preprint2014arXiv

Markov tail chains

The extremes of a univariate Markov chain with regulary varying stationary marginal distribution and asymptotically linear behavior are known to exhibit a multiplicative random walk structure called the tail chain. In this paper, we extend this fact to Markov chains with multivariate regularly varying marginal distribution in R^d. We analyze both the forward and the backward tail process and show that they mutually determine each other through a kind of adjoint relation. In a broader setting, it will be seen that even for non-Markovian underlying processes a Markovian forward tail chain always implies that the backward tail chain is Markovian as well. We analyze the resulting class of limiting processes in detail. Applications of the theory yield the asymptotic distribution of both the past and the future of univariate and multivariate stochastic difference equations conditioned on an extreme event.

preprint2014arXiv

Max-factor individual risk models with application to credit portfolios

Individual risk models need to capture possible correlations as failing to do so typically results in an underestimation of extreme quantiles of the aggregate loss. Such dependence modelling is particularly important for managing credit risk, for instance, where joint defaults are a major cause of concern. Often, the dependence between the individual loss occurrence indicators is driven by a small number of unobservable factors. Conditional loss probabilities are then expressed as monotone functions of linear combinations of these hidden factors. However, combining the factors in a linear way allows for some compensation between them. Such diversification effects are not always desirable and this is why the present work proposes a new model replacing linear combinations with maxima. These max-factor models give more insight into which of the factors is dominant.

preprint2014arXiv

Nonparametric estimation of extremal dependence

There is an increasing interest to understand the dependence structure of a random vector not only in the center of its distribution but also in the tails. Extreme-value theory tackles the problem of modelling the joint tail of a multivariate distribution by modelling the marginal distributions and the dependence structure separately. For estimating dependence at high levels, the stable tail dependence function and the spectral measure are particularly convenient. These objects also lie at the basis of nonparametric techniques for modelling the dependence among extremes in the max-domain of attraction setting. In case of asymptotic independence, this setting is inadequate, and more refined tail dependence coefficients exist, serving, among others, to discriminate between asymptotic dependence and independence. Throughout, the methods are illustrated on financial data.

preprint2014arXiv

On the asymptotic distribution of the mean absolute deviation about the mean

The mean absolute deviation about the mean is an alternative to the standard deviation for measuring dispersion in a sample or in a population. For stationary, ergodic time series with a finite first moment, an asymptotic expansion for the sample mean absolute deviation is proposed. The expansion yields the asymptotic distribution of the sample mean absolute deviation under a wide range of settings, allowing for serial dependence or an infinite second moment.

preprint2014arXiv

Semiparametric Gaussian copula models: Geometry and efficient rank-based estimation

We propose, for multivariate Gaussian copula models with unknown margins and structured correlation matrices, a rank-based, semiparametrically efficient estimator for the Euclidean copula parameter. This estimator is defined as a one-step update of a rank-based pilot estimator in the direction of the efficient influence function, which is calculated explicitly. Moreover, finite-dimensional algebraic conditions are given that completely characterize efficiency of the pseudo-likelihood estimator and adaptivity of the model with respect to the unknown marginal distributions. For correlation matrices structured according to a factor model, the pseudo-likelihood estimator turns out to be semiparametrically efficient. On the other hand, for Toeplitz correlation matrices, the asymptotic relative efficiency of the pseudo-likelihood estimator can be as low as 20%. These findings are confirmed by Monte Carlo simulations. We indicate how our results can be extended to joint regression models.

preprint2014arXiv

Statistics for Tail Processes of Markov Chains

At high levels, the asymptotic distribution of a stationary, regularly varying Markov chain is conveniently given by its tail process. The latter takes the form of a geometric random walk, the increment distribution depending on the sign of the process at the current state and on the flow of time, either forward or backward. Estimation of the tail process provides a nonparametric approach to analyze extreme values. A duality between the distributions of the forward and backward increments provides additional information that can be exploited in the construction of more efficient estimators. The large-sample distribution of such estimators is derived via empirical process theory for cluster functionals. Their finite-sample performance is evaluated via Monte Carlo simulations involving copula-based Markov models and solutions to stochastic recurrence equations. The estimators are applied to stock price data to study the absence or presence of symmetries in the succession of large gains and losses.

preprint2014arXiv

When uniform weak convergence fails: Empirical processes for dependence functions and residuals via epi- and hypographs

In the past decades, weak convergence theory for stochastic processes has become a standard tool for analyzing the asymptotic properties of various statistics. Routinely, weak convergence is considered in the space of bounded functions equipped with the supremum metric. However, there are cases when weak convergence in those spaces fails to hold. Examples include empirical copula and tail dependence processes and residual empirical processes in linear regression models in case the underlying distributions lack a certain degree of smoothness. To resolve the issue, a new metric for locally bounded functions is introduced and the corresponding weak convergence theory is developed. Convergence with respect to the new metric is related to epi- and hypo-convergence and is weaker than uniform convergence. Still, for continuous limits, it is equivalent to locally uniform convergence, whereas under mild side conditions, it implies $L^p$ convergence. For the examples mentioned above, weak convergence with respect to the new metric is established in situations where it does not occur with respect to the supremum distance. The results are applied to obtain asymptotic properties of resampling procedures and goodness-of-fit tests.

preprint2013arXiv

Nonparametric estimation of the tree structure of a nested Archimedean copula

One of the features inherent in nested Archimedean copulas, also called hierarchical Archimedean copulas, is their rooted tree structure. A nonparametric, rank-based method to estimate this structure is presented. The idea is to represent the target structure as a set of trivariate structures, each of which can be estimated individually with ease. Indeed, for any three variables there are only four possible rooted tree structures and, based on a sample, a choice can be made by performing comparisons between the three bivariate margins of the empirical distribution of the three variables. The set of estimated trivariate structures can then be used to build an estimate of the target structure. The advantage of this estimation method is that it does not require any parametric assumptions concerning the generator functions at the nodes of the tree.

preprint2012arXiv

A Euclidean likelihood estimator for bivariate tail dependence

The spectral measure plays a key role in the statistical modeling of multivariate extremes. Estimation of the spectral measure is a complex issue, given the need to obey a certain moment condition. We propose a Euclidean likelihood-based estimator for the spectral measure which is simple and explicitly defined, with its expression being free of Lagrange multipliers. Our estimator is shown to have the same limit distribution as the maximum empirical likelihood estimator of J. H. J. Einmahl and J. Segers, Annals of Statistics 37(5B), 2953--2989 (2009). Numerical experiments suggest an overall good performance and identical behavior to the maximum empirical likelihood estimator. We illustrate the method in an extreme temperature data analysis.

preprint2012arXiv

A functional limit theorem for dependent sequences with infinite variance stable limits

Under an appropriate regular variation condition, the affinely normalized partial sums of a sequence of independent and identically distributed random variables converges weakly to a non-Gaussian stable random variable. A functional version of this is known to be true as well, the limit process being a stable Lévy process. The main result in the paper is that for a stationary, regularly varying sequence for which clusters of high-threshold excesses can be broken down into asymptotically independent blocks, the properly centered partial sum process still converges to a stable Lévy process. Due to clustering, the Lévy triple of the limit process can be different from the one in the independent case. The convergence takes place in the space of càdlàg functions endowed with Skorohod's $M_1$ topology, the more usual $J_1$ topology being inappropriate as the partial sum processes may exhibit rapid successions of jumps within temporal clusters of large values, collapsing in the limit to a single jump. The result rests on a new limit theorem for point processes which is of independent interest. The theory is applied to moving average processes, squared GARCH(1,1) processes and stochastic volatility models.

preprint2012arXiv

An M-estimator for tail dependence in arbitrary dimensions

Consider a random sample in the max-domain of attraction of a multivariate extreme value distribution such that the dependence structure of the attractor belongs to a parametric model. A new estimator for the unknown parameter is defined as the value that minimizes the distance between a vector of weighted integrals of the tail dependence function and their empirical counterparts. The minimization problem has, with probability tending to one, a unique, global solution. The estimator is consistent and asymptotically normal. The spectral measures of the tail dependence models to which the method applies can be discrete or continuous. Examples demonstrate the applicability and the performance of the method.

preprint2012arXiv

Asymptotics of empirical copula processes under non-restrictive smoothness assumptions

Weak convergence of the empirical copula process is shown to hold under the assumption that the first-order partial derivatives of the copula exist and are continuous on certain subsets of the unit hypercube. The assumption is non-restrictive in the sense that it is needed anyway to ensure that the candidate limiting process exists and has continuous trajectories. In addition, resampling methods based on the multiplier central limit theorem, which require consistent estimation of the first-order derivatives, continue to be valid. Under certain growth conditions on the second-order partial derivatives that allow for explosive behavior near the boundaries, the almost sure rate in Stute's representation of the empirical copula process can be recovered. The conditions are verified, for instance, in the case of the Gaussian copula with full-rank correlation matrix, many Archimedean copulas, and many extreme-value copulas.

preprint2012arXiv

Max-stable models for multivariate extremes

Multivariate extreme-value analysis is concerned with the extremes in a multivariate random sample, that is, points of which at least some components have exceptionally large values. Mathematical theory suggests the use of max-stable models for univariate and multivariate extremes. A comprehensive account is given of the various ways in which max-stable models are described. Furthermore, a construction device is proposed for generating parametric families of max-stable distributions. Although the device is not new, its role as a model generator seems not yet to have been fully exploited.

preprint2012arXiv

Nonparametric Bayesian Inference on Bivariate Extremes

The tail of a bivariate distribution function in the domain of attraction of a bivariate extreme-value distribution may be approximated by the one of its extreme-value attractor. The extreme-value attractor has margins that belong to a three-parameter family and a dependence structure which is characterised by a spectral measure, that is a probability measure on the unit interval with mean equal to one half. As an alternative to parametric modelling of the spectral measure, we propose an infinite-dimensional model which is at the same time manageable and still dense within the class of spectral measures. Inference is done in a Bayesian framework, using the censored-likelihood approach. In particular, we construct a prior distribution on the class of spectral measures and develop a trans-dimensional Markov chain Monte Carlo algorithm for numerical computations. The method provides a bivariate predictive density which can be used for predicting the extreme outcomes of the bivariate distribution. In a practical perspective, this is useful for computing rare event probabilities and extreme conditional quantiles. The methodology is validated by simulations and applied to a data-set of Danish fire insurance claims.

preprint2012arXiv

Nonparametric estimation of pair-copula constructions with the empirical pair-copula

A pair-copula construction is a decomposition of a multivariate copula into a structured system, called regular vine, of bivariate copulae or pair-copulae. The standard practice is to model these pair-copulae parametrically, which comes at the cost of a large model risk, with errors propagating throughout the vine structure. The empirical pair-copula proposed in the paper provides a nonparametric alternative still achieving the parametric convergence rate. It can be used as a basis for inference on dependence measures, for selecting and pruning the vine structure, and for hypothesis tests concerning the form of the pair-copulae.

preprint2011arXiv

Large-sample tests of extreme-value dependence for multivariate copulas

Starting from the characterization of extreme-value copulas based on max-stability, large-sample tests of extreme-value dependence for multivariate copulas are studied. The two key ingredients of the proposed tests are the empirical copula of the data and a multiplier technique for obtaining approximate p-values for the derived statistics. The asymptotic validity of the multiplier approach is established, and the finite-sample performance of a large number of candidate test statistics is studied through extensive Monte Carlo experiments for data sets of dimension two to five. In the bivariate case, the rejection rates of the best versions of the tests are compared with those of the test of Ghoudi, Khoudraji and Rivest (1998) recently revisited by Ben Ghorbal, Genest and Neslehova (2009). The proposed procedures are illustrated on bivariate financial data and trivariate geological data.

preprint2011arXiv

Measuring Association between Random Vectors

This paper suggests five measures of association between two random vectors X = (X_1, ..., X_p) and Y = (Y_1, ..., Y_q). They are copula based and therefore invariant with respect to the marginal distributions of the components X_i and Y_j. The measures capture positive as well as negative association of X and Y. In case p = q = 1 they reduce to Spearman's rho. Various properties of these new measures are investigated. Nonparametric estimators, based on ranks, for the measures are derived and their small sample behaviour is investigated by simulation. The measures are applied to characterise strength and direction of association of bond and stock indices of five countries over time.

preprint2011arXiv

Nonparametric estimation of multivariate extreme-value copulas

Extreme-value copulas arise in the asymptotic theory for componentwise maxima of independent random samples. An extreme-value copula is determined by its Pickands dependence function, which is a function on the unit simplex subject to certain shape constraints that arise from an integral transform of an underlying measure called spectral measure. Multivariate extensions are provided of certain rank-based nonparametric estimators of the Pickands dependence function. The shape constraint that the estimator should itself be a Pickands dependence function is enforced by replacing an initial estimator by its best least-squares approximation in the set of Pickands dependence functions having a discrete spectral measure supported on a sufficiently fine grid. Weak convergence of the standardized estimators is demonstrated and the finite-sample performance of the estimators is investigated by means of a simulation experiment.

preprint2010arXiv

Regularly varying time series in Banach spaces

When a spatial process is recorded over time and the observation at a given time instant is viewed as a point in a function space, the result is a time series taking values in a Banach space. To study the spatio-temporal extremal dynamics of such a time series, the latter is assumed to be jointly regularly varying. This assumption is shown to be equivalent to convergence in distribution of the rescaled time series conditionally on the event that at a given moment in time it is far away from the origin. The limit is called the tail process or the spectral process depending on the way of rescaling. These processes provide convenient starting points to study, for instance, joint survival functions, tail dependence coefficients, extremograms, extremal indices, and point processes of extremes. The theory applies to linear processes composed of infinite sums of linearly transformed independent random elements whose common distribution is regularly varying.