Source author record

Philippe Naveau

Philippe Naveau appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

10works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

10 published item(s)

preprint2026arXiv

A parsimonious tail compliant multiscale statistical model for aggregated rainfall

Modeling rainfall intensity distributions across aggregation scales (from sub-hourly to weekly) is essential for hydrological risk analysis and IDF curves. Aggregation naturally imposes mathematical constraints: return levels must be ordered by time scale, as daily accumulations necessarily exceed sub-daily ones. From a statistical perspective, each aggregation step should ideally not require additional parameters, yet parsimonious models describing the full distribution remain scarce, as most literature focuses on seasonal block maxima. In this study, we propose a parsimonious framework to model all rainfall intensities (low to large) across scales. We utilize the Extended Generalized Pareto Distribution (EGPD), which aligns with extreme value theory for both tails while remaining flexible for the bulk of the distribution. We establish a general result on the behavior of EGPD variables under various aggregation procedures. To overcome the difficulty of direct likelihood inference, we link the EGPD class to Poisson compound sums. This allows the use of the Panjer algorithm for efficient composite likelihood evaluation. Our approach ensures that return levels do not cross across scales and enables estimation for return periods below annual or seasonal levels. We demonstrate the method using sub-hourly series from six French stations with diverse climates. Only eight parameters are needed per station to capture scales from six minutes to three days. IDF curves above and below the annual scale are provided.

preprint2022arXiv

Stable sums to infer high return levels of multivariate rainfall time series

Heavy rainfall distributional modeling is essential in any impact studies linked to the water cycle, e.g.\ flood risks. Still, statistical analyses that both take into account the temporal and multivariate nature of extreme rainfall are rare, and often, a complex de-clustering step is needed to make extreme rainfall temporally independent. A natural question is how to bypass this de-clustering in a multivariate context. To address this issue, we introduce the stable sums method. Our goal is to incorporate time and space extreme dependencies in the analysis of heavy tails. To reach our goal, we build on large deviations of regularly varying stationary time series. Numerical experiments demonstrate that our novel approach enhances return levels inference in two ways. First, it is robust concerning time dependencies. We implement it alike on independent and dependent observations. In the univariate setting, it improves the accuracy of confidence intervals compared to the main estimators requiring temporal de-clustering. Second, it thoughtfully integrates the spatial dependencies. In simulation, the multivariate stable sums method has a smaller mean squared error than its component-wise implementation. We apply our method to infer high return levels of daily fall precipitation amounts from a national network of weather stations in France.

preprint2020arXiv

Climate extreme event attribution using multivariate peaks-over-thresholds modeling and counterfactual theory

Numerical climate models are complex and combine a large number of physical processes. They are key tools in quantifying the relative contribution of potential anthropogenic causes (e.g., the current increase in greenhouse gases) on high impact atmospheric variables like heavy rainfall. These so-called climate extreme event attribution problems are particularly challenging in a multivariate context, that is, when the atmospheric variables are measured on a possibly high-dimensional grid. In this paper, we leverage two statistical theories to assess causality in the context of multivariate extreme event attribution. As we consider an event to be extreme when at least one of the components of the vector of interest is large, extreme-value theory justifies, in an asymptotical sense, a multivariate generalized Pareto distribution to model joint extremes. Under this class of distributions, we derive and study probabilities of necessary and sufficient causation as defined by the counterfactual theory of Pearl. To increase causal evidence, we propose a dimension reduction strategy based on the optimal linear projection that maximizes such causation probabilities. Our approach is tested on simulated examples and applied to weekly winter maxima precipitation outputs of the French CNRM from the recent CMIP6 experiment.

preprint2020arXiv

Spatial Modeling of Heavy Precipitation by Coupling Weather Station Recordings and Ensemble Forecasts with Max-Stable Processes

Due to complex physical phenomena, the distribution of heavy rainfall events is difficult to model spatially. Physically based numerical models can often provide physically coherent spatial patterns, but may miss some important precipitation features like heavy rainfall intensities. Measurements at ground-based weather stations, however, supply adequate rainfall intensities, but most national weather recording networks are often spatially too sparse to capture rainfall patterns adequately. To bring the best out of these two sources of information, climatologists and hydrologists have been seeking models that can efficiently merge different types of rainfall data. One inherent difficulty is to capture the appropriate multivariate dependence structure among rainfall maxima. For this purpose, multivariate extreme value theory suggests the use of a max-stable process. Such a process can be represented by a max-linear combination of independent copies of a hidden stochastic process weighted by a Poisson point process. In practice, the choice of this hidden process is non-trivial, especially if anisotropy, non-stationarity and nugget effects are present in the spatial data at hand. By coupling forecast ensemble data from the French national weather service (Météo-France) with local observations, we construct and compare different types of data driven max-stable processes that are parsimonious in parameters, easy to simulate and capable of reproducing nugget effects and spatial non-stationarities. We also compare our new method with classical approaches from spatial extreme value theory such as Brown-Resnick processes.

preprint2016arXiv

Detecting distributional changes in samples of independent block maxima using probability weighted moments

The analysis of seasonal or annual block maxima is of interest in fields such as hydrology, climatology or meteorology. In connection with the celebrated method of block maxima, we study several tests that can be used to assess whether the available series of maxima is identically distributed. It is assumed that block maxima are independent but not necessarily generalized extreme value distributed. The asymptotic null distributions of the test statistics are investigated and the practical computation of approximate p-values is addressed. Extensive Monte-Carlo simulations show the adequate finite-sample behavior of the studied tests for a large number of realistic data generating scenarios. Illustrations on several environmental datasets conclude the work.

preprint2015arXiv

DADA: Data Assimilation for the Detection and Attribution of Weather- and Climate-related Events

We describe a new approach allowing for systematic causal attribution of weather and climate-related events, in near-real time. The method is purposely designed to facilitate its implementation at meteorological centers by relying on data treatments that are routinely performed when numerically forecasting the weather. Namely, we show that causal attribution can be obtained as a by-product of so-called data assimilation procedures that are run on a daily basis to update the meteorological model with new atmospheric observations; hence, the proposed methodology can take advantage of the powerful computational and observational capacity of weather forecasting centers. We explain the theoretical rationale of this approach and sketch the most prominent features of a "data assimilation-based detection and attribution" (DADA) procedure. The proposal is illustrated in the context of the classical three-variable Lorenz model with additional forcing. Several theoretical and practical research questions that need to be addressed to make the proposal readily operational within weather forecasting centers are finally laid out.

preprint2014arXiv

A frailty-contagion model for multi-site hourly precipitation driven by atmospheric covariates

Accurate stochastic simulations of hourly precipitation are needed for impact studies at local spatial scales. Statistically, hourly precipitation data represent a difficult challenge. They are non-negative, skewed, heavy tailed, contain a lot of zeros (dry hours) and they have complex temporal structures (e.g., long persistence of dry episodes). Inspired by frailty-contagion approaches used in finance and insurance, we propose a multi-site precipitation simulator that, given appropriate regional atmospheric variables, can simultaneously handle dry events and heavy rainfall periods. One advantage of our model is its conceptual simplicity in its dynamical structure. In particular, the temporal variability is represented by a common factor based on a few classical atmospheric covariates like temperatures, pressures and others. Our inference approach is tested on simulated data and applied on measurements made in the northern part of French Brittany.

preprint2013arXiv

Approximating the conditional density given large observed values via a multivariate extremes framework, with application to environmental data

Phenomena such as air pollution levels are of greatest interest when observations are large, but standard prediction methods are not specifically designed for large observations. We propose a method, rooted in extreme value theory, which approximates the conditional distribution of an unobserved component of a random vector given large observed values. Specifically, for $\mathbf{Z}=(Z_1,...,Z_d)^T$ and $\mathbf{Z}_{-d}=(Z_1,...,Z_{d-1})^T$, the method approximates the conditional distribution of $[Z_d|\mathbf{Z}_{-d}=\mathbf{z}_{-d}]$ when $|\mathbf{z}_{-d}|>r_*$. The approach is based on the assumption that $\mathbf{Z}$ is a multivariate regularly varying random vector of dimension $d$. The conditional distribution approximation relies on knowledge of the angular measure of $\mathbf{Z}$, which provides explicit structure for dependence in the distribution's tail. As the method produces a predictive distribution rather than just a point predictor, one can answer any question posed about the quantity being predicted, and, in particular, one can assess how well the extreme behavior is represented. Using a fitted model for the angular measure, we apply our method to nitrogen dioxide measurements in metropolitan Washington DC. We obtain a predictive distribution for the air pollutant at a location given the air pollutant's measurements at four nearby locations and given that the norm of the vector of the observed measurements is large.

preprint2007arXiv

Detecting spatial patterns with the cumulant function. Part I: The theory

In climate studies, detecting spatial patterns that largely deviate from the sample mean still remains a statistical challenge. Although a Principal Component Analysis (PCA), or equivalently a Empirical Orthogonal Functions (EOF) decomposition, is often applied on this purpose, it can only provide meaningful results if the underlying multivariate distribution is Gaussian. Indeed, PCA is based on optimizing second order moments quantities and the covariance matrix can only capture the full dependence structure for multivariate Gaussian vectors. Whenever the application at hand can not satisfy this normality hypothesis (e.g. precipitation data), alternatives and/or improvements to PCA have to be developed and studied. To go beyond this second order statistics constraint that limits the applicability of the PCA, we take advantage of the cumulant function that can produce higher order moments information. This cumulant function, well-known in the statistical literature, allows us to propose a new, simple and fast procedure to identify spatial patterns for non-Gaussian data. Our algorithm consists in maximizing the cumulant function. To illustrate our approach, its implementation for which explicit computations are obtained is performed on three family of of multivariate random vectors. In addition, we show that our algorithm corresponds to selecting the directions along which projected data display the largest spread over the marginal probability density tails.

preprint2007arXiv

Detecting spatial patterns with the cumulant function. Part II: An application to El Nino

The spatial coherence of a measured variable (e.g. temperature or pressure) is often studied to determine the regions where this variable varies the most or to find teleconnections, i.e. correlations between specific regions. While usual methods to find spatial patterns, such as Principal Components Analysis (PCA), are constrained by linear symmetries, the dependence of variables such as temperature or pressure at different locations is generally nonlinear. In particular, large deviations from the sample mean are expected to be strongly affected by such nonlinearities. Here we apply a newly developed nonlinear technique (Maxima of Cumulant Function, MCF) for the detection of typical spatial patterns that largely deviate from the mean. In order to test the technique and to introduce the methodology, we focus on the El Nino/Southern Oscillation and its spatial patterns. We find nonsymmetric temperature patterns corresponding to El Nino and La Nina, and we compare the results of MCF with other techniques, such as the symmetric solutions of PCA, and the nonsymmetric solutions of Nonlinear PCA (NLPCA). We found that MCF solutions are more reliable than the NLPCA fits, and can capture mixtures of principal components. Finally, we apply Extreme Value Theory on the temporal variations extracted from our methodology. We find that the tails of the distribution of extreme temperatures during La Nina episodes is bounded, while the tail during El Ninos is less likely to be bounded. This implies that the mean spatial patterns of the two phases are asymmetric, as well as the behaviour of their extremes.