Source author record

Stéphane Guerrier

Stéphane Guerrier appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2022arXiv

A penalized two-pass regression to predict stock returns with time-varying risk premia

We develop a penalized two-pass regression with time-varying factor loadings. The penalization in the first pass enforces sparsity for the time-variation drivers while also maintaining compatibility with the no-arbitrage restrictions by regularizing appropriate groups of coefficients. The second pass delivers risk premia estimates to predict equity excess returns. Our Monte Carlo results and our empirical results on a large cross-sectional data set of US individual stocks show that penalization without grouping can yield to nearly all estimated time-varying models violating the no-arbitrage restrictions. Moreover, our results demonstrate that the proposed method reduces the prediction errors compared to a penalized approach without appropriate grouping or a time-invariant factor model.

preprint2020arXiv

Asymptotically Optimal Bias Reduction for Parametric Models

An important challenge in statistical analysis concerns the control of the finite sample bias of estimators. This problem is magnified in high-dimensional settings where the number of variables $p$ diverges with the sample size $n$, as well as for nonlinear models and/or models with discrete data. For these complex settings, we propose to use a general simulation-based approach and show that the resulting estimator has a bias of order $\mathcal{O}(0)$, hence providing an asymptotically optimal bias reduction. It is based on an initial estimator that can be slightly asymptotically biased, making the approach very generally applicable. This is particularly relevant when classical estimators, such as the maximum likelihood estimator, can only be (numerically) approximated. We show that the iterative bootstrap of Kuk (1995) provides a computationally efficient approach to compute this bias reduced estimator. We illustrate our theoretical results in simulation studies for which we develop new bias reduced estimators for the logistic regression, with and without random effects. These estimators enjoy additional properties such as robustness to data contamination and to the problem of separability.

preprint2020arXiv

Robust Two-Step Wavelet-Based Inference for Time Series Models

Complex time series models such as (the sum of) ARMA$(p,q)$ models with additional noise, random walks, rounding errors and/or drifts are increasingly used for data analysis in fields such as biology, ecology, engineering and economics where the length of the observed signals can be extremely large. Performing inference on and/or prediction from these models can be highly challenging for several reasons: (i) the data may contain outliers that can adversely affect the estimation procedure; (ii) the computational complexity can become prohibitive when models include more than just a few parameters and/or the time series are large; (iii) model building and/or selection adds another layer of (computational) complexity to the previous task; and (iv) solutions that address (i), (ii) and (iii) simultaneously do not exist in practice. For this reason, this paper aims at jointly addressing these challenges by proposing a general framework for robust two-step estimation based on a bounded influence M-estimator of the wavelet variance. In this perspective, we first develop the conditions for the joint asymptotic normality of the latter estimator thereby providing the necessary tools to perform (direct) inference for scale-based analysis of signals. Taking advantage of the model-independent weights of this first-step estimator that are computed only once, we then develop the asymptotic properties of two-step robust estimators using the framework of the Generalized Method of Wavelet Moments (GMWM), hence defining the Robust GMWM (RGMWM) that we then use for robust model estimation and inference in a computationally efficient manner even for large time series. Simulation studies illustrate the good finite sample performance of the RGMWM estimator and applied examples highlight the practical relevance of the proposed approach.

preprint2016arXiv

Fast and Robust Parametric Estimation for Time Series and Spatial Models

We present a new framework for robust estimation and inference on second-order stationary time series and random fields. This framework is based on the Generalized Method of Wavelet Moments which uses the wavelet variance to achieve parameter estimation for complex models. Using an M-estimator of the wavelet variance, this method can be made robust therefore allowing to estimate the parameters of a wide range of time series and spatial models when the data suffers from outliers or different forms of contamination. The paper presents a series of simulation studies as well as a range of applications where this new approach can be considered as a computationally efficient, numerically stable and robust method which performs at least as well as existing methods in bounding the influence of outliers on the estimation procedure.

preprint2016arXiv

On the Identifiability of Latent Models for Dependent Data

The condition of parameter identifiability is essential for the consistency of all estimators and is often challenging to prove. As a consequence, this condition is often assumed for simplicity although this may not be straightforward to assume for a variety of model settings. In this paper we deal with a particular class of models that we refer to as "latent" models which can be defined as models made by the sum of underlying models, such as a variety of linear state-space models for time series. These models are of great importance in many fields, from ecology to engineering, and in this paper we prove the identifiability of a wide class of (second-order stationary) latent time series and spatial models and discuss what this implies for some extremum estimators, thereby reducing the conditions for their consistency to some very basic regularity conditions. Finally, a specific focus is given to the Generalized Method of Wavelet Moments estimator which is also able to estimate intrinsically second-order stationary models.

preprint2016arXiv

The gmwm R package: a comprehensive tool for time series analysis from state-space models to robustness

The gmwm R package for inference on time series models is mainly based on the quantity called wavelet variance which is derived from a wavelet decomposition of a time series. This quantity provides a means to summarize and graphically represent the features of time series in order to identify possible models. Moreover, it is used as a moment condition for model estimation through the generalized method of wavelet moments. Based on the latter method, this package not only provides an alternative method to estimate classical ARMA models but also delivers a general framework for the robust estimation of many time series models as well as a quick and efficient estimation of many linear state-space models.

preprint2016arXiv

Wavelet Variance for Random Fields: an M-Estimation Framework

We present a general M-estimation framework for inference on the wavelet variance. This framework generalizes the results on the scale-wise properties of the standard estimator and extends them to deliver the joint asymptotic properties of the estimated wavelet variance vector. Moreover, this is achieved by extending the estimation of the wavelet variance to multidimensional random fields and by stating the necessary conditions for these properties to hold when the size of the wavelet variance vector goes to infinity with the sample size. Finally, these results generally hold when using bounded estimating functions thereby delivering a robust framework for the estimation of this quantity which improves over existing methods both in terms of asymptotic properties and in terms of its finite sample performance. The proposed estimator is investigated in simulation studies and different applications highlighting its good properties.

preprint2015arXiv

A Paradigmatic Regression Algorithm for Gene Selection Problems

Motivation: Gene selection has become a common task in most gene expression studies. The objective of such research is often to identify the smallest possible set of genes that can still achieve good predictive performance. The problem of assigning tumours to a known class is a particularly important example that has received considerable attention in the last ten years. Many of the classification methods proposed recently require some form of dimension-reduction of the problem. These methods provide a single model as an output and, in most cases, rely on the likelihood function in order to achieve variable selection. Results: We propose a prediction-based objective function that can be tailored to the requirements of practitioners and can be used to assess and interpret a given problem. The direct optimization of such a function can be very difficult because the problem is potentially discontinuous and nonconvex. We therefore propose a general procedure for variable selection that resembles importance sampling to explore the feature space. Our proposal compares favorably with competing alternatives when applied to two cancer data sets in that smaller models are obtained for better or at least comparable classification errors. Furthermore by providing a set of selected models instead of a single one, we construct a network of possible models for a target prediction accuracy level.