Source author record

Theofanis Sapatinas

Theofanis Sapatinas appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

17works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

17 published item(s)

preprint2020arXiv

Testing Equality of Autocovariance Operators for Functional Time Series

We consider strictly stationary stochastic processes of Hilbert space-valued random variables and focus on fully functional tests for the equality of the lag-zero autocovariance operators of several independent functional time series. A moving block bootstrap-based testing procedure is proposed which generates pseudo random elements that satisfy the null hypothesis of interest. It is based on directly bootstrapping the time series of tensor products which overcomes some common difficulties associated with applications of the bootstrap to related testing problems. The suggested methodology can be potentially applied to a broad range of test statistics of the hypotheses of interest. As an example, we establish validity for approximating the distribution under the null of a test statistic based on the Hilbert-Schmidt distance of the corresponding sample lag-zero autocovariance operators, and show consistency under the alternative. As a prerequisite, we prove a central limit theorem for the moving block bootstrap procedure applied to the sample autocovariance operator which is of interest on its own. The finite sample size and power performance of the suggested moving block bootstrap-based testing procedure is illustrated through simulations and an application to a real-life dataset is discussed.

preprint2020arXiv

Testing equality of spectral density operators for functional linear processes

The problem of testing equality of the entire second order structure of two independent functional linear processes is considered. A fully functional $L^2$-type test is developed which evaluates, over all frequencies, the Hilbert-Schmidt distance between the estimated spectral density operators of the two processes. The asymptotic behavior of the test statistic is investigated and its limiting distribution under the null hypothesis is derived. Furthermore, a novel frequency domain bootstrap method is developed which approximates more accurately the distribution of the test statistic under the null than the large sample Gaussian approximation obtained. Asymptotic validity of the bootstrap procedure is established and consistency of the bootstrap-based test under the alternative is proved. Numerical simulations show that, even for small samples, the bootstrap-based test has very good size and power behavior. An application to meteorological functional time series is also presented.

preprint2016arXiv

Bootstrap-Based K-Sample Testing For Functional Data

We investigate properties of a bootstrap-based methodology for testing hypotheses about equality of certain characteristics of the distributions between different populations in the context of functional data. The suggested testing methodology is simple and easy to implement. It resamples the original dataset in such a way that the null hypothesis of interest is satisfied and it can be potentially applied to a wide range of testing problems and test statistics of interest. Furthermore, it can be utilized to the case where more than two populations of functional data are considered. We illustrate the bootstrap procedure by considering the important problems of testing the equality of mean functions or the equality of covariance functions (resp. covariance operators) between two populations. Theoretical results that justify the validity of the suggested bootstrap-based procedure are established. Furthermore, simulation results demonstrate very good size and power performances in finite sample situations, including the case of testing problems and/or sample sizes where asymptotic considerations do not lead to satisfactory approximations. A real-life dataset analyzed in the literature is also examined.

preprint2015arXiv

A unified treatment for non-asymptotic and asymptotic approaches to minimax signal detection

We are concerned with minimax signal detection. In this setting, we discuss non-asymptotic and asymptotic approaches through a unified treatment. In particular, we consider a Gaussian sequence model that contains classical models as special cases, such as, direct, well-posed inverse and ill-posed inverse problems. Working with certain ellipsoids in the space of squared-summable sequences of real numbers, with a ball of positive radius removed, we compare the construction of lower and upper bounds for the minimax separation radius (non-asymptotic approach) and the minimax separation rate (asymptotic approach) that have been proposed in the literature. Some additional contributions, bringing into light links between non-asymptotic and asymptotic approaches to minimax signal, are also presented. An example of a mildly ill-posed inverse problem is used for illustrative purposes. In particular, it is shown that tools used to derive `asymptotic' results can be exploited to draw `non-asymptotic' conclusions, and vice-versa.

preprint2015arXiv

Minimax Goodness-of-Fit Testing in Ill-Posed Inverse Problems with Partially Unknown Operators

We consider a Gaussian sequence model that contains ill-posed inverse problems as special cases. We assume that the associated operator is partially unknown in the sense that its singular functions are known and the corresponding singular values are unknown but observed with Gaussian noise. For the considered model, we study the minimax goodness-of-fit testing problem. Working with certain ellipsoids in the space of squared-summable sequences of real numbers, with a ball of positive radius removed, we obtain lower and upper bounds for the minimax separation radius in the non-asymptotic framework, i.e., for fixed values of the involved noise levels. Examples of mildly and severely ill-posed inverse problems with ellipsoids of ordinary-smooth and super-smooth sequences are examined in detail and minimax rates of goodness-of-fit testing are obtained for illustrative purposes.

preprint2014arXiv

Multichannel Deconvolution with Long Range Dependence: Upper bounds on the $L^p$-risk $(1 \le p < \infty)$

We consider multichannel deconvolution in a periodic setting with long-memory errors under three different scenarios for the convolution operators, i.e., super-smooth, regular-smooth and box-car convolutions. We investigate global performances of linear and hard-thresholded non-linear wavelet estimators for functions over a wide range of Besov spaces and for a variety of loss functions defining the risk. In particular, we obtain upper bounds on convergence rates using the $L^p$-risk $(1 \le p < \infty)$. Contrary to the case where the errors follow independent Brownian motions, it is demonstrated that multichannel deconvolution with errors that follow independent fractional Brownian motions with different Hurst parameters results in a much more involved situation. An extensive finite-sample numerical study is performed to supplement the theoretical findings.

preprint2013arXiv

Multichannel Deconvolution with Long-Range Dependence: A Minimax Study

We consider the problem of estimating the unknown response function in the multichannel deconvolution model with long-range dependent Gaussian errors. We do not limit our consideration to a specific type of long-range dependence rather we assume that the errors should satisfy a general assumption in terms of the smallest and larger eigenvalues of their covariance matrices. We derive minimax lower bounds for the quadratic risk in the proposed multichannel deconvolution model when the response function is assumed to belong to a Besov ball and the blurring function is assumed to possess some smoothness properties, including both regular-smooth and super-smooth convolutions. Furthermore, we propose an adaptive wavelet estimator of the response function that is asymptotically optimal (in the minimax sense), or near-optimal within a logarithmic factor, in a wide range of Besov balls. It is shown that the optimal convergence rates depend on the balance between the smoothness parameter of the response function, the kernel parameters of the blurring function, the long memory parameters of the errors, and how the total number of observations is distributed among the total number of channels. Some examples of inverse problems in mathematical physics where one needs to recover initial or boundary conditions on the basis of observations from a noisy solution of a partial differential equation are used to illustrate the application of the theory we developed. The optimal convergence rates and the adaptive estimators we consider extend the ones studied by Pensky and Sapatinas (2009, 2010) for independent and identically distributed Gaussian errors to the case of long-range dependent Gaussian errors.

preprint2012arXiv

Minimax signal detection in ill-posed inverse problems

Ill-posed inverse problems arise in various scientific fields. We consider the signal detection problem for mildly, severely and extremely ill-posed inverse problems with $l^q$-ellipsoids (bodies), $q\in(0,2]$, for Sobolev, analytic and generalized analytic classes of functions under the Gaussian white noise model. We study both rate and sharp asymptotics for the error probabilities in the minimax setup. By construction, the derived tests are, often, nonadaptive. Minimax rate-optimal adaptive tests of rather simple structure are also constructed.

preprint2012arXiv

Nonparametric adaptive time-dependent multivariate function estimation

We consider the nonparametric estimation problem of time-dependent multivariate functions observed in a presence of additive cylindrical Gaussian white noise of a small intensity. We derive minimax lower bounds for the $L^2$-risk in the proposed spatio-temporal model as the intensity goes to zero, when the underlying unknown response function is assumed to belong to a ball of appropriately constructed inhomogeneous time-dependent multivariate functions, motivated by practical applications. Furthermore, we propose both non-adaptive linear and adaptive non-linear wavelet estimators that are asymptotically optimal (in the minimax sense) in a wide range of the so-constructed balls of inhomogeneous time-dependent multivariate functions. The usefulness of the suggested adaptive nonlinear wavelet estimator is illustrated with the help of simulated and real-data examples.

preprint2012arXiv

Nonparametric Regression Estimation Based on Spatially Inhomogeneous Data: Minimax Global Convergence Rates and Adaptivity

We consider the nonparametric regression estimation problem of recovering an unknown response function f on the basis of spatially inhomogeneous data when the design points follow a known compactly supported density g with a finite number of well separated zeros. In particular, we consider two different cases: when g has zeros of a polynomial order and when g has zeros of an exponential order. These two cases correspond to moderate and severe data losses, respectively. We obtain asymptotic minimax lower bounds for the global risk of an estimator of f and construct adaptive wavelet nonlinear thresholding estimators of f which attain those minimax convergence rates (up to a logarithmic factor in the case of a zero of a polynomial order), over a wide range of Besov balls. The spatially inhomogeneous ill-posed problem that we investigate is inherently more difficult than spatially homogeneous problems like, e.g., deconvolution. In particular, due to spatial irregularity, assessment of minimax global convergence rates is a much harder task than the derivation of minimax local convergence rates studied recently in the literature. Furthermore, the resulting estimators exhibit very different behavior and minimax global convergence rates in comparison with the solution of spatially homogeneous ill-posed problems. For example, unlike in deconvolution problem, the minimax global convergence rates are greatly influenced not only by the extent of data loss but also by the degree of spatial homogeneity of f. Specifically, even if 1/g is not integrable, one can recover f as well as in the case of an equispaced design (in terms of minimax global convergence rates) when it is homogeneous enough since the estimator is "borrowing strength" in the areas where f is adequately sampled.

preprint2012arXiv

Short-Term Load Forecasting: The Similar Shape Functional Time Series Predictor

We introduce a novel functional time series methodology for short-term load forecasting. The prediction is performed by means of a weighted average of past daily load segments, the shape of which is similar to the expected shape of the load segment to be predicted. The past load segments are identified from the available history of the observed load segments by means of their closeness to a so-called reference load segment, the later being selected in a manner that captures the expected qualitative and quantitative characteristics of the load segment to be predicted. Weak consistency of the suggested functional similar shape predictor is established. As an illustration, we apply the suggested functional time series forecasting methodology to historical daily load data in Cyprus and compare its performance to that of a recently proposed alternative functional time series methodology for short-term load forecasting.

preprint2011arXiv

Multichannel Boxcar Deconvolution with Growing Number of Channels

We consider the problem of estimating the unknown response function in the multichannel deconvolution model with a boxcar-like kernel which is of particular interest in signal processing. It is known that, when the number of channels is finite, the precision of reconstruction of the response function increases as the number of channels $M$ grow (even when the total number of observations $n$ for all channels $M$ remains constant) and this requires that the parameter of the channels form a Badly Approximable $M$-tuple. Recent advances in data collection and recording techniques made it of urgent interest to study the case when the number of channels $M=M_n$ grow with the total number of observations $n$. However, in real-life situations, the number of channels $M = M_n$ usually refers to the number of physical devices and, consequently, may grow to infinity only at a slow rate as $n \rightarrow \infty$. When $M=M_n$ grows slowly as $n$ increases, we develop a procedure for the construction of a Badly Approximable $M$-tuple on a specified interval, of a non-asymptotic length, together with a lower bound associated with this $M$-tuple, which explicitly shows its dependence on $M$ as $M$ is growing. This result is further used for the evaluation of the $L^2$-risk of the suggested adaptive wavelet thresholding estimator of the unknown response function and, furthermore, for the choice of the optimal number of channels $M$ which minimizes the $L^2$-risk.

preprint2011arXiv

Some new approaches to infinite divisibility

Using an approach based, amongst other things, on Proposition 1 of Kaluza (1928), Goldie (1967) and, using a different approach based especially on zeros of polynomials, Steutel (1967) have proved that each nondegenerate distribution function (d.f.) $F$ (on $\RR$, the real line), satisfying $F(0-) = 0$ and $F(x) = F(0) + (1-F(0)) G(x)$, $x > 0$, where $G$ is the d.f. corresponding to a mixture of exponential distributions, is infinitely divisible. Indeed, Proposition 1 of Kaluza (1928) implies that any nondegenerate discrete probability distribution ${p_x: x= 0,1, ...}$ that is log-convex or, in particular, completely monotone, is compound geometric, and, hence, infinitely divisible. Steutel (1970), Shanbhag & Sreehari (1977) and Steutel & van Harn (2004, Chapter VI) have given certain extensions or variations of one or more of these results. Following a modified version of the C.R. Rao et al. (2009, Section 4) approach based on the Wiener-Hopf factorization, we establish some further results of significance to the literature on infinite divisibility.

preprint2010arXiv

On convergence rates equivalency and sampling strategies in functional deconvolution models

Using the asymptotical minimax framework, we examine convergence rates equivalency between a continuous functional deconvolution model and its real-life discrete counterpart over a wide range of Besov balls and for the $L^2$-risk. For this purpose, all possible models are divided into three groups. For the models in the first group, which we call uniform, the convergence rates in the discrete and the continuous models coincide no matter what the sampling scheme is chosen, and hence the replacement of the discrete model by its continuous counterpart is legitimate. For the models in the second group, to which we refer as regular, one can point out the best sampling strategy in the discrete model, but not every sampling scheme leads to the same convergence rates; there are at least two sampling schemes which deliver different convergence rates in the discrete model (i.e., at least one of the discrete models leads to convergence rates that are different from the convergence rates in the continuous model). The third group consists of models for which, in general, it is impossible to devise the best sampling strategy; we call these models irregular. We formulate the conditions when each of these situations takes place. In the regular case, we not only point out the number and the selection of sampling points which deliver the fastest convergence rates in the discrete model but also investigate when, in the case of an arbitrary sampling scheme, the convergence rates in the continuous model coincide or do not coincide with the convergence rates in the discrete model. We also study what happens if one chooses a uniform, or a more general pseudo-uniform, sampling scheme which can be viewed as an intuitive replacement of the continuous model.

preprint2009arXiv

Minimax Goodness-of-Fit Testing in Multivariate Nonparametric Regression

We consider an unknown response function $f$ defined on $Δ=[0,1]^d$, $1\le d\le\infty$, taken at $n$ random uniform design points and observed with Gaussian noise of known variance. Given a positive sequence $r_n\to 0$ as $n\to\infty$ and a known function $f_0 \in L_2(Δ)$, we propose, under general conditions, a unified framework for the goodness-of-fit testing problem for testing the null hypothesis $H_0: f=f_0$ against the alternative $H_1: f\in\CF, \|f-f_0\|\ge r_n$, where $\CF$ is an ellipsoid in the Hilbert space $ L_2(Δ)$ with respect to the tensor product Fourier basis and $\|\cdot\|$ is the norm in $ L_2(Δ)$. We obtain both rate and sharp asymptotics for the error probabilities in the minimax setup. The derived tests are inherently non-adaptive. Several illustrative examples are presented. In particular, we consider functions belonging to ellipsoids arising from the well-known multidimensional Sobolev and tensor product Sobolev norms as well as from the less-known Sloan-Wo$\rm\acute{z}$niakowski norm and a norm constructed from multivariable analytic functions on the complex strip. Some extensions of the suggested minimax goodness-of-fit testing methodology, covering the cases of general design schemes with a known product probability density function, unknown variance, other basis functions and adaptivity of the suggested tests, are also briefly discussed.

preprint2009arXiv

Moment properties of multivariate infinitely divisible laws and criteria for self-decomposability

Ramachandran (1969, Theorem 8) has shown that for any univariate infinitely divisible distribution and any positive real number $α$, an absolute moment of order $α$ relative to the distribution exists (as a finite number) if and only if this is so for a certain truncated version of the corresponding L$\acute{\rm e}$vy measure. A generalized version of this result in the case of multivariate infinitely divisible distributions, involving the concept of g-moments, is given by Sato (1999, Theorem 25.3). We extend Ramachandran's theorem to the multivariate case, keeping in mind the immediate requirements under appropriate assumptions of cumulant studies of the distributions referred to; the format of Sato's theorem just referred to obviously varies from ours and seems to be having a different agenda. Also, appealing to a further criterion based on the L$\acute{\rm e}$vy measure, we identify in a certain class of multivariate infinitely divisible distributions the distributions that are self-decomposable; this throws new light on structural aspects of certain multivariate distributions such as the multivariate generalized hyperbolic distributions studied by Barndorff-Nielsen (1977) and others. Various points of relevance to the study are also addressed through specific examples.