Source author record

Evarist Giné

Evarist Giné appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

10works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

10 published item(s)

preprint2012arXiv

Rates of contraction for posterior distributions in $\bolds{L^r}$-metrics, $\bolds{1\le r\le\infty}$

The frequentist behavior of nonparametric Bayes estimates, more specifically, rates of contraction of the posterior distributions to shrinking $L^r$-norm neighborhoods, $1\le r\le\infty$, of the unknown parameter, are studied. A theorem for nonparametric density estimation is proved under general approximation-theoretic assumptions on the prior. The result is applied to a variety of common examples, including Gaussian process, wavelet series, normal mixture and histogram priors. The rates of contraction are minimax-optimal for $1\le r\le2$, but deteriorate as $r$ increases beyond 2. In the case of Gaussian nonparametric regression a Gaussian prior is devised for which the posterior contracts at the optimal rate in all $L^r$-norms, $1\le r\le\infty$.

preprint2011arXiv

Adaptive estimation of a distribution function and its density in sup-norm loss by wavelet and spline projections

Given an i.i.d. sample from a distribution $F$ on $\mathbb{R}$ with uniformly continuous density $p_0$, purely data-driven estimators are constructed that efficiently estimate $F$ in sup-norm loss and simultaneously estimate $p_0$ at the best possible rate of convergence over Hölder balls, also in sup-norm loss. The estimators are obtained by applying a model selection procedure close to Lepski's method with random thresholds to projections of the empirical measure onto spaces spanned by wavelets or $B$-splines. The random thresholds are based on suprema of Rademacher processes indexed by wavelet or spline projection kernels. This requires Bernstein-type analogs of the inequalities in Koltchinskii [Ann. Statist. 34 (2006) 2593-2656] for the deviation of suprema of empirical processes from their Rademacher symmetrizations.

preprint2010arXiv

Confidence bands in density estimation

Given a sample from some unknown continuous density $f:\mathbb{R}\to\mathbb{R}$, we construct adaptive confidence bands that are honest for all densities in a "generic" subset of the union of $t$-Hölder balls, $0<t\le r$, where $r$ is a fixed but arbitrary integer. The exceptional ("nongeneric") set of densities for which our results do not hold is shown to be nowhere dense in the relevant Hölder-norm topologies. In the course of the proofs we also obtain limit theorems for maxima of linear wavelet and kernel density estimators, which are of independent interest.

preprint2010arXiv

On the estimation of smooth densities by strict probability densities at optimal rates in sup-norm

It is shown that the variable bandwidth density estimator proposed by McKay (1993a and b) following earlier findings by Abramson (1982) approximates density functions in $C^4(\mathbb R^d)$ at the minimax rate in the supremum norm over bounded sets where the preliminary density estimates on which they are based are bounded away from zero. A somewhat more complicated estimator proposed by Jones McKay and Hu (1994) to approximate densities in $C^6(\mathbb R)$ is also shown to attain minimax rates in sup norm over the same kind of sets. These estimators are strict probability densities.

preprint2010arXiv

Uniform asymptotics for kernel density estimators with variable bandwidths

It is shown that the Hall, Hu and Marron [Hall, P., Hu, T., and Marron J.S. (1995), Improved Variable Window Kernel Estimates of Probability Densities, {\it Annals of Statistics}, 23, 1--10] modification of Abramson's [Abramson, I. (1982), On Bandwidth Variation in Kernel Estimates - A Square-root Law, {\it Annals of Statistics}, 10, 1217--1223] variable bandwidth kernel density estimator satisfies the optimal asymptotic properties for estimating densities with four uniformly continuous derivatives, uniformly on bounded sets where the preliminary estimator of the density is bounded away from zero.

preprint2006arXiv

Concentration inequalities and asymptotic results for ratio type empirical processes

Let $\mathcal{F}$ be a class of measurable functions on a measurable space $(S,\mathcal{S})$ with values in $[0,1]$ and let \[P_n=n^{-1}\sum_{i=1}^nδ_{X_i}\] be the empirical measure based on an i.i.d. sample $(X_1,...,X_n)$ from a probability distribution $P$ on $(S,\mathcal{S})$. We study the behavior of suprema of the following type: \[\sup_{r_n<σ_Pf\leq δ_n}\frac{|P_nf-Pf|}{ϕ(σ_Pf)},\] where $σ_Pf\ge\operatorname {Var}^{1/2}_Pf$ and $ϕ$ is a continuous, strictly increasing function with $ϕ(0)=0$. Using Talagrand's concentration inequality for empirical processes, we establish concentration inequalities for such suprema and use them to derive several results about their asymptotic behavior, expressing the conditions in terms of expectations of localized suprema of empirical processes. We also prove new bounds for expected values of sup-norms of empirical processes in terms of the largest $σ_Pf$ and the $L_2(P)$ norm of the envelope of the function class, which are especially suited for estimating localized suprema. With this technique, we extend to function classes most of the known results on ratio type suprema of empirical processes, including some of Alexander's results for VC classes of sets. We also consider applications of these results to several important problems in nonparametric statistics and in learning theory (including general excess risk bounds in empirical risk minimization and their versions for $L_2$-regression and classification and ratio type bounds for margin distributions in classification).

preprint2006arXiv

Empirical graph Laplacian approximation of Laplace--Beltrami operators: Large sample results

Let ${M}$ be a compact Riemannian submanifold of ${{\bf R}^m}$ of dimension $\scriptstyle{d}$ and let ${X_1,...,X_n}$ be a sample of i.i.d. points in ${M}$ with uniform distribution. We study the random operators $$ Δ_{h_n,n}f(p):=\frac{1}{nh_n^{d+2}}\sum_{i=1}^n K(\frac{p-X_i}{h_n})(f(X_i)-f(p)), p\in M $$ where ${K(u):={\frac{1}{(4π)^{d/2}}}e^{-\|u\|^2/4}}$ is the Gaussian kernel and ${h_n\to 0}$ as ${n\to\infty.}$ Such operators can be viewed as graph laplacians (for a weighted graph with vertices at data points) and they have been used in the machine learning literature to approximate the Laplace-Beltrami operator of ${M,}$ ${Δ_Mf}$ (divided by the Riemannian volume of the manifold). We prove several results on a.s. and distributional convergence of the deviations ${Δ_{h_n,n}f(p)-{\frac{1}{|μ|}}Δ_Mf(p)}$ for smooth functions ${f}$ both pointwise and uniformly in ${f}$ and ${p}$ (here ${|μ|=μ(M)}$ and $μ$ is the Riemannian volume measure). In particular, we show that for any class ${\cal F}$ of three times differentiable functions on ${M}$ with uniformly bounded derivatives $$ \sup_{p\in M}\sup_{f\in F}\Big|Δ_{h_n,p}f(p)-\frac{1}{|μ|}Δ_Mf(p)\Big|= O\Big(\sqrt{\frac{\log(1/h_n)}{nh_n^{d+2}}}\Big) a.s. $$ as soon as $$ nh_n^{d+2}/\log h_n^{-1}\to \infty and nh^{d+4}_n/\log h_n^{-1}\to 0, $$ and also prove asymptotic normality of ${Δ_{h_n,p}f(p)-{\frac{1}{|μ|}}Δ_Mf(p)}$ (functional CLT) for a fixed ${p\in M}$ and uniformly in ${f}.$

preprint2006arXiv

High Dimensional Probability

About forty years ago it was realized by several researchers that the essential features of certain objects of Probability theory, notably Gaussian processes and limit theorems, may be better understood if they are considered in settings that do not impose structures extraneous to the problems at hand. For instance, in the case of sample continuity and boundedness of Gaussian processes, the essential feature is the metric or pseudometric structure induced on the index set by the covariance structure of the process, regardless of what the index set may be. This point of view ultimately led to the Fernique-Talagrand majorizing measure characterization of sample boundedness and continuity of Gaussian processes, thus solving an important problem posed by Kolmogorov. Similarly, separable Banach spaces provided a minimal setting for the law of large numbers, the central limit theorem and the law of the iterated logarithm, and this led to the elucidation of the minimal (necessary and/or sufficient) geometric properties of the space under which different forms of these theorems hold. However, in light of renewed interest in Empirical processes, a subject that has considerably influenced modern Statistics, one had to deal with a non-separable Banach space, namely $\mathcal{L}_{\infty}$. With separability discarded, the techniques developed for Gaussian processes and for limit theorems and inequalities in separable Banach spaces, together with combinatorial techniques, led to powerful inequalities and limit theorems for sums of independent bounded processes over general index sets, or, in other words, for general empirical processes.

preprint1999arXiv

The LIL for canonical U-statistics of order 2

Let X,X_1,X_2,... be independent identically distributed random variables and let h(x,y)=h(y,x) be a measurable function of two variables. It is shown that the bounded law of the iterated logarithm, $\limsup_n (n\log\log n)^{-1}|\sum_{1<= i< j<= n}h(X_i,X_j)|<\infty$ a.s., holds if and only if the following three conditions are satisfied: h is canonical for the law of X (that is Eh(X,y)=0 for almost y) and there exists $C<\infty$ such that, both, $E\min(h^2(X_1,X_2),u)<C\log\log u$ for all large u and $sup\{Eh(X_1,X_2)f(X_1)g(X_2):|f(X)|_2<1,\|g(X)\|_2<1, \|f\|_\infty<\infty, \|g\|_\infty<\infty\}< C$.