Source author record

Iain M. Johnstone

Iain M. Johnstone appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2022arXiv

An edge CLT for the log determinant of Wigner ensembles

We derive a Central Limit Theorem (CLT) for $\log \left\vert\det \left( W_{N}-E_{N}\right)\right\vert,$ where $W_{N}$ is a Wigner matrix, and $E_{N}$ is local to the edge of the semi-circle law. Precisely, $E_N=2+N^{-2/3}σ_N$ with $σ_N$ being either a constant (possibly negative), or a sequence of positive real numbers, slowly diverging to infinity so that $σ_N \ll \log^{2} N$. We also extend our CLT to cover spiked Wigner matrices. Our interest in the CLT is motivated by its applications to statistical testing in critically spiked models and to the fluctuations of the free energy in the spherical Sherrington-Kirkpatrick model of statistical physics.

preprint2020arXiv

Tracy-Widom at each edge of real covariance and MANOVA estimators

We study the sample covariance matrix for real-valued data with general population covariance, as well as MANOVA-type covariance estimators in variance components models under null hypotheses of global sphericity. In the limit as matrix dimensions increase proportionally, the asymptotic spectra of such estimators may have multiple disjoint intervals of support, possibly intersecting the negative half line. We show that the distribution of the extremal eigenvalue at each regular edge of the support has a GOE Tracy-Widom limit. Our proof extends a comparison argument of Ji Oon Lee and Kevin Schnelli, replacing a continuous Green function flow by a discrete Lindeberg swapping scheme.

preprint2015arXiv

Exact minimax estimation of the predictive density in sparse Gaussian models

We consider estimating the predictive density under Kullback-Leibler loss in an $\ell_0$ sparse Gaussian sequence model. Explicit expressions of the first order minimax risk along with its exact constant, asymptotically least favorable priors and optimal predictive density estimates are derived. Compared to the sparse recovery results involving point estimation of the normal mean, new decision theoretic phenomena are seen. Suboptimal performance of the class of plug-in density estimates reflects the predictive nature of the problem and optimal strategies need diversification of the future risk. We find that minimax optimal strategies lie outside the Gaussian family but can be constructed with threshold predictive density estimates. Novel minimax techniques involving simultaneous calibration of the sparsity adjustment and the risk diversification mechanisms are used to design optimal predictive density estimates.

preprint2015arXiv

Roy's Largest Root Test Under Rank-One Alternatives

Roy's largest root is a common test statistic in multivariate analysis, statistical signal processing and allied fields. Despite its ubiquity, provision of accurate and tractable approximations to its distribution under the alternative has been a longstanding open problem. Assuming Gaussian observations and a rank one alternative, or concentrated non-centrality, we derive simple yet accurate approximations for the most common low-dimensional settings. These include signal detection in noise, multiple response regression, multivariate analysis of variance and canonical correlation analysis. A small noise perturbation approach, perhaps underused in statistics, leads to simple combinations of standard univariate distributions, such as central and non-central $χ^2$ and $F$. Our results allow approximate power and sample size calculations for Roy's test for rank one effects, which is precisely where it is most powerful.

preprint2014arXiv

Adaptation in a class of linear inverse problems

We consider the linear inverse problem of estimating an unknown signal $f$ from noisy measurements on $Kf$ where the linear operator $K$ admits a wavelet-vaguelette decomposition (WVD). We formulate the problem in the Gaussian sequence model and propose estimation based on complexity penalized regression on a level-by-level basis. We adopt squared error loss and show that the estimator achieves exact rate-adaptive optimality as $f$ varies over a wide range of Besov function classes.

preprint2014arXiv

Joint density of eigenvalues in spiked multivariate models

The classical methods of multivariate analysis are based on the eigenvalues of one or two sample covariance matrices. In many applications of these methods, for example to high dimensional data, it is natural to consider alternative hypotheses which are a low rank departure from the null hypothesis. For rank one alternatives, this note provides a representation for the joint eigenvalue density in terms of a single contour integral. This will be of use for deriving approximate distributions for likelihood ratios and linear statistics used in testing.

preprint2014arXiv

Local Asymptotic Normality of the spectrum of high-dimensional spiked F-ratios

We consider two types of spiked multivariate F distributions: a scaled distribution with the scale matrix equal to a rank-one perturbation of the identity, and a distribution with trivial scale, but rank-one non-centrality. The norm of the rank-one matrix (spike) parameterizes the joint distribution of the eigenvalues of the corresponding F matrix. We show that, for a spike located above a phase transition threshold, the asymptotic behavior of the log ratio of the joint density of the eigenvalues of the F matrix to their joint density under a local deviation from this value depends only on the largest eigenvalue $λ_{1}$. Furthermore, $λ_{1}$ is asymptotically normal, and the statistical experiment of observing all the eigenvalues of the F matrix converges in the Le Cam sense to a Gaussian shift experiment that depends on the asymptotic mean and variance of $λ_{1}$. In particular, the best statistical inference about a sufficiently large spike in the local asymptotic regime is based on the largest eigenvalue only. As a by-product of our analysis, we establish joint asymptotic normality of a few of the largest eigenvalues of the multi-spiked F matrix when the corresponding spikes are above the phase transition threshold.

preprint2012arXiv

Augmented sparse principal component analysis for high dimensional data

We study the problem of estimating the leading eigenvectors of a high-dimensional population covariance matrix based on independent Gaussian observations. We establish lower bounds on the rates of convergence of the estimators of the leading eigenvectors under $l^q$-sparsity constraints when an $l^2$ loss function is used. We also propose an estimator of the leading eigenvectors based on a coordinate selection scheme combined with PCA and show that the proposed estimator achieves the optimal rate of convergence under a sparsity regime. Moreover, we establish that under certain scenarios, the usual PCA achieves the minimax convergence rate.

preprint2012arXiv

Fast approach to the Tracy-Widom law at the edge of GOE and GUE

We study the rate of convergence for the largest eigenvalue distributions in the Gaussian unitary and orthogonal ensembles to their Tracy-Widom limits. We show that one can achieve an $O(N^{-2/3})$ rate with particular choices of the centering and scaling constants. The arguments here also shed light on more complicated cases of Laguerre and Jacobi ensembles, in both unitary and orthogonal versions. Numerical work shows that the suggested constants yield reasonable approximations, even for surprisingly small values of N.

preprint2012arXiv

Minimax bounds for sparse PCA with noisy high-dimensional data

We study the problem of estimating the leading eigenvectors of a high-dimensional population covariance matrix based on independent Gaussian observations. We establish a lower bound on the minimax risk of estimators under the $l_2$ loss, in the joint limit as dimension and sample size increase to infinity, under various models of sparsity for the population eigenvectors. The lower bound on the risk points to the existence of different regimes of sparsity of the eigenvectors. We also propose a new method for estimating the eigenvectors by a two-stage coordinate selection scheme.

preprint2012arXiv

On the within-family Kullback-Leibler risk in Gaussian Predictive models

We consider estimating the predictive density under Kullback-Leibler loss in a high-dimensional Gaussian model. Decision theoretic properties of the within-family prediction error -- the minimal risk among estimates in the class $\mathcal{G}$ of all Gaussian densities are discussed. We show that in sparse models, the class $\mathcal{G}$ is minimax sub-optimal. We produce asymptotically sharp upper and lower bounds on the within-family prediction errors for various subfamilies of $\mathcal{G}$. Under mild regularity conditions, in the sub-family where the covariance structure is represented by a single data dependent parameter $\Shat=\dhat \cdot I$, the Kullback-Leiber risk has a tractable decomposition which can be subsequently minimized to yield optimally flattened predictive density estimates. The optimal predictive risk can be explicitly expressed in terms of the corresponding mean square error of the location estimate, and so, the role of shrinkage in the predictive regime can be determined based on point estimation theory results. Our results demonstrate that some of the decision theoretic parallels between predictive density estimation and point estimation regimes can be explained by second moment based concentration properties of the quadratic loss.

preprint2010arXiv

Approximate null distribution of the largest root in multivariate analysis

The greatest root distribution occurs everywhere in classical multivariate analysis, but even under the null hypothesis the exact distribution has required extensive tables or special purpose software. We describe a simple approximation, based on the Tracy--Widom distribution, that in many cases can be used instead of tables or software, at least for initial screening. The quality of approximation is studied, and its use illustrated in a variety of setttings.