Source author record

Yvik Swan

Yvik Swan appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

26works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

26 published item(s)

preprint2022arXiv

A note on one-dimensional Poincaré inequalities by Stein-type integration

We study the weighted Poincaré constant $C(p,w)$ of a probability density $p$ with weight function $w$ using integration methods inspired by Stein's method. We obtain a new version of the Chen-Wang variational formula which, as a byproduct, yields simple upper and lower bounds on $C(p,w)$ in terms of the so-called Stein kernel of $p$. We also iterate these variational formulas so as to build sequences of nested intervals containing the Poincaré constant, sequences of functions converging to said constant, as well as sequences of functions converging to the solutions of the corresponding spectral problem. Our results rely on the properties of a pseudo inverse operator of the classical Sturm-Liouville operator. We illustrate our methods on a variety of examples: Gaussian functionals, weighted Gaussian, beta, gamma, Subbotin, and Weibull distributions.

preprint2022arXiv

Stein's Method Meets Computational Statistics: A Review of Some Recent Developments

Stein's method compares probability distributions through the study of a class of linear operators called Stein operators. While mainly studied in probability and used to underpin theoretical statistics, Stein's method has led to significant advances in computational statistics in recent years. The goal of this survey is to bring together some of these recent developments and, in doing so, to stimulate further research into the successful field of Stein's method and statistics. The topics we discuss include tools to benchmark and compare sampling methods such as approximate Markov chain Monte Carlo, deterministic alternatives to sampling methods, control variate techniques, parameter estimation and goodness-of-fit testing.

preprint2016arXiv

On the rate of convergence in de Finetti's representation theorem

A consequence of de Finetti's representation theorem is that for every infinite sequence of exchangeable 0-1 random variables $(X_k)_{k\geq1}$, there exists a probability measure $μ$ on the Borel sets of $[0,1]$ such that $\bar X_n = n^{-1} \sum_{i=1}^n X_i$ converges weakly to $μ$. For a wide class of probability measures $μ$ having smooth density on $(0,1)$, we give bounds of order $1/n$ with explicit constants for the Wasserstein distance between the law of $\bar X_n$ and $μ$. This extends a recent result {by} Goldstein and Reinert \cite{goldstein2013stein} regarding the distance between the scaled number of white balls drawn in a Pólya-Eggenberger urn and its limiting distribution. We prove also that, in the most general cases, the distance between the law of $\bar X_n$ and $μ$ is bounded below by $1/n$ and above by $1/\sqrt{n}$ (up to some multiplicative constants). For every $δ\in [1/2,1]$, we give an example of an exchangeable sequence such that this distance is of order $1/n^δ$.

preprint2016arXiv

One step futher: an explicit solution to Robbins' problem when $n=4$

Fix some $n \in \mathbb{N}$ and let $X_1, X_2,\dots, X_n$ be independent random variables drawn from the uniform distribution on $[0,1]$. A decision maker is shown the variables sequentially and, after each observation, must decide whether or not to keep the current one, with payoff the overall rank of the selected observation. Decisions are final: no recall is allowed, no regret is tolerated. The objective is to act in such a way as to minimise the expected payoff. In this note we give the explicit solution to this problem, known as Robbins' problem of optimal stopping, when $n=4$.

preprint2016arXiv

Stein's method for comparison of univariate distributions

We propose a new general version of Stein's method for univariate distributions. In particular we propose a canonical definition of the Stein operator of a probability distribution {which is based on a linear difference or differential-type operator}. The resulting Stein identity highlights the unifying theme behind the literature on Stein's method (both for continuous and discrete distributions). Viewing the Stein operator as an operator acting on pairs of functions, we provide an extensive toolkit for distributional comparisons. Several abstract approximation theorems are provided. Our approach is illustrated for comparison of several pairs of distributions : normal vs normal, sums of independent Rademacher vs normal, normal vs Student, and maximum of random variables vs exponential, Frechet and Gumbel.

preprint2016arXiv

Stein's method on the second Wiener chaos : 2-Wasserstein distance

In the first part of the paper we use a new Fourier technique to obtain a Stein characterizations for random variables in the second Wiener chaos. We provide the connection between this result and similar conclusions that can be derived using Malliavin calculus. We also introduce a new form of discrepancy which we use, in the second part of the paper, to provide bounds on the 2-Wasserstein distance between linear combinations of independent centered random variables. Our method of proof is entirely original. In particular it does not rely on estimation of bounds on solutions of the so-called Stein equations at the heart of Stein's method. We provide several applications, and discuss comparison with recent similar results on the same topic.

preprint2016arXiv

Stein's method, many interacting worlds and quantum mechanics

Hall, Deckert and Wiseman (2014) recently proposed that quantum theory can be understood as the continuum limit of a deterministic theory in which there is a large, but finite, number of classical "worlds." A resulting Gaussian limit theorem for particle positions in the ground state, agreeing with quantum theory, was conjectured in Hall, Deckert and Wiseman (2014) and proven by McKeague and Levin (2016) using Stein's method. In this article we propose new connections between Stein's method and Many Interacting Worlds (MIW) theory. In particular, we show that quantum position probability densities for higher energy levels beyond the ground state arise as distributional fixed points in a new generalization of Stein's method. These are then used to obtain a rate of distributional convergence for conjectured particle positions in the first energy level above the ground state to the (two-sided) Maxwell distribution; new techniques must be developed for this setting where the usual "density approach" Stein solution (see Chatterjee and Shao (2011)) has a singularity.

preprint2015arXiv

Distances between nested densities and a measure of the impact of the prior in Bayesian statistics

In this paper we propose tight upper and lower bounds for the Wasserstein distance between any two {univariate continuous distributions} with probability densities $p_1$ and $p_2$ having nested supports. These explicit bounds are expressed in terms of the derivative of the likelihood ratio $p_1/p_2$ as well as the Stein kernel $τ_1$ of $p_1$. The method of proof relies on a new variant of Stein's method which manipulates Stein operators. We give several applications of these bounds. Our main application is in Bayesian statistics : we derive explicit data-driven bounds on the Wasserstein distance between the posterior distribution based on a given prior and the no-prior posterior based uniquely on the sampling distribution. This is the first finite sample result confirming the well-known fact that with well-identified parameters and large sample sizes, reasonable choices of prior distributions will have only minor effects on posterior inferences if the data are benign.

preprint2014arXiv

Integration by parts and representation of information functionals

We introduce a new formalism for computing expectations of functionals of arbitrary random vectors, by using generalised integration by parts formulae. In doing so we extend recent representation formulae for the score function introduced in Nourdin, Peccati and Swan (JFA, to appear) and also provide a new proof of a central identity first discovered in Guo, Shamai, and Verd{ú} (IEEE Trans. Information Theory, 2005). We derive a representation for the standardized Fisher information of sums of i.i.d. random vectors which use our identities to provide rates of convergence in information theoretic central limit theorems (both in Fisher information distance and in relative entropy).

preprint2014arXiv

Maximum likelihood characterization of distributions

A famous characterization theorem due to C.F. Gauss states that the maximum likelihood estimator (MLE) of the parameter in a location family is the sample mean for all samples of all sample sizes if and only if the family is Gaussian. There exist many extensions of this result in diverse directions, most of them focussing on location and scale families. In this paper, we propose a unified treatment of this literature by providing general MLE characterization theorems for one-parameter group families (with particular attention on location and scale parameters). In doing so, we provide tools for determining whether or not a given such family is MLE-characterizable, and, in case it is, we define the fundamental concept of minimal necessary sample size at which a given characterization holds. Many of the cornerstone references on this topic are retrieved and discussed in the light of our findings, and several new characterization theorems are provided. Of particular interest is that one part of our work, namely the introduction of so-called equivalence classes for MLE characterizations, is a modernized version of Daniel Bernoulli's viewpoint on maximum likelihood estimation.

preprint2013arXiv

Entropy and the fourth moment phenomenon

We develop a new method for bounding the relative entropy of a random vector in terms of its Stein factors. Our approach is based on a novel representation for the score function of smoothly perturbed random variables, as well as on the de Bruijn's identity of information theory. When applied to sequences of functionals of a general Gaussian field, our results can be combined with the Carbery-Wright inequality in order to yield multidimensional entropic rates of convergence that coincide, up to a logarithmic factor, with those achievable in smooth distances (such as the 1-Wasserstein distance). In particular, our findings settle the open problem of proving a quantitative version of the multidimensional fourth moment theorem for random vectors having chaotic components, with explicit rates of convergence in total variation that are independent of the order of the associated Wiener chaoses. The results proved in the present paper are outside the scope of other existing techniques, such as for instance the multidimensional Stein's method for normal approximations.

preprint2013arXiv

Local Pinsker inequalities via Stein's discrete density approach

Pinsker's inequality states that the relative entropy $d_{\mathrm{KL}}(X, Y)$ between two random variables $X$ and $Y$ dominates the square of the total variation distance $d_{\mathrm{TV}}(X,Y)$ between $X$ and $Y$. In this paper we introduce generalized Fisher information distances $\mathcal{J}(X, Y)$ between discrete distributions $X$ and $Y$ and prove that these also dominate the square of the total variation distance. To this end we introduce a general discrete Stein operator for which we prove a useful covariance identity. We illustrate our approach with several examples. Whenever competitor inequalities are available in the literature, the constants in ours are at least as good, and, in several cases, better.

preprint2013arXiv

On Hodges and Lehmann's "$6/π$ result"

While the asymptotic relative efficiency (ARE) of Wilcoxon rank-based tests for location and regression with respect to their parametric Student competitors can be arbitrarily large, Hodges and Lehmann (1961) have shown that the ARE of the same Wilcoxon tests with respect to their van der Waerden or normal-score counterparts is bounded from above by $6/π\approx 1.910$. In this paper, we revisit that result, and investigate similar bounds for statistics based on Student scores. We also consider the serial version of this ARE. More precisely, we study the ARE, under various densities, of the Spearman-Wald-Wolfowitz and Kendall rank-based autocorrelations with respect to the van der Waerden or normal-score ones used to test (ARMA) serial dependence alternatives.

preprint2013arXiv

Parametric Stein operators and variance bounds

Stein operators are differential operators which arise within the so-called Stein's method for stochastic approximation. We propose a new mechanism for constructing such operators for arbitrary (continuous or discrete) parametric distributions with continuous dependence on the parameter. We provide explicit general expressions for location, scale and skewness families. We also provide a general expression for discrete distributions. For specific choices of target distributions (including the Gaussian, Gamma and Poisson) we compare the operators hereby obtained with those provided by the classical approaches from the literature on Stein's method. We use properties of our operators to provide upper and lower variance bounds (only lower bounds in the discrete case) on functionals $h(X)$ of random variables $X$ following parametric distributions. These bounds are expressed in terms of the first two moments of the derivatives (or differences) of $h$. We provide general variance bounds for location, scale and skewness families and apply our bounds to specific examples (namely the Gaussian, exponential, Gamma and Poisson distributions). The results obtained via our techniques are systematically competitive with, and sometimes improve on, the best bounds available in the literature.

preprint2013arXiv

Stein's density approach and information inequalities

We provide a new perspective on Stein's so-called density approach by introducing a new operator and characterizing class which are valid for a much wider family of probability distributions on the real line. We prove an elementary factorization property of this operator and propose a new Stein identity which we use to derive information inequalities in terms of what we call the \emph{generalized Fisher information distance}. We provide explicit bounds on the constants appearing in these inequalities for several important cases. We conclude with a comparison between our results and known results in the Gaussian case, hereby improving on several known inequalities from the literature.

preprint2012arXiv

Efficient ANOVA for directional data

In this paper we tackle the ANOVA problem for directional data (with particular emphasis on geological data) by having recourse to the Le Cam methodology usually reserved for linear multivariate analysis. We construct locally and asymptotically most stringent parametric tests for ANOVA for directional data within the class of rotationally symmetric distributions. We turn these parametric tests into semi-parametric ones by (i) using a studentization argument (which leads to what we call pseudo-FvML tests) and by (ii) resorting to the invariance principle (which leads to efficient rank-based tests). Within each construction the semi-parametric tests inherit optimality under a given distribution (the FvML distribution in the first case, any rotationally symmetric distribution in the second) from their parametric antecedents and also improve on the latter by being valid under the whole class of rotationally symmetric distributions. Asymptotic relative efficiencies are calculated and the finite-sample behavior of the proposed tests is investigated by means of a Monte Carlo simulation. We conclude by applying our findings on a real-data example involving geological data.

preprint2012arXiv

One-Step R-Estimation in Linear Models with Stable Errors

Classical estimation techniques for linear models either are inconsistent, or perform rather poorly, under $α$-stable error densities; most of them are not even rate-optimal. In this paper, we propose an original one-step R-estimation method and investigate its asymptotic performances under stable densities. Contrary to traditional least squares, the proposed R-estimators remain root-$n$ consistent (the optimal rate) under the whole family of stable distributions, irrespective of their asymmetry and tail index. While parametric stable-likelihood estimation, due to the absence of a closed form for stable densities, is quite cumbersome, our method allows us to construct estimators reaching the parametric efficiency bounds associated with any prescribed values $(α_0, \ b_0)$ of the tail index $α$ and skewness parameter $b$, while preserving root-$n$ consistency under any $(α, \ b)$ as well as under usual light-tailed densities. The method furthermore avoids all forms of multidimensional argmin computation. Simulations confirm its excellent finite-sample performances.

preprint2012arXiv

Optimal R-Estimation of a Spherical Location

In this paper, we provide $R$-estimators of the location of a rotationally symmetric distribution on the unit sphere of $\R^k$. In order to do so we first prove the local asymptotic normality property of a sequence of rotationally symmetric models; this is a non standard result due to the curved nature of the unit sphere. We then construct our estimators by adapting the Le Cam one-step methodology to spherical statistics and ranks. We show that they are asymptotically normal under any rotationally symmetric distribution and achieve the efficiency bound under a specific density. Their small sample behavior is studied via a Monte Carlo simulation and our methodology is illustrated on geological data.

preprint2011arXiv

A note on the normal approximation error for randomly weighted self-normalized sums

Let $\bX=\{X_n\}_{n\geq 1}$ and $\bY=\{Y_n\}_{n\geq 1}$ be two independent random sequences. We obtain rates of convergence to the normal law of randomly weighted self-normalized sums $$ ψ_n(\bX,\bY)=\sum_{i=1}^nX_iY_i/V_n,\quad V_n=\sqrt{Y_1^2+...+Y_n^2}. $$ These rates are seen to hold for the convergence of a number of important statistics, such as for instance Student's $t$-statistic or the empirical correlation coefficient.

preprint2011arXiv

A Stochastic Analysis of Table Tennis

We establish a general formula for the distribution of the score in table tennis. We use this formula to derive the probability distribution (and hence the expectation and variance) of the number of rallies necessary to achieve any given score. We use these findings to investigate the dependence of these quantities on the different parameters involved (number of points needed to win a set, number of consecutive serves, etc.), with particular focus on the rule change imposed in 2001 by the International Table Tennis Federation (ITTF). Finally we briefly indicate how our results can lead to more efficient estimation techniques of individual players' abilities.

preprint2011arXiv

A unified approach to Stein characterizations

This article deals with Stein characterizations of probability distributions. We provide a general framework for interpreting these in terms of the parameters of the underlying distribution. In order to do so we introduce two concepts (a class of functions and an operator) which generalize those which were developed in the 70's by Charles Stein and Louis Chen for characterizing the Gaussian and the Poisson distributions. Our methodology (i) allows for writing many (if not all) known univariate Stein characterizations, (ii) permits to identify clearly minimal conditions under which these results hold and (iii) provides a straightforward tool for constructing new Stein characterizations. Our parametric interpretation of Stein characterizations also raises a number of questions which we outline at the end of the paper.

preprint2011arXiv

On a connection between Stein characterizations and Fisher information

We generalize the so-called density approach to Stein characterizations of probability distributions. We prove an elementary factorization property of the resulting Stein operator in terms of a generalized (standardized) score function. We use this result to connect Stein characterizations with information distances such as the generalized (standardized) Fisher information.

preprint2010arXiv

A Stochastic Analysis of some Two-Person Sports

We consider two-person sports where each rally is initiated by a \emph{server}, the other player (the \emph{receiver}) becoming the server when he/she wins a rally. Historically, these sports used a scoring based on the \emph{side-out scoring system}, in which points are only scored by the server. Recently, however, some federations have switched to the \emph{rally-point scoring system} in which a point is scored on every rally. As various authors before us, we study how much this change affects the game. Our approach is based on a \emph{rally-level analysis} of the process through which, besides the well-known probability distribution of the scores, we also obtain the distribution of the number of rallies. This yields a comprehensive knowledge of the process at hand, and allows for an in-depth comparison of both scoring systems. In particular, our results {help} to explain why the transition from one scoring system to the other has more important implications than those predicted from game-winning probabilities alone. Some of our findings are quite surprising, and unattainable through Monte Carlo experiments. Our results are of high practical relevance to international federations and local tournament organizers alike, and also open the way to efficient estimation of the rally-winning probabilities, which should have a significant impact on the quality of ranking procedures.