Source author record

Abraham Wyner

Abraham Wyner appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

2works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

2 published item(s)

preprint2016arXiv

Yule's "Nonsense Correlation" Solved!

In this paper, we resolve a longstanding open statistical problem. The problem is to mathematically confirm Yule's 1926 empirical finding of "nonsense correlation" (\cite{Yule}). We do so by analytically determining the second moment of the empirical correlation coefficient \beqn θ:= \frac{\int_0^1W_1(t)W_2(t) dt - \int_0^1W_1(t) dt \int_0^1 W_2(t) dt}{\sqrt{\int_0^1 W^2_1(t) dt - \parens{\int_0^1W_1(t) dt}^2} \sqrt{\int_0^1 W^2_2(t) dt - \parens{\int_0^1W_2(t) dt}^2}}, \eeqn of two {\em independent} Wiener processes, $W_1,W_2$. Using tools from Fred- holm integral equation theory, we successfully calculate the second moment of $θ$ to obtain a value for the standard deviation of $θ$ of nearly .5. The "nonsense" correlation, which we call "volatile" correlation, is volatile in the sense that its distribution is heavily dispersed and is frequently large in absolute value. It is induced because each Wiener process is "self-correlated" in time. This is because a Wiener process is an integral of pure noise and thus its values at different time points are correlated. In addition to providing an explicit formula for the second moment of $θ$, we offer implicit formulas for higher moments of $θ$.

preprint2013arXiv

Data Analysis with Bayesian Networks: A Bootstrap Approach

In recent years there has been significant progress in algorithms and methods for inducing Bayesian networks from data. However, in complex data analysis problems, we need to go beyond being satisfied with inducing networks with high scores. We need to provide confidence measures on features of these networks: Is the existence of an edge between two nodes warranted? Is the Markov blanket of a given node robust? Can we say something about the ordering of the variables? We should be able to address these questions, even when the amount of data is not enough to induce a high scoring network. In this paper we propose Efron's Bootstrap as a computationally efficient approach for answering these questions. In addition, we propose to use these confidence measures to induce better structures from the data, and to detect the presence of latent variables.