Source author record

Lancelot F. James

Lancelot F. James appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
3topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2015arXiv

Generalized Mittag Leffler distributions arising as limits in preferential attachment models

For $0<α<1,$ and $θ>-α,$ let $(S^{-α}_{α,θ+r})_{\{r\ge 0\}}$ denote an increasing(decreasing) sequence of variables forming a time inhomogeneous Markov chain whose marginal distributions are equivalent to generalized Mittag Leffler distributions. We exploit the property that such a sequence may be connected with the two parameter $(α,θ)$ family of Poisson Dirichlet distributions. We demonstrate that the sequences serve as limits in certain types of preferential attachment models. As one illustrative application, we describe the explicit joint limiting distribution of scaled degree sequences arising under a class of linear weighted preferential attachment models as treated in Móri (2005), with weight $β>-1.$ When $β=0$ this corresponds to the Barbasi-Albert preferential attachment model. We then construct sequences of nested $(α,θ)$ Chinese restaurant partitions of $[n]$. From this, we identify and analyze relevant quantities that may be thought of as mimics for vectors of degree sequences, or differences in tree lengths. We also describe connections to a wide class of continuous time coalescent processes that can be seen as a variation of stochastic flows of bridges related to generalized Fleming-Viot models. Under a change of measure our results suggest the possibilities for identification of limiting distributions related to consistent families of nested Gibbs partitions of $[n]$ that would otherwise be difficult by methods using moments or Laplace transforms. In this regard, we focus on special simplifications obtained in the case of $α=1/2.$ That is to say, limits derived from a $\mathrm{PD}(1/2|t)$ distribution. Throughout we present some distributional results that are relevant to various settings. We describe nestings across the $α$ parameter in section 6

preprint2015arXiv

Scaled subordinators and generalizations of the Indian buffet process

We study random families of subsets of $\mathbb{N}$ that are similar to exchangeable random partitions, but do not require constituent sets to be disjoint: Each element of ${\mathbb{N}}$ may be contained in multiple subsets. One class of such objects, known as Indian buffet processes, has become a popular tool in machine learning. Based on an equivalence between Indian buffet and scale-invariant Poisson processes, we identify a random scaling variable whose role is similar to that played in exchangeable partition models by the total mass of a random measure. Analogous to the construction of exchangeable partitions from normalized subordinators, random families of sets can be constructed from randomly scaled subordinators. Coupling to a heavy-tailed scaling variable induces a power law on the number of sets containing the first $n$ elements. Several examples, with properties desirable in applications, are derived explicitly. A relationship to exchangeable partitions is made precise as a correspondence between scaled subordinators and Poisson-Kingman measures, generalizing a result of Arratia, Barbour and Tavare on scale-invariant processes.

preprint2014arXiv

Poisson Latent Feature Calculus for Generalized Indian Buffet Processes

The purpose of this work is to describe a unified, and indeed simple, mechanism for non-parametric Bayesian analysis, construction and generative sampling of a large class of latent feature models which one can describe as generalized notions of Indian Buffet Processes(IBP). This is done via the Poisson Process Calculus as it now relates to latent feature models. The IBP was ingeniously devised by Griffiths and Ghahramani in (2005) and its generative scheme is cast in terms of customers entering sequentially an Indian Buffet restaurant and selecting previously sampled dishes as well as new dishes. In this metaphor dishes corresponds to latent features, attributes, preferences shared by individuals. The IBP, and its generalizations, represent an exciting class of models well suited to handle high dimensional statistical problems now common in this information age. The IBP is based on the usage of conditionally independent Bernoulli random variables, coupled with completely random measures acting as Bayesian priors, that are used to create sparse binary matrices. This Bayesian non-parametric view was a key insight due to Thibaux and Jordan (2007). One way to think of generalizations is to to use more general random variables. Of note in the current literature are models employing Poisson and Negative-Binomial random variables. However, unlike their closely related counterparts, generalized Chinese restaurant processes, the ability to analyze IBP models in a systematic and general manner is not yet available. The limitations are both in terms of knowledge about the effects of different priors and in terms of models based on a wider choice of random variables. This work will not only provide a thorough description of the properties of existing models but also provide a simple template to devise and analyze new models.

preprint2013arXiv

Stick-breaking PG(α,ζ)-Generalized Gamma Processes

This work centers around results related to Proposition 21 of Pitman and Yor's (1997) paper on the two parameter Poisson Dirichlet distribution indexed by (α,θ) for 0<α<1, also α=0, and θ>-α, denoted PD(α,θ). We develop explicit stick-breaking representations for a class that contains the PD(α,θ) for the range θ=0, and θ>0, we call PG(α,ζ). We also construct a larger class, EPG(α,ζ), containing the entire range. These classes are indexed by α, and an arbitrary non-negative random variable ζ. The bulk of this work focuses on investigating various properties of this larger class, EPG(α,ζ), which lead to connections to other work in the literature. In particular, we develop completely explicit stick-breaking representations for this entire class via size biased sampling as described in Perman, Pitman and Yor (1992). This represents the first case outside of the PD(α,θ) where one obtains explicit results for the entire range of α, the EPG are within the larger class of mass partitions generated by conditioning on the total mass of an α-stable subordinator. Furthermore Markov chains are derived which establish links between Markov chains derived from stick-breaking(insertion/deletion), as described in Perman, Pitman and Yor and Markov chains derived from successive usage of dual coagulation fragmentation operators described in Bertoin and Goldschmidt and Dong, Goldschmidt and Martin. Which have connections to certain types of fragmentation trees and coalescents appearing in the recent literature. Our results are also suggestive of new models and tools, for applications in Bayesian Nonparametrics/Machine Learning, where PD(α,θ) bridges are often referred to as Pitman-Yor processes.

preprint2011arXiv

Bayesian nonparametric estimation and consistency of mixed multinomial logit choice models

This paper develops nonparametric estimation for discrete choice models based on the mixed multinomial logit (MMNL) model. It has been shown that MMNL models encompass all discrete choice models derived under the assumption of random utility maximization, subject to the identification of an unknown distribution $G$. Noting the mixture model description of the MMNL, we employ a Bayesian nonparametric approach, using nonparametric priors on the unknown mixing distribution $G$, to estimate choice probabilities. We provide an important theoretical support for the use of the proposed methodology by investigating consistency of the posterior distribution for a general nonparametric prior on the mixing distribution. Consistency is defined according to an $L_1$-type distance on the space of choice probabilities and is achieved by extending to a regression model framework a recent approach to strong consistency based on the summability of square roots of prior probabilities. Moving to estimation, slightly different techniques for non-panel and panel data models are discussed. For practical implementation, we describe efficient and relatively easy-to-use blocked Gibbs sampling procedures. These procedures are based on approximations of the random probability measure by classes of finite stick-breaking processes. A simulation study is also performed to investigate the performance of the proposed methods.

preprint2011arXiv

Quantile clocks

Quantile clocks are defined as convolutions of subordinators $L$, with quantile functions of positive random variables. We show that quantile clocks can be chosen to be strictly increasing and continuous and discuss their practical modeling advantages as business activity times in models for asset prices. We show that the marginal distributions of a quantile clock, at each fixed time, equate with the marginal distribution of a single subordinator. Moreover, we show that there are many quantile clocks where one can specify $L$, such that their marginal distributions have a desired law in the class of generalized $s$-self decomposable distributions, and in particular the class of self-decomposable distributions. The development of these results involves elements of distribution theory for specific classes of infinitely divisible random variables and also decompositions of a gamma subordinator, that is of independent interest. As applications, we construct many price models that have continuous trajectories, exhibit volatility clustering and have marginal distributions that are equivalent to those of quite general exponential Lévy price models. In particular, we provide explicit details for continuous processes whose marginals equate with the popular VG, CGMY and NIG price models. We also show how to perfectly sample the marginal distributions of more general classes of convoluted subordinators when $L$ is in a sub-class of generalized gamma convolutions, which is relevant for pricing of European style options.

preprint2010arXiv

Coag-Frag duality for a class of stable Poisson-Kingman mixtures

Exchangeable sequences of random probability measures (partitions of mass) and their corresponding exchangeable bridges play an important role in a variety of areas in probability, statistics and related areas, including Bayesian statistics, physics, finance and machine learning. An area of theoretical as well as practical interest, is the study of coagulation and fragmentation operators on partitions of mass. In this regard, an interesting but formidable question is the identification of operators and distributional families on mass partitions that exhibit interesting duality relations. In this paper we identify duality relations for a large sub-class of mixed Poisson-Kingman models generated by a stable subordinator. Our results are natural generalizations of the duality relations developed in Pitman, Bertoin and Goldschmidt, and Dong, Goldschmidt and Martin for the two-parameter Poisson Dirichlet family. These results are deduced from results for corresponding bridges.

preprint2010arXiv

Dirichlet mean identities and laws of a class of subordinators

An interesting line of research is the investigation of the laws of random variables known as Dirichlet means. However, there is not much information on interrelationships between different Dirichlet means. Here, we introduce two distributional operations, one of which consists of multiplying a mean functional by an independent beta random variable, the other being an operation involving an exponential change of measure. These operations identify relationships between different means and their densities. This allows one to use the often considerable analytic work on obtaining results for one Dirichlet mean to obtain results for an entire family of otherwise seemingly unrelated Dirichlet means. Additionally, it allows one to obtain explicit densities for the related class of random variables that have generalized gamma convolution distributions and the finite-dimensional distribution of their associated Lévy processes. The importance of this latter statement is that Lévy processes now commonly appear in a variety of applications in probability and statistics, but there are relatively few cases where the relevant densities have been described explicitly. We demonstrate how the technique allows one to obtain the finite-dimensional distribution of several interesting subordinators which have recently appeared in the literature.

preprint2010arXiv

Lamperti-type laws

This paper explores various distributional aspects of random variables defined as the ratio of two independent positive random variables where one variable has an $α$-stable law, for $0<α<1$, and the other variable has the law defined by polynomially tilting the density of an $α$-stable random variable by a factor $θ>-α$. When $θ=0$, these variables equate with the ratio investigated by Lamperti [Trans. Amer. Math. Soc. 88 (1958) 380--387] which, remarkably, was shown to have a simple density. This variable arises in a variety of areas and gains importance from a close connection to the stable laws. This rationale, and connection to the $\operatorname {PD}(α,θ)$ distribution, motivates the investigations of its generalizations which we refer to as Lamperti-type laws. We identify and exploit links to random variables that commonly appear in a variety of applications. Namely Linnik, generalized Pareto and $z$-distributions. In each case we obtain new results that are of potential interest. As some highlights, we then use these results to (i) obtain integral representations and other identities for a class of generalized Mittag--Leffler functions, (ii) identify explicitly the Lévy density of the semigroup of stable continuous state branching processes (CSBP) and hence corresponding limiting distributions derived in Slack and in Zolotarev [Z. Wahrsch. Verw. Gebiete 9 (1968) 139--145, Teor. Veroyatn. Primen. 2 (1957) 256--266], which are related to the recent work by Berestycki, Berestycki and Schweinsberg, and Bertoin and LeGall [Ann. Inst. H. Poincaré Probab. Statist. 44 (2008) 214--238, Illinois J. Math. 50 (2006) 147--181] on beta coalescents. (iii) We obtain explicit results for the occupation time of generalized Bessel bridges and some interesting stochastic equations for $\operatorname {PD}(α,θ)$-bridges. In particular we obtain the best known results for the density of the time spent positive of a Bessel bridge of dimension $2-2α$.

preprint2010arXiv

On the posterior distribution of classes of random means

The study of properties of mean functionals of random probability measures is an important area of research in the theory of Bayesian nonparametric statistics. Many results are now known for random Dirichlet means, but little is known, especially in terms of posterior distributions, for classes of priors beyond the Dirichlet process. In this paper, we consider normalized random measures with independent increments (NRMI's) and mixtures of NRMI. In both cases, we are able to provide exact expressions for the posterior distribution of their means. These general results are then specialized, leading to distributional results for means of two important particular cases of NRMI's and also of the two-parameter Poisson--Dirichlet process.

preprint2007arXiv

New Dirichlet Mean Identities

An important line of research is the investigation of the laws of random variables known as Dirichlet means as discussed in Cifarelli and Regazzini(1990). However there is not much information on inter-relationships between different Dirichlet means. Here we introduce two distributional operations, which consist of multiplying a mean functional by an independent beta random variable and an operation involving an exponential change of measure. These operations identify relationships between different means and their densities. This allows one to use the often considerable analytic work to obtain results for one Dirichlet mean to obtain results for an entire family of otherwise seemingly unrelated Dirichlet means. Additionally, it allows one to obtain explicit densities for the related class of random variables that have generalized gamma convolution distributions, and the finite-dimensional distribution of their associated Lévy processes. This has implications in, for instance, the explicit description of Bayesian nonparametric prior and posterior models, and more generally in a variety of applications in probability and statistics involving Levy processes.