Source author record

Piet Groeneboom

Piet Groeneboom appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

20works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

20 published item(s)

preprint2023arXiv

Estimation of the incubation time distribution in the singly and doubly interval censored model

We analyze nonparametric estimators for the distribution function of the incubation time in the singly and doubly interval censoring model. The classical approach is to use parametric families like Weibull, log-normal or gamma distributions in the estimation procedure. We propose nonparametric estimates which stay closer to the data than the classical parametric methods. We also give explicit limit distributions for discrete versions of the models and apply this to compute confidence intervals. The methods complement the analysis of the continuous model. R scripts for computation of the estimates are provided on https://github.com/pietg/incubationtime.

preprint2015arXiv

Nonparametric confidence intervals for monotone functions

We study nonparametric isotonic confidence intervals for monotone functions. In Banerjee and Wellner (2001) pointwise confidence intervals, based on likelihood ratio tests for the restricted and unrestricted MLE in the current status model, are introduced. We extend the method to the treatment of other models with monotone functions, and demonstrate our method by a new proof of the results in Banerjee and Wellner (2001) and also by constructing confidence intervals for monotone densities, for which still theory had to be developed. For the latter model we prove that the limit distribution of the LR test under the null hypothesis is the same as in the current status model. We compare the confidence intervals, so obtained, with confidence intervals using the smoothed maximum likelihood estimator (SMLE), using bootstrap methods. The `Lagrange-modified' cusum diagrams, developed here, are an essential tool both for the computation of the restricted MLEs and for the development of the theory for the confidence intervals, based on the LR tests.

preprint2014arXiv

Maximum smoothed likelihood estimators for the interval censoring model

We study the maximum smoothed likelihood estimator (MSLE) for interval censoring, case 2, in the so-called separated case. Characterizations in terms of convex duality conditions are given and strong consistency is proved. Moreover, we show that, under smoothness conditions on the underlying distributions and using the usual bandwidth choice in density estimation, the local convergence rate is $n^{-2/5}$ and the limit distribution is normal, in contrast with the rate $n^{-1/3}$ of the ordinary maximum likelihood estimator.

preprint2013arXiv

Chernoff's distribution and differential equations of parabolic and Airy type

We give a direct derivation of the distribution of the maximum and the location of the maximum of one-sided and two-sided Brownian motion with a negative parabolic drift. The argument uses a relation between integrals of special functions, in particular involving integrals with respect to functions which can be called "incomplete Scorer functions". The relation is proved by showing that both integrals, as a function of two parameters, satisfy the same extended heat equation, and the maximum principle is used to show that these solution must therefore have the stated relation. Once this relation is established, a direct derivation of the distribution of the maximum and location of the maximum of Brownian motion minus a parabola is possible, leading to a considerable shortening of the original proofs.

preprint2013arXiv

Likelihood ratio type two-sample tests for current status data

We introduce fully nonparametric two-sample tests for testing the null hypothesis that the samples come from the same distribution if the values are only indirectly given via current status censoring. The tests are based on the likelihood ratio principle and allow the observation distributions to be different for the two samples, in contrast with earlier proposals for this situation. A bootstrap method is given for determining critical values and asymptotic theory is developed. A simulation study, using Weibull distributions, is presented to compare the power behavior of the tests with the power of other nonparametric tests in this situation.

preprint2013arXiv

Testing equality of functions under monotonicity constraints

We consider the problem of testing equality of functions $f_j:[0,1]\to \mathbb{R}$ for $j=1,2,...,J$ the basis of $J$ independent samples from possibly different distributions under the assumption that the functions are monotone. We provide a uniform approach that covers testing equality of monotone regression curves, equality of monotone densities and equality of monotone hazards in the random censorship model. Two test statistics are proposed based on $L_1$-distances. We show that both statistics are asymptotically normal and we provide bootstrap implementations, which are shown to have critical regions with asymptotic level $α$.

preprint2013arXiv

The bivariate current status model

For the univariate current status and, more generally, the interval censoring model, distribution theory has been developed for the maximum likelihood estimator (MLE) and smoothed maximum likelihood estimator (SMLE) of the unknown distribution function, see, e.g., [12], [7], [4], [5], [6], [10], [11] and [8]. For the bivariate current status and interval censoring models distribution theory of this type is still absent and even the rate at which we can expect reasonable estimators to converge is unknown. We define a purely discrete plug-in estimator of the distribution function which locally converges at rate n^{1/3} and derive its (normal) limit distribution. Unlike the MLE or SMLE, this estimator is not a proper distribution function. Since the estimator is purely discrete, it demonstrates that the n^{1/3} convergence rate is in principle possible for the MLE, but whether this actually holds for the MLE is still an open problem. If the cube root n rate holds for the MLE, this would mean that the local 1-dimensional rate of the MLE continues to hold in dimension 2, a (perhaps) somewhat surprising result. The simulation results do not seem to be in contradiction with this assumption, however. We compare the behavior of the plug-in estimator with the behavior of the MLE on a sieve and the SMLE in a simulation study. This indicates that the plug-in estimator and the SMLE have a smaller variance but a larger bias than the sieved MLE. The SMLE is conjectured to have a n^{1/3}-rate of convergence if we use bandwidths of order n^{-1/6}. We derive its (normal) limit distribution, using this assumption. Finally, we demonstrate the behavior of the MLE and SMLE for the bivariate interval censored data of [1], which have been discussed by many authors, see e.g., [18], [3], [2] and [15].

preprint2013arXiv

Vertices of the least concave majorant of Brownian motion with parabolic drift

It was shown in Groeneboom (1983) that the least concave majorant of one-sided Brownian motion without drift can be characterized by a jump process with independent increments, which is the inverse of the process of slopes of the least concave majorant. This result can be used to prove the result of Sparre Andersen (1954) that the number of vertices of the smallest concave majorant of the empirical distribution function of a sample of size n from the uniform distribution on [0,1] is asymptotically normal, with an asymptotic expectation and variance which are both of order log n. A similar (Markovian) inverse jump process was introduced in Groeneboom (1989), in an analysis of the least concave majorant of two-sided Brownian motion with a parabolic drift. This process is quite different from the process for one-sided Brownian motion without drift: the number of vertices in a (corresponding slopes) interval has an expectation proportional to the length of the interval and the variance of the number of vertices in such an interval is about half the size of the expectation, if the length of the interval tends to infinity. We prove an asymptotic normality result for the number of vertices in an increasing interval, which translates into a corresponding result for the least concave majorant of an empirical distribution function of a sample of size n, generated by a strictly concave distribution function. In this case the number of vertices is of order cube root n, and the variance is again about half the size of the asymptotic expectation. As a side result we obtain some interesting relations between the first moments of the number of vertices, the square of the location of the maximum of Brownian motion minus a parabola, the value of the maximum itself, the squared slope of the least concave majorant at zero, and the value of the least concave majorant at zero.

preprint2011arXiv

A maximum smoothed likelihood estimator in the current status continuous mark model

We consider the problem of estimating the joint distribution function of the event time and a continuous mark variable based on censored data. More specifically, the event time is subject to current status censoring and the continuous mark is only observed in case inspection takes place after the event time. The nonparametric maximum likelihood estimator (MLE) in this model is known to be inconsistent. We propose and study an alternative likelihood based estimator, maximizing a smoothed log-likelihood, hence called a maximum smoothed likelihood estimator (MSLE). This estimator is shown to be well defined and consistent, and a simple algorithm is described that can be used to compute it. The MSLE is compared with other estimators in a small simulation study.

preprint2011arXiv

Convex hulls of uniform samples from a convex polygon

In Groeneboom (1988) a central limit theorem for the number of vertices of the convex hull of a uniform sample from the interior of convex polygon was derived. In the unpublished preprint Nagaev and Khamdamov (1991) (in Russian) a central limit result for the joint distribution of the number of vertices and the remaining area is given, using a coupling of the sample process near the border of the polygon with a Poisson point process as in Groeneboom (1988), and representing the remaining area in the Poisson approximation as a union of a doubly infinite sequence of independent standard exponential random variables. We derive this representation from the representation in Groeneboom (1988) and also prove the central limit result of Nagaev and Khamdamov (1991), using this representation. The relation between the variances of the asymptotic normal distributions of number of vertices and the area, established in Nagaev and Khamdamov (1991), corresponds to a relation between the actual sample variances of the number of vertices and the remaining area in Buchta (2005). We show how these asymptotic results all follow from one simple guiding principle. This corrects at the same time the scaling constants in Nagaev (1995) and Cabo and Groeneboom (1994).

preprint2011arXiv

Isotonic L_2-projection test for local monotonicity of a hazard

We introduce a new test statistic for testing the null hypothesis that the sampling distribution has an increasing hazard rate on a specified interval [0,a]. It is based on a comparison of the empirical distribution function with an isotonic estimate, using the restriction that the hazard is increasing, and measures the excursions of the empirical distribution above the isotonic estimate, due to local non-monotonicity. It is proved in the companion paper Groeneboom and Jongbloed (2011a) that the test statistic is asymptotically normal if the hazard is strictly increasing on the interval [0,a] and certain regularity conditions are satisfied. We discuss a bootstrap method for computing the critical values and compare the test, thus obtained, with other proposals in a simulation study.

preprint2011arXiv

Smooth and non-smooth estimates of a monotone hazard

We discuss a number of estimates of the hazard under the assumption that the hazard is monotone on an interval [0,a]. The usual isotonic least squares estimators of the hazard are inconsistent at the boundary points 0 and a. We use penalization to obtain uniformly consistent estimators. Moreover, we determine the optimal penalization constants, extending related work in this direction by Woodroofe and Sun (1993) and Woodroofe and Sun (1999). Two methods of obtaining smooth monotone estimates based on a non-smooth monotone estimator are discussed. One is based on kernel smoothing, the other on penalization.

preprint2011arXiv

Smooth plug-in inverse estimators in the current status continuous mark model

We consider the problem of estimating the joint distribution function of the event time and a continuous mark variable when the event time is subject to interval censoring case 1 and the continuous mark variable is only observed in case the event occurred before the time of inspection. The nonparametric maximum likelihood estimator in this model is known to be inconsistent. We study two alternative smooth estimators, based on the explicit (inverse) expression of the distribution function of interest in terms of the density of the observable vector. We derive the pointwise asymptotic distribution of both estimators.

preprint2011arXiv

Testing monotonicity of a hazard: asymptotic distribution theory

Two new test statistics are introduced to test the null hypotheses that the sampling distribution has an increasing hazard rate on a specified interval [0,a]. These statistics are empirical L_1-type distances between the isotonic estimates, which use the monotonicity constraint, and either the empirical distribution function or the empirical cumulative hazard. They measure the excursions of the empirical estimates with respect to the isotonic estimates, due to local non-monotonicity. Asymptotic normality of the test statistics, if the hazard is strictly increasing on [0,a], is established under mild conditions. This is done by first approximating the global empirical distance by an distance with respect to the underlying distribution function. The resulting integral is treated as sum of increasingly many local integrals to which a CLT can be applied. The behavior of the local integrals is determined by a canonical process: the difference between the stochastic process x -> W(x)+x^2 where W is standard two-sided Brownian Motion, and its greatest convex minorant.

preprint2011arXiv

The remaining area of the convex hull of a Poisson process

In Cabo and Groeneboom (1994) the remaining area of the left-lower convex hull of a Poisson point process with intensity one in the first quadrant of the plane was analyzed, using the methods of Groeneboom (1988), giving formulas for the expectation and variance of the remaining area for a finite interval of slopes of the boundary of the convex hull. However, the time inversion argument of Groeneboom (1988) was not correctly applied in Cabo and Groeneboom (1994), leading to an incorrect scaling constant for the variance. The purpose of this note is to show how the correct application of the time inversion argument gives the right expression, which is in accordance with results in Nagaev and Khamdamov (1991) and Buchta (2003).

preprint2011arXiv

The tail of the maximum of Brownian motion minus a parabola

We analyze the tail behavior of the maximum N of Brownian motion minus a parabola and give an asymptotic expansion for P(N>x) as x tends to infinity. This extends a first order result on the tail behavior, which can be deduced from Huesler and Piterbarg (1999). We also point out the relation between certain results in Groeneboom (2010) and Janson, Louchard and Martin-Löf (2010).

preprint2010arXiv

Maximum smoothed likelihood estimation and smoothed maximum likelihood estimation in the current status model

We consider the problem of estimating the distribution function, the density and the hazard rate of the (unobservable) event time in the current status model. A well studied and natural nonparametric estimator for the distribution function in this model is the nonparametric maximum likelihood estimator (MLE). We study two alternative methods for the estimation of the distribution function, assuming some smoothness of the event time distribution. The first estimator is based on a maximum smoothed likelihood approach. The second method is based on smoothing the (discrete) MLE of the distribution function. These estimators can be used to estimate the density and hazard rate of the event time distribution based on the plug-in principle.