Source author record

Geert Molenberghs

Geert Molenberghs appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

6works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

6 published item(s)

preprint2016arXiv

Approximate central limit theorems

We refine the classical Lindeberg-Feller central limit theorem by obtaining asymptotic bounds on the Kolmogorov distance, the Wasserstein distance, and the parametrized Prokhorov distances in terms of a Lindeberg index. We thus obtain more general approximate central limit theorems, which roughly state that the row-wise sums of a triangular array are approximately asymptotically normal if the array approximately satisfies Lindeberg's condition. This allows us to continue to provide information in non-standard settings in which the classical central limit theorem fails to hold. Stein's method plays a key role in the development of this theory.

preprint2015arXiv

A permutational-splitting sample procedure to quantify expert opinion on clusters of chemical compounds using high-dimensional data

Expert opinion plays an important role when selecting promising clusters of chemical compounds in the drug discovery process. We propose a method to quantify these qualitative assessments using hierarchical models. However, with the most commonly available computing resources, the high dimensionality of the vectors of fixed effects and correlated responses renders maximum likelihood unfeasible in this scenario. We devise a reliable procedure to tackle this problem and show, using theoretical arguments and simulations, that the new methodology compares favorably with maximum likelihood, when the latter option is available. The approach was motivated by a case study, which we present and analyze.

preprint2013arXiv

Dynamic Predictions with Time-Dependent Covariates in Survival Analysis using Joint Modeling and Landmarking

A key question in clinical practice is accurate prediction of patient prognosis. To this end, nowadays, physicians have at their disposal a variety of tests and biomarkers to aid them in optimizing medical care. These tests are often performed on a regular basis in order to closely follow the progression of the disease. In this setting it is of medical interest to optimally utilize the recorded information and provide medically-relevant summary measures, such as survival probabilities, that will aid in decision making. In this work we present and compare two statistical techniques that provide dynamically-updated estimates of survival probabilities, namely landmark analysis and joint models for longitudinal and time-to-event data. Special attention is given to the functional form linking the longitudinal and event time processes, and to measures of discrimination and calibration in the context of dynamic prediction.

preprint2013arXiv

The surface nitrogen abundance of a massive star in relation to its oscillations, rotation, and magnetic field

We have composed a sample of 68 massive stars in our galaxy whose projected rotational velocity, effective temperature and gravity are available from high-precision spectroscopic measurements. The additional seven observed variables considered here are their surface nitrogen abundance, rotational frequency, magnetic field strength, and the amplitude and frequency of their dominant acoustic and gravity mode of oscillation. Multiple linear regression to estimate the nitrogen abundance combined with principal components analysis, after addressing the incomplete and truncated nature of the data, reveals that the effective temperature and the frequency of the dominant acoustic oscillation mode are the only two significant predictors for the nitrogen abundance, while the projected rotational velocity and the rotational frequency have no predictive power. The dominant gravity mode and the magnetic field strength are correlated with the effective temperature but have no predictive power for the nitrogen abundance. Our findings are completely based on observations and their proper statistical treatment and call for a new strategy in evaluating the outcome of stellar evolution computations.

preprint2011arXiv

A Family of Generalized Linear Models for Repeated Measures with Normal and Conjugate Random Effects

Non-Gaussian outcomes are often modeled using members of the so-called exponential family. Notorious members are the Bernoulli model for binary data, leading to logistic regression, and the Poisson model for count data, leading to Poisson regression. Two of the main reasons for extending this family are (1) the occurrence of overdispersion, meaning that the variability in the data is not adequately described by the models, which often exhibit a prescribed mean--variance link, and (2) the accommodation of hierarchical structure in the data, stemming from clustering in the data which, in turn, may result from repeatedly measuring the outcome, for various members of the same family, etc. The first issue is dealt with through a variety of overdispersion models, such as, for example, the beta-binomial model for grouped binary data and the negative-binomial model for counts. Clustering is often accommodated through the inclusion of random subject-specific effects. Though not always, one conventionally assumes such random effects to be normally distributed. While both of these phenomena may occur simultaneously, models combining them are uncommon. This paper proposes a broad class of generalized linear models accommodating overdispersion and clustering through two separate sets of random effects. We place particular emphasis on so-called conjugate random effects at the level of the mean for the first aspect and normal random effects embedded within the linear predictor for the second aspect, even though our family is more general. The binary, count and time-to-event cases are given particular emphasis. Apart from model formulation, we present an overview of estimation methods, and then settle for maximum likelihood estimation with analytic--numerical integration. Implications for the derivation of marginal correlations functions are discussed. The methodology is applied to data from a study in epileptic seizures, a clinical trial in toenail infection named onychomycosis and survival data in children with asthma.