Source author record

Suojin Wang

Suojin Wang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

3works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

3 published item(s)

preprint2022arXiv

The impact of clustering binary data on relative risk towards a study of inferential methods

In epidemiological cohort studies, the relative risk (also known as risk ratio) is a major measure of association to summarize the results of two treatments or exposures. Generally, it measures the relative change in disease risk as a result of treatment application. Standard approaches to estimating relative risk available in common software packages may produce biased inference when applied to correlated binary data collected from longitudinal or clustered studies. In recent years, several methods for estimating the risk ratio for correlated binary data have been published, some of which maintain a well-controlled coverage probability but do not maintain an appropriate interval width or the interval location to measure the balance between distal and mesial noncoverage probabilities accurately or, vice versa. This paper develops efficient and straightforward inference procedures for estimating a confidence interval for risk ratio based on a hybrid method. In general, the hybrid method combines two separate confidence intervals for two single risk rates to form a hybrid confidence interval for their ratio. Additionally, we propose the procedures for constructing a confidence interval for risk ratio that directly extends recently recommended methods for correlated binary data by building on the concepts of the design effect and effective sample sizes typically used in representative sample surveys. In order to investigate the performance of these proposed methods, we conduct an extensive simulation study. To demonstrate the utility of our proposed methods, we present three examples from real-life applications, comparing the side effects of low-dose tricyclic antidepressants with a placebo, the efficacy of the treatment group in a teratological experiment, and the efficiency of the active drugs in curing infection for clinical trials.

preprint2020arXiv

Simultaneous confidence bands for nonparametric regression with partially missing covariates

In this paper, we consider a weighted local linear estimator based on the inverse selection probability for nonparametric regression with missing covariates at random. The asymptotic distribution of the maximal deviation between the estimator and the true regression function is derived and an asymptotically accurate simultaneous confidence band is constructed. The estimator for the regression function is shown to be oracally efficient in the sense that it is uniformly indistinguishable from that when the selection probabilities are known. Finite sample performance is examined via simulation studies which support our asymptotic theory. The proposed method is demonstrated via an analysis of a data set from the Canada 2010/2011 Youth Student Survey.

preprint2015arXiv

Semiparametric GEE analysis in partially linear single-index models for longitudinal data

In this article, we study a partially linear single-index model for longitudinal data under a general framework which includes both the sparse and dense longitudinal data cases. A semiparametric estimation method based on a combination of the local linear smoothing and generalized estimation equations (GEE) is introduced to estimate the two parameter vectors as well as the unknown link function. Under some mild conditions, we derive the asymptotic properties of the proposed parametric and nonparametric estimators in different scenarios, from which we find that the convergence rates and asymptotic variances of the proposed estimators for sparse longitudinal data would be substantially different from those for dense longitudinal data. We also discuss the estimation of the covariance (or weight) matrices involved in the semiparametric GEE method. Furthermore, we provide some numerical studies including Monte Carlo simulation and an empirical application to illustrate our methodology and theory.