Source author record

Tiandong Wang

Tiandong Wang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

9works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

9 published item(s)

preprint2022arXiv

An Efficient Algorithm for Generating Directed Networks with Predetermined Assortativity Measures

Assortativity coefficients are important metrics to analyze both directed and undirected networks. In general, it is not guaranteed that the fitted model will always agree with the assortativity coefficients in the given network, and the structure of directed networks is more complicated than the undirected ones. Therefore, we provide a remedy by proposing a degree-preserving rewiring algorithm, called DiDPR, for generating directed networks with given directed assortativity coefficients. We construct the joint edge distribution of the target network by accounting for the four directed assortativity coefficients simultaneously, provided that they are attainable, and obtain the desired network by solving a convex optimization problem.Our algorithm also helps check the attainability of the given assortativity coefficients. We assess the performance of the proposed algorithm by simulation studies with focus on two different network models, namely Erdös--Rényi and preferential attachment random networks. We then apply the algorithm to a Facebook wall post network as a real data example. The codes for implementing our algorithm are publicly available in R package wdnet.

preprint2022arXiv

Preferential Attachment with Reciprocity: Properties and Estimation

Reciprocity in social networks helps understand information exchange between two individuals, and indicates interaction patterns between pairs of users. A recent study indicates the reciprocity coefficient of a classical directed preferential attachment (PA) model does not match empirical evidence. In this paper, we extend the classical 3-scenario directed PA model by adding an additional parameter that controls the probability of creating a reciprocal edge. Our proposed model also allows edge creation between two existing nodes, making it a more realistic choice for fitting to real datasets. In addition to analysis of the theoretical properties of this PA model with reciprocity, we provide and compare two estimation procedures for the fitting of the extended model to both simulated and real datasets. The fitted models provide a good match with the empirical tail distributions of both in- and out-degrees. Other mismatched diagnostics suggest that further generalization of the model is warranted.

preprint2022arXiv

Random Networks with Heterogeneous Reciprocity

Users of social networks display diversified behavior and online habits. For instance, a user's tendency to reply to a post can depend on the user and the person posting. For convenience, we group users into aggregated behavioral patterns, focusing here on the tendency to reply to or reciprocate messages. The reciprocity feature in social networks reflects the information exchange among users. We study the properties of a preferential attachment model with heterogeneous reciprocity levels, give the growth rate of model edge counts, and prove convergence of empirical degree frequencies to a limiting distribution. This limiting distribution is not only multivariate regularly varying, but also has the property of hidden regular variation.

preprint2021arXiv

Directed Hybrid Random Networks Mixing Preferential Attachment with Uniform Attachment Mechanisms

Motivated by the complexity of network data, we propose a directed hybrid random network that mixes preferential attachment (PA) rules with uniform attachment (UA) rules. When a new edge is created, with probability $p\in [0,1]$, it follows the PA rule. Otherwise, this new edge is added between two uniformly chosen nodes. Such mixture makes the in- and out-degrees of a fixed node grow at a slower rate, compared to the pure PA case, thus leading to lighter distributional tails. Useful inference methods for the proposed hybrid model are then provided and applied to both synthetic and real datasets. We see that with extra flexibility given by the parameter $p$, the hybrid random network provides a better fit to real-world scenarios, where lighter tails from in- and out-degrees are observed.

preprint2020arXiv

A Directed Preferential Attachment Model with Poisson Measurement

When modeling a directed social network, one choice is to use the traditional preferential attachment model, which generates power-law tail distributions. In a traditional directed preferential attachment, every new edge is added sequentially into the network. However, for real datasets, it is common to only have coarse timestamps available, which means several new edges are created at the same timestamp. Previous analyses on the evolution of social networks reveal that after reaching a stable phase, the growth of edge counts in a network follows a non-homogeneous Poisson process with a constant rate across the day but varying rates from day to day. Taking such empirical observations into account, we propose a modified preferential attachment model with Poisson measurement, and study its asymptotic behavior. This modified model is then fitted to real datasets, and we see it provides a better fit than the traditional one.

preprint2020arXiv

Functional data analysis: An application to COVID-19 data in the United States

The COVID-19 pandemic so far has caused huge negative impacts on different areas all over the world, and the United States (US) is one of the most affected countries. In this paper, we use methods from the functional data analysis to look into the COVID-19 data in the US. We explore the modes of variation of the data through a functional principal component analysis (FPCA), and study the canonical correlation between confirmed and death cases. In addition, we run a cluster analysis at the state level so as to investigate the relation between geographical locations and the clustering structure. Lastly, we consider a functional time series model fitted to the cumulative confirmed cases in the US, and make forecasts based on the dynamic FPCA. Both point and interval forecasts are provided, and the methods for assessing the accuracy of the forecasts are also included.

preprint2020arXiv

On a minimum distance procedure for threshold selection in tail analysis

Power-law distributions have been widely observed in different areas of scientific research. Practical estimation issues include how to select a threshold above which observations follow a power-law distribution and then how to estimate the power-law tail index. A minimum distance selection procedure (MDSP) is proposed in Clauset et al. (2009) and has been widely adopted in practice, especially in the analyses of social networks. However, theoretical justifications for this selection procedure remain scant. In this paper, we study the asymptotic behavior of the selected threshold and the corresponding power-law index given by the MDSP. We find that the MDSP tends to choose too high a threshold level and leads to Hill estimates with large variances and root mean squared errors for simulated data with Pareto-like tails.

preprint2016arXiv

Multivariate Regular Variation of Discrete Mass Functions with Applications to Preferential Attachment Networks

Regular variation of a multivariate measure with a Lebesgue density implies the regular variation of its density provided the density satisfies some regularity conditions. Unlike the univariate case, the converse also requires regularity conditions. We extend these arguments to discrete mass functions and their associated measures using the concept that the the mass function can be embedded in a continuous density function. We give two different conditions, monotonicity and convergence on the unit sphere, both of which can make the discrete function embeddable. Our results are then applied to the preferential attachment network model, and we conclude that the joint mass function of in- and out-degree is embeddable and thus regularly varying.

preprint2015arXiv

Asymptotic Normality of In- and Out-Degree Counts in a Preferential Attachment Model

Preferential attachment in a directed scale-free graph is widely used to model the evolution of social networks. Statistical analyses of social networks often relies on node based data rather than conventional repeated sampling. For our directed edge model with preferential attachment, we prove asymptotic normality of node counts based on a martingale construction and a martingale central limit theorem. This helps justify estimation methods based on the statistics of node counts which have specified in-degree and out-degree.