Source author record

Irene Crimaldi

Irene Crimaldi appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

19works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

19 published item(s)

preprint2026arXiv

Triggered urn models for frequently asked questions (FAQ)

We investigate a nonclassic urn model with triggers that increase the number of colors. The scheme has emerged as a model for web services that set up frequently asked questions (FAQ). We present a thorough asymptotic analysis of the FAQ urn scheme in generality that covers a large number of special cases, such as Simon urn. For instance, we consider time dependent triggering probabilities. We identify regularity conditions on these probabilities that classify the schemes into those where the number of colors in the urn remains almost surely finite or increases to infinity and conditions that tell us whether all the existing colors are observed infinitely often or not. We determine the rank curve, too. In view of the broad generality of the trigger probabilities, a spectrum of limit distributions appears, from central limit theorems to Poisson approximation, to power-laws, revealing connections to Heap's exponent and Zipf's law. A combinatorial approach to the Simon urn is presented to indicate the possibility of such exact analysis, which is important for short-term predictions. Extensive simulations on real datasets (from Amazon sales) as well as computer-generated data clearly indicate that the asymptotic and exact theory developed agrees with practice.

preprint2022arXiv

Statistical test for an urn model with random multidrawing and random addition

We complete the study of the model introduced in [11]. It is a two-color urn model with multiple drawing and random (non-balanced) time-dependent reinforcement matrix. The number of sampled balls at each time-step is random. We identify the exact rates at which the number of balls of each color grows to infinity and define two strongly consistent estimators for the limiting reinforcement averages. Then we prove a Central Limit Theorem, which allows to design a statistical test for such averages.

preprint2021arXiv

Generalized Rescaled Polya urn and its statistical applications

We introduce the Generalized Rescaled Polya (GRP) urn, that provides a generative model for a chi-squared test of goodness of fit for the long-term probabilities of clustered data, with independence between clusters and correlation, due to a reinforcement mechanism, inside each cluster. We apply the proposed test to a data set of Twitter posts about COVID-19 pandemic: in a few words, for a classical chi-squared test the data result strongly significant for the rejection of the null hypothesis (the daily long-run sentiment rate remains constant), but, taking into account the correlation among data, the introduced test leads to a different conclusion. Beside the statistical application, we point out that the GRP urn is a simple variant of the standard Eggenberger-Polya urn, that, with suitable choices of the parameters, shows "local" reinforcement, almost sure convergence of the empirical mean to a deterministic limit and different asymptotic behaviours of the predictive mean. Moreover, the study of this model provides the opportunity to analyze stochastic approximation dynamics, that are unusual in the related literature.

preprint2021arXiv

The Rescaled Polya Urn and the Wright-Fisher process with mutation

In [arXiv:1906.10951 (forthcoming on Advances in Applied Probability),arXiv:2011.05933 (published on PLOS ONE)] the authors introduce, study and apply a new variant of the Eggenberger-Polya urn, called the "Rescaled" Polya urn, which, for a suitable choice of the model parameters, is characterized by the following features: (i) a "local" reinforcement, i.e. a reinforcement mechanism mainly based on the last observations, (ii) a random persistent fluctuation of the predictive mean, and (iii) a long-term almost sure convergence of the empirical mean to a deterministic limit, together with a chi-squared goodness of fit result for the limit probabilities. In this work, motivated by some empirical evidences in [arXiv:2011.05933 (published on PLOS ONE)], we show that the multidimensional Wright-Fisher diffusion with mutation can be obtained as a suitable limit of the predictive means associated to a family of rescaled Polya urns

preprint2020arXiv

Interacting non-linear reinforced stochastic processes: synchronization and no-synchronization

'Rich get richer' rule comforts previously often chosen actions. What is happening to the evolution of individual inclinations to choose an action when agents do interact ? Interaction tends to homogenize while each individual dynamics tends to reinforce its own position. Interacting stochastic systems of reinforced processes were recently considered in many papers, where the asymptotic behavior was proven to exhibit a.s. synchronization. We consider in this paper models where, even if interaction among agents is present, absence of synchronization may happen due to the choice of an individual non-linear reinforcement. We show how these systems can naturally be considered as models for coordination games, technological or opinion dynamics.

preprint2019arXiv

Collaboration and followership: a stochastic model for activities in social networks

In this work we investigate how future actions are influenced by the previous ones, in the specific contexts of scientific collaborations and friendships on social networks. We are not interested in modeling the process of link formation between the agents themselves, we instead describe the activity of the agents, providing a model for the formation of the bipartite network of actions and their features. Therefore we only require to know the chronological order in which the actions are performed, and not the order in which the agents are observed. Moreover, the total number of possible features is not specified a priori but is allowed to increase along time, and new actions can independently show some new entry features or exhibit some of the old ones. The choice of the old features is driven by a degree-fitness method. With this term we mean that the probability that a new action shows one of the old features does not solely depend on the "popularity" of that feature (i.e. the number of previous actions showing it), but is also affected by some individual traits of the agents or the features themselves, synthesized in certain quantities, called "fitnesses" or "weights", that can have different forms and different meaning according to the specific setting considered. We show some theoretical properties of the model and provide statistical tools for the parameters' estimation. The model has been tested on three different datasets and the numerical results are provided and discussed.

preprint2018arXiv

Interacting reinforced stochastic processes: statistical inference based on the weighted empirical means

This work deals with a system of interacting reinforced stochastic processes, where each process $X^j=(X_{n,j})_n$ is located at a vertex $j$ of a finite weighted direct graph, and it can be interpreted as the sequence of "actions" adopted by an agent $j$ of the network. The interaction among the dynamics of these processes depends on the weighted adjacency matrix $W$ associated to the underlying graph: indeed, the probability that an agent $j$ chooses a certain action depends on its personal "inclination" $Z_{n,j}$ and on the inclinations $Z_{n,h}$, with $h\neq j$, of the other agents according to the entries of $W$. The best known example of reinforced stochastic process is the Polya urn. The present paper characterizes the asymptotic behavior of the weighted empirical means $N_{n,j}=\sum_{k=1}^n q_{n,k} X_{k,j}$, proving their almost sure synchronization and some central limit theorems in the sense of stable convergence. By means of a more sophisticated decomposition of the considered processes adopted here, these findings complete and improve some asymptotic results for the personal inclinations $Z^j=(Z_{n,j})_n$ and for the empirical means $\overline{X}^j=(\sum_{k=1}^n X_{k,j}/n)_n$ given in recent papers (e.g. [arXiv:1705.02126, Bernoulli, Forth.]; [arXiv:1607.08514, Ann. Appl. Probab., 27(6):3787-3844, 2017]; [arXiv:1602.06217, Stochastic Process. Appl., 129(1):70-101, 2019]). Our work is motivated by the aim to understand how the different rates of convergence of the involved stochastic processes combine and, from an applicative point of view, by the construction of confidence intervals for the common limit inclination of the agents and of a test statistics to make inference on the matrix $W$, based on the weighted empirical means. In particular, we answer a research question posed in [arXiv:1705.02126, Bernoulli, Forth.]

preprint2016arXiv

Homophily and Triadic Closure in Evolving Social Networks

We present a new network model accounting for multidimensional assortativity. Each node is characterized by a number of features and the probability of a link between two nodes depends on common features. We do not fix a priori the total number of possible features. The bipartite network of the nodes and the features evolves according to a stochastic dynamics that depends on three parameters that respectively regulate the preferential attachment in the transmission of the features to the nodes, the number of new features per node, and the power-law behavior of the total number of observed features. Our model also takes into account a mechanism of triadic closure. We provide theoretical results and statistical estimators for the parameters of the model. We validate our approach by means of simulations and an empirical analysis of a network of scientific collaborations.

preprint2016arXiv

Synchronization and functional central limit theorems for interacting reinforced random walks

We obtain Central Limit Theorems in Functional form for a class of time-inhomogeneous interacting random walks on the simplex of probability measures over a finite set. Due to a reinforcement mechanism, the increments of the walks are correlated, forcing their convergence to the same, possibly random, limit. Random walks of this form have been introduced in the context of urn models and in stochastic approximation. We also propose an application to opinion dynamics in a random network evolving via preferential attachment. We study, in particular, random walks interacting through a mean-field rule and compare the rate they converge to their limit with the rate of synchronization, i.e. the rate at which their mutual distances converge to zero. Under certain conditions, synchronization is faster than convergence.

preprint2015arXiv

Asymptotics for randomly reinforced urns with random barriers

An urn contains black and red balls. Let $Z_n$ be the proportion of black balls at time $n$ and $0\leq L<U\leq 1$ random barriers. At each time $n$, a ball $b_n$ is drawn. If $b_n$ is black and $Z_{n-1}<U$, then $b_n$ is replaced together with a random number $B_n$ of black balls. If $b_n$ is red and $Z_{n-1}>L$, then $b_n$ is replaced together with a random number $R_n$ of red balls. Otherwise, no additional balls are added, and $b_n$ alone is replaced. In this paper, we assume $R_n=B_n$. Then, under mild conditions, it is shown that $Z_n\overset{a.s.}\longrightarrow Z$ for some random variable $Z$, and \begin{gather*} D_n:=\sqrt{n}\,(Z_n-Z)\longrightarrow\mathcal{N}(0,σ^2)\quad\text{conditionally a.s.} \end{gather*} where $σ^2$ is a certain random variance. Almost sure conditional convergence means that \begin{gather*} P\bigl(D_n\in\cdot\mid\mathcal{G}_n\bigr)\overset{weakly}\longrightarrow\mathcal{N}(0,\,σ^2)\quad\text{a.s.} \end{gather*} where $P\bigl(D_n\in\cdot\mid\mathcal{G}_n\bigr)$ is a regular version of the conditional distribution of $D_n$ given the past $\mathcal{G}_n$. Thus, in particular, one obtains $D_n\longrightarrow\mathcal{N}(0,σ^2)$ stably. It is also shown that $L<Z<U$ a.s. and $Z$ has non-atomic distribution.

preprint2015arXiv

Central limit theorems for a hypergeometric randomly reinforced urn

We consider a variant of the randomly reinforced urn where more balls can be simultaneously drawn out and balls of different colors can be simultaneously added. More precisely, at each time-step, the conditional distribution of the number of extracted balls of a certain color given the past is assumed to be hypergeometric. We prove some central limit theorems in the sense of stable convergence and of almost sure conditional convergence, which are stronger than convergence in distribution. The proven results provide asymptotic confidence intervals for the limit proportion, whose distribution is generally unknown. Moreover, we also consider the case of more urns subjected to some random common factors.

preprint2015arXiv

Central limit theorems for an Indian buffet model with random weights

The three-parameter Indian buffet process is generalized. The possibly different role played by customers is taken into account by suitable (random) weights. Various limit theorems are also proved for such generalized Indian buffet process. Let $L_n$ be the number of dishes experimented by the first $n$ customers, and let $\overline{K}_n=(1/n)\sum_{i=1}^nK_i$ where $K_i$ is the number of dishes tried by customer $i$. The asymptotic distributions of $L_n$ and $\overline{K}_n$, suitably centered and scaled, are obtained. The convergence turns out to be stable (and not only in distribution). As a particular case, the results apply to the standard (i.e., nongeneralized) Indian buffet process.

preprint2015arXiv

Fluctuation Theorems for Synchronization of Interacting Polya's urns

We consider a model of N two-colors urns in which the reinforcement of each urn depends also on the content of all the other urns. This interaction is of mean-field type and it is tuned by a parameter $α$ in [0,1]; in particular, for $α=0$ the N urns behave as N independent Polya's urns. As shown in [9], for $α>0$ urns synchronize, in the sense that the fraction of balls of a given color converges a.s. to the same (random) limit in all urns. In this paper we study fluctuations around this synchronized regime. The scaling of these fluctuations depends on the parameter $α$. In particular the standard scaling $t^{-1/2}$ appears only for $α>1/2$. For $α\geq 1/2$ we also determine the limit distribution of the rescaled fluctuations. We use the notion of stable convergence, which is stronger than convergence in distribution.

preprint2014arXiv

A Network Model characterized by a Latent Attribute Structure with Competition

The quest for a model that is able to explain, describe, analyze and simulate real-world complex networks is of uttermost practical as well as theoretical interest. In this paper we introduce and study a network model that is based on a latent attribute structure: each node is characterized by a number of features and the probability of the existence of an edge between two nodes depends on the features they share. Features are chosen according to a process of Indian-Buffet type but with an additional random "fitness" parameter attached to each node, that determines its ability to transmit its own features to other nodes. As a consequence, a node's connectivity does not depend on its age alone, so also "young" nodes are able to compete and succeed in acquiring links. One of the advantages of our model for the latent bipartite "node-attribute" network is that it depends on few parameters with a straightforward interpretation. We provide some theoretical, as well experimental, results regarding the power-law behaviour of the model and the estimation of the parameters. By experimental data, we also show how the proposed model for the attribute structure naturally captures most local and global properties (e.g., degree distributions, connectivity and distance distributions) real networks exhibit. keyword: Complex network, social network, attribute matrix, Indian Buffet process

preprint2014arXiv

Cluster analysis of weighted bipartite networks: a new copula-based approach

In this work we are interested in identifying clusters of "positional equivalent" actors, i.e. actors who play a similar role in a system. In particular, we analyze weighted bipartite networks that describes the relationships between actors on one side and features or traits on the other, together with the intensity level to which actors show their features. The main contribution of our work is twofold. First, we develop a methodological approach that takes into account the underlying multivariate dependence among groups of actors. The idea is that positions in a network could be defined on the basis of the similar intensity levels that the actors exhibit in expressing some features, instead of just considering relationships that actors hold with each others. Second, we propose a new clustering procedure that exploits the potentiality of copula functions, a mathematical instrument for the modelization of the stochastic dependence structure. Our clustering algorithm can be applied both to binary and real-valued matrices. We validate it with simulations and applications to real-world data.

preprint2013arXiv

A generalized telegraph process with velocity driven by random trials

We consider a random trial-based telegraph process, which describes a motion on the real line with two constant velocities along opposite directions. At each epoch of the underlying counting process the new velocity is determined by the outcome of a random trial. Two schemes are taken into account: Bernoulli trials and classical Pólya urn trials. We investigate the probability law of the process and the mean of the velocity of the moving particle. We finally discuss two cases of interest: (i) the case of Bernoulli trials and intertimes having exponential distributions with linear rates (in which, interestingly, the process exhibits a logistic stationary density with non-zero mean), and (ii) the case of Pólya trials and intertimes having first Gamma and then exponential distributions with constant rates.

preprint2012arXiv

The Evolution of Complex Networks: A New Framework

We introduce a new framework for the analysis of the dynamics of networks, based on randomly reinforced urn (RRU) processes, in which the weight of the edges is determined by a reinforcement mechanism. We rigorously explain the empirical evidence that in many real networks there is a subset of "dominant edges" that control a major share of the total weight of the network. Furthermore, we introduce a new statistical procedure to study the evolution of networks over time, assessing if a given instance of the nework is taken at its steady state or not. Our results are quite general, since they are not based on a particular probability distribution or functional form of the weights. We test our model in the context of the International Trade Network, showing the existence of a core of dominant links and determining its size.

preprint2010arXiv

Rate of convergence of predictive distributions for dependent data

This paper deals with empirical processes of the type \[C_n(B)=\sqrt{n}\{μ_n(B)-P(X_{n+1}\in B\mid X_1,...,X_n)\},\] where $(X_n)$ is a sequence of random variables and $μ_n=(1/n)\sum_{i=1}^nδ_{X_i}$ the empirical measure. Conditions for $\sup_B|C_n(B)|$ to converge stably (in particular, in distribution) are given, where $B$ ranges over a suitable class of measurable sets. These conditions apply when $(X_n)$ is exchangeable or, more generally, conditionally identically distributed (in the sense of Berti et al. [Ann. Probab. 32 (2004) 2029--2052]). By such conditions, in some relevant situations, one obtains that $\sup_B|C_n(B)|\stackrel{P}{\to}0$ or even that $\sqrt{n}\sup_B|C_n(B)|$ converges a.s. Results of this type are useful in Bayesian statistics.