Catalog footprint

What is connected

66works
30topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

66 published item(s)

preprint2022arXiv

Equivalence of information production and generalized entropies in complex processes

Complex systems that are characterized by strong correlations and fat-tailed distribution functions have been argued to be incompatible within the framework of Boltzmann-Gibbs entropy. As an alternative, so-called generalized entropies were proposed and intensively studied. Here we show that this incompatibility is a misconception. For a broad class of processes, Boltzmann entropy the log multiplicity remains the valid entropy concept, however, for non-i.i.d., non-multinomial, and non-ergodic processes, Boltzmann entropy is not of Shannon form. The correct form of Boltzmann entropy can be shown to be identical with generalized entropies. We derive this result for all processes that can be mapped reversibly to adjoint representations where processes are i.i.d.. In these representations the information production is given by the Shannon entropy. We proof that over the original sampling space this yields functionals that are identical to generalized entropies. The problem of constructing adequate context-sensitive entropy functionals therefore can be translated into the much simpler problem of finding adjoint representations. The method provides a comprehensive framework for a statistical physics of strongly correlated systems and complex processes.

preprint2022arXiv

Loss of sustainability in scientific work

For decades the number of scientific publications has been rapidly increasing, effectively out-dating knowledge at a tremendous rate. Only few scientific milestones remain relevant and continuously attract citations. Here we quantify how long scientific work remains being utilized, how long it takes before today's work is forgotten, and how milestone papers differ from those forgotten. To answer these questions, we study the complete temporal citation network of all American Physical Society journals. We quantify the probability of attracting citations for individual publications based on age and the number of citations they have received in the past. We capture both aspects, the forgetting and the tendency to cite already popular works, in a microscopic generative model for the dynamics of scientific citation networks. We find that the probability of citing a specific paper declines with age as a power law with an exponent of $α\sim -1.4$. Whenever a paper in its early years can be characterized by a scaling exponent above a critical value, $α_c$, the paper is likely to become "ever-lasting". We validate the model with out-of-sample predictions, with an accuracy of up to 90% (AUC $\sim 0.9$). The model also allows us to estimate an expected citation landscape of the future, predicting that 95% of papers cited in 2050 have yet to be published. The exponential growth of articles, combined with a power-law type of forgetting and papers receiving fewer and fewer citations on average, suggests a worrying tendency toward information overload and raises concerns about scientific publishing's long-term sustainability.

preprint2022arXiv

Propagation of disruptions in supply networks of essential goods: A population-centered perspective of systemic risk

The Covid-19 pandemic drastically emphasized the fragility of national and international supply networks (SNs),leading to significant supply shortages of essential goods for people, such as food and medical equipment. Severe disruptions that propagate along complex SNs can expose the population of entire regions or even countries to these risks. A lack of both, data and quantitative methodology, has hitherto hindered us to empirically quantify the vulnerability of the population to disruptions. Here we develop a data-driven simulation methodology to locally quantify actual supply losses for the population that result from the cascading of supply disruptions. We demonstrate the method on a large food SN of a European country including 22,938 business premises, 44,355 supply links and 116 local administrative districts. We rank the business premises with respect to their criticality for the districts' population with the proposed systemic risk index, SRIcrit, to identify around 30 premises that -- in case of their failure -- are expected to cause critical supply shortages in sizable fractions of the population. The new methodology is immediately policy relevant as a fact-driven and generalizable crisis management tool. This work represents a starting point for quantitatively studying SN disruptions focused on the well-being of the population.

preprint2022arXiv

Specialization in Criminal Careers

We use a comprehensive longitudinal dataset on criminal acts over five years in a European country to study specialization in criminal careers. We cluster crime categories by their relative co-occurrence within criminal careers, deriving a natural, data-based taxonomy of criminal specialization. Defining specialists as active criminals who stay within one category of offending behavior, we study their socio-demographic attributes, geographic range, and positions in their collaboration networks, relative to their generalist counterparts. In comparison to generalists, specialists tend to be older, more likely to be female, operate within a smaller geographic range, and collaborate in smaller, more tightly-knit local networks. We observe that specialists are more intensely embedded in criminal networks and find evidence that specialization indeed reflects division of labor and organization.

preprint2021arXiv

Thermodynamics of structure-forming systems

Structure-forming systems are ubiquitous in nature, ranging from atoms building molecules to self-assembly of colloidal amphibolic particles. The understanding of the underlying thermodynamics of such systems remains an important problem. Here we derive the entropy for structure-forming systems that differs from Boltzmann-Gibbs entropy by a term that explicitly captures clustered states. For large systems and low concentrations, the approach is equivalent to the grand-canonical ensemble; for small systems, we find significant deviations. We derive the detailed fluctuation theorem and Crooks' work fluctuation theorem for structure-forming systems. The connection to the theory of particle self-assembly is discussed. We apply the results to several physical systems. We present the phase diagram for patchy particles described by the Kern-Frenkel potential. We show that the Curie-Weiss model with molecule structures exhibits a first-order phase transition.

preprint2020arXiv

Boosting test-efficiency by pooled testing strategies for SARS-CoV-2

In the current COVID19 crisis many national healthcare systems are confronted with an acute shortage of tests for confirming SARS-CoV-2 infections. For low overall infection levels in the population, pooling of samples can drastically amplify the testing efficiency. Here we present a formula to estimate the optimal pooling size, the efficiency gain (tested persons per test), and the expected upper bound of missed infections in the pooled testing, all as a function of the populationwide infection levels and the false negative/positive rates of the currently used PCR tests. Assuming an infection level of 0.1 % and a false negative rate of 2 %, the optimal pool size is about 32, the efficiency gain is about 15 tested persons per test. For an infection level of 1 % the optimal pool size is 11, the efficiency gain is 5.1 tested persons per test. For an infection level of 10 % the optimal pool size reduces to about 4, the efficiency gain is about 1.7 tested persons per test. For infection levels of 30 % and higher there is no more benefit from pooling. To see to what extent replicates of the pooled tests improve the estimate of the maximal number of missed infections, we present all results for 1, 3, and 5 replicates.

preprint2020arXiv

Information geometry of scaling expansions of non-exponentially growing configuration spaces

Many stochastic complex systems are characterized by the fact that their configuration space doesn't grow exponentially as a function of the degrees of freedom. The use of scaling expansions is a natural way to measure the asymptotic growth of the configuration space volume in terms of the scaling exponents of the system. These scaling exponents can, in turn, be used to define universality classes that uniquely determine the statistics of a system. Every system belongs to one of these classes. Here we derive the information geometry of scaling expansions of sample spaces. In particular, we present the deformed logarithms and the metric in a systematic and coherent way. We observe a phase transition for the curvature. The phase transition can be well measured by the characteristic length r, corresponding to a ball with radius 2r having the same curvature as the statistical manifold. Increasing characteristic length with respect to the size of the system is associated with sub-exponential sample space growth is associated with strongly constrained and correlated complex systems. Decreasing of the characteristic length corresponds to super-exponential sample space growth that occurs for example in systems that develop structure as they evolve. Constant curvature means exponential sample space growth that is associated with multinomial statistics, and traditional Boltzmann-Gibbs, or Shannon statistics applies. This allows us to characterize transitions between statistical manifolds corresponding to different families of probability distributions.

preprint2020arXiv

Quantifying exaptation in scientific evolution

Rediscovering a new function for something can be just as important as the discovery itself. In 1982, Stephen Jay Gould and Elisabeth Vrba named this phenomenon Exaptation to describe a radical shift in the function of a specific trait during biological evolution. While exaptation is thought to be a fundamental mechanism for generating adaptive innovations, diversity, and sophisticated features, relatively little effort has been made to quantify exaptation outside the topic of biological evolution. We think that this concept provides a useful framework for characterising the emergence of innovations in science. This article explores the notion that exaptation arises from the usage of scientific ideas in domains other than the area that they were originally applied to. In particular, we adopt a normalised entropy and an inverse participation ratio as observables that reveal and quantify the concept of exaptation. We identify distinctive patterns of exaptation and expose specific examples of papers that display those patterns. Our approach represents a first step towards the quantification of exaptation phenomena in the context of scientific evolution.

preprint2020arXiv

The effect of social balance on social fragmentation

With the availability of cell phones, internet, social media etc. the interconnectedness of people within most societies has increased drastically over the past three decades. Across the same timespan, we are observing the phenomenon of increasing levels of fragmentation in society into relatively small and isolated groups that have been termed filter bubbles, or echo chambers. These pose a number of threats to open societies, in particular, a radicalisation in political, social or cultural issues, and a limited access to facts. In this paper we show that these two phenomena might be tightly related. We study a simple stochastic co-evolutionary model of a society of interacting people. People are not only able to update their opinions within their social context, but can also update their social links from collaborative to hostile, and vice versa. The latter is implemented such that social balance is realised. We find that there exists a critical level of interconnectedness, above which society fragments into small sub-communities that are positively linked within and hostile towards other groups. We argue that the existence of a critical communication density is a universal phenomenon in all societies that exhibit social balance. The necessity arises from the underlying mathematical structure of a phase transition phenomenon that is known from the theory of a kind of disordered magnets called spin glasses. We discuss the consequences of this phase transition for social fragmentation in society.

preprint2020arXiv

Why are most COVID-19 infection curves linear?

Many countries have passed their first COVID-19 epidemic peak. Traditional epidemiological models describe this as a result of non-pharmaceutical interventions that pushed the growth rate below the recovery rate. In this new phase of the pandemic many countries show an almost linear growth of confirmed cases for extended time-periods. This new containment regime is hard to explain by traditional models where infection numbers either grow explosively until herd immunity is reached, or the epidemic is completely suppressed (zero new cases). Here we offer an explanation of this puzzling observation based on the structure of contact networks. We show that for any given transmission rate there exists a critical number of social contacts, $D_c$, below which linear growth and low infection prevalence must occur. Above $D_c$ traditional epidemiological dynamics takes place, as e.g. in SIR-type models. When calibrating our corresponding model to empirical estimates of the transmission rate and the number of days being contagious, we find $D_c\sim 7.2$. Assuming realistic contact networks with a degree of about 5, and assuming that lockdown measures would reduce that to household-size (about 2.5), we reproduce actual infection curves with a remarkable precision, without fitting or fine-tuning of parameters. In particular we compare the US and Austria, as examples for one country that initially did not impose measures and one that responded with a severe lockdown early on. Our findings question the applicability of standard compartmental models to describe the COVID-19 containment phase. The probability to observe linear growth in these is practically zero.

preprint2019arXiv

The role of mainstreamness and interdisciplinarity for the relevance of scientific papers

There is demand from science funders, industry, and the public that science should become more risk-taking, more out-of-the-box, and more interdisciplinary. Is it possible to tell how interdisciplinary and out-of-the-box scientific papers are, or which papers are mainstream? Here we use the bibliographic coupling network, derived from all physics papers that were published in the Physical Review journals in the past century, to try to identify them as mainstream, out-of-the-box, or interdisciplinary. We show that the network clusters into scientific fields. The position of individual papers with respect to these clusters allows us to estimate their degree of mainstreamness or interdisciplinary. We show that over the past decades the fraction of mainstream papers increases, the fraction of out-of-the-box decreases, and the fraction of interdisciplinary papers remains constant. Studying the rewards of papers, we find that in terms of absolute citations, both, mainstream and interdisciplinary papers are rewarded. In the long run, mainstream papers perform less than interdisciplinary ones in terms of citation rates. We conclude that to avoid a trend towards mainstreamness a new incentive scheme is necessary.

preprint2016arXiv

Analytical computation of frequency distributions of path-dependent processes by means of a non-multinomial maximum entropy approach

Path-dependent stochastic processes are often non-ergodic and observables can no longer be computed within the ensemble picture. The resulting mathematical difficulties pose severe limits to the analytical understanding of path-dependent processes. Their statistics is typically non-multinomial in the sense that the multiplicities of the occurrence of states is not a multinomial factor. The maximum entropy principle is tightly related to multinomial processes, non-interacting systems, and to the ensemble picture; It loses its meaning for path-dependent processes. Here we show that an equivalent to the ensemble picture exists for path-dependent processes, such that the non-multinomial statistics of the underlying dynamical process, by construction, is captured correctly in a functional that plays the role of a relative entropy. We demonstrate this for self-reinforcing Pólya urn processes, which explicitly generalise multinomial statistics. We demonstrate the adequacy of this constructive approach towards non-multinomial pendants of entropy by computing frequency and rank distributions of Pólya urn processes. We show how microscopic update rules of a path-dependent process allow us to explicitly construct a non-multinomial entropy functional, that, when maximized, predicts the time-dependent distribution function.

preprint2016arXiv

Basel III capital surcharges for G-SIBs fail to control systemic risk and can cause pro-cyclical side effects

In addition to constraining bilateral exposures of financial institutions, there are essentially two options for future financial regulation of systemic risk (SR): First, financial regulation could attempt to reduce the financial fragility of global or domestic systemically important financial institutions (G-SIBs or D-SIBs), as for instance proposed in Basel III. Second, future financial regulation could attempt strengthening the financial system as a whole. This can be achieved by re-shaping the topology of financial networks. We use an agent-based model (ABM) of a financial system and the real economy to study and compare the consequences of these two options. By conducting three "computer experiments" with the ABM we find that re-shaping financial networks is more effective and efficient than reducing leverage. Capital surcharges for G-SIBs can reduce SR, but must be larger than those specified in Basel III in order to have a measurable impact. This can cause a loss of efficiency. Basel III capital surcharges for G-SIBs can have pro-cyclical side effects.

preprint2016arXiv

Disentangling genetic and environmental risk factors for individual diseases from multiplex comorbidity networks

Most disorders are caused by a combination of multiple genetic and/or environmental factors. If two diseases are caused by the same molecular mechanism, they tend to co-occur in patients. Here we provide a quantitative method to disentangle how much genetic or environmental risk factors contribute to the pathogenesis of 358 individual diseases, respectively. We pool data on genetic, pathway-based, and toxicogenomic disease-causing mechanisms with disease co-occurrence data obtained from almost two million patients. From this data we construct a multilayer network where nodes represent disorders that are connected by links that either represent phenotypic comorbidity of the patients or the involvement of a certain molecular mechanism. From the similarity of phenotypic and mechanism-based networks for each disorder we derive measure that allows us to quantify the relative importance of various molecular mechanisms for a given disease. We find that most diseases are dominated by genetic risk factors, while environmental influences prevail for disorders such as depressions, cancers, or dermatitis. Almost never we find that more than one type of mechanisms is involved in the pathogenesis of diseases.

preprint2016arXiv

Dynamical origins of the community structure of multi-layer societies

Social structures emerge as a result of individuals managing a variety of different of social relationships. Societies can be represented as highly structured dynamic multiplex networks. Here we study the dynamical origins of the specific community structures of a large-scale social multiplex network of a human society that interacts in a virtual world of a massive multiplayer online game. There we find substantial differences in the community structures of different social actions, represented by the various network layers in the multiplex. Community size distributions are either similar to a power-law or appear to be centered around a size of 50 individuals. To understand these observations we propose a voter model that is built around the principle of triadic closure. It explicitly models the co-evolution of node- and link-dynamics across different layers of the multiplex. Depending on link- and node fluctuation rates, the model exhibits an anomalous shattered fragmentation transition, where one layer fragments from one large component into many small components. The observed community size distributions are in good agreement with the predicted fragmentation in the model. We show that the empirical pairwise similarities of network layers, in terms of link overlap and degree correlations, practically coincide with the model. This suggests that several detailed features of the fragmentation in societies can be traced back to the triadic closure processes.

preprint2016arXiv

Elimination of systemic risk in financial networks by means of a systemic risk transaction tax

Financial markets are exposed to systemic risk (SR), the risk that a major fraction of the system ceases to function, and collapses. It has recently become possible to quantify SR in terms of underlying financial networks where nodes represent financial institutions, and links capture the size and maturity of assets (loans), liabilities, and other obligations, such as derivatives. We demonstrate that it is possible to quantify the share of SR that individual liabilities within a financial network contribute to the overall SR. We use empirical data of nationwide interbank liabilities to show that the marginal contribution to overall SR of liabilities for a given size varies by a factor of a thousand. We propose a tax on individual transactions that is proportional to their marginal contribution to overall SR. If a transaction does not increase SR it is tax-free. With an agent-based model (CRISIS macro-financial model) we demonstrate that the proposed "Systemic Risk Tax" (SRT) leads to a self-organised restructuring of financial networks that are practically free of SR. The SRT can be seen as an insurance for the public against costs arising from cascading failure. ABM predictions are shown to be in remarkable agreement with the empirical data and can be used to understand the relation of credit risk and SR.

preprint2016arXiv

Stripping syntax from complexity: An information-theoretical perspective on complex systems

Claude Shannons information theory (1949) has had a revolutionary impact on communication science. A crucial property of his framework is that it decouples the meaning of a message from the mechanistic details from the actual communication process itself, which opened the way to solve long-standing communication problems. Here we argue that a similar impact could be expected by applying information theory in the context of complexity science to answer long-standing, cross-domain questions about the nature of complex systems. This happens by decoupling the domain-specific model details (e.g., neuronal networks, ecosystems, flocks of birds) from the cross-domain phenomena that characterize complex systems (e.g., criticality, robustness, tipping points). This goes beyond using information theory as a non-linear correlation measure, namely it allows describing a complex system entirely in terms of the storage, transfer, and modification of informational bits. After all, a phenomenon that does not depend on model details should best be studied in a framework that strips away all such details. We highlight the first successes of information-theoretic descriptions in the recent complexity literature, and emphasize that this type of research is still in its infancy. Finally we sketch how such an information-theoretic description may even lead to a new type of universality among complex systems, with a potentially tremendous impact. The goal of this perspective article is to motivate a paradigm shift in the young field of complexity science using a lesson learnt in communication science.

preprint2015arXiv

Systemic trade-risk of critical resources

In the wake of the 2008 financial crisis the role of strongly interconnected markets in fostering systemic instability has been increasingly acknowledged. Trade networks of commodities are susceptible to deleterious cascades of supply shocks that increase systemic trade-risks and pose a threat to geopolitical stability. On a global and a regional level we show that supply risk, scarcity, and price volatility of non-fuel mineral resources are intricately connected with the structure of the world-trade network of or spanned by these resources. On the global level we demonstrate that the scarcity of a resource, as measured by its trade volume compared to extractable reserves, is closely related to the susceptibility of the trade network with respect to cascading shocks. On the regional level we find that to some extent the region-specific price volatility and supply risk can be understood by centrality measures that capture systemic trade-risk. The resources associated with the highest systemic trade-risk indicators are often those that are produced as byproducts of major metals. We identify significant shortcomings in the management of systemic trade-risk, in particular in the EU.

preprint2015arXiv

The multi-layer network nature of systemic risk and its implications for the costs of financial crises

The inability to see and quantify systemic financial risk comes at an immense social cost. Systemic risk in the financial system arises to a large extent as a consequence of the interconnectedness of its institutions, which are linked through networks of different types of financial contracts, such as credit, derivatives, foreign exchange and securities. The interplay of the various exposure networks can be represented as layers in a financial multi-layer network. In this work we quantify the daily contributions to systemic risk from four layers of the Mexican banking system from 2007-2013. We show that focusing on a single layer underestimates the total systemic risk by up to 90%. By assigning systemic risk levels to individual banks we study the systemic risk profile of the Mexican banking system on all market layers. This profile can be used to quantify systemic risk on a national level in terms of nation-wide expected systemic losses. We show that market-based systemic risk indicators systematically underestimate expected systemic losses. We find that expected systemic losses are up to a factor four higher now than before the financial crisis of 2007-2008. We find that systemic risk contributions of individual transactions can be up to a factor of thousand higher than the corresponding credit risk, which creates huge risks for the public. We find an intriguing non-linear effect whereby the sum of systemic risk of all layers underestimates the total risk. The method presented here is the first objective data driven quantification of systemic risk on national scales that reveal its true levels.

preprint2015arXiv

Understanding scaling through history-dependent processes with collapsing sample space

History-dependent processes are ubiquitous in natural and social systems. Many such stochastic processes, especially those that are associated with complex systems, become more constrained as they unfold, meaning that their sample-space, or their set of possible outcomes, reduces as they age. We demonstrate that these sample-space reducing (SSR) processes necessarily lead to Zipf's law in the rank distributions of their outcomes. We show that by adding noise to SSR processes the corresponding rank distributions remain exact power-laws, $p(x)\sim x^{-λ}$, where the exponent directly corresponds to the mixing ratio of the SSR process and noise. This allows us to give a precise meaning to the scaling exponent in terms of the degree to how much a given process reduces its sample-space as it unfolds. Noisy SSR processes further allow us to explain a wide range of scaling exponents in frequency distributions ranging from $α= 2$ to $\infty$. We discuss several applications showing how SSR processes can be used to understand Zipf's law in word frequencies, and how they are related to diffusion processes in directed networks, or ageing processes such as in fragmentation processes. SSR processes provide a new alternative to understand the origin of scaling in complex systems without the recourse to multiplicative, preferential, or self-organised critical processes.

preprint2015arXiv

Understanding Zipf's law of word frequencies through sample-space collapse in sentence formation

The formation of sentences is a highly structured and history-dependent process. The probability of using a specific word in a sentence strongly depends on the 'history' of word-usage earlier in that sentence. We study a simple history-dependent model of text generation assuming that the sample-space of word usage reduces along sentence formation, on average. We first show that the model explains the approximate Zipf law found in word frequencies as a direct consequence of sample-space reduction. We then empirically quantify the amount of sample-space reduction in the sentences of ten famous English books, by analysis of corresponding word-transition tables that capture which words can follow any given word in a text. We find a highly nested structure in these transition tables and show that this `nestedness' is tightly related to the power law exponents of the observed word frequency distributions. With the proposed model it is possible to understand that the nestedness of a text can be the origin of the actual scaling exponent, and that deviations from the exact Zipf law can be understood by variations of the degree of nestedness on a book-by-book basis. On a theoretical level we are able to show that in case of weak nesting, Zipf's law breaks down in a fast transition. Unlike previous attempts to understand Zipf's law in language the sample-space reducing model is not based on assumptions of multiplicative, preferential, or self-organised critical mechanisms behind language formation, but simply used the empirically quantifiable parameter 'nestedness' to understand the statistics of word frequencies.

preprint2014arXiv

Behavioral and Network Origins of Wealth Inequality: Insights from a Virtual World

Almost universally, wealth is not distributed uniformly within societies or economies. Even though wealth data have been collected in various forms for centuries, the origins for the observed wealth-disparity and social inequality are not yet fully understood. Especially the impact and connections of human behavior on wealth could so far not be inferred from data. Here we study wealth data from the virtual economy of the massive multiplayer online game (MMOG) Pardus. This data not only contains every player's wealth at every point in time, but also all actions of every player over a timespan of almost a decade. We find that wealth distributions in the virtual world are very similar to those in western countries. In particular we find an approximate exponential for low wealth and a power-law tail. The Gini index is found to be $g=0.65$, which is close to the indices of many Western countries. We find that wealth-increase rates depend on the time when players entered the game. Players that entered the game early on tend to have remarkably higher wealth-increase rates than those who joined later. Studying the players' positions within their social networks, we find that the local position in the trade network is most relevant for wealth. Wealthy people have high in- and out-degree in the trade network, relatively low nearest-neighbor degree and a low clustering coefficient. Wealthy players have many mutual friendships and are socially well respected by others, but spend more time on business than on socializing. We find that players that are not organized within social groups with at least three members are significantly poorer on average. We observe that high `political' status and high wealth go hand in hand. Wealthy players have few personal enemies, but show animosity towards players that behave as public enemies.

preprint2014arXiv

Detection of the elite structure in a virtual multiplex social system by means of a generalized $K$-core

Elites are subgroups of individuals within a society that have the ability and means to influence, lead, govern, and shape societies. Members of elites are often well connected individuals, which enables them to impose their influence to many and to quickly gather, process, and spread information. Here we argue that elites are not only composed of highly connected individuals, but also of intermediaries connecting hubs to form a cohesive and structured elite-subgroup at the core of a social network. For this purpose we present a generalization of the $K$-core algorithm that allows to identify a social core that is composed of well-connected hubs together with their `connectors'. We show the validity of the idea in the framework of a virtual world defined by a massive multiplayer online game, on which we have complete information of various social networks. Exploiting this multiplex structure, we find that the hubs of the generalized $K$-core identify those individuals that are high social performers in terms of a series of indicators that are available in the game. In addition, using a combined strategy which involves the generalized $K$-core and the recently introduced $M$-core, the elites of the different 'nations' present in the game are perfectly identified as modules of the generalized $K$-core. Interesting sudden shifts in the composition of the elite cores are observed at deep levels. We show that elite detection with the traditional $K$-core is not possible in a reliable way. The proposed method might be useful in a series of more general applications, such as community detection.

preprint2014arXiv

Fractal multi-level organisation of human groups in a virtual world

Humans are fundamentally social. They have progressively dominated their environment by the strength and creativity provided by and within their grouping. It is well recognised that human groups are highly structured, and the anthropological literature has loosely classified them according to their size and function, such as support cliques, sympathy groups, bands, cognitive groups, tribes, linguistic groups and so on. Recently, combining data on human grouping patterns in a comprehensive and systematic study, Zhou et al. identified a quantitative discrete hierarchy of group sizes with a preferred scaling ratio close to $3$, which was later confirmed for hunter-gatherer groups and for other mammalian societies. Using high precision large scale Internet-based social network data, we extend these early findings on a very large data set. We analyse the organisational structure of a complete, multi-relational, large social multiplex network of a human society consisting of about 400,000 odd players of a massive multiplayer online game for which we know all about the group memberships of every player. Remarkably, the online players exhibit the same type of structured hierarchical layers as the societies studied by anthropologists, where each of these layers is three to four times the size of the lower layer. Our findings suggest that the hierarchical organisation of human society is deeply nested in human psychology.

preprint2014arXiv

How multiplicity determines entropy and the derivation of the maximum entropy principle for complex systems

The maximum entropy principle (MEP) is a method for obtaining the most likely distribution functions of observables from statistical systems, by maximizing entropy under constraints. The MEP has found hundreds of applications in ergodic and Markovian systems in statistical mechanics, information theory, and statistics. For several decades there exists an ongoing controversy whether the notion of the maximum entropy principle can be extended in a meaningful way to non-extensive, non-ergodic, and complex statistical systems and processes. In this paper we start by reviewing how Boltzmann-Gibbs-Shannon entropy is related to multiplicities of independent random processes. We then show how the relaxation of independence naturally leads to the most general entropies that are compatible with the first three Shannon-Khinchin axioms, the (c,d)-entropies. We demonstrate that the MEP is a perfectly consistent concept for non-ergodic and complex statistical systems if their relative entropy can be factored into a generalized multiplicity and a constraint term. The problem of finding such a factorization reduces to finding an appropriate representation of relative entropy in a linear basis. In a particular example we show that path-dependent random processes with memory naturally require specific generalized entropies. The example is the first exact derivation of a generalized entropy from the microscopic properties of a path-dependent random process.

preprint2014arXiv

Instrumentational complexity of music genres and why simplicity sells

Listening habits are strongly influenced by two opposing aspects, the desire for variety and the demand for uniformity in music. In this work we quantify these two notions in terms of musical instrumentation and production technologies that are typically involved in crafting popular music. We assign a "complexity value" to each music style. A style is complex if it shows the property of having both high variety and low uniformity in instrumentation. We find a strong inverse relation between variety and uniformity of music styles that is remarkably stable over the last half century. Individual styles, however, show dramatic changes in their "complexity" during that period. Styles like "new wave" or "disco" quickly climbed towards higher complexity in the 70s and fell back to low complexity levels shortly afterwards, whereas styles like "folk rock" remained at constant high complexity levels. We show that changes in the complexity of a style are related to its number of sales and to the number of artists contributing to that style. As a style attracts a growing number of artists, its instrumentational variety usually increases. At the same time the instrumentational uniformity of a style decreases, i.e. a unique stylistic and increasingly complex expression pattern emerges. In contrast, album sales of a given style typically increase with decreasing complexity. This can be interpreted as music becoming increasingly formulaic once commercial or mainstream success sets in.

preprint2014arXiv

Interevent time distributions of human multi-level activity in a virtual world

Studying human behaviour in virtual environments provides extraordinary opportunities for a quantitative analysis of social phenomena with levels of accuracy that approach those of the natural sciences. In this paper we use records of player activities in the massive multiplayer online game Pardus over 1,238 consecutive days, and analyze dynamical features of sequences of actions of players. We build on previous work were temporal structures of human actions of the same type were quantified, and extend provide an empirical understanding of human actions of different types. This study of multi-level human activity can be seen as a dynamic counterpart of static multiplex network analysis. We show that the interevent time distributions of actions in the Pardus universe follow highly non-trivial distribution functions, from which we extract action-type specific characteristic "decay constants". We discuss characteristic features of interevent time distributions, including periodic patterns on different time scales, bursty dynamics, and various functional forms on different time scales. We comment on gender differences of players in emotional actions, and find that while male and female act similarly when performing some positive actions, females are slightly faster for negative actions. We also observe effects on the age of players: more experienced players are generally faster in making decisions about engaging and terminating in enmity and friendship, respectively.

preprint2014arXiv

Leverage-induced systemic risk under Basle II and other credit risk policies

We use a simple agent based model of value investors in financial markets to test three credit regulation policies. The first is the unregulated case, which only imposes limits on maximum leverage. The second is Basle II and the third is a hypothetical alternative in which banks perfectly hedge all of their leverage-induced risk with options. When compared to the unregulated case both Basle II and the perfect hedge policy reduce the risk of default when leverage is low but increase it when leverage is high. This is because both regulation policies increase the amount of synchronized buying and selling needed to achieve deleveraging, which can destabilize the market. None of these policies are optimal for everyone: Risk neutral investors prefer the unregulated case with low maximum leverage, banks prefer the perfect hedge policy, and fund managers prefer the unregulated case with high maximum leverage. No one prefers Basle II.

preprint2014arXiv

Physical forces between humans and how humans attract and repel each other based on their social interactions in an online world

Physical interactions between particles are the result of the exchange of gauge bosons. Human interactions are mediated by the exchange of messages, goods, money, promises, hostilities, etc. While in the physical world interactions and their associated forces have immediate dynamical consequences (Newton's law) the situation is not clear for human interactions. Here we study the acceleration between humans who interact through the exchange of messages, goods and hostilities in a massive multiplayer online game. For this game we have complete information about all interactions (exchange events) between about 1/2 million players, and about their trajectories (positions) in a metric space of the game universe at any point in time. We derive the interaction potentials for communication, trade and attacks and show that they are harmonic in nature. Individuals who exchange messages and trade goods generally attract each other and start to separate immediately after exchange events stop. The interaction potential for attacks mirrors the usual "hit-and-run" tactics of aggressive players. By measuring interaction intensities as a function of distance, velocity and acceleration, we show that "forces" between players are directly related to the number of exchange events. The power-law of the likelihood for interactions vs. distance is in accordance with previous real world empirical work. We show that the obtained potentials can be understood with a simple model assuming an exchange-driven force in combination with a distance dependent exchange rate.

preprint2014arXiv

Physiologically motivated multiplex Kuramoto model describes phase diagram of cortical activity

We derive a two-layer multiplex Kuramoto model from weakly coupled Wilson-Cowan oscillators on a cortical network with inhibitory synaptic time delays. Depending on the coupling strength and a phase shift parameter, related to cerebral blood flow and GABA concentration, respectively, we numerically identify three macroscopic phases: unsynchronized, synchronized, and chaotic dynamics. These correspond to physiological background-, epileptic seizure-, and resting-state cortical activity, respectively. We also observe frequency suppression at the transition from resting-state to seizure activity.

preprint2014arXiv

Spreading of diseases through comorbidity networks across life and gender

The state of health of patients is typically not characterized by a single disease alone but by multiple (comorbid) medical conditions. These comorbidities may depend strongly on age and gender. We propose a specific phenomenological comorbidity network of human diseases that is based on medical claims data of the entire population of Austria. The network is constructed from a two-layer multiplex network, where in one layer the links represent the conditional probability for a comorbidity, and in the other the links contain the respective statistical significance. We show that the network undergoes dramatic structural changes across the lifetime of patients.Disease networks for children consist of a single, strongly inter-connected cluster. During adolescence and adulthood further disease clusters emerge that are related to specific classes of diseases, such as circulatory, mental, or genitourinary disorders.For people above 65 these clusters start to merge and highly connected hubs dominate the network. These hubs are related to hypertension, chronic ischemic heart diseases, and chronic obstructive pulmonary diseases. We introduce a simple diffusion model to understand the spreading of diseases on the disease network at the population level. For the first time we are able to show that patients predominantly develop diseases which are in close network-proximity to disorders that they already suffer. The model explains more than 85 % of the variance of all disease incidents in the population. The presented methodology could be of importance for anticipating age-dependent disease-profiles for entire populations, and for validation and of prevention schemes.

preprint2014arXiv

The weak core and the structure of elites in social multiplex networks

Recent approaches on elite identification highlighted the important role of {\em intermediaries}, by means of a new definition of the core of a multiplex network, the {\em generalised} $K$-core. This newly introduced core subgraph crucially incorporates those individuals who, in spite of not being very connected, maintain the cohesiveness and plasticity of the core. Interestingly, it has been shown that the performance on elite identification of the generalised $K$-core is sensibly better that the standard $K$-core. Here we go further: Over a multiplex social system, we isolate the community structure of the generalised $K$-core and we identify the weakly connected regions acting as bridges between core communities, ensuring the cohesiveness and connectivity of the core region. This gluing region is the {\em Weak core} of the multiplex system. We test the suitability of our method on data from the society of 420.000 players of the Massive Multiplayer Online Game {\em Pardus}. Results show that the generalised $K$-core displays a clearly identifiable community structure and that the weak core gluing the core communities shows very low connectivity and clustering. Nonetheless, despite its low connectivity, the weak core forms a unique, cohesive structure. In addition, we find that members populating the weak core have the best scores on social performance, when compared to the other elements of the generalised $K$-core. The weak core provides a new angle on understanding the social structure of elites, highlighting those subgroups of individuals whose role is to glue different communities in the core.

preprint2014arXiv

To bail-out or to bail-in? Answers from an agent-based model

Since beginning of the 2008 financial crisis almost half a trillion euros have been spent to financially assist EU member states in taxpayer-funded bail-outs. These crisis resolutions are often accompanied by austerity programs causing political and social friction on both domestic and international levels. The question of how to resolve failing financial institutions under which economic preconditions is therefore a pressing and controversial issue of vast political importance. In this work we employ an agent-based model to study the economic and financial ramifications of three highly relevant crisis resolution mechanisms. To establish the validity of the model we show that it reproduces a series of key stylized facts if the financial and real economy. The distressed institution can either be closed via a purchase & assumption transaction, it can be bailed-out using taxpayer money, or it may be bailed-in in a debt-to-equity conversion. We find that for an economy characterized by low unemployment and high productivity the optimal crisis resolution with respect to financial stability and economic productivity is to close the distressed institution. For economies in recession with high unemployment the bail-in tool provides the most efficient crisis resolution mechanism. Under no circumstances do taxpayer-funded bail-out schemes outperform bail-ins with private sector involvement.

preprint2013arXiv

DebtRank-transparency: Controlling systemic risk in financial networks

Banks in the interbank network can not assess the true risks associated with lending to other banks in the network, unless they have full information on the riskiness of all the other banks. These risks can be estimated by using network metrics (for example DebtRank) of the interbank liability network which is available to Central Banks. With a simple agent based model we show that by increasing transparency by making the DebtRank of individual nodes (banks) visible to all nodes, and by imposing a simple incentive scheme, that reduces interbank borrowing from systemically risky nodes, the systemic risk in the financial network can be drastically reduced. This incentive scheme is an effective regulation mechanism, that does not reduce the efficiency of the financial network, but fosters a more homogeneous distribution of risk within the system in a self-organized critical way. We show that the reduction of systemic risk is to a large extent due to the massive reduction of cascading failures in the transparent system. An implementation of this minimal regulation scheme in real financial networks should be feasible from a technical point of view.

preprint2013arXiv

Generalized (c,d)-entropy and aging random walks

Complex systems are often inherently non-ergodic and non-Markovian for which Shannon entropy loses its applicability. In particular accelerating, path-dependent, and aging random walks offer an intuitive picture for these non-ergodic and non-Markovian systems. It was shown that the entropy of non-ergodic systems can still be derived from three of the Shannon-Khinchin axioms, and by violating the fourth -- the so-called composition axiom. The corresponding entropy is of the form $S_{c,d} \sim \sum_i Γ(1+d,1-c\ln p_i)$ and depends on two system-specific scaling exponents, $c$ and $d$. This entropy contains many recently proposed entropy functionals as special cases, including Shannon and Tsallis entropy. It was shown that this entropy is relevant for a special class of non-Markovian random walks. In this work we generalize these walks to a much wider class of stochastic systems that can be characterized as `aging' systems. These are systems whose transition rates between states are path- and time-dependent. We show that for particular aging walks $S_{c,d}$ is again the correct extensive entropy. Before the central part of the paper we review the concept of $(c,d)$-entropy in a self-contained way.

preprint2013arXiv

How women organize social networks different from men

Superpositions of social networks, such as communication, friendship, or trade networks, are called multiplex networks, forming the structural backbone of human societies. Novel datasets now allow quantification and exploration of multiplex networks. Here we study gender-specific differences of a multiplex network from a complete behavioral dataset of an online-game society of about 300,000 players. On the individual level females perform better economically and are less risk-taking than males. Males reciprocate friendship requests from females faster than vice versa and hesitate to reciprocate hostile actions of females. On the network level females have more communication partners, who are less connected than partners of males. We find a strong homophily effect for females and higher clustering coefficients of females in trade and attack networks. Cooperative links between males are under-represented, reflecting competition for resources among males. These results confirm quantitatively that females and males manage their social networks in substantially different ways.

preprint2013arXiv

Quantifying age- and gender-related diabetes comorbidity risks using nation-wide big claims data

Currently emerging "big data" techniques are reshaping medical science into a data science. Medical claims data allow assessing an entire nation's health state in a quantitative way, in particular with regard to the occurrences and consequences of chronic and pandemic diseases like diabetes. We develop a quantitative, statistical approach to test for associations between the incidence of type 1 or type 2 diabetes and any possible other disease as provided by the ICD10 diagnosis codes using a complete set of Austrian inpatient data. With a new co-occurrence analysis the relative risks for each possible comorbidity are studied as a function of patient age and gender, a temporal analysis investigates whether the onset of diabetes typically precedes or follows the onset of the other disease. The samples is always of maximal size, i.e. contains all patients with that comorbidity within the country. The present study is an equivalent of almost 40,000 studies, all with maximum patient number available. Out of more than thousand possible associations, 123 comorbid diseases for type 1 or type 2 diabetes are identified at high significance levels. Well known diabetic comorbidities are recovered, such as retinopathies, hypertension, chronic kidney diseases, etc. This validates the method. Additionally, a number of comorbidities are identified which have only been recognized to a lesser extent, for example epilepsy, sepsis, or mental disorders. The temporal evolution, age, and gender-dependence of these comorbidities are discussed. The new statistical-network methodology developed here can be readily applied to other chronic diseases.

preprint2013arXiv

Statistical detection of systematic election irregularities

Democratic societies are built around the principle of free and fair elections, that each citizen's vote should count equal. National elections can be regarded as large-scale social experiments, where people are grouped into usually large numbers of electoral districts and vote according to their preferences. The large number of samples implies certain statistical consequences for the polling results which can be used to identify election irregularities. Using a suitable data collapse, we find that vote distributions of elections with alleged fraud show a kurtosis of hundred times more than normal elections on certain levels of data aggregation. As an example we show that reported irregularities in recent Russian elections are indeed well explained by systematic ballot stuffing and develop a parametric model quantifying to which extent fraudulent mechanisms are present. We show that if specific statistical properties are present in an election, the results do not represent the will of the people. We formulate a parametric test detecting these statistical properties in election results. Remarkably, this technique produces similar outcomes irrespective of the data resolution and thus allows for cross-country comparisons.

preprint2013arXiv

The Transformation-Groupoid Structure of the q-Gaussian Family

The q-Gaussian function emerges naturally in various applications of statistical mechanics of non-ergodic and complex systems. In particular it was shown that in the theory of binary processes with correlations, the q-Gaussian can appear as a limiting distribution. Further, there exist several problems and situations where, depending on procedural or algorithmic details of data-processing, q-Gaussian distributions may yield distinct values of q, where one value is larger, the other smaller than one. To relate such pairs of q-Gaussians it would be convenient to map such distributions onto one another, ideally in a way, that any value of q can be mapped uniquely to any other value q'. So far a (duality) map from q -> q'=(7-5q)/(5-3q) was found, mapping q from the interval q\in [-\infty, 1] -> q'\in [1, 5/3]. Here we complete the theory of transformations of q-Gaussians by deriving a general map γ_{qq'}, that transforms normalizable q-Gaussian distributions onto one another for which q and q' are in the range of [1,3). By combining this with the previous result, a mapping from any value of q \in [-\infty,3) is possible to any other value q'\in [-\infty,3). We show that the action of γ_{qq'} on the set of q-Gaussian distributions is a transformation groupoid.

preprint2013arXiv

Triadic closure dynamics drives scaling-laws in social multiplex networks

Social networks exhibit scaling-laws for several structural characteristics, such as the degree distribution, the scaling of the attachment kernel, and the clustering coefficients as a function of node degree. A detailed understanding if and how these scaling laws are inter-related is missing so far, let alone whether they can be understood through a common, dynamical principle. We propose a simple model for stationary network formation and show that the three mentioned scaling relations follow as natural consequences of triadic closure. The validity of the model is tested on multiplex data from a well studied massive multiplayer online game. We find that the three scaling exponents observed in the multiplex data for the friendship, communication and trading networks can simultaneously be explained by the model. These results suggest that triadic closure could be identified as one of the fundamental dynamical principles in social multiplex network formation.

preprint2012arXiv

A self-organized model for cell-differentiation based on variations of molecular decay rates

Systemic properties of living cells are the result of molecular dynamics governed by so-called genetic regulatory networks (GRN). These networks capture all possible features of cells and are responsible for the immense levels of adaptation characteristic to living systems. At any point in time only small subsets of these networks are active. Any active subset of the GRN leads to the expression of particular sets of molecules (expression modes). The subsets of active networks change over time, leading to the observed complex dynamics of expression patterns. Understanding of this dynamics becomes increasingly important in systems biology and medicine. While the importance of transcription rates and catalytic interactions has been widely recognized in modeling genetic regulatory systems, the understanding of the role of degradation of biochemical agents (mRNA, protein) in regulatory dynamics remains limited. Recent experimental data suggests that there exists a functional relation between mRNA and protein decay rates and expression modes. In this paper we propose a model for the dynamics of successions of sequences of active subnetworks of the GRN. The model is able to reproduce key characteristics of molecular dynamics, including homeostasis, multi-stability, periodic dynamics, alternating activity, differentiability, and self-organized critical dynamics. Moreover the model allows to naturally understand the mechanism behind the relation between decay rates and expression modes. The model explains recent experimental observations that decay-rates (or turnovers) vary between differentiated tissue-classes at a general systemic level and highlights the role of intracellular decay rate control mechanisms in cell differentiation.

preprint2012arXiv

Bistatic scattering characterization of a three-dimensional broadband cloaking structure

Here we present the results of full experimental characterization of broadband cloaking of a finite-sized metallic cylinder at X-band. The cloaking effect is characterized by measuring the bistatic scattering patterns of uncloaked and cloaked objects in free space and then comparing these with each other. The results of the measurements demonstrate a broadband cloaking effect and are in good agreement with numerical predictions.

preprint2012arXiv

Generalized entropies and logarithms and their duality relations

For statistical systems that violate one of the four Shannon-Khinchin axioms, entropy takes a more general form than the Boltzmann-Gibbs entropy. The framework of superstatistics allows one to formulate a maximum entropy principle with these generalized entropies, making them useful for understanding distribution functions of non-Markovian or non-ergodic complex systems. For such systems where the composability axiom is violated there exist only two ways to implement the maximum entropy principle, one using escort probabilities, the other not. The two ways are connected through a duality. Here we show that this duality fixes a unique escort probability, which allows us to derive a complete theory of the generalized logarithms that naturally arise from the violation of this axiom. We then show how the functional forms of these generalized logarithms are related to the asymptotic scaling behavior of the entropy.

preprint2011arXiv

Emergence of good conduct, scaling and Zipf laws in human behavioral sequences in an online world

We study behavioral action sequences of players in a massive multiplayer online game. In their virtual life players use eight basic actions which allow them to interact with each other. These actions are communication, trade, establishing or breaking friendships and enmities, attack, and punishment. We measure the probabilities for these actions conditional on previous taken and received actions and find a dramatic increase of negative behavior immediately after receiving negative actions. Similarly, positive behavior is intensified by receiving positive actions. We observe a tendency towards anti-persistence in communication sequences. Classifying actions as positive (good) and negative (bad) allows us to define binary 'world lines' of lives of individuals. Positive and negative actions are persistent and occur in clusters, indicated by large scaling exponents alpha~0.87 of the mean square displacement of the world lines. For all eight action types we find strong signs for high levels of repetitiveness, especially for negative actions. We partition behavioral sequences into segments of length n (behavioral `words' and 'motifs') and study their statistical properties. We find two approximate power laws in the word ranking distribution, one with an exponent of kappa-1 for the ranks up to 100, and another with a lower exponent for higher ranks. The Shannon n-tuple redundancy yields large values and increases in terms of word length, further underscoring the non-trivial statistical properties of behavioral sequences. On the collective, societal level the timeseries of particular actions per day can be understood by a simple mean-reverting log-normal model.

preprint2011arXiv

Empirical confirmation of creative destruction from world trade data

We show that world trade network datasets contain empirical evidence that the dynamics of innovation in the world economy follows indeed the concept of creative destruction, as proposed by J.A. Schumpeter more than half a century ago. National economies can be viewed as complex, evolving systems, driven by a stream of appearance and disappearance of goods and services. Products appear in bursts of creative cascades. We find that products systematically tend to co-appear, and that product appearances lead to massive disappearance events of existing products in the following years. The opposite - disappearances followed by periods of appearances - is not observed. This is an empirical validation of the dominance of cascading competitive replacement events on the scale of national economies, i.e. creative destruction. We find a tendency that more complex products drive out less complex ones, i.e. progress has a direction. Finally we show that the growth trajectory of a country's product output diversity can be understood by a recently proposed evolutionary model of Schumpeterian economic dynamics.

preprint2011arXiv

Experimental characterization of a broadband transmission-line cloak in free space

The cloaking efficiency of a finite-size cylindrical transmission-line cloak operating in the X-band is verified with bistatic free space measurements. The cloak is designed and optimized with numerical full-wave simulations. The reduction of the total scattering width of a metal object, enabled by the cloak, is clearly observed from the bistatic free space measurements. The numerical and experimental results are compared resulting in good agreement with each other.

preprint2011arXiv

Generalized entropies and the transformation group of superstatistics

Superstatistics describes statistical systems that behave like superpositions of different inverse temperatures $β$, so that the probability distribution is $p(ε_i) \propto \int_{0}^{\infty} f(β) e^{-βε_i}dβ$, where the `kernel' $f(β)$ is nonnegative and normalized ($\int f(β)d β=1$). We discuss the relation between this distribution and the generalized entropic form $S=\sum_i s(p_i)$. The first three Shannon-Khinchin axioms are assumed to hold. It then turns out that for a given distribution there are two different ways to construct the entropy. One approach uses escort probabilities and the other does not; the question of which to use must be decided empirically. The two approaches are related by a duality. The thermodynamic properties of the system can be quite different for the two approaches. In that connection we present the transformation laws for the superstatistical distributions under macroscopic state changes. The transformation group is the Euclidean group in one dimension.

preprint2011arXiv

The blogosphere as an excitable social medium: Richter's and Omori's Law in media coverage

We study the dynamics of public media attention by monitoring the content of online blogs. Social and media events can be traced by the propagation of word frequencies of related keywords. Media events are classified as exogenous - where blogging activity is triggered by an external news item - or endogenous where word frequencies build up within a blogging community without external influences. We show that word occurrences show statistical similarities to earthquakes. The size distribution of media events follows a Gutenberg-Richter law, the dynamics of media attention before and after the media event follows Omori's law. We present further empirical evidence that for media events of endogenous origin the overall public reception of the event is correlated with the behavior of word frequencies at the beginning of the event, and is to a certain degree predictable. These results may imply that the process of opinion formation in a human society might be related to effects known from excitable media.

preprint2011arXiv

Understanding mobility in a social petri dish

Despite the recent availability of large data sets on human movements, a full understanding of the rules governing motion within social systems is still missing, due to incomplete information on the socio-economic factors and to often limited spatio-temporal resolutions. Here we study an entire society of individuals, the players of an online-game, with complete information on their movements in a network-shaped universe and on their social and economic interactions. Such a "socio-economic laboratory" allows to unveil the intricate interplay of spatial constraints, social and economic factors, and patterns of mobility. We find that the motion of individuals is not only constrained by physical distances, but also strongly shaped by the presence of socio-economic areas. These regions can be recovered perfectly by community detection methods solely based on the measured human dynamics. Moreover, we uncover that long-term memory in the time-order of visited locations is the essential ingredient for modeling the trajectories.

preprint2011arXiv

What do generalized entropies look like? An axiomatic approach for complex, non-ergodic systems

Shannon and Khinchin showed that assuming four information theoretic axioms the entropy must be of Boltzmann-Gibbs type, $S=-\sum_i p_i \log p_i$. Here we note that in physical systems one of these axioms may be violated. For non-ergodic systems the so called separation axiom (Shannon-Khinchin axiom 4) will in general not be valid. We show that when this axiom is violated the entropy takes a more general form, $S_{c,d}\propto \sum_i ^W Γ(d+1, 1- c \log p_i)$, where $c$ and $d$ are scaling exponents and $Γ(a,b)$ is the incomplete gamma function. The exponents $(c,d)$ define equivalence classes for all interacting and non interacting systems and unambiguously characterize any statistical system in its thermodynamic limit. The proof is possible because of two newly discovered scaling laws which any entropic form has to fulfill, if the first three Shannon-Khinchin axioms hold. $(c,d)$ can be used to define equivalence classes of statistical systems. A series of known entropies can be classified in terms of these equivalence classes. We show that the corresponding distribution functions are special forms of Lambert-${\cal W}$ exponentials containing -- as special cases -- Boltzmann, stretched exponential and Tsallis distributions (power-laws). In the derivation we assume trace form entropies, $S=\sum_i g(p_i)$, with $g$ some function, however more general entropic forms can be classified along the same scaling analysis.

preprint2011arXiv

When do generalized entropies apply? How phase space volume determines entropy

We show how the dependence of phase space volume $Ω(N)$ of a classical system on its size $N$ uniquely determines its extensive entropy. We give a concise criterion when this entropy is not of Boltzmann-Gibbs type but has to assume a {\em generalized} (non-additive) form. We show that generalized entropies can only exist when the dynamically (statistically) relevant fraction of degrees of freedom in the system vanishes in the thermodynamic limit. These are systems where the bulk of the degrees of freedom is frozen and is practically statistically inactive. Systems governed by generalized entropies are therefore systems whose phase space volume effectively collapses to a lower-dimensional 'surface'. We explicitly illustrate the situation for binomial processes and argue that generalized entropies could be relevant for self organized critical systems such as sand piles, for spin systems which form meta-structures such as vortices, domains, instantons, etc., and for problems associated with anomalous diffusion.

preprint2010arXiv

A comprehensive classification of complex statistical systems and an ab-initio derivation of their entropy and distribution functions

To characterize strongly interacting statistical systems within a thermodynamical framework - complex systems in particular - it might be necessary to introduce generalized entropies, $S_g$. A series of such entropies have been proposed in the past, mainly to accommodate important empirical distribution functions to a maximum ignorance principle. Until now the understanding of the fundamental origin of these entropies and its deeper relations to complex systems is limited. Here we explore this questions from first principles. We start by observing that the 4th Khinchin axiom (separability axiom) is violated by strongly interacting systems in general and ask about the consequences of violating the 4th axiom while assuming the first three Khinchin axioms (K1-K3) to hold and $S_g=\sum_ig(p_i)$. We prove by simple scaling arguments that under these requirements {\em each} statistical system is uniquely characterized by a distinct pair of scaling exponents $(c,d)$ in the large size limit. The exponents define equivalence classes for all interacting and non interacting systems. This allows to derive a unique entropy, $S_{c,d}\propto \sum_i Γ(d+1, 1- c \ln p_i)$, which covers all entropies which respect K1-K3 and can be written as $S_g=\sum_ig(p_i)$. Known entropies can now be classified within these equivalence classes. The corresponding distribution functions are special forms of Lambert-$W$ exponentials containing as special cases Boltzmann, stretched exponential and Tsallis distributions (power-laws) -- all widely abundant in nature. This is, to our knowledge, the first {\em ab initio} justification for the existence of generalized entropies. Even though here we assume $S_g=\sum_ig(p_i)$, we show that more general entropic forms can be classified along the same lines.

preprint2010arXiv

Leverage Causes Fat Tails and Clustered Volatility

We build a simple model of leveraged asset purchases with margin calls. Investment funds use what is perhaps the most basic financial strategy, called "value investing", i.e. systematically attempting to buy underpriced assets. When funds do not borrow, the price fluctuations of the asset are normally distributed and uncorrelated across time. All this changes when the funds are allowed to leverage, i.e. borrow from a bank, to purchase more assets than their wealth would otherwise permit. During good times competition drives investors to funds that use more leverage, because they have higher profits. As leverage increases price fluctuations become heavy tailed and display clustered volatility, similar to what is observed in real markets. Previous explanations of fat tails and clustered volatility depended on "irrational behavior", such as trend following. Here instead this comes from the fact that leverage limits cause funds to sell into a falling market: A prudent bank makes itself locally safer by putting a limit to leverage, so when a fund exceeds its leverage limit, it must partially repay its loan by selling the asset. Unfortunately this sometimes happens to all the funds simultaneously when the price is already falling. The resulting nonlinear feedback amplifies large downward price movements. At the extreme this causes crashes, but the effect is seen at every time scale, producing a power law of price disturbances. A standard (supposedly more sophisticated) risk control policy in which individual banks base leverage limits on volatility causes leverage to rise during periods of low volatility, and to contract more quickly when volatility gets high, making these extreme fluctuations even worse.

preprint2010arXiv

Living on the edge of chaos: minimally nonlinear models of genetic regulatory dynamics

Linearized catalytic reaction equations modeling e.g. the dynamics of genetic regulatory networks under the constraint that expression levels, i.e. molecular concentrations of nucleic material are positive, exhibit nontrivial dynamical properties, which depend on the average connectivity of the reaction network. In these systems the inflation of the edge of chaos and multi-stability have been demonstrated to exist. The positivity constraint introduces a nonlinearity which makes chaotic dynamics possible. Despite the simplicity of such minimally nonlinear systems, their basic properties allow to understand fundamental dynamical properties of complex biological reaction networks. We analyze the Lyapunov spectrum, determine the probability to find stationary oscillating solutions, demonstrate the effect of the nonlinearity on the effective in- and out-degree of the active interaction network and study how the frequency distributions of oscillatory modes of such system depend on the average connectivity.

preprint2010arXiv

Peer-review in a world with rational scientists: Toward selection of the average

One of the virtues of peer review is that it provides a self-regulating selection mechanism for scientific work, papers and projects. Peer review as a selection mechanism is hard to evaluate in terms of its efficiency. Serious efforts to understand its strengths and weaknesses have not yet lead to clear answers. In theory peer review works if the involved parties (editors and referees) conform to a set of requirements, such as love for high quality science, objectiveness, and absence of biases, nepotism, friend and clique networks, selfishness, etc. If these requirements are violated, what is the effect on the selection of high quality work? We study this question with a simple agent based model. In particular we are interested in the effects of rational referees, who might not have any incentive to see high quality work other than their own published or promoted. We find that a small fraction of incorrect (selfish or rational) referees can drastically reduce the quality of the published (accepted) scientific standard. We quantify the fraction for which peer review will no longer select better than pure chance. Decline of quality of accepted scientific work is shown as a function of the fraction of rational and unqualified referees. We show how a simple quality-increasing policy of e.g. a journal can lead to a loss in overall scientific quality, and how mutual support-networks of authors and referees deteriorate the system.

preprint2009arXiv

Evolutionary dynamics from a variational principle

We demonstrate with a thought experiment that fitness-based population dynamical approaches to evolution are not able to make quantitative, falsifiable predictions about the long-term behavior of evolutionary systems. A key characteristic of evolutionary systems is the ongoing endogenous production of new species. These novel entities change the conditions for already existing species. Even {\em Darwin's Demon}, a hypothetical entity with exact knowledge of the abundance of all species and their fitness functions at a given time, could not pre-state the impact of these novelties on established populations. We argue that fitness is always {\it a posteriori} knowledge -- it measures but does not explain why a species has reproductive success or not. To overcome these conceptual limitations, a variational principle is proposed in a spin-model-like setup of evolutionary systems. We derive a functional which is minimized under the most general evolutionary formulation of a dynamical system, i.e. evolutionary trajectories causally emerge as a minimization of a functional. This functional allows the derivation of analytic solutions of the asymptotic diversity for stochastic evolutionary systems within a mean-field approximation. We test these approximations by numerical simulations of the corresponding model and find good agreement in the position of phase transitions in diversity curves. The model is further able to reproduce stylized facts of timeseries from several man-made and natural evolutionary systems. Light will be thrown on how species and their fitness landscapes dynamically co-evolve.

preprint2009arXiv

Limit distributions of scale-invariant probabilistic models of correlated random variables with the q-Gaussian as an explicit example

Extremization of the Boltzmann-Gibbs (BG) entropy under appropriate norm and width constraints yields the Gaussian distribution. Also, the basic solutions of the standard Fokker-Planck (FP) equation (related to the Langevin equation with additive noise), as well as the Central Limit Theorem attractors, are Gaussians. The simplest stochastic model with such features is N to infinity independent binary random variables, as first proved by de Moivre and Laplace. What happens for strongly correlated random variables? Such correlations are often present in physical situations as e.g. systems with long range interactions or memory. Frequently q-Gaussians become observed. This is typically so if the Langevin equation includes multiplicative noise, or the FP equation to be nonlinear. Scale-invariance, i.e. exchangeable binary stochastic processes, allow a systematical analysis of the relation between correlations and non-Gaussian distributions. In particular, a generalized stochastic model yielding q-Gaussians for all q (including q>1) was missing. This is achieved here by using the Laplace-de Finetti representation theorem, which embodies strict scale-invariance of interchangeable random variables. We demonstrate that strict scale invariance together with q-Gaussianity mandates the associated extensive entropy to be BG.

preprint2009arXiv

Measuring social dynamics in a massive multiplayer online game

Quantification of human group-behavior has so far defied an empirical, falsifiable approach. This is due to tremendous difficulties in data acquisition of social systems. Massive multiplayer online games (MMOG) provide a fascinating new way of observing hundreds of thousands of simultaneously socially interacting individuals engaged in virtual economic activities. We have compiled a data set consisting of practically all actions of all players over a period of three years from a MMOG played by 300,000 people. This large-scale data set of a socio-economic unit contains all social and economic data from a single and coherent source. Players have to generate a virtual income through economic activities to `survive' and are typically engaged in a multitude of social activities offered within the game. Our analysis of high-frequency log files focuses on three types of social networks, and tests a series of social-dynamics hypotheses. In particular we study the structure and dynamics of friend-, enemy- and communication networks. We find striking differences in topological structure between positive (friend) and negative (enemy) tie networks. All networks confirm the recently observed phenomenon of network densification. We propose two approximate social laws in communication networks, the first expressing betweenness centrality as the inverse square of the overlap, the second relating communication strength to the cube of the overlap. These empirical laws provide strong quantitative evidence for the Weak ties hypothesis of Granovetter. Further, the analysis of triad significance profiles validates well-established assertions from social balance theory. We find overrepresentation (underrepresentation) of complete (incomplete) triads in networks of positive ties, and vice versa for networks of negative ties...

preprint2009arXiv

Schumpeterian economic dynamics as a quantifiable minimum model of evolution

We propose a simple quantitative model of Schumpeterian economic dynamics. New goods and services are endogenously produced through combinations of existing goods. As soon as new goods enter the market they may compete against already existing goods, in other words new products can have destructive effects on existing goods. As a result of this competition mechanism existing goods may be driven out from the market - often causing cascades of secondary defects (Schumpeterian gales of destruction). The model leads to a generic dynamics characterized by phases of relative economic stability followed by phases of massive restructuring of markets - which could be interpreted as Schumpeterian business `cycles'. Model timeseries of product diversity and productivity reproduce several stylized facts of economics timeseries on long timescales such as GDP or business failures, including non-Gaussian fat tailed distributions, volatility clustering etc. The model is phrased in an open, non-equilibrium setup which can be understood as a self organized critical system. Its diversity dynamics can be understood by the time-varying topology of the active production networks.

preprint2008arXiv

To how many politicians should government be left?

The quality of governance of institutions, corporations and countries depends on the ability of efficient decision making within the respective boards or cabinets. Opinion formation processes within groups are size dependent. It is often argued - as now e.g. in the discussion of the future size of the European Commission - that decision making bodies of a size beyond 20 become strongly inefficient. We report empirical evidence that the performance of national governments declines with increasing membership and undergoes a qualitative change in behavior at a particular group size. We use recent UNDP, World Bank and CIA data on overall government efficacy, i.e. stability, the quality of policy formulation as well as human development indices of individual countries and relate it to the country's cabinet size. We are able to understand our findings through a simple physical model of opinion dynamics in groups.

preprint2006arXiv

Nonextensive statistical mechanics and complex scale-free networks

One explanation for the impressive recent boom in network theory might be that it provides a promising tool for an understanding of complex systems. Network theory is mainly focusing on discrete large-scale topological structures rather than on microscopic details of interactions of its elements. This viewpoint allows to naturally treat collective phenomena which are often an integral part of complex systems, such as biological or socio-economical phenomena. Much of the attraction of network theory arises from the discovery that many networks, natural or man-made, seem to exhibit some sort of universality, meaning that most of them belong to one of three classes: {\it random}, {\it scale-free} and {\it small-world} networks. Maybe most important however for the physics community is, that due to its conceptually intuitive nature, network theory seems to be within reach of a full and coherent understanding from first principles ...

preprint2006arXiv

Transport on Complex Networks: Flow, Jamming and Optimization

Many transport processes on networks depend crucially on the underlying network geometry, although the exact relationship between the structure of the network and the properties of transport processes remain elusive. In this paper we address this question by using numerical models in which both structure and dynamics are controlled systematically. We consider the traffic of information packets that include driving, searching and queuing. We present the results of extensive simulations on two classes of networks; a correlated cyclic scale-free network and an uncorrelated homogeneous weakly clustered network. By measuring different dynamical variables in the free flow regime we show how the global statistical properties of the transport are related to the temporal fluctuations at individual nodes (the traffic noise) and the links (the traffic flow). We then demonstrate that these two network classes appear as representative topologies for optimal traffic flow in the regimes of low density and high density traffic, respectively. We also determine statistical indicators of the pre-jamming regime on different network geometries and discuss the role of queuing and dynamical betweenness for the traffic congestion. The transition to the jammed traffic regime at a critical posting rate on different network topologies is studied as a phase transition with an appropriate order parameter. We also address several open theoretical problems related to the network dynamics.

preprint2005arXiv

Statistical Indicators of Collective Behavior and Functional Clusters in Gene Networks of Yeast

We analyze gene expression time-series data of yeast S. cerevisiae measured along two full cell-cycles. We quantify these data by using q-exponentials, gene expression ranking and a temporal mean-variance analysis. We construct gene interaction networks based on correlation coefficients and study the formation of the corresponding giant components and minimum spanning trees. By coloring genes according to their cell function we find functional clusters in the correlation networks and functional branches in the associated trees. Our results suggest that a percolation point of functional clusters can be identified on these gene expression correlation networks.

preprint2005arXiv

Statistical mechanics of scale-free networks at a critical point: Complexity without irreversibility?

Based on a rigorous extension of classical statistical mechanics to networks, we study a specific microscopic network Hamiltonian. The form of this Hamiltonian is derived from the assumption that individual nodes increase/decrease their utility by linking to nodes with a higher/lower degree than their own. We interpret utility as an equivalent to energy in physical systems and discuss the temperature dependence of the emerging networks. We observe the existence of a critical temperature $T_c$ where total energy (utility) and network-architecture undergo radical changes. Along this topological transition we obtain scale-free networks with complex hierarchical topology. In contrast to models for scale-free networks introduced so far, the scale-free nature emerges within equilibrium, with a clearly defined microcanonical ensemble and the principle of detailed balance strictly fulfilled. This provides clear evidence that 'complex' networks may arise without irreversibility. The results presented here should find a wide variety of applications in socio-economic statistical systems.

preprint2003arXiv

Information Super-Diffusion on Structured Networks

We study diffusion of information packets on several classes of structured networks. Packets diffuse from a randomly chosen node to a specified destination in the network. As local transport rules we consider random diffusion and an improved local search method. Numerical simulations are performed in the regime of stationary workloads away from the jamming transition. We find that graph topology determines the properties of diffusion in a universal way, which is reflected by power-laws in the transit-time and velocity distributions of packets. With the use of multifractal scaling analysis and arguments of non-extensive statistics we find that these power-laws are compatible with super-diffusive traffic for random diffusion and for improved local search. We are able to quantify the role of network topology on overall transport efficiency. Further, we demonstrate the implications of improved transport rules and discuss the importance of matching (global) topology with (local) transport rules for the optimal function of networks. The presented model should be applicable to a wide range of phenomena ranging from Internet traffic to protein transport along the cytoskeleton in biological cells.