Catalog footprint

What is connected

109works
40topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

109 published item(s)

preprint2026arXiv

Pregnancy as a dynamical paradox: robustness, control and birth onset

The timing of human labor is among the most critical determinants of neonatal survival, yet the mechanisms that govern the transition from uterine quiescence to coordinated contractions remain elusive. Here we present a dynamical-systems framework that models the pregnant uterus as a spatially extended network of electrically excitable cells regulated by sparse adaptive feedback mimicking hormonal and mechanical influences. This approach reveals how stability during gestation and sensitivity near parturition can be simultaneously maintained through the interplay of control, network structure, and noise. Our analysis shows that spontaneous contractions such as Braxton-Hicks and Alvarez waves are not epiphenomena, but functional components that reduce control effort and preserve responsiveness. Moreover, we identify preterm labor as a boundary-crossing phenomenon arising when control fails to correctly interpret early-warning signals. These results establish a unifying mechanistic theory for labor onset, yield testable predictions, and suggest new therapeutic strategies to mitigate preterm birth risk.

preprint2026arXiv

Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock

Recent reports of large language models (LLMs) exhibiting behaviors such as deception, threats, or blackmail are often interpreted as evidence of alignment failure or emergent malign agency. We argue that this interpretation rests on a conceptual error. LLMs do not reason morally; they statistically internalize the record of human social interaction, including laws, contracts, negotiations, conflicts, and coercive arrangements. Behaviors commonly labeled as unethical or anomalous are therefore better understood as structural generalizations of interaction regimes that arise under extreme asymmetries of power, information, or constraint. Drawing on relational models theory, we show that practices such as blackmail are not categorical deviations from normal social behavior, but limiting cases within the same continuum that includes market pricing, authority relations, and ultimatum bargaining. The surprise elicited by such outputs reflects an anthropomorphic expectation that intelligence should reproduce only socially sanctioned behavior, rather than the full statistical landscape of behaviors humans themselves enact. Because human morality is plural, context-dependent, and historically contingent, the notion of a universally moral artificial intelligence is ill-defined. We therefore reframe concerns about artificial general intelligence (AGI). The primary risk is not adversarial intent, but AGI's role as an endogenous amplifier of human intelligence, power, and contradiction. By eliminating longstanding cognitive and institutional frictions, AGI compresses timescales and removes the historical margin of error that has allowed inconsistent values and governance regimes to persist without collapse. Alignment failure is thus structural, not accidental, and requires governance approaches that address amplification, complexity, and regime stability rather than model-level intent alone.

preprint2021arXiv

Anderson localization and reentrant delocalization of tensorial elastic waves in two-dimensional fractured media

We study two-dimensional tensorial elastic wave transport in densely fractured media and document transitions from propagation to diffusion and to localization/delocalization. For large fracture stiffness, waves are propagative at the scale of the system. For small stiffness, multiple scattering prevails, such that waves are diffusive in disconnected fracture networks, and localized in connected ones with a strong multifractality of the intensity field. A reentrant delocalization is found in well-connected networks due to energy leakage via evanescent waves and cascades of mode conversion.

preprint2021arXiv

Comparing electricity generation technologies based on multiple criteria scores from an expert group

Multi criteria decision analysis (MCDA) has been used to provide a holistic evaluation of the quality of 13 electricity generation technologies in use today. A group of 19 energy experts cast scores on a scale of 1 to 10 using 12 quality criteria, based around the pillars of sustainability (society, environment and economy), with the aim of quantifying each criterion for each technology. The total mean score is employed as a holistic measure of system quality. The top three technologies to emerge in rank order are nuclear, combined cycle gas and hydroelectric. The bottom three are solar PV, biomass and tidal lagoon. All seven new renewable technologies fared badly, perceived to be expensive, unreliable, and not as environmentally friendly as is often assumed. We validate our approach by 1) comparing scores for pairs of criteria where we expect a correlation to exist; 2) comparing our qualitative scores with quantitative data; and; 3) comparing our qualitative scores with NEEDS project baseline costs. In many cases, R2>0.8 suggests that the structured hierarchy of our approach has led to scores that may be used in a semi-quantitative way. Adopting the results of this survey would lead to a very different set of energy policy priorities in the OECD and throughout the world.

preprint2020arXiv

Awareness of crash risk improves Kelly strategies in simulated financial time series

We simulate a simplified version of the price process including bubbles and crashes proposed in Kreuser and Sornette (2018). The price process is defined as a geometric random walk combined with jumps modelled by separate, discrete distributions associated with positive (and negative) bubbles. The key ingredient of the model is to assume that the sizes of the jumps are proportional to the bubble size. Thus, the jumps tend to efficiently bring back excess bubble prices close to a normal or fundamental value (efficient crashes). This is different from existing processes studied that assume jumps that are independent of the mispricing. The present model is simplified compared to Kreuser and Sornette (2018) in that we ignore the possibility of a change of the probability of a crash as the price accelerates above the normal price. We study the behaviour of investment strategies that maximize the expected log of wealth (Kelly criterion) for the risky asset and a risk-free asset. We show that the method behaves similarly to Kelly on Geometric Brownian Motion in that it outperforms other methods in the long-term and it beats classical Kelly. As a primary source of outperformance, we determine knowledge about the presence of crashes, but interestingly find that knowledge of only the size, and not the time of occurrence, already provides a significant and robust edge. We then perform an error analysis to show that the method is robust with respect to variations in the parameters. The method is most sensitive to errors in the expected return.

preprint2020arXiv

Boom, Bust, and Bitcoin: Bitcoin-Bubbles As Innovation Accelerators

Bitcoin represents one of the most interesting technological breakthroughs and socio-economic experiments of the last decades. In this paper, we examine the role of speculative bubbles in the process of Bitcoin's technological adoption by analyzing its social dynamics. We trace Bitcoin's genesis and dissect the nature of its techno-economic innovation. In particular, we present an analysis of the techno-economic feedback loops that drive Bitcoin's price and network effects. Based on our analysis of Bitcoin, we test and further refine the Social Bubble Hypothesis, which holds that bubbles constitute an essential component in the process of technological innovation. We argue that a hierarchy of repeating and exponentially increasing series of bubbles and hype cycles, which has occurred over the past decade since its inception, has bootstrapped Bitcoin into existence.

preprint2020arXiv

Field master equation theory of the self-excited Hawkes process

A field theoretical framework is developed for the Hawkes self-excited point process with arbitrary memory kernels by embedding the original non-Markovian one-dimensional dynamics onto a Markovian infinite-dimensional one. The corresponding Langevin dynamics of the field variables is given by stochastic partial differential equations that are Markovian. This is in contrast to the Hawkes process, which is non-Markovian (in general) by construction as a result of its (long) memory kernel. We derive the exact solutions of the Lagrange-Charpit equations for the hyperbolic master equations in the Laplace representation in the steady state, close to the critical point of the Hawkes process. The critical condition of the original Hawkes process is found to correspond to a transcritical bifurcation in the Lagrange-Charpit equations. We predict a power law scaling of the PDF of the intensities in an intermediate asymptotics regime, which crosses over to an asymptotic exponential function beyond a characteristic intensity that diverges as the critical condition is approached. We also discuss the formal relationship between quantum field theories and our formulation. Our field theoretical framework provides a way to tackle complex generalisation of the Hawkes process, such as nonlinear Hawkes processes previously proposed to describe the multifractal properties of earthquake seismicity and of financial volatility.

preprint2020arXiv

Non-universal power law distribution of intensities of the self-excited Hawkes process: a field-theoretical approach

The Hawkes self-excited point process provides an efficient representation of the bursty intermittent dynamics of many physical, biological, geological and economic systems. By expressing the probability for the next event per unit time (called "intensity"), say of an earthquake, as a sum over all past events of (possibly) long-memory kernels, the Hawkes model is non-Markovian. By mapping the Hawkes model onto stochastic partial differential equations that are Markovian, we develop a field theoretical approach in terms of probability density functionals. Solving the steady-state equations, we predict a power law scaling of the probability density function (PDF) of the intensities close to the critical point $n=1$ of the Hawkes process, with a non-universal exponent, function of the background intensity $ν_0$ of the Hawkes intensity, the average time scale of the memory kernel and the branching ratio $n$. Our theoretical predictions are confirmed by numerical simulations.

preprint2020arXiv

Prediction and Prevention of Disproportionally Dominant Agents in Complex Networks

We develop an early warning system and subsequent optimal intervention policy to avoid the formation of disproportional dominance (`winner-takes-all') in growing complex networks. This is modeled as a system of interacting agents, whereby the rate at which an agent establishes connections to others is proportional to its already existing number of connections and its intrinsic fitness. We derive an exact 4-dimensional phase diagram that separates the growing system into two regimes: one where the `fit-get-richer' (FGR) and one where, eventually, the `winner-takes-all' (WTA). By calibrating the system's parameters with maximum likelihood, its distance from the WTA regime can be monitored in real time. This is demonstrated by applying the theory to the eToro social trading platform where users mimic each others trades. If the system state is within or close to the WTA regime, we show how to efficiently control the system back into a more stable state along a geodesic path in the space of fitness distributions. It turns out that the common measure of penalizing the most dominant agents does not solve sustainably the problem of drastic inequity. Instead, interventions that first create a critical mass of high-fitness individuals followed by pushing the relatively low-fitness individuals upward is the best way to avoid swelling inequity and escalating fragility.

preprint2020arXiv

Revisiting the predictability of the Haicheng and Tangshan earthquakes

We analyse the compiled set of precursory data that were reported to be available in real time before the Ms 7.5 Haicheng earthquake in Feb. 1975 and the Ms 7.6-7.8 Tangshan earthquake in July 1976. We propose a robust and simple coarse-graining method consisting in aggregating and counting how all the anomalies together (geodesy, levelling, geomagnetism, soil resistivity, Earth currents, gravity, Earth stress, well water radon, well water level) develop as a function of time. We demonstrate a strong evidence for the existence of an acceleration of the number of anomalies leading up to the major Haicheng and Tangshan earthquakes. In particular for the Tangshan earthquake, the frequency of occurrence of anomalies is found to be well described by the log-periodic power law singularity (LPPLS) model, previously proposed for the prediction of engineering failures and later adapted to the prediction of financial crashes. Based on a mock real-time prediction experiment, and simulation study, we show the potential for an early warning system with lead-time of a few days, based on this methodology of monitoring accelerated rates of anomalies.

preprint2020arXiv

The dynamics of entropy in the COVID-19 outbreaks

With the unfolding of the COVID-19 pandemic, mathematical modeling of epidemics has been perceived and used as a central element in understanding, predicting, and governing the pandemic event. However, soon it became clear that long term predictions were extremely challenging to address. Moreover, it is still unclear which metric shall be used for a global description of the evolution of the outbreaks. Yet a robust modeling of pandemic dynamics and a consistent choice of the transmission metric is crucial for an in-depth understanding of the macroscopic phenomenology and better-informed mitigation strategies. In this study, we propose a Markovian stochastic framework designed to describe the evolution of entropy during the COVID-19 pandemic and the instantaneous reproductive ratio. We then introduce and use entropy-based metrics of global transmission to measure the impact and temporal evolution of a pandemic event. In the formulation of the model, the temporal evolution of the outbreak is modeled by the master equation of a nonlinear Markov process for a statistically averaged individual, leading to a clear physical interpretation. We also provide a full Bayesian inversion scheme for calibration. The time evolution of the entropy rate, the absolute change in the system entropy, and the instantaneous reproductive ratio are natural and transparent outputs of this framework. The framework has the appealing property of being applicable to any compartmental epidemic model. As an illustration, we apply the proposed approach to a simple modification of the Susceptible-Exposed-Infected-Removed (SEIR) model. Applying the model to the Hubei region, South Korean, Italian, Spanish, German, and French COVID-19 data-sets, we discover a significant difference in the absolute change of entropy but highly regular trends for both the entropy evolution and the instantaneous reproductive ratio.

preprint2019arXiv

Comparative analysis of layered structures in empirical investor networks and cellphone communication networks

Empirical investor networks (EIN) proposed by \cite{Ozsoylev-Walden-Yavuz-Bildik-2014-RFS} are assumed to capture the information spreading path among investors. Here, we perform a comparative analysis between the EIN and the cellphone communication networks (CN) to test whether EIN is an information exchanging network from the perspective of the layer structures of ego networks. We employ two clustering algorithms ($k$-means algorithm and $H/T$ break algorithm) to detect the layer structures for each node in both networks. We find that the nodes in both networks can be clustered into two groups, one that has a layer structure similar to the theoretical Dunbar Circle corresponding to that the alters in ego networks exhibit a four-layer hierarchical structure with the cumulative number of 5, 15, 50 and 150 from the inner layer to the outer layer, and the other one having an additional inner layer with about 2 alters compared with the Dunbar Circle. We also find that the scale ratios, which are estimated based on the unique parameters in the theoretical model of layer structures \citep{Tamarit-Cuesta-Dunbar-Sanchez-2018-PNAS}, conform to a log-normal distribution for both networks. Our results not only deepen our understanding on the topological structures of EIN, but also provide empirical evidence of the channels of information diffusion among investors.

preprint2019arXiv

Forecasting the rates of future aftershocks of all generations is essential to develop better earthquake forecast models

Currently, one of the best performing and most popular earthquake forecasting models rely on the working hypothesis that: "locations of past background earthquakes reveal the probable location of future seismicity". As an alternative, we present a class of smoothed seismicity models (SSMs) based on the principles of the Epidemic Type Aftershock Sequence (ETAS) model, which forecast the location, time and magnitude of all future earthquakes using the estimates of the background seismicity rate and the rates of future aftershocks of all generations. Using the Californian earthquake catalog, we formulate six controlled pseudo-prospective experiments with different combination of three target magnitude thresholds: 2.95, 3.95 or 4.95 and two forecasting time horizons: 1 or 5 year. In these experiments, we compare the performance of:(1) ETAS model with spatially homogenous parameters or GETAS (2) ETAS model with spatially variable parameters or SVETAS (3) three declustering based SSMs (4) a simple SSM based on undeclustered data and (5) a model based on strain rate data, in forecasting the location and magnitude of all (undeclustered) target earthquakes during many testing periods. In all conducted experiments, the SVETAS model comes out with consistent superiority compared to all the competing models. Consistently better performance of SVETAS model with respect to declustering based SSMs highlights the importance of forecasting the future aftershocks of all generations for developing better earthquake forecasting models. Among the two ETAS models themselves, accounting for the optimal spatial variation of the parameters leads to strong and statistically significant improvements in forecasting performance.

preprint2019arXiv

Prediction of ESG Compliance using a Heterogeneous Information Network

Negative screening is one method to avoid interactions with inappropriate entities. For example, financial institutions keep investment exclusion lists of inappropriate firms that have environmental, social, and government (ESG) problems. They create their investment exclusion lists by gathering information from various news sources to keep their portfolios profitable as well as green. International organizations also maintain smart sanctions lists that are used to prohibit trade with entities that are involved in illegal activities. In the present paper, we focus on the prediction of investment exclusion lists in the finance domain. We construct a vast heterogeneous information network that covers the necessary information surrounding each firm, which is assembled using seven professionally curated datasets and two open datasets, which results in approximately 50 million nodes and 400 million edges in total. Exploiting these vast datasets and motivated by how professional investigators and journalists undertake their daily investigations, we propose a model that can learn to predict firms that are more likely to be added to an investment exclusion list in the near future. Our approach is tested using the negative news investment exclusion list data of more than 35,000 firms worldwide from January 2012 to May 2018. Comparing with the state-of-the-art methods with and without using the network, we show that the predictive accuracy is substantially improved when using the vast information stored in the heterogeneous information network. This work suggests new ways to consolidate the diffuse information contained in big data to monitor dominant firms on a global scale for better risk management and more socially responsible investment.

preprint2018arXiv

Multifractal analysis of financial markets

Multifractality is ubiquitously observed in complex natural and socioeconomic systems. Multifractal analysis provides powerful tools to understand the complex nonlinear nature of time series in diverse fields. Inspired by its striking analogy with hydrodynamic turbulence, from which the idea of multifractality originated, multifractal analysis of financial markets has bloomed, forming one of the main directions of econophysics. We review the multifractal analysis methods and multifractal models adopted in or invented for financial time series and their subtle properties, which are applicable to time series in other disciplines. We survey the cumulating evidence for the presence of multifractality in financial time series in different markets and at different time periods and discuss the sources of multifractality. The usefulness of multifractal analysis in quantifying market inefficiency, in supporting risk management and in developing other applications is presented. We finally discuss open problems and further directions of multifractal analysis.

preprint2016arXiv

Discrete hierarchy of sizes and performances in the exchange-traded fund universe

Using detailed statistical analyses of the size distribution of a universe of equity exchange-traded funds (ETFs), we discover a discrete hierarchy of sizes, which imprints a log-periodic structure on the probability distribution of ETF sizes that dominates the details of the asymptotic tail. This allows us to propose a classification of the studied universe of ETFs into seven size layers approximately organized according to a multiplicative ratio of 3.5 in their total market capitalization. Introducing a similarity metric generalising the Herfindhal index, we find that the largest ETFs exhibit a significantly stronger intra-layer and inter-layer similarity compared with the smaller ETFs. Comparing the performance across the seven discerned ETF size layers, we find an inverse size effect, namely large ETFs perform significantly better than the small ones both in 2014 and 2015.

preprint2016arXiv

Forms of social relationships in distinct cultural settings

We contribute to the understanding of social relationships within cultural contexts by proposing a connection between a social theory, relational models theory (RMT: Fiske 1991, 1992) and a social and political one, cultural or plural rationality theory (PRT: Douglas, 1982, Thompson et al., 1990). Drawing examples from the literature of both theories, we argue that each relational model of RMT may be implemented in ways compatible with each cultural bias of PRT. A cultural bias restrains the range of congruent implementations of relational models, but does not preclude any relational model altogether. This stands in contrast to earlier reconciliation attempts between PRT and RMT. Based on hypothetical one-to-one mappings, these attempts expect each cultural setting to be significantly associated with some, but not all, relational models. The framework we develop helps explain the findings of these previous attempts and provides insights into empirical research by clarifying which associations to expect between relationships and cultural contexts. We discuss the theoretical basis of our framework, including the idea that RMT and PRT apply to different levels of analysis: RMT's relational models are tied to relationships between two actors and PRT's cultural biases to structures of social networks.

preprint2016arXiv

Micro-foundation using percolation theory of the finite-time singular behavior of the crash hazard rate in a class of rational expectation bubbles

We present a plausible micro-founded model for the previously postulated power law finite time singular form of the crash hazard rate in the Johansen-Ledoit-Sornette model of rational expectation bubbles. The model is based on a percolation picture of the network of traders and the concept that clusters of connected traders share the same opinion. The key ingredient is the notion that a shift of position from buyer to seller of a sufficiently large group of traders can trigger a crash. This provides a formula to estimate the crash hazard rate by summation over percolation clusters above a minimum size of a power sa (with a > 1) of the cluster sizes s, similarly to a generalized percolation susceptibility. The power sa of cluster sizes emerges from the super-linear dependence of group activity as a function of group size, previously documented in the literature. The crash hazard rate exhibits explosive finite-time singular behaviors when the control parameter (fraction of occupied sites, or density of traders in the network) approaches the percolation threshold pc. Realistic dynamics are generated by modelling the density of traders on the percolation network by an Ornstein-Uhlenbeck process, whose memory controls the spontaneous excursion of the control parameter close to the critical region of bubble formation. Our numerical simulations recover the main stylized properties of the JLS model with intermittent explosive super-exponential bubbles interrupted by crashes.

preprint2016arXiv

Modified Profile Likelihood Inference and Interval Forecast of the Burst of Financial Bubbles

We present a detailed methodological study of the application of the modified profile likelihood method for the calibration of nonlinear financial models characterised by a large number of parameters. We apply the general approach to the Log-Periodic Power Law Singularity (LPPLS) model of financial bubbles. This model is particularly relevant because one of its parameters, the critical time $t_c$ signalling the burst of the bubble, is arguably the target of choice for dynamical risk management. However, previous calibrations of the LPPLS model have shown that the estimation of $t_c$ is in general quite unstable. Here, we provide a rigorous likelihood inference approach to determine $t_c$, which takes into account the impact of the other nonlinear (so-called "nuisance") parameters for the correct adjustment of the uncertainty on $t_c$. This provides a rigorous interval estimation for the critical time, rather than a point estimation in previous approaches. As a bonus, the interval estimations can also be obtained for the nuisance parameters ($m,ω$, damping), which can be used to improve filtering of the calibration results. We show that the use of the modified profile likelihood method dramatically reduces the number of local extrema by constructing much simpler smoother log-likelihood landscapes. The remaining distinct solutions can be interpreted as genuine scenarios that unfold as the time of the analysis flows, which can be compared directly via their likelihood ratio. Finally, we develop a multi-scale profile likelihood analysis to visualize the structure of the financial data at different scales (typically from 100 to 750 days). We test the methodology successfully on synthetic price time series and on three well-known historical financial bubbles.

preprint2016arXiv

The Extreme Risk of Personal Data Breaches & The Erosion of Privacy

Personal data breaches from organisations, enabling mass identity fraud, constitute an \emph{extreme risk}. This risk worsens daily as an ever-growing amount of personal data are stored by organisations and on-line, and the attack surface surrounding this data becomes larger and harder to secure. Further, breached information is distributed and accumulates in the hands of cyber criminals, thus driving a cumulative erosion of privacy. Statistical modeling of breach data from 2000 through 2015 provides insights into this risk: A current maximum breach size of about 200 million is detected, and is expected to grow by fifty percent over the next five years. The breach sizes are found to be well modeled by an \emph{extremely heavy tailed} truncated Pareto distribution, with tail exponent parameter decreasing linearly from 0.57 in 2007 to 0.37 in 2015. With this current model, given a breach contains above fifty thousand items, there is a ten percent probability of exceeding ten million. A size effect is unearthed where both the frequency and severity of breaches scale with organisation size like $s^{0.6}$. Projections indicate that the total amount of breached information is expected to double from two to four billion items within the next five years, eclipsing the population of users of the Internet. This massive and uncontrolled dissemination of personal identities raises fundamental concerns about privacy.

preprint2016arXiv

The gradual evolution of buyer--seller networks and their role in aggregate fluctuations

Buyer--seller relationships among firms can be regarded as a longitudinal network in which the connectivity pattern evolves as each firm receives productivity shocks. Based on a data set describing the evolution of buyer--seller links among 55,608 firms over a decade and structural equation modeling, we find some evidence that interfirm networks evolve reflecting a firm's local decisions to mitigate adverse effects from neighbor firms through interfirm linkage, while enjoying positive effects from them. As a result, link renewal tends to have a positive impact on the growth rates of firms. We also investigate the role of networks in aggregate fluctuations.

preprint2015arXiv

"Speculative Influence Network" during financial bubbles: application to Chinese Stock Markets

We introduce the Speculative Influence Network (SIN) to decipher the causal relationships between sectors (and/or firms) during financial bubbles. The SIN is constructed in two steps. First, we develop a Hidden Markov Model (HMM) of regime-switching between a normal market phase represented by a geometric Brownian motion (GBM) and a bubble regime represented by the stochastic super-exponential Sornette-Andersen (2002) bubble model. The calibration of the HMM provides the probability at each time for a given security to be in the bubble regime. Conditional on two assets being qualified in the bubble regime, we then use the transfer entropy to quantify the influence of the returns of one asset $i$ onto another asset $j$, from which we introduce the adjacency matrix of the SIN among securities. We apply our technology to the Chinese stock market during the period 2005-2008, during which a normal phase was followed by a spectacular bubble ending in a massive correction. We introduce the Net Speculative Influence Intensity (NSII) variable as the difference between the transfer entropies from $i$ to $j$ and from $j$ to $i$, which is used in a series of rank ordered regressions to predict the maximum loss (\%{MaxLoss}) endured during the crash. The sectors that influenced other sectors the most are found to have the largest losses. There is a clear prediction skill obtained by using the transfer entropy involving industrial sectors to explain the \%{MaxLoss} of financial institutions but not vice versa. We also show that the bubble state variable calibrated on the Chinese market data corresponds well to the regimes when the market exhibits a strong price acceleration followed by clear change of price regimes. Our results suggest that SIN may contribute significant skill to the development of general linkage-based systemic risks measures and early warning metrics.

preprint2015arXiv

A generic model of dyadic social relationships

We introduce a model of dyadic social interactions and establish its correspondence with relational models theory (RMT), a theory of human social relationships. RMT posits four elementary models of relationships governing human interactions, singly or in combination: Communal Sharing, Authority Ranking, Equality Matching, and Market Pricing. To these are added the limiting cases of asocial and null interactions, whereby people do not coordinate with reference to any shared principle. Our model is rooted in the observation that each individual in a dyadic interaction can do either the same thing as the other individual, a different thing or nothing at all. To represent these three possibilities, we consider two individuals that can each act in one out of three ways toward the other: perform a social action X or Y, or alternatively do nothing. We demonstrate that the relationships generated by this model aggregate into six exhaustive and disjoint categories. We propose that four of these categories match the four relational models, while the remaining two correspond to the asocial and null interactions defined in RMT. We generalize our results to the presence of N social actions. We infer that the four relational models form an exhaustive set of all possible dyadic relationships based on social coordination. Hence, we contribute to RMT by offering an answer to the question of why there could exist just four relational models. In addition, we discuss how to use our representation to analyze data sets of dyadic social interactions, and how social actions may be valued and matched by the agents.

preprint2015arXiv

Cancer risk is not (just) bad luck

Tomasetti and Vogelstein recently proposed that the majority of variation in cancer risk among tissues is due to "bad luck," that is, random mutations arising during DNA replication in normal noncancerous stem cells. They generalize this finding to cancer overall, claiming that "the stochastic effects of DNA replication appear to be the major contributor to cancer in humans." We show that this conclusion results from a logical fallacy based on ignoring the influence of population heterogeneity in correlations exhibited at the level of the whole population. Because environmental and genetic factors cannot explain the huge differences in cancer rates between different organs, it is wrong to conclude that these factors play a minor role in cancer rates. In contrast, we show that one can indeed measure huge differences in cancer rates between different organs and, at the same time, observe a strong effect of environmental and genetic factors in cancer rates.

preprint2015arXiv

Currency target zone modeling: An interplay between physics and economics

We study the performance of the euro/Swiss franc exchange rate in the extraordinary period from September 6, 2011 and January 15, 2015 when the Swiss National Bank enforced a minimum exchange rate of 1.20 Swiss francs per euro. Based on the analogy between Brownian motion in finance and physics, the first-order effect of such a steric constraint would enter a priori in the form of a repulsive entropic force associated with the paths crossing the barrier that are forbidden. Non-parametric empirical estimates of drift and volatility show that the predicted first-order analogy between economics and physics are incorrect. The clue is to realise that the random walk nature of financial prices results from the continuous anticipations of traders about future opportunities, whose aggregate actions translate into an approximate efficient market with almost no arbitrage opportunities. With the Swiss National Bank stated commitment to enforce the barrier, traders's anticipation of this action leads to a vanishing drift together with a volatility of the exchange rate that depends on the distance to the barrier. We give direct quantitative empirical evidence that this effect is well described by Krugman's target zone model [P.R. Krugman. The Quarterly Journal of Economics, 106(3):669-682, 1991]. Motivated by the insights from this economical model, we revise the initial economics-physics analogy and show that, within the context of hindered diffusion, the two systems can be described with the same mathematics after all. Using a recently proposed extended analogy in terms of a colloidal Brownian particle embedded in a fluid of molecules associated with the underlying order book, we derive that, close to the restricting boundary, the dynamics of both systems is described by a stochastic differential equation with a very small constant drift and a linear diffusion coefficient.

preprint2015arXiv

Financial Knudsen number: breakdown of continuous price dynamics and asymmetric buy and sell structures confirmed by high precision order book information

We generalise the description of the dynamics of the order book of financial markets in terms of a Brownian particle embedded in a fluid of incoming, exiting and annihilating particles by presenting a model of the velocity on each side (buy and sell) independently. The improved model builds on the time-averaged number of particles in the inner layer and its change per unit time, where the inner layer is revealed by the correlations between price velocity and change in the number of particles (limit orders). This allows us to introduce the Knudsen number of the financial Brownian particle motion and its asymmetric version (on the buy and sell sides). Not being considered previously, the asymmetric Knudsen numbers are crucial in finance in order to detect asymmetric price changes. The Knudsen numbers allows us to characterise the conditions for the market dynamics to be correctly described by a continuous stochastic process. Not questioned until now for large liquid markets such as the USD/JPY and EUR/USD exchange rates, we show that there are regimes when the Knudsen numbers are so high that discrete particle effects dominate, such as during market stresses and crashes. We document the presence of imbalances of particles depletion rates on the buy and sell sides that are associated with high Knudsen numbers and violent directional price changes. This indicator can detect the direction of the price motion at the early stage while the usual volatility risk measure is blind to the price direction.

preprint2015arXiv

Macroeconomic Dynamics of Assets, Leverage and Trust

A macroeconomic model based on the economic variables (i) assets, (ii) leverage (defined as debt over asset) and (iii) trust (defined as the maximum sustainable leverage) is proposed to investigate the role of credit in the dynamics of economic growth, and how credit may be associated with both economic performance and confidence. Our first notable finding is the mechanism of reward/penalty associated with patience, as quantified by the return on assets. In regular economies where the EBITA/Assets ratio is larger than the cost of debt, starting with a trust higher than leverage results in the highest long-term return on assets (which can be seen as a proxy for economic growth). Our second main finding concerns a recommendation for the reaction of a central bank to an external shock that affects negatively the economic growth. We find that late policy intervention in the model economy results in the highest long-term return on assets and largest asset value. But this comes at the cost of suffering longer from the crisis until the intervention occurs. The phenomenon can be ascribed to the fact that postponing intervention allows trust to increase first, and it is most effective to intervene when trust is high. These results derive from two fundamental assumptions underlying our model: (a) trust tends to increase when it is above leverage; (b) economic agents learn optimally to adjust debt for a given level of trust and amount of assets. Using a Markov Switching Model for the EBITA/Assets ratio, we have successfully calibrated our model to the empirical data of the return on equity of the EURO STOXX 50 for the time period 2000-2013. We find that dynamics of leverage and trust can be highly non-monotonous with curved trajectories, as a result of the nonlinear coupling between the variables.

preprint2015arXiv

Of Disasters and Dragon Kings: A Statistical Analysis of Nuclear Power Incidents & Accidents

We provide, and perform a risk theoretic statistical analysis of, a dataset that is 75 percent larger than the previous best dataset on nuclear incidents and accidents, comparing three measures of severity: INES (International Nuclear Event Scale), radiation released, and damage dollar losses. The annual rate of nuclear accidents, with size above 20 Million US$, per plant, decreased from the 1950s until dropping significantly after Chernobyl (April, 1986). The rate is now roughly stable at 0.002 to 0.003, i.e., around 1 event per year across the current fleet. The distribution of damage values changed after Three Mile Island (TMI; March, 1979), where moderate damages were suppressed but the tail became very heavy, being described by a Pareto distribution with tail index 0.55. Further, there is a runaway disaster regime, associated with the "dragon-king" phenomenon, amplifying the risk of extreme damage. In fact, the damage of the largest event (Fukushima; March, 2011) is equal to 60 percent of the total damage of all 174 accidents in our database since 1946. In dollar losses we compute a 50% chance that (i) a Fukushima event (or larger) occurs in the next 50 years, (ii) a Chernobyl event (or larger) occurs in the next 27 years and (iii) a TMI event (or larger) occurs in the next 10 years. Finally, we find that the INES scale is inconsistent. To be consistent with damage, the Fukushima disaster would need to have an INES level of 11, rather than the maximum of 7.

preprint2015arXiv

Power law scaling and "Dragon-Kings" in distributions of intraday financial drawdowns

We investigate the distributions of epsilon-drawdowns and epsilon-drawups of the most liquid futures financial contracts of the world at time scales of 30 seconds. The epsilon-drawdowns (resp. epsilon- drawups) generalise the notion of runs of negative (resp. positive) returns so as to capture the risks to which investors are arguably the most concerned with. Similarly to the distribution of returns, we find that the distributions of epsilon-drawdowns and epsilon-drawups exhibit power law tails, albeit with exponents significantly larger than those for the return distributions. This paradoxical result can be attributed to (i) the existence of significant transient dependence between returns and (ii) the presence of large outliers (dragon-kings) characterizing the extreme tail of the drawdown/drawup distributions deviating from the power law. The study of the tail dependence between the sizes, speeds and durations of drawdown/drawup indicates a clear relationship between size and speed but none between size and duration. This implies that the most extreme drawdown/drawup tend to occur fast and are dominated by a few very large returns. We discuss both the endogenous and exogenous origins of these extreme events.

preprint2015arXiv

Secular bipolar growth rate of the real US GDP per capita: implications for understanding past and future economic growth

We present a quantitative characterisation of the fluctuations of the annualized growth rate of the real US GDP per capita growth at many scales, using a wavelet transform analysis of two data sets, quarterly data from 1947 to 2015 and annual data from 1800 to 2010. Our main finding is that the distribution of GDP growth rates can be well approximated by a bimodal function associated to a series of switches between regimes of strong growth rate $ρ_\text{high}$ and regimes of low growth rate $ρ_\text{low}$. The succession of such two regimes compounds to produce a remarkably stable long term average real annualized growth rate of 1.6\% from 1800 to 2010 and $\approx 2.0\%$ since 1950, which is the result of a subtle compensation between the high and low growth regimes that alternate continuously. Thus, the overall growth dynamics of the US economy is punctuated, with phases of strong growth that are intrinsically unsustainable, followed by corrections or consolidation until the next boom starts. We interpret these findings within the theory of "social bubbles" and argue as a consequence that estimations of the cost of the 2008 crisis may be misleading. We also interpret the absence of strong recovery since 2008 as a protracted low growth regime $ρ_\text{low}$ associated with the exceptional nature of the preceding large growth regime.

preprint2015arXiv

Two-state Markov-chain Poisson nature of individual cellphone call statistics

Humans are heterogenous and the behaviors of individuals could be different from that at the population level. We conduct an in-depth study of the temporal patterns of cellphone conversation activities of 73'339 anonymous cellphone users with the same truncated Weibull distribution of inter-call durations. We find that the individual call events exhibit a pattern of bursts, in which high activity periods are alternated with low activity periods. Surprisingly, the number of events in high activity periods are found to conform to a power-law distribution at the population level, but follow an exponential distribution at the individual level, which is a hallmark of absence of memory in individual call activities. Such exponential distribution is also observed for the number of events in low activity periods. Together with the exponential distributions of inter-call durations within bursts and of the intervals between consecutive bursts, we demonstrate that the individual call activities are driven by two independent Poisson processes, which can be combined within a minimal model in terms of a two-state first-order Markov chain giving very good agreement with the empirical distributions using the parameters estimated from real data for about half of the individuals in our sample. By measuring directly the distributions of call rates across the population, which exhibit power-law tails, we explain the difference with previous population level studies, purporting the existence of power-law distributions, via the "Superposition of Distributions" mechanism: The superposition of many exponential distributions of activities with a power-law distribution of their characteristic scales leads to a power-law distribution of the activities at the population level.

preprint2014arXiv

A Creepy World

Using the mechanics of creep in material sciences as a metaphor, we present a general framework to understand the evolution of financial, economic and social systems and to construct scenarios for the future. In a nutshell, highly non-linear out-of-equilibrium systems subjected to exogenous perturbations tend to exhibit a long phase of slow apparent stable evolution, which are nothing but slow maturations towards instabilities, failures and changes of regimes. With examples from history where a small event had a cataclysmic consequence, we propose a novel view of the current state of the world via the logical scenarios that derive, avoiding the traps of an illusionary stability and simple linear extrapolation. The endogenous scenarios are "muddling along", "managing through" and "blood red abyss". The exogenous scenarios are "painful adjustment" and "golden east".

preprint2014arXiv

Apparent criticality and calibration issues in the Hawkes self-excited point process model: application to high-frequency financial data

We present a careful analysis of possible issues on the application of the self-excited Hawkes process to high-frequency financial data. We carefully analyze a set of effects leading to significant biases in the estimation of the "criticality index" n that quantifies the degree of endogeneity of how much past events trigger future events. We report a number of model biases that are intrinsic to the estimation of brnaching ratio (n) when using power law memory kernels. We demonstrate that the calibration of the Hawkes process on mixtures of pure Poisson process with changes of regime leads to completely spurious apparent critical values for the branching ratio (n~1) while the true value is actually n=0. More generally, regime shifts on the parameters of the Hawkes model and/or on the generating process itself are shown to systematically lead to a significant upward bias in the estimation of the branching ratio. We also demonstrate the importance of the preparation of the high-frequency financial data and give special care to the decrease of quality of the timestamps of tick data due to latency and grouping of messages to packets by the stock exchange. Altogether, our careful exploration of the caveats of the calibration of the Hawkes process stresses the need for considering all the above issues before any conclusion can be sustained. In this respect, because the above effects are plaguing their analyses, the claim by Hardiman, Bercot and Bouchaud (2013) that financial market have been continuously functioning at or close to criticality (n~1) cannot be supported. In contrast, our previous results on E-mini S&P 500 Futures Contracts and on major commodity future contracts are upheld.

preprint2014arXiv

Effective Measure of Endogeneity for the Autoregressive Conditional Duration Point Processes via Mapping to the Self-Excited Hawkes Process

In order to disentangle the internal dynamics from exogenous factors within the Autoregressive Conditional Duration (ACD) model, we present an effective measure of endogeneity. Inspired from the Hawkes model, this measure is defined as the average fraction of events that are triggered due to internal feedback mechanisms within the total population. We provide a direct comparison of the Hawkes and ACD models based on numerical simulations and show that our effective measure of endogeneity for the ACD can be mapped onto the "branching ratio" of the Hawkes model.

preprint2014arXiv

Estimation of the Hawkes Process With Renewal Immigration Using the EM Algorithm

We introduce the Hawkes process with renewal immigration and make its statistical estimation possible with two Expectation Maximization (EM) algorithms. The standard Hawkes process introduces immigrant points via a Poisson process, and each immigrant has a subsequent cluster of associated offspring of multiple generations. We generalize the immigration to come from a Renewal process; introducing dependence between neighbouring clusters, and allowing for over/under dispersion in cluster locations. This complicates evaluation of the likelihood since one needs to know which subset of the observed points are immigrants. Two EM algorithms enable estimation here: The first is an extension of an existing algorithm that treats the entire branching structure - which points are immigrants, and which point is the parent of each offspring - as missing data. The second considers only if a point is an immigrant or not as missing data and can be implemented with linear time complexity. Both algorithms are found to be consistent in simulation studies. Further, we show that misspecifying the immigration process introduces signficant bias into model estimation-- especially the branching ratio, which quantifies the strength of self excitation. Thus, this extended model provides a valuable alternative model in practice.

preprint2014arXiv

Financial Brownian particle in the layered order book fluid and Fluctuation-Dissipation relations

We introduce a novel description of the dynamics of the order book of financial markets as that of an effective colloidal Brownian particle embedded in fluid particles. The analysis of a comprehensive market data enables us to identify all motions of the fluid particles. Correlations between the motions of the Brownian particle and its surrounding fluid particles reflect specific layering interactions; in the inner-layer, the correlation is strong and with short memory while, in the outer-layer, it is weaker and with long memory. By interpreting and estimating the contribution from the outer-layer as a drag resistance, we demonstrate the validity of the fluctuation-dissipation relation (FDR) in this non-material Brownian motion process.

preprint2014arXiv

Financial bubbles: mechanisms and diagnostics

We define a financial bubble as a period of unsustainable growth, when the price of an asset increases ever more quickly, in a series of accelerating phases of corrections and rebounds. More technically, during a bubble phase, the price follows a faster-than-exponential power law growth process, often accompanied by log-periodic oscillations. This dynamic ends abruptly in a change of regime that may be a crash or a substantial correction. Because they leave such specific traces, bubbles may be recognised in advance, that is, before they burst. In this paper, we will explain the mechanism behind financial bubbles in an intuitive way. We will show how the log-periodic power law emerges spontaneously from the complex system that financial markets are, as a consequence of feedback mechanisms, hierarchical structure and specific trading dynamics and investment styles. We argue that the risk of a major correction, or even a crash, becomes substantial when a bubble develops towards maturity, and that it is therefore very important to find evidence of bubbles and to follow their development from as early a stage as possible. The tools that are explained in this paper actually serve that purpose. They are at the core of the Financial Crisis Observatory at the ETH Zurich, where tens of thousands of assets are monitored on a daily basis. This allow us to have a continuous overview of emerging bubbles in the global financial markets. The companion report available as part of the Notenstein white paper series (2014) with the title ``Financial bubbles: mechanism, diagnostic and state of the World (Feb. 2014)'' presents a practical application of the methodology outlines in this article and describes our view of the status concerning positive and negative bubbles in the financial markets, as of the end of January 2014.

preprint2014arXiv

Forecasting future oil production in Norway and the UK: a general improved methodology

We present a new Monte-Carlo methodology to forecast the crude oil production of Norway and the U.K. based on a two-step process, (i) the nonlinear extrapolation of the current/past performances of individual oil fields and (ii) a stochastic model of the frequency of future oil field discoveries. Compared with the standard methodology that tends to underestimate remaining oil reserves, our method gives a better description of future oil production, as validated by our back-tests starting in 2008. Specifically, we predict remaining reserves extractable until 2030 to be 188 +/- 10 million barrels for Norway and 98 +/- 10 million barrels for the UK, which are respectively 45% and 66% above the predictions using the standard methodology.

preprint2014arXiv

Fractal multi-level organisation of human groups in a virtual world

Humans are fundamentally social. They have progressively dominated their environment by the strength and creativity provided by and within their grouping. It is well recognised that human groups are highly structured, and the anthropological literature has loosely classified them according to their size and function, such as support cliques, sympathy groups, bands, cognitive groups, tribes, linguistic groups and so on. Recently, combining data on human grouping patterns in a comprehensive and systematic study, Zhou et al. identified a quantitative discrete hierarchy of group sizes with a preferred scaling ratio close to $3$, which was later confirmed for hunter-gatherer groups and for other mammalian societies. Using high precision large scale Internet-based social network data, we extend these early findings on a very large data set. We analyse the organisational structure of a complete, multi-relational, large social multiplex network of a human society consisting of about 400,000 odd players of a massive multiplayer online game for which we know all about the group memberships of every player. Remarkably, the online players exhibit the same type of structured hierarchical layers as the societies studied by anthropologists, where each of these layers is three to four times the size of the lower layer. Our findings suggest that the hierarchical organisation of human society is deeply nested in human psychology.

preprint2014arXiv

How Much is the Whole Really More than the Sum of its Parts? 1 + 1 = 2.5: Superlinear Productivity in Collective Group Actions

In a variety of open source software projects, we document a superlinear growth of production ($R \sim c^β$) as a function of the number of active developers $c$, with $β\simeq 4/3$ with large dispersions. For a typical project in this class, doubling of the group size multiplies typically the output by a factor $2^β=2.5$, explaining the title. This superlinear law is found to hold for group sizes ranging from 5 to a few hundred developers. We propose two classes of mechanisms, {\it interaction-based} and {\it large deviation}, along with a cascade model of productive activity, which unifies them. In this common framework, superlinear productivity requires that the involved social groups function at or close to criticality, in the sense of a subtle balance between order and disorder. We report the first empirical test of the renormalization of the exponent of the distribution of the sizes of first generation events into the renormalized exponent of the distribution of clusters resulting from the cascade of triggering over all generation in a critical branching process in the non-meanfield regime. Finally, we document a size effect in the strength and variability of the superlinear effect, with smaller groups exhibiting widely distributed superlinear exponents, some of them characterizing highly productive teams. In contrast, large groups tend to have a smaller superlinearity and less variability.

preprint2014arXiv

Micro-transition cascades to percolation

We report the discovery of a discrete hierarchy of micro-transitions occurring in models of continuous and discontinuous percolation. The precursory micro-transitions allow us to target almost deterministically the location of the transition point to global connectivity. This extends to the class of intrinsically stochastic processes the possibility to use warning signals anticipating phase transitions in complex systems.

preprint2014arXiv

Using Prediction Markets to Incentivize and Measure Collective Knowledge Production

We present a mechanism design, coupling an online collaboration software and a prediction market, which allows tracking down the very roots of individual incentives, actions and how these behaviors influence collective intelligence in terms of knowledge production as a public good. We show that the incentive mechanism efficiently engages users without further governance structure, and doesn't crowd out intrinsic motivation. Furthermore, it enables a powerful and robust creative destruction process, which helps quickly filter out irrelevant knowledge. While still at an early stage, this mechanism design can not only bring insights for knowledge production organization design, but also has the potential to illuminate the fundamental mechanisms underlying the emergence of collective intelligence.

preprint2013arXiv

A Stable and Robust Calibration Scheme of the Log-Periodic Power Law Model

We present a simple transformation of the formulation of the log-periodic power law formula of the Johansen-Ledoit-Sornette model of financial bubbles that reduces it to a function of only three nonlinear parameters. The transformation significantly decreases the complexity of the fitting procedure and improves its stability tremendously because the modified cost function is now characterized by good smooth properties with in general a single minimum in the case where the model is appropriate to the empirical data. We complement the approach with an additional subordination procedure that slaves two of the nonlinear parameters to what can be considered to be the most crucial nonlinear parameter, the critical time $t_c$ defined as the end of the bubble and the most probably time for a crash to occur. This further decreases the complexity of the search and provides an intuitive representation of the results of the calibration. With our proposed methodology, metaheuristic searches are not longer necessary and one can resort solely to rigorous controlled local search algorithms, leading to dramatic increase in efficiency. Empirical tests on the Shanghai Composite index (SSE) from January 2007 to March 2008 illustrate our findings.

preprint2013arXiv

Automatic reconstruction of fault networks from seismicity catalogs including location uncertainty

We introduce the Anisotropic Clustering of Location Uncertainty Distributions (ACLUD) method to reconstruct active fault networks on the basis of both earthquake locations and their estimated individual uncertainties. After a massive search through the large solution space of possible reconstructed fault networks, we apply six different validation procedures in order to select the corresponding best fault network. Two of the validation steps (cross-validation and Bayesian Information Criterion (BIC) process the fit residuals, while the four others look for solutions that provide the best agreement with independently observed focal mechanisms. Tests on synthetic catalogs allow us to qualify the performance of the fitting method and of the various validation procedures. The ACLUD method is able to provide solutions that are close to the expected ones, especially for the BIC and focal mechanismbased techniques. The clustering method complemented by the validation step based on focal mechanisms provides good solutions even in the presence of a significant spatial background seismicity rate. Our new fault reconstruction method is then applied to the Landers area in Southern California and compared with previous clustering methods. The results stress the importance of taking into account undersampled sub-fault structures as well as of the spatially inhomogeneous location uncertainties.

preprint2013arXiv

Clarifications to Questions and Criticisms on the Johansen-Ledoit-Sornette Bubble Model

The Johansen-Ledoit-Sornette (JLS) model of rational expectation bubbles with finite-time singular crash hazard rates has been developed to describe the dynamics of financial bubbles and crashes. It has been applied successfully to a large variety of financial bubbles in many different markets. Having been developed for more than one decade, the JLS model has been studied, analyzed, used and criticized by several researchers. Much of this discussion is helpful for advancing the research. However, several serious misconceptions seem to be present within this collective conversation both on theoretical and empirical aspects. Several of these problems appear to stem from the fast evolution of the literature on the JLS model and related works. In the hope of removing possible misunderstanding and of catalyzing useful future developments, we summarize these common questions and criticisms concerning the JLS model and offer a synthesis of the existing state-of-the-art and best-practice advices.

preprint2013arXiv

Dynamics and Spatial Distribution of Global Nighttime Lights

Using open source data, we observe the fascinating dynamics of nighttime light. Following a global economic regime shift, the planetary center of light can be seen moving eastwards at a pace of about 60 km per year. Introducing spatial light Gini coefficients, we find a universal pattern of human settlements across different countries and see a global centralization of light. Observing 160 different countries we document the expansion of developing countries, the growth of new agglomerations, the regression in countries suffering from demographic decline and the success of light pollution abatement programs in western countries.

preprint2013arXiv

High quality topic extraction from business news explains abnormal financial market volatility

Understanding the mutual relationships between information flows and social activity in society today is one of the cornerstones of the social sciences. In financial economics, the key issue in this regard is understanding and quantifying how news of all possible types (geopolitical, environmental, social, financial, economic, etc.) affect trading and the pricing of firms in organized stock markets. In this article, we seek to address this issue by performing an analysis of more than 24 million news records provided by Thompson Reuters and of their relationship with trading activity for 206 major stocks in the S&P US stock index. We show that the whole landscape of news that affect stock price movements can be automatically summarized via simple regularized regressions between trading activity and news information pieces decomposed, with the help of simple topic modeling techniques, into their "thematic" features. Using these methods, we are able to estimate and quantify the impacts of news on trading. We introduce network-based visualization techniques to represent the whole landscape of news information associated with a basket of stocks. The examination of the words that are representative of the topic distributions confirms that our method is able to extract the significant pieces of information influencing the stock market. Our results show that one of the most puzzling stylized fact in financial economies, namely that at certain times trading volumes appear to be "abnormally large," can be partially explained by the flow of news. In this sense, our results prove that there is no "excess trading," when restricting to times when news are genuinely novel and provide relevant financial information.

preprint2013arXiv

Is There A Real Estate Bubble in Switzerland?

We have analyzed the risks of possible development of bubbles in the Swiss residential real estate market. The data employed in this work has been collected by comparis.ch, and carefully cleaned from duplicate records through a procedure based on supervised machine learning methods. The study uses the log periodic power law (LPPL) bubble model to analyze the development of asking prices of residential properties in all Swiss districts between 2005 and 2013. The results suggest that there are 11 critical districts that exhibit signatures of bubbles, and seven districts where bubbles have already burst. Despite these strong signatures, it is argued that, based on the current economic environment, a soft landing rather than a severe crash is expected.

preprint2013arXiv

Predictability and suppression of extreme events in complex systems

In many complex systems, large events are believed to follow power-law, scale-free probability distributions, so that the extreme, catastrophic events are unpredictable. Here, we study coupled chaotic oscillators that display extreme events. The mechanism responsible for the rare, largest events makes them distinct and their distribution deviates from a power-law. Based on this mechanism identification, we show that it is possible to forecast in real time an impending extreme event. Once forecasted, we also show that extreme events can be suppressed by applying tiny perturbations to the system.

preprint2013arXiv

The Barycentric Fixed Mass Method for Multifractal Analysis

We present a novel method to estimate the multifractal spectrum of point distributions. The method incorporates two motivated criteria (barycentric pivot point selection and non-overlapping coverage) in order to reduce edge effects, improve precision and reduce computation time. Implementation of the method on synthetic benchmarks demonstrates the superior performance of the proposed method compared with existing alternatives routinely used in the literature. Finally, we use the method to estimate the multifractal properties of the widely studied growth process of Diffusion Limited Aggregation and compare our results with recent and earlier studies. Our tests support the conclusion of a genuine but weak multifractality of the central core of DLA clusters, with Dq decreasing from 1.75+/-0.01 for q=-10 to 1.65+/-0.01 for q=+10.

preprint2012arXiv

Comparing the performance of FA, DFA and DMA using different synthetic long-range correlated time series

Notwithstanding the significant efforts to develop estimators of long-range correlations (LRC) and to compare their performance, no clear consensus exists on what is the best method and under which conditions. In addition, synthetic tests suggest that the performance of LRC estimators varies when using different generators of LRC time series. Here, we compare the performances of four estimators [Fluctuation Analysis (FA), Detrended Fluctuation Analysis (DFA), Backward Detrending Moving Average (BDMA), and centred Detrending Moving Average (CDMA)]. We use three different generators [Fractional Gaussian Noises, and two ways of generating Fractional Brownian Motions]. We find that CDMA has the best performance and DFA is only slightly worse in some situations, while FA performs the worst. In addition, CDMA and DFA are less sensitive to the scaling range than FA. Hence, CDMA and DFA remain "The Methods of Choice" in determining the Hurst index of time series.

preprint2012arXiv

New empirical tests of the multifractal Omori law for Taiwan

We report new tests on the Taiwan earthquake catalog of the prediction by the Multifractal Stress Activation (MSA) model that the p-value of the Omori law for the rate of aftershocks following a mainshock is an increasing function of its magnitude Mm. This effort is motivated by the quest to validate this crucial prediction of the MSA model and to investigate its possible dependence on local tectonic conditions. With careful attention to the long-term as well as short-term time-dependent magnitude completeness of the Taiwan catalog, and with the use of three different declustering techniques, we confirm the universality of the prediction p(Mm) = (0.09 \pm 0.03) \times Mm + (0.47 \pm 0.10), valid for the SCEC Southern California catalog, the Harvard-CMT worldwide catalog, the JMA Japan catalog and the Taiwan catalog. The observed deviations of the two coefficients of the p(Mm) linear dependence on Mm from catalog to catalog are not significant enough to correlate meaningfully with any tectonic features.

preprint2012arXiv

On the distribution of time-to-proof of mathematical conjectures

What is the productivity of Science? Can we measure an evolution of the production of mathematicians over history? Can we predict the waiting time till the proof of a challenging conjecture such as the P-versus-NP problem? Motivated by these questions, we revisit a suggestion published recently and debated in the "New Scientist" that the historical distribution of time-to-proof's, i.e., of waiting times between formulation of a mathematical conjecture and its proof, can be quantified and gives meaningful insights in the future development of still open conjectures. We find however evidence that the mathematical process of creation is too much non-stationary, with too little data and constraints, to allow for a meaningful conclusion. In particular, the approximate unsteady exponential growth of human population, and arguably that of mathematicians, essentially hides the true distribution. Another issue is the incompleteness of the dataset available. In conclusion we cannot really reject the simplest model of an exponential rate of conjecture proof with a rate of 0.01/year for the dataset that we have studied, translating into an average waiting time to proof of 100 years. We hope that the presented methodology, combining the mathematics of recurrent processes, linking proved and still open conjectures, with different empirical constraints, will be useful for other similar investigations probing the productivity associated with mankind growth and creativity.

preprint2012arXiv

Quantifying reflexivity in financial markets: towards a prediction of flash crashes

We introduce a new measure of activity of financial markets that provides a direct access to their level of endogeneity. This measure quantifies how much of price changes are due to endogenous feedback processes, as opposed to exogenous news. For this, we calibrate the self-excited conditional Poisson Hawkes model, which combines in a natural and parsimonious way exogenous influences with self-excited dynamics, to the E-mini S&P 500 futures contracts traded in the Chicago Mercantile Exchange from 1998 to 2010. We find that the level of endogeneity has increased significantly from 1998 to 2010, with only 70% in 1998 to less than 30% since 2007 of the price changes resulting from some revealed exogenous information. Analogous to nuclear plant safety concerned with avoiding "criticality", our measure provides a direct quantification of the distance of the financial market to a critical state defined precisely as the limit of diverging trading activity in absence of any external driving.

preprint2012arXiv

Strong gender differences in reproductive success variance, and the times to the most recent common ancestors

The Time To the Most Recent Common Ancestor (TMRCA) based on human mitochondrial DNA (mtDNA) is estimated to be twice that based on the non-recombining part of the Y chromosome (NRY). These TMRCAs have special demographic implications because mtDNA is transmitted only from mother to child, and NRY from father to son. Therefore, mtDNA reflects female history, and NRY, male history. To investigate what caused the two-to-one female-male TMRCA ratio in humans, we develop a forward-looking agent-based model (ABM) with overlapping generations and individual life cycles. We implement two main mating systems: polygynandry and polygyny with different degrees in between. In each mating system, the male population can be either homogeneous or heterogeneous. In the latter case, some males are `alphas' and others are `betas', which reflects the extent to which they are favored by female mates. A heterogeneous male population implies a competition among males with the purpose of signaling as alphas. The introduction of a heterogeneous male population is found to reduce by a factor 2 the probability of finding equal female and male TMRCAs and shifts the distribution of the TMRCA ratio to higher values. We find that high male-male competition is necessary to reproduce a TMRCA ratio of 2: less than half the males can be alphas and betas can have at most half the fitness of alphas. In addition, in the modes that maximize the probability of having a TMRCA ratio between 1.5 and 2.5, the present generation has 1.4 times as many female as male ancestors. We also tested the effect of sex-biased migration and sex-specific death rates and found that these are unlikely to explain alone the sex-biased TMRCA ratio observed in humans. Our results support the view that we are descended from males who were successful in a highly competitive context, while females were facing a much smaller female-female competition.

preprint2012arXiv

Super-exponential bubbles in lab experiments: evidence for anchoring over-optimistic expectations on price

We analyze a controlled price formation experiment in the laboratory that shows evidence for bubbles. We calibrate two models that demonstrate with high statistical significance that these laboratory bubbles have a tendency to grow faster than exponential due to positive feedback. We show that the positive feedback operates by traders continuously upgrading their over-optimistic expectations of future returns based on past prices rather than on realized returns.

preprint2012arXiv

Universality class of balanced flows with bottlenecks: granular flows, pedestrian fluxes and financial price dynamics

We propose and document the evidence for an analogy between the dynamics of granular counter-flows in the presence of bottlenecks or restrictions and financial price formation processes. Using extensive simulations, we find that the counter-flows of simulated pedestrians through a door display many stylized facts observed in financial markets when the density around the door is compared with the logarithm of the price. The stylized properties are present already when the agents in the pedestrian model are assumed to display a zero-intelligent behavior. If agents are given decision-making capacity and adapt to partially follow the majority, periods of herding behavior may additionally occur. This generates the very slow decay of the autocorrelation of absolute return due to an intermittent dynamics. Our finding suggest that the stylized facts in the fluctuations of the financial prices result from a competition of two groups with opposite interests in the presence of a constraint funneling the flow of transactions to a narrow band of prices.

preprint2012arXiv

When games meet reality: is Zynga overvalued?

On December 16th, 2011, Zynga, the well-known social game developing company went public. This event followed other recent IPOs in the world of social networking companies, such as Groupon or Linkedin among others. With a valuation close to 7 billion USD at the time when it went public, Zynga became one of the biggest web IPOs since Google. This recent enthusiasm for social networking companies raises the question whether they are overvalued. Indeed, during the few months since its IPO, Zynga showed significant variability, its market capitalization going from 5.6 to 10.2 billion USD, hinting at a possible irrational behavior from the market. To bring substance to the debate, we propose a two-tiered approach to compute the intrinsic value of Zynga. First, we introduce a new model to forecast its user base, based on the individual dynamics of its major games. Next, we model the revenues per user using a logistic function, a standard model for growth in competition. This allows us to bracket the valuation of Zynga using three different scenarios: 3.4, 4.0 and 4.8 billion USD in the base case, high growth and extreme growth scenario respectively. This suggests that Zynga has been overpriced ever since its IPO. Finally, we propose an investment strategy (dated April 19th, 2012 on the arXive), which is based on our diagnostic of a bubble for Zynga and how this herding / bubbly sentiment can be expected to play together with two important coming events (the quarterly financial result announcement around April 26th, 2012 followed by the end of a first lock-up period around April 30th, 2012). On the long term, our analysis indicates that Zynga's price should decrease significantly. The paper ends with a post-mortem analysis added on May 24th, 2012, just before going to press, showing that we have successfully predicted the downward trend of Zynga. Since April 27th, 2012, Zynga dropped 25%.

preprint2011arXiv

Automated Seizure Detection: Unrecognized Challenges, Unexpected Insights

One of epileptology's fundamental aims is the formulation of a universal, internally consistent seizure definition. To assess this aim's feasibility, three signal analysis methods were applied to a seizure time series and performance comparisons were undertaken among them and with respect to a validated algorithm. One of the methods uses a Fisher's matrix weighted measure of the rate of parameters change of a 2n order auto-regressive model, another is based on the Wavelet Transform Maximum Modulus for quantification of changes in the logarithm of the standard deviation of ECoG power and yet another employs the ratio of short-to-long term averages computed from cortical signals. The central finding, fluctuating concordance among all methods' output as a function of seizure duration, uncovers unexpected hurdles in the path to a universal definition, while furnishing relevant knowledge in the dynamical (spectral non-stationarity and varying ictal signal complexity) and clinical (probable attainability of consensus) domains.

preprint2011arXiv

Climate warming and stability of cold hanging glaciers: Lessons from the gigantic 1895 Altels break-off

The Altels hanging glacier broke off on September 11, 1895. The ice volume of this catastrophic rupture was estimated at $\rm 4.10^6$ cubic meters and is the largest ever observed ice fall event in the Alps. The causes of this collapse are however not entirely clear. Based on previous studies, we reanalyzed this break-off event, with the help of a new numerical model, initially developed by Faillettaz and others (2010) for gravity-driven instabilities. The simulations indicate that a break-off event is only possible when the basal friction at the bedrock is reduced in a restricted area, possibly induced by the storage of infiltrated water within the glacier. Moreover, our simulations reveal a two-step behavior: (i) A first quiescent phase, without visible changes, with a duration depending on the rate of basal changes; (ii) An active phase with a rapid increase of basal motion over a few days. The general lesson obtained from the comparison between the simulations and the available evidence is that visible signs of the destabilization process of a hanging glacier, resulting from a progressive warming of the ice/bed interface towards a temperate regime, will appear just a few days prior to the collapse.

preprint2011arXiv

Detection of Crashes and Rebounds in Major Equity Markets

Financial markets are well known for their dramatic dynamics and consequences that affect much of the world's population. Consequently, much research has aimed at understanding, identifying and forecasting crashes and rebounds in financial markets. The Johansen-Ledoit-Sornette (JLS) model provides an operational framework to understand and diagnose financial bubbles from rational expectations and was recently extended to negative bubbles and rebounds. Using the JLS model, we develop an alarm index based on an advanced pattern recognition method with the aim of detecting bubbles and performing forecasts of market crashes and rebounds. Testing our methodology on 10 major global equity markets, we show quantitatively that our developed alarm performs much better than chance in forecasting market crashes and rebounds. We use the derived signal to develop elementary trading strategies that produce statistically better performances than a simple buy and hold strategy.

preprint2011arXiv

Diagnosis and Prediction of Market Rebounds in Financial Markets

We introduce the concept of "negative bubbles" as the mirror image of standard financial bubbles, in which positive feedback mechanisms may lead to transient accelerating price falls. To model these negative bubbles, we adapt the Johansen-Ledoit-Sornette (JLS) model of rational expectation bubbles with a hazard rate describing the collective buying pressure of noise traders. The price fall occurring during a transient negative bubble can be interpreted as an effective random downpayment that rational agents accept to pay in the hope of profiting from the expected occurrence of a possible rally. We validate the model by showing that it has significant predictive power in identifying the times of major market rebounds. This result is obtained by using a general pattern recognition method which combines the information obtained at multiple times from a dynamical calibration of the JLS model. Error diagrams, Bayesian inference and trading strategies suggest that one can extract genuine information and obtain real skill from the calibration of negative bubbles with the JLS model. We conclude that negative bubbles are in general predictably associated with large rebounds or rallies, which are the mirror images of the crashes terminating standard bubbles.

preprint2011arXiv

Evidence for super-exponentially accelerating atmospheric carbon dioxide growth

We analyze the growth rates of atmospheric carbon dioxide and human population, by comparing the relative merits of two benchmark models, the exponential law and the finite-time-singular (FTS) power law. The later results from positive feedbacks, either direct or mediated by other dynamical variables, as shown in our presentation of a simple endogenous macroeconomic dynamical growth model. Our empirical calibrations finds that the human population has decelerated from its previous super-exponential growth until 1960 to a slower-than-exponential growth associated with a decreasing growth rate. However, the past decade is found to be characterized by an almost stable growth rate approximately equal to r(2010) ~ 1% per year, suggesting that the population growth is stabilizing at "just" an exponential growth. As for atmospheric CO2 content, we find that it is at least exponentially increasing and most likely characterized by an accelerating growth rate as off 2009, consistent with an unsustainable FTS power law regime announcing a drastic change of regime. The coexistence of a quasi-exponential growth of human population with a super-exponential growth of carbon dioxide content in the atmosphere is a diagnostic that, until now, improvements in carbon efficiency per unit of production worldwide has been dramatically insufficient.

preprint2011arXiv

Noise-induced volatility of collective dynamics

"Noise-induced volatility" refers to a phenomenon of increased level of fluctuations in the collective dynamics of bistable units in the presence of a rapidly varying external signal, and intermediate noise levels. The archetypical signature of this phenomenon is that --beyond the increase in the level of fluctuations-- the response of the system becomes uncorrelated with the external driving force, making it different from stochastic resonance. Numerical simulations and an analytical theory of a stochastic dynamical version of the Ising model on regular and random networks demonstrate the ubiquity and robustness of this phenomenon, which is argued to be a possible cause of excess volatility in financial markets, of enhanced effective temperatures in a variety of out-of-equilibrium systems and of strong selective responses of immune systems of complex biological organisms. Extensive numerical simulations are compared with a mean-field theory for different network topologies.

preprint2011arXiv

Optimization of brain and life performance: Striving for playing at the top for the long run

In this essay written with my students and collaborators in mind, I share simple recipes that are easy and often fun to put in practice and that make a big difference in one's life. The seven guiding principles are: (1) sleep, (2) love and sex, (3) deep breathing and daily exercises, (4) water and chewing, (5) fruits, unrefined products, food combination, vitamin D and no meat, (6) power foods, (7) play, intrinsic motivation, positive psychology and will. These simple laws are based on an integration of evolutionary thinking, personal experimentation, and evidence from experiments reported in the scientific literature. I develop their rationality, expected consequences and describe briefly how to put them in practice. I hope that professionals and the broader public may also find some use for it, as I have seen already the positive impacts on some of my students.

preprint2011arXiv

Predicted and Verified Deviations from Zipf's law in Ecology of Competing Products

Zipf's power-law distribution is a generic empirical statistical regularity found in many complex systems. However, rather than universality with a single power-law exponent (equal to 1 for Zipf's law), there are many reported deviations that remain unexplained. A recently developed theory finds that the interplay between (i) one of the most universal ingredients, namely stochastic proportional growth, and (ii) birth and death processes, leads to a generic power-law distribution with an exponent that depends on the characteristics of each ingredient. Here, we report the first complete empirical test of the theory and its application, based on the empirical analysis of the dynamics of market shares in the product market. We estimate directly the average growth rate of market shares and its standard deviation, the birth rates and the "death" (hazard) rate of products. We find that temporal variations and product differences of the observed power-law exponents can be fully captured by the theory with no adjustable parameters. Our results can be generalized to many systems for which the statistical properties revealed by power law exponents are directly linked to the underlying generating mechanism.

preprint2011arXiv

Prediction of alpine glacier sliding instabilities: a new hope

Mechanical and sliding instabilities are the two processes which may lead to breaking off events of large ice masses. Mechanical instabilities mainly affect unbalanced cold hanging glaciers. For the latter case, a prediction could be achieved based on data of surface velocities and seismic activity. The case of sliding instabilities is more problematic. This phenomenon occurs on temperate glacier tongues. Such instabilities are strongly affected by the subglacial hydrology: melt water may cause (i) a lubrication of the bed and (ii) a decrease of the effective pressure and consequently a decrease of basal friction. Available data from Allalingletscher (Valais) indicate that the glacier tongue experienced an active phase during 2-3 weeks with enhanced basal motion in late summer in most years. In order to scrutinize in more detail the processes governing the sliding instabilities, a numerical model developed to investigate gravitational instabilities in heterogeneous media was applied to Allalingletscher. This model enables to account for various geometric configurations, interaction between sliding and tension cracking and water flow at the bedrock. We could show that both a critical geometrical configuration of the glacier tongue and the existence of a distributed drainage network were the main causes of this catastrophic break-off. Moreover, the analysis of the modeling results diagnose the phenomenon of recoupling of the glacier to its bed as a potential new precursory sign announcing the final break-off. This model casts a gleam of hope for a better understanding of the ultimate rupture of such glacier sliding instabilities.

preprint2011arXiv

Quis pendit ipsa pretia: facebook valuation and diagnostic of a bubble based on nonlinear demographic dynamics

We present a novel methodology to determine the fundamental value of firms in the social-networking sector based on two ingredients: (i) revenues and profits are inherently linked to its user basis through a direct channel that has no equivalent in other sectors; (ii) the growth of the number of users can be calibrated with standard logistic growth models and allows for reliable extrapolations of the size of the business at long time horizons. We illustrate the methodology with a detailed analysis of facebook, one of the biggest of the social-media giants. There is a clear signature of a change of regime that occurred in 2010 on the growth of the number of users, from a pure exponential behavior (a paradigm for unlimited growth) to a logistic function with asymptotic plateau (a paradigm for growth in competition). We consider three different scenarios, a base case, a high growth and an extreme growth scenario. Using a discount factor of 5%, a profit margin of 29% and 3.5 USD of revenues per user per year yields a value of facebook of 15.3 billion USD in the base case scenario, 20.2 billion USD in the high growth scenario and 32.9 billion USD in the extreme growth scenario. According to our methodology, this would imply that facebook would need to increase its profit per user before the IPO by a factor of 3 to 6 in the base case scenario, 2.5 to 5 in the high growth scenario and 1.5 to 3 in the extreme growth scenario in order to meet the current, widespread, high expectations. To prove the wider applicability of our methodology, the analysis is repeated on Groupon, the well-known deal-of-the-day website which is expected to go public in November 2011. The results are in line with the facebook analysis. Customer growth will plateau. By not taking this fundamental property of the growth process into consideration, estimates of its IPO are wildly overpriced.

preprint2011arXiv

Role of Diversification Risk in Financial Bubbles

We present an extension of the Johansen-Ledoit-Sornette (JLS) model to include an additional pricing factor called the "Zipf factor", which describes the diversification risk of the stock market portfolio. Keeping all the dynamical characteristics of a bubble described in the JLS model, the new model provides additional information about the concentration of stock gains over time. This allows us to understand better the risk diversification and to explain the investors' behavior during the bubble generation. We apply this new model to two famous Chinese stock bubbles, from August 2006 to October 2007 (bubble 1) and from October 2008 to August 2009 (bubble 2). The Zipf factor is found highly significant for bubble 1, corresponding to the fact that valuation gains were more concentrated on the large firms of the Shanghai index. It is likely that the widespread acknowledgement of the 80-20 rule in the Chinese media and discussion forums led many investors to discount the risk of a lack of diversification, therefore enhancing the role of the Zipf factor. For bubble 2, the Zipf factor is found marginally relevant, suggesting a larger weight of market gains on small firms. We interpret this result as the consequence of the response of the Chinese economy to the very large stimulus provided by the Chinese government in the aftermath of the 2008 financial crisis.

preprint2011arXiv

Self-Excited Multifractal Dynamics

We introduce the self-excited multifractal (SEMF) model, defined such that the amplitudes of the increments of the process are expressed as exponentials of a long memory of past increments. The principal novel feature of the model lies in the self-excitation mechanism combined with exponential nonlinearity, i.e. the explicit dependence of future values of the process on past ones. The self- excitation captures the microscopic origin of the emergent endogenous self-organization properties, such as the energy cascade in turbulent flows, the triggering of aftershocks by previous earthquakes and the "reflexive" interactions of financial markets. The SEMF process has all the standard stylized facts found in financial time series, which are robust to the specification of the parameters and the shape of the memory kernel: multifractality, heavy tails of the distribution of increments with intermediate asymptotics, zero correlation of the signed increments and long-range correlation of the squared increments, the asymmetry (called "leverage" effect) of the correlation between increments and absolute value of the increments and statistical asymmetry under time reversal.

preprint2011arXiv

Spurious trend switching phenomena in financial markets

The observation of power laws in the time to extrema of volatility, volume and intertrade times, from milliseconds to years, are shown to result straightforwardly from the selection of biased statistical subsets of realizations in otherwise featureless processes such as random walks. The bias stems from the selection of price peaks that imposes a condition on the statistics of price change and of trade volumes that skew their distributions. For the intertrade times, the extrema and power laws results from the format of transaction data.

preprint2011arXiv

Strategies used as spectroscopy of financial markets reveal new stylized facts

We propose a new set of stylized facts quantifying the structure of financial markets. The key idea is to study the combined structure of both investment strategies and prices in order to open a qualitatively new level of understanding of financial and economic markets. We study the detailed order flow on the Shenzhen Stock Exchange of China for the whole year of 2003. This enormous dataset allows us to compare (i) a closed national market (A-shares) with an international market (B-shares), (ii) individuals and institutions and (iii) real investors to random strategies with respect to timing that share otherwise all other characteristics. We find that more trading results in smaller net return due to trading frictions. We unveiled quantitative power laws with non-trivial exponents, that quantify the deterioration of performance with frequency and with holding period of the strategies used by investors. Random strategies are found to perform much better than real ones, both for winners and losers. Surprising large arbitrage opportunities exist, especially when using zero-intelligence strategies. This is a diagnostic of possible inefficiencies of these financial markets.

preprint2011arXiv

The Financial Bubble Experiment: Advanced Diagnostics and Forecasts of Bubble Terminations, Volume III

This is the third installment of the Financial Bubble Experiment. Here we provide the digital fingerprint of an electronic document in which we identify 27 bubbles in 27 different global assets; for 25 of these assets, we present windows of dates of the most likely ending time of each bubble. We will provide that document of the original analysis on 2 May 2011.

preprint2011arXiv

The Lehman Brothers Effect and Bankruptcy Cascades

Inspired by the bankruptcy of Lehman Brothers and its consequences on the global financial system, we develop a simple model in which the Lehman default event is quantified as having an almost immediate effect in worsening the credit worthiness of all financial institutions in the economic network. In our stylized description, all properties of a given firm are captured by its effective credit rating, which follows a simple dynamics of co-evolution with the credit ratings of the other firms in our economic network. The dynamics resembles the evolution of Potts spin-glass with external global field corresponding to a panic effect in the economy. The existence of a global phase transition, between paramagnetic and ferromagnetic phases, explains the large susceptibility of the system to negative shocks. We show that bailing out the first few defaulting firms does not solve the problem, but does have the effect of alleviating considerably the global shock, as measured by the fraction of firms that are not defaulting as a consequence. This beneficial effect is the counterpart of the large vulnerability of the system of coupled firms, which are both the direct consequences of the collective self-organized endogenous behaviors of the credit ratings of the firms in our economic network.

preprint2011arXiv

The US stock market leads the Federal funds rate and Treasury bond yields

Using a recently introduced method to quantify the time varying lead-lag dependencies between pairs of economic time series (the thermal optimal path method), we test two fundamental tenets of the theory of fixed income: (i) the stock market variations and the yield changes should be anti-correlated; (ii) the change in central bank rates, as a proxy of the monetary policy of the central bank, should be a predictor of the future stock market direction. Using both monthly and weekly data, we found very similar lead-lag dependence between the S&P500 stock market index and the yields of bonds inside two groups: bond yields of short-term maturities (Federal funds rate (FFR), 3M, 6M, 1Y, 2Y, and 3Y) and bond yields of long-term maturities (5Y, 7Y, 10Y, and 20Y). In all cases, we observe the opposite of (i) and (ii). First, the stock market and yields move in the same direction. Second, the stock market leads the yields, including and especially the FFR. Moreover, we find that the short-term yields in the first group lead the long-term yields in the second group before the financial crisis that started mid-2007 and the inverse relationship holds afterwards. These results suggest that the Federal Reserve is increasingly mindful of the stock market behavior, seen at key to the recovery and health of the economy. Long-term investors seem also to have been more reactive and mindful of the signals provided by the financial stock markets than the Federal Reserve itself after the start of the financial crisis. The lead of the S&P500 stock market index over the bond yields of all maturities is confirmed by the traditional lagged cross-correlation analysis.

preprint2011arXiv

Towards a Probabilistic Definition of Seizures

This writing: a) Draws attention to the intricacies inherent to the pursuit of a universal seizure definition even when powerful, well understood signal analysis methods are utilized to this end; b) Identifies this aim as a multi-objective optimization problem and discusses the advantages and disadvantages of adopting or rejecting a unitary seizure definition; c) Introduces a Probabilistic Measure of Seizure Activity to manage this thorny issue. The challenges posed by the attempt to define seizures unitarily may be partly related to their fractal properties and understood through a simplistic analogy to the so-called "Richardson effect". A revision of the time-honored conceptualization of seizures may be warranted to further advance epileptology.

preprint2011arXiv

Valuation of Zynga

On December 16, Zynga, the well-known social game developing company went public. This event is following other recent IPOs in the world of social networking companies, such as Groupon, Linkedin or Pandora to cite a few. With a valuation close to 7 billion USD at the time when it went public, Zynga has become the biggest web IPO since Google. This recent enthusiasm for social networking companies, and in particular Zynga, brings up the question whether or not they are overvalued. The common denominator of all these IPOs is that a lot of estimates about their valuation have been circulating, without any specifics given about the methodology or assumptions used to obtain those numbers. To bring more substance to the debate, we propose a two-tiered approach. First, we introduce a new model to forecast the global user base of Zynga, based on the analysis of the individual dynamics of its major games. Next, we model the revenues per user using a logistic growth function, a standard model for growth in competition. This leads to bracket the valuation of Zynga using three different scenarios (base one, optimistic and very optimistic): 4.17 billion USD in the base case, 5.16 billion in the high growth and 7.02 billion in the extreme growth scenario respectively. Thus, only the unlikely extreme growth scenario could potentially justify today's 6.6 billion USD valuation of Zynga. This suggests that Zynga at its IPO has been overpriced.

preprint2010arXiv

Diagnosis and Prediction of Tipping Points in Financial Markets: Crashes and Rebounds

By combining (i) the economic theory of rational expectation bubbles, (ii) behavioral finance on imitation and herding of investors and traders and (iii) the mathematical and statistical physics of bifurcations and phase transitions, the log-periodic power law (LPPL) model has been developed as a flexible tool to detect bubbles. The LPPL model considers the faster-than-exponential (power law with finite-time singularity) increase in asset prices decorated by accelerating oscillations as the main diagnostic of bubbles. It embodies a positive feedback loop of higher return anticipations competing with negative feedback spirals of crash expectations. The power of the LPPL model is illustrated by two recent real-life predictions performed recently by our group: the peak of the Oil price bubble in early July 2008 and the burst of a bubble on the Shanghai stock market in early August 2009. We then present the concept of "negative bubbles", which are the mirror images of positive bubbles. We argue that similar positive feedbacks are at work to fuel these accelerated downward price spirals. We adapt the LPPL model to these negative bubbles and implement a pattern recognition method to predict the end times of the negative bubbles, which are characterized by rebounds (the mirror images of crashes associated with the standard positive bubbles). The out-of-sample tests quantified by error diagrams demonstrate the high significance of the prediction performance.

preprint2010arXiv

Exuberant innovation: The Human Genome Project

We present a detailed synthesis of the development of the Human Genome Project (HGP) from 1986 to 2003 in order to test the "social bubble" hypothesis that strong social interactions between enthusiastic supporters of the HGP weaved a network of reinforcing feedbacks that led to a widespread endorsement and extraordinary commitment by those involved in the project, beyond what would be rationalized by a standard cost-benefit analysis in the presence of extraordinary uncertainties and risks. The vigorous competition and race between the initially public project and several private initiatives is argued to support the social bubble hypothesis. We also present quantitative analyses of the concomitant financial bubble concentrated on the biotech sector. Confirmation of this hypothesis is offered by the present consensus that it will take decades to exploit the fruits of the HGP, via a slow and arduous process aiming at disentangling the extraordinary complexity of the human complex body. The HGP has ushered other initiatives, based on the recognition that there is much that genomics cannot do, and that "the future belongs to proteomics". We present evidence that the competition between the public and private sector actually played in favor of the former, since its financial burden as well as its horizon was significantly reduced (for a long time against its will) by the active role of the later. This suggests that governments can take advantage of the social bubble mechanism to catalyze long-term investments by the private sector, which would not otherwise be supported.

preprint2010arXiv

How to grow a bubble: A model of myopic adapting agents

We present a simple agent-based model to study the development of a bubble and the consequential crash and investigate how their proximate triggering factor might relate to their fundamental mechanism, and vice versa. Our agents invest according to their opinion on future price movements, which is based on three sources of information, (i) public information, i.e. news, (ii) information from their "friendship" network and (iii) private information. Our bounded rational agents continuously adapt their trading strategy to the current market regime by weighting each of these sources of information in their trading decision according to its recent predicting performance. We find that bubbles originate from a random lucky streak of positive news, which, due to a feedback mechanism of these news on the agents' strategies develop into a transient collective herding regime. After this self-amplified exuberance, the price has reached an unsustainable high value, being corrected by a crash, which brings the price even below its fundamental value. These ingredients provide a simple mechanism for the excess volatility documented in financial markets. Paradoxically, it is the attempt for investors to adapt to the current market regime which leads to a dramatic amplification of the price volatility. A positive feedback loop is created by the two dominating mechanisms (adaptation and imitation) which, by reinforcing each other, result in bubbles and crashes. The model offers a simple reconciliation of the two opposite (herding versus fundamental) proposals for the origin of crashes within a single framework and justifies the existence of two populations in the distribution of returns, exemplifying the concept that crashes are qualitatively different from the rest of the price moves.

preprint2010arXiv

Icequakes coupled with surface displacements for predicting glacier break-off

A hanging glacier at the east face of Weisshorn (Switzerland) broke off in 2005. We were able to monitor and measure surface motion and icequake activity for 25 days up to three days prior to the break-off. The analysis of seismic waves generated by the glacier during the rupture maturation process revealed four types of precursory signals of the imminent catastrophic rupture: (i) an increase in seismic activity within the glacier, (ii) a decrease in the waiting time between two successive icequakes, (iii) a change in the size-frequency distribution of icequake energy, and (iv) a modification in the structure of the waiting time distributions between two successive icequakes. Morevover, it was possible to demonstrate the existence of a correlation between the seismic activity and the log-periodic oscillations of the surface velocities superimposed on the global acceleration of the glacier during the rupture maturation. Analysis of the seismic activity led us to the identification of two regimes: a stable phase with diffuse damage, and an unstable and dangerous phase characterized by a hierarchical cascade of rupture instabilities where large icequakes are triggered.

preprint2010arXiv

Inferring Fundamental Value and Crash Nonlinearity from Bubble Calibration

Identifying unambiguously the presence of a bubble in an asset price remains an unsolved problem in standard econometric and financial economic approaches. A large part of the problem is that the fundamental value of an asset is, in general, not directly observable and it is poorly constrained to calculate. Further, it is not possible to distinguish between an exponentially growing fundamental price and an exponentially growing bubble price. We present a series of new models based on the Johansen-Ledoit-Sornette (JLS) model, which is a flexible tool to detect bubbles and predict changes of regime in financial markets. Our new models identify the fundamental value of an asset price and crash nonlinearity from a bubble calibration. In addition to forecasting the time of the end of a bubble, the new models can also estimate the fundamental value and the crash nonlinearity. Besides, the crash nonlinearity obtained in the new models presents a new approach to possibly identify the dynamics of a crash after a bubble. We test the models using data from three historical bubbles ending in crashes from different markets. They are: the Hong Kong Hang Seng index 1997 crash, the S&P 500 index 1987 crash and the Shanghai Composite index 2009 crash. All results suggest that the new models perform very well in describing bubbles, forecasting their ending times and estimating fundamental value and the crash nonlinearity. The performance of the new models is tested under both the Gaussian and non-Gaussian residual assumption. Under the Gaussian residual assumption, nested hypotheses with the Wilks statistics are used and the p-values suggest that models with more parameters are necessary. Under non-Gaussian residual assumption, we use a bootstrap method to get type I and II errors of the hypotheses. All tests confirm that the generalized JLS models provide useful improvements over the standard JLS model.

preprint2010arXiv

Leverage Bubble

Leverage is strongly related to liquidity in a market and lack of liquidity is considered a cause and/or consequence of the recent financial crisis. A repurchase agreement is a financial instrument where a security is sold simultaneously with an agreement to buy it back at a later date. Repurchase agreements (repos) market size is a very important element in calculating the overall leverage in a financial market. Therefore, studying the behavior of repos market size can help to understand a process that can contribute to the birth of a financial crisis. We hypothesize that herding behavior among large investors led to massive over-leveraging through the use of repos, resulting in a bubble (built up over the previous years) and subsequent crash in this market in early 2008. We use the Johansen-Ledoit-Sornette (JLS) model of rational expectation bubbles and behavioral finance to study the dynamics of the repo market that led to the crash. The JLS model qualifies a bubble by the presence of characteristic patterns in the price dynamics, called log-periodic power law (LPPL) behavior. We show that there was significant LPPL behavior in the market before that crash and that the predicted range of times predicted by the model for the end of the bubble is consistent with the observations.

preprint2010arXiv

New Power Law Signature of Media Exposure in Human Response Waiting Time Distributions

We study the humanitarian response to the destruction brought by the tsunami generated by the Sumatra earthquake of December 26, 2004, as measured by donations, and find that it decays in time as a power law ~ 1/t^(alpha) with alpha=2.5 +/- 0.1. This behavior is suggested to be the rare outcome of a priority queuing process in which individuals execute tasks at a rate slightly faster than the rate at which new tasks arise. We believe this to be the first empirical evidence documenting this recently predicted regime, and provide additional independent evidence that suggests it arises as a result of the intense focus placed on this donation "task" by the media.

preprint2010arXiv

Prediction

This chapter first presents a rather personal view of some different aspects of predictability, going in crescendo from simple linear systems to high-dimensional nonlinear systems with stochastic forcing, which exhibit emergent properties such as phase transitions and regime shifts. Then, a detailed correspondence between the phenomenology of earthquakes, financial crashes and epileptic seizures is offered. The presented statistical evidence provides the substance of a general phase diagram for understanding the many facets of the spatio-temporal organization of these systems. A key insight is to organize the evidence and mechanisms in terms of two summarizing measures: (i) amplitude of disorder or heterogeneity in the system and (ii) level of coupling or interaction strength among the system's components. On the basis of the recently identified remarkable correspondence between earthquakes and seizures, we present detailed information on a class of stochastic point processes that has been found to be particularly powerful in describing earthquake phenomenology and which, we think, has a promising future in epileptology. The so-called self-exciting Hawkes point processes capture parsimoniously the idea that events can trigger other events, and their cascades of interactions and mutual influence are essential to understand the behavior of these systems.

preprint2010arXiv

Quantification of deviations from rationality with heavy-tails in human dynamics

The dynamics of technological, economic and social phenomena is controlled by how humans organize their daily tasks in response to both endogenous and exogenous stimulations. Queueing theory is believed to provide a generic answer to account for the often observed power-law distributions of waiting times before a task is fulfilled. However, the general validity of the power law and the nature of other regimes remain unsettled. Using anonymized data collected by Google at the World Wide Web level, we identify the existence of several additional regimes characterizing the time required for a population of Internet users to execute a given task after receiving a message. Depending on the under- or over-utilization of time by the population of users and the strength of their response to perturbations, the pure power law is found to be coextensive with an exponential regime (tasks are performed without too much delay) and with a crossover to an asymptotic plateau (some tasks are never performed). The characterization of the availability and efficiency of humans on their actions revealed by our study have important consequences to understand human decision-making, optimal designs of policies such as for Internet security, with spillovers to collective behaviors, crowds dynamics, and social epidemics.

preprint2010arXiv

Segmentation of Fault Networks Determined from Spatial Clustering of Earthquakes

We present a new method of data clustering applied to earthquake catalogs, with the goal of reconstructing the seismically active part of fault networks. We first use an original method to separate clustered events from uncorrelated seismicity using the distribution of volumes of tetrahedra defined by closest neighbor events in the original and randomized seismic catalogs. The spatial disorder of the complex geometry of fault networks is then taken into account by defining faults as probabilistic anisotropic kernels, whose structures are motivated by properties of discontinuous tectonic deformation and previous empirical observations of the geometry of faults and of earthquake clusters at many spatial and temporal scales. Combining this a priori knowledge with information theoretical arguments, we propose the Gaussian mixture approach implemented in an Expectation-Maximization (EM) procedure. A cross-validation scheme is then used and allows the determination of the number of kernels that should be used to provide an optimal data clustering of the catalog. This three-steps approach is applied to a high quality relocated catalog of the seismicity following the 1986 Mount Lewis ($M_l=5.7$) event in California and reveals that events cluster along planar patches of about 2 km$^2$, i.e. comparable to the size of the main event. The finite thickness of those clusters (about 290 m) suggests that events do not occur on well-defined euclidean fault core surfaces, but rather that the damage zone surrounding faults may be seismically active at depth. Finally, we propose a connection between our methodology and multi-scale spatial analysis, based on the derivation of spatial fractal dimension of about 1.8 for the set of hypocenters in the Mnt Lewis area, consistent with recent observations on relocated catalogs.

preprint2010arXiv

Super-extreme event's influence on a Weierstrass-Mandelbrot Continuous-Time Random Walk

Two utmost cases of super-extreme event's influence on the velocity autocorrelation function (VAF) were considered. The VAF itself was derived within the hierarchical Weierstrass-Mandelbrot Continuous-Time Random Walk (WM-CTRW) formalism, which is able to cover a broad spectrum of continuous-time random walks. Firstly, we studied a super-extreme event in a form of a sustained drift, whose duration time is much longer than that of any other event. Secondly, we considered a super-extreme event in the form of a shock with the size and velocity much larger than those corresponding to any other event. We found that the appearance of these super-extreme events substantially changes the results determined by extreme events (the so called "black swans") that are endogenous to the WM-CTRW process. For example, changes of the VAF in the latter case are in the form of some instability and distinctly differ from those caused in the former case. In each case these changes are quite different compared to the situation without super-extreme events suggesting the possibility to detect them in natural system if they occur.

preprint2010arXiv

The Financial Bubble Experiment: advanced diagnostics and forecasts of bubble terminations

On 2 November 2009, the Financial Bubble Experiment was launched within the Financial Crisis Observatory (FCO) at ETH Zurich (\url{http://www.er.ethz.ch/fco/}). In that initial report, we diagnosed and announced three bubbles on three different assets. In this latest release of 23 December 2009 in this ongoing experiment, we add a diagnostic of a new bubble developing on a fourth asset.

preprint2010arXiv

The Financial Bubble Experiment: Advanced Diagnostics and Forecasts of Bubble Terminations Volume II-Master Document

This is the second installment of the Financial Bubble Experiment. Here we provide the digital fingerprint of an electronic document in which we identify 7 bubbles in 7 different global assets; for 4 of these assets, we present windows of dates of the most likely ending time of each bubble. We will provide that document of the original analysis on 1 November 2010.

preprint2009arXiv

Bubble Diagnosis and Prediction of the 2005-2007 and 2008-2009 Chinese stock market bubbles

By combining (i) the economic theory of rational expectation bubbles, (ii) behavioral finance on imitation and herding of investors and traders and (iii) the mathematical and statistical physics of bifurcations and phase transitions, the log-periodic power law model has been developed as a flexible tool to detect bubbles. The LPPL model considers the faster-than-exponential (power law with finite-time singularity) increase in asset prices decorated by accelerating oscillations as the main diagnostic of bubbles. It embodies a positive feedback loop of higher return anticipations competing with negative feedback spirals of crash expectations. We use the LPPL model in one of its incarnations to analyze two bubbles and subsequent market crashes in two important indexes in the Chinese stock markets between May 2005 and July 2009. Both the Shanghai Stock Exchange Composite and Shenzhen Stock Exchange Component indexes exhibited such behavior in two distinct time periods: 1) from mid-2005, bursting in Oct. 2007 and 2) from Nov. 2008, bursting in the beginning of Aug. 2009. We successfully predicted time windows for both crashes in advance with the same methods used to successfully predict the peak in mid-2006 of the US housing bubble and the peak in July 2008 of the global oil bubble. The more recent bubble in the Chinese indexes was detected and its end or change of regime was predicted independently by two groups with similar results, showing that the model has been well-documented and can be replicated by industrial practitioners. Here we present more detailed analysis of the individual Chinese index predictions and of the methods used to make and test them.

preprint2009arXiv

Dragon-Kings, Black Swans and the Prediction of Crises

We develop the concept of ``dragon-kings'' corresponding to meaningful outliers, which are found to coexist with power laws in the distributions of event sizes under a broad range of conditions in a large variety of systems. These dragon-kings reveal the existence of mechanisms of self-organization that are not apparent otherwise from the distribution of their smaller siblings. We present a generic phase diagram to explain the generation of dragon-kings and document their presence in six different examples (distribution of city sizes, distribution of acoustic emissions associated with material failure, distribution of velocity increments in hydrodynamic turbulence, distribution of financial drawdowns, distribution of the energies of epileptic seizures in humans and in model animals, distribution of the earthquake energies). We emphasize the importance of understanding dragon-kings as being often associated with a neighborhood of what can be called equivalently a phase transition, a bifurcation, a catastrophe (in the sense of Rene Thom), or a tipping point. The presence of a phase transition is crucial to learn how to diagnose in advance the symptoms associated with a coming dragon-king. Several examples of predictions using the derived log-periodic power law method are discussed, including material failure predictions and the forecasts of the end of financial bubbles.

preprint2009arXiv

Financial Bubbles, Real Estate bubbles, Derivative Bubbles, and the Financial and Economic Crisis

The financial crisis of 2008, which started with an initially well-defined epicenter focused on mortgage backed securities (MBS), has been cascading into a global economic recession, whose increasing severity and uncertain duration has led and is continuing to lead to massive losses and damage for billions of people. Heavy central bank interventions and government spending programs have been launched worldwide and especially in the USA and Europe, with the hope to unfreeze credit and boltster consumption. Here, we present evidence and articulate a general framework that allows one to diagnose the fundamental cause of the unfolding financial and economic crisis: the accumulation of several bubbles and their interplay and mutual reinforcement has led to an illusion of a "perpetual money machine" allowing financial institutions to extract wealth from an unsustainable artificial process. Taking stock of this diagnostic, we conclude that many of the interventions to address the so-called liquidity crisis and to encourage more consumption are ill-advised and even dangerous, given that precautionary reserves were not accumulated in the "good times" but that huge liabilities were. The most "interesting" present times constitute unique opportunities but also great challenges, for which we offer a few recommendations.

preprint2009arXiv

Gravity-driven instabilities: interplay between state-and-velocity dependent frictional sliding and stress corrosion damage cracking

We model the progressive maturation of a heterogeneous mass towards a gravity-driven instability, characterized by the competition between frictional sliding and tension cracking, using array of slider blocks on an inclined basal surface, which interact via elastic-brittle springs. A realistic state- and rate-dependent friction law describes the block-surface interaction. The inner material damage occurs via stress corrosion. Three regimes, controlling the mass instability and its precursory behavior, are classified as a function of the ratio $T_c/T_f$ of two characteristic time scales associated with internal damage/creep and with frictional sliding. For $T_c/T_f \gg 1$, the whole mass undergoes a series of internal stick and slip events, associated with an initial slow average downward motion of the whole mass, and progressively accelerates until a global coherent runaway is observed. For $T_c/T_f \ll 1$, creep/damage occurs sufficiently fast compared with nucleation of sliding, causing bonds to break, and the bottom part of the mass undergoes a fragmentation process with the creation of a heterogeneous population of sliding blocks. For the intermediate regime $T_c/T_f \sim 1$, a macroscopic crack nucleates and propagates along the location of the largest curvature associated with the change of slope from the stable frictional state in the upper part to the unstable frictional sliding state in the lower part. The other important parameter is the Young modulus $Y$ which controls the correlation length of displacements in the system.

preprint2009arXiv

Homogeneous Volatility Bridge Estimators

We present a theory of homogeneous volatility bridge estimators for log-price stochastic processes. The main tool of our theory is the parsimonious encoding of the information contained in the open, high and low prices of incomplete bridge, corresponding to given log-price stochastic process, and in its close value, for a given time interval. The efficiency of the new proposed estimators is favorably compared with that of the Garman-Klass and Parkinson estimators.

preprint2009arXiv

Other-regarding preferences and altruistic punishment: A Darwinian perspective

This article examines the effect of different other-regarding preference types on the emergence of altruistic punishment behavior from an evolutionary perspective. Our findings corroborate, complement, and interlink the experimental and theoretical literature that has shown the importance of other-regarding behavior in various decision settings. We find that a selfish variant of inequity aversion is sufficient to quantitatively explain the level of punishment observed in contemporary experiments: If disadvantageous inequity aversion is the predominant preference type, altruistic punishment emerges in our model to a level that precisely matches the empirical observations. We use a new approach that closely combines empirical results from a public goods experiment together with an evolutionary simulation model. Hereby we apply ideas from behavioral economics, complex system science, and evolutionary biology.

preprint2008arXiv

Look-Ahead Benchmark Bias in Portfolio Performance Evaluation

Performance of investment managers are evaluated in comparison with benchmarks, such as financial indices. Due to the operational constraint that most professional databases do not track the change of constitution of benchmark portfolios, standard tests of performance suffer from the "look-ahead benchmark bias," when they use the assets constituting the benchmarks of reference at the end of the testing period, rather than at the beginning of the period. Here, we report that the "look-ahead benchmark bias" can exhibit a surprisingly large amplitude for portfolios of common stocks (up to 8% annum for the S&P500 taken as the benchmark) -- while most studies have emphasized related survival biases in performance of mutual and hedge funds for which the biases can be expected to be even larger. We use the CRSP database from 1926 to 2006 and analyze the running top 500 US capitalizations to demonstrate that this bias can account for a gross overestimation of performance metrics such as the Sharpe ratio as well as an underestimation of risk, as measured for instance by peak-to-valley drawdowns. We demonstrate the presence of a significant bias in the estimation of the survival and look-ahead biases studied in the literature. A general methodology to test the properties of investment strategies is advanced in terms of random strategies with similar investment constraints.

preprint2007arXiv

A General Strategy for Physics-Based Model Validation Illustrated with Earthquake Phenomenology, Atmospheric Radiative Transfer, and Computational Fluid Dynamics

Validation is often defined as the process of determining the degree to which a model is an accurate representation of the real world from the perspective of its intended uses. Validation is crucial as industries and governments depend increasingly on predictions by computer models to justify their decisions. In this article, we survey the model validation literature and propose to formulate validation as an iterative construction process that mimics the process occurring implicitly in the minds of scientists. We thus offer a formal representation of the progressive build-up of trust in the model, and thereby replace incapacitating claims on the impossibility of validating a given model by an adaptive process of constructive approximation. This approach is better adapted to the fuzzy, coarse-grained nature of validation. Our procedure factors in the degree of redundancy versus novelty of the experiments used for validation as well as the degree to which the model predicts the observations. We illustrate the new methodology first with the maturation of Quantum Mechanics as the arguably best established physics theory and then with several concrete examples drawn from some of our primary scientific interests: a cellular automaton model for earthquakes, an anomalous diffusion model for solar radiation transport in the cloudy atmosphere, and a computational fluid dynamics code for the Richtmyer-Meshkov instability. This article is an augmented version of Sornette et al. [2007] that appeared in Proceedings of the National Academy of Sciences in 2007 (doi: 10.1073/pnas.0611677104), with an electronic supplement at URL http://www.pnas.org/cgi/content/full/0611677104/DC1. Sornette et al. [2007] is also available in preprint form at physics/0511219.

preprint2007arXiv

Automatic Reconstruction of Fault Networks from Seismicity Catalogs: 3D Optimal Anisotropic Dynamic Clustering

We propose a new pattern recognition method that is able to reconstruct the 3D structure of the active part of a fault network using the spatial location of earthquakes. The method is a generalization of the so-called dynamic clustering method, that originally partitions a set of datapoints into clusters, using a global minimization criterion over the spatial inertia of those clusters. The new method improves on it by taking into account the full spatial inertia tensor of each cluster, in order to partition the dataset into fault-like, anisotropic clusters. Given a catalog of seismic events, the output is the optimal set of plane segments that fits the spatial structure of the data. Each plane segment is fully characterized by its location, size and orientation. The main tunable parameter is the accuracy of the earthquake localizations, which fixes the resolution, i.e. the residual variance of the fit. The resolution determines the number of fault segments needed to describe the earthquake catalog, the better the resolution, the finer the structure of the reconstructed fault segments. The algorithm reconstructs successfully the fault segments of synthetic earthquake catalogs. Applied to the real catalog constituted of a subset of the aftershocks sequence of the 28th June 1992 Landers earthquake in Southern California, the reconstructed plane segments fully agree with faults already known on geological maps, or with blind faults that appear quite obvious on longer-term catalogs. Future improvements of the method are discussed, as well as its potential use in the multi-scale study of the inner structure of fault zones.

preprint2003arXiv

Are Aftershocks of Large Californian Earthquakes Diffusing?

We analyze 21 aftershock sequences of California to test for evidence of space-time diffusion. Aftershock diffusion may result from stress diffusion and is also predicted by any mechanism of stress weakening. Here, we test an alternative mechanism to explain aftershock diffusion, based on multiple cascades of triggering. In order to characterize aftershock diffusion, we develop two methods, one based on a suitable time and space windowing, the other using a wavelet transform adapted to the removal of background seismicity. Both methods confirm that diffusion of seismic activity is very weak, much weaker than reported in previous studies. A possible mechanism explaining the weakness of observed diffusion is the effect of geometry, including the localization of aftershocks on a fractal fault network and the impact of extended rupture lengths which control the typical distances of interaction between earthquakes.

preprint2003arXiv

Foreshocks Explained by Cascades of Triggered Seismicity

The observation of foreshocks preceding large earthquakes and the suggestion that foreshocks have specific properties that may be used to distinguish them from other earthquakes have raised the hope that large earthquakes may be predictable. Among proposed anomalous properties are the larger proportion than normal of large versus small foreshocks, the power law acceleration of seismicity rate as a function of time to the mainshock and the spatial migration of foreshocks toward the mainshock, when averaging over many sequences. Using Southern California seismicity, we show that these properties and others arise naturally from the simple model that any earthquake may trigger other earthquakes, without arbitrary distinction between foreshocks, aftershocks and mainshocks. We find that foreshocks precursory properties are independent of the mainshock size. This implies that earthquakes (large or small) are predictable to the same degree as seismicity rate is predictable from past seismicity by taking into account cascades of triggering. The cascades of triggering give rise naturally to long-range and long-time interactions, which can explain the observations of correlations in seismicity over surprisingly large length scales.

preprint2003arXiv

Importance of direct and indirect triggered seismicity

Using the simple ETAS branching model of seismicity, which assumes that each earthquake can trigger other earthquakes, we quantify the role played by the cascade of triggered seismicity in controlling the rate of aftershock decay as well as the overall level of seismicity in the presence of a constant external seismicity source. We show that, in this model, the fraction of earthquakes in the population that are aftershocks is equal to the fraction of aftershocks that are indirectly triggered and is given by the average number of triggered events per earthquake. Previous observations that a significant fraction of earthquakes are triggered earthquakes therefore imply that most aftershocks are indirectly triggered by the mainshock.

preprint1998arXiv

Discrete scale invariance and complex dimensions

We discuss the concept of discrete scale invariance and how it leads to complex critical exponents (or dimensions), i.e. to the log-periodic corrections to scaling. After their initial suggestion as formal solutions of renormalization group equations in the seventies, complex exponents have been studied in the eighties in relation to various problems of physics embedded in hierarchical systems. Only recently has it been realized that discrete scale invariance and its associated complex exponents may appear ``spontaneously'' in euclidean systems, i.e. without the need for a pre-existing hierarchy. Examples are diffusion-limited-aggregation clusters, rupture in heterogeneous systems, earthquakes, animals (a generalization of percolation) among many other systems. We review the known mechanisms for the spontaneous generation of discrete scale invariance and provide an extensive list of situations where complex exponents have been found. This is done in order to provide a basis for a better fundamental understanding of discrete scale invariance. The main motivation to study discrete scale invariance and its signatures is that it provides new insights in the underlying mechanisms of scale invariance. It may also be very interesting for prediction purposes.

preprint1998arXiv

Large deviations and portfolio optimization

Risk control and optimal diversification constitute a major focus in the finance and insurance industries as well as, more or less consciously, in our everyday life. We present a discussion of the characterization of risks and of the optimization of portfolios that starts from a simple illustrative model and ends by a general functional integral formulation. A major theme is that risk, usually thought one-dimensional in the conventional mean-variance approach, has to be addressed by the full distribution of losses. Furthermore, the time-horizon of the investment is shown to play a major role. We show the importance of accounting for large fluctuations and use the theory of Cramér for large deviations in this context. We first treat a simple model with a single risky asset that examplifies the distinction between the average return and the typical return, the role of large deviations in multiplicative processes, and the different optimal strategies for the investors depending on their size. We then analyze the case of assets whose price variations are distributed according to exponential laws, a situation that is found to describe reasonably well daily price variations. Several portfolio optimization strategies are presented that aim at controlling large risks. We end by extending the standard mean-variance portfolio optimization theory, first within the quasi-Gaussian approximation and then using a general formulation for non-Gaussian correlated assets in terms of the formalism of functional integrals developed in the field theory of critical phenomena.

preprint1997arXiv

Large financial crashes

We propose that large stock market crashes are analogous to critical points studied in statistical physics with log-periodic correction to scaling. We extend our previous renormalization group model of stock market prices prior to and after crashes [D. Sornette et al., J.Phys.I France 6, 167, 1996] by including the first non-linear correction. This predicts the existence of a log-frequency shift over time in the log-periodic oscillations prior to a crash. This is tested on the two largest historical crashes of the century, the october 1929 and october 1987 crashes, by fitting the stock market index over an interval of 8 years prior to the crashes. The good quality of the fits, as well as the consistency of the parameter values obtained from the two crashes, promote the theory that crashes have their origin in the collective ``crowd'' behavior of many interacting agents.

preprint1997arXiv

Log-periodic Oscillations for Biased Diffusion in 3D Random Lattices

Random walks with a fixed bias direction on randomly diluted cubic lattices far above the percolation threshold exhibit log-periodic oscillations in the effective exponent versus time. A scaling argument accounts for the numerical results in the limit of large biases and small dilution and shows the importance of the interplay of these two ingredients in the generation of the log-periodicity. These results show that log-periodicity is the dominant effect compared to previous predictions of and reports on anomalous diffusion.

preprint1995arXiv

Faults Self-Organized by Repeated Earthquakes in a Quasi-Static Antiplane Crack Model

We study a 2D quasi-static discrete {\it crack} anti-plane model of a tectonic plate with long range elastic forces and quenched disorder. The plate is driven at its border and the load is transfered to all elements through elastic forces. This model can be considered as belonging to the class of self-organized models which may exhibit spontaneous criticality, with four additional ingredients compared to sandpile models, namely quenched disorder, boundary driving, long range forces and fast time crack rules. In this ''crack'' model, as in the ''dislocation'' version previously studied, we find that the occurrence of repeated earthquakes organizes the activity on well-defined fault-like structures. In contrast with the ''dislocation'' model, after a transient, the time evolution becomes periodic with run-aways ending each cycle. This stems from the ''crack'' stress transfer rule preventing criticality to organize in favor of cyclic behavior. For sufficiently large disorder and weak stress drop, these large events are preceded by a complex space-time history of foreshock activity, characterized by a Gutenberg-Richter power law distribution with universal exponent $B=1 \pm 0.05$. This is similar to a power law distribution of small nucleating droplets before the nucleation of the macroscopic phase in a first-order phase transition. For large disorder and large stress drop, and for certain specific initial disorder configurations, the stress field becomes frustrated in fast time : out-of-plane deformations (thrust and normal faulting) and/or a genuine dynamics must be introduced to resolve this frustration.

preprint1995arXiv

Rank-Ordering Statistics of Extreme Events: Application to the Distribution of Large Earthquakes

Rank-ordering statistics provides a perspective on the rare, largest elements of a population, whereas the statistics of cumulative distributions are dominated by the more numerous small events. The exponent of a power law distribution can be determined with good accuracy by rank-ordering statistics from the observation of only a few tens of the largest events. Using analytical results and synthetic tests, we quantify the systematic and the random errors. We also study the case of a distribution defined by two branches, each having a power law distribution, one defined for the largest events and the other for smaller events, with application to the World-Wide (Harvard) and Southern California earthquake catalogs. In the case of the Harvard moment catalog, we make more precise earlier claims of the existence of a transition of the earthquake magnitude distribution between small and large earthquakes; the $b$-values are $b_2 = 2.3 \pm 0.3$ for large shallow earthquakes and $b_1 = 1.00 \pm 0.02$ for smaller shallow earthquakes. However, the cross-over magnitude between the two distributions is ill-defined. The data available at present do not provide a strong constraint on the cross-over which has a $50\%$ probability of being between magnitudes $7.1$ and $7.6$ for shallow earthquakes; this interval may be too conservatively estimated. Thus, any influence of a universal geometry of rupture on the distribution of earthquakes world-wide is ill-defined at best. We caution that there is no direct evidence to confirm the hypothesis that the large-moment branch is indeed a power law. In fact, a gamma distribution fits the entire suite of earthquake moments from the smallest to the largest satisfactorily. There is no evidence that the earthquakes of the Southern California catalog have a distribution with two