Source author record

Frank Raischel

Frank Raischel appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

17works
14topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

17 published item(s)

preprint2016arXiv

Record breaking bursts during the compressive failure of porous materials

An accurate understanding of the interplay between random and deterministic processes in generating extreme events is of critical importance in many fields, from forecasting extreme meteorological events to the catastrophic failure of materials and in the Earth. Here we investigate the statistics of record-breaking events in the time series of crackling noise generated by local rupture events during the compressive failure of porous materials. The events are generated by computer simulations of the uni-axial compression of cylindrical samples in a discrete element model of sedimentary rocks that closely resemble those of real experiments. The number of records grows initially as a decelerating power law of the number of events, followed by an acceleration immediately prior to failure. We demonstrate the existence of a characteristic record rank k^* which separates the two regimes of the time evolution. Up to this rank deceleration occurs due to the effect of random disorder. Record breaking then accelerates towards macroscopic failure, when physical interactions leading to spatial and temporal correlations dominate the location and timing of local ruptures. Sub-sequences of bursts between consecutive records are characterized by a power law size distribution with an exponent which decreases as failure is approached. High rank records are preceded by bursts of increasing size and waiting time between consecutive events and they are followed by a relaxation process. As a reference, surrogate time series are generated by reshuffling the crackling bursts. The record statistics of the uncorrelated surrogates agrees very well with the corresponding predictions of independent identically distributed random variables, which confirms that the temporal and spatial correlation of cracking bursts are responsible for the observed unique behaviour.

preprint2015arXiv

From empirical data to continuous Markov processes: a systematic approach

We present an approach for testing for the existence of continuous generators of discrete stochastic transition matrices. Typically, the known approaches to ascertain the existence of continuous Markov processes are based in the assumption that only time-homogeneous generators exist. Here, a systematic extension to time-inhomogeneity is presented, based in new mathematical propositions incorporating necessary and sufficient conditions, which are then implemented computationally and applied to numerical data. A discussion concerning the bridging between rigorous mathematical results on the existence of generators to its computational implementation. Our detection algorithm shows to be effective in more than $80\%$ of tested matrices, typically $90\%$ to $95\%$, and for those an estimate of the (non-homogeneous) generator matrix follows. We also solve the embedding problem analytically for the particular case of three-dimensional circulant matrices. Finally, a discussion of possible applications of our framework to problems in different fields is briefly addressed.

preprint2015arXiv

Parameter-free resolution of the superposition of stochastic signals

This paper presents a direct method to obtain the deterministic and stochastic contribution of the sum of two independent sets of stochastic processes, one of which is composed by Ornstein-Uhlenbeck processes and the other being a general (non-linear) Langevin process. The method is able to distinguish between all stochastic process, retrieving their corresponding stochastic evolution equations. This framework is based on a recent approach for the analysis of multidimensional Langevin-type stochastic processes in the presence of strong measurement (or observational) noise, which is here extended to impose neither constraints nor parameters and extract all coefficients directly from the empirical data sets. Using synthetic data, it is shown that the method yields satisfactory results.

preprint2015arXiv

Uncovering the evolution of non-stationary stochastic variables: the example of asset volume-price fluctuations

We present a framework for describing the evolution of stochastic observables having a non-stationary distribution of values. The framework is applied to empirical volume-prices from assets traded at the New York stock exchange. Using Kullback-Leibler divergence we evaluate the best model out from four biparametric models standardly used in the context of financial data analysis. In our present data sets we conclude that the inverse $Γ$-distribution is a good model, particularly for the distribution tail of the largest volume-price fluctuations. Extracting the time-series of the corresponding parameter values we show that they evolve in time as stochastic variables themselves. For the particular case of the parameter controlling the volume-price distribution tail we are able to extract an Ornstein-Uhlenbeck equation which describes the fluctuations of the largest volume-prices observed in the data. Finally, we discuss how to bridge from the stochastic evolution of the distribution parameters to the stochastic evolution of the (non-stationary) observable and put our conclusions into perspective for other applications in geophysics and biology.

preprint2014arXiv

A thermostatistical approach to scale-free networks

We describe an ensemble of growing scale-free networks in an equilibrium framework, providing insight into why the exponent of empirical scale-free networks in nature is typically robust. In an analogy to thermostatistics, to describe the canonical and microcanonical ensembles, we introduce a functional, whose maximum corresponds to a scale-free configuration. We then identify the equivalents to energy, Zeroth-law, entropy and heat capacity for scale-free networks. Discussing the merging of scale-free networks, we also establish an exact relation to predict their final "equilibrium" degree exponent. All analytic results are complemented with Monte Carlo simulations. Our approach illustrates the possibility to apply the tools of equilibrium statistical physics to study the properties of growing networks, and it also supports the recent arguments on the complementarity between equilibrium and nonequilibrium systems.

preprint2014arXiv

Are credit ratings time-homogeneous and Markov?

We introduce a simple approach for testing the reliability of homogeneous generators and the Markov property of the stochastic processes underlying empirical time series of credit ratings. We analyze open access data provided by Moody's and show that the validity of these assumptions - existence of a homogeneous generator and Markovianity - is not always guaranteed. Our analysis is based on a comparison between empirical transition matrices aggregated over fixed time windows and candidate transition matrices generated from measurements taken over shorter periods. Ratings are widely used in credit risk, and are a key element in risk assessment; our results provide a tool for quantifying confidence in predictions extrapolated from these time series.

preprint2014arXiv

Daily pollution forecast using optimal meteorological data at synoptic and local scales

We present a simple framework to easily pre-select the most essential data for accurately forecasting the concentration of the pollutant PM$_{10}$, based on pollutants observations for the years 2002 until 2006 in the metropolitan region of Lisbon, Portugal. Starting from a broad panoply of different data sets collected at several meteorological stations, we apply a forward stepwise regression procedure that enables us not only to identify the most important variables for forecasting the pollutant but also to rank them in order of importance. We argue the importance of this variable ranking, showing that the ranking is very sensitive to the urban spot where measurements are taken. Having this pre-selection, we then present the potential of linear and non-linear neural network models when applied to the concentration of pollutant PM$_{10}$. Similarly to previous studies for other pollutants, our validation results show that non-linear models in average perform as well or worse as linear models for PM$_{10}$. Finally, we also address the influence of Circulation Weather Types, characterizing synoptic scale circulation patterns and the concentration of pollutants.

preprint2014arXiv

From human mobility to renewable energies: Big data analysis to approach worldwide multiscale phenomena

We address and discuss recent trends in the analysis of big data sets, with the emphasis on studying multiscale phenomena. Applications of big data analysis in different scientific fields are described and two particular examples of multiscale phenomena are explored in more detail. The first one deals with wind power production at the scale of single wind turbines, the scale of entire wind farms and also at the scale of a whole country. Using open source data we show that the wind power production has an intermittent character at all those three scales, with implications for defining adequate strategies for stable energy production. The second example concerns the dynamics underlying human mobility, which presents different features at different scales. For that end, we analyze $12$-month data of the Eduroam database within Portuguese universities, and find that, at the smallest scales, typically within a set of a few adjacent buildings, the characteristic exponents of average displacements are different from the ones found at the scale of one country or one continent.

preprint2014arXiv

Modeling and analysis of cyclic inhomogeneous Markov processes: a wind turbine case study

A method is proposed to reconstruct a cyclic time-inhomogeneous Markov pro- cess from measured data. First, a time-inhomogeneous Markov model is fit to the data, taken here from measurements on a wind turbine. From the time-dependent transition matrices, the time-dependent Kramers-Moyal coefficients of the corresponding stochastic process are computed. Further applications of this method are discussed.

preprint2014arXiv

Modeling the functional network of primary intercellular Ca$^{2+}$ wave propagation in astrocytes and its application to study drug effects

We introduce a simple procedure of multivariate signal analysis to uncover the functional connectivity among cells composing a living tissue and describe how to apply it for extracting insight on the effect of drugs in the tissue. The procedure is based on the covariance matrix of time resolved activity signals. By determining the time-lag that maximizes covariance, one derives the weight of the corresponding connection between cells. Introducing simple constraints, it is possible to conclude whether pairs of cells are functionally connected and in which direction. After testing the method against synthetic data we apply it to study intercellular propagation of Ca$^{2+}$ waves in astrocytes following an external stimulus, with the aim of uncovering the functional cellular connectivity network. Our method proves to be particularly suited for this type of networking signal propagation where signals are pulse-like and have short time-delays, and is shown to be superior to standard methods, namely a multivariate Granger algorithm. Finally, based the statistical analysis of the connection weight distribution, we propose simple measures for assessing the impact of drugs on the functional connectivity between cells.

preprint2014arXiv

Optimal models of extreme volume-prices are time-dependent

We present evidence that the best model for empirical volume-price distributions is not always the same and it strongly depends in (i) the region of the volume-price spectrum that one wants to model and (ii) the period in time that is being modelled. To show these two features we analyze stocks of the New York stock market with four different models: Gamma, inverse-gamma, log-normal, and Weibull distributions. To evaluate the accuracy of each model we use standard relative deviations as well as the Kullback-Leibler distance and introduce an additional distance particularly suited to evaluate how accurate are the models for the distribution tails (large volume-price). Finally we put our findings in perspective and discuss how they can be extended to other situations in finance engineering.

preprint2014arXiv

Principal wind turbines for a conditional portfolio approach to wind farms

We introduce a measure for estimating the best risk-return relation of power production in wind farms within a given time-lag, conditioned to the velocity field. The velocity field is represented by a scalar that weighs the influence of the velocity at each wind turbine at present and previous time-steps for the present "state" of the wind field. The scalar measure introduced is a linear combination of the few turbines, that most influence the overall power production. This quantity is then used as the condition for computing a conditional expected return and corresponding risk associated to the future total power output.

preprint2014arXiv

Stochastic Evolution of Stock Market Volume-Price Distributions

Using available data from the New York stock market (NYSM) we test four different bi-parametric models to fit the correspondent volume-price distributions at each $10$-minute lag: the Gamma distribution, the inverse Gamma distribution, the Weibull distribution and the log-normal distribution. The volume-price data, which measures market capitalization, appears to follow a specific statistical pattern, other than the evolution of prices measured in similar studies. We find that the inverse Gamma model gives a superior fit to the volume-price evolution than the other models. We then focus on the inverse Gamma distribution as a model for the NYSM data and analyze the evolution of the pair of distribution parameters as a stochastic process. Assuming that the evolution of these parameters is governed by coupled Langevin equations, we derive the corresponding drift and diffusion coefficients, which then provide insight for understanding the mechanisms underlying the evolution of the stock market.

preprint2013arXiv

Air quality prediction using optimal neural networks with stochastic variables

We apply recent methods in stochastic data analysis for discovering a set of few stochastic variables that represent the relevant information on a multivariate stochastic system, used as input for artificial neural networks models for air quality forecast. We show that using these derived variables as input variables for training the neural networks it is possible to significantly reduce the amount of input variables necessary for the neural network model, without considerably changing the predictive power of the model. The reduced set of variables including these derived variables is therefore proposed as optimal variable set for training neural networks models in forecasting geophysical and weather properties. Finally, we briefly discuss other possible applications of such optimized neural network models.

preprint2013arXiv

Uncovering wind turbine properties through two-dimensional stochastic modeling of wind dynamics

Using a method for stochastic data analysis, borrowed from statistical physics, we analyze synthetic data from a Markov chain model that reproduces measurements of wind speed and power production in a wind park in Portugal. We first show that our analysis retrieves indeed the power performance curve, which yields the relationship between wind speed and power production and we discuss how this procedure can be extended for extracting unknown functional relationships between pairs of physical variables in general. Second, we show how specific features, such as the rated speed of the wind turbine or the descriptive wind speed statistics, can be related with the equations describing the evolution of power production and wind speed at single wind turbines.

preprint2012arXiv

Searching for optimal variables in real multivariate stochastic data

By implementing a recent technique for the determination of stochastic eigendirections of two coupled stochastic variables, we investigate the evolution of fluctuations of NO2 concentrations at two monitoring stations in the city of Lisbon, Portugal. We analyze the stochastic part of the measurements recorded at the monitoring stations by means of a method where the two concentrations are considered as stochastic variables evolving according to a system of coupled stochastic differential equations. Analysis of their structure allows for transforming the set of measured variables to a set of derived variables, one of them with reduced stochasticity. For the specific case of NO2 concentration measures, the set of derived variables are well approximated by a global rotation of the original set of measured variables. We conclude that the stochastic sources at each station are independent from each other and typically have amplitudes of the order of the deterministic contributions. Such findings show significant limitations when predicting such quantities. Still, we briefly discuss how predictive power can be increased in general in the light of our methods.

preprint2010arXiv

Quantitative analysis of numerical estimates for the permeability of porous media from lattice-Boltzmann simulations

During the last decade, lattice-Boltzmann (LB) simulations have been improved to become an efficient tool for determining the permeability of porous media samples. However, well known improvements of the original algorithm are often not implemented. These include for example multirelaxation time schemes or improved boundary conditions, as well as different possibilities to impose a pressure gradient. This paper shows that a significant difference of the calculated permeabilities can be found unless one uses a carefully selected setup. We present a detailed discussion of possible simulation setups and quantitative studies of the influence of simulation parameters. We illustrate our results by applying the algorithm to a Fontainebleau sandstone and by comparing our benchmark studies to other numerical permeability measurements in the literature.