Source author record

Mikhail Kanevski

Mikhail Kanevski appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
11topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2022arXiv

Spatio-temporal estimation of wind speed and wind power using machine learning: predictions, uncertainty and technical potential

The growth of wind generation capacities in the past decades has shown that wind energy can contribute to the energy transition in many parts of the world. Being highly variable and complex to model, the quantification of the spatio-temporal variation of wind power and the related uncertainty is highly relevant for energy planners. Machine Learning has become a popular tool to perform wind-speed and power predictions. However, the existing approaches have several limitations. These include (i) insufficient consideration of spatio-temporal correlations in wind-speed data, (ii) a lack of existing methodologies to quantify the uncertainty of wind speed prediction and its propagation to the wind-power estimation, and (iii) a focus on less than hourly frequencies. To overcome these limitations, we introduce a framework to reconstruct a spatio-temporal field on a regular grid from irregularly distributed wind-speed measurements. After decomposing data into temporally referenced basis functions and their corresponding spatially distributed coefficients, the latter are spatially modelled using Extreme Learning Machines. Estimates of both model and prediction uncertainties, and of their propagation after the transformation of wind speed into wind power, are then provided without any assumptions on distribution patterns of the data. The methodology is applied to the study of hourly wind power potential on a grid of 250 by 250 squared meters for turbines of 100 meters hub height in Switzerland, generating the first dataset of its type for the country. The potential wind power generation is combined with the available area for wind turbine installations to yield an estimate of the technical potential for wind power in Switzerland. The wind power estimate presented here represents an important input for planners to support the design of future energy systems with increased wind power generation.

preprint2021arXiv

Advanced analysis of temporal data using Fisher-Shannon information: theoretical development and application in geosciences

Complex non-linear time series are ubiquitous in geosciences. Quantifying complexity and non-stationarity of these data is a challenging task, and advanced complexity-based exploratory tool are required for understanding and visualizing such data. This paper discusses the Fisher-Shannon method, from which one can obtain a complexity measure and detect non-stationarity, as an efficient data exploration tool. The state-of-the-art studies related to the Fisher-Shannon measures are collected, and new analytical formulas for positive unimodal skewed distributions are proposed. Case studies on both synthetic and real data illustrate the usefulness of the Fisher-Shannon method, which can find application in different domains including time series discrimination and generation of times series features for clustering, modeling and forecasting. The paper is accompanied with Python and R libraries for the non-parametric estimation of the proposed measures.

preprint2021arXiv

Spatio-temporal evolution of global surface temperature distributions

Climate is known for being characterised by strong non-linearity and chaotic behaviour. Nevertheless, few studies in climate science adopt statistical methods specifically designed for non-stationary or non-linear systems. Here we show how the use of statistical methods from Information Theory can describe the non-stationary behaviour of climate fields, unveiling spatial and temporal patterns that may otherwise be difficult to recognize. We study the maximum temperature at two meters above ground using the NCEP CDAS1 daily reanalysis data, with a spatial resolution of 2.5 by 2.5 degree and covering the time period from 1 January 1948 to 30 November 2018. The spatial and temporal evolution of the temperature time series are retrieved using the Fisher Information Measure, which quantifies the information in a signal, and the Shannon Entropy Power, which is a measure of its uncertainty -- or unpredictability. The results describe the temporal behaviour of the analysed variable. Our findings suggest that tropical and temperate zones are now characterized by higher levels of entropy. Finally, Fisher-Shannon Complexity is introduced and applied to study the evolution of the daily maximum surface temperature distributions.

preprint2019arXiv

Analysis of air pollution time series using complexity-invariant distance and information measures

Air pollution is known to be a major threat for human and ecosystem health. A proper understanding of the factors generating pollution and of the behavior of air pollution in time is crucial to support the development of effective policies aiming at the reduction of pollutant concentration. This paper considers the hourly time series of three pollutants, namely NO$_2$, O$_3$ and PM$_{2.5}$, collected on sixteen measurement stations in Switzerland. The air pollution patterns due to the location of measurement stations and their relationship with anthropogenic activities, and specifically land use, are studied using two approaches: Fisher-Shannon information plane and complexity-invariant distance between time series. A clustering analysis is used to recognize within the measurements of a same pollutant group of stations behaving in a similar way. The results clearly demonstrate the relationship between the air pollution probability densities and land use activities.

preprint2018arXiv

Fisher-Shannon complexity analysis of high-frequency urban wind speed time series

1Hz wind time series recorded at different levels (from 1.5 to 25.5 meters) in an urban area are investigated by using the Fisher-Shannon (FS) analysis. FS analysis is a well known method to get insight of the complex behavior of nonlinear systems, by quantifying the order/disorder properties of time series. Our findings reveal that the FS complexity, defined as the product between the Fisher Information Measure and the Shannon entropy power, decreases with the height of the anemometer from the ground, suggesting a height-dependent variability in the order/disorder features of the high frequency wind speed measured in urban layouts. Furthermore, the correlation between the FS complexity of wind speed and the daily variance of the ambient temperature shows similar decrease with the height of the wind sensor. Such correlation is larger for the lower anemometers, indicating that ambient temperature is an important forcing of the wind speed variability in the vicinity of the ground.

preprint2016arXiv

Spatial Patterns of Wind Speed Distributions in Switzerland

This paper presents an initial exploration of high frequency records of extreme wind speed in two steps. The first consists in finding the suitable extreme distribution for $120$ measuring stations in Switzerland, by comparing three known distributions: Weibull, Gamma, and Generalized extreme value. This comparison serves as a basis for the second step which applies a spatial modelling by using Extreme Learning Machine. The aim is to model distribution parameters by employing a high dimensional input space of topographical information. The knowledge of probability distribution gives a comprehensive information and a global overview of wind phenomena. Through this study, a flexible and a simple modelling approach is presented, which can be generalized to almost extreme environmental data for risk assessment and to model renewable energy.

preprint2015arXiv

A New Estimator of Intrinsic Dimension Based on the Multipoint Morisita Index

The size of datasets has been increasing rapidly both in terms of number of variables and number of events. As a result, the empty space phenomenon and the curse of dimensionality complicate the extraction of useful information. But, in general, data lie on non-linear manifolds of much lower dimension than that of the spaces in which they are embedded. In many pattern recognition tasks, learning these manifolds is a key issue and it requires the knowledge of their true intrinsic dimension. This paper introduces a new estimator of intrinsic dimension based on the multipoint Morisita index. It is applied to both synthetic and real datasets of varying complexities and comparisons with other existing estimators are carried out. The proposed estimator turns out to be fairly robust to sample size and noise, unaffected by edge effects, able to handle large datasets and computationally efficient.

preprint2013arXiv

Multifractal portrayal of the Swiss population

Fractal geometry is a fundamental approach for describing the complex irregularities of the spatial structure of point patterns. The present research characterizes the spatial structure of the Swiss population distribution in the three Swiss geographical regions (Alps, Plateau and Jura) and at the entire country level. These analyses were carried out using fractal and multifractal measures for point patterns, which enabled the estimation of the spatial degree of clustering of a distribution at different scales. The Swiss population dataset is presented on a grid of points and thus it can be modelled as a "point process" where each point is characterized by its spatial location (geometrical support) and a number of inhabitants (measured variable). The fractal characterization was performed by means of the box-counting dimension and the multifractal analysis was conducted through the Renyi's generalized dimensions and the multifractal spectrum. Results showed that the four population patterns are all multifractals and present different clustering behaviours. Applying multifractal and fractal methods at different geographical regions and at different scales allowed us to quantify and describe the dissimilarities between the four structures and their underlying processes. This paper is the first Swiss geodemographic study applying multifractal methods using high resolution data.