Source author record

Sándor Baran

Sándor Baran appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

18works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

18 published item(s)

preprint2021arXiv

Calibration of wind speed ensemble forecasts for power generation

In the last decades wind power became the second largest energy source in the EU covering 16% of its electricity demand. However, due to its volatility, accurate short range wind power predictions are required for successful integration of wind energy into the electrical grid. Accurate predictions of wind power require accurate hub height wind speed forecasts, where the state of the art method is the probabilistic approach based on ensemble forecasts obtained from multiple runs of numerical weather prediction models. Nonetheless, ensemble forecasts are often uncalibrated and might also be biased, thus require some form of post-processing to improve their predictive performance. We propose a novel flexible machine learning approach for calibrating wind speed ensemble forecasts, which results in a truncated normal predictive distribution. In a case study based on 100m wind speed forecasts produced by the operational ensemble prediction system of the Hungarian Meteorological Service, the forecast skill of this method is compared with the predictive performance of three different ensemble model output statistics approaches and the raw ensemble forecasts. We show that compared with the raw ensemble, post-processing always improves the calibration of probabilistic and accuracy of point forecasts and from the four competing methods the novel machine learning based approach results in the best overall performance.

preprint2021arXiv

Truncated generalized extreme value distribution based EMOS model for calibration of wind speed ensemble forecasts

In recent years, ensemble weather forecasting have become a routine at all major weather prediction centres. These forecasts are obtained from multiple runs of numerical weather prediction models with different initial conditions or model parametrizations. However, ensemble forecasts can often be underdispersive and also biased, so some kind of post-processing is needed to account for these deficiencies. One of the most popular state of the art statistical post-processing techniques is the ensemble model output statistics (EMOS), which provides a full predictive distribution of the studied weather quantity. We propose a novel EMOS model for calibrating wind speed ensemble forecasts, where the predictive distribution is a generalized extreme value (GEV) distribution left truncated at zero (TGEV). The truncation corrects the disadvantage of the GEV distribution based EMOS models of occasionally predicting negative wind speed values, without affecting its favorable properties. The new model is tested on four data sets of wind speed ensemble forecasts provided by three different ensemble prediction systems, covering various geographical domains and time periods. The forecast skill of the TGEV EMOS model is compared with the predictive performance of the truncated normal, log-normal and GEV methods and the raw and climatological forecasts as well. The results verify the advantageous properties of the novel TGEV EMOS approach.

preprint2018arXiv

Statistical post-processing of dual-resolution ensemble forecasts

The computational cost as well as the probabilistic skill of ensemble forecasts depends on the spatial resolution of the numerical weather prediction model and the ensemble size. Periodically, e.g. when more computational resources become available, it is appropriate to reassess the balance between resolution and ensemble size. Recently, it has been proposed to investigate this balance in the context of dual-resolution ensembles, which use members with two different resolutions to make probabilistic forecasts. This study investigates whether statistical post-processing of such dual-resolution ensemble forecasts changes the conclusions regarding the optimal dual-resolution configuration. Medium-range dual-resolution ensemble forecasts of 2-metre temperature have been calibrated using ensemble model output statistics. The forecasts are produced with ECMWF's Integrated Forecast System and have horizontal resolutions between 18 km and 45 km. The ensemble sizes range from 8 to 254 members. The forecasts are verified with SYNOP station data. Results show that score differences between various single and dual-resolution configurations are strongly reduced by statistical post-processing. Therefore, the benefit of some dual-resolution configurations over single resolution configurations appears to be less pronounced than for raw forecasts. Moreover, the ranking of the ensemble configurations can be affected by the statistical post-processing.

preprint2018arXiv

Statistical post-processing of hydrological forecasts using Bayesian model averaging

Accurate and reliable probabilistic forecasts of hydrological quantities like runoff or water level are beneficial to various areas of society. Probabilistic state-of-the-art hydrological ensemble prediction models are usually driven with meteorological ensemble forecasts. Hence, biases and dispersion errors of the meteorological forecasts cascade down to the hydrological predictions and add to the errors of the hydrological models. The systematic parts of these errors can be reduced by applying statistical post-processing. For a sound estimation of predictive uncertainty and an optimal correction of systematic errors, statistical post-processing methods should be tailored to the particular forecast variable at hand. Former studies have shown that it can make sense to treat hydrological quantities as bounded variables. In this paper, a doubly truncated Bayesian model averaging (BMA) method, which allows for flexible post-processing of (multi-model) ensemble forecasts of water level, is introduced. A case study based on water level for a gauge of river Rhine, reveals a good predictive skill of doubly truncated BMA compared both to the raw ensemble and the reference ensemble model output statistics approach.

preprint2016arXiv

Censored and shifted gamma distribution based EMOS model for probabilistic quantitative precipitation forecasting

Recently all major weather prediction centres provide forecast ensembles of different weather quantities which are obtained from multiple runs of numerical weather prediction models with various initial conditions and model parametrizations. However, ensemble forecasts often show an underdispersive character and may also be biased, so that some post-processing is needed to account for these deficiencies. Probably the most popular modern post-processing techniques are the ensemble model output statistics (EMOS) and the Bayesian model averaging (BMA) which provide estimates of the density of the predictable weather quantity. In the present work an EMOS method for calibrating ensemble forecasts of precipitation accumulation is proposed, where the predictive distribution follows a censored and shifted gamma (CSG) law with parameters depending on the ensemble members. The CSG EMOS model is tested on ensemble forecasts of 24 h precipitation accumulation of the eight-member University of Washington mesoscale ensemble and on the 11 member ensemble produced by the operational Limited Area Model Ensemble Prediction System of the Hungarian Meteorological Service. The predictive performance of the new EMOS approach is compared with the fit of the raw ensemble, the generalized extreme value (GEV) distribution based EMOS model and the gamma BMA method. According to the results, the proposed CSG EMOS model slightly outperforms the GEV EMOS approach in terms of calibration of probabilistic and accuracy of point forecasts and shows significantly better predictive skill that the raw ensemble and the BMA model.

preprint2015arXiv

Mixture EMOS model for calibrating ensemble forecasts of wind speed

Ensemble model output statistics (EMOS) is a statistical tool for post-processing forecast ensembles of weather variables obtained from multiple runs of numerical weather prediction models in order to produce calibrated predictive probability density functions (PDFs). The EMOS predictive PDF is given by a parametric distribution with parameters depending on the ensemble forecasts. We propose an EMOS model for calibrating wind speed forecasts based on weighted mixtures of truncated normal (TN) and log-normal (LN) distributions where model parameters and component weights are estimated by optimizing the values of proper scoring rules over a rolling training period. The new model is tested on wind speed forecasts of the 50 member European Centre for Medium-Range Weather Forecasts ensemble, the 11 member Aire Limitée Adaptation dynamique Développement International-Hungary Ensemble Prediction System ensemble of the Hungarian Meteorological Service and the eight-member University of Washington mesoscale ensemble, and its predictive performance is compared to that of various benchmark EMOS models based on single parametric families and combinations thereof. The results indicate improved calibration of probabilistic and accuracy of point forecasts in comparison with the raw ensemble and climatological forecasts. The mixture EMOS model significantly outperforms the TN and LN EMOS methods, moreover, it provides better calibrated forecasts than the TN-LN combination model and offers an increased flexibility while avoiding covariate selection problems.

preprint2015arXiv

Optimal designs for the methane flux in troposphere

The understanding of methane emission and methane absorption plays a central role both in the atmosphere and on the surface of the Earth. Several important ecological processes, e.g., ebullition of methane and its natural microergodicity request better designs for observations in order to decrease variability in parameter estimation. Thus, a crucial fact, before the measurements are taken, is to give an optimal design of the sites where observations should be collected in order to stabilize the variability of estimators. In this paper we introduce a realistic parametric model of covariance and provide theoretical and numerical results on optimal designs. For parameter estimation D-optimality, while for prediction integrated mean square error and entropy criteria are used. We illustrate applicability of obtained benchmark designs for increasing/measuring the efficiency of the engineering designs for estimation of methane rate in various temperature ranges and under different correlation parameters. We show that in most situations these benchmark designs have higher efficiency.

preprint2014arXiv

Joint probabilistic forecasting of wind speed and temperature using Bayesian model averaging

Ensembles of forecasts are typically employed to account for the forecast uncertainties inherent in predictions of future weather states. However, biases and dispersion errors often present in forecast ensembles require statistical post-processing. Univariate post-processing models such as Bayesian model averaging (BMA) have been successfully applied for various weather quantities. Nonetheless, BMA and many other standard post-processing procedures are designed for a single weather variable, thus ignoring possible dependencies among weather quantities. In line with recently upcoming research to develop multivariate post-processing procedures, e.g., BMA for bivariate wind vectors, or flexible procedures applicable for multiple weather quantities of different types, a bivariate BMA model for joint calibration of wind speed and temperature forecasts is proposed based on the bivariate truncated normal distribution. It extends the univariate truncated normal BMA model designed for post-processing ensemble forecast of wind speed by adding a normally distributed temperature component with a covariance structure representing the dependency among the two weather quantities. The method is applied to wind speed and temperature forecasts of the eight-member University of Washington mesoscale ensemble and of the eleven-member ALADIN-HUNEPS ensemble of the Hungarian Meteorological Service and its predictive performance is compared to that of the general Gaussian copula method. The results indicate improved calibration of probabilistic and accuracy of point forecasts in comparison to the raw ensemble and the overall performance of this model is able to keep up with that of the Gaussian copula method.

preprint2014arXiv

Log-normal distribution based EMOS models for probabilistic wind speed forecasting

Ensembles of forecasts are obtained from multiple runs of numerical weather forecasting models with different initial conditions and typically employed to account for forecast uncertainties. However, biases and dispersion errors often occur in forecast ensembles, they are usually under-dispersive and uncalibrated and require statistical post-processing. We present an Ensemble Model Output Statistics (EMOS) method for calibration of wind speed forecasts based on the log-normal (LN) distribution, and we also show a regime-switching extension of the model which combines the previously studied truncated normal (TN) distribution with the LN. Both presented models are applied to wind speed forecasts of the eight-member University of Washington mesoscale ensemble, of the fifty-member ECMWF ensemble and of the eleven-member ALADIN-HUNEPS ensemble of the Hungarian Meteorological Service, and their predictive performances are compared to those of the TN and general extreme value (GEV) distribution based EMOS methods and to the TN-GEV mixture model. The results indicate improved calibration of probabilistic and accuracy of point forecasts in comparison to the raw ensemble and to climatological forecasts. Further, the TN-LN mixture model outperforms the traditional TN method and its predictive performance is able to keep up with the models utilizing the GEV distribution without assigning mass to negative values.

preprint2013arXiv

Comparison of BMA and EMOS statistical calibration methods for temperature and wind speed ensemble weather prediction

The evolution of the weather can be described by deterministic numerical weather forecasting models. Multiple runs of these models with different initial conditions and/or model physics result in forecast ensembles which are used for estimating the distribution of future atmospheric variables. However, these ensembles are usually under-dispersive and uncalibrated, so post-processing is required. In the present work we compare different versions of Bayesian Model Averaging (BMA) and Ensemble Model Output Statistics (EMOS) post-processing methods in order to calibrate 2m temperature and 10m wind speed forecasts of the operational ALADIN Limited Area Model Ensemble Prediction System of the Hungarian Meteorological Service. We show that compared to the raw ensemble both post-processing methods improve the calibration of probabilistic and accuracy of point forecasts and that the best BMA method slightly outperforms the EMOS technique.

preprint2013arXiv

Optimal designs for parameters of shifted Ornstein-Uhlenbeck sheets measured on monotonic sets

Measurement on sets with a specific geometric shape can be of interest for many important applications (e.g. measurement along the isotherms in structural engineering). In the present paper the properties of optimal designs for estimating the parameters of shifted Ornstein-Uhlenbeck sheets, that is Gaussian two-variable random fields with exponential correlation structures, are investigated when the processes are observed on monotonic sets. Substantial differences are demonstrated between the cases when one is interested only in trend parameters and when the whole parameter set is of interest. The theoretical results are illustrated by computer experiments and simulated examples from the field of structure engineering. From the design point of view the most interesting finding of the paper is the loss of efficiency of the regular grid design compared to the optimal monotonic design.

preprint2013arXiv

Probabilistic temperature forecasting with statistical calibration in Hungary

Weather forecasting is mostly based on the outputs of deterministic numerical weather forecasting models. Multiple runs of these models with different initial conditions result in forecast ensembles which is are used for estimating the distribution of future atmospheric variables. However, these ensembles are usually under-dispersive and uncalibrated, so post-processing is required. In the present work Bayesian Model Averaging (BMA) is applied for calibrating ensembles of temperature forecasts produced by the operational Limited Area Model Ensemble Prediction System of the Hungarian Meteorological Service (HMS). We describe two possible BMA models for temperature data of the HMS and show that BMA post-processing significantly improves calibration and probabilistic forecasts although the accuracy of point forecasts is rather unchanged.

preprint2013arXiv

Probabilistic wind speed forecasting in Hungary

Prediction of various weather quantities is mostly based on deterministic numerical weather forecasting models. Multiple runs of these models with different initial conditions result ensembles of forecasts which are applied for estimating the distribution of future weather quantities. However, the ensembles are usually under-dispersive and uncalibrated, so post-processing is required. In the present work Bayesian Model Averaging (BMA) is applied for calibrating ensembles of wind speed forecasts produced by the operational Limited Area Model Ensemble Prediction System of the Hungarian Meteorological Service (HMS). We describe two possible BMA models for wind speed data of the HMS and show that BMA post-processing significantly improves the calibration and precision of forecasts.

preprint2013arXiv

Probabilistic wind speed forecasting using Bayesian model averaging with truncated normal components

Bayesian model averaging (BMA) is a statistical method for post-processing forecast ensembles of atmospheric variables, obtained from multiple runs of numerical weather prediction models, in order to create calibrated predictive probability density functions (PDFs). The BMA predictive PDF of the future weather quantity is the mixture of the individual PDFs corresponding to the ensemble members and the weights and model parameters are estimated using ensemble members and validating observation from a given training period. In the present paper we introduce a BMA model for calibrating wind speed forecasts, where the components PDFs follow truncated normal distribution with cut-off at zero, and apply it to the ALADIN-HUNEPS ensemble of the Hungarian Meteorological Service. Three parameter estimation methods are proposed and each of the corresponding models outperforms the traditional gamma BMA model both in calibration and in accuracy of predictions. Moreover, since here the maximum likelihood estimation of the parameters does not require numerical optimization, modelling can be performed much faster than in case of gamma mixtures.

preprint2012arXiv

Testing stability in a spatial unilateral autoregressive model

Least squares estimator of the stability parameter $\varrho := |α| + |β|$ for a spatial unilateral autoregressive process $X_{k,\ell}=αX_{k-1,\ell}+βX_{k,\ell-1}+\varepsilon_{k,\ell}$ is investigated. Asymptotic normality with a scaling factor $n^{5/4}$ is shown in the unstable case, i.e., when $\varrho = 1$, in contrast to the AR(p) model $X_k=α_1 X_{k-1}+... +α_p X_{k-p}+ \varepsilon_k$, where the least squares estimator of the stability parameter $\varrho :=α_1 + ... + α_p$ is not asymptotically normal in the unstable, i.e., in the unit root case.

preprint2011arXiv

Parameter estimation in a spatial unit root autoregressive model

Spatial unilateral autoregressive model $X_{k,\ell}=αX_{k-1,\ell}+βX_{k,\ell-1}+γX_{k-1,\ell-1}+ε_{k,\ell}$ is investigated in the unit root case, that is when the parameters are on the boundary of the domain of stability that forms a tetrahedron with vertices $(1,1,-1), \ (1,-1,1),\ (-1,1,1)$ and $(-1,-1,-1)$. It is shown that the limiting distribution of the least squares estimator of the parameters is normal and the rate of convergence is $n$ when the parameters are in the faces or on the edges of the tetrahedron, while on the vertices the rate is $n^{3/2}$.

preprint2011arXiv

Parameter estimation in linear regression driven by a Gaussian sheet

The problem of estimating the parameters of a linear regression model $Z(s,t)=m_1g_1(s,t)+ \cdots + m_pg_p(s,t)+U(s,t)$ based on observations of $Z$ on a spatial domain $G$ of special shape is considered, where the driving process $U$ is a Gaussian random field and $g_1, \ldots, g_p$ are known functions. Explicit forms of the maximum likelihood estimators of the parameters are derived in the cases when $U$ is either a Wiener or a stationary or nonstationary Ornstein-Uhlenbeck sheet. Simulation results are also presented, where the driving random sheets are simulated with the help of their Karhunen-Loève expansions.

preprint2010arXiv

On the variances of a spatial unit root model

The asymptotic properties of the variances of the spatial autoregressive model $X_{k,\ell}=αX_{k-1,\ell}+βX_{k,\ell-1}+γX_{k-1,\ell-1}+ε_{k,\ell}$ are investigated in the unit root case, that is when the parameters are on the boundary of domain of stability that forms a tetrahedron in $[-1,1]^3$. The limit of the variance of $n^{-\varrho}X_{[ns],[nt]}$ is determined, where on the interior of the faces of the domain of stability $\varrho=1/4$, on the edges $\varrho =1/2$, while on the vertices $\varrho =1$.