Source author record

Istvan Szunyogh

Istvan Szunyogh appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2020arXiv

Combining Machine Learning with Knowledge-Based Modeling for Scalable Forecasting and Subgrid-Scale Closure of Large, Complex, Spatiotemporal Systems

We consider the commonly encountered situation (e.g., in weather forecasting) where the goal is to predict the time evolution of a large, spatiotemporally chaotic dynamical system when we have access to both time series data of previous system states and an imperfect model of the full system dynamics. Specifically, we attempt to utilize machine learning as the essential tool for integrating the use of past data into predictions. In order to facilitate scalability to the common scenario of interest where the spatiotemporally chaotic system is very large and complex, we propose combining two approaches:(i) a parallel machine learning prediction scheme; and (ii) a hybrid technique, for a composite prediction system composed of a knowledge-based component and a machine-learning-based component. We demonstrate that not only can this method combining (i) and (ii) be scaled to give excellent performance for very large systems, but also that the length of time series data needed to train our multiple, parallel machine learning components is dramatically less than that necessary without parallelization. Furthermore, considering cases where computational realization of the knowledge-based component does not resolve subgrid-scale processes, our scheme is able to use training data to incorporate the effect of the unresolved short-scale dynamics upon the resolved longer-scale dynamics ("subgrid-scale closure").

preprint2011arXiv

Ensemble regional data assimilation using joint states

We propose a data assimilation scheme that produces the analyses for a global and an embedded limited area model simultaneously, considering forecast information from both models. The purpose of the proposed approach is twofold. First, we expect that the global analysis will benefit from incorporation of information from the higher resolution limited area model. Second, our method is expected to produce a limited area analysis that is more strongly constrained by the large scale flow than a conventional limited area analysis. The proposed scheme minimizes a cost function in which the control variable is the joint state of the global and the limited area models. In addition, the cost function includes a constraint term that penalizes large differences between the global and the limited area state estimates. The proposed approach is tested by idealized experiments, using `toy' models introduced by Lorenz in 2005. The results of these experiments suggest that the proposed approach improves the global analysis within and near the limited area domain and the regional analysis near the lateral boundaries. These analysis improvements lead to forecast improvements in both the global and the limited area models.

preprint2011arXiv

On the propagation of information and the use of localization in ensemble Kalman filtering

Several localized versions of the ensemble Kalman filter have been proposed. Although tests applying such schemes have proven them to be extremely promising, a full basic understanding of the rationale and limitations of localization is currently lacking. It is one of the goals of this paper to contribute toward addressing this issue. The second goal is to elucidate the role played by chaotic wave dynamics in the propagation of information and the resulting impact on forecasts. To accomplish these goals, the principal tool used here will be analysis and interpretation of numerical experiments on a toy atmospheric model introduced by Lorenz in 2005. Propagation of the wave packets of this model is shown. It is found that, when an ensemble Kalman filter scheme is employed, the spatial correlation function obtained at each forecast cycle by averaging over the background ensemble members is short ranged, and this is in strong contrast to the much longer range correlation function obtained by averaging over states from free evolution of the model. Propagation of the effects of observations made in one region on forecasts in other regions is studied. The error covariance matrices from the analyses with localization and without localization are compared. From this study, major characteristics of the localization process and information propagation are extracted and summarized.

preprint2007arXiv

Comparison between Local Ensemble Transform Kalman Filter and PSAS in the NASA finite volume GCM: perfect model experiments

This paper explores the potential of Local Ensemble Transform Kalman Filter (LETKF) by comparing the performance of LETKF with an operational 3D-Var assimilation system, Physical-Space Statistical Analysis System (PSAS), under a perfect model scenario. The comparison is carried out on the finite volume Global Circulation Model (fvGCM) with 72 grid points zonally, 46 grid points meridionally and 55 vertical levels. With only forty ensemble members, LETKF obtains an analysis and forecasts with lower RMS errors than those from PSAS. The performance of LETKF is further improved, especially over the oceans, by assimilating simulated temperature observations from rawinsondes and conventional surface pressure observations instead of geopotential heights. An initial decrease of the forecast errors in the NH observed in PSAS but not in LETKF suggests that the PSAS analysis is less balanced. The observed advantage of LETKF over PSAS is due to the ability of the forty-member ensemble from LETKF to capture flow-dependent errors and thus create a good estimate of the true background uncertainty. Furthermore, localization makes LETKF highly parallel and efficient, requiring only 5 minutes per analysis in a cluster of 20 PCs with forty ensemble members.