Source author record

David Bolin

David Bolin appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

18works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

18 published item(s)

preprint2025arXiv

Forecasting the Term Structure of Interest Rates with SPDE-Based Models

The Dynamic Nelson--Siegel (DNS) model is a widely used framework for term structure forecasting. We propose a novel extension that models DNS residuals as a Gaussian random field, capturing dependence across both time and maturity. The residual field is represented via a stochastic partial differential equation (SPDE), enabling flexible covariance structures and scalable Bayesian inference through sparse precision matrices. We consider a range of SPDE specifications, including stationary, non-stationary, anisotropic, and nonseparable models. The SPDE--DNS model is estimated in a Bayesian framework using the integrated nested Laplace approximation (INLA), jointly inferring latent DNS factors and the residual field. Empirical results show that the SPDE-based extensions improve both point and probabilistic forecasts relative to standard benchmarks. When applied in a mean--variance bond portfolio framework, the forecasts generate economically meaningful utility gains, measured as performance fees relative to a Bayesian DNS benchmark under monthly rebalancing. Importantly, incorporating the structured SPDE residual substantially reduces cross-maturity and intertemporal dependence in the remaining measurement error, bringing it closer to white noise. These findings highlight the advantages of combining DNS with SPDE-driven residual modeling for flexible, interpretable, and computationally efficient yield curve forecasting.

preprint2022arXiv

Fast Bayesian estimation of brain activation with cortical surface and subcortical fMRI data using EM

Analysis of brain imaging scans is critical to understanding the way the human brain functions, which can be leveraged to treat injuries and conditions that affect the quality of life for a significant portion of the human population. In particular, functional magnetic resonance imaging (fMRI) scans give detailed data on a living subject at high spatial and temporal resolutions. Due to the high cost involved in the collection of these scans, robust methods of analysis are of critical importance in order to produce meaningful inference. Bayesian methods in particular allow for the inclusion of expected behavior from prior study into an analysis, increasing the power of the results while circumventing problems that arise in classical analyses, including the effects of smoothing results and sensitivity to multiple comparison testing corrections. Recent development of a surface-based spatial Bayesian general linear model for cortical surface fMRI (cs-fMRI) data provides the desired power increase in task fMRI data using stochastic partial differential equation (SPDE) priors. This model relies on the computational efficiencies of the integrated nested Laplace approximation (INLA) to perform powerful analyses that have been validated to outperform classical analyses. In this article, we develop an exact Bayesian analysis method for the GLM, employing an expectation-maximization (EM) algorithm to find maximum a posteriori (MAP) estimates of task-based regressors on cs-fMRI and subcortical fMRI data while using minimal computational resources. Our proposed method is compared to the INLA implementation of the Bayesian GLM, as well as a classical GLM on simulated data. A validation of the method on data from the Human Connectome Project is also provided.

preprint2022arXiv

Local scale invariance and robustness of proper scoring rules

Averages of proper scoring rules are often used to rank probabilistic forecasts. In many cases, the individual terms in these averages are based on observations and forecasts from different distributions. We show that some of the most popular proper scoring rules, such as the continuous ranked probability score (CRPS), give more importance to observations with large uncertainty which can lead to unintuitive rankings. To describe this issue, we define the concept of local scale invariance for scoring rules. A new class of generalized proper kernel scoring rules is derived and as a member of this class we propose the scaled CRPS (SCRPS). This new proper scoring rule is locally scale invariant and therefore works in the case of varying uncertainty. Like CRPS it is computationally available for output from ensemble forecasts, and does not require the ability to evaluate densities of forecasts. We further define robustness of scoring rules, show why this also is an important concept for average scores, and derive new proper scoring rules that are robust against outliers. The theoretical findings are illustrated in three different applications from spatial statistics, stochastic volatility models, and regression for count data.

preprint2020arXiv

A spatial template independent component analysis model for subject-level brain network estimation and inference

Independent component analysis is commonly applied to functional magnetic resonance imaging (fMRI) data to extract independent components (ICs) representing functional brain networks. While ICA produces reliable group-level estimates, single-subject ICA often produces noisy results. Template ICA (tICA) is a hierarchical ICA model using empirical population priors to produce reliable subject-level IC estimates. However, this and other hierarchical ICA models assume unrealistically that subject effects are spatially independent. Here, we propose spatial template ICA (stICA), which incorporates spatial process priors into tICA. This results in greater estimation efficiency of ICs and subject effects. Additionally, the joint posterior distribution can be used to identify engaged areas using an excursions set approach. By leveraging spatial dependencies and avoiding massive multiple comparisons, stICA has high power to detect true effects. We derive an efficient expectation-maximization algorithm to obtain maximum likelihood estimates of the model parameters and posterior moments of the latent fields. Based on analysis of simulated data and fMRI data from the Human Connectome Project, we find that stICA produces estimates that are more accurate and reliable than benchmark approaches, and identifies larger and more reliable areas of engagement. The algorithm is quite tractable, achieving convergence within 7 hours in our fMRI analysis.

preprint2020arXiv

Deformed SPDE models with an application to spatial modeling of significant wave height

A non-stationary Gaussian random field model is developed based on a combination of the stochastic partial differential equation (SPDE) approach and the classical deformation method. With the deformation method, a stationary field is defined on a domain which is deformed so that the field becomes non-stationary. We show that if the stationary field is a Mat'ern field defined as a solution to a fractional SPDE, the resulting non-stationary model can be represented as the solution to another fractional SPDE on the deformed domain. By defining the model in this way, the computational advantages of the SPDE approach can be combined with the deformation method's more intuitive parameterisation of non-stationarity. In particular it allows for independent control over the non-stationary practical correlation range and the variance, which has not been possible with previously proposed non-stationary SPDE models. The model is tested on spatial data of significant wave height, a characteristic of ocean surface conditions which is important when estimating the wear and risks associated with a planned journey of a ship. The model parameters are estimated to data from the north Atlantic using a maximum likelihood approach. The fitted model is used to compute wave height exceedance probabilities and the distribution of accumulated fatigue damage for ships traveling a popular shipping route. The model results agree well with the data, indicating that the model could be used for route optimization in naval logistics.

preprint2020arXiv

The COST IRACON Geometry-based Stochastic Channel Model for Vehicle-to-Vehicle Communication in Intersections

Vehicle-to-vehicle (V2V) wireless communications can improve traffic safety at road intersections and enable congestion avoidance. However, detailed knowledge about the wireless propagation channel is needed for the development and realistic assessment of V2V communication systems. We present a novel geometry-based stochastic MIMO channel model with support for frequencies in the band of 5.2-6.2 GHz. The model is based on extensive high-resolution measurements at different road intersections in the city of Berlin, Germany. We extend existing models, by including the effects of various obstructions, higher order interactions, and by introducing an angular gain function for the scatterers. Scatterer locations have been identified and mapped to measured multi-path trajectories using a measurement-based ray tracing method and a subsequent RANSAC algorithm. The developed model is parameterized, and using the measured propagation paths that have been mapped to scatterer locations, model parameters are estimated. The time variant power fading of individual multi-path components is found to be best modeled by a Gamma process with an exponential autocorrelation. The path coherence distance is estimated to be in the range of 0-2 m. The model is also validated against measurement data, showing that the developed model accurately captures the behavior of the measured channel gain, Doppler spread, and delay spread. This is also the case for intersections that have not been used when estimating model parameters.

preprint2019arXiv

Multivariate type G Matérn stochastic partial differential equation random fields

For many applications with multivariate data, random field models capturing departures from Gaussianity within realisations are appropriate. For this reason, we formulate a new class of multivariate non-Gaussian models based on systems of stochastic partial differential equations with additive type G noise whose marginal covariance functions are of Matérn type. We consider four increasingly flexible constructions of the noise, where the first two are similar to existing copula-based models. In contrast to these, the latter two constructions can model non-Gaussian spatial data without replicates. Computationally efficient methods for likelihood-based parameter estimation and probabilistic prediction are proposed, and the flexibility of the suggested models is illustrated by numerical examples and two statistical applications.

preprint2016arXiv

Efficient adaptive MCMC through precision estimation

A novel adaptive Markov chain Monte Carlo algorithm is presented. The algorithm utilizes sparsity in the partial correlation structure of a density to efficiently estimate the covariance matrix through the Cholesky factor of the precision matrix. The algorithm also utilizes the sparsity to sample efficiently from both MALA and Metropolis Hasting random walk proposals. Further, an algorithm that estimates the partial correlation structure of a density is proposed. Combining this with the Cholesky factor estimation algorithm results in an efficient black-box AMCMC method that can be used for general densities with unknown dependency structure. The method is compared with regular empirical covariance adaption for two examples. In both examples, the proposed method's covariance estimates converge faster to the true covariance matrix and the computational cost for each iteration is lower.

preprint2016arXiv

Fast Bayesian whole-brain fMRI analysis with spatial 3D priors

Spatial whole-brain Bayesian modeling of task-related functional magnetic resonance imaging (fMRI) is a great computational challenge. Most of the currently proposed methods therefore do inference in subregions of the brain separately or do approximate inference without comparison to the true posterior distribution. A popular such method, which is now the standard method for Bayesian single subject analysis in the SPM software, is introduced in Penny et al. (2005b). The method processes the data slice-by-slice and uses an approximate variational Bayes (VB) estimation algorithm that enforces posterior independence between activity coefficients in different voxels. We introduce a fast and practical Markov chain Monte Carlo (MCMC) scheme for exact inference in the same model, both slice-wise and for the whole brain using a 3D prior on activity coefficients. The algorithm exploits sparsity and uses modern techniques for efficient sampling from high-dimensional Gaussian distributions, leading to speed-ups without which MCMC would not be a practical option. Using MCMC, we are for the first time able to evaluate the approximate VB posterior against the exact MCMC posterior, and show that VB can lead to spurious activation. In addition, we develop an improved VB method that drops the assumption of independent voxels a posteriori. This algorithm is shown to be much faster than both MCMC and the original VB for large datasets, with negligible error compared to the MCMC posterior.

preprint2016arXiv

Quantifying the uncertainty of contour maps

Contour maps are widely used to display estimates of spatial fields. Instead of showing the estimated field, a contour map only shows a fixed number of contour lines for different levels. However, despite the ubiquitous use of these maps, the uncertainty associated with them has been given a surprisingly small amount of attention. We derive measures of the statistical uncertainty, or quality, of contour maps, and use these to decide an appropriate number of contour lines, that relates to the uncertainty in the estimated spatial field. For practical use in geostatistics and medical imaging, computational methods are constructed, that can be applied to Gaussian Markov random fields, and in particular be used in combination with integrated nested Laplace approximations for latent Gaussian models. The methods are demonstrated on simulated data and an application to temperature estimation is presented.

preprint2016arXiv

Spatially adaptive covariance tapering

Covariance tapering is a popular approach for reducing the computational cost of spatial prediction and parameter estimation for Gaussian process models. However, tapering can have poor performance when the process is sampled at spatially irregular locations or when non-stationary covariance models are used. This work introduces an adaptive tapering method in order to improve the performance of tapering in these problematic cases. This is achieved by introducing a computationally convenient class of compactly supported non-stationary covariance functions, combined with a new method for choosing spatially varying taper ranges. Numerical experiments are used to show that the performance of both kriging prediction and parameter estimation can be improved by allowing for spatially varying taper ranges. However, although adaptive tapering outperforms regular tapering, simply dividing the data into blocks and ignoring the dependence between the blocks is often a better method for parameter estimation.

preprint2016arXiv

Statistical Modeling and Estimation of Censored Pathloss Data

Pathloss is typically modeled using a log-distance power law with a large-scale fading term that is log-normal. However, the received signal is affected by the dynamic range and noise floor of the measurement system used to sound the channel, which can cause measurement samples to be truncated or censored. If the information about the censored samples are not included in the estimation method, as in ordinary least squares estimation, it can result in biased estimation of both the pathloss exponent and the large scale fading. This can be solved by applying a Tobit maximum-likelihood estimator, which provides consistent estimates for the pathloss parameters. This letter provides information about the Tobit maximum-likelihood estimator and its asymptotic variance under certain conditions.

preprint2016arXiv

Whole-brain substitute CT generation using Markov random field mixture models

Computed tomography (CT) equivalent information is needed for attenuation correction in PET imaging and for dose planning in radiotherapy. Prior work has shown that Gaussian mixture models can be used to generate a substitute CT (s-CT) image from a specific set of MRI modalities. This work introduces a more flexible class of mixture models for s-CT generation, that incorporates spatial dependency in the data through a Markov random field prior on the latent field of class memberships associated with a mixture model. Furthermore, the mixture distributions are extended from Gaussian to normal inverse Gaussian (NIG), allowing heavier tails and skewness. The amount of data needed to train a model for s-CT generation is of the order of 100 million voxels. The computational efficiency of the parameter estimation and prediction methods are hence paramount, especially when spatial dependency is included in the models. A stochastic Expectation Maximization (EM) gradient algorithm is proposed in order to tackle this challenge. The advantages of the spatial model and NIG distributions are evaluated with a cross-validation study based on data from 14 patients. The study show that the proposed model enhances the predictive quality of the s-CT images by reducing the mean absolute error with 17.9%. Also, the distribution of CT values conditioned on the MR images are better explained by the proposed model as evaluated using continuous ranked probability scores.

preprint2013arXiv

Non-Gaussian Matérn fields with an application to precipitation modeling

The recently proposed non-Gaussian Matérn random field models, generated through Stochastic Partial differential equations (SPDEs), are extended by considering the class of Generalized Hyperbolic processes as noise forcings. The models are also extended to the standard geostatistical setting where irregularly spaced observations are modeled using measurement errors and covariates. A maximum likelihood estimation technique based on the Monte Carlo Expectation Maximization (MCEM) algorithm is presented, and it is shown how the model can be used to do predictions at unobserved locations. Finally, an application to precipitation data over the United States for two month in 1997 is presented, and the performance of the non-Gaussian models is compared with standard Gaussian and transformed Gaussian models through cross-validation.

preprint2012arXiv

Excursion and contour uncertainty regions for latent Gaussian models

An interesting statistical problem is to find regions where some studied process exceeds a certain level. Estimating such regions so that the probability for exceeding the level in the entire set is equal to some predefined value is a difficult problem that occurs in several areas of applications ranging from brain imaging to astrophysics. In this work, a method for solving this problem, as well as the related problem of finding uncertainty regions for contour curves, for latent Gaussian models is proposed. The method is based on using a parametric family for the excursion sets in combination with a sequential importance sampling method for estimating joint probabilities. The accuracy of the method is investigated using simulated data and two environmental applications are presented. In the first application, areas where the air pollution in the Piemonte region in northern Italy exceeds the daily limit value, set by the European Union for human health protection, are estimated. In the second application, regions in the African Sahel that experienced an increase in vegetation after the drought period in the early 1980s are estimated.

preprint2012arXiv

Spatial Matérn fields driven by non-Gaussian noise

The article studies non-Gaussian extensions of a recently discovered link between certain Gaussian random fields, expressed as solutions to stochastic partial differential equations (SPDEs), and Gaussian Markov random fields. The focus is on non-Gaussian random fields with Matérn covariance functions, and in particular we show how the SPDE formulation of a Laplace moving average model can be used to obtain an efficient simulation method as well as an accurate parameter estimation technique for the model. This should be seen as a demonstration of how these techniques can be used, and generalizations to more general SPDEs are readily available.

preprint2011arXiv

How do Markov approximations compare with other methods for large spatial data sets?

The Matérn covariance function is a popular choice for modeling dependence in spatial environmental data. Standard Matérn covariance models are, however, often computationally infeasible for large data sets. In this work, recent results for Markov approximations of Gaussian Matérn fields based on Hilbert space approximations are extended using wavelet basis functions. These Markov approximations are compared with two of the most popular methods for efficient covariance approximations; covariance tapering and the process convolution method. The results show that, for a given computational cost, the Markov methods have a substantial gain in accuracy compared with the other methods.

preprint2011arXiv

Spatial models generated by nested stochastic partial differential equations, with an application to global ozone mapping

A new class of stochastic field models is constructed using nested stochastic partial differential equations (SPDEs). The model class is computationally efficient, applicable to data on general smooth manifolds, and includes both the Gaussian Matérn fields and a wide family of fields with oscillating covariance functions. Nonstationary covariance models are obtained by spatially varying the parameters in the SPDEs, and the model parameters are estimated using direct numerical optimization, which is more efficient than standard Markov Chain Monte Carlo procedures. The model class is used to estimate daily ozone maps using a large data set of spatially irregular global total column ozone data.