Source author record

Alexey Miroshnikov

Alexey Miroshnikov appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

9works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

9 published item(s)

preprint2022arXiv

Wasserstein-based fairness interpretability framework for machine learning models

The objective of this article is to introduce a fairness interpretability framework for measuring and explaining the bias in classification and regression models at the level of a distribution. In our work, we measure the model bias across sub-population distributions in the model output using the Wasserstein metric. To properly quantify the contributions of predictors, we take into account the favorability of both the model and predictors with respect to the non-protected class. The quantification is accomplished by the use of transport theory, which gives rise to the decomposition of the model bias and bias explanations to positive and negative contributions. To gain more insight into the role of favorability and allow for additivity of bias explanations, we adapt techniques from cooperative game theory.

preprint2016arXiv

Stability of fully discrete variational schemes for elastodynamics with a polyconvex stored energy

In this article we develop a fully discrete variational scheme that approximates the equations of three dimensional elastodynamics with polyconvex stored energy. The fully discrete scheme is based on a time-discrete variational scheme developed by S.~Demoulini, D.~M.~A.~Stuart and A.~E.~Tzavaras (2001). We show that the fully discrete scheme is unconditionally stable. The proof of stability is based on a relative entropy estimation for the fully discrete approximates.

preprint2015arXiv

BayesSummaryStatLM: An R package for Bayesian Linear Models for Big Data and Data Science

Recent developments in data science and big data research have produced an abundance of large data sets that are too big to be analyzed in their entirety, due to limits on either computer memory or storage capacity. Here, we introduce our R package 'BayesSummaryStatLM' for Bayesian linear regression models with Markov chain Monte Carlo implementation that overcomes these limitations. Our Bayesian models use only summary statistics of data as input; these summary statistics can be calculated from subsets of big data and combined over subsets. Thus, complete data sets do not need to be read into memory in full, which removes any physical memory limitations of a user. Our package incorporates the R package 'ff' and its functions for reading in big data sets in chunks while simultaneously calculating summary statistics. We describe our Bayesian linear regression models, including several choices of prior distributions for unknown model parameters, and illustrate capabilities and features of our R package using both simulated and real data sets.

preprint2015arXiv

Parallel Markov Chain Monte Carlo for Non-Gaussian Posterior Distributions

Recent developments in big data and analytics research have produced an abundance of large data sets that are too big to be analyzed in their entirety, due to limits on computer memory or storage capacity. To address these issues, communication-free parallel Markov chain Monte Carlo (MCMC) methods have been developed for Bayesian analysis of big data. These methods partition data into manageable subsets, perform independent Bayesian MCMC analysis on each subset, and combine the subset posterior samples to estimate the full data posterior. Current approaches to combining subset posterior samples include sample averaging, weighted averaging, and kernel smoothing techniques. Although these methods work well for Gaussian posteriors, they are not well-suited to non-Gaussian posterior distributions. Here, we develop a new direct density product method for combining subset marginal posterior samples to estimate full data marginal posterior densities. Using a commonly-implemented distance metric, we show in simulation studies of Bayesian models with non-Gaussian posteriors that our method outperforms the existing methods in approximating the full data marginal posteriors. Since our method estimates only marginal densities, there is no limitation on the number of model parameters analyzed. Our procedure is suitable for Bayesian models with unknown parameters with fixed dimension in continuous parameter spaces.

preprint2014arXiv

A variational approximation scheme for radial polyconvex elasticity that preserves the positivity of Jacobians

We consider the equations describing the dynamics of radial motions for isotropic elastic materials; these form a system of non-homogeneous conservation laws. We construct a variational approximation scheme that decreases the total mechanical energy and at the same time leads to physically realizable motions that avoid interpenetration of matter.

preprint2014arXiv

Convergence of Variational Approximation Schemes for Elastodynamics with Polyconvex Energy

We consider a variational scheme developed by S. Demoulini, D. M. A. Stuart and A. E. Tzavaras [Arch. Ration. Mech. Anal. 157 (2001)] that approximates the equations of three dimensional elastodynamics with polyconvex stored energy. We establish the convergence of the time-continuous interpolates constructed in the scheme to a solution of polyconvex elastodynamics before shock formation. The proof is based on a relative entropy estimation for the time-discrete approximants in an environment of Lp-theory bounds, and provides an error estimate for the approximation before the formation of shocks.

preprint2014arXiv

Motile Geobacter Dechlorinators Migrate into a Model Source Zone of Trichloroethene Dense Non-aqueous Phase Liquid: Experimental Evaluation and Modeling

Microbial migration towards a trichloroethene (TCE) dense non-aqueous phase liquid (DNAPL) could facilitate the bioaugmentation of TCE DNAPL source zones. This study characterized the motility of the Geobacter dechlorinators in a TCE to cisdichloroethene dechlorinating KB-1TM subculture. No chemotaxis towards or away from TCE was found using an agarose in-plug bridge method. A second experiment placed an inoculated aqueous layer on top of a sterile sand layer and showed that Geobacter migrated several centimeters in the sand layer in just $7$ days. A random motility coefficient for Geobacter in water of $0.24 \pm 0.02$ cm$^2$ day$^{-1}$ was fitted. A third experiment used a diffusion-cell setup with a $5.5$ cm central sand layer separating a DNAPL from an aqueous top layer as a model source zone to examine the effect of random motility on TCE DNAPL dissolution. With top layer inoculation, Geobacter quickly colonized the sand layer, thereby enhancing the initial TCE DNAPL dissolution flux. After $19$ days, the DNAPL dissolution enhancement was only $24\%$ lower than with an homogenous inoculation of the sand layer. A diffusion-motility model was developed to describe dechlorination and migration in the diffusion-cells. This model suggested that the fast colonization of the sand layer by Geobacter was due to the combination of random motility and growth on TCE.

preprint2014arXiv

parallelMCMCcombine: An R Package for Bayesian Methods for Big Data and Analytics

Recent advances in big data and analytics research have provided a wealth of large data sets that are too big to be analyzed in their entirety, due to restrictions on computer memory or storage size. New Bayesian methods have been developed for large data sets that are only large due to large sample sizes; these methods partition big data sets into subsets, and perform independent Bayesian Markov chain Monte Carlo analyses on the subsets. The methods then combine the independent subset posterior samples to estimate a posterior density given the full data set. These approaches were shown to be effective for Bayesian models including logistic regression models, Gaussian mixture models and hierarchical models. Here, we introduce the R package parallelMCMCcombine which carries out four of these techniques for combining independent subset posterior samples. We illustrate each of the methods using a Bayesian logistic regression model for simulation data and a Bayesian Gamma model for real data; we also demonstrate features and capabilities of the R package. The package assumes the user has carried out the Bayesian analysis and has produced the independent subposterior samples outside of the package. The methods are primarily suited to models with unknown parameters of fixed dimension that exist in continuous parameter spaces. We envision this tool will allow researchers to explore the various methods for their specific applications, and will assist future progress in this rapidly developing field.

preprint2014arXiv

Relative entropy in hyperbolic relaxation for balance laws

We present a general framework for the approximation of systems of hyperbolic balance laws. The novelty of the analysis lies in the construction of suitable relaxation systems and the derivation of a delicate estimate on the relative entropy. We provide a direct proof of convergence in the smooth regime for a wide class of physical systems. We present results for systems arising in materials science, where the presence of source terms presents a number of additional challenges and requires delicate treatment. Our analysis is in the spirit of the framework introduced by Tzavaras [A. Tzavaras, Commun. Math. Sci., 3-2, 2005] for systems of hyperbolic conservation laws.