Source author record

Alona Kryshchenko

Alona Kryshchenko appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2022arXiv

Semi-supervised Nonnegative Matrix Factorization for Document Classification

We propose new semi-supervised nonnegative matrix factorization (SSNMF) models for document classification and provide motivation for these models as maximum likelihood estimators. The proposed SSNMF models simultaneously provide both a topic model and a model for classification, thereby offering highly interpretable classification results. We derive training methods using multiplicative updates for each new model, and demonstrate the application of these models to single-label and multi-label document classification, although the models are flexible to other supervised learning tasks such as regression. We illustrate the promise of these models and training methods on document classification datasets (e.g., 20 Newsgroups, Reuters).

preprint2020arXiv

COVID-19 Literature Topic-Based Search via Hierarchical NMF

A dataset of COVID-19-related scientific literature is compiled, combining the articles from several online libraries and selecting those with open access and full text available. Then, hierarchical nonnegative matrix factorization is used to organize literature related to the novel coronavirus into a tree structure that allows researchers to search for relevant literature based on detected topics. We discover eight major latent topics and 52 granular subtopics in the body of literature, related to vaccines, genetic structure and modeling of the disease and patient studies, as well as related diseases and virology. In order that our tool may help current researchers, an interactive website is created that organizes available literature using this hierarchical structure.

preprint2015arXiv

benchNGS : An approach to benchmark short reads alignment tools

In the last decade a number of algorithms and associated software have been developed to align next generation sequencing (NGS) reads with relevant reference genomes. The accuracy of these programs may vary significantly, especially when the NGS reads are quite different from the available reference genome. We propose a benchmark to assess accuracy of short reads mapping based on the pre-computed global alignment of related genome sequences. In this paper we propose a benchmark to assess accuracy of the short reads mapping based on the pre-computed global alignment of closely related genome sequences. We outline the method and also present a short report of an experiment performed on five popular alignment tools based on the pairwise alignments of Escherichia coli O6 CFT073 genome with genomes of seven other bacteria.

preprint2015arXiv

Nonparametric estimation of a mixing distribution for a family of linear stochastic dynamical systems

In this paper we develop a nonparametric maximum likelihood estimate of the mixing distribution of the parameters of a linear stochastic dynamical system. This includes, for example, pharmacokinetic population models with process and measurement noise that are linear in the state vector, input vector and the process and measurement noise vectors. Most research in mixing distributions only considers measurement noise. The advantages of the models with process noise are that, in addition to the measurements errors, the uncertainties in the model itself are taken into the account. For example, for deterministic pharmacokinetic models, errors in dose amounts, administration times, and timing of blood samples are typically not included. For linear stochastic models, we use linear Kalman-Bucy filtering to calculate the likelihood of the observations and then employ a nonparametric adaptive grid algorithm to find the nonparametric maximum likelihood estimate of the mixing distribution. We then use the directional derivatives of the estimated mixing distribution to show that the result found attains a global maximum. A simple example using a one compartment pharmacokinetic linear stochastic model is given. In addition to population pharmacokinetics, this research also applies to empirical Bayes estimation.