Source author record

Martin Jankowiak

Martin Jankowiak appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

14works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

14 published item(s)

preprint2022arXiv

Fast Bayesian Variable Selection in Binomial and Negative Binomial Regression

Bayesian variable selection is a powerful tool for data analysis, as it offers a principled method for variable selection that accounts for prior information and uncertainty. However, wider adoption of Bayesian variable selection has been hampered by computational challenges, especially in difficult regimes with a large number of covariates or non-conjugate likelihoods. Generalized linear models for count data, which are prevalent in biology, ecology, economics, and beyond, represent an important special case. Here we introduce an efficient MCMC scheme for variable selection in binomial and negative binomial regression that exploits Tempered Gibbs Sampling (Zanella and Roberts, 2019) and that includes logistic regression as a special case. In experiments we demonstrate the effectiveness of our approach, including on cancer data with seventeen thousand covariates.

preprint2022arXiv

Scalable Cross Validation Losses for Gaussian Process Models

We introduce a simple and scalable method for training Gaussian process (GP) models that exploits cross-validation and nearest neighbor truncation. To accommodate binary and multi-class classification we leverage Pòlya-Gamma auxiliary variables and variational inference. In an extensive empirical comparison with a number of alternative methods for scalable GP regression and classification, we find that our method offers fast training and excellent predictive performance. We argue that the good predictive performance can be traced to the non-parametric nature of the resulting predictive distributions as well as to the cross-validation loss, which provides robustness against model mis-specification.

preprint2022arXiv

Surrogate Likelihoods for Variational Annealed Importance Sampling

Variational inference is a powerful paradigm for approximate Bayesian inference with a number of appealing properties, including support for model learning and data subsampling. By contrast MCMC methods like Hamiltonian Monte Carlo do not share these properties but remain attractive since, contrary to parametric methods, MCMC is asymptotically unbiased. For these reasons researchers have sought to combine the strengths of both classes of algorithms, with recent approaches coming closer to realizing this vision in practice. However, supporting data subsampling in these hybrid methods can be a challenge, a shortcoming that we address by introducing a surrogate likelihood that can be learned jointly with other variational parameters. We argue theoretically that the resulting algorithm permits the user to make an intuitive trade-off between inference fidelity and computational cost. In an extensive empirical comparison we show that our method performs well in practice and that it is well-suited for black-box inference in probabilistic programming frameworks.

preprint2020arXiv

A Unified Stochastic Gradient Approach to Designing Bayesian-Optimal Experiments

We introduce a fully stochastic gradient based approach to Bayesian optimal experimental design (BOED). Our approach utilizes variational lower bounds on the expected information gain (EIG) of an experiment that can be simultaneously optimized with respect to both the variational and design parameters. This allows the design process to be carried out through a single unified stochastic gradient ascent procedure, in contrast to existing approaches that typically construct a pointwise EIG estimator, before passing this estimator to a separate optimizer. We provide a number of different variational objectives including the novel adaptive contrastive estimation (ACE) bound. Finally, we show that our gradient-based approaches are able to provide effective design optimization in substantially higher dimensional settings than existing approaches.

preprint2020arXiv

Functional Tensors for Probabilistic Programming

It is a significant challenge to design probabilistic programming systems that can accommodate a wide variety of inference strategies within a unified framework. Noting that the versatility of modern automatic differentiation frameworks is based in large part on the unifying concept of tensors, we describe a software abstraction for integration --functional tensors-- that captures many of the benefits of tensors, while also being able to describe continuous probability distributions. Moreover, functional tensors are a natural candidate for generalized variable elimination and parallel-scan filtering algorithms that enable parallel exact inference for a large family of tractable modeling motifs. We demonstrate the versatility of functional tensors by integrating them into the modeling frontend and inference backend of the Pyro programming language. In experiments we show that the resulting framework enables a large variety of inference strategies, including those that mix exact and approximate inference.

preprint2020arXiv

Variational Bayesian Optimal Experimental Design

Bayesian optimal experimental design (BOED) is a principled framework for making efficient use of limited experimental resources. Unfortunately, its applicability is hampered by the difficulty of obtaining accurate estimates of the expected information gain (EIG) of an experiment. To address this, we introduce several classes of fast EIG estimators by building on ideas from amortized variational inference. We show theoretically and empirically that these estimators can provide significant gains in speed and accuracy over previous approaches. We further demonstrate the practicality of our approach on a number of end-to-end experiments.

preprint2014arXiv

Constraining CP-violating Higgs Sectors at the LHC using gluon fusion

We investigate the constraints that the LHC can set on a 126 GeV Higgs boson that is an admixture of CP eigenstates. Traditional analyses rely on Higgs couplings to massive vector bosons, which are suppressed for CP-odd couplings, so that these analyses have limited sensitivity. Instead we focus on Higgs production in gluon fusion, which occurs at the same order in the strong coupling for both CP-even and -odd couplings. We study the Higgs plus two jet final state followed by Higgs decay into a pair of tau leptons. We show that using the 8 TeV dataset it is possible to rule out the pure CP-odd hypothesis in this channel alone at nearly 95\% C.L, assuming that the Higgs is CP-even. We also provide projected limits for the 14 TeV LHC run.

preprint2014arXiv

Jet Substructure Templates: Data-driven QCD Backgrounds for Fat Jet Searches

QCD is often the dominant background to new physics searches for which jet substructure provides a useful handle. Due to the challenges associated with modeling this background, data-driven approaches are necessary. This paper presents a novel method for determining QCD predictions using templates -- probability distribution functions for jet substructure properties as a function of kinematic inputs. Templates can be extracted from a control region and then used to compute background distributions in the signal region. Using Monte Carlo, we illustrate the procedure with two case studies and show that the template approach effectively models the relevant QCD background. This work strongly motivates the application of these techniques to LHC data.

preprint2013arXiv

Learning How to Count: A High Multiplicity Search for the LHC

We introduce a search technique that is sensitive to a broad class of signals with large final state multiplicities. Events are clustered into large radius jets and jet substructure techniques are used to count the number of subjets within each jet. The search consists of a cut on the total number of subjets in the event as well as the summed jet mass and missing energy. Two different techniques for counting subjets are described and expected sensitivities are presented for eight benchmark signals. These signals exhibit diverse phenomenology, including 2-step cascade decays, direct three body decays, and multi-top final states. We find improved sensitivity to these signals as compared to previous high multiplicity searches as well as a reduced reliance on missing energy requirements. One benefit of this approach is that it allows for natural data driven estimates of the QCD background.

preprint2013arXiv

LHC probes the hidden sector

In this note we establish LHC limits on a variety of benchmark models for hidden sector physics using 2011 and 2012 data. First, we consider a "hidden" U(1) gauge boson under which all Standard Model particles are uncharged at tree-level and which interacts with the visible sector either via kinetic mixing or higher dimensional operators. Second, we constrain scalar and pseudo-scalar particles interacting with the Standard Model via dimension five operators and Yukawa interactions, in particular including so-called axion-like particles. In both cases we consider several different final states, including photons, electrons, muons and taus, establishing new constraints for a range of GeV to TeV scale masses. Finally, we also comment on particles with electric charges smaller than e that arise from hidden sector matter.

preprint2012arXiv

Angular Scaling in Jets

We introduce a jet shape observable defined for an ensemble of jets in terms of two-particle angular correlations and a resolution parameter R. This quantity is infrared and collinear safe and can be interpreted as a scaling exponent for the angular distribution of mass inside the jet. For small R it is close to the value 2 as a consequence of the approximately scale invariant QCD dynamics. For large R it is sensitive to non-perturbative effects. We describe the use of this correlation function for tests of QCD, for studying underlying event and pile-up effects, and for tuning Monte Carlo event generators.

preprint2011arXiv

Jet Substructure Without Trees

We present an alternative approach to identifying and characterizing jet substructure. An angular correlation function is introduced that can be used to extract angular and mass scales within a jet without reference to a clustering algorithm. This procedure gives rise to a number of useful jet observables. As an application, we construct a top quark tagging algorithm that is competitive with existing methods.

preprint2010arXiv

Nearly Supersymmetric Dark Atoms

Theories of dark matter that support bound states are an intriguing possibility for the identity of the missing mass of the Universe. This article proposes a class of models of supersymmetric composite dark matter where the interactions with the Standard Model communicate supersymmetry breaking to the dark sector. In these models supersymmetry breaking can be treated as a perturbation on the spectrum of bound states. Using a general formalism, the spectrum with leading supersymmetry effects is computed without specifying the details of the binding dynamics. The interactions of the composite states with the Standard Model are computed and several benchmark models are described. General features of non-relativistic supersymmetric bound states are emphasized.