Source author record

David Jensen

David Jensen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

29works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

29 published item(s)

preprint2026arXiv

The Imperative for Grand Challenges in Computing

Computing is an indispensable component of nearly all technologies and is ubiquitous for vast segments of society. It is also essential to discoveries and innovations in most disciplines. However, while past grand challenges in science have involved computing as one of the tools to address the challenge, these challenges have not been principally about computing. Why has the computing community not yet produced challenges at the scale of grandeur that we see in disciplines such as physics, astronomy, or engineering? How might we go about identifying similarly grand challenges? What are the grand challenges of computing that transcend our discipline's traditional boundaries and have the potential to dramatically improve our understanding of the world and positively shape the future of our society? There is a significant benefit in us, as a field, taking a more intentional approach to "grand challenges." We are seeking challenge problems that are sufficiently compelling as to both ignite the imagination of computer scientists and draw researchers from other disciplines to computational challenges. This paper emphasizes the importance, now more than ever, of defining and pursuing grand challenges in computing as a field, and being intentional about translation and realizing its impacts on science and society. Building on lessons from prior grand challenges, the paper explores the nature of a grand challenge today emphasizing both scale and impact, and how the community may tackle such a grand challenge, given a rapidly changing innovation ecosystem in computing. The paper concludes with a call to action for our community to come together to define grand challenges in computing for the next decade and beyond.

preprint2020arXiv

Brill-Noether theory for curves of a fixed gonality

We prove a generalization of the Brill-Noether theorem for the variety of special divisors $W^r_d(C)$ on a general curve $C$ of prescribed gonality. Our main theorem gives a closed formula for the dimension of $W^r_d(C)$. We build on previous work of Pflueger, who used an analysis of the tropical divisor theory of special chains of cycles to give upper bounds on the dimensions of Brill--Noether varieties on such curves. We prove his conjecture, that this upper bound is achieved for a general curve. Our methods introduce logarithmic stable maps as a systematic tool in Brill-Noether theory. A precise relation between the divisor theory on chains of cycles and the corresponding tropical maps theory is exploited to prove new regeneration theorems for linear series with negative Brill-Noether number. The strategy involves blending an analysis of obstruction theories for logarithmic stable maps with the geometry of Berkovich curves. To show the utility of these methods, we provide a short new derivation of lifting for special divisors on a chain of cycles with generic edge lengths, proved using different techniques by Cartwright, Jensen, and Payne. A crucial technical result is a new realizability theorem for tropical stable maps in obstructed geometries, generalizing a well-known theorem of Speyer on genus one curves to arbitrary genus.

preprint2020arXiv

Causal Inference using Gaussian Processes with Structured Latent Confounders

Latent confounders---unobserved variables that influence both treatment and outcome---can bias estimates of causal effects. In some cases, these confounders are shared across observations, e.g. all students taking a course are influenced by the course's difficulty in addition to any educational interventions they receive individually. This paper shows how to semiparametrically model latent confounders that have this structure and thereby improve estimates of causal effects. The key innovations are a hierarchical Bayesian model, Gaussian processes with structured latent confounders (GP-SLC), and a Monte Carlo inference algorithm for this model based on elliptical slice sampling. GP-SLC provides principled Bayesian uncertainty estimates of individual treatment effect with minimal assumptions about the functional forms relating confounders, covariates, treatment, and outcome. Finally, this paper shows GP-SLC is competitive with or more accurate than widely used causal inference techniques on three benchmark datasets, including the Infant Health and Development Program and a dataset showing the effect of changing temperatures on state-wide energy consumption across New England.

preprint2020arXiv

Exploratory Not Explanatory: Counterfactual Analysis of Saliency Maps for Deep Reinforcement Learning

Saliency maps are frequently used to support explanations of the behavior of deep reinforcement learning (RL) agents. However, a review of how saliency maps are used in practice indicates that the derived explanations are often unfalsifiable and can be highly subjective. We introduce an empirical approach grounded in counterfactual reasoning to test the hypotheses generated from saliency maps and assess the degree to which they correspond to the semantics of RL environments. We use Atari games, a common benchmark for deep RL, to evaluate three types of saliency maps. Our results show the extent to which existing claims about Atari games can be evaluated and suggest that saliency maps are best viewed as an exploratory tool rather than an explanatory tool.

preprint2020arXiv

Text and Causal Inference: A Review of Using Text to Remove Confounding from Causal Estimates

Many applications of computational social science aim to infer causal conclusions from non-experimental data. Such observational data often contains confounders, variables that influence both potential causes and potential effects. Unmeasured or latent confounders can bias causal estimates, and this has motivated interest in measuring potential confounders from observed text. For example, an individual's entire history of social media posts or the content of a news article could provide a rich measurement of multiple confounders. Yet, methods and applications for this problem are scattered across different communities and evaluation practices are inconsistent. This review is the first to gather and categorize these examples and provide a guide to data-processing and evaluation decisions. Despite increased attention on adjusting for confounding using text, there are still many open problems, which we highlight in this paper.

preprint2020arXiv

Tropical Methods in Hurwitz-Brill-Noether Theory

Splitting type loci are the natural generalizations of Brill-Noether varieties for curves with a distinguished map to the projective line. We give a tropical proof of a theorem of H. Larson, showing that splitting type loci have the expected dimension for general elements of the Hurwitz space. Our proof uses an explicit description of splitting type loci on a certain family of tropical curves. We further show that these tropical splitting type loci are connected in codimension one, and describe an algorithm for computing their cardinality when they are zero-dimensional. We provide a conjecture for the numerical class of splitting type loci, which we confirm in a number of cases.

preprint2016arXiv

A simplicial approach to effective divisors in $\overline{M}_{0,n}$

We study the Cox ring and monoid of effective divisor classes of $\overline{M}_{0,n} = Bl\mathbb{P}^{n-3}$, over a ring R. We provide a bijection between elements of the Cox ring, not divisible by any exceptional divisor section, and pure-dimensional singular simplicial complexes on {1,...,n-1} with nonzero weights in R satisfying a zero-tension condition. This leads to a combinatorial criterion, satisfied by many triangulations of closed manifolds, for a divisor class to be among the minimal generators for the effective monoid. For classes obtained as the strict transform of quadrics, we present a complete classification of minimal generators, generalizing to all n the well-known Keel-Vermeire classes for n=6. We use this classification to construct new divisors with interesting properties for all n > 6.

preprint2016arXiv

Causal Discovery for Manufacturing Domains

Yield and quality improvement is of paramount importance to any manufacturing company. One of the ways of improving yield is through discovery of the root causal factors affecting yield. We propose the use of data-driven interpretable causal models to identify key factors affecting yield. We focus on factors that are measured in different stages of production and testing in the manufacturing cycle of a product. We apply causal structure learning techniques on real data collected from this line. Specifically, the goal of this work is to learn interpretable causal models from observational data produced by manufacturing lines. Emphasis has been given to the interpretability of the models to make them actionable in the field of manufacturing. We highlight the challenges presented by assembly line data and propose ways to alleviate them.We also identify unique characteristics of data originating from assembly lines and how to leverage them in order to improve causal discovery. Standard evaluation techniques for causal structure learning shows that the learned causal models seem to closely represent the underlying latent causal relationship between different factors in the production process. These results were also validated by manufacturing domain experts who found them promising. This work demonstrates how data mining and knowledge discovery can be used for root cause analysis in the domain of manufacturing and connected industry.

preprint2016arXiv

Evaluating Causal Models by Comparing Interventional Distributions

The predominant method for evaluating the quality of causal models is to measure the graphical accuracy of the learned model structure. We present an alternative method for evaluating causal models that directly measures the accuracy of estimated interventional distributions. We contrast such distributional measures with structural measures, such as structural Hamming distance and structural intervention distance, showing that structural measures often correspond poorly to the accuracy of estimated interventional distributions. We use a number of real and synthetic datasets to illustrate various scenarios in which structural measures provide misleading results with respect to algorithm selection and parameter tuning, and we recommend that distributional measures become the new standard for evaluating causal models.

preprint2015arXiv

Tropical independence II: The maximal rank conjecture for quadrics

Building on our earlier results on tropical independence and shapes of divisors in tropical linear series, we give a tropical proof of the maximal rank conjecture for quadrics. We also prove a tropical analogue of Max Noether's theorem on quadrics containing a canonically embedded curve, and state a combinatorial conjecture about tropical independence on chains of loops that implies the maximal rank conjecture for algebraic curves.

preprint2014arXiv

Learning to Generate Networks

We investigate the problem of learning to generate complex networks from data. Specifically, we consider whether deep belief networks, dependency networks, and members of the exponential random graph family can learn to generate networks whose complex behavior is consistent with a set of input examples. We find that the deep model is able to capture the complex behavior of small networks, but that no model is able capture this behavior for networks with more than a handful of nodes.

preprint2014arXiv

Online Dating Recommendations: Matching Markets and Learning Preferences

Recommendation systems for online dating have recently attracted much attention from the research community. In this paper we proposed a two-side matching framework for online dating recommendations and design an LDA model to learn the user preferences from the observed user messaging behavior and user profile features. Experimental results using data from a large online dating website shows that two-sided matching improves significantly the rate of successful matches by as much as 45%. Finally, using simulated matchings we show that the the LDA model can correctly capture user preferences.

preprint2014arXiv

Reasoning about Independence in Probabilistic Models of Relational Data

We extend the theory of d-separation to cases in which data instances are not independent and identically distributed. We show that applying the rules of d-separation directly to the structure of probabilistic models of relational data inaccurately infers conditional independence. We introduce relational d-separation, a theory for deriving conditional independence facts from relational models. We provide a new representation, the abstract ground graph, that enables a sound, complete, and computationally efficient method for answering d-separation queries about relational models, and we present empirical results that demonstrate effectiveness.

preprint2014arXiv

Refining the Semantics of Social Influence

With the proliferation of network data, researchers are increasingly focusing on questions investigating phenomena occurring on networks. This often includes analysis of peer-effects, i.e., how the connections of an individual affect that individual's behavior. This type of influence is not limited to direct connections of an individual (such as friends), but also to individuals that are connected through longer paths (for example, friends of friends, or friends of friends of friends). In this work, we identify an ambiguity in the definition of what constitutes the extended neighborhood of an individual. This ambiguity gives rise to different semantics and supports different types of underlying phenomena. We present experimental results, both on synthetic and real networks, that quantify differences among the sets of extended neighbors under different semantics. Finally, we provide experimental evidence that demonstrates how the use of different semantics affects model selection.

preprint2013arXiv

A Sound and Complete Algorithm for Learning Causal Models from Relational Data

The PC algorithm learns maximally oriented causal Bayesian networks. However, there is no equivalent complete algorithm for learning the structure of relational models, a more expressive generalization of Bayesian networks. Recent developments in the theory and representation of relational models support lifted reasoning about conditional independence. This enables a powerful constraint for orienting bivariate dependencies and forms the basis of a new algorithm for learning structure. We present the relational causal discovery (RCD) algorithm that learns causal relational models. We prove that RCD is sound and complete, and we present empirical results that demonstrate effectiveness.

preprint2013arXiv

Identifying Independence in Relational Models

The rules of d-separation provide a framework for deriving conditional independence facts from model structure. However, this theory only applies to simple directed graphical models. We introduce relational d-separation, a theory for deriving conditional independence in relational models. We provide a sound, complete, and computationally efficient method for relational d-separation, and we present empirical results that demonstrate effectiveness.

preprint2013arXiv

Log canonical models and variation of GIT for genus four canonical curves

We discuss GIT for canonically embedded genus four curves and the connection to the Hassett-Keel program. A canonical genus four curve is a complete intersection of a quadric and a cubic, and, in contrast to the genus three case, there is a family of GIT quotients that depend on a choice of linearization. We discuss the corresponding VGIT problem and show that the resulting spaces give the final steps in the Hassett-Keel program for genus four curves.

preprint2012arXiv

The geometry of the ball quotient model of the moduli space of genus four curves

S. Kondo has constructed a ball quotient compactification for the moduli space of non-hyperelliptic genus four curves. In this paper, we show that this space essentially coincides with a GIT quotient of the Chow variety of canonically embedded genus four curves. More specifically, we give an explicit description of this GIT quotient, and show that the birational map from this space to Kondo's space is resolved by the blow-up of a single point. This provides a modular interpretation of the points in the boundary of Kondo's space. Connections with the slope nine space in the Hassett-Keel program are also discussed.

preprint2011arXiv

Birational Contractions of $\bar{M}_{3,1}$ and $\bar{M}_{4,1}$

We study the birational geometry of $\bar{M}_{3,1}$ and $\bar{M}_{4,1}$. In particular, we pose a pointed analogue of the Slope Conjecture and prove it in these low-genus cases. Using variation of GIT, we construct birational contractions of these spaces in which certain divisors of interest -- the pointed Brill-Noether divisors -- are contracted. As a consequence, we see that these pointed Brill-Noether divisors generate extremal rays of the effective cones for these spaces.