Source author record

Carlo Berzuini

Carlo Berzuini appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2022arXiv

Bayesian Mendelian randomization testing of interval causal null hypotheses: ternary decision rules and loss function calibration

Our approach to Mendelian Randomization (MR) analysis is designed to increase reproducibility of causal effect "discoveries" by: (i) using a Bayesian approach to inference; (ii) replacing the point null hypothesis with a region of practical equivalence consisting of values of negligible magnitude for the effect of interest, while exploiting the ability of Bayesian analysis to quantify the evidence of the effect falling inside/outside the region; (iii) rejecting the usual binary decision logic in favour of a ternary logic where the hypothesis test may result in either an acceptance or a rejection of the null, while also accommodating an "uncertain" outcome. We present an approach to calibration of the proposed method via loss function, which we use to compare our approach with a frequentist one. We illustrate the method with the aid of a study of the causal effect of obesity on risk of juvenile myocardial infarction.

preprint2022arXiv

Discovery methods for systematic analysis of causal molecular networks in modern omics datasets

With the increasing availability and size of multi-omics datasets, investigating the casual relationships between molecular phenotypes has become an important aspect of exploring underlying biology and genetics. This paper aims to introduce and review the available methods for building large-scale causal molecular networks that have been developed in the past decade. Existing methods have their own strengths and limitations so there is no one best approach, and it is instead down to the discretion of the researcher. This review also aims to discuss some of the current limitations to biological interpretation of these networks, and important factors to consider for future studies on molecular networks.

preprint2020arXiv

Mendelian Randomization with Incomplete Exposure Data: a Bayesian Approach

We expand Mendelian Randomization (MR) methodology to deal with randomly missing data on either the exposure or the outcome variable, and furthermore with data from nonindependent individuals (eg components of a family). Our method rests on the Bayesian MR framework proposed by Berzuini et al (2018), which we apply in a study of multiplex Multiple Sclerosis (MS) Sardinian families to characterise the role of certain plasma proteins in MS causation. The method is robust to presence of pleiotropic effects in an unknown number of instruments, and is able to incorporate inter-individual kinship information. Introduction of missing data allows us to overcome the bias introduced by the (reverse) effect of treatment (in MS cases) on level of protein. From a substantive point of view, our study results confirm recent suspicion that an increase in circulating IL12A and STAT4 protein levels does not cause an increase in MS risk, as originally believed, suggesting that these two proteins may not be suitable drug targets for MS.

preprint2014arXiv

Stochastic Mechanistic Interaction

We propose a fully probabilistic formulation of the notion of mechanistic interaction (interaction in some fundamental mechanistic sense) between the effects of putative (possibly continuous) causal factors A and B on a binary outcome variable Y indicating 'survival' vs 'failure'. We define mechanistic interaction in terms of departure from a generalized 'noisy OR' model, under which the multiplicative causal effect of A (resp., B) on the probability of failure cannot be enhanced by manipulating B (resp., A). We present conditions under which mechanistic interaction in the above sense can be assessed via simple tests on excess risk or superadditivity, in a possibly retrospective regime of observation. These conditions are defined in terms of generalized conditional independence relationships (generalised because they may involve non-stochastic 'regime indicators') that can often be checked on a graphical representation of the problem. Inference about mechanistic interaction between direct, or path-specific, causal effects can be accommodated in the proposed framework. The method is illustrated with the aid of a study in experimental psychology.

preprint2013arXiv

Temporal Reasoning with Probabilities

In this paper we explore representations of temporal knowledge based upon the formalism of Causal Probabilistic Networks (CPNs). Two different ?continuous-time? representations are proposed. In the first, the CPN includes variables representing ?event-occurrence times?, possibly on different time scales, and variables representing the ?state? of the system at these times. In the second, the CPN describes the influences between random variables with values in () representing dates, i.e. time-points associated with the occurrence of relevant events. However, structuring a system of inter-related dates as a network where all links commit to a single specific notion of cause and effect is in general far from trivial and leads to severe difficulties. We claim that we should recognize explicitly different kinds of relation between dates, such as ?cause?, ?inhibition?, ?competition?, etc., and propose a method whereby these relations are coherently embedded in a CPN using additional auxiliary nodes corresponding to "instrumental" variables. Also discussed, though not covered in detail, is the topic concerning how the quantitative specifications to be inserted in a temporal CPN can be learned from specific data.

preprint2011arXiv

Direct genetic effects and their estimation from matched case-control data

In genetic association studies, a single marker is often associated with multiple, correlated phenotypes (e.g., obesity and cardiovascular disease, or nicotine dependence and lung cancer). A pervasive question is then whether that marker has independent effects on all phenotypes. In this article, we address this question by assessing whether there is a direct genetic effect on one phenotype that is not mediated through the other phenotypes. In particular, we investigate how to identify and estimate such direct genetic effects on the basis of (matched) case-control data. We discuss conditions under which such effects are identifiable from the available (matched) case-control data. We find that direct genetic effects are sometimes estimable via standard regression methods, and sometimes via a more general G-estimation method, which has previously been proposed for random samples and unmatched case-control studies (Vansteelandt, 2009) and is here extended to matched case-control studies. The results are used to assess whether the FTO gene is associated with myocardial infarction other than via an effect on obesity.

preprint2010arXiv

Deep determinism and the assessment of mechanistic interaction between categorical and continuous variables

Our aim is to detect mechanistic interaction between the effects of two causal factors on a binary response, as an aid to identifying situations where the effects are mediated by a common mechanism. We propose a formalization of mechanistic interaction which acknowledges asymmetries of the kind "factor A interferes with factor B, but not viceversa". A class of tests for mechanistic interaction is proposed, which works on discrete or continuous causal variables, in any combination. Conditions under which these tests can be applied under a generic regime of data collection, be it interventional or observational, are discussed in terms of conditional independence assumptions within the framework of Augmented Directed Graphs. The scientific relevance of the method and the practicality of the graphical framework are illustrated with the aid of two studies in coronary artery disease. Our analysis relies on the "deep determinism" assumption that there exists some relevant set V - possibly unobserved - of "context variables", such that the response Y is a deterministic function of the values of V and of the causal factors of interest. Caveats regarding this assumption in real studies are discussed.