Source author record

Roberto Fontana

Roberto Fontana appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

13works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

13 published item(s)

preprint2026arXiv

Characterization of multi-way binary tables with uniform margins and fixed correlations

In many applications involving binary variables, only pairwise dependence measures, such as correlations, are available. However, for multi-way tables involving more than two variables, these quantities do not uniquely determine the joint distribution, but instead define a family of admissible distributions that share the same pairwise dependence while potentially differing in higher-order interactions. In this paper, we introduce a geometric framework to describe the entire feasible set of such joint distributions with uniform margins. We show that this admissible set forms a convex polytope, analyze its symmetry properties, and characterize its extreme rays. These extremal distributions provide fundamental insights into how higher-order dependence structures may vary while preserving the prescribed pairwise information. Unlike traditional methods for table generation, which return a single table, our framework makes it possible to explore and understand the full admissible space of dependence structures, enabling more flexible choices for modeling and simulation. We illustrate the usefulness of our theoretical results through examples and a real case study on rater agreement.

preprint2022arXiv

High dimensional Bernoulli distributions: algebraic representation and applications

The main contribution of this paper is to find a representation of the class $\mathcal{F}_d(p)$ of multivariate Bernoulli distributions with the same mean $p$ that allows us to find its generators analytically in any dimension. We map $\mathcal{F}_d(p)$ to an ideal of points and we prove that the class $\mathcal{F}_d(p)$ can be generated from a finite set of simple polynomials. We present two applications. Firstly, we show that polynomial generators help to find extremal points of the convex polytope $\mathcal{F}_d(p)$ in high dimensions. Secondly, we solve the problem of determining the lower bounds in the convex order for sums of multivariate Bernoulli distributions with given margins, but with an unspecified dependence structure.

preprint2022arXiv

Robustness against data loss with Algebraic Statistics

The paper describes an algorithm that, given an initial design $\mathcal{F}_n$ of size $n$ and a linear model with $p$ parameters, provides a sequence $\mathcal{F}_n \supset \ldots \supset \mathcal{F}_{n-k} \supset \ldots \supset \mathcal{F}_p$ of nested \emph{robust} designs. The sequence is obtained by the removal, one by one, of the runs of $\mathcal{F}_n$ till a $p$-run \emph{saturated} design $\mathcal{F}_p$ is obtained. The potential impact of the algorithm on real applications is high. The initial fraction $\mathcal{F}_n$ can be of any type and the output sequence can be used to organize the experimental activity. The experiments can start with the runs corresponding to $\mathcal{F}_p$ and continue adding one run after the other (from $\mathcal{F}_{n-k}$ to $\mathcal{F}_{n-k+1}$) till the initial design $\mathcal{F}_n$ is obtained. In this way, if for some unexpected reasons the experimental activity must be stopped before the end when only $n-k$ runs are completed, the corresponding $\mathcal{F}_{n-k}$ has a high value of robustness for $k \in \{1, \ldots, n-p\}$. The algorithm uses the circuit basis, a special representation of the kernel of a matrix with integer entries. The effectiveness of the algorithm is demonstrated through the use of simulations.

preprint2021arXiv

Exchangeable Bernoulli distributions: high dimensional simulation, estimate and testing

We explore the class of exchangeable Bernoulli distributions building on their geometrical structure. Exchangeable Bernoulli probability mass functions are points in a convex polytope and we have found analytical expressions for their extremal generators. The geometrical structure turns out to be crucial to simulate high dimensional and negatively correlated binary data. Furthermore, for a wide class of statistical indices and measures of a probability mass function we are able to find not only their sharp bounds in the class, but also their distribution across the class. Estimate and testing are also addressed.

preprint2016arXiv

Graphical models for studying museum networks: the Abbonamento Musei Torino Piemonte

Probabilistic graphical models are a powerful tool to represent real-word phenomena and to learn network structures starting from data. This paper applies graphical models in a new framework to study association rules driven by consumer choices in a network of museums. The network consists of the museums participating in the program of Abbonamento Musei Torino Piemonte, which is a yearly subscription managed by the Associazione Torino Città Capitale Europea. Consumers are card-holders, who are allowed to entry to all the museums in the network for one year. We employ graphical models to highlight associations among the museums driven by card-holder visiting behaviour. We use both undirected graphs to investigate the strength of the network and directed graphs to highlight asimmetry in the association rules.

preprint2016arXiv

Simulations on the combinatorial structure of D-optimal designs

In this work we present the results of several simulations on main-effect factorial designs. The goal of such simulations is to investigate the connections between the $D$-optimality of a design and its geometrical structure. By means of a combinatorial object, namely the circuit basis of the design matrix, we show that it is possible to define a simple index that exhibits strong connections with the $D$-optimality.

preprint2015arXiv

Aberration in qualitative multilevel designs

Generalized Word Length Pattern (GWLP) is an important and widely-used tool for comparing fractional factorial designs. We consider qualitative factors, and we code their levels using the roots of the unity. We write the GWLP of a fraction ${\mathcal F}$ using the polynomial indicator function, whose coefficients encode many properties of the fraction. We show that the coefficient of a simple or interaction term can be written using the counts of its levels. This apparently simple remark leads to major consequence, including a convolution formula for the counts. We also show that the mean aberration of a term over the permutation of its levels provides a connection with the variance of the level counts. Moreover, using mean aberrations for symmetric $s^m$ designs with $s$ prime, we derive a new formula for computing the GWLP of ${\mathcal F}$. It is computationally easy, does not use complex numbers and also provides a clear way to interpret the GWLP. As case studies, we consider non-isomorphic orthogonal arrays that have the same GWLP. The different distributions of the mean aberrations suggest that they could be used as a further tool to discriminate between fractions.

preprint2015arXiv

Generalized Minimum Aberration mixed-level orthogonal arrays A general approach based on sequential integer quadratically constrained quadratic programming

Orthogonal Fractional Factorial Designs and in particular Orthogonal Arrays are frequently used in many fields of application, including medicine, engineering and agriculture. In this paper we present a methodology and an algorithm to find an orthogonal array, of given size and strength, that satisfies the generalized minimum aberration criterion. The methodology is based on the joint use of polynomial counting functions, complex coding of levels and algorithms for quadratic optimization and puts no restriction on the number of levels of each factor.

preprint2014arXiv

$D$-optimal saturated designs: a simulation study

In this work we focus on saturated $D$-optimal designs. Using recent results, we identify $D$-optimal designs with the solutions of an optimization problem with linear constraints. We introduce new objective functions based on the geometric structure of the design and we compare them with the classical $D$-efficiency criterion. We perform a simulation study. In all the test cases we observe that designs with high values of $D$-efficiency have also high values of the new objective functions.

preprint2013arXiv

A Characterization of Saturated Designs for Factorial Experiments

In this paper we study saturated fractions of factorial designs under the perspective of Algebraic Statistics. We define a criterion to check whether a fraction is saturated or not with respect to a given model. The proposed criterion is based purely on combinatorial objects. Our technique is particularly useful when several fractions are needed. We also show how to generate random saturated fractions with given projections, by applying the theory of Markov bases for contingency tables.

preprint2013arXiv

Random generation of optimal saturated designs

Efficient algorithms for searching for optimal saturated designs are widely available. They maximize a given efficiency measure (such as D-optimality) and provide an optimum design. Nevertheless, they do not guarantee a \emph{global} optimal design. Indeed, they start from an initial random design and find a local optimal design. If the initial design is changed the optimum found will, in general, be different. A natural question arises. Should we stop at the design found or should we run the algorithm again in search of a better design? This paper uses very recent methods and software for discovery probability to support the decision to continue or stop the sampling. A software tool written in SAS has been developed.

preprint2012arXiv

Saturated fractions of two-factor designs

In this paper we study saturated fractions of a two-factor design under the simple effect model. In particular, we define a criterion to check whether a given fraction is saturated or not, and we compute the number of saturated fractions. All proofs are constructive and can be used as actual methods to build saturated fractions. Moreover, we show how the theory of Markov bases for contingency tables can be applied to two-factor designs for moving between the designs with given margins.