Source author record

Sylvie Huet

Sylvie Huet appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

10works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

10 published item(s)

preprint2014arXiv

Nonparametric species richness estimation under convexity constraint

We consider the estimation of the total number $N$ of species based on the abundances of species that have been observed. We adopt a non parametric approach where the true abundance distribution $p$ is only supposed to be convex. From this assumption, we propose a definition for convex abundance distributions. We use a least-squares estimate of the truncated version of $p$ under the convexity constraint. We deduce two estimators of the total number of species, the asymptotic distribution of which are derived. We propose three different procedures, including a bootstrap one, to obtain a confidence interval for $N$. The performances of the estimators are assessed in a simulation study and compared with competitors. The proposed method is illustrated on several examples.

preprint2014arXiv

Openness leads to opinion stability and narrowness to volatility

We propose a new opinion dynamic model based on the experiments and results of Wood et al (1996). We consider pairs of individuals discussing on two attitudinal dimensions, and we suppose that one dimension is important, the other secondary. The dynamics are mainly ruled by the level of agreement on the main dimension. If two individuals are close on the main dimension, then they attract each other on the main and on the secondary dimensions, whatever their disagreement on the secondary dimension. If they are far from each other on the main dimension, then too much proximity on the secondary dimension is uncomfortable, and generates rejection on this dimension. The proximity is defined by comparing the opinion distance with a threshold called attraction threshold on the main dimension and rejection threshold on the secondary dimension. With such dynamics, a population with opinions initially uniformly drawn evolves to a set of clusters, inside which secondary opinions fluctuate more or less depending on threshold values. We observe that a low attraction threshold favours fluctuations on the secondary dimension, especially when the rejection threshold is high. The opinion evolutions of the model can be related to some stylised facts.

preprint2014arXiv

Rejection Mechanism in 2D Bounded Confidence Provides more Conformity

We add a rejection mechanism (negative influence) into a two-dimensions bounded confidence model. The principle is that one shifts aways from a close attitude of one's interlocutor, when there is a strong disagreement on the other attitude. The model shows metastable clusters, which maintain themselves through opposite influences of competitor clusters. Our analysis and first experiments support the hypothesis that the number of clusters grows linearly with the inverse of the uncertainty, whereas this growth is quadratic in the bounded confidence model.

preprint2012arXiv

Estimation of a convex discrete distribution

Non-parametric estimation of a convex discrete distribution may be of interest in several applications, such as the estimation of species abundance distribution in ecology. In this paper we study the least squares estimator of a discrete distribution under the constraint of convexity. We show that this estimator exists and is unique, and that it always outperforms the classical empirical estimator in terms of the $\ell_{2}$-distance. We provide an algorithm for its computation, based on the support reduction algorithm. We compare its performance to those of the empirical estimator, on the basis of a simulation study.

preprint2012arXiv

Graph selection with GGMselect

Applications on inference of biological networks have raised a strong interest in the problem of graph estimation in high-dimensional Gaussian graphical models. To handle this problem, we propose a two-stage procedure which first builds a family of candidate graphs from the data, and then selects one graph among this family according to a dedicated criterion. This estimation procedure is shown to be consistent in a high-dimensional setting, and its risk is controlled by a non-asymptotic oracle-like inequality. The procedure is tested on a real data set concerning gene expression data, and its performances are assessed on the basis of a large numerical study. The procedure is implemented in the R-package GGMselect available on the CRAN.

preprint2012arXiv

High-dimensional regression with unknown variance

We review recent results for high-dimensional sparse linear regression in the practical case of unknown variance. Different sparsity settings are covered, including coordinate-sparsity, group-sparsity and variation-sparsity. The emphasis is put on non-asymptotic analyses and feasible procedures. In addition, a small numerical study compares the practical performance of three schemes for tuning the Lasso estimator and some references are collected for some more general models, including multivariate regression and nonparametric regression.

preprint2012arXiv

The Leviathan model: Absolute dominance, generalised distrust, small worlds and other patterns emerging from combining vanity with opinion propagation

We propose an opinion dynamics model that combines processes of vanity and opinion propagation. The interactions take place between randomly chosen pairs. During an interaction, the agents propagate their opinions about themselves and about other people they know. Moreover, each individual is subject to vanity: if her interlocutor seems to value her highly, then she increases her opinion about this interlocutor. On the contrary she tends to decrease her opinion about those who seem to undervalue her. The combination of these dynamics with the hypothesis that the opinion propagation is more efficient when coming from highly valued individuals, leads to different patterns when varying the parameters. For instance, for some parameters the positive opinion links between individuals generate a small world network. In one of the patterns, absolute dominance of one agent alternates with a state of generalised distrust, where all agents have a very low opinion of all the others (including themselves). We provide some explanations of the mechanisms behind these emergent behaviors and finally propose a discussion about their interest

preprint2011arXiv

A commuting network model: going to the bulk

The influence of commuting in socio-economic dynamics increases constantly. Analysing and modelling the networks formed by commuters to help decision-making regarding the land-use has become crucial. This paper presents a simple spatial interaction simulated model with only one parameter. The proposed algorithm considers each individual who wants to commute, starting from their living place to all their workplaces. It decides where the location of the workplace following the classical rule inspired from the gravity law consisting in a compromise between the job offers and the distance to the jobs. The further away the job offer is, the more important it must be in order to be considered. Inversely, only the quantity of offers is important for the decision when these offers are close. The paper also presents a comparative analysis of the structure of the commuting networks of the four European regions to which we apply our model. The model is calibrated and validated on these regions. Results from the analysis shows that the model is very efficient in reproducing most of the statistical properties of the network given by the data sources.

preprint2011arXiv

Estimator selection in the Gaussian setting

We consider the problem of estimating the mean $f$ of a Gaussian vector $Y$ with independent components of common unknown variance $σ^{2}$. Our estimation procedure is based on estimator selection. More precisely, we start with an arbitrary and possibly infinite collection $\FF$ of estimators of $f$ based on $Y$ and, with the same data $Y$, aim at selecting an estimator among $\FF$ with the smallest Euclidean risk. No assumptions on the estimators are made and their dependencies with respect to $Y$ may be unknown. We establish a non-asymptotic risk bound for the selected estimator. As particular cases, our approach allows to handle the problems of aggregation and model selection as well as those of choosing a window and a kernel for estimating a regression function, or tuning the parameter involved in a penalized criterion. We also derive oracle-type inequalities when $\FF$ consists of linear estimators. For illustration, we carry out two simulation studies. One aims at comparing our procedure to cross-validation for choosing a tuning parameter. The other shows how to implement our approach to solve the problem of variable selection in practice.

preprint2009arXiv

An iterative approach for generating statistically realistic populations of households

Background: Many different simulation frameworks, in different topics, need to treat realistic datasets to initialize and calibrate the system. A precise reproduction of initial states is extremely important to obtain reliable forecast from the model. Methodology/Principal Findings: This paper proposes an algorithm to create an artificial population where individuals are described by their age, and are gathered in households respecting a variety of statistical constraints (distribution of household types, sizes, age of household head, difference of age between partners and among parents and children). Such a population is often the initial state of microsimulation or (agent) individual-based models. To get a realistic distribution of households is often very important, because this distribution has an impact on the demographic evolution. Usual techniques from microsimulation approach cross different sources of aggregated data for generating individuals. In our case the number of combinations of different households (types, sizes, age of participants) makes it computationally difficult to use directly such methods. Hence we developed a specific algorithm to make the problem more easily tractable. Conclusions/Significance: We generate the populations of two pilot municipalities in Auvergne region (France), to illustrate the approach. The generated populations show a good agreement with the available statistical datasets (not used for the generation) and are obtained in a reasonable computational time.