Source author record

Viet Chi Tran

Viet Chi Tran appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

34works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

34 published item(s)

preprint2025arXiv

Measure estimation on a manifold explored by a diffusion process

From the observation of a diffusion path $(X_t)_{t\in [0,T]}$ on a compact connected $d$-dimensional manifold $\mathcal{M}$ without boundary, we consider the problem of estimating the stationary measure $μ$ of the process. Wang and Zhu (2023) showed that for the Wasserstein metric $\mathcal{W}_2$ and for $d\geq 5$, the convergence rate of $T^{-1/(d-2)}$ is attained by the occupation measure of the path $(X_t)_{t\in [0,T]}$ when $(X_t)_{t\in [0,T]}$ is a Langevin diffusion. We extend their result in several directions. First, we show that the rate of convergence holds for a large class of diffusion paths, whose generators are uniformly elliptic. Second, the regularity of the density $p$ of the stationary measure $μ$ with respect to the volume measure of $\mathcal{M}$ can be leveraged to obtain faster estimators: when $p$ belongs to a Sobolev space of order $\ell\geq 2$, smoothing the occupation measure by convolution with a kernel yields an estimator whose rate of convergence is of order $T^{-(\ell+1)/(2\ell+d-2)}$. We further show that this rate is the minimax rate of estimation for this problem.

preprint2023arXiv

Thick trace at infinity for the Hyperbolic Radial Spanning Tree

Since the works of Howard and Newman (2001), it is known that in straight radial rooted trees, with probability 1, infinite paths all have an asymptotic direction and each asymptotic direction is reached by (at least) an infinite path. Moreover, there exists a set of 'exceptionnal' directions reached by (at least) two infinite paths which is random, dense and only countable in dimension 2. Howard and Newman's method says nothing about (random) directions reached by more than two infinite paths and, in particular, if such 'very exceptionnal' directions exist in dimension 2. In this paper, we prove that the answer is no for the hyperbolic Radial Spanning Tree (RST): in dimension 2, this tree does not contain 3 infinite paths with the same (random) asymptotic direction with probability one. Turned in another way, this means that there is no infinite but thin subtree in the hyperbolic RST, i.e. whose infinite paths would all have the same asymptotic direction. We actually prove a stronger result in dimension $d+1$, $d\geq 1$, stating that any infinite subtree of the hyperbolic RST a.s. generates a thick trace at infinity, i.e. the set of asymptotic directions reached by its infinite paths has a positive measure.

preprint2022arXiv

Algebraic two-level measure trees

With the algebraic trees, Löhr and Winter (2021) introduced a generalization of the notion of graph-theoretic trees to account for potentially uncountable structures. The tree structure is given by the map which assigns to each triple of points their branch point. No edge length or distance is considered. One can equip a tree with a natural topology and a probability measure on the Borel-$σ$-field, defining in this way an algebraic measure tree. The main result of Löhr and Winter is to provide with the sample shape convergence a compact topology on the space of binary algebraic measure trees. This was proved by encoding the latter with triangulations of the circle. In the present paper, we extend this result to a two level setup. Motivated by the study of hierarchical systems with two levels in biology, such as host-parasite populations, we equip algebraic trees with a probability measure on the set of probability measures. To show the compactness of the space of binary algebraic two-level measure trees, we enrich the encoding of these trees by triangulations of the circle, by adding a two-level measure on the circle line. As an application, we define the two-level algebraic Kingman tree, that is the random algebraic two-level measure tree obtained from the nested Kingman coalescent.

preprint2022arXiv

Filling the gap between individual-based evolutionary models and Hamilton-Jacobi equations

We consider a stochastic model for the evolution of a discrete population structured by a trait with values on a finite grid of the torus, and with mutation and selection. Traits are vertically inherited unless a mutation occurs, and influence the birth and death rates. We focus on a parameter scaling where population is large, individual mutations are small but not rare, and the grid mesh for the trait values is much smaller than the size of mutation steps. When considering the evolution of the population in a long time scale, the contribution of small sub-populations may strongly influence the dynamics. Our main result quantifies the asymptotic dynamics of sub-population sizes on a logarithmic scale. We establish that under the parameter scaling the logarithm of the stochastic population size process, conveniently normalized, converges to the unique viscosity solution of a Hamilton-Jacobi equation. Such Hamilton-Jacobi equations have already been derived from parabolic integro-differential equations and have been widely developed in the study of adaptation of quantitative traits. Our work provides a justification of this framework directly from a stochastic individual based model, leading to a better understanding of the results obtained within this approach. The proof makes use of almost sure maximum principles and careful controls of the martingale parts.

preprint2022arXiv

Time reversal of spinal processes for linear and non-linear branching processes near stationarity

We consider a stochastic individual-based population model with competition, trait-structure affecting reproduction and survival, and changing environment. The changes of traits are described by jump processes, and the dynamics can be approximated in large population by a non-linear PDE with a non-local mutation operator. Using the fact that this PDE admits a non-trivial stationary solution, we can approximate the non-linear stochastic population process by a linear birth-death process where the interactions are frozen, as long as the population remains close to this equilibrium. This allows us to derive, when the population is large, the equation satisfied by the ancestral lineage of an individual uniformly sampled at a fixed time $T$, which is the path constituted of the traits of the ancestors of this individual in past times $t\leq T$. This process is a time inhomogeneous Markov process, but we show that the time reversal of this process possesses a very simple structure (e.g. time-homogeneous and independent of $T$). This extends recent results where the authors studied a similar model with a Laplacian operator but where the methods essentially relied on the Gaussian nature of the mutations.

preprint2020arXiv

COVID-19 pandemic control: balancing detection policy and lockdown intervention under ICU sustainability

We consider here an extended SIR model, including several features of the recent COVID-19 outbreak: in particular the infected and recovered individuals can either be detected (+) or undetected (-) and we also integrate an intensive care unit (ICU) capacity. Our model enables a tractable quantitative analysis of the optimal policy for the control of the epidemic dynamics using both lockdown and detection intervention levers. With parametric specification based on literature on COVID-19, we investigate the sensitivities of various quantities on the optimal strategies, taking into account the subtle trade-off between the sanitary and the socio-economic cost of the pandemic, together with the limited capacity level of ICU. We identify the optimal lockdown policy as an intervention structured in 4 successive phases: First a quick and strong lockdown intervention to stop the exponential growth of the contagion; second a short transition phase to reduce the prevalence of the virus; third a long period with full ICU capacity and stable virus prevalence; finally a return to normal social interactions with disappearance of the virus. The optimal scenario hereby avoids the second wave of infection, provided the lockdown is released sufficiently slowly. We also provide optimal intervention measures with increasing ICU capacity, as well as optimization over the effort on detection of infectious and immune individuals. Whenever massive resources are introduced to detect infected individuals, the pressure on social distancing can be released, whereas the impact of detection of immune individuals reveals to be more moderate.

preprint2020arXiv

Inference with selection, varying population size and evolving population structure: Application of ABC to a forward-backward coalescent process with interactions

Genetic data are often used to infer demographic history and changes or detect genes under selection. Inferential methods are commonly based on models making various strong assumptions: demography and population structures are supposed \textit{a priori} known, the evolution of the genetic composition of a population does not affect demography nor population structure, and there is no selection nor interaction between and within genetic strains. In this paper, we present a stochastic birth-death model with competitive interactions and asexual reproduction. We develop an inferential procedure for ecological, demographic and genetic parameters. We first show how genetic diversity and genealogies are related to birth and death rates, and to how individuals compete within and between strains. {This leads us to propose an original model of phylogenies, with trait structure and interactions, that allows multiple merging}. Second, we develop an Approximate Bayesian Computation framework to use our model for analyzing genetic data. We apply our procedure to simulated data from a toy model, and to real data by analyzing the genetic diversity of microsatellites on Y-chromosomes sampled from Central Asia human populations in order to test whether different social organizations show significantly different fertility.

preprint2020arXiv

Nonparametric adaptive estimation of order 1 Sobol indices in stochastic models, with an application to Epidemiology

Global sensitivity analysis is a set of methods aiming at quantifying the contribution of an uncertain input parameter of the model (or combination of parameters) on the variability of the response. We consider here the estimation of the Sobol indices of order 1 which are commonly-used indicators based on a decomposition of the output's variance. In a deterministic framework, when the same inputs always give the same outputs, these indices are usually estimated by replicated simulations of the model. In a stochastic framework, when the response given a set of input parameters is not unique due to randomness in the model, metamodels are often used to approximate the mean and dispersion of the response by deterministic functions. We propose a new non-parametric estimator without the need of defining a metamodel to estimate the Sobol indices of order 1. The estimator is based on warped wavelets and is adaptive in the regularity of the model. The convergence of the mean square error to zero, when the number of simulations of the model tend to infinity, is computed and an elbow effect is shown, depending on the regularity of the model. Applications in Epidemiology are carried to illustrate the use of non-parametric estimators.

preprint2020arXiv

Renewal in Hawkes processes with self-excitation and inhibition

This paper investigates Hawkes processes on the positive real line exhibiting both self-excitation and inhibition. Each point of this point process impacts its future intensity by the addition of a signed reproduction function. The case of a nonnegative reproduction function corresponds to self-excitation, and has been widely investigated in the literature. In particular, there exists a cluster representation of the Hawkes process which allows to apply results known for Galton-Watson trees. In the present paper, we establish limit theorems for Hawkes process with signed reproduction functions by using renewal techniques. We notably prove exponential concentration inequalities, and thus extend results of Reynaud-Bouret and Roy (2007) which were proved for nonnegative reproduction functions using this cluster representation which is no longer valid in our case. An important step for this is to establish the existence of exponential moments for renewal times of M/G/infinity queues that appear naturally in our problem. These results have their own interest, independently of the original problem for the Hawkes processes.

preprint2020arXiv

Statistical deconvolution of the free Fokker-Planck equation at fixed time

We are interested in reconstructing the initial condition of a non-linear partial differential equation (PDE), namely the Fokker-Planck equation, from the observation of a Dyson Brownian motion at a given time $t>0$. The Fokker-Planck equation describes the evolution of electrostatic repulsive particle systems, and can be seen as the large particle limit of correctly renormalized Dyson Brownian motions. The solution of the Fokker-Planck equation can be written as the free convolution of the initial condition and the semi-circular distribution. We propose a nonparametric estimator for the initial condition obtained by performing the free deconvolution via the subordination functions method. This statistical estimator is original as it involves the resolution of a fixed point equation, and a classical deconvolution by a Cauchy distribution. This is due to the fact that, in free probability, the analogue of the Fourier transform is the R-transform, related to the Cauchy transform. In past literature, there has been a focus on the estimation of the initial conditions of linear PDEs such as the heat equation, but to the best of our knowledge, this is the first time that the problem is tackled for a non-linear PDE. The convergence of the estimator is proved and the integrated mean square error is computed, providing rates of convergence similar to the ones known for non-parametric deconvolution methods. Finally, a simulation study illustrates the good performances of our estimator.

preprint2020arXiv

Statistical inference for epidemic processes in a homogeneous community (Part IV of the book Stochastic Epidemic Models and Inference)

This document is the Part IV of the book 'Stochastic Epidemic Models with Inference' edited by Tom Britton and Etienne Pardoux. It is written by Catherine Larédo, with the contribution of Viet Chi Tran for the Chapter 4. Epidemic data present challenging statistical problems, starting from the recurrent issue of handling missing information. We review methods such as MCMC, ABC or methods based on diffusion approximations. Plan of this document: 1) Observations and Asymptotic Frameworks; 2) Inference for Markov Chain Epidemic Models; 3) Inference Based on the Diffusion Approximation of Epidemic Models; 4) Inference for Continuous Time SIR models.

preprint2020arXiv

The 2d-directed spanning forest converges to the Brownian web

The two-dimensional directed spanning forest (DSF) introduced by Baccelli and Bordenave is a planar directed forest whose vertex set is given by a homogeneous Poisson point process $\mathcal{N}$ on $\mathbb{R}^2$. If the DSF has direction $-e_y$, the ancestor $h(u)$ of a vertex $u \in \mathcal{N}$ is the nearest Poisson point (in the $L_2$ distance) having strictly larger $y$-coordinate. This construction induces complex geometrical dependencies. In this paper we show that the collection of DSF paths, properly scaled, converges in distribution to the Brownian web (BW). This verifies a conjecture made by Baccelli and Bordenave in 2007.

preprint2016arXiv

Inferring $R_0$ in emerging epidemics - the effect of common population structure is small

When controlling an emerging outbreak of an infectious disease it is essential to know the key epidemiological parameters, such as the basic reproduction number $R_0$ and the control effort required to prevent a large outbreak. These parameters are estimated from the observed incidence of new cases and information about the infectious contact structures of the population in which the disease spreads. However, the relevant infectious contact structures for new, emerging infections are often unknown or hard to obtain. Here we show that for many common true underlying heterogeneous contact structures, the simplification to neglect such structures and instead assume that all contacts are made homogeneously in the whole population, results in conservative estimates for $R_0$ and the required control effort. This means that robust control policies can be planned during the early stages of an outbreak, using such conservative estimates of the required control effort.

preprint2015arXiv

A statistical network analysis of the HIV/AIDS epidemics in Cuba

The Cuban contact-tracing detection system set up in 1986 allowed the reconstruction and analysis of the sexual network underlying the epidemic (5,389 vertices and 4,073 edges, giant component of 2,386 nodes and 3,168 edges), shedding light onto the spread of HIV and the role of contact-tracing. Clustering based on modularity optimization provides a better visualization and understanding of the network, in combination with the study of covariates. The graph has a globally low but heterogeneous density, with clusters of high intraconnectivity but low interconnectivity. Though descriptive, our results pave the way for incorporating structure when studying stochastic SIR epidemics spreading on social networks.

preprint2015arXiv

Dynamic modelling of hepatitis C virus transmission among people who inject drugs: a methodological review

Equipment sharing among people who inject drugs (PWID) is a key risk factor in infection by hepatitis C virus (HCV). Both the effectiveness and cost-effectiveness of interventions aimed at reducing HCV transmission in this population (such as opioid substitution therapy, needle exchange programs or improved treatment) are difficult to evaluate using field surveys. Ethical issues and complicated access to the PWID population make it difficult to gather epidemiological data. In this context, mathematical modelling of HCV transmission is a useful alternative for comparing the cost and effectiveness of various interventions. Several models have been developed in the past few years. They are often based on strong hypotheses concerning the population structure. This review presents compartmental and individual-based models in order to underline their strengths and limits in the context of HCV infection among PWID. The final section discusses the main results of the papers.

preprint2015arXiv

Impact of a treatment as prevention strategy on hepatitis C virus transmission and on morbidity in people who inject drugs

Background: Highly effective direct-acting antiviral (DAA) regimens (90% efficacy) are becoming available for hepatitis C virus (HCV) treatment. This therapeutic revolution leads us to consider possibility of eradicating the virus. However, for this, an effective cascade of care is required. Methods: In the context of the incoming DAAs, we used a dynamic individual-based model including a model of the people who inject drugs (PWID) social network to simulate the impact of improved testing, linkage to care, and adherence to treatment, and of modified treatment recommendation on the transmission and on the morbidity of HCV in PWID in France. Results: Under the current incidence and cascade of care, with treatment initiated at fibrosis stage $\ge$F2, the HCV prevalence decreased from 42.8% to 24.9% [95% confidence interval 24.8%--24.9%] after 10 years. Changing treatment initiation criteria to treat from F0 was the only intervention leading to a substantial additional decrease in the prevalence, which fell to 11.6% [11.6%--11.7%] at 10 years. Combining this change with improved testing, linkage to care, and adherence to treatment decreased HCV prevalence to 7% [7%--7.1%] at 10 years and avoided 15.3% [14.0%-16.6%] and 29.0% [27.9%--30.1%] of cirrhosis complications over 10 and 40 years respectively. Conclusion: A high decrease in viral transmission occurs only when treatment is initiated before liver disease progresses to severe stages, suggesting that systematic treatment in PWID, where incidence remains high, would be beneficial. However, eradication will be difficult to achieve.

preprint2015arXiv

The effect of competition and horizontal trait inheritance on invasion, fixation and polymorphism

Horizontal transfer (HT) of heritable information or `traits' (carried by genetic elements, endosymbionts, or culture) is widespread among living organisms. Yet current ecological and evolutionary theory addressing HT is limited. We present a modeling framework for the dynamics of two populations that compete for resources and exchange horizontally (transfer) an otherwise vertically inherited trait. Competition influences individual demographics, affecting population size, which feeds back on the dynamics of transfer. We capture this feedback with a stochastic individual-based model, from which we derive a deterministic approximation for large populations. The interaction between horizontal transfer and competition makes possible the stable (or bi-stable) polymorphic maintenance of deleterious traits (including costly plasmids). When transfer rates are of a general density-dependent form, transfer stochasticity contributes strongly to population fluctuations. For an initially rare trait, we describe the probabilistic dynamics of invasion and fixation. Acceleration of fixation by HT is faster when competition is weak in the resident population. Thus, HT can have a major impact on the distribution of mutational effects that are fixed, and our model provides a basis for a general theory of the influence of HT on eco-evolutionary dynamics and adaptation.

preprint2013arXiv

On Computer-Intensive Simulation and Estimation Methods for Rare Event Analysis in Epidemic Models

This article focuses, in the context of epidemic models, on rare events that may possibly correspond to crisis situations from the perspective of Public Health. In general, no close analytic form for their occurrence probabilities is available and crude Monte-Carlo procedures fail. We show how recent intensive computer simulation techniques, such as interacting branching particle methods, can be used for estimation purposes, as well as for generating model paths that correspond to realizations of such events. Applications of these simulation-based methods to several epidemic models are also considered and discussed thoroughly.

preprint2012arXiv

A new proof for the convergence of an individual based model to the Trait substitution sequence

We consider a continuous time stochastic individual based model for a population structured only by an inherited vector trait and with logistic interactions. We consider its limit in a context from adaptive dynamics: the population is large, the mutations are rare and we view the process in the timescale of mutations. Using averaging techniques due to Kurtz (1992), we give a new proof of the convergence of the individual based model to the trait substitution sequence of Metz et al. (1992) first worked out by Dieckman and Law (1996) and rigorously proved by Champagnat (2006): rigging the model such that "invasion implies substitution", we obtain in the limit a process that jumps from one population equilibrium to another when mutations occur and invade the population.

preprint2012arXiv

Daphnias: from the individual based model to the large population equation

The class of deterministic 'Daphnia' models treated by Diekmann et al. (J Math Biol 61: 277-318, 2010) has a long history going back to Nisbet and Gurney (Theor Pop Biol 23: 114-135, 1983) and Diekmann et al. (Nieuw Archief voor Wiskunde 4: 82-109, 1984). In this note, we formulate the individual based models (IBM) supposedly underlying those deterministic models. The models treat the interaction between a general size-structured consumer population ('Daphnia') and an unstructured resource ('algae'). The discrete, size and age-structured Daphnia population changes through births and deaths of its individuals and throught their aging and growth. The birth and death rates depend on the sizes of the individuals and on the concentration of the algae. The latter is supposed to be a continuous variable with a deterministic dynamics that depends on the Daphnia population. In this model setting we prove that when the Daphnia population is large, the stochastic differential equation describing the IBM can be approximated by the delay equation featured in (Diekmann et al., l.c.).

preprint2012arXiv

Extinction probabilities for a distylous plant population modeled by an inhomogeneous random walk on the positive quadrant

In this paper, we study a flower population in which self-reproduction is not permitted. Individuals are diploid, {that is, each cell contains two sets of chromosomes}, and {distylous, that is, two alleles, A and a, can be found at the considered locus S}. Pollen and ovules of flowers with the same genotype at locus S cannot mate. This prevents the pollen of a given flower to fecundate its {own} stigmata. Only genotypes AA and Aa can be maintained in the population, so that the latter can be described by a random walk in the positive quadrant whose components are the number of individuals of each genotype. This random walk is not homogeneous and its transitions depend on the location of the process. We are interested in the computation of the extinction probabilities, {as} extinction happens when one of the axis is reached by the process. These extinction probabilities, which depend on the initial condition, satisfy a doubly-indexed recurrence equation that cannot be solved directly. {Our contribution is twofold : on the one hand, we obtain an explicit, though intricate, solution through the study of the PDE solved by the associated generating function. On the other hand, we provide numerical results comparing stochastic and deterministic approximations of the extinction probabilities.

preprint2012arXiv

Large graph limit for an SIR process in random network with heterogeneous connectivity

We consider an SIR epidemic model propagating on a configuration model network, where the degree distribution of the vertices is given and where the edges are randomly matched. The evolution of the epidemic is summed up into three measure-valued equations that describe the degrees of the susceptible individuals and the number of edges from an infectious or removed individual to the set of susceptibles. These three degree distributions are sufficient to describe the course of the disease. The limit in large population is investigated. As a corollary, this provides a rigorous proof of the equations obtained by Volz [Mathematical Biology 56 (2008) 293--310].

preprint2012arXiv

Limit theorems for Markov processes indexed by continuous time Galton--Watson trees

We study the evolution of a particle system whose genealogy is given by a supercritical continuous time Galton--Watson tree. The particles move independently according to a Markov process and when a branching event occurs, the offspring locations depend on the position of the mother and the number of offspring. We prove a law of large numbers for the empirical measure of individuals alive at time t. This relies on a probabilistic interpretation of its intensity by mean of an auxiliary process. The latter has the same generator as the Markov process along the branches plus additional jumps, associated with branching events of accelerated rate and biased distribution. This comes from the fact that choosing an individual uniformly at time t favors lineages with more branching events and larger offspring number. The central limit theorem is considered on a special case. Several examples are developed, including applications to splitting diffusions, cellular aging, branching Lévy processes.

preprint2012arXiv

Nonlinear historical superprocess approximations for population models with past dependence

We are interested in the evolving genealogy of a birth and death process with trait structure and ecological interactions. Traits are hereditarily transmitted from a parent to its offspring unless a mutation occurs. The dynamics may depend on the trait of the ancestors and on its past and allows interactions between individuals through their lineages. We define an interacting historical particle process describing the genealogies of the living individuals; it takes values in the space of point measures on an infinite dimensional càdlàg path space. This individual-based process can be approximated by a nonlinear historical superprocess, under the assumptions of large populations, small individuals and allometric demographies. Because of the interactions, the branching property fails and we use martingale problems and fine couplings between our population and independent branching particles. Our convergence theorem is illustrated by two examples of current interest in biology. The first one relates the biodiversity history of a population and its phylogeny, while the second treats a spatial model with competition between individuals through their past trajectories.

preprint2011arXiv

A general stochastic model for sporophytic self-incompatibility

Disentangling the processes leading populations to extinction is a major topic in ecology and conservation biology. The difficulty to find a mate in many species is one of these processes. Here, we investigate the impact of self-incompatibility in flowering plants, where several inter-compatible classes of individuals exist but individuals of the same class cannot mate. We model pollen limitation through different relationships between mate availability and fertilization success. After deriving a general stochastic model, we focus on the simple case of distylous plant species where only two classes of individuals exist. We first study the dynamics of such a species in a large population limit and then, we look for an approximation of the extinction probability in small populations. This leads us to consider inhomogeneous random walks on the positive quadrant. We compare the dynamics of distylous species to self-fertile species with and without inbreeding depression, to obtain the conditions under which self-incompatible species could be less sensitive to extinction while they can suffer more pollen limitation.

preprint2011arXiv

Level sets estimation and Vorob'ev expectation of random compact sets

The issue of a "mean shape" of a random set $X$ often arises, in particular in image analysis and pattern detection. There is no canonical definition but one possible approach is the so-called Vorob'ev expectation $\E_V(X)$, which is closely linked to quantile sets. In this paper, we propose a consistent and ready to use estimator of $\E_V(X)$ built from independent copies of $X$ with spatial discretization. The control of discretization errors is handled with a mild regularity assumption on the boundary of $X$: a not too large 'box counting' dimension. Some examples are developed and an application to cosmological data is presented.

preprint2011arXiv

New first trimester crown-rump length's equations optimized by structured data collection from a French general population

--- Objectives --- Prior to foetal karyotyping, the likelihood of Down's syndrome is often determined combining maternal age, serum free beta-HCG, PAPP-A levels and embryonic measurements of crown-rump length and nuchal translucency for gestational ages between 11 and 13 weeks. It appeared important to get a precise knowledge of these scan parameters' normal values during the first trimester. This paper focused on crown-rump length. --- METHODS --- 402 pregnancies from in-vitro fertilization allowing a precise estimation of foetal ages (FA) were used to determine the best model that describes crown-rump length (CRL) as a function of FA. Scan measures by a single operator from 3846 spontaneous pregnancies representative of the general population from Northern France were used to build a mathematical model linking FA and CRL in a context as close as possible to normal scan screening used in Down's syndrome likelihood determination. We modeled both CRL as a function of FA and FA as a function of CRL. For this, we used a clear methodology and performed regressions with heteroskedastic corrections and robust regressions. The results were compared by cross-validation to retain the equations with the best predictive power. We also studied the errors between observed and predicted values. --- Results --- Data from 513 spontaneous pregnancies allowed to model CRL as a function of age of foetal age. The best model was a polynomial of degree 2. Datation with our equation that models spontaneous pregnancies from a general population was in quite agreement with objective datations obtained from 402 IVF pregnancies and thus support the validity of our model. The most precise measure of CRL was when the SD was minimal (1.83mm), for a CRL of 23.6 mm where our model predicted a 49.4 days of foetal age. Our study allowed to model the SD from 30 to 90 days of foetal age and offers the opportunity of using Zscores in the future to detect growth abnormalities. --- Conclusion --- With powerful statistical tools we report a good modeling of the first trimester embryonic growth in the general population allowing a better knowledge of the date of fertilization useful in the ultrasound screening of Down's syndrome. The optimal period to measure CRL and predict foetal age was 49.4 days (9 weeks of gestational age). Our results open the way to the detection of foetal growth abnormalities using CRL Zscores throughout the first trimester.

preprint2011arXiv

Slow and fast scales for superprocess limits of age-structured populations

A superprocess limit for an interacting birth-death particle system modelling a population with trait and physical age-structures is established. Traits of newborn offspring are inherited from the parents except when mutations occur, while ages are set to zero. Because of interactions between individuals, standard approaches based on the Laplace transform do not hold. We use a martingale problem approach and a separation of the slow (trait) and fast (age) scales. While the trait marginals converge in a pathwise sense to a superprocess, the age distributions, on another time scale, average to equilibria that depend on traits. The convergence of the whole process depending on trait and age, only holds for finite-dimensional time-marginals. We apply our results to the study of examples illustrating different cases of trade-off between competition and senescence.

preprint2011arXiv

The 2d-Directed Spanning Forest is almost surely a tree

We consider the Directed Spanning Forest (DSF) constructed as follows: given a Poisson point process N on the plane, the ancestor of each point is the nearest vertex of N having a strictly larger abscissa. We prove that the DSF is actually a tree. Contrary to other directed forests of the literature, no Markovian process can be introduced to study the paths in our DSF. Our proof is based on a comparison argument between surface and perimeter from percolation theory. We then show that this result still holds when the points of N belonging to an auxiliary Boolean model are removed. Using these results, we prove that there is no bi-infinite paths in the DSF.

preprint2011arXiv

The Evolution of the Cuban HIV/AIDS Network

An individual detected as HIV positive in Cuba is asked to provide a list of his/her sexual contacts for the previous 2 years. This allows one to gather detailed information on the spread of the HIV epidemic. Here we study the evolution of the sexual contact graph of detected individuals and also the directed graph of HIV infections. The study covers the Cuban HIV epidemic between the years 1986 and 2004 inclusive and is motivated by an earlier study on the static properties of the network at the end of 2004. We use a variety of advanced graph algorithms to paint a picture of the growth of the epidemic, including an examination of diameters, geodesic distances, community structure and centrality amongst others characteristics. The analysis contrasts the HIV network with other real networks, and graphs generated using the configuration model. We find that the early epidemic starts in the heterosexual population and then grows mainly through MSM (Men having Sex with Men) contact. The epidemic exhibits a giant component which is shown to have degenerate chains of vertices and after 1989, diameters are larger than that expected by the equivalent configuration model graphs. In 1997 there is an significant increase in the detection rate from 73 to 256 detections/year covering mainly MSMs which results in a rapid increase of distances and diameters in the giant component.

preprint2010arXiv

Branching Feller diffusion for cell division with parasite infection

We describe the evolution of the quantity of parasites in a population of cells which divide in continuous-time. The quantity of parasites in a cell follows a Feller diffusion, which is splitted randomly between the two daughter cells when a division occurs. The cell division rate may depend on the quantity of parasites inside the cell and we are interested in the cases of constant or monotone division rate. We first determine the asymptotic behavior of the quantity of parasites in a cell line, which follows a Feller diffusion with multiplicative jumps. We then consider the evolution of the infection of the cell population and give criteria to determine whether the proportion of infected cells goes to zero (recovery) or if a positive proportion of cells becomes largely infected (proliferation of parasites inside the cells).

preprint2010arXiv

HIV with contact-tracing: a case study in Approximate Bayesian Computation

Missing data is a recurrent issue in epidemiology where the infection process may be partially observed. Approximate Bayesian Computation, an alternative to data imputation methods such as Markov Chain Monte Carlo integration, is proposed for making inference in epidemiological models. It is a likelihood-free method that relies exclusively on numerical simulations. ABC consists in computing a distance between simulated and observed summary statistics and weighting the simulations according to this distance. We propose an original extension of ABC to path-valued summary statistics, corresponding to the cumulated number of detections as a function of time. For a standard compartmental model with Suceptible, Infectious and Recovered individuals (SIR), we show that the posterior distributions obtained with ABC and MCMC are similar. In a refined SIR model well-suited to the HIV contact-tracing data in Cuba, we perform a comparison between ABC with full and binned detection times. For the Cuban data, we evaluate the efficiency of the detection system and predict the evolution of the HIV-AIDS disease. In particular, the percentage of undetected infectious individuals is found to be of the order of 40%.