Source author record

Ramon Grima

Ramon Grima appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

26works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

26 published item(s)

preprint2022arXiv

Bayesian learning of effective chemical master equations in crowded intracellular conditions

Biochemical reactions inside living cells often occur in the presence of crowders -- molecules that do not participate in the reactions but influence the reaction rates through excluded volume effects. However the standard approach to modelling stochastic intracellular reaction kinetics is based on the chemical master equation (CME) whose propensities are derived assuming no crowding effects. Here, we propose a machine learning strategy based on Bayesian Optimisation utilising synthetic data obtained from spatial cellular automata (CA) simulations (that explicitly model volume-exclusion effects) to learn effective propensity functions for CMEs. The predictions from a small CA training data set can then be extended to the whole range of parameter space describing physiologically relevant levels of crowding by means of Gaussian Process regression. We demonstrate the method on an enzyme-catalyzed reaction and a genetic feedback loop, showing good agreement between the time-dependent distributions of molecule numbers predicted by the effective CME and CA simulations.

preprint2020arXiv

Exact solution of stochastic gene expression models with bursting, cell cycle and replication dynamics

The bulk of stochastic gene expression models in the literature do not have an explicit description of the age of a cell within a generation and hence they cannot capture events such as cell division and DNA replication. Instead, many models incorporate cell cycle implicitly by assuming that dilution due to cell division can be described by an effective decay reaction with first-order kinetics. If it is further assumed that protein production occurs in bursts then the stationary protein distribution is a negative binomial. Here we seek to understand how accurate these implicit models are when compared with more detailed models of stochastic gene expression. We derive the exact stationary solution of the chemical master equation describing bursty protein dynamics, binomial partitioning at mitosis, age-dependent transcription dynamics including replication, and random interdivision times sampled from Erlang or more general distributions; the solution is different for single lineage and population snapshot settings. We show that protein distributions are well approximated by the solution of implicit models (a negative binomial) when the mean number of mRNAs produced per cycle is low and the cell cycle length variability is large. When these conditions are not met, the distributions are either almost bimodal or else display very flat regions near the mode and cannot be described by implicit models. We also show that for genes with low transcription rates, the size of protein noise has a strong dependence on the replication time, it is almost independent of cell cycle variability for lineage measurements and increases with cell cycle variability for population snapshot measurements. In contrast for large transcription rates, the size of protein noise is independent of replication time and increases with cell cycle variability for both lineage and population measurements.

preprint2020arXiv

Small protein number effects in stochastic models of autoregulated bursty gene expression

A stochastic model of autoregulated bursty gene expression by Kumar et al. [Phys. Rev. Lett. 113, 268105 (2014)] has been exactly solved in steady-state conditions under the implicit assumption that protein numbers are sufficiently large such that fluctuations in protein numbers due to reversible protein-promoter binding can be ignored. Here we derive an alternative model that takes into account these fluctuations and hence can be used to study low protein number effects. The exact steady-state protein number distributions is derived as a sum of Gaussian hypergeometric functions. We use the theory to study how promoter switching rates and the type of feedback influence the size of protein noise and noise-induced bistability. Furthermore we show that our model predictions for the protein number distribution are significantly different from those of Kumar et al. when the protein mean is small, gene switching is fast, and protein binding is faster than unbinding.

preprint2020arXiv

Steady-state fluctuations of a genetic feedback loop with fluctuating rate parameters using the unified colored noise approximation

A common model of stochastic auto-regulatory gene expression describes promoter switching via cooperative protein binding, effective protein production in the active state and dilution of proteins. Here we consider an extension of this model whereby colored noise with a short correlation time is added to the reaction rate parameters -- we show that when the size and timescale of the noise is appropriately chosen it accounts for fast reactions that are not explicitly modelled, e.g., in models with no mRNA description, fluctuations in the protein production rate can account for rapid multiple stages of nuclear mRNA processing which precede translation in eukaryotes. We show how the unified colored noise approximation can be used to derive expressions for the protein number distribution that is in good agreement with stochastic simulations. We find that even when the noise in the rate parameters is small, the protein distributions predicted by our model can be significantly different than models assuming constant reaction rates.

preprint2019arXiv

Parameter estimation for biochemical reaction networks using Wasserstein distances

We present a method for estimating parameters in stochastic models of biochemical reaction networks by fitting steady-state distributions using Wasserstein distances. We simulate a reaction network at different parameter settings and train a Gaussian process to learn the Wasserstein distance between observations and the simulator output for all parameters. We then use Bayesian optimization to find parameters minimizing this distance based on the trained Gaussian process. The effectiveness of our method is demonstrated on the three-stage model of gene expression and a genetic feedback loop for which moment-based methods are known to perform poorly. Our method is applicable to any simulator model of stochastic reaction networks, including Brownian Dynamics.

preprint2019arXiv

Stochastic modeling of auto-regulatory genetic feedback loops: a review and comparative study

Auto-regulatory feedback loops are one of the most common network motifs. A wide variety of stochastic models have been constructed to understand how the fluctuations in protein numbers in these loops are influenced by the kinetic parameters of the main biochemical steps. These models differ according to (i) which sub-cellular processes are explicitly modelled; (ii) the modelling methodology employed (discrete, continuous or hybrid); (iii) whether they can be analytically solved for the steady-state distribution of protein numbers. We discuss the assumptions and properties of the main models in the literature, summarize our current understanding of the relationship between them and highlight some of the insights gained through modelling.

preprint2016arXiv

An exact solution to Brownian dynamics of a reversible bimolecular reaction in one dimension

Brownian dynamics is a popular fine-grained method for simulating systems of interacting particles, such as chemical reactions. Though the method is simple to simulate, it is generally assumed that the dynamics is impossible to solve exactly and analytically, aside from some trivial systems. We here give the first exact analytical solution to a non-trivial Brownian dynamics system: the reaction $A+B\xrightleftharpoons[]{}C$ in equilibrium in one-dimensional periodic space. The solution is a function of the particles' diffusion coefficients, radii, length of space and unbinding distance.

preprint2016arXiv

Cox process representation and inference for stochastic reaction-diffusion processes

Complex behaviour in many systems arises from the stochastic interactions of spatially distributed particles or agents. Stochastic reaction-diffusion processes are widely used to model such behaviour in disciplines ranging from biology to the social sciences, yet they are notoriously difficult to simulate and calibrate to observational data. Here we use ideas from statistical physics and machine learning to provide a solution to the inverse problem of learning a stochastic reaction-diffusion process from data. Our solution relies on a non-trivial connection between stochastic reaction-diffusion processes and spatio-temporal Cox processes, a well-studied class of models from computational statistics. This connection leads to an efficient and flexible algorithm for parameter inference and model selection. Our approach shows excellent accuracy on numeric and real data examples from systems biology and epidemiology. Our work provides both insights into spatio-temporal stochastic systems, and a practical solution to a long-standing problem in computational modelling.

preprint2016arXiv

Molecular finite-size effects in stochastic models of equilibrium chemical systems

The reaction-diffusion master equation (RDME) is a standard modelling approach for understanding stochastic and spatial chemical kinetics. An inherent assumption is that molecules are point-like. Here we introduce the crowded reaction-diffusion master equation (cRDME) which takes into account volume exclusion effects on stochastic kinetics due to a finite molecular radius. We obtain an exact closed form solution of the RDME and of the cRDME for a general chemical system in equilibrium conditions. The difference between the two solutions increases with the ratio of molecular diameter to the compartment length scale. We show that an increase in molecular crowding can (i) lead to deviations from the classical inverse square root law for the noise-strength; (ii) flip the skewness of the probability distribution from right to left-skewed; (iii) shift the equilibrium of bimolecular reactions so that more product molecules are formed; (iv) strongly modulate the Fano factors and coefficients of variation. These crowding-induced effects are found to be particularly pronounced for chemical species not involved in chemical conservation laws.Finally we show that statistics obtained using the vRDME are in good agreement with those obtained from Brownian dynamics with excluded volume interactions.

preprint2016arXiv

The breakdown of the reaction-diffusion master equation with non-elementary rates

The chemical master equation (CME) is the exact mathematical formulation of chemical reactions occurring in a dilute and well-mixed volume. The reaction-diffusion master equation (RDME) is a stochastic description of reaction-diffusion processes on a spatial lattice, assuming well-mixing only on the length scale of the lattice. It is clear that, for the sake of consistency, the solution of the RDME of a chemical system should converge to the solution of the CME of the same system in the limit of fast diffusion: indeed, this has been tacitly assumed in most literature concerning the RDME. We show that, in the limit of fast diffusion, the RDME indeed converges to a master equation, but not necessarily the CME. We introduce a class of propensity functions, such that if the RDME has propensities exclusively of this class then the RDME converges to the CME of the same system; while if the RDME has propensities not in this class then convergence is not guaranteed. These are revealed to be elementary and non-elementary propensities respectively. We also show that independent of the type of propensity, the RDME converges to the CME in the simultaneous limit of fast diffusion and large volumes. We illustrate our results with some simple example systems, and argue that the RDME cannot be an accurate description of systems with non-elementary rates.

preprint2015arXiv

Approximate probability distributions of the master equation

Master equations are common descriptions of mesoscopic systems. Analytical solutions to these equations can rarely be obtained. We here derive an analytical approximation of the time-dependent probability distribution of the master equation using orthogonal polynomials. The solution is given in two alternative formulations: a series with continuous and a series with discrete support both of which can be systematically truncated. While both approximations satisfy the system size expansion of the master equation, the continuous distribution approximations become increasingly negative and tend to oscillations with increasing truncation order. In contrast, the discrete approximations rapidly converge to the underlying non-Gaussian distributions. The theory is shown to lead to particularly simple analytical expressions for the probability distributions of molecule numbers in metabolic reactions and gene expression systems.

preprint2015arXiv

Comparison of different moment-closure approximations for stochastic chemical kinetics

In recent years moment-closure approximations (MA) of the chemical master equation have become a popular method for the study of stochastic effects in chemical reaction systems. Several different MA methods have been proposed and applied in the literature, but it remains unclear how they perform with respect to each other. In this paper we study the normal, Poisson, log-normal and central-moment-neglect MAs by applying them to understand the stochastic properties of chemical systems whose deterministic rate equations show the properties of bistability, ultrasensitivity and oscillatory behaviour. Our results suggest that the normal MA is favourable over the other studied MAs. In particular we found that (i) the size of the region of parameter space where a closure gives physically meaningful results, e.g. positive mean and variance, is considerably larger for the normal closure than for the other three closures; (ii) the accuracy of the predictions of the four closures (relative to simulations using the stochastic simulation algorithm) is comparable in those regions of parameter space where all closures give physically meaningful results; (iii) the Poisson and log-normal MAs are not uniquely defined for systems involving conservation laws in molecule numbers. We also describe the new software package MOCA which enables the automated numerical analysis of various MA methods in a graphical user interface and which was used to perform the comparative analysis presented in this paper. MOCA allows the user to develop novel closure methods and can treat polynomial, non-polynomial, as well as time-dependent propensity functions, thus being applicable to virtually any chemical reaction system.

preprint2015arXiv

Distribution approximations for the chemical master equation: comparison of the method of moments and the system size expansion

The stochastic nature of chemical reactions involving randomly fluctuating population sizes has lead to a growing research interest in discrete-state stochastic models and their analysis. A widely-used approach is the description of the temporal evolution of the system in terms of a chemical master equation (CME). In this paper we study two approaches for approximating the underlying probability distributions of the CME. The first approach is based on an integration of the statistical moments and the reconstruction of the distribution based on the maximum entropy principle. The second approach relies on an analytical approximation of the probability distribution of the CME using the system size expansion, considering higher-order terms than the linear noise approximation. We consider gene expression networks with unimodal and multimodal protein distributions to compare the accuracy of the two approaches. We find that both methods provide accurate approximations to the distributions of the CME while having different benefits and limitations in applications.

preprint2015arXiv

Model reduction for stochastic chemical systems with abundant species

Biochemical processes typically involve many chemical species, some in abundance and some in low molecule numbers. Here we first identify the rate constant limits under which the concentrations of a given set of species will tend to infinity (the abundant species) while the concentrations of all other species remains constant (the non-abundant species). Subsequently we prove that in this limit, the fluctuations in the molecule numbers of non-abundant species are accurately described by a hybrid stochastic description consisting of a chemical master equation coupled to deterministic rate equations. This is a reduced description when compared to the conventional chemical master equation which describes the fluctuations in both abundant and non-abundant species. We show that the reduced master equation can be solved exactly for a number of biochemical networks involving gene expression and enzyme catalysis, whose conventional chemical master equation description is analytically impenetrable. We use the linear noise approximation to obtain approximate expressions for the difference between the variance of fluctuations in the non-abundant species as predicted by the hybrid approach and by the conventional chemical master equation. Furthermore we show that surprisingly, irrespective of any separation in the mean molecule numbers of various species, the conventional and hybrid master equations exactly agree for a class of chemical systems.

preprint2015arXiv

Stochastic Simulation of Biomolecular Networks in Dynamic Environments

Simulation of biomolecular networks is now indispensable for studying biological systems, from small reaction networks to large ensembles of cells. Here we present a novel approach for stochastic simulation of networks embedded in the dynamic environment of the cell and its surroundings. We thus sample trajectories of the stochastic process described by the chemical master equation with time-varying propensities. A comparative analysis shows that existing approaches can either fail dramatically, or else can impose impractical computational burdens due to numerical integration of reaction propensities, especially when cell ensembles are studied. Here we introduce the Extrande method which, given a simulated time course of dynamic network inputs, provides a conditionally exact and several orders-of-magnitude faster simulation solution. The new approach makes it feasible to demonstrate, using decision-making by a large population of quorum sensing bacteria, that robustness to fluctuations from upstream signaling places strong constraints on the design of networks determining cell fate. Our approach has the potential to significantly advance both understanding of molecular systems biology and design of synthetic circuits.

preprint2015arXiv

The linear-noise approximation and the chemical master equation exactly agree up to second-order moments for a class of chemical systems

It is well known that the linear-noise approximation (LNA) exactly agrees with the chemical master equation, up to second-order moments, for chemical systems composed of zero and first-order reactions. Here we show that this is also a property of the LNA for a subset of chemical systems with second-order reactions. This agreement is independent of the number of interacting molecules.

preprint2015arXiv

Validity conditions for moment closure approximations in stochastic chemical kinetics

Approximations based on moment-closure (MA) are commonly used to obtain estimates of the mean molecule numbers and of the variance of fluctuations in the number of molecules of chemical systems. The advantage of this approach is that it can be far less computationally expensive than exact stochastic simulations of the chemical master equation. Here we numerically study the conditions under which the MA equations yield results reflecting the true stochastic dynamics of the system. We show that for bistable and oscillatory chemical systems with deterministic initial conditions, the solution of the MA equations can be interpreted as a valid approximation to the true moments of the CME, only when the steady-state mean molecule numbers obtained from the chemical master equation fall within a certain finite range. The same validity criterion for monostable systems implies that the steady-state mean molecule numbers obtained from the chemical master equation must be above a certain threshold. For mean molecule numbers outside of this range of validity, the MA equations lead to either qualitatively wrong oscillatory dynamics or to unphysical predictions such as negative variances in the molecule numbers or multiple steady-state moments of the stationary distribution as the initial conditions are varied. Our results clarify the range of validity of the MA approach and show that pitfalls in the interpretation of the results can only be overcome through the systematic comparison of the solutions of the MA equations of a certain order with those of higher orders.

preprint2014arXiv

Noise-induced multistability in chemical systems: Discrete vs Continuum modeling

The noisy dynamics of chemical systems is commonly studied using either the chemical master equation (CME) or the chemical Fokker-Planck equation (CFPE). The latter is a continuum approximation of the discrete CME approach. We here show that the CFPE may fail to capture the CME's prediction of noise-induced multistability. In particular we find a simple chemical system for which the CME's marginal probability distribution changes from unimodal to multimodal as the system-size decreases below a critical value, while the CFPE's marginal probability distribution is unimodal for all physically meaningful system sizes.

preprint2014arXiv

System size expansion using Feynman rules and diagrams

Few analytical methods exist for quantitative studies of large fluctuations in stochastic systems. In this article, we develop a simple diagrammatic approach to the Chemical Master Equation that allows us to calculate multi-time correlation functions which are accurate to a any desired order in van Kampen's system size expansion. Specifically, we present a set of Feynman rules from which this diagrammatic perturbation expansion can be constructed algorithmically. We then apply the methodology to derive in closed form the leading order corrections to the linear noise approximation of the intrinsic noise power spectrum for general biochemical reaction networks. Finally, we illustrate our results by describing noise-induced oscillations in the Brusselator reaction scheme which are not captured by the common linear noise approximation.

preprint2014arXiv

The complex chemical Langevin equation

The chemical Langevin equation (CLE) is a popular simulation method to probe the stochastic dynamics of chemical systems. The CLE's main disadvantage is its break down in finite time due to the problem of evaluating square roots of negative quantities whenever the molecule numbers become sufficiently small. We show that this issue is not a numerical integration problem, rather in many systems it is intrinsic to all representations of the CLE. Various methods of correcting the CLE have been proposed which avoid its break down. We show that these methods introduce undesirable artefacts in the CLE's predictions. In particular, for unimolecular systems, these correction methods lead to CLE predictions for the mean concentrations and variance of fluctuations which disagree with those of the chemical master equation. We show that, by extending the domain of the CLE to complex space, break down is eliminated, and the CLE's accuracy for unimolecular systems is restored. Although the molecule numbers are generally complex, we show that the "complex CLE" predicts real-valued quantities for the mean concentrations, the moments of intrinsic noise, power spectra and first passage times, hence admitting a physical interpretation. It is also shown to provide a more accurate approximation of the chemical master equation of simple biochemical circuits involving bimolecular reactions than the various corrected forms of the real-valued CLE, the linear-noise approximation and a commonly used two moment-closure approximation.

preprint2013arXiv

Computation of biochemical pathway fluctuations beyond the linear noise approximation using iNA

The linear noise approximation is commonly used to obtain intrinsic noise statistics for biochemical networks. These estimates are accurate for networks with large numbers of molecules. However it is well known that many biochemical networks are characterized by at least one species with a small number of molecules. We here describe version 0.3 of the software intrinsic Noise Analyzer (iNA) which allows for accurate computation of noise statistics over wide ranges of molecule numbers. This is achieved by calculating the next order corrections to the linear noise approximation's estimates of variance and covariance of concentration fluctuations. The efficiency of the methods is significantly improved by automated just-in-time compilation using the LLVM framework leading to a fluctuation analysis which typically outperforms that obtained by means of exact stochastic simulations. iNA is hence particularly well suited for the needs of the computational biology community.

preprint2013arXiv

Single molecule enzymology a la Michaelis-Menten

In the past one hundred years, deterministic rate equations have been successfully used to infer enzyme-catalysed reaction mechanisms and to estimate rate constants from reaction kinetics experiments conducted in vitro. In recent years, sophisticated experimental techniques have been developed that allow the measurement of enzyme- catalysed and other biopolymer-mediated reactions inside single cells at the single molecule level. Time course data obtained by these methods are considerably noisy because molecule numbers within cells are typically quite small. As a consequence, the interpretation and analysis of single cell data requires stochastic methods, rather than deterministic rate equations. Here we concisely review both experimental and theoretical techniques which enable single molecule analysis with particular emphasis on the major developments in the field of theoretical stochastic enzyme kinetics, from its inception in the mid-twentieth century to its modern day status. We discuss the differences between stochastic and deterministic rate equation models, how these depend on enzyme molecule numbers and substrate inflow into the reaction compartment and how estimation of rate constants from single cell data is possible using recently developed stochastic approaches.

preprint2012arXiv

Rigorous elimination of fast stochastic variables from the linear noise approximation using projection operators

The linear noise approximation (LNA) offers a simple means by which one can study intrinsic noise in monostable biochemical networks. Using simple physical arguments, we have recently introduced the slow-scale LNA (ssLNA) which is a reduced version of the LNA under conditions of timescale separation. In this paper, we present the first rigorous derivation of the ssLNA using the projection operator technique and show that the ssLNA follows uniquely from the standard LNA under the same conditions of timescale separation as those required for the deterministic quasi-steady state approximation. We also show that the large molecule number limit of several common stochastic model reduction techniques under timescale separation conditions constitutes a special case of the ssLNA.

preprint2011arXiv

How accurate are the non-linear chemical Fokker-Planck and chemical Langevin equations?

The chemical Fokker-Planck equation and the corresponding chemical Langevin equation are commonly used approximations of the chemical master equation. These equations are derived from an uncontrolled, second-order truncation of the Kramers-Moyal expansion of the chemical master equation and hence their accuracy remains to be clarified. We use the system-size expansion to show that chemical Fokker-Planck estimates of the mean concentrations and of the variance of the concentration fluctuations about the mean are accurate to order $Ω^{-3/2}$ for reaction systems which do not obey detailed balance and at least accurate to order $Ω^{-2}$ for systems obeying detailed balance, where $Ω$ is the characteristic size of the system. Hence the chemical Fokker-Planck equation turns out to be more accurate than the linear-noise approximation of the chemical master equation (the linear Fokker-Planck equation) which leads to mean concentration estimates accurate to order $Ω^{-1/2}$ and variance estimates accurate to order $Ω^{-3/2}$. This higher accuracy is particularly conspicuous for chemical systems realized in small volumes such as biochemical reactions inside cells. A formula is also obtained for the approximate size of the relative errors in the concentration and variance predictions of the chemical Fokker-Planck equation, where the relative error is defined as the difference between the predictions of the chemical Fokker-Planck equation and the master equation divided by the prediction of the master equation. For dimerization and enzyme-catalyzed reactions, the errors are typically less than few percent even when the steady-state is characterized by merely few tens of molecules.

preprint2011arXiv

Limitations of the stochastic quasi-steady-state approximation in open biochemical reaction networks

The application of the quasi-steady-state approximation to the Michaelis-Menten reaction embedded in large open chemical reaction networks is a popular model reduction technique in deterministic and stochastic simulations of biochemical reactions inside cells. It is frequently assumed that the predictions of the reduced master equations obtained using the stochastic quasi-steady-state approach are in very good agreement with the predictions of the full master equations, provided the conditions for the validity of the deterministic quasi-steady-state approximation are fulfilled. We here use the linear-noise approximation to show that this assumption is not generally justified for the Michaelis-Menten reaction with substrate input, the simplest example of an open embedded enzyme reaction. The reduced master equation approach is found to considerably overestimate the size of intrinsic noise at low copy numbers of molecules. A simple formula is obtained for the relative error between the predictions of the reduced and full master equations for the variance of the substrate concentration fluctuations. The maximum error is reached when modeling moderately or highly efficient enzymes, in which case the error is approximately 30%. The theoretical predictions are validated by stochastic simulations using experimental parameter values for enzymes involved in proteolysis, gluconeogenesis and fermentation.

preprint2011arXiv

Stochastic theory of large-scale enzyme-reaction networks: Finite copy number corrections to rate equation models

Chemical reactions inside cells occur in compartment volumes in the range of atto- to femtolitres. Physiological concentrations realized in such small volumes imply low copy numbers of interacting molecules with the consequence of considerable fluctuations in the concentrations. In contrast, rate equation models are based on the implicit assumption of infinitely large numbers of interacting molecules, or equivalently, that reactions occur in infinite volumes at constant macroscopic concentrations. In this article we compute the finite-volume corrections (or equivalently the finite copy number corrections) to the solutions of the rate equations for chemical reaction networks composed of arbitrarily large numbers of enzyme-catalyzed reactions which are confined inside a small sub-cellular compartment. This is achieved by applying a mesoscopic version of the quasi-steady state assumption to the exact Fokker-Planck equation associated with the Poisson Representation of the chemical master equation. The procedure yields impressively simple and compact expressions for the finite-volume corrections. We prove that the predictions of the rate equations will always underestimate the actual steady-state substrate concentrations for an enzyme-reaction network confined in a small volume. In particular we show that the finite-volume corrections increase with decreasing sub-cellular volume, decreasing Michaelis-Menten constants and increasing enzyme saturation. The magnitude of the corrections depends sensitively on the topology of the network. The predictions of the theory are shown to be in excellent agreement with stochastic simulations for two types of networks typically associated with protein methylation and metabolism.