Source author record

Jayanth R. Banavar

Jayanth R. Banavar appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

21works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

21 published item(s)

preprint2021arXiv

Local sequence-structure relationships in proteins

We seek to understand the interplay between amino acid sequence and local structure in proteins. Are some amino acids unique in their ability to fit harmoniously into certain local structures? What is the role of sequence in sculpting the putative native state folds from myriad possible conformations? In order to address these questions, we represent the local structure of each C-alpha atom of a protein by just two angles, theta and mu, and we analyze a set of more than 4000 protein structures from the PDB. We use a hierarchical clustering scheme to divide the 20 amino acids into six distinct groups based on their similarity to each other in fitting local structural space. We present the results of a detailed analysis of patterns of amino acid specificity in adopting local structural conformations and show that the sequence-structure correlation is not very strong compared to a random assignment of sequence to structure. Yet, our analysis may be useful to determine an effective scoring rubric for quantifying the match of an amino acid to its putative local structure.

preprint2021arXiv

Proteins -- a celebration of consilience

Proteins are the common constituents of all living cells. They are molecular machines that interact with each other as well as with other cell products and carry out a dizzying array of functions with distinction. These interactions follow from their native state structures and therefore understanding sequence-structure relationships is of fundamental importance. What is quite remarkable about proteins is that their understanding necessarily straddles several disciplines. The importance of geometry in defining protein native state structure, the constraints placed on protein behavior by mathematics and physics, the need for proteins to obey the laws of quantum chemistry, and the rich role of evolution and biology all come together in defining protein science. Here we review ideas from the literature and present an interdisciplinary framework that aims to marry ideas from Plato and Darwin and demonstrates an astonishing consilience between disciplines in describing proteins. We discuss the consequences of this framework on protein behavior.

preprint2020arXiv

True scale-free networks hidden by finite size effects

We analyze about two hundred naturally occurring networks with distinct dynamical origins to formally test whether the commonly assumed hypothesis of an underlying scale-free structure is generally viable. This has recently been questioned on the basis of statistical testing of the validity of power law distributions of network degrees by contrasting real data. Specifically, we analyze by finite-size scaling analysis the datasets of real networks to check whether purported departures from the power law behavior are due to the finiteness of the sample size. In this case, power laws would be recovered in the case of progressively larger cutoffs induced by the size of the sample. We find that a large number of the networks studied follow a finite size scaling hypothesis without any self-tuning. This is the case of biological protein interaction networks, technological computer and hyperlink networks, and informational networks in general. Marked deviations appear in other cases, especially infrastructure and transportation but also social networks. We conclude that underlying scale invariance properties of many naturally occurring networks are extant features often clouded by finite-size effects due to the nature of the sample data.

preprint2016arXiv

The geometry of coexistence in large ecosystems

The role of species interactions in controlling the interplay between the stability of an ecosystem and its biodiversity is still not well understood. The ability of ecological communities to recover after a small perturbation of the species abundances (local asymptotic stability) has been well studied, whereas the likelihood of a community to persist when the interactions are altered (structural stability) has received much less attention. Our goal is to understand the effects of diversity, interaction strenghts and ecological network structure on the volume of parameter space leading to feasible equilibria, i.e., ones in which all populations have positive abundances. We develop a geometrical framework to study the range of conditions necessary for feasible coexistence in both mutualistic and consumer-resource systems. Using analytical and numerical methods, we show that feasibility is determined by just a handful of quantities describing the interactions, yielding a nontrivial complexity-feasibility relationship. Analyzing more than 100 empirical networks, we show that the range of coexistence conditions in mutualistic systems can be analytically predicted by means of a null model of random interactions, whereas food webs are characterized by smaller coexistence domains than those expected by chance. Finally, we characterize the geometric shape of the feasibility domain, thereby identifying the direction of perturbations that are more likely to cause extinctions. Interestingly, the structure of mutualistic interactions leads to very heterogeneous responses to perturbations, making those systems more fragile than expected by chance.

preprint2015arXiv

Phase diagram of the ground states of DNA condensates

Phase diagram of the ground states of DNA in a bad solvent is studied for a semi-flexible polymer model with a generalized local elastic bending potential characterized by a nonlinearity parameter $x$ and effective self-attraction promoting compaction. $x=1$ corresponds to the worm-like chain model. Surprisingly, the phase diagram as well as the transition lines between the ground states are found to be a function of $x$. The model provides a simple explanation for the results of prior experimental and computational studies and makes predictions for the specific geometries of the ground states. The results underscore the impact of the form of the microscopic bending energy at macroscopic observable scales.

preprint2015arXiv

Statistical Mechanics of Ecological Systems: Neutral Theory and Beyond

The simplest theories often have much merit and many limitations, and in this vein, the value of Neutral Theory (NT) has been the subject of much debate over the past 15 years. NT was proposed at the turn of the century by Stephen Hubbell to explain pervasive patterns observed in the organization of ecosystems. Its originally tepid reception among ecologists contrasted starkly with the excitement it caused among physicists and mathematicians. Indeed, NT spawned several theoretical studies that attempted to explain empirical data and predicted trends of quantities that had not yet been studied. While there are a few reviews of NT oriented towards ecologists, our goal here is to review the quantitative results of NT and its extensions for physicists who are interested in learning what NT is, what its successes are and what important problems remain unresolved. Furthermore, we hope that this review could also be of interest to theoretical ecologists because many potentially interesting results are buried in the vast NT literature. We propose to make these more accessible by extracting them and presenting them in a logical fashion. We conclude the review by discussing how one might introduce realistic non-neutral elements into the current models.

preprint2015arXiv

Uncovering low-dimensional, miR-based signatures of acute myeloid and lymphoblastic leukemias with a machine-learning-driven network approach

Complex phenotypic differences among different acute leukemias cannot be fully captured by analyzing the expression levels of one single molecule, such as a miR, at a time, but requires systematic analysis of large sets of miRs. While a popular approach for analysis of such datasets is principal component analysis (PCA), this method is not designed to optimally discriminate different phenotypes. Moreover, PCA and other low-dimensional representation methods yield linear or non-linear combinations of all measured miRs. Global human miR expression was measured in AML, B-ALL, and T-ALL cell lines and patient RNA samples. By systematically applying support vector machines to all measured miRs taken in dyad and triad groups, we built miR networks using cell line data and validated our findings with primary patient samples. All the coordinately transcribed members of the miR-23a cluster (which includes also miR-24 and miR-27a), known to function as tumor suppressors of acute leukemias, appeared in the AML, B-ALL and T-ALL centric networks. Subsequent qRT-PCR analysis showed that the most connected miR in the B-ALL-centric network, miR-708, is highly and specifically expressed in B-ALLs, suggesting that miR-708 might serve as a biomarker for B-ALL. This approach is systematic, quantitative, scalable, and unbiased. Rather than a single signature, our approach yields a network of signatures reflecting the redundant nature of biological signaling pathways. The network representation allows for visual analysis of all signatures by an expert and for future integration of additional information. Furthermore, each signature involves only small sets of miRs, such as dyads and triads, which are well suited for in depth validation through laboratory experiments such as loss- and gain-of-function assays designed to drive changes in leukemia cell survival, proliferation and differentiation.

preprint2014arXiv

From toroidal to rod-like condensates of semiflexible polymers

The competition between toroidal and rod-like conformations as possible ground states for DNA condensation is studied as a function of the stiffness, the length of the DNA and the form of the long-range interactions between neighboring molecules, using analytical theory supported by Monte Carlo simulations. Both conformations considered are characterized by a local nematic order with hexagonal packing symmetry of neighboring DNA molecules, but differ in global configuration of the chain and the distribution of its curvature as it wraps around to form a condensate. The long-range interactions driving the DNA condensation are assumed to be of the form pertaining to the attractive depletion potential as well as the attractive counterion induced soft potential. In the stiffness-length plane we find a transition between rod-like to toroid condensate for increasing stiffness at a fixed chain length $L$. Strikingly, the transition line is found to have a $L^{1/3}$ dependence irrespective of the details of the long-range interactions between neighboring molecules. When realistic DNA parameters are used, our description reproduces rather well some of the experimental features observed in DNA condensates.

preprint2014arXiv

Information-based fitness and the emergence of criticality in living systems

Empirical evidence suggesting that living systems might operate in the vicinity of critical points, at the borderline between order and disorder, has proliferated in recent years, with examples ranging from spontaneous brain activity to flock dynamics. However, a well-founded theory for understanding how and why interacting living systems could dynamically tune themselves to be poised in the vicinity of a critical point is lacking. Here we employ tools from statistical mechanics and information theory to show that complex adaptive or evolutionary systems can be much more efficient in coping with diverse heterogeneous environmental conditions when operating at criticality. Analytical as well as computational evolutionary and adaptive models vividly illustrate that a community of such systems dynamically self-tunes close to a critical state as the complexity of the environment increases while they remain non-critical for simple and predictable environments. A more robust convergence to criticality emerges in co-evolutionary and co-adaptive set-ups in which individuals aim to represent other agents in the community with fidelity, thereby creating a collective critical ensemble and providing the best possible trade-off between accuracy and flexibility. Our approach provides a parsimonious and general mechanism for the emergence of critical-like behavior in living systems needing to cope with complex environments or trying to efficiently coordinate themselves as an ensemble.

preprint2014arXiv

Spatial maximum entropy modeling from presence/absence tropical forest data

Understanding the assembly of ecosystems to estimate the number of species at different spatial scales is a challenging problem. Until now, maximum entropy approaches have lacked the important feature of considering space in an explicit manner. We propose a spatially explicit maximum entropy model suitable to describe spatial patterns such as the species area relationship and the endemic area relationship. Starting from the minimal information extracted from presence/absence data, we compare the behavior of two models considering the occurrence or lack thereof of each species and information on spatial correlations. Our approach uses the information at shorter spatial scales to infer the spatial organization at larger ones. We also hypothesize a possible ecological interpretation of the effective interaction we use to characterize spatial clustering.

preprint2013arXiv

Emergence of structural and dynamical properties of ecological mutualistic networks

Mutualistic networks are formed when the interactions between two classes of species are mutually beneficial. They are important examples of cooperation shaped by evolution. Mutualism between animals and plants plays a key role in the organization of ecological communities. Such networks in ecology have generically evolved a nested architecture independent of species composition and latitude - specialists interact with proper subsets of the nodes with whom generalists interact. Despite sustained efforts to explain observed network structure on the basis of community-level stability or persistence, such correlative studies have reached minimal consensus. Here we demonstrate that nested interaction networks could emerge as a consequence of an optimization principle aimed at maximizing the species abundance in mutualistic communities. Using analytical and numerical approaches, we show that because of the mutualistic interactions, an increase in abundance of a given species results in a corresponding increase in the total number of individuals in the community, as also the nestedness of the interaction matrix. Indeed, the species abundances and the nestedness of the interaction matrix are correlated by an amount that depends on the strength of the mutualistic interactions. Nestedness and the observed spontaneous emergence of generalist and specialist species occur for several dynamical implementations of the variational principle under stationary conditions. Optimized networks, while remaining stable, tend to be less resilient than their counterparts with randomly assigned interactions. In particular, we analytically show that the abundance of the rarest species is directly linked to the resilience of the community. Our work provides a unifying framework for studying the emergent structural and dynamical properties of ecological mutualistic networks.

preprint2013arXiv

From Cellular Characteristics to Disease Diagnosis: Uncovering Phenotypes with Supercells

Cell heterogeneity and the inherent complexity due to the interplay of multiple molecular processes within the cell pose difficult challenges for current single-cell biology. We introduce an approach that identifies a disease phenotype from multiparameter single-cell measurements, which is based on the concept of "supercell statistics", a single-cell-based averaging procedure followed by a machine learning classification scheme. We are able to assess the optimal tradeoff between the number of single cells averaged and the number of measurements needed to capture phenotypic differences between healthy and diseased patients, as well as between different diseases that are difficult to diagnose otherwise. We apply our approach to two kinds of single-cell datasets, addressing the diagnosis of a premature aging disorder using images of cell nuclei, as well as the phenotypes of two non-infectious uveitides (the ocular manifestations of Behçet's disease and sarcoidosis) based on multicolor flow cytometry. In the former case, one nuclear shape measurement taken over a group of 30 cells is sufficient to classify samples as healthy or diseased, in agreement with usual laboratory practice. In the latter, our method is able to identify a minimal set of 5 markers that accurately predict Behçet's disease and sarcoidosis. This is the first time that a quantitative phenotypic distinction between these two diseases has been achieved. To obtain this clear phenotypic signature, about one hundred CD8+ T cells need to be measured. Beyond these specific cases, the approach proposed here is applicable to datasets generated by other kinds of state-of-the-art and forthcoming single-cell technologies, such as multidimensional mass cytometry, single-cell gene expression, and single-cell full genome sequencing techniques.

preprint2013arXiv

Understanding Health and Disease with Multidimensional Single-Cell Methods

Current efforts in the biomedical sciences and related interdisciplinary fields are focused on gaining a molecular understanding of health and disease, which is a problem of daunting complexity that spans many orders of magnitude in characteristic length scales, from small molecules that regulate cell function to cell ensembles that form tissues and organs working together as an organism. In order to uncover the molecular nature of the emergent properties of a cell, it is essential to measure multiple cell components simultaneously in the same cell. In turn, cell heterogeneity requires multiple cells to be measured in order to understand health and disease in the organism. This review summarizes current efforts towards a data-driven framework that leverages single-cell technologies to build robust signatures of healthy and diseased phenotypes. While some approaches focus on multicolor flow cytometry data and other methods are designed to analyze high-content image-based screens, we emphasize the so-called Supercell/SVM paradigm (recently developed by the authors of this review and collaborators) as a unified framework that captures mesoscopic-scale emergence to build reliable phenotypes. Beyond their specific contributions to basic and translational biomedical research, these efforts illustrate, from a larger perspective, the powerful synergy that might be achieved from bringing together methods and ideas from statistical physics, data mining, and mathematics to solve the most pressing problems currently facing the life sciences.

preprint2012arXiv

Absence of Detailed Balance in Ecology

Living systems are typically characterized by irreversible processes. A condition equivalent to the reversibility is the detailed balance, whose absence is an obstacle for analytically solving ecological models. We revisit a promising model with an elegant field-theoretic analytic solution and show that the theoretical analysis is invalid because of an implicit assumption of detailed balance. A signature of the difficulties is evident in the inconsistencies appearing in the many-point correlation functions and in the analytical formula for the species area relationship.

preprint2012arXiv

An allometry-based approach for understanding forest structure, predicting tree-size distribution and assessing the degree of disturbance

Tree-size distribution is one of the most investigated subjects in plant population biology. The forestry literature reports that tree-size distribution trajectories vary across different stands and/or species, while the metabolic scaling theory suggests that the tree number scales universally as -2 power of diameter. Here, we propose a simple functional scaling model in which these two opposing results are reconciled. Basic principles related to crown shape, energy optimization and the finite size scaling approach were used to define a set of relationships based on a single parameter, which allows us to predict the slope of the tree-size distributions in a steady state condition. We tested the model predictions on four temperate mountain forests. Plots (4 ha each, fully mapped) were selected with different degrees of human disturbance (semi-natural stands vs. formerly managed). Results showed that the size distribution range successfully fitted by the model is related to the degree of forest disturbance: in semi-natural forests the range is wide, while in formerly managed forests, the agreement with the model is confined to a very restricted range. We argue that simple allometric relationships, at individual level, shape the structure of the whole forest community.

preprint2012arXiv

Protein sequence and structure: Is one more fundamental than the other?

We argue that protein native state structures reside in a novel "phase" of matter which confers on proteins their many amazing characteristics. This phase arises from the common features of all globular proteins and is characterized by a sequence-independent free energy landscape with relatively few low energy minima with funnel-like character. The choice of a sequence that fits well into one of these predetermined structures facilitates rapid and cooperative folding. Our model calculations show that this novel phase facilitates the formation of an efficient route for sequence design starting from random peptides.

preprint2012arXiv

Spontaneously Broken Neutral Symmetry in an Ecological System

Spontaneous symmetry breaking plays a fundamental role in many areas of condensed matter and particle physics. A fundamental problem in ecology is the elucidation of the mechanisms responsible for biodiversity and stability. Neutral theory, which makes the simplifying assumption that all individuals (such as trees in a tropical forest) --regardless of the species they belong to-- have the same prospect of reproduction, death, etc., yields gross patterns that are in accord with empirical data. We explore the possibility of birth and death rates that depend on the population density of species while treating the dynamics in a species-symmetric manner. We demonstrate that the dynamical evolution can lead to a stationary state characterized simultaneously by both biodiversity and spontaneously broken neutral symmetry.

preprint2010arXiv

Self-Similarity and Scaling in Forest Communities

Ecological communities exhibit pervasive patterns and inter-relationships between size, abundance, and the availability of resources. We use scaling ideas to develop a unified, model-independent framework for understanding the distribution of tree sizes, their energy use and spatial distribution in tropical forests. We demonstrate that the scaling of the tree crown at the individual level drives the forest structure when resources are fully used. Our predictions match perfectly with the scaling behaviour of an exactly solvable self-similar model of a forest and are in good accord with empirical data. The range, over which pure power law behaviour is observed, depends on the available amount of resources. The scaling framework can be used for assessing the effects of natural and anthropogenic disturbances on ecosystem structure and functionality.

preprint2010arXiv

System size expansion for systems with an absorbing state

The well known van Kampen system size expansion, while of rather general applicability, is shown to fail to reproduce some qualitative features of the time evolution for systems with an absorbing state, apart from a transient initial time interval. We generalize the van Kampen ansatz by introducing a new prescription leading to non-Gaussian fluctuations around the absorbing state. The two expansion predictions are explicitly compared for the infinite range voter model with speciation as a paradigmatic model with an absorbing state. The new expansion, both for a finite size system in the large time limit and at finite time in the large size limit, converges to to the exact solution as obtained in a numerical implementation using the Gillespie algorithm. Furthermore, the predicted lifetime distribution is shown to have the correct asymptotic behavior.

preprint2009arXiv

First-principles design of nanomachines

Learning from nature's amazing molecular machines, globular proteins, we present a framework for the predictive design of nano-machines. We show that the crucial ingredients for a chain molecule to behave as a machine are its inherent anisotropy and the coupling between the local Frenet coordinate reference frames of nearby monomers. We demonstrate that, even in the absence of heterogeneity, protein-like behavior is obtained for a simple chain molecule made up of just thirty hard spheres. This chain spontaneously switches between two distinct geometries, a single helix and a dual helix, merely due to thermal fluctuations.

preprint2009arXiv

Probing Noise in Gene Expression and Protein Production

We derive exact solutions of simplified models for the temporal evolution of the protein concentration within a cell population arbitrarily far from the stationary state. We show that monitoring the dynamics can assist in modeling and understanding the nature of the noise and its role in gene expression and protein production. We introduce a new measure, the cell turnover distribution, which can be used to probe the phase of transcription of DNA into messenger RNA.