Source author record

Nathan A. Baker

Nathan A. Baker appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

11works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

11 published item(s)

preprint2020arXiv

Data-driven molecular modeling with the generalized Langevin equation

The complexity of molecular dynamics simulations necessitates dimension reduction and coarse-graining techniques to enable tractable computation. The generalized Langevin equation (GLE) describes coarse-grained dynamics in reduced dimensions. In spite of playing a crucial role in non-equilibrium dynamics, the memory kernel of the GLE is often ignored because it is difficult to characterize and expensive to solve. To address these issues, we construct a data-driven rational approximation to the GLE. Building upon previous work leveraging the GLE to simulate simple systems, we extend these results to more complex molecules, whose many degrees of freedom and complicated dynamics require approximation methods. We demonstrate the effectiveness of our approximation by testing it against exact methods and comparing observables such as autocorrelation and transition rates.

preprint2020arXiv

Towards quantum computing for high-energy excited states in molecular systems: quantum phase estimations of core-level states

This paper explores the utility of the quantum phase estimation (QPE) in calculating high-energy excited states characterized by promotions of electrons occupying inner energy shells. These states have been intensively studied over the last few decades especially in supporting the experimental effort at light sources. Results obtained with the QPE are compared with various high-accuracy many-body techniques developed to describe core-level states. The feasibility of the quantum phase estimator in identifying classes of challenging shake-up states characterized by the presence of higher-order excitation effects is also discussed.

preprint2016arXiv

An ISA-Tab specification for protein titration data exchange

Data curation presents a challenge to all scientific disciplines to ensure public availability and reproducibility of experimental data. Standards for data preservation and exchange are central to addressing this challenge: the Investigation-Study-Assay Tabular (ISA-Tab) project has developed a widely used template for such standards in biological research. This paper describes the application of ISA-Tab to protein titration data. Despite the importance of titration experiments for understanding protein structure, stability, and function and for testing computational approaches to protein electrostatics, no such mechanism currently exists for sharing and preserving biomolecular titration data. We have adapted the ISA-Tab template to provide a structured means of supporting experimental structural chemistry data with a particular emphasis on the calculation and measurement of pKa values. This activity has been performed as part of the broader pKa Cooperative effort, leveraging data that has been collected and curated by the Cooperative members. In this article, we present the details of this specification and its application to a broad range of pKa and electrostatics data obtained for multiple protein systems. The resulting curated data is publicly available at http://pkacoop.org.

preprint2016arXiv

Bayesian Model Averaging for Ensemble-Based Estimates of Solvation Free Energies

This paper applies the Bayesian Model Averaging (BMA) statistical ensemble technique to estimate small molecule solvation free energies. There is a wide range of methods available for predicting solvation free energies, ranging from empirical statistical models to ab initio quantum mechanical approaches. Each of these methods is based on a set of conceptual assumptions that can affect predictive accuracy and transferability. Using an iterative statistical process, we have selected and combined solvation energy estimates using an ensemble of 17 diverse methods from the fourth Statistical Assessment of Modeling of Proteins and Ligands (SAMPL) blind prediction study to form a single, aggregated solvation energy estimate. The ensemble design process evaluates the statistical information in each individual method as well as the performance of the aggregate estimate obtained from the ensemble as a whole. Methods that possess minimal or redundant information are pruned from the ensemble and the evaluation process repeats until aggregate predictive performance can no longer be improved. We show that this process results in a final aggregate estimate that outperforms all individual methods by reducing estimate errors by as much as 91% to 1.2 kcal/mol accuracy. We also compare our iterative refinement approach to other statistical ensemble approaches and demonstrate that this iterative process reduces estimate errors by as much as 61%. This work provides a new approach for accurate solvation free energy prediction and lays the foundation for future work on aggregate models that can balance computational cost with prediction accuracy.

preprint2016arXiv

Continuum Electrostatics Approaches to Calculating p$K_a$s and $E_m$s in Proteins

Proteins change their charge state through protonation and redox reactions as well as through binding charged ligands. The free energy of these reactions are dominated by solvation and electrostatic energies and modulated by protein conformational relaxation in response to the ionization state changes. Although computational methods for calculating these interactions can provide very powerful tools for predicting protein charge states, they include several critical approximations of which users should be aware. This chapter discusses the strengths, weaknesses, and approximations of popular computational methods for predicting charge states and understanding their underlying electrostatic interactions. The goal of this chapter is to inform users about applications and potential caveats of these methods as well as outline directions for future theoretical and computational research.

preprint2016arXiv

Energy Minimization of Discrete Protein Titration State Models Using Graph Theory

There are several applications in computational biophysics which require the optimization of discrete interacting states; e.g., amino acid titration states, ligand oxidation states, or discrete rotamer angles. Such optimization can be very time-consuming as it scales exponentially in the number of sites to be optimized. In this paper, we describe a new polynomial-time algorithm for optimization of discrete states in macromolecular systems. This algorithm was adapted from image processing and uses techniques from discrete mathematics and graph theory to restate the optimization problem in terms of "maximum flow-minimum cut" graph analysis. The interaction energy graph, a graph in which vertices (amino acids) and edges (interactions) are weighted with their respective energies, is transformed into a flow network in which the value of the minimum cut in the network equals the minimum free energy of the protein, and the cut itself encodes the state that achieves the minimum free energy. Because of its deterministic nature and polynomial-time performance, this algorithm has the potential to allow for the ionization state of larger proteins to be discovered.

preprint2016arXiv

Multi-shell model of ion-induced nucleic acid condensation

We present a semi-quantitative model of condensation of short nucleic acid (NA) duplexes induced by tri-valent cobalt(III) hexammine (CoHex) ions. The model is based on partitioning of bound counterion distribution around singleNA duplex into "external" and "internal" ion binding shells distinguished by the proximity to duplex helical axis. In the aggregated phase the shells overlap, which leads to significantly increased attraction of CoHex ions in these overlaps with the neighboring duplexes. The duplex aggregation free energy is decomposed into attractive and repulsive components in such a way that they can be represented by simple analytical expressions with parameters derived from molecular dynamic (MD) simulations and numerical solutions of Poisson equation. The short-range interactions described by the attractive term depend on the fractions of bound ions in the overlapping shells and affinity of CoHex to the "external" shell of nearly neutralized duplex. The repulsive components of the free energy are duplex configurational entropy loss upon the aggregation and the electrostatic repulsion of the duplexes that remains after neutralization by bound CoHex ions. The estimates of the aggregation free energy are consistent with the experimental range of NA duplex condensation propensities, including the unusually poor condensation of RNA structures and subtle sequence effects upon DNA condensation. The model predicts that, in contrast to DNA, RNA duplexes may condense into tighter packed aggregates with a higher degree of duplex neutralization. The model also predicts that longer NA fragments will condense more readily than shorter ones. The ability of this model to explain experimentally observed trends in NA condensation, lends support to proposed NA condensation picture based on the multivalent "ion binding shells".

preprint2015arXiv

Enhancing Sparsity of Hermite Polynomial Expansions by Iterative Rotations

Compressive sensing has become a powerful addition to uncertainty quantification in recent years. This paper identifies new bases for random variables through linear mappings such that the representation of the quantity of interest is more sparse with new basis functions associated with the new random variables. This sparsity increases both the efficiency and accuracy of the compressive sensing-based uncertainty quantification method. Specifically, we consider rotation-based linear mappings which are determined iteratively for Hermite polynomial expansions. We demonstrate the effectiveness of the new method with applications in solving stochastic partial differential equations and high-dimensional ($\mathcal{O}(100)$) problems.

preprint2015arXiv

Numerical calculation of protein-ligand binding rates through solution of the Smoluchowski equation using smooth particle hydrodynamics

Background. The calculation of diffusion-controlled ligand binding rates is important for understanding enzyme mechanisms as well as designing enzyme inhibitors. We demonstrate the accuracy and effectiveness of a Lagrangian particle-based method, smoothed particle hydrodynamics (SPH), to study diffusion in biomolecular systems by numerically solving the time-dependent Smoluchowski equation for continuum diffusion. Results. The numerical method is first verified in simple systems and then applied to the calculation of ligand binding to an acetylcholinesterase monomer. Unlike previous studies, a reactive Robin boundary condition (BC), rather than the absolute absorbing (Dirichlet) boundary condition, is considered on the reactive boundaries. This new boundary condition treatment allows for the analysis of enzymes with "imperfect" reaction rates. Rates for inhibitor binding to mAChE are calculated at various ionic strengths and compared with experiment and other numerical methods. We find that imposition of the Robin BC improves agreement between calculated and experimental reaction rates. Conclusions. Although this initial application focuses on a single monomer system, our new method provides a framework to explore broader applications of SPH in larger-scale biomolecular complexes by taking advantage of its Lagrangian particle-based nature.

preprint2015arXiv

Smoothed Dissipative Particle Dynamics model for mesoscopic multiphase flows in the presence of thermal fluctuations

Thermal fluctuations cause perturbations of fluid-fluid interfaces and highly nonlinear hydrodynamics in multiphase flows. In this work, we develop a novel multiphase smoothed dissipative particle dynamics model. This model accounts for both bulk hydrodynamics and interfacial fluctuations. Interfacial surface tension is modeled by imposing a pairwise force between SDPD particles. We show that the relationship between the model parameters and surface tension, previously derived under the assumption of zero thermal fluctuation, is accurate for fluid systems at low temperature but overestimates the surface tension for intermediate and large thermal fluctuations. To analyze the effect of thermal fluctuations on surface tension, we construct a coarse-grained Euler lattice model based on the mean field theory and derive a semi-analytical formula to directly relate the surface tension to model parameters for a wide range of temperatures and model resolutions. We demonstrate that the present method correctly models the dynamic processes, such as bubble coalescence and capillary spectra across the interface.

preprint2015arXiv

The role of correlation and solvation in ion interactions with B-DNA

The ionic atmospheres around nucleic acids play important roles in biological function. Large-scale explicit solvent simulations coupled to experimental assays such as anomalous small-angle X-ray scattering (ASAXS) can provide important insights into the structure and energetics of such atmospheres but are time- and resource-intensive. In this paper, we use classical density functional theory (cDFT) to explore the balance between ion-DNA, ion-water, and ion-ion interactions in ionic atmospheres of RbCl, SrCl$_2$, and CoHexCl$_3$ (cobalt hexammine chloride) around a B-form DNA molecule. The accuracy of the cDFT calculations was assessed by comparison between simulated and experimental ASAXS curves, demonstrating that an accurate model should take into account ion-ion correlation and ion hydration forces, DNA topology, and the discrete distribution of charges on DNA strands. As expected, these calculations revealed significant differences between monovalent, divalent, and trivalent cation distributions around DNA. About half of the DNA-bound Rb$^+$ ions penetrate into the minor groove of the DNA and half adsorb on the DNA strands. The fraction of cations in the minor groove decreases for the larger Sr$^{2+}$ ions and becomes zero for CoHex$^{3+}$ ions, which all adsorb on the DNA strands. The distribution of CoHex$^{3+}$ ions is mainly determined by Coulomb and steric interactions, while ion-correlation forces play a central role in the monovalent Rb$^+$ distribution and a combination of ion-correlation and hydration forces affect the Sr$^{2+}$ distribution around DNA.