Source author record

Aaron Smith

Aaron Smith appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

28works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

28 published item(s)

preprint2025arXiv

RIGEL: Feedback-regulated cloud-scale star formation efficiency in a simulated dwarf galaxy merger

Major mergers of galaxies are likely to trigger bursty star formation activities. The accumulation of dense gas and the boost of star formation efficiency (SFE) are considered to be the two main drivers of starbursts. However, it remains unclear how each process operates on the scale of individual star-forming clouds. Here, we present a high-resolution (2 Msun) RHD simulation of a gas-rich dwarf galaxy merger using the RIGEL model to investigate how mergers affect the properties of the structure of dense star-forming gas and the cloud-scale SFE. We tracked the evolution of sub-virial dense clouds in the simulation by mapping them across successive snapshots taken at intervals of 0.2 Myr. We find that the merger triggers a 130 fold increase in the SFR and shortens the galaxy-wide gas-depletion time by two orders of magnitude compared to those in two isolated galaxies. However, the depletion time of individual clouds and their lifetime distribution remained unchanged over the simulation period. The cloud life cycles and cloud-scale SFE are determined by the local stellar feedback rather than such environmental factors as tidal fields regardless of the merger process, and the integrated SFE ($ε_{\rm int}$) of clouds in complex environments remains well-described by an $ε_{\rm int}-Σ_{\rm tot}$ relation found in idealized isolated-cloud experiments. During the peak of the starburst, the media SFE was lower by only 0.17-0.33 dex compared to the value when the galaxies were not interacting. The merger boosts the SFR through the accumulation and compression of dense gas fueling star formation. Strong tidal torques assemble $>10^5$ Msun clouds, which seed massive star clusters. The average separation between star-forming clouds decreases during the merger, which in turn decreases the cloud--cluster spatial de-correlation from >1 kpc to 0.1 kpc depicted in tuning fork diagrams.

preprint2025arXiv

The THESAN-ZOOM project: Star-formation efficiencies in high-redshift galaxies

Recent JWST observations hint at unexpectedly intense cosmic star-formation in the early Universe, often attributed to enhanced star-formation efficiencies (SFEs). Here, we analyze the SFE in THESAN-ZOOM, a novel zoom-in radiation-hydrodynamic simulation campaign of high-redshift ($z \gtrsim 3$) galaxies employing a state-of-the-art galaxy formation model resolving the multiphase interstellar medium (ISM). The halo-scale SFE ($ε^{\ast}_{\rm halo}$) - the fraction of baryons accreted by a halo that are converted to stars - follows a double power-law dependence on halo mass, with a mild redshift evolution above $M_{\rm halo} \gtrsim 10^{9.5}\,{\rm M}_{\odot}$. The power-law slope is roughly $1/3$ at large halo masses, consistent with expectations when gas outflows are momentum-driven. At lower masses, the slope is roughly $2/3$ and is more aligned with the energy-driven outflow scenario. $ε^{\ast}_{\rm halo}$ is a factor of $2-3$ larger than commonly assumed in empirical galaxy-formation models at $M_{\rm halo} \lesssim 10^{11}\,{\rm M}_{\odot}$. On galactic (kpc) scales, the Kennicutt-Schmidt (KS) relation of neutral gas is universal in THESAN-ZOOM, following $Σ_{\rm SFR} \propto Σ_{\rm gas}^2$, indicative of a turbulent energy balance in the ISM maintained by stellar feedback. The rise of $ε^{\ast}_{\rm halo}$ with halo mass can be traced primarily to increasing gas surface densities in massive galaxies, while the underlying KS relation and neutral, star-forming gas fraction remain unchanged. Although the increase in $ε^{\ast}_{\rm halo}$ with redshift is relatively modest, it is sufficient to explain the large observed number density of UV-bright galaxies at $z \gtrsim 12$. However, reproducing the brightest sources at $M_{\rm UV} \lesssim -21$ may require extrapolating the SFE beyond the halo mass range directly covered by THESAN-ZOOM.

preprint2023arXiv

There and back again: understanding the critical properties of backsplash galaxies

Backsplash galaxies are galaxies that once resided inside a cluster, and have migrated back oustide as they move towards the apocentre of their orbit. The kinematic properties of these galaxies are well understood, thanks to the significant study of backsplashers in dark matter-only simulations, but their intrinsic properties are not well constrained due to modelling uncertainties in sub-grid physics, ram pressure stripping, dynamical friction, and tidal forces. In this paper, we use the IllustrisTNG300-1 simulation, with a baryonic resolution of $M_{\rm b} \approx 1.1\times 10^7$ M$_\odot$, to study backsplash galaxies around 1302 isolated galaxy clusters with mass $10^{13.0} < M_{\rm 200,mean} / {\rm M}_\odot< 10^{15.5}$. We employ a decision tree classifier to extract features of galaxies that make them likely to be backsplash galaxies, compared to nearby field galaxies, and find that backsplash galaxies have low gas fractions, high mass-to-light ratios, large stellar sizes, and low black hole occupation fractions. We investigate in detail the origins of these large sizes, and hypothesise their origins are linked to the tidal environments in the cluster. We show that the black hole recentreing scheme employed in many cosmological simulations leads to the loss of black holes from galaxies accreted into clusters, and suggest improvements to these models. Generally, we find that backsplash galaxies are a useful population to test and understand numerical galaxy formation models due to their challenging environments and evolutionary pathways that interact with poorly constrained physics.

preprint2022arXiv

H-alpha emission in local galaxies: star formation, time variability and the diffuse ionized gas

The nebular recombination line H$α$ is widely used as a star-formation rate (SFR) indicator in the local and high-redshift Universe. We present a detailed H$α$ radiative transfer study of high-resolution isolated Milky-Way and Large Magellanic Cloud simulations that include radiative transfer, non-equilibrium thermochemistry, and dust evolution. We focus on the spatial morphology and temporal variability of the H$α$ emission, and its connection to the underlying gas and star formation properties. The H$α$ and H$β$ radial and vertical surface brightness profiles are in excellent agreement with observations of nearby galaxies. We find that the fraction of H$α$ emission from collisional excitation amounts to $f_{\rm col}\sim5-10\%$, only weakly dependent on radius and vertical height, and that scattering boosts the H$α$ luminosity by $\sim40\%$. The dust correction via the Balmer decrement works well (intrinsic H$α$ emission recoverable within $25\%$), though the dust attenuation law depends on the amount of attenuation itself both on spatially resolved and integrated scales. Important for the understanding of the H$α$-SFR connection is the dust and helium absorption of ionizing radiation (Lyman continuum [LyC] photons), which are about $f_{\rm abs}\approx28\%$ and $f_{\rm He}\approx9\%$, respectively. Together with an escape fraction of $f_{\rm esc}\approx6\%$, this reduces the available budget for hydrogen line emission by nearly half ($f_{\rm H}\approx57\%$). We discuss the impact of the diffuse ionized gas, showing - among other things - that the extraplanar H$α$ emission is powered by LyC photons escaping the disc. Future applications of this framework to cosmological (zoom-in) simulations will assist in the interpretation of spectroscopy of high-redshift galaxies with the upcoming James Webb Space Telescope.

preprint2022arXiv

Rapid Convergence of Informed Importance Tempering

Informed Markov chain Monte Carlo (MCMC) methods have been proposed as scalable solutions to Bayesian posterior computation on high-dimensional discrete state spaces, but theoretical results about their convergence behavior in general settings are lacking. In this article, we propose a class of MCMC schemes called informed importance tempering (IIT), which combine importance sampling and informed local proposals, and derive generally applicable spectral gap bounds for IIT estimators. Our theory shows that IIT samplers have remarkable scalability when the target posterior distribution concentrates on a small set. Further, both our theory and numerical experiments demonstrate that the informed proposal should be chosen with caution: the performance of some proposals may be very sensitive to the shape of the target distribution. We find that the "square-root proposal weighting" scheme tends to perform well in most settings.

preprint2022arXiv

Rate-optimal refinement strategies for local approximation MCMC

Many Bayesian inference problems involve target distributions whose density functions are computationally expensive to evaluate. Replacing the target density with a local approximation based on a small number of carefully chosen density evaluations can significantly reduce the computational expense of Markov chain Monte Carlo (MCMC) sampling. Moreover, continual refinement of the local approximation can guarantee asymptotically exact sampling. We devise a new strategy for balancing the decay rate of the bias due to the approximation with that of the MCMC variance. We prove that the error of the resulting local approximation MCMC (LA-MCMC) algorithm decays at roughly the expected $1/\sqrt{T}$ rate, and we demonstrate this rate numerically. We also introduce an algorithmic parameter that guarantees convergence given very weak tail bounds, significantly strengthening previous convergence results. Finally, we apply LA-MCMC to a computationally intensive Bayesian inverse problem arising in groundwater hydrology.

preprint2022arXiv

The THESAN project: predictions for multi-tracer line intensity mapping in the epoch of reionization

Line intensity mapping (LIM) is rapidly emerging as a powerful technique to study galaxy formation and cosmology in the high-redshift Universe. We present LIM estimates of select spectral lines originating from the interstellar medium (ISM) of galaxies and 21 cm emission from neutral hydrogen gas in the Universe using the large volume, high resolution THESAN reionization simulations. A combination of sub-resolution photo-ionization modelling for HII regions and Monte Carlo radiative transfer calculations is employed to estimate the dust-attenuated spectral energy distributions (SEDs) of high-redshift galaxies ($z\gtrsim5.5$). We show that the derived photometric properties such as the ultraviolet (UV) luminosity function and the UV continuum slopes match observationally inferred values, demonstrating the accuracy of the SED modelling. We provide fits to the luminosity--star formation rate relation (L-SFR) for the brightest emission lines and find that important differences exist between the derived scaling relations and the widely used low-$z$ ones because the interstellar medium of reionization era galaxies is generally less metal-enriched than in their low redshift counterparts. We use these relations to construct line intensity maps of nebular emission lines and cross correlate with the 21 cm emission. Interestingly, the wavenumber at which the correlation switches sign ($k_\mathrm{transition}$) depends heavily on the reionization model and to a lesser extent on the targeted emission line, which is consistent with the picture that $k_\mathrm{transition}$ probes the typical sizes of ionized regions. The derived scaling relations and intensity maps represent a timely state-of-the-art framework for forecasting and interpreting results from current and upcoming LIM experiments.

preprint2022arXiv

The THESAN project: properties of the intergalactic medium and its connection to Reionization-era galaxies

The high-redshift intergalactic medium (IGM) and the primeval galaxy population are rapidly becoming the new frontier of extra-galactic astronomy. We investigate the IGM properties and their connection to galaxies at $z\geq5.5$ under different assumptions for the ionizing photon escape and the nature of dark matter, employing our novel THESAN radiation-hydrodynamical simulation suite, designed to provide a comprehensive picture of the emergence of galaxies in a full reionization context. Our simulations have realistic `late' reionization histories, match available constraints on global IGM properties and reproduce the recently-observed rapid evolution of the mean free path of ionizing photons. We additionally examine high-z Lyman-$α$ transmission. The optical depth evolution is consistent with data, and its distribution suggests an even-later reionization than simulated, although with a strong sensitivity to the source model. We show that the effects of these two unknowns can be disentangled by characterising the spectral shape and separation of Lyman-$α$ transmission regions, opening up the possibility to observationally constrain both. For the first time in simulations, THESAN reproduces the modulation of the Lyman-$α$ flux as a function of galaxy distance, demonstrating the power of coupling a realistic galaxy formation model with proper radiation-hydrodynamics. We find this feature to be extremely sensitive on the timing of reionization, while being relatively insensitive to the source model. Overall, THESAN produces a realistic IGM and galaxy population, providing a robust framework for future analysis of the high-z Universe.

preprint2020arXiv

Drift, Minorization, and Hitting Times

The "drift-and-minorization" method, introduced and popularized in (Rosenthal, 1995; Meyn and Tweedie, 1994; Meyn and Tweedie, 2012), remains the most popular approach for bounding the convergence rates of Markov chains used in statistical computation. This approach requires estimates of two quantities: the rate at which a single copy of the Markov chain "drifts" towards a fixed "small set", and a "minorization condition" which gives the worst-case time for two Markov chains started within the small set to couple with moderately large probability. In this paper, we build on (Oliveira, 2012; Peres and Sousi, 2015) and our work (Anderson, Duanmu, Smith, 2019a; Anderson, Duanmu, Smith, 2019b) to replace the "minorization condition" with an alternative "hitting condition" that is stated in terms of only one Markov chain, and illustrate how this can be used to obtain similar bounds that can be easier to use.

preprint2020arXiv

Fast and memory-optimal dimension reduction using Kac's walk

In this work, we analyze dimension reduction algorithms based on the Kac walk and discrete variants. (1) For $n$ points in $\mathbb{R}^{d}$, we design an optimal Johnson-Lindenstrauss (JL) transform based on the Kac walk which can be applied to any vector in time $O(d\log{d})$ for essentially the same restriction on $n$ as in the best-known transforms due to Ailon and Liberty [SODA, 2008], and Bamberger and Krahmer [arXiv, 2017]. Our algorithm is memory-optimal, and outperforms existing algorithms in regimes when $n$ is sufficiently large and the distortion parameter is sufficiently small. In particular, this confirms a conjecture of Ailon and Chazelle [STOC, 2006] in a stronger form. (2) The same construction gives a simple transform with optimal Restricted Isometry Property (RIP) which can be applied in time $O(d\log{d})$ for essentially the same range of sparsity as in the best-known such transform due to Ailon and Rauhut [Discrete Comput. Geom., 2014]. (3) We show that by fixing the angle in the Kac walk to be $π/4$ throughout, one obtains optimal JL and RIP transforms with almost the same running time, thereby confirming -- up to a $\log\log{d}$ factor -- a conjecture of Avron, Maymounkov, and Toledo [SIAM J. Sci. Comput., 2010]. Our moment-based analysis of this modification of the Kac walk may also be of independent interest.

preprint2020arXiv

Resonant-line radiative transfer within power-law density profiles

Star-forming regions in galaxies are surrounded by vast reservoirs of gas capable of both emitting and absorbing Lyman-alpha (Lya) radiation. Observations of Lya emitters and spatially extended Lya haloes indeed provide insights into the formation and evolution of galaxies. However, due to the complexity of resonant scattering, only a few analytic solutions are known in the literature. We discuss several idealized but physically motivated scenarios to extend the existing formalism to new analytic solutions, enabling quantitative predictions about the transport and diffusion of Lya photons. This includes a closed form solution for the radiation field and derived quantities including the emergent flux, peak locations, energy density, average internal spectrum, number of scatters, outward force multiplier, trapping time, and characteristic radius. To verify our predictions, we employ a robust gridless Monte Carlo radiative transfer (GMCRT) method, which is straightforward to incorporate into existing ray-tracing codes but requires modifications to opacity-based calculations, including dynamical core-skipping acceleration schemes. We primarily focus on power-law density and emissivity profiles, however both the analytic and numerical methods can be generalized to other cases. Such studies provide additional intuition and understanding regarding the connection between the physical environments and observational signatures of galaxies throughout the Universe.

preprint2016arXiv

Elementary Bounds On Mixing Times for Decomposable Markov Chains

Many finite-state reversible Markov chains can be naturally decomposed into "projection" and "restriction" chains. In this paper we provide bounds on the total variation mixing times of the original chain in terms of the mixing properties of these related chains. This paper is in the tradition of existing bounds on Poincare and log-Sobolev constants of Markov chains in terms of similar decompositions [JSTV04, MR02, MR06, MY09]. Our proofs are simple, relying largely on recent results relating hitting and mixing times of reversible Markov chains [PS13, Oli12]. We describe situations in which our results give substantially better bounds than those obtained by applying existing decomposition results and provide examples for illustration.

preprint2016arXiv

Evidence for a direct collapse black hole in the Lyman-alpha source CR7

Throughout the epoch of reionization the most luminous Lyα emitters are capable of ionizing their own local bubbles. The CR7 galaxy at $z \approx 6.6$ stands out for its combination of exceptionally bright Lyα and HeII 1640 Angstrom line emission but absence of metal lines. As a result CR7 may be the first viable candidate host of a young primordial starburst or direct collapse black hole. High-resolution spectroscopy reveals a +160 km s$^{-1}$ velocity offset between the Lyα and HeII line peaks while the spatial extent of the Lyα emitting region is $\sim 16$ kpc. The observables are indicative of an outflow signature produced by a strong central source. We present one-dimensional radiation-hydrodynamics simulations incorporating accurate Lyα feedback and ionizing radiation to investigate the nature of the CR7 source. We find that a Population III star cluster with $10^5$ K blackbody emission ionizes its environment too efficiently to generate a significant velocity offset. However, a massive black hole with a nonthermal Compton-thick spectrum is able to reproduce the Lyα signatures as a result of higher photon trapping and longer potential lifetime. For both sources, Lyα radiation pressure turns out to be dynamically important.

preprint2016arXiv

Kac's Walk on $n$-sphere mixes in $n\log n$ steps

Determining the mixing time of Kac's random walk on the sphere $\mathrm{S}^{n-1}$ is a long-standing open problem. We show that the total variation mixing time of Kac's walk on $\mathrm{S}^{n-1}$ is between $\frac{1}{2} \, n \log(n)$ and $200 \,n \log(n)$. Our bound is thus optimal up to a constant factor, improving on the best-known upper bound of $O(n^{5} \log(n)^{2})$ due to Jiang. Our main tool is a `non-Markovian' coupling recently introduced by the second author for obtaining the convergence rates of certain high dimensional Gibbs samplers in continuous state spaces.

preprint2016arXiv

Lyman-alpha radiation hydrodynamics of galactic winds before cosmic reionization

The dynamical impact of Lyman-alpha (Lyα) radiation pressure on galaxy formation depends on the rate and duration of momentum transfer between Lyα photons and neutral hydrogen gas. Although photon trapping has the potential to multiply the effective force, ionizing radiation from stellar sources may relieve the Lyα pressure before appreciably affecting the kinematics of the host galaxy or efficiently coupling Lyα photons to the outflow. We present self-consistent Lyα radiation-hydrodynamics simulations of high-$z$ galaxy environments by coupling the Cosmic Lyα Transfer code (COLT) with spherically symmetric Lagrangian frame hydrodynamics. The accurate but computationally expensive Monte-Carlo radiative transfer calculations are feasible under the one-dimensional approximation. The initial starburst drives an expanding shell of gas from the centre and in certain cases Lyα feedback significantly enhances the shell velocity. Radiative feedback alone is capable of ejecting baryons into the intergalactic medium (IGM) for protogalaxies with a virial mass of $M_{\rm vir} \lesssim 10^8~{\rm M}_\odot$. We compare the Lyα signatures of Population III stars with $10^5$ K blackbody emission to that of direct collapse black holes with a nonthermal Compton-thick spectrum and find substantial differences if the Lyα spectra are shaped by gas pushed by Lyα radiation-driven winds. For both sources, the flux emerging from the galaxy is reprocessed by the IGM such that the observed Lyα luminosity is reduced significantly and the time-averaged velocity offset of the Lyα peak is shifted redward.

preprint2016arXiv

On the Mixing Time of Kac's Walk and Other High-Dimensional Gibbs Samplers with Constraints

Determining the total variation mixing time of Kac's random walk on the special orthogonal group $\mathrm{SO}(n)$ has been a long-standing open problem. In this paper, we construct a novel non-Markovian coupling for bounding this mixing time. The analysis of our coupling entails controlling the smallest singular value of a certain random matrix with highly dependent entries. The dependence of the entries in our matrix makes it not-amenable to existing techniques in random matrix theory. To circumvent this difficulty, we extend some recent bounds on the smallest singular values of matrices with independent entries to our setting. These bounds imply that the mixing time of Kac's walk on the group $\mathrm{SO}(n)$ is between $C_{1} n^{2}$ and $C_{2} n^{4} \log(n)$ for some explicit constants $0 < C_{1}, C_{2} < \infty$, substantially improving on the bound of $O(n^{5} \log(n)^{2})$ by Jiang. Our methods may also be applied to other high dimensional Gibbs samplers with constraints and thus are of independent interest. In addition to giving analytical bounds on the mixing time, our approach allows us to compute rigorous estimates of the mixing time by simulating the eigenvalues of a random matrix.

preprint2016arXiv

Parallel Markov Chain Monte Carlo via Spectral Clustering

As it has become common to use many computer cores in routine applications, finding good ways to parallelize popular algorithms has become increasingly important. In this paper, we present a parallelization scheme for Markov chain Monte Carlo (MCMC) methods based on spectral clustering of the underlying state space, generalizing earlier work on parallelization of MCMC methods by state space partitioning. We show empirically that this approach speeds up MCMC sampling for multimodal distributions and that it can be usefully applied in greater generality than several related algorithms. Our algorithm converges under reasonable conditions to an `optimal' MCMC algorithm. We also show that our approach can be asymptotically far more efficient than naive parallelization, even in situations such as completely flat target distributions where no unique optimal algorithm exists. Finally, we combine theoretical and empirical bounds to provide practical guidance on the choice of tuning parameters.

preprint2016arXiv

The Use of a Single Pseudo-Sample in Approximate Bayesian Computation

We analyze the computational efficiency of approximate Bayesian computation (ABC), which approximates a likelihood function by drawing pseudo-samples from the associated model. For the rejection sampling version of ABC, it is known that multiple pseudo-samples cannot substantially increase (and can substantially decrease) the efficiency of the algorithm as compared to employing a high-variance estimate based on a single pseudo-sample. We show that this conclusion also holds for a Markov chain Monte Carlo version of ABC, implying that it is unnecessary to tune the number of pseudo-samples used in ABC-MCMC. This conclusion is in contrast to particle MCMC methods, for which increasing the number of particles can provide large gains in computational efficiency.

preprint2015arXiv

Ergodicity of Approximate MCMC Chains with Applications to Large Data Sets

In many modern applications, difficulty in evaluating the posterior density makes performing even a single MCMC step slow. This difficulty can be caused by intractable likelihood functions, but also appears for routine problems with large data sets. Many researchers have responded by running approximate versions of MCMC algorithms. In this note, we develop quantitative bounds for showing the ergodicity of these approximate samplers. We then use these bounds to study the bias-variance trade-off of approximate MCMC algorithms. We apply our results to simple versions of recently proposed algorithms, including a variant of the "austerity" framework of Korratikara et al.

preprint2015arXiv

Mixing times for a constrained Ising process on the torus at low density

We study a kinetically constrained Ising process (KCIP) associated with a graph G and density parameter p; this process is an interacting particle system with state space $\{0,1\}^{G}$. The stationary distribution of the KCIP Markov chain is the Binomial($|G|, p$) distribution on the number of particles, conditioned on having at least one particle. The `constraint' in the name of the process refers to the rule that a vertex cannot change its state unless it has at least one neighbour in state `1'. The KCIP has been proposed by statistical physicists as a model for the glass transition, and more recently as a simple algorithm for data storage in computer networks. In this note, we study the mixing time of this process on the torus $G = \mathbb{Z}_{L}^{d}$, $d \geq 3$, in the low-density regime $p = \frac{c}{n}$ for arbitrary $0 < c < 1$; this regime is the subject of a conjecture of Aldous and is natural in the context of computer networks. Our results provide a counterexample to Aldous' conjecture, suggest a natural modifcation of the conjecture, and show that this modifcation is correct up to logarithmic factors. The methods developed in this paper also provide a strategy for tackling Aldous' conjecture for other graphs.

preprint2015arXiv

The Lyman-alpha signature of the first galaxies

We present the Cosmic Lyman-$α$ Transfer code (COLT), a massively parallel Monte-Carlo radiative transfer code, to simulate Lyman-$α$ (Ly$α$) resonant scattering through neutral hydrogen as a probe of the first galaxies. We explore the interaction of centrally produced Ly$α$ radiation with the host galactic environment. Ly$α$ photons emitted from the luminous starburst region escape with characteristic features in the line profile depending on the density distribution, ionization structure, and bulk velocity fields. For example, anisotropic ionization exhibits a tall peak close to line centre with a skewed tail that drops off gradually. Idealized models of first galaxies explore the effect of mass, anisotropic H II regions, and radiation pressure driven winds on Ly$α$ observables. We employ mesh refinement to resolve critical structures. We also post-process an ab initio cosmological simulation and examine images captured at various escape distances within the 1 Mpc$^3$ comoving volume. Finally, we discuss the emergent spectra and surface brightness profiles of these objects in the context of high-$z$ observations. The first galaxies will likely be observed through the red damping wing of the Ly$α$ line. Observations will be biased toward galaxies with an intrinsic red peak located far from line centre that reside in extensive H II super bubbles, which allows Hubble flow to sufficiently redshift photons away from line centre and facilitate transmission through the intergalactic medium (IGM). Even with gravitational lensing to boost the luminosity this preliminary work indicates that Ly$α$ emission from stellar clusters within haloes of $M_{\rm vir}<10^9~{\rm M}_\odot$ is generally too faint to be detected by the James Webb Space Telescope (JWST).

preprint2014arXiv

A Gibbs sampler on the $n$-simplex

We determine the mixing time of a simple Gibbs sampler on the unit simplex, confirming a conjecture of Aldous. The upper bound is based on a two-step coupling, where the first step is a simple contraction argument and the second step is a non-Markovian coupling. We also present a MCMC-based perfect sampling algorithm based on our proof which can be applied with Gibbs samplers that are harder to analyze.

preprint2014arXiv

Finite Sample Properties of Adaptive Markov Chains via Curvature

Adaptive Markov chains are an important class of Monte Carlo methods for sampling from probability distributions. The time evolution of adaptive algorithms depends on past samples, and thus these algorithms are non-Markovian. Although there has been previous work establishing conditions for their ergodicity, not much is known theoretically about their finite sample properties. In this paper, using a notion of discrete Ricci curvature for Markov kernels introduced by Ollivier, we establish concentration inequalities and finite sample bounds for a class of adaptive Markov chains. After establishing some general results, we give quantitative bounds for `multi-level' adaptive algorithms such as the equi-energy sampler. We also provide the first rigorous proofs that the finite sample properties of an equi-energy sampler are superior to those of related parallel tempering and Metropolis-Hastings samplers after a learning period comparable to their mixing times.

preprint2013arXiv

Comparison Theory for Markov Chains on Different State Spaces and Application to Random Walk on Derangements

Let $X_{t}$ and $Y_{t}$ be two Markov chains, on state spaces $Ω\subset \hatΩ$. In this paper, we discuss how to prove bounds on the spectrum of $X_{t}$ based on bounds on the spectrum of $Y_{t}$. This generalizes work of Diaconis, Saloff-Coste, Yuen and others on comparison of chains in the case $Ω= \hatΩ$. The main tool is the extension of functions from the smaller space to the larger, which allows comparison of the entire spectrum of the two chains. The theory is used to give quick analyses of several chains without symmetry. The main application is to a `random transposition' walk on derangements.

preprint2012arXiv

The Cutoff Phenomenon for Random Birth and Death Chains

For any distribution $π$ with support equal to $[n] = \{1, 2,..., n \}$, we study the set $\mathcal{A}_π$ of tridiagonal stochastic matrices $K$ satisfying $π(i) K[i,j] = π(j) K[j,i]$ for all $i, j \in [n]$. These matrices correspond to birth and death chains with stationary distribution $π$. We study matrices $K$ drawn uniformly from $\mathcal{A}_π$, following the work of Diaconis and Wood on the case $π(i) = \frac{1}{n}$. We analyze a `block sampler' version of their algorithm for drawing from $\mathcal{A}_π$ at random, and use results from this analysis to draw conclusions about typical matrices. The main result is a soft argument comparing cutoff for sequences of random birth and death chains to cutoff for a special family of birth and death chains with the same stationary distributions.

preprint2012arXiv

The Impact of Assuming Flatness in the Determination of Neutrino Properties from Cosmological Data

Cosmological data have provided new constraints on the number of neutrino species and the neutrino mass. However these constraints depend on assumptions related to the underlying cosmology. Since a correlation is expected between the number of effective neutrinos N_{eff}, the neutrino mass \sum m_ν, and the curvature of the universe Ω_k, it is useful to investigate the current constraints in the framework of a non-flat universe. In this paper we update the constraints on neutrino parameters by making use of the latest cosmic microwave background (CMB) data from the ACT and SPT experiments and consider the possibility of a universe with non-zero curvature. We first place new constraints on N_{eff} and Ω_k, with N_{eff} = 4.03 +/- 0.45 and 10^3 Ω_k = -4.46 +/- 5.24. Thus, even when Ω_k is allowed to vary, N_{eff} = 3 is still disfavored with 95% confidence. We then investigate the correlation between neutrino mass and curvature that shifts the 95% upper limit of \sum m_ν< 0.45 eV to \sum m_ν< 0.95 eV. Thus, the impact of assuming flatness in neutrino cosmology is significant and an essential consideration with future experiments.

preprint2011arXiv

Three-Axis Measurement and Cancellation of Background Magnetic Fields to less than 50 uG in a Cold Atom Experiment

Many experiments involving cold and ultracold atomic gases require very precise control of magnetic fields that couple to and drive the atomic spins. Examples include quantum control of atomic spins, quantum control and quantum simulation in optical lattices, and studies of spinor Bose condensates. This makes accurate cancellation of the (generally time dependent) background magnetic field a critical factor in such experiments. We describe a technique that uses the atomic spins themselves to measure DC and AC components of the background field independently along three orthogonal axes, with a resolution of a few tens of uG in a bandwidth of ~1 kHz. Once measured, the background field can be cancelled with three pairs of compensating coils driven by arbitrary waveform generators. In our laboratory, the magnetic field environment is sufficiently stable for the procedure to reduce the field along each axis to less than ~50 uG rms, corresponding to a suppression of the AC part by about one order of magnitude. This suggests our approach can provide access to a new low-field regime in cold-atom experiments.