Source author record

Xiaohu Yang

Xiaohu Yang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

77works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

77 published item(s)

preprint2026arXiv

DepRadar: Agentic Coordination for Context Aware Defect Impact Analysis in Deep Learning Libraries

Deep learning libraries like Transformers and Megatron are now widely adopted in modern AI programs. However, when these libraries introduce defects, ranging from silent computation errors to subtle performance regressions, it is often challenging for downstream users to assess whether their own programs are affected. Such impact analysis requires not only understanding the defect semantics but also checking whether the client code satisfies complex triggering conditions involving configuration flags, runtime environments, and indirect API usage. We present DepRadar, an agent coordination framework for fine grained defect and impact analysis in DL library updates. DepRadar coordinates four specialized agents across three steps: 1. the PR Miner and Code Diff Analyzer extract structured defect semantics from commits or pull requests, 2. the Orchestrator Agent synthesizes these signals into a unified defect pattern with trigger conditions, and 3. the Impact Analyzer checks downstream programs to determine whether the defect can be triggered. To improve accuracy and explainability, DepRadar integrates static analysis with DL-specific domain rules for defect reasoning and client side tracing. We evaluate DepRadar on 157 PRs and 70 commits across two representative DL libraries. It achieves 90% precision in defect identification and generates high quality structured fields (average field score 1.6). On 122 client programs, DepRadar identifies affected cases with 90% recall and 80% precision, substantially outperforming other baselines.

preprint2026arXiv

Safety at One Shot: Patching Fine-Tuned LLMs with A Single Instance

Fine-tuning safety-aligned large language models (LLMs) can substantially compromise their safety. Previous approaches require many safety samples or calibration sets, which not only incur significant computational overhead during realignment but also lead to noticeable degradation in model utility. Contrary to this belief, we show that safety alignment can be fully recovered with only a single safety example, without sacrificing utility and at minimal cost. Remarkably, this recovery is effective regardless of the number of harmful examples used in fine-tuning or the size of the underlying model, and convergence is achieved within just a few epochs. Furthermore, we uncover the low-rank structure of the safety gradient, which explains why such efficient correction is possible. We validate our findings across five safety-aligned LLMs and multiple datasets, demonstrating the generality of our approach.

preprint2026arXiv

SolContractEval: A Benchmark for Evaluating Contract-Level Solidity Code Generation

The rise of blockchain has brought smart contracts into mainstream use, creating a demand for smart contract generation tools. While large language models (LLMs) excel at generating code in general-purpose languages, their effectiveness on Solidity, the primary language for smart contracts, remains underexplored. Solidity constitutes only a small portion of typical LLM training data and differs from general-purpose languages in its version-sensitive syntax and limited flexibility. These factors raise concerns about the reliability of existing LLMs for Solidity code generation. Critically, existing evaluations, focused on isolated functions and synthetic inputs, fall short of assessing models' capabilities in real-world contract development. To bridge this gap, we introduce SolContractEval, the first contract-level benchmark for Solidity code generation. It comprises 124 tasks drawn from real on-chain contracts across nine major domains. Each task input, consisting of complete context dependencies, a structured contract framework, and a concise task prompt, is independently annotated and cross-validated by experienced developers. To enable precise and automated evaluation of functional correctness, we also develop a dynamic evaluation framework based on historical transaction replay. Building on SolContractEval, we perform a systematic evaluation of six mainstream LLMs. We find that Claude-3.7-Sonnet achieves the highest overall performance, though evaluated models underperform relative to their capabilities on class-level generation tasks in general-purpose programming languages. Second, current models perform better on tasks that follow standard patterns but struggle with complex logic and inter-contract dependencies. Finally, they exhibit limited understanding of Solidity-specific features and contextual dependencies.

preprint2026arXiv

Understanding and Preserving Safety in Fine-Tuned LLMs

Fine-tuning is an essential and pervasive functionality for applying large language models (LLMs) to downstream tasks. However, it has the potential to substantially degrade safety alignment, e.g., by greatly increasing susceptibility to jailbreak attacks, even when the fine-tuning data is entirely harmless. Despite garnering growing attention in defense efforts during the fine-tuning stage, existing methods struggle with a persistent safety-utility dilemma: emphasizing safety compromises task performance, whereas prioritizing utility typically requires deep fine-tuning that inevitably leads to steep safety declination. In this work, we address this dilemma by shedding new light on the geometric interaction between safety- and utility-oriented gradients in safety-aligned LLMs. Through systematic empirical analysis, we uncover three key insights: (I) safety gradients lie in a low-rank subspace, while utility gradients span a broader high-dimensional space; (II) these subspaces are often negatively correlated, causing directional conflicts during fine-tuning; and (III) the dominant safety direction can be efficiently estimated from a single sample. Building upon these novel insights, we propose safety-preserving fine-tuning (SPF), a lightweight approach that explicitly removes gradient components conflicting with the low-rank safety subspace. Theoretically, we show that SPF guarantees utility convergence while bounding safety drift. Empirically, SPF consistently maintains downstream task performance and recovers nearly all pre-trained safety alignment, even under adversarial fine-tuning scenarios. Furthermore, SPF exhibits robust resistance to both deep fine-tuning and dynamic jailbreak attacks. Together, our findings provide new mechanistic understanding and practical guidance toward always-aligned LLM fine-tuning.

preprint2025arXiv

Introduction to the Chinese Space Station Survey Telescope (CSST)

The Chinese Space Station Survey Telescope (CSST) is an upcoming Stage-IV sky survey telescope, distinguished by its large field of view (FoV), high image quality, and multi-band observation capabilities. It can simultaneously conduct precise measurements of the Universe by performing multi-color photometric imaging and slitless spectroscopic surveys. The CSST is equipped with five scientific instruments, i.e. Multi-band Imaging and Slitless Spectroscopy Survey Camera (SC), Multi-Channel Imager (MCI), Integral Field Spectrograph (IFS), Cool Planet Imaging Coronagraph (CPI-C), and THz Spectrometer (TS). Using these instruments, CSST is expected to make significant contributions and discoveries across various astronomical fields, including cosmology, galaxies and active galactic nuclei (AGN), the Milky Way and nearby galaxies, stars, exoplanets, Solar System objects, astrometry, and transients and variable sources. This review aims to provide a comprehensive overview of the CSST instruments, observational capabilities, data products, and scientific potential.

preprint2022arXiv

\textsc{The Three Hundred} project: The \textsc{Gizmo-Simba} run

We introduce \textsc{Gizmo-Simba}, a new suite of galaxy cluster simulations within \textsc{The Three Hundred} project. \textsc{The Three Hundred} consists of zoom re-simulations of 324 clusters with $M_{200}\gtrsim 10^{14.8}M_\odot$ drawn from the MultiDark-Planck $N$-body simulation, run using several hydrodynamic and semi-analytic codes. The \textsc{Gizmo-Simba} suite adds a state-of-the-art galaxy formation model based on the highly successful {\sc Simba} simulation, mildly re-calibrated to match $z=0$ cluster stellar properties. Comparing to \textsc{The Three Hundred} zooms run with \textsc{Gadget-X}, we find intrinsic differences in the evolution of the stellar and gas mass fractions, BCG ages, and galaxy colour-magnitude diagrams, with \textsc{Gizmo-Simba} generally providing a good match to available data at $z \approx 0$. \textsc{Gizmo-Simba}'s unique black hole growth and feedback model yields agreement with the observed BH scaling relations at the intermediate-mass range and predicts a slightly different slope at high masses where few observations currently lie. \textsc{Gizmo-Simba} provides a new and novel platform to elucidate the co-evolution of galaxies, gas, and black holes within the densest cosmic environments.

preprint2022arXiv

An Extended Halo-based Group/Cluster finder: application to the DESI legacy imaging surveys DR8

We extend the halo-based group finder developed by \citet[][]{Yang2005a} to use data {\it simultaneously} with either photometric or spectroscopic redshifts. A mock galaxy redshift survey constructed from a high-resolution N-body simulation is used to evaluate the performance of this extended group finder. For galaxies with magnitude ${\rm z\le 21}$ and redshift $0<z\le 1.0$ in the DESI legacy imaging surveys (the Legacy Surveys), our group finder successfully identifies more than 60\% of the members in about $90\%$ of halos with mass $\ga 10^{12.5}\msunh$. Detected groups with mass $\ga 10^{12.0}\msunh$ have a purity (the fraction of true groups) greater than 90\%. The halo mass assigned to each group has an uncertainty of about 0.2 dex at the high mass end $\ga 10^{13.5}\msunh$ and 0.40 dex at the low mass end. Groups with more than 10 members have a redshift accuracy of $\sim 0.008$. We apply this group finder to the Legacy Surveys DR8 and find 5.2 Million groups with at least 3 members. About 387,000 of these groups have at least 10 members. The resulting catalog containing 3D coordinates, richness, halo masses, and total group luminosities, is made publicly available.

preprint2022arXiv

Cross-correlation of Planck CMB lensing with DESI galaxy groups

We measure the cross-correlation between galaxy groups constructed from DESI Legacy Imaging Survey DR8 and \emph{Planck} CMB lensing, over overlapping sky area of 16876 $\rm deg^2$. The detections are significant and consistent with the expected signal of the large-scale structure of the universe, over group samples of various redshift, mass, richness $N_{\rm g}$ and over various scale cuts. The overall S/N is 40 for a conservative sample with $N_{\rm g}\geq 5$, and increases to $50$ for the sample with $N_{\rm g}\geq 2$. Adopting the \emph{Planck} 2018 cosmology, we constrain the density bias of groups with $N_{\rm g}\geq 5$ as $b_{\rm g}=1.31\pm 0.10$, $2.22\pm 0.10$, $3.52\pm 0.20$ at $0.1<z\leq 0.33$, $0.33<z\leq 0.67$, $0.67<z\leq1$ respectively. The group catalog provides the estimation of group halo mass and therefore allows us to detect the dependence of bias on group mass with high significance. It also allows us to compare the measured bias with the theoretically predicted one using the estimated group mass. We find excellent agreement for the two high redshift bins. However, it is lower than the theory by $\sim 3σ$ for the lowest redshift bin. Another interesting finding is the significant impact of the thermal Sunyaev Zel'dovich (tSZ). It contaminates the galaxy group-CMB lensing cross-correlation at $\sim 30\%$ level, and must be deprojected first in CMB lensing reconstruction.

preprint2022arXiv

Defect Identification, Categorization, and Repair: Better Together

Just-In-Time defect prediction (JIT-DP) models can identify defect-inducing commits at check-in time. Even though previous studies have achieved a great progress, these studies still have the following limitations: 1) useful information (e.g., semantic information and structure information) are not fully used; 2) existing work can only predict a commit as buggy one or clean one without more information about what type of defect it is; 3) a commit may involve changes in many files, which cause difficulty in locating the defect; 4) prior studies treat defect identification and defect repair as separate tasks, none aims to handle both tasks simultaneously. In this paper, to handle aforementioned limitations, we propose a comprehensive defect prediction and repair framework named CompDefect, which can identify whether a changed function (a more fine-grained level) is defect-prone, categorize the type of defect, and repair such a defect automatically if it falls into several scenarios, e.g., defects with single statement fixes, or those that match a small set of defect templates. Generally, the first two tasks in CompDefect are treated as a multiclass classification task, while the last one is treated as a sequence generation task. The whole input of CompDefect consists of three parts (exampled with positive functions): the clean version of a function (i.e., the version before defect introduced), the buggy version of a function and the fixed version of a function. In multiclass classification task, CompDefect categorizes the type of defect via multiclass classification with the information in both the clean version and the buggy version. In code sequence generation task, CompDefect repairs the defect once identified or keeps it unchanged.

preprint2022arXiv

ELUCID VII: Using Constrained Hydro Simulations to Explore the Gas Component of the Cosmic Web

Using reconstructed initial conditions in the SDSS survey volume, we carry out constrained hydrodynamic simulations in three regions representing different types of the cosmic web: the Coma cluster of galaxies; the SDSS great wall; and a large low-density region at $z\sim 0.05$. These simulations, which include star formation and stellar feedback but no AGN formation and feedback, are used to investigate the properties and evolution of intergalactic and intra-cluster media. About half of the warm-hot intergalactic gas is associated with filaments in the local cosmic web. Gas in the outskirts of massive filaments and halos can be heated significantly by accretion shocks generated by mergers of filaments and halos, respectively, and there is a tight correlation between gas temperature and the strength of the local tidal field. The simulations also predict some discontinuities associated with shock fronts and contact edges, which can be tested using observations of the thermal SZ effect and X-rays. A large fraction of the sky is covered by Ly$α$ and OVI absorption systems, and most of the OVI systems and low-column density HI systems are associated with filaments in the cosmic web. The constrained simulations, which follow the formation and heating history of the observed cosmic web, provide an important avenue to interpret observational data. With full information about the origin and location of the cosmic gas to be observed, such simulations can also be used to develop observational strategies.

preprint2022arXiv

Elucidating Galaxy Assembly Bias in SDSS

We investigate the level of galaxy assembly bias in the Sloan Digital Sky Survey (SDSS) main galaxy sample using ELUCID, a state-of-the-art constrained simulation that accurately reconstructed the initial density perturbations within the SDSS volume. On top of the ELUCID haloes, we develop an extended HOD model that includes the assembly bias of central and satellite galaxies, parameterized as $\mathcal{Q}_\mathrm{cen}$ and $\mathcal{Q}_\mathrm{sat}$, respectively, to predict a suite of one- and two-point observables. In particular, our fiducial constraint employs the probability distribution of the galaxy number counts measured on $8\,\mathrm{Mpc}\,h^{-1}$ scales $N_8^g$ and the projected cross-correlation functions of quintiles of galaxies selected by $N_8^g$ with our entire galaxy sample. We perform extensive tests of the efficacy of our method by fitting the same observables to mock data using both constrained and non-constrained simulations. We discover that in many cases the level of cosmic variance between the two simulations can produce biased constraints that lead to an erroneous detection of galaxy assembly bias if the non-constrained simulation is used. When applying our method to the SDSS data, the ELUCID reconstruction effectively removes an otherwise strong degeneracy between cosmic variance and galaxy assembly bias in SDSS, enabling us to derive an accurate and stringent constraint on the latter. Our fiducial ELUCID constraint, for galaxies above a stellar mass threshold $M_*{=}10^{10.2}\,h^{-2}\,M_\odot$, is $\mathcal{Q}_\mathrm{cen}{=}{-}0.09\pm{0.05}$ and $\mathcal{Q}_\mathrm{sat}{=}0.09\pm{0.10}$, indicating no evidence for a significant~($>2σ$) galaxy assembly bias in the local Universe probed by SDSS. Finally, our method provides a promising path to the robust modelling of the galaxy-halo connection within future surveys like DESI and PFS.

preprint2022arXiv

First measurement of the characteristic depletion radius of dark matter haloes from weak lensing

We use weak lensing observations to make the first measurement of the characteristic depletion radius, one of the three radii that characterize the region where matter is being depleted by growing haloes. The lenses are taken from the halo catalog produced by the extended halo-based group/cluster finder applied to DESI Legacy Imaging Surveys DR9, while the sources are extracted from the DECaLS DR8 imaging data with the Fourier_Quad pipeline. We study halo masses $12 < \log ( M_{\rm grp} ~[{\rm M_{\odot}}/h] ) \leq 15.3$ within redshifts $0.2 \leq z \leq 0.3$. The virial and splashback radii are also measured and used to test the original findings on the depletion region. When binning haloes by mass, we find consistency between most of our measurements and predictions from the CosmicGrowth simulation, with exceptions to the lowest mass bins. The characteristic depletion radius is found to be roughly $2.5$ times the virial radius and $1.7 - 3$ times the splashback radius, in line with an approximately universal outer density profile, and the average enclosed density within the characteristic depletion radius is found to be roughly $29$ times the mean matter density of the Universe in our sample. When binning haloes by both mass and a proxy for halo concentration, we do not detect a significant variation of the depletion radius with concentration, on which the simulation prediction is also sensitive to the choice of concentration proxy. We also confirm that the measured splashback radius varies with concentration differently from simulation predictions.

preprint2022arXiv

Groups and protocluster candidates in the CLAUDS and HSC-SSP joint deep surveys

Using the extended halo-based group finder developed by Yang et al. (2021), which is able to deal with galaxies via spectroscopic and photometric redshifts simultaneously, we construct galaxy group and candidate protocluster catalogs in a wide redshift range ($0 < z < 6$) from the joint CFHT Large Area $U$-band Deep Survey (CLAUDS) and Hyper Suprime-Cam Subaru Strategic Program (HSC-SSP) deep data set. Based on a selection of 5,607,052 galaxies with $i$-band magnitude $m_{i} < 26$ and a sky coverage of $34.41\ {\rm deg}^2$, we identify a total of 2,232,134 groups, within which 402,947 groups have at least three member galaxies. We have visually checked and discussed the general properties of those richest groups at redshift $z>2.0$. By checking the galaxy number distributions within a $5-7\ h^{-1}\mathrm{Mpc}$ projected separation and a redshift difference $Δz \le 0.1$ around those richest groups at redshift $z>2$, we identified a list of 761, 343 and 43 protocluster candidates in the redshift bins $2\leq z<3$, $3\leq z<4$ and $z \geq 4$, respectively. In general, these catalogs of galaxy groups and protocluster candidates will provide useful environmental information in probing galaxy evolution along the cosmic time.

preprint2022arXiv

Massive Star-Forming Galaxies Have Converted Most of Their Halo Gas into Stars

In the local Universe, the efficiency for converting baryonic gas into stars is very low. In dark matter halos where galaxies form and evolve, the average efficiency varies with galaxy stellar mass and has a maximum of about twenty percent for Milky-Way-like galaxies. The low efficiency at higher mass is believed to be produced by some quenching processes, such as the feedback from active galactic nuclei. We perform an analysis of weak lensing and satellite kinematics for SDSS central galaxies. Our results reveal that the efficiency is much higher, more than sixty percent, for a large population of massive star-forming galaxies around $10^{11}M_{\odot}$. This suggests that these galaxies acquired most of the gas in their halos and converted it into stars without being affected significantly by quenching processes. This population of galaxies is not reproduced in current galaxy formation models, indicating that our understanding of galaxy formation is incomplete. The implications of our results on circumgalactic media, star formation quenching and disc galaxy rotation curves are discussed. We also examine systematic uncertainties in halo-mass and stellar-mass measurements that might influence our results.

preprint2022arXiv

The Universal Specific Merger Rate of Dark Matter Halos

We employ a set of high resolution N-body simulations to study the merger rate of dark matter halos. We define a specific merger rate by normalizing the average number of mergers per halo with the logarithmic mass growth change of the hosts at the time of accretion. Based on the simulation results, we find that this specific merger rate, $\mathrm{d}N_{\mathrm{merge}}(ξ|M,z)/\mathrm{d}ξ/\mathrm{d}\log M(z)$, has a universal form, which is only a function of the mass ratio of merging halo pairs, $ξ$, and does not depend on the host halo mass, $M$, or redshift, $z$, over a wide range of masses ($10^{12}\lesssim M \lesssim10^{14}\,M_\odot/h$) and merger ratios ($ξ\ge 1e-2$). We further test with simulations of different $Ω_m$ and $σ_8$, and get the same specific merger rate. The universality of the specific merger rate shows that halos in the universe are built up self-similarly, with a universal composition in the mass contributions and an absolute merger rate that grows in proportion to the halo mass growth. As a result, the absolute merger rate relates with redshift and cosmology only through the halo mass variable, whose evolution can be readily obtained from the universal mass accretion history (MAH) model of \cite{2009ApJ...707..354Z}. Lastly, we show that this universal specific merger rate immediately predicts an universal un-evolved subhalo mass function that is independent on the redshift, MAH or the final halo mass, and vice versa.

preprint2022arXiv

What to expect from dynamical modelling of cluster haloes II. Investigating dynamical state indicators with Random Forest

We investigate the importances of various dynamical features in predicting the dynamical state (DS) of galaxy clusters, based on the Random Forest (RF) machine learning approach. We use a large sample of galaxy clusters from the Three Hundred Project of hydrodynamical zoomed-in simulations, and construct dynamical features from the raw data as well as from the corresponding mock maps in the optical, X-ray, and Sunyaev-Zel'dovich (SZ) channels. Instead of relying on the impurity based feature importance of the RF algorithm, we directly use the out-of-bag (OOB) scores to evaluate the importances of individual features and different feature combinations. Among all the features studied, we find the virial ratio, $η$, to be the most important single feature. The features calculated directly from the simulations and in 3-dimensions carry more information on the DS than those constructed from the mock maps. Compared with the features based on X-ray or SZ maps, features related to the centroid positions are more important. Despite the large number of investigated features, a combination of up to three features of different types can already saturate the score of the prediction. Lastly, we show that the most sensitive feature $η$ is strongly correlated with the well-known half-mass bias in dynamical modelling. Without a selection in DS, cluster halos have an asymmetric distribution in $η$, corresponding to an overall positive half-mass bias. Our work provides a quantitative reference for selecting the best features to discriminate the DS of galaxy clusters in both simulations and observations.

preprint2021arXiv

The Color Gradients of the Globular Cluster Systems in M87 and M49

Combining data from the ACS Virgo Cluster Survey (ACSVCS) and the Next Generation Virgo cluster Survey (NGVS), we extend previous studies of color gradients of the globular cluster (GC) systems of the two most massive galaxies in the Virgo cluster, M87 and M49, to radii of $\sim 15~R_e$ ($\sim 200$ kpc for M87 and $\sim 250$ kpc for M49). We find significant negative color gradients, i.e., becoming bluer with increasing distance, out to these large radii. The gradients are driven mainly by the outwards decrease of the ratio of red to blue GC numbers. The color gradients are also detected out to $\sim 15~R_e$ in the red and blue sub-populations of GCs taken separately. In addition, we find a negative color gradient when we consider the satellite low-mass elliptical galaxies as a system, i.e., the satellite galaxies closer to the center of the host galaxy usually have redder color indices, both for their stars and GCs. According to the "two phase" formation scenario of massive early-type galaxies, the host galaxy accretes stars and GCs from low-mass satellite galaxies in the second phase. So the accreted GC system naturally inherits the negative color gradient present in the satellite population. This can explain why the color gradient of the GC system can still be observed at large radii after multiple minor mergers.

preprint2020arXiv

Detection of missing baryons in galaxy groups with kinetic Sunyaev-Zel'dovich effect

We present the detection of the kinetic Sunyaev-Zel'dovich effect (kSZE) signals from groups of galaxies as a function of halo mass down to $\log (M_{500}/{\rm M_\odot}) \sim 12.3$, using the {\it Planck} CMB maps and stacking about $40,000$ galaxy systems with known positions, halo masses, and peculiar velocities. The signals from groups of different mass are constrained simultaneously to take care of projection effects of nearby halos. The total kSZE flux within halos estimated implies that the gas fraction in halos is about the universal baryon fraction, even in low-mass halos, indicating that the `missing baryons' are found. Various tests performed show that our results are robust against systematic effects, such as contamination by infrared/radio sources and background variations, beam-size effects and contributions from halo exteriors. Combined with the thermal Sunyaev-Zel'dovich effect, our results indicate that the `missing baryons' associated with galaxy groups are contained in warm-hot media with temperatures between $10^5$ and $10^6\,{\rm K}$.

preprint2020arXiv

Detection of missing baryons in galaxy groups with kinetic Sunyaev-Zel'dovich effect

We present the detection of the kinetic Sunyaev-Zel'dovich effect (kSZE) signals from groups of galaxies as a function of halo mass down to $\log (M_{500}/{\rm M_\odot}) \sim 12.3$, using the {\it Planck} CMB maps and stacking about $40,000$ galaxy systems with known positions, halo masses, and peculiar velocities. The signals from groups of different mass are constrained simultaneously to take care of projection effects of nearby halos. The total kSZE flux within halos estimated implies that the gas fraction in halos is about the universal baryon fraction, even in low-mass halos, indicating that the `missing baryons' are found. Various tests performed show that our results are robust against systematic effects, such as contamination by infrared/radio sources and background variations, beam-size effects and contributions from halo exteriors. Combined with the thermal Sunyaev-Zel'dovich effect, our results indicate that the `missing baryons' associated with galaxy groups are contained in warm-hot media with temperatures between $10^5$ and $10^6\,{\rm K}$.

preprint2020arXiv

Observing the Effects of Galaxy Interactions on the Circumgalactic Medium

We continue our empirical study of the emission line flux originating in the cool ($T\sim10^4$ K) gas that populates the halos of galaxies and their environments. Specifically, we present results obtained for a sample of galaxy pairs with a range of projected separations, {\bf $10 < {S_p/\rm kpc} < 200$}, and mass ratios $<$ 1:5, intersected by 5,443 SDSS lines of sight at projected radii of 10 to 50 kpc from either or both of the two galaxies. We find significant enhancement in H$α$ emission and a moderate enhancement in [N {\small II}]6583 emission for low mass pairs (mean stellar mass per galaxy, $\overline{\rm M}_*, <10^{10.4} {\rm M}_\odot$) relative to the results from a control sample. This enhanced H$α$ emission comes almost entirely from sight lines located between the galaxies, consistent with a short-term, interaction-driven origin for the enhancement. We find no enhancement in H$α$ emission, but significant enhancement in [N {\small II}]6583 emission for high mass ($\overline{\rm M}_* >10^{10.4}{\rm M}_\odot$) pairs. Furthermore, we find a dependence of the emission line properties on the galaxy pair mass ratio such that those with a mass ratio below 1:2.5 have enhanced [N {\small II}]6583 and those with a mass ratio between 1:2.5 and 1:5 do not. In all cases, departures from the control sample are only detected for close pairs ($S_p <$ 100 kpc). Attributing an elevated [N {\small II}]6583/H$α$ ratio to shocks, we infer that shocks play a role in determining the CGM properties for close pairs that are among the more massive and have mass ratios closer to 1:1.

preprint2020arXiv

Populating HI gas in dark matter halos: I. method

We combine data from the Sloan Digital Sky Survey (SDSS) and the Arecibo Legacy Fast ALFA Survey (ALFALFA) to establish an empirical model for the HI gas content within dark matter halos. A cross-match between our SDSS DR7 galaxy group sample and the ALFALFA HI sources provides a catalog of 16,520 HI-galaxy pairs within 14,270 galaxy groups (halos). Using these matched pairs, we model the HI gas mass distributions within halos using two components: 1) {\it in situ} galaxy relations that involve the HI masses, colors $({\rm g-r})$ and stellar masses 2) an {\it ex situ} dependence of the HI mass on the halo mass/environment. We find that if we solely use galaxy associated scaling relations to predict the HI gas distribution (solely component 1), the number of HI detections is significantly over-predicted with respect the ALFALFA observations. We introduce a concept for the survival of the HI masses/members within halos of different masses labelled as the `efficiency' factor, in order to describe the probability that a halo has in retaining its HI detections. Taking the above consideration into account we construct a `halo based HI mass model' which does not only predict the HI masses of galaxies, but also yields similar number, stellar, halo mass and satellite fraction distributions to the HI detections retrieved from observational data.

preprint2020arXiv

Predictive Models in Software Engineering: Challenges and Opportunities

Predictive models are one of the most important techniques that are widely applied in many areas of software engineering. There have been a large number of primary studies that apply predictive models and that present well-preformed studies and well-desigeworks in various research domains, including software requirements, software design and development, testing and debugging and software maintenance. This paper is a first attempt to systematically organize knowledge in this area by surveying a body of 139 papers on predictive models. We describe the key models and approaches used, classify the different models, summarize the range of key application areas, and analyze research results. Based on our findings, we also propose a set of current challenges that still need to be addressed in future work and provide a proposed research road map for these opportunities.

preprint2020arXiv

Probing Primordial Chirality with Galaxy Spins

Chiral symmetry is maximally violated in weak interactions, and such microscopic asymmetries in the early Universe might leave observable imprints on astrophysical scales without violating the cosmological principle. In this Letter, we propose a helicity measurement to detect primordial chiral violation. We point out that observations of halo-galaxy angular momentum directions (spins), which are frozen in during the galaxy formation process, provide a fossil chiral observable. From the clustering mode of large scale structure of the Universe, we construct a spin mode in Lagrangian space and show in simulations that it is a good probe of halo-galaxy spins. In standard model, a strong symmetric correlation between the left and right helical components of this spin mode and galaxy spins is expected. Measurements of these correlations will be sensitive to chiral breaking, providing a direct test of chiral symmetry breaking in the early Universe.

preprint2020arXiv

Relating the structure of dark matter halos to their assembly and environment

We use a large $N$-body simulation to study the relation of the structural properties of dark matter halos to their assembly history and environment. The complexity of individual halo assembly histories can be well described by a small number of principal components (PCs), which, compared to formation times, provide a more complete description of halo assembly histories and have a stronger correlation with halo structural properties. Using decision trees built with the random ensemble method, we find that about $60\%$, $10\%$, and $20\%$ of the variances in halo concentration, axis ratio, and spin, respectively, can be explained by combining four dominating predictors: the first PC of the assembly history, halo mass, and two environment parameters. Halo concentration is dominated by halo assembly. The local environment is found to be important for the axis ratio and spin but is degenerate with halo assembly. The small percentages of the variance in the axis ratio and spin that are explained by known assembly and environmental factors suggest that the variance is produced by many nuanced factors and should be modeled as such. The relations between halo intrinsic properties and environment are weak compared to their variances, with the anisotropy of the local tidal field having the strongest correlation with halo properties. Our method of dimension reduction and regression can help simplify the characterization of the halo population and clarify the degeneracy among halo properties.

preprint2020arXiv

The Breakdown Scale of HI Bias Linearity

The 21 cm intensity mapping experiments promise to obtain the large-scale distribution of HI gas at the post-reionization epoch. In order to reveal the underlying matter density fluctuations from the HI mapping, it is important to understand how HI gas traces the matter density distribution. Both nonlinear halo clustering and nonlinear effects modulating HI gas in halos may determine the scale below which the HI bias deviates from linearity. We employ three approaches to generate the mock HI density from a large-scale N-body simulation at low redshifts, and demonstrate that the assumption of HI linearity is valid at the scale corresponding to the first peak of baryon acoustic oscillations, but breaks down at $k \gtrsim 0.1\,h\, {\rm Mpc}^{-1}$. The nonlinear effects of halo clustering and HI content modulation counteract each other at small scales, and their competition results in a model-dependent "sweet-spot" redshift near $z$=1 where the HI bias is scale-independent down to small scales. We also find that the linear HI bias scales approximately linearly with redshift for $z\le 3$.

preprint2020arXiv

The intrinsic SFRF and sSFRF of galaxies: comparing SDSS observation with IllustrisTNG simulation

The star formation rate function (SFRF) and specific star formation rate function (sSFRF) from the observation are impacted by the Eddington bias, due to the uncertainties on the estimated SFR. We develop a novel method to correct the Eddington bias and obtained the intrinsic SFRF and sSFRF from the Sloan Digital Sky Survey Data Release 7. The intrinsic SFRF is in good agreement with measurements from previous data in the literature that relied on UV SFRs but its high star-forming end is slightly lower than those IR and radio tracers. We demonstrate that the intrinsic sSFRF from SDSS has a bi-modal form with the one peak found at ${\rm sSFR \sim 10^{-9.7} yr^{-1}}$ representing the star-forming objects while the other peak is found at ${\rm sSFR \sim 10^{-12} yr^{-1}}$ representing the quenched population. Furthermore, we compare our observations with the predictions from the IllustrisTNG and Illustris simulations and show that the ``TNG'' model performs much better than its predecessor. However, we show that the simulated SFRF and cosmic star formation density (CSFRD) of TNG simulations are highly dependent on resolution, reflecting the limitations of the model and today state-of-the-art simulations. We demonstrate that the bi-modal, two peaked sSFRF implied by the SDSS observations does not appear in TNG regardless of the adopted box-size or resolution. This tension reflects the need for inclusion of an additional efficient quenching mechanism to the TNG model.

preprint2020arXiv

The parameter-free Finger-Of-God model and its application to 21cm intensity mapping

Using the galaxy catalog built from ELUCID N-body simulation and the semi-analytical galaxy formation model, we have built a mock HI intensity mapping map. We have implemented the Finger-of-God (FoG) effect in the map by considering the galaxy HI gas velocity dispersion. By comparing the HI power spectrum in the redshift space with the measurement from IllustrisTNG simulation, we have found that such FoG effect can explain the discrepancy between current mock map built from N-body simulation and Illustris TNG simulation. Then we built a parameter-free FoG model and a shot-noise model to calculate the HI power spectrum. We found that our model can accurately fit both the monopole and quadrupole moments of the HI matter power spectrum. Our method of building the mock HI intensity map and the parameter-free FoG model will be widely useful for the up-coming 21cm intensity mapping experiments, such as CHIME, Tianlai, BINGO, FAST and SKA. It is also crucial for us to study the non-linear effects in 21cm intensity mapping.

preprint2020arXiv

The Three Hundred Project: the stellar and gas profiles

Using the catalogues of galaxy clusters from The Three Hundred project, modelled with both hydrodynamic simulations, (Gadget-X and Gadget-MUSIC), and semi-analytic models (SAMs), we study the scatter and self-similarity of the profiles and distributions of the baryonic components of the clusters: the stellar and gas mass, metallicity, the stellar age, gas temperature, and the (specific) star formation rate. Through comparisons with observational results, we find that the shape and the scatter of the gas density profiles matches well the observed trends including the reduced scatter at large radii which is a signature of self-similarity suggested in previous studies. One of our simulated sets, Gadget-X, reproduces well the shape of the observed temperature profile, while Gadget-MUSIC has a higher and flatter profile in the cluster centre and a lower and steeper profile at large radii. The gas metallicity profiles from both simulation sets, despite following the observed trend, have a relatively lower normalisation. The cumulative stellar density profiles from SAMs are in better agreement with the observed result than both hydrodynamic simulations which show relatively higher profiles. The scatter in these physical profiles, especially in the cluster centre region, shows a dependence on the cluster dynamical state and on the cool-core/non-cool-core dichotomy. The stellar age, metallicity and (s)SFR show very large scatter, which are then presented in 2D maps. We also do not find any clear radial dependence of these properties. However, the brightest central galaxies have distinguishable features compared to the properties of the satellite galaxies.

preprint2020arXiv

UV & U-band luminosity functions from CLAUDS and HSC-SSP -- I. Using four million galaxies to simultaneously constrain the very faint and bright regimes to $z \sim 3$

We constrain the rest-frame FUV (1546Å), NUV (2345Å) and U-band (3690Å) luminosity functions (LFs) and luminosity densities (LDs) with unprecedented precision from $z\sim0.2$ to $z\sim3$ (FUV, NUV) and $z\sim2$ (U-band). Our sample of over 4.3 million galaxies, selected from the CFHT Large Area $U$-band Deep Survey (CLAUDS) and HyperSuprime-Cam Subaru Strategic Program (HSC-SSP) data lets us probe the very faint regime (down to $M_\mathrm{FUV},M_\mathrm{NUV},M_\mathrm{U} \simeq -15$ at low redshift) while simultaneously detecting very rare galaxies at the bright end down to comoving densities $<10^{-5}$ Mpc$^{-3}$. Our FUV and NUV LFs are well fitted by single Schechter functions, with faint-end slopes that are very stable up to $z\sim2$. We confirm, but self-consistently and with much better precision than previous studies, that the LDs at all three wavelengths increase rapidly with lookback time to $z\sim1$, and then much more slowly at $1<z<2$--$3$. Evolution of the FUV and NUV LFs and LDs at $z<1$ is driven almost entirely by the fading of the characteristic magnitude, $M^\star_{UV}$, while at $z>1$ it is due to the evolution of both $M^\star_{UV}$ and the characteristic number density $ϕ^\star_{UV}$. In contrast, the U-band LF has an excess of faint galaxies and is fitted with a double-Schechter form; $M^\star_\mathrm{U}$, both $ϕ^\star_\mathrm{U}$ components, and the bright-end slope evolve throughout $0.2<z<2$, while the faint-end slope is constant over at least the measurable $0.05<z<0.6$. We present tables of our Schechter parameters and LD measurements that can be used for testing theoretical galaxy evolution models and forecasting future observations.

preprint2019arXiv

The Dearth of Difference between Central and Satellite Galaxies III. Environmental Dependence of Mass-Size and Mass-Structure Relations

As demonstrated in Paper I, the quenching properties of central and satellite galaxies are quite similar as long as both stellar mass and halo mass are controlled. Here we extend the analysis to the size and bulge-to-total light ratio (B/T) of galaxies. In general central galaxies have size-stellar mass and B/T-stellar mass relations different from satellites. However, the differences are eliminated when halo mass is controlled. We also study the dependence of size and B/T on halo-centric distance and find a transitional stellar mass (M$_{*,t}$) at given halo mass (M$_h$), which is about one fifth of the mass of the central galaxies in halos of mass M$_h$. The transitional stellar masses for size, B/T and quenched fraction are similar over the whole halo mass range, suggesting a connection between the quenching of star formation and the structural evolution of galaxies. Our analysis further suggests that the classification based on the transitional stellar mass is more fundamental than the central-satellite dichotomy, and provide a more reliable way to understand the environmental effects on galaxy properties. We compare the observational results with the hydro-dynamical simulation, EAGLE and the semi-analytic model, L-GALAXIES. The EAGLE simulation successfully reproduces the similarities of size for centrals and satellites and even M$_{*,t}$, while L-GALAXIES fails to recover the observational results.

preprint2019arXiv

Toward accurate measurement of property-dependent galaxy clustering I. Comparison of the Vmax method and the "shuffled" method

Galaxy clustering provides insightful clues to our understanding of galaxy formation and evolution, as well as the universe. The redshift assignment for the random sample is one of the key steps to measure the galaxy clustering accurately. In this paper, by virtue of the mock galaxy catalogs, we investigate the effect of two redshift assignment methods on the measurement of galaxy two-point correlation functions (hereafter 2PCFs), the Vmax method and the "shuffled" method. We found that the shuffled method significantly underestimates both of the projected 2PCFs and the two-dimensional 2PCFs in redshift space. While the Vmax method does not show any notable bias on the 2PCFs for volume-limited samples. For flux-limited samples, the bias produced by the Vmax method is less than half of the shuffled method on large scales. Therefore, we strongly recommend the Vmax method to assign redshifts to random samples in the future galaxy clustering analysis.

preprint2016arXiv

An empirical model to form and evolve galaxies in dark matter halos

Based on the star formation histories (SFH) of galaxies in halos of different masses, we develop an empirical model to grow galaxies in dark mattet halos. This model has very few ingredients, any of which can be associated to observational data and thus be efficiently assessed. By applying this model to a very high resolution cosmological $N$-body simulation, we predict a number of galaxy properties that are a very good match to relevant observational data. Namely, for both centrals and satellites, the galaxy stellar mass function (SMF) up to redshift $z\simeq4$ and the conditional stellar mass functions (CSMF) in the local universe are in good agreement with observations. In addition, the 2-point correlation is well predicted in the different stellar mass ranges explored by our model. Furthermore, after applying stellar population synthesis models to our stellar composition as a function of redshift, we find that the luminosity functions in $^{0.1}u$, $^{0.1}g$, $^{0.1}r$, $^{0.1}i$ and $^{0.1}z$ bands agree quite well with the SDSS observational results down to an absolute magnitude at about -17.0. The SDSS conditional luminosity functions (CLF) itself is predicted well. Finally, the cold gas is derived from the star formation rate (SFR) to predict the HI gas mass within each mock galaxy. We find a remarkably good match to observed HI-to-stellar mass ratios. These features ensure that such galaxy/gas catalogs can be used to generate reliable mock redshift surveys.

preprint2016arXiv

ELUCID - Exploring the Local Universe with reConstructed Initial Density field III: Constrained Simulation in the SDSS Volume

A method we developed recently for the reconstruction of the initial density field in the nearby Universe is applied to the Sloan Digital Sky Survey Data Release 7. A high-resolution N-body constrained simulation (CS) of the reconstructed initial condition, with $3072^3$ particles evolved in a 500 Mpc/h box, is carried out and analyzed in terms of the statistical properties of the final density field and its relation with the distribution of SDSS galaxies. We find that the statistical properties of the cosmic web and the halo populations are accurately reproduced in the CS. The galaxy density field is strongly correlated with the CS density field, with a bias that depend on both galaxy luminosity and color. Our further investigations show that the CS provides robust quantities describing the environments within which the observed galaxies and galaxy systems reside. Cosmic variance is greatly reduced in the CS so that the statistical uncertainties can be controlled effectively even for samples of small volumes.

preprint2016arXiv

Galaxy groups in the 2MASS Redshift Survey

A galaxy group catalog is constructed from the 2MASS Redshift Survey (2MRS) with the use of a halo-based group finder. The halo mass associated with a group is estimated using a `GAP' method based on the luminosity of the central galaxy and its gap with other member galaxies. Tests using mock samples shows that this method is reliable, particularly for poor systems containing only a few members. On average 80% of all the groups have completeness >0.8, and about 65% of the groups have zero contamination. Halo masses are estimated with a typical uncertainty $\sim 0.35\,{\rm dex}$. The application of the group finder to the 2MRS gives 29,904 groups from a total of 43,246 galaxies at $z \leq 0.08$, with 5,286 groups having two or more members. Some basic properties of this group catalog is presented, and comparisons are made with other groups catalogs in overlap regions. With a depth to $z\sim 0.08$ and uniformly covering about 91% of the whole sky, this group catalog provides a useful data base to study galaxies in the local cosmic web, and to reconstruct the mass distribution in the local Universe.

preprint2016arXiv

LAMOST observations in the Kepler field. Analysis of the stellar parameters measured with the LASP based on the low-resolution spectra

All of the 14 subfields of the Kepler field have been observed at least once with the Large Sky Area Multi-Object Fiber Spectroscopic Telescope (LAMOST, Xinglong Observatory, China) during the 2012-2014 observation seasons. There are 88,628 reduced spectra with SNR$_g$ (signal-to-noise ratio in g band) $\geq$ 6 after the first round (2012-2014) of observations for the LAMOST-Kepler project (LK-project). By adopting the upgraded version of the LAMOST Stellar Parameter pipeline (LASP), we have determined the atmospheric parameters ($T_{\rm eff}$ , $\log g$, and $\rm [Fe/H]$) and heliocentric radial velocity $v_{\rm rad}$ for 51,406 stars with 61,226 spectra. Compared with atmospheric parameters derived from both high-resolution spectroscopy and asteroseismology method for common stars in Huber et al. (2014), an external calibration of LASP atmospheric parameters was made, leading to the determination of external errors for the giants and dwarfs, respectively. Multiple spectroscopic observations for the same objects of the LK-project were used to estimate the internal uncertainties of the atmospheric parameters as a function of SNR$_g$ with the unbiased estimation method. The LASP atmospheric parameters were calibrated based on both the external and internal uncertainties for the giants and dwarfs, respectively. A general statistical analysis of the stellar parameters leads to discovery of 106 candidate metal-poor stars, 9 candidate very metal-poor stars, and 18 candidate high-velocity stars. Fitting formulae were obtained segmentally for both the calibrated atmospheric parameters of the LK-project and the KIC parameters with the common stars. The calibrated atmospheric parameters and radial velocities of the LK-project will be useful for studying stars in the Kepler field.

preprint2016arXiv

Mapping the real space distributions of galaxies in SDSS DR7: I. Two Point Correlation Functions

Using a method to correct redshift space distortion (RSD) for individual galaxies, we mapped the real space distributions of galaxies in the Sloan Digital Sky Survey (SDSS) Data Release 7 (DR7). We use an ensemble of mock catalogs to demonstrate the reliability of our method. Here as the first paper in a series, we mainly focus on the two point correlation function (2PCF) of galaxies. Overall the 2PCF measured in the reconstructed real space for galaxies brighter than $^{0.1}{\rm M}_r-5\log h=-19.0$ agrees with the direct measurement to an accuracy better than the measurement error due to cosmic variance, if the reconstruction uses the correct cosmology. Applying the method to the SDSS DR7, we construct a real space version of the main galaxy catalog, which contains 396,068 galaxies in the North Galactic Cap with redshifts in the range $0.01 \leq z \leq 0.12$. The Sloan Great Wall, the largest known structure in the nearby Universe, is not as dominant an over-dense structure as appears to be in redshift space. We measure the 2PCFs in reconstructed real space for galaxies of different luminosities and colors. All of them show clear deviations from single power-law forms, and reveal clear transitions from 1-halo to 2-halo terms. A comparison with the corresponding 2PCFs in redshift space nicely demonstrates how RSDs boost the clustering power on large scales (by about $40-50\%$ at scales $\sim 10 h^{-1}{\rm {Mpc}}$) and suppress it on small scales (by about $70-80\%$ at a scale of $0.3 h^{-1}{\rm {Mpc}}$).

preprint2015arXiv

Assessing Colour-dependent Occupation Statistics Inferred from Galaxy Group Catalogues

We investigate the ability of current implementations of galaxy group finders to recover colour-dependent halo occupation statistics. To test the fidelity of group catalogue inferred statistics, we run three different group finders used in the literature over a mock that includes galaxy colours in a realistic manner. Overall, the resulting mock group catalogues are remarkably similar, and most colour-dependent statistics are recovered with reasonable accuracy. However, it is also clear that certain systematic errors arise as a consequence of correlated errors in group membership determination, central/satellite designation, and halo mass assignment. We introduce a new statistic, the halo transition probability (HTP), which captures the combined impact of all these errors. As a rule of thumb, errors tend to equalize the properties of distinct galaxy populations (i.e. red vs. blue galaxies or centrals vs. satellites), and to result in inferred occupation statistics that are more accurate for red galaxies than for blue galaxies. A statistic that is particularly poorly recovered from the group catalogues is the red fraction of central galaxies as function of halo mass. Group finders do a good job in recovering galactic conformity, but also have a tendency to introduce weak conformity when none is present. We conclude that proper inference of colour-dependent statistics from group catalogues is best achieved using forward modelling (i.e., running group finders over mock data), or by implementing a correction scheme based on the HTP, as long as the latter is not too strongly model-dependent.

preprint2015arXiv

Star Formation and Stellar Mass Assembly in Dark Matter Halos: From Giants to Dwarfs

The empirical model of Lu et al. 2014 is updated with recent data and used to study galaxy star formation and assembly histories. At $z > 2$, the predicted galaxy stellar mass functions are steep, and a significant amount of star formation is hosted by low-mass haloes that may be missed in current observations. Most of the stars in cluster centrals formed earlier than $z\approx 2$ but have been assembled much later. Milky Way mass galaxies have had on-going star formation without significant mergers since $z\approx 2$, and are thus free of significant (classic) bulges produced by major mergers. In massive clusters, stars bound in galaxies and scattered in the halo form a homogeneous population that is old and with solar metallicity. In contrast, in Milky Way mass systems the two components form two distinct populations, with halo stars being older and poorer in metals by a factor of $\approx 3$. Dwarf galaxies in haloes with $M_{\rm h} < 10^{11}h^{-1}M_{\odot}$ have experienced a star formation burst accompanied by major mergers at $z > 2$, followed by a nearly constant star formation rate after $z = 1$. The early burst leaves a significant old stellar population that is distributed in spheroids.

preprint2015arXiv

Using member galaxy luminosities as halo mass proxies of galaxy groups

Reliable halo mass estimation for a given galaxy system plays an important role both in cosmology and galaxy formation studies. Here we set out to find the way that can improve the halo mass estimation for those galaxy systems with limited brightest member galaxies been observed. Using four mock galaxy samples constructed from semi-analytical formation models, the subhalo abundance matching method and the conditional luminosity functions, respectively, we find that the luminosity gap between the brightest and the subsequent brightest member galaxies in a halo (group) can be used to significantly reduce the scatter in the halo mass estimation based on the luminosity of the brightest galaxy alone. Tests show that these corrections can significantly reduce the scatter in the halo mass estimations by $\sim 50\%$ to $\sim 70\%$ in massive halos depending on which member galaxies are considered. Comparing to the traditional ranking method, we find that this method works better for groups with less than five members, or in observations with very bright magnitude cut.

preprint2014arXiv

An Empirical Model for the Star Formation History in Dark Matter Halos

We develop an empirical approach to infer the star formation rate in dark matter halos from the galaxy stellar mass function (SMF) at different redshifts and the local cluster galaxy luminosity function (CGLF), which has a steeper faint end relative to the SMF of local galaxies. As satellites are typically old galaxies which have been accreted earlier, this feature can cast important constraint on the formation of low-mass galaxies at high-redshift. The evolution of the SMFs suggests the star formation in high mass halos ($>10^{12}M_{\odot}h^{-1}$) has to be boosted at high redshift beyond what is expected from a simple scaling of the dynamical time. The faint end of the CGLF implies a characteristic redshift $z_c\approx2$ above which the star formation rate in low mass halos with masses $< 10^{11}M_{\odot}h^{-1}$ must be enhanced relative to that at lower z. This is not directly expected from the standard stellar feedback models. Also, this enhancement leads to some interesting predictions, for instance, a significant old stellar population in present-day dwarf galaxies with $M_* < 10^8 M_{\odot}h^{-2}$ and steep slopes of high redshift stellar mass and star formation rate functions.

preprint2014arXiv

Connections between galaxy mergers and Starburst: evidence from local Universe

Major mergers and interactions between gas-rich galaxies with comparable masses are thought to be the main triggers of starburst. In this work, we study, for a large stellar mass range, the interaction rate of the starburst galaxies in the local universe. We focus independently on central and satellite star forming galaxies extracted from the Sloan Digital Sky Survey. Here the starburst galaxies are selected in the star formation rate (SFR) stellar mass plane with SFR five times larger than the median value found for "star forming" galaxies of the same stellar mass. Through visual inspection of their images together with close companions determined using spectroscopic redshifts, we find that ~50% of the "starburst" populations show evident merger features, i.e., tidal tails, bridges between galaxies, double cores and close companions. In contrast, in the control sample we selected from the normal star forming galaxies, only ~19% of galaxies are associated with evident mergers. The interaction rates may increase by ~5% for the starburst sample and 2% for the control sample if close companions determined using photometric redshifts are considered. The contrast of the merger rate between the two samples strengthens the hypothesis that mergers and interactions are indeed the main causes of starburst.

preprint2014arXiv

ELUCID - Exploring the Local Universe with reConstructed Initial Density field I: Hamiltonian Markov Chain Monte Carlo Method with Particle Mesh Dynamics

Simulating the evolution of the local universe is important for studying galaxies and the intergalactic medium in a way free of cosmic variance. Here we present a method to reconstruct the initial linear density field from an input non-linear density field, employing the Hamiltonian Markov Chain Monte Carlo (HMC) algorithm combined with Particle Mesh (PM) dynamics. The HMC+PM method is applied to cosmological simulations, and the reconstructed linear density fields are then evolved to the present day with N-body simulations. The constrained simulations so obtained accurately reproduce both the amplitudes and phases of the input simulations at various $z$. Using a PM model with a grid cell size of 0.75 Mpc/h and 40 time-steps in the HMC can recover more than half of the phase information down to a scale k~0.85 h/Mpc at high z and to k~3.4 h/Mpc at z=0, which represents a significant improvement over similar reconstruction models in the literature, and indicates that our model can reconstruct the formation histories of cosmic structures over a large dynamical range. Adopting PM models with higher spatial and temporal resolutions yields even better reconstructions, suggesting that our method is limited more by the availability of computer resource than by principle. Dynamic models of structure evolution adopted in many earlier investigations can induce non-Gaussianity in the reconstructed linear density field, which in turn can cause large systematic deviations in the predicted halo mass function. Such deviations are greatly reduced or absent in our reconstruction.

preprint2014arXiv

LAMOST observations in the Kepler field

The Large Sky Area Multi-Object Fiber Spectroscopic Telescope (LAMOST) at the Xinglong observatory in China is a new 4-m telescope equipped with 4,000 optical fibers. In 2010, we initiated the LAMOST-Kepler project. We requested to observe the full field-of-view of the nominal Kepler mission with the LAMOST to collect low-resolution spectra for as many objects from the KIC10 catalogue as possible. So far, 12 of the 14 requested LAMOST fields have been observed resulting in more than 68,000 low-resolution spectra. Our preliminary results show that the stellar parameters derived from the LAMOST spectra are in good agreement with those found in the literature based on high-resolution spectroscopy. The LAMOST data allows to distinguish dwarfs from giants and can provide the projected rotational velocity for very fast rotators.

preprint2014arXiv

Spin alignments of spiral galaxies within the large-scale structure from SDSS DR7

Using a sample of spiral galaxies selected from the Sloan Digital Sky Survey Data Release 7 (SDSS DR7) and Galaxy Zoo 2 (GZ2), we investigate the alignment of spin axes of spiral galaxies with their surrounding large scale structure, which is characterized by the large-scale tidal field reconstructed from the data using galaxy groups above a certain mass threshold. We find that the spin axes of only have weak tendency to be aligned with (or perpendicular to) the intermediate (or minor) axis of the local tidal tensor. The signal is the strongest in a \cluster environment where all the three eigenvalues of the local tidal tensor are positive. Compared to the alignments between halo spins and local tidal field obtained in N-body simulations, the above observational results are in best agreement with those for the spins of inner regions of halos, suggesting that the disk material traces the angular momentum of dark matter halos in the inner regions.

preprint2014arXiv

The statistical nature of the brightest group galaxies

We examine the statistical properties of the brightest group galaxies (BGGs) using a complete spectroscopic sample of groups/clusters of galaxies selected from the Data Release 7 of the Sloan Digital Sky Survey. We test whether BGGs and other bright members of groups are consistent with an ordered population among the total population of group galaxies. We find that the luminosity distributions of BGGs do not follow the predictions from the order statistics (OS). The average luminosities of BGGs are systematically brighter than OS predictions. On the other hand, by properly taking into account the brightening effect of the BGGs, the luminosity distributions of the second brightest galaxies are in excellent agreement with the expectations of OS. The brightening of BGGs relative to the OS expectation is consistent with a scenario that the BGGs on average have over-grown about 20 percent masses relative to the other member galaxies. The growth ($ΔM$) is not stochastic but correlated with the magnitude gap ($G_{1,2}$) between the brightest and the second brightest galaxy. The growth ($ΔM$) is larger for the groups having more prominent BGGs (larger $G_{1,2}$) and averagely contributes about 30 percent of the final $G_{1,2}$ of the groups of galaxies.

preprint2013arXiv

Alignments of galaxies within cosmic filaments from SDSS DR7

Using a sample of galaxy groups selected from the Sloan Digital Sky Survey Data Release 7 (SDSS DR7), we examine the alignment between the orientation of galaxies and their surrounding large scale structure in the context of the cosmic web. The latter is quantified using the large-scale tidal field, reconstructed from the data using galaxy groups above a certain mass threshold. We find that the major axes of galaxies in filaments tend to be preferentially aligned with the directions of the filaments, while galaxies in sheets have their major axes preferentially aligned parallel to the plane of the sheets. The strength of this alignment signal is strongest for red, central galaxies, and in good agreement with that of dark matter halos in N-body simulations. This suggests that red, central galaxies are well aligned with their host halos, in quantitative agreement with previous studies based on the spatial distribution of satellite galaxies. There is a luminosity and mass dependence that brighter and more massive galaxies in filaments and sheets have stronger alignment signals. We also find that the orientation of galaxies is aligned with the eigenvector associated with the smallest eigenvalue of the tidal tensor. These observational results indicate that galaxy formation is affected by large-scale environments, and strongly suggests that galaxies are aligned with each other over scales comparable to those of sheets and filaments in the cosmic web.

preprint2013arXiv

Constraining the Star Formation Histories in Dark Matter Halos: I. Central Galaxies

Using the self-consistent modeling of the conditional stellar mass functions across cosmic time by Yang et al. (2012), we make model predictions for the star formation histories (SFHs) of {\it central} galaxies in halos of different masses. The model requires the following two key ingredients: (i) mass assembly histories of central and satellite galaxies, and (ii) local observational constraints of the star formation rates of central galaxies as function of halo mass. We obtain a universal fitting formula that describes the (median) SFH of central galaxies as function of halo mass, galaxy stellar mass and redshift. We use this model to make predictions for various aspects of the star formation rates of central galaxies across cosmic time. Our main findings are the following. (1) The specific star formation rate (SSFR) at high $z$ increases rapidly with increasing redshift [$\propto (1+z)^{2.5}$] for halos of a given mass and only slowly with halo mass ($\propto M_h^{0.12}$) at a given $z$, in almost perfect agreement with the specific mass accretion rate of dark matter halos. (2) The ratio between the star formation rate (SFR) in the main-branch progenitor and the final stellar mass of a galaxy peaks roughly at a constant value, $\sim 10^{-9.3} h^2 {\rm yr}^{-1}$, independent of halo mass or the final stellar mass of the galaxy. However, the redshift at which the SFR peaks increases rapidly with halo mass. (3) More than half of the stars in the present-day Universe were formed in halos with $10^{11.1}\msunh < M_h < 10^{12.3}\msunh$ in the redshift range $0.4 < z < 1.9$. (4) ... [abridged]

preprint2013arXiv

Constraining the substructure of dark matter haloes with galaxy-galaxy lensing

With galaxy groups constructed from the Sloan Digital Sky Survey (SDSS), we analyze the expected galaxy-galaxy lensing signals around satellite galaxies residing in different host haloes and located at different halo-centric distances. We use Markov Chain Monte Carlo (MCMC) method to explore the potential constraints on the mass and density profile of subhaloes associated with satellite galaxies from SDSS-like surveys and surveys similar to the Large Synoptic Survey Telescope (LSST). Our results show that for SDSS-like surveys, we can only set a loose constraint on the mean mass of subhaloes. With LSST-like surveys, however, both the mean mass and the density profile of subhaloes can be well constrained.

preprint2013arXiv

Detection of galaxy assembly bias

Assembly bias describes the finding that the clustering of dark matter haloes depends on halo formation time at fixed halo mass. In this paper, we analyse the influence of assembly bias on galaxy clustering using both semi-analytical models (SAMs) and observational data. At fixed stellar mass, SAMs predict that the clustering of {\it central} galaxies depends on the specific star formation rate (sSFR), with more passive galaxies having a higher clustering amplitude. We find similar trends using SDSS group catalogues, and verify that these are not affected by possible biases due to the group finding algorithm. Low mass central galaxies reside in narrow bins of halo mass, so the observed trends of higher clustering amplitude for galaxies with lower sSFR is not driven by variations of the parent halo mass. We argue that the clustering dependence on sSFR represent a direct detection of assembly bias. In addition, contrary to what expected based on clustering of dark matter haloes, we find that low-mass central galaxies in SAMs with larger host halo mass have a {\it lower} clustering amplitude than their counter-parts residing in lower mass haloes. This results from the fact that, at fixed stellar mass, assembly bias has a stronger influence on clustering than the dependence on the parent halo mass.

preprint2013arXiv

First Galaxy-Galaxy Lensing Measurement of Satellite Halo Mass in the CFHT Stripe-82 Survey

We select satellite galaxies from the galaxy group catalog constructed with the SDSS spectroscopic galaxies and measure the tangential shear around these galaxies with source catalog extracted from CFHT/MegaCam Stripe-82 Survey to constrain the mass of subhalos associated with them. The lensing signal is measured around satellites in groups with masses in the range [10^{13}, 5x10^{14}]h^{-1}M_{sun}, and is found to agree well with theoretical expectation. Fitting the data with a truncated NFW profile, we obtain an average subhalo mass of log M_{sub}= 11.68 \pm 0.67 for satellites whose projected distances to central galaxies are in the range [0.1, 0.3] h^{-1}Mpc, and log M_{sub}= 11.68 \pm 0.76 for satellites with projected halo-centric distance in [0.3, 0.5] h^{-1}Mpc. The best-fit subhalo masses are comparable to the truncated subhalo masses assigned to satellite galaxies using abundance matching and about 5 to 10 times higher than the average stellar mass of the lensing satellite galaxies.

preprint2013arXiv

Measuring the X-ray luminosities of SDSS DR7 clusters from RASS

We use ROSAT All Sky Survey (RASS) broadband X-ray images and the optical clusters identified from SDSS DR7 to estimate the X-ray luminosities around $\sim 65,000$ candidate clusters with masses $\ga 10^{13}\msunh$ based on an Optical to X-ray (OTX) code we develop. We obtain a catalogue with X-ray luminosity for each cluster. This catalog contains 817 clusters (473 at redshift $z\le 0.12$) with $S/N> 3$ in X-ray detection. We find about $65\%$ of these X-ray clusters have their most massive member located near the X-ray flux peak; for the rest $35\%$, the most massive galaxy is separated from the X-ray peak, with the separation following a distribution expected from a NFW profile. We investigate a number of correlations between the optical and X-ray properties of these X-ray clusters, and find that: the cluster X-ray luminosity is correlated with the stellar mass (luminosity) of the clusters, as well as with the stellar mass (luminosity) of the central galaxy and the mass of the halo, but the scatter in these correlations is large. Comparing the properties of X-ray clusters of similar halo masses but having different X-ray luminosities, we find that massive halos with masses $\ga 10^{14}\msunh$ contain a larger fraction of red satellite galaxies when they are brighter in X-ray. ... A cluster catalog containing the optical properties of member galaxies and the X-ray luminosity is available at {\it http://gax.shao.ac.cn/data/Group.html}.

preprint2013arXiv

Nonlinearities in modified gravity cosmology. II. Impacts of modified gravity on the halo properties

The statistics of dark matter halos is an essential component of understanding the nonlinear evolution in modified gravity cosmology. Based on a series of modified gravity N-body simulations, we investigate the halo mass function, concentration and bias. We model the impact of modified gravity by a single parameter ζ, which determines the enhancement of particle acceleration with respect to GR, given the identical mass distribution (ζ=1 in GR). We select snapshot redshifts such that the linear matter power spectra of different gravity models are identical, in order to isolate the impact of gravity beyond modifying the linear growth rate. At the baseline redshift corresponding to z_S=1.2 in the standard ΛCDM, for a 10% deviation from GR(|ζ-1|=0.1), the measured halo mass function can differ by about 5-10%, the halo concentration by about 10-20%, while the halo bias differs significantly less. These results demonstrate that the halo mass function and/or the halo concentration are sensitive to the nature of gravity and may be used to make interesting constraints along this line.

preprint2013arXiv

Reconstructing the Initial Density Field of the Local Universe: Method and Test with Mock Catalogs

Our research objective in this paper is to reconstruct an initial linear density field, which follows the multivariate Gaussian distribution with variances given by the linear power spectrum of the current CDM model and evolves through gravitational instability to the present-day density field in the local Universe. For this purpose, we develop a Hamiltonian Markov Chain Monte Carlo method to obtain the linear density field from a posterior probability function that consists of two components: a prior of a Gaussian density field with a given linear spectrum, and a likelihood term that is given by the current density field. The present-day density field can be reconstructed from galaxy groups using the method developed in Wang et al. (2009a). Using a realistic mock SDSS DR7, obtained by populating dark matter haloes in the Millennium simulation with galaxies, we show that our method can effectively and accurately recover both the amplitudes and phases of the initial, linear density field. To examine the accuracy of our method, we use $N$-body simulations to evolve these reconstructed initial conditions to the present day. The resimulated density field thus obtained accurately matches the original density field of the Millennium simulation in the density range 0.3 <= rho/rho_mean <= 20 without any significant bias. Especially, the Fourier phases of the resimulated density fields are tightly correlated with those of the original simulation down to a scale corresponding to a wavenumber of ~ 1 h/Mpc, much smaller than the translinear scale, which corresponds to a wavenumber of ~ 0.15 h\Mpc.

preprint2012arXiv

Bulk flow of halos in ΛCDM simulation

Analysis of the Pangu N-body simulation validates that the bulk flow of halos follows a Maxwellian distribution which variance is consistent with the prediction of the linear theory of structure formation. We propose that the consistency between the observed bulk velocity and theories should be examined at the effective scale of the radius of a spherical top-hat window function yielding the same smoothed velocity variance in linear theory as the sample window function does. We compared some recently estimated bulk flows from observational samples with the prediction of the ΛCDM model we used; some results deviate from expectation at a level of ~ 3σbut the discrepancy is not as severe as previously claimed. We show that bulk flow is only weakly correlated with the dipole of the internal mass distribution, the alignment angle between the mass dipole and the bulk flow has a broad distribution peaked at ~ 30-50 deg., and also that the bulk flow shows little dependence on the mass of the halos used in the estimation. In a simulation of box size 1Gpc/h, for a cell of radius 100 Mpc/h the maximal bulk velocity is >500 km/s, dipoles of the environmental mass outside the cell are not tightly aligned with the bulk flow, but are rather located randomly around it with separation angles ~ 20-40 deg. In the fastest cell there is a slightly smaller number of low-mass halos; however halos inside are clustered more strongly at scales > ~ 20 Mpc/h, which might be a significant feature since the correlation between bulk flow and halo clustering actually increases in significance beyond such scales.

preprint2012arXiv

Cosmological Constraints from a Combination of Galaxy Clustering & Lensing -- II. Fisher Matrix Analysis

We quantify the accuracy with which the cosmological parameters characterizing the energy density of matter (Ω_m), the amplitude of the power spectrum of matter fluctuations (σ_8), the energy density of neutrinos (Ω_ν) and the dark energy equation of state (w_0) can be constrained using data from large galaxy redshift surveys. We advocate a joint analysis of the abundance of galaxies, galaxy clustering, and the galaxy-galaxy weak lensing signal in order to simultaneously constrain the halo occupation statistics (i.e., galaxy bias) and the cosmological parameters of interest. We parameterize the halo occupation distribution of galaxies in terms of the conditional luminosity function and use the analytical framework of the halo model described in our companion paper (van den Bosch et al. 2012), to predict the relevant observables. By performing a Fisher matrix analysis, we show that a joint analysis of these observables, even with the precision with which they are currently measured from the Sloan Digital Sky Survey, can be used to obtain tight constraints on the cosmological parameters, fully marginalized over uncertainties in galaxy bias. We demonstrate that the cosmological constraints from such an analysis are nearly uncorrelated with the halo occupation distribution constraints, thus, minimizing the systematic impact of any imperfections in modeling the halo occupation statistics on the cosmological constraints. In fact, we demonstrate that the constraints from such an analysis are both complementary to and competitive with existing constraints on these parameters from a number of other techniques, such as cluster abundances, cosmic shear and/or baryon acoustic oscillations, thus paving the way to test the concordance cosmological model.

preprint2012arXiv

Cosmological Constraints from a Combination of Galaxy Clustering and Lensing -- I. Theoretical Framework

We present a new method that simultaneously solves for cosmology and galaxy bias on non-linear scales. The method uses the halo model to analytically describe the (non-linear) matter distribution, and the conditional luminosity function (CLF) to specify the halo occupation statistics. For a given choice of cosmological parameters, this model can be used to predict the galaxy luminosity function, as well as the two-point correlation functions of galaxies, and the galaxy-galaxy lensing signal, both as function of scale and luminosity. In this paper, the first in a series, we present the detailed, analytical model, which we test against mock galaxy redshift surveys constructed from high-resolution numerical $N$-body simulations. We demonstrate that our model, which includes scale-dependence of the halo bias and a proper treatment of halo exclusion, reproduces the 3-dimensional galaxy-galaxy correlation and the galaxy-matter cross-correlation (which can be projected to predict the observables) with an accuracy better than 10 (in most cases 5) percent. Ignoring either of these effects, as is often done, results in systematic errors that easily exceed 40 percent on scales of $\sim 1 h^{-1}\Mpc$, where the data is typically most accurate. Finally, since the projected correlation functions of galaxies are never obtained by integrating the redshift space correlation function along the line-of-sight out to infinity, simply because the data only cover a finite volume, they are still affected by residual redshift space distortions (RRSDs). Ignoring these, as done in numerous studies in the past, results in systematic errors that easily exceed 20 perent on large scales ($r_\rmp \gta 10 h^{-1}\Mpc$). We show that it is fairly straightforward to correct for these RRSDs, to an accuracy better than $\sim 2$ percent, using a mildly modified version of the linear Kaiser formalism.

preprint2012arXiv

Cosmological Constraints from a Combination of Galaxy Clustering and Lensing -- III. Application to SDSS Data

We simultaneously constrain cosmology and galaxy bias using measurements of galaxy abundances, galaxy clustering and galaxy-galaxy lensing taken from the Sloan Digital Sky Survey. We use the conditional luminosity function (which describes the halo occupation statistics as function of galaxy luminosity) combined with the halo model (which describes the non-linear matter field in terms of its halo building blocks) to describe the galaxy-dark matter connection. We explicitly account for residual redshift space distortions in the projected galaxy-galaxy correlation functions, and marginalize over uncertainties in the scale dependence of the halo bias and the detailed structure of dark matter haloes. Under the assumption of a spatially flat, vanilla ΛCDM cosmology, we focus on constraining the matter density, Ωm, and the normalization of the matter power spectrum, σ8, and we adopt WMAP7 priors for the spectral index, the Hubble parameter, and the baryon density. We obtain that \Omegam = 0.278_{-0.026}^{+0.023} and σ8 = 0.763_{-0.049}^{+0.064} (95% CL). These results are robust to uncertainties in the radial number density distribution of satellite galaxies, while allowing for non-Poisson satellite occupation distributions results in a slightly lower value for σ8 (0.744_{-0.047}^{+0.056}). These constraints are in excellent agreement (at the 1σ level) with the cosmic microwave background constraints from WMAP. This demonstrates that the use of a realistic and accurate model for galaxy bias, down to the smallest non-linear scales currently observed in galaxy surveys, leads to results perfectly consistent with the vanilla ΛCDM cosmology.

preprint2012arXiv

Cross identification between X-ray and Optical Clusters of Galaxies in the SDSS DR7 Field

We use the ROSAT all sky survey X-ray cluster catalogs and the optical SDSS DR7 galaxy and group catalogs to cross-identify X-ray clusters with their optical counterparts, resulting in a sample of 201 X-ray clusters in the sky coverage of SDSS DR7. We investigate various correlations between the optical and X-ray properties of these X-ray clusters, and find that the following optical properties are correlated with the X-ray luminosity: the central galaxy luminosity, the central galaxy mass, the characteristic group luminosity ($\propto \Lx^{0.43}$), the group stellar mass ($\propto \Lx^{0.46}$), with typical 1-$σ$ scatter of $\sim 0.67$ in $\log \Lx$. Using the observed number distribution of X-ray clusters, we obtain an unbiased scaling relation between the X-ray luminosity, the central galaxy stellar mass and the characteristic satellite stellar mass as ${\log L_X} = -0.26 + 2.90 [\log (M_{\ast, c} + 0.26 M_{\rm sat}) -12.0]$ (and in terms of luminosities, as ${\log L_X} = -0.15 + 2.38 [\log (L_{c} + 0.72 L_{\rm sat}) -12.0]$). We find that the systematic difference between different halo mass estimations, e.g., using the ranking of characteristic group stellar mass or using the X-ray luminosity scaling relation can be used to constrain cosmology. Comparing the properties of groups of similar stellar mass (or optical luminosities) and redshift that are X-ray luminous or under-luminous, we find that X-ray luminous groups have more faint satellite galaxies and higher red fraction in their satellites. The cross-identified X-ray clusters together with their optical properties are provided in Appendix B.

preprint2012arXiv

Evolution of the Galaxy - Dark Matter Connection and the Assembly of Galaxies in Dark Matter Halos

We present a new model to describe the galaxy-dark matter connection across cosmic time, which unlike the popular subhalo abundance matching technique is self-consistent in that it takes account of the facts that (i) subhalos are accreted at different times, and (ii) the properties of satellite galaxies may evolve after accretion. Using observations of galaxy stellar mass functions out to $z \sim 4$, the conditional stellar mass function at $z\sim 0.1$ obtained from SDSS galaxy group catalogues, and the two-point correlation function (2PCF) of galaxies at $z \sim 0.1$ as function of stellar mass, we constrain the relation between galaxies and dark matter halos over the entire cosmic history from $z \sim 4$ to the present. This relation is then used to predict the median assembly histories of different stellar mass components within dark matter halos (central galaxies, satellite galaxies, and halo stars). We also make predictions for the 2PCFs of high-$z$ galaxies as function of stellar mass. Our main findings are the following: (i) Our model reasonably fits all data within the observational uncertainties, indicating that the $Λ$CDM concordance cosmology is consistent with a wide variety of data regarding the galaxy population across cosmic time. (ii) ... [abridged]

preprint2012arXiv

Internal kinematics of groups of galaxies in the Sloan Digital Sky Survey data release 7

We present measurements of the velocity dispersion profile (VDP) for galaxy groups in the final data release of the Sloan Digital Sky Survey (SDSS). For groups of given mass we estimate the redshift-space cross-correlation function (CCF) with respect to a reference galaxy sample, xi(r_p, pi), the projected CCF, w_p(r_p), and the real-space CCF, xi(r). The VDP is then extracted from the redshift distortion in xi(r_p, pi), by comparing xi(r_p, pi) with xi(r). We find that the velocity dispersion (VD) within virial radius (R_200) shows a roughly flat profile, with a slight increase at radii below ~0.3 R_200 for high mass systems. The average VD within the virial radius, sigma_v, is a strongly increasing function of central galaxy mass. We apply the same methodology to N-body simulations with the concordance Lambda cold dark matter cosmology but different values of the density fluctuation parameter sigma_8, and we compare the results to the SDSS results. We show that the sigma_v-M_* relation from the data provides stringent constraints on both sigma_8 and sigma_ms, the dispersion in log M_* of central galaxies at fixed halo mass. Our best-fitting model suggests sigma_8 = 0.86 +/- 0.03 and sigma_ms = 0.16 +/- 0.03. The slightly higher value of sigma_8 compared to the WMAP7 result might be due to a smaller matter density parameter assumed in our simulations. Our VD measurements also provide a direct measure of the dark matter halo mass for central galaxies of different luminosities and masses, in good agreement with the results obtained by Mandelbaum et al. (2006) from stacking the gravitational lensing signals of the SDSS galaxies.

preprint2012arXiv

Measures of Galaxy Environment - I. What is "Environment"?

The influence of a galaxy's environment on its evolution has been studied and compared extensively in the literature, although differing techniques are often used to define environment. Most methods fall into two broad groups: those that use nearest neighbours to probe the underlying density field and those that use fixed apertures. The differences between the two inhibit a clean comparison between analyses and leave open the possibility that, even with the same data, different properties are actually being measured. In this work we apply twenty published environment definitions to a common mock galaxy catalogue constrained to look like the local Universe. We find that nearest neighbour-based measures best probe the internal densities of high-mass haloes, while at low masses the inter-halo separation dominates and acts to smooth out local density variations. The resulting correlation also shows that nearest neighbour galaxy environment is largely independent of dark matter halo mass. Conversely, aperture-based methods that probe super-halo scales accurately identify high-density regions corresponding to high mass haloes. Both methods show how galaxies in dense environments tend to be redder, with the exception of the largest apertures, but these are the strongest at recovering the background dark matter environment. We also warn against using photometric redshifts to define environment in all but the densest regions. When considering environment there are two regimes: the 'local environment' internal to a halo best measured with nearest neighbour and 'large-scale environment' external to a halo best measured with apertures. This leads to the conclusion that there is no universal environment measure and the most suitable method depends on the scale being probed.

preprint2012arXiv

Simulation of relativistically colliding laser-generated electron flows

The plasma dynamics resulting from the simultaneous impact, of two equal, ultra-intense laser pulses, in two spatially separated spots, onto a dense target is studied via particle-in-cell (PIC) simulations. The simulations show that electrons accelerated to relativistic speeds, cross the target and exit at its rear surface. Most energetic electrons are bound to the rear surface by the ambipolar electric field and expand along it. Their current is closed by a return current in the target, and this current configuration generates strong surface magnetic fields. The two electron sheaths collide at the midplane between the laser impact points. The magnetic repulsion between the counter-streaming electron beams separates them along the surface normal direction, before they can thermalize through other beam instabilities. This magnetic repulsion is also the driving mechanism for the beam-Weibel (filamentation) instability, which is thought to be responsible for magnetic field growth close to the internal shocks of gamma-ray burst (GRB) jets. The relative strength of this repulsion compared to the competing electrostatic interactions, which is evidenced by the simulations, suggests that the filamentation instability can be examined in an experimental setting.

preprint2012arXiv

The Galaxy-Dark Matter Connection: A Cosmological Perspective

We present a method that uses observations of galaxies to simultaneously constrain cosmological parameters and the galaxy-dark matter connection (aka halo occupation statistics). The latter describes how galaxies are distributed over dark matter haloes, and is an imprint of the poorly understood physics of galaxy formation. A generic problem of using galaxies to constrain cosmology is that galaxies are a biased tracer of the mass distribution, and this bias is generally unknown. The great advantage of simultaneously constraining cosmology and halo occupation statistics is that this effectively allows cosmological constraints marginalized over the uncertainties regarding galaxy bias. Not only that, it also yields constraints on the galaxy-dark matter connection, this time properly marginalized over cosmology, which is of great value to inform theoretical models of galaxy formation. We use a combination of the analytical halo model and the conditional luminosity function to describe the galaxy-dark matter connection, which we use to model the abundance, clustering and galaxy-galaxy lensing properties of the galaxy population. We use a Fisher matrix analysis to gauge the complementarity of these different observables, and present some preliminary results from an analysis based on data from the Sloan Digital Sky Survey. Our results are complementary to and perfectly consistent with the results from the 7 year data release of the WMAP mission, strengthening the case for a true 'concordance' cosmology.

preprint2011arXiv

An analytical model for the accretion of dark matter subhalos

An analytical model is developed for the mass function of cold dark matter subhalos at the time of accretion and for the distribution of their accretion times. Our model is based on the model of Zhao et al. (2009) for the median assembly histories of dark matter halos, combined with a simple log-normal distribution to describe the scatter in the main-branch mass at a given time for halos of the same final mass. Our model is simple, and can be used to predict the un-evolved subhalo mass function, the mass function of subhalos accreted at a given time, the accretion-time distribution of subhalos of a given initial mass, and the frequency of major mergers as a function of time. We test our model using high-resolution cosmological $N$-body simulations, and find that our model predictions match the simulation results remarkably well. Finally, we discuss the implications of our model for the evolution of subhalos in their hosts and for the construction of a self-consistent model to link galaxies and dark matter halos at different cosmic times.

preprint2011arXiv

Are Brightest Halo Galaxies Central Galaxies?

It is generally assumed that the central galaxy in a dark matter halo, that is, the galaxy with the lowest specific potential energy, is also the brightest halo galaxy (BHG), and that it resides at rest at the centre of the dark matter potential well. This central galaxy paradigm (CGP) is an essential assumption made in various fields of astronomical research. In this paper we test the validity of the CGP using a large galaxy group catalogue constructed from the Sloan Digital Sky Survey. For each group we compute two statistics, ${\cal R}$ and ${\cal S}$, which quantify the offsets of the line-of-sight velocities and projected positions of brightest group galaxies relative to the other group members. By comparing the cumulative distributions of $|{\cal R}|$ and $|{\cal S}|$ to those obtained from detailed mock group catalogues, we rule out the null-hypothesis that the CGP is correct. Rather, the data indicate that in a non-zero fraction $f_{\rm BNC}(M)$ of all haloes of mass $M$ the BHG is not the central galaxy, but instead, a satellite galaxy. In particular, we find that $f_{\rm BNC}$ increases from $\sim 0.25$ in low mass haloes ($10^{12} h^{-1} {\rm M_{\odot}} \leq M \lsim 2 \times 10^{13} h^{-1}{\rm M_{\odot}}$) to $\sim 0.4$ in massive haloes ($M \gsim 5 \times 10^{13} h^{-1} {\rm M_{\odot}}$). We show that these values of $f_{\rm BNC}$ are uncomfortably high compared to predictions from halo occupation statistics and from semi-analytical models of galaxy formation. We end by discussing various implications of a non-zero $f_{\rm BNC}(M)$, with an emphasis on the halo masses inferred from satellite kinematics.

preprint2011arXiv

Probing Hot Gas in Galaxy Groups through the Sunyaev-Zeldovich Effect

We investigate the potential of exploiting the Sunyaev-Zeldovich effect (SZE) to study the properties of hot gas in galaxy groups. It is shown that, with upcoming SZE surveys, one can stack SZE maps around galaxy groups of similar halo masses selected from large galaxy redshift surveys to study the hot gas in halos represented by galaxy groups. We use various models for the hot halo gas to study how the expected SZE signals are affected by gas fraction, equation of state, halo concentration, and cosmology. Comparing the model predictions with the sensitivities expected from the SPT, ACT and Planck surveys shows that a SPT-like survey can provide stringent constraints on the hot gas properties for halos with masses M ~> 10^{13} h^{-1}Msun. We also explore the idea of using the cross correlation between hot gas and galaxies of different luminosity to probe the hot gas in dark matter halos without identifying galaxy groups to represent dark halos. Our results show that, with a galaxy survey as large as the Sloan Digital Sky Survey and with the help of the conditional luminosity function (CLF) model, one can obtain stringent constraints on the hot gas properties in halos with masses down to 10^{13} h^{-1}Msun. Thus, the upcoming SZE surveys should provide a very promising avenue to probe the hot gas in relatively low-mass halos where the majority of L*-galaxies reside.

preprint2011arXiv

Properties of fossil groups in cosmological simulations and galaxy formation models

It has been a long-standing question whether fossil groups are just sampling the tail of the distribution of ordinary groups, or whether they are a physically distinct class of objects, characterized by an unusual and special formation history. To study this question, we here investigate fossil groups identified in the hydrodynamical simulations of the GIMIC project, which consists of resimulations of five regions in the Millennium Simulation (MS) that are characterized by different large-scale densities, ranging from a deep void to a proto-cluster region. For comparison, we also consider semi-analytic models built on top of the MS, as well as a conditional luminosity function approach. We identify galaxies in the GIMIC simulations as groups of stars and use a spectral synthesis code to derive their optical properties. The X-ray luminosity of the groups is estimated in terms of the thermal bremsstrahlung emission of the gas in the host halos, neglecting metallicity effects. We focus on comparing the properties of fossil groups in the theoretical models and observational results, highlighting the differences between them, and trying to identify possible dependencies on environment for which our approach is particularly well set-up. We find that the optical fossil fraction in all of our theoretical models declines with increasing halo mass, and there is no clear environmental dependence. Combining the optical and X-ray selection criteria for fossil groups, the halo mass dependence of the fossil groups seen in optical vanishes. Over the GIMIC halo mass range we resolve best, 9.0\times1012 \sim 4.0\times1013 h-1 M, the central galaxies in the fossil groups show similar properties as those in ordinary groups, in terms of age, metallicity, color, concentration, and mass-to-light ratio. [abridged]

preprint2011arXiv

Reconstructing the Cosmic Velocity and Tidal Fields with Galaxy Groups Selected from the Sloan Digital Sky Survey

[abridge]Cosmic velocity and tidal fields are important for the understanding of the cosmic web and the environments of galaxies, and can also be used to constrain cosmology. In this paper, we reconstruct these two fields in SDSS volume from dark matter halos represented by galaxy groups. Detailed mock catalogues are used to test the reliability of our method against uncertainties arising from redshift distortions, survey boundaries, and false identifications of groups by our group finder. We find that both the velocity and tidal fields, smoothed on a scale of ~2Mpc/h, can be reliably reconstructed in the inner region (~66%) of the survey volume. The reconstructed tidal field is used to split the cosmic web into clusters, filaments, sheets, and voids, depending on the sign of the eigenvalues of tidal tensor. The reconstructed velocity field nicely shows how the flows are diverging from the centers of voids, and converging onto clusters, while sheets and filaments have flows that are convergent along one and two directions, respectively. We use the reconstructed velocity field and the Zel'dovich approximation to predict the mass density field in the SDSS volume as function of redshift, and find that the mass distribution closely follows the galaxy distribution even on small scales. We find a large-scale bulk flow of about 117km/s in a very large volume, equivalent to a sphere with a radius of ~170Mpc/h, which seems to be produced by the massive structures associated with the SDSS Great Wall. Finally, we discuss potential applications of our reconstruction to study the environmental effects of galaxy formation, to generate initial conditions for simulations of the local Universe, and to constrain cosmological models. The velocity, tidal and density fields in the SDSS volume, specified on a Cartesian grid with a spatial resolution of ~700kpc/h, are available from the authors upon request.

preprint2010arXiv

Genus statistics using the Delaunay tessellation field estimation method: (I) tests with the Millennium Simulation and the SDSS DR7

We study the topology of cosmic large-scale structure through the genus statistics, using galaxy catalogues generated from the Millennium Simulation and observational data from the latest Sloan Digital Sky Survey Data Release (SDSS DR7). We introduce a new method for constructing galaxy density fields and for measuring the genus statistics of its isodensity surfaces. It is based on a Delaunay tessellation field estimation (DTFE) technique that allows the definition of a piece-wise continuous density field and the exact computation of the topology of its polygonal isodensity contours, without introducing any free numerical parameter. Besides this new approach, we also employ the traditional approaches of smoothing the galaxy distribution with a Gaussian of fixed width, or by adaptively smoothing with a kernel that encloses a constant number of neighboring galaxies. Our results show that the Delaunay-based method extracts the largest amount of topological information. Unlike the traditional approach for genus statistics, it is able to discriminate between the different theoretical galaxy catalogues analyzed here, both in real space and in redshift space, even though they are based on the same underlying simulation model. In particular, the DTFE approach detects with high confidence a discrepancy of one of the semi-analytic models studied here compared with the SDSS data, while the other models are found to be consistent.

preprint2010arXiv

Nonlinearities in modified gravity cosmology I: signatures of modified gravity in the nonlinear matter power spectrum

A large fraction of cosmological information on dark energy and gravity is encoded in the nonlinear regime. Precision cosmology thus requires precision modeling of nonlinearities in general dark energy and modified gravity models. We modify the Gadget-2 code and run a series of N-body simulations on modified gravity cosmology to study the nonlinearities. The modified gravity model that we investigate in the present paper is characterized by a single parameter ζ, which determines the enhancement of particle acceleration with respect to general relativity (GR), given the identical mass distribution (ζ= 1 in GR). The first nonlinear statistics we investigate is the nonlinear matter power spectrum at k < 3h/Mpc, which is the relevant range for robust weak lensing power spectrum modeling at l < 2000. In this study, we focus on the relative difference in the nonlinear power spectra at corresponding redshifts where different gravity models have the same linear power spectra. This particular statistics highlights the imprint of modified gravity in the nonlinear regime and the importance to include the nonlinear regime in testing GR. By design, it is less susceptible to the sample variance and numerical artifacts. We adopt a mass assignment method based on wavelet to improve the power spectrum measurement. We run a series of tests to determine the suitable simulation specifications (particle number, box size and initial redshift). We find that, the nonlinear power spectra can differ by ~30% for 10% deviation from GR (|ζ-1| = 0.1) where the rms density fluctuations reach 10. This large difference, on one hand, shows the richness of information on gravity in the corresponding scales, and on the other hand, invalidates simple extrapolations of some existing fitting formulae to modified gravity cosmology.

preprint2010arXiv

Satellite Kinematics III: Halo Masses of Central Galaxies in SDSS

We use the kinematics of satellite galaxies that orbit around the central galaxy in a dark matter halo to infer the scaling relations between halo mass and central galaxy properties. Using galaxies from the Sloan Digital Sky Survey, we investigate the halo mass-luminosity relation (MLR) and the halo mass-stellar mass relation (MSR) of central galaxies. In particular, we focus on the dependence of these scaling relations on the colour of the central galaxy. We find that red central galaxies on average occupy more massive haloes than blue central galaxies of the same luminosity. However, at fixed stellar mass there is no appreciable difference in the average halo mass of red and blue centrals, especially for M* $\lsim$ 10^{10.5} h^{-2} Msun. This indicates that stellar mass is a better indicator of halo mass than luminosity. Nevertheless, we find that the scatter in halo masses at fixed stellar mass is non-negligible for both red and blue centrals. It increases as a function of stellar mass for red centrals but shows a fairly constant behaviour for blue centrals. We compare the scaling relations obtained in this paper with results from other independent studies of satellite kinematics, with results from a SDSS galaxy group catalog, from galaxy-galaxy weak lensing measurements, and from subhalo abundance matching studies. Overall, these different techniques yield MLRs and MSRs in fairly good agreement with each other (typically within a factor of two), indicating that we are converging on an accurate and reliable description of the galaxy-dark matter connection. We briefly discuss some of the remaining discrepancies among the various methods.

preprint2010arXiv

The Stellar Mass Components of Galaxies: Comparing Semi-Analytical Models with Observation

We compare the stellar masses of central and satellite galaxies predicted by three independent semianalytical models with observational results obtained from a large galaxy group catalogue constructed from the Sloan Digital Sky Survey. In particular, we compare the stellar mass functions of centrals and satellites, the relation between total stellar mass and halo mass, and the conditional stellar mass functions, which specify the average number of galaxies of stellar mass M_* that reside in a halo of mass M_h. The semi-analytical models only predict the correct stellar masses of central galaxies within a limited mass range and all models fail to reproduce the sharp decline of stellar mass with decreasing halo mass observed at the low mass end. In addition, all models over-predict the number of satellite galaxies by roughly a factor of two. The predicted stellar mass in satellite galaxies can be made to match the data by assuming that a significant fraction of satellite galaxies are tidally stripped and disrupted, giving rise to a population of intra-cluster stars in their host halos. However, the amount of intra-cluster stars thus predicted is too large compared to observation. This suggests that current galaxy formation models still have serious problems in modeling star formation in low-mass halos.

preprint2009arXiv

Galaxy Groups in the SDSS DR4: III. the luminosity and stellar mass functions

Using a large galaxy group catalogue constructed from the Sloan Digital Sky Survey Data Release 4 (SDSS DR4) with an adaptive halo-based group finder, we investigate the luminosity and stellar mass functions for different populations of galaxies (central versus satellite; red versus blue; and galaxies in groups of different masses) and for groups themselves. The conditional stellar mass function (CSMF), which describes the stellar distribution of galaxies in halos of a given mass for central and satellite galaxies can be well modeled with a log-normal distribution and a modified Schechter form, respectively. On average, there are about 3 times as many central galaxies as satellites. Among the satellite population, there are in general more red galaxies than blue ones. For the central population, the luminosity function is dominated by red galaxies at the massive end, and by blue galaxies at the low mass end. At the very low-mass end ($M_\ast \la 10^9 h^{-2}\Msun$), however, there is a marked increase in the number of red centrals. We speculate that these galaxies are located close to large halos so that their star formation is truncated by the large-scale environments. The stellar-mass function of galaxy groups is well described by a double power law, with a characteristic stellar mass at $\sim 4\times 10^{10}h^{-2}\Msun$. Finally, we use the observed stellar mass function of central galaxies to constrain the stellar mass - halo mass relation for low mass halos, and obtain $M_{\ast, c}\propto M_h^{4.9}$ for $M_h \ll 10^{11} \msunh$.

preprint2009arXiv

Stellar Ages and Metallicities of Central and Satellite Galaxies: Implications for Galaxy Formation and Evolution

Using a large SDSS galaxy group catalogue, we study how the stellar ages and metallicities of central and satellite galaxies depend on stellar mass and halo mass. We find that satellites are older and metal-richer than centrals of the same stellar mass. In addition, the slopes of the age-stellar mass and metallicity-stellar mass relations are found to become shallower in denser environments. This is due to the fact that the average age and metallicity of low mass satellite galaxies increase with the mass of the halo in which they reside. A comparison with the semi-analytical model of Wang et al. (2008) shows that it succesfully reproduces the fact that satellites are older than centrals of the same stellar mass and that the age difference increases with the halo mass of the satellite. This is a consequence of strangulation, which leaves the stellar populations of satellites to evolve passively, while the prolonged star formation activity of centrals keeps their average ages younger. The resulting age offset is larger in more massive environments because their satellites were accreted earlier. The model fails, however, in reproducing the halo mass dependence of the metallicities of low mass satellites, yields metallicity-stellar mass and age-stellar mass relations that are too shallow, and predicts that satellite galaxies have the same metallicities as centrals of the same stellar mass, in disagreement with the data. We argue that these discrepancies are likely to indicate the need to (i) modify the recipes of both supernova feedback and AGN feedback, (ii) use a more realistic description of strangulation, and (iii) include a proper treatment of the tidal stripping, heating and destruction of satellite galaxies. [Abridged]

preprint2009arXiv

The nature of red dwarf galaxies

Using dark matter halos traced by galaxy groups selected from the Sloan Digital Sky Survey Data Release 4, we find that about 1/4 of the faint galaxies ($\rmag >-17.05$, hereafter dwarfs) that are the central galaxies in their own halo are not blue and star forming, as expected in standard models of galaxy formation, but are red. In contrast, this fraction is about 1/2 for dwarf satellite galaxies. Many red dwarf galaxies are physically associated with more massive halos. In total, about $\sim 45$% of red dwarf galaxies reside in massive halos as satellites, while another $\sim 25$% have a spatial distribution that is much more concentrated towards their nearest massive haloes than other dwarf galaxies. We use mock catalogs to show that the reddest population of non-satellite dwarf galaxies are distributed within about 3 times the virial radii of their nearest massive halos. We suggest that this population of dwarf galaxies are hosted by low-mass halos that have passed through their massive neighbors, and that the same environmental effects that cause satellite galaxies to become red are also responsible for the red colors of this population of galaxies. We do not find any significant radial dependence of the population of dwarf galaxies with the highest concentrations, suggesting that the mechanisms operating on these galaxies affect color more than structure. However, over 30% of dwarf galaxies are red and isolated and their origin remains unknown.

preprint2008arXiv

The subhalo - satellite connection and the fate of disrupted satellite galaxies

In the standard paradigm, satellite galaxies are believed to be associated with the population of dark matter subhalos. In this paper, we use the conditional stellar mass functions of {\it satellite galaxies} obtained from a large galaxy group catalogue together with models of the subhalo mass functions to explore the fraction and fate of stripped stars from satellites in galaxy groups and clusters of different masses. The majority of the stripped stars in massive halos are predicted to end up as intra-cluster stars, and the predicted amounts of the intra-cluster component as a function of the velocity dispersion of galaxy system match well the observational results obtained by Gonzalez et al. (2007). The fraction of the mass in the stripped stars to that remain bound in the central and satellite galaxies is the highest ($\sim 40%$ of the total stellar mass) in halos with masses $M_h\sim 10^{14}\msunh$. If all these stars end up in the intra-cluster component (Max), or maximum of them are accreted into the central galaxy (Min), then we can predict that a maximum $\sim 19%$ and a minimum $\sim 5%$ of the total stars in the whole universe are in terms of the diffused intra-cluster component. In the former case, in massive halos with $M_h \sim 10^{15} \msunh$, the stellar mass of the intra-cluster component is roughly 6 times as large as that of the central galaxy. This factor decreases to $\sim 2$, 1 and 0.1 in halos with $M_h \sim 10^{14}$, $10^{13}$, and $10^{12} \msunh$, respectively. The total amount of stars stripped from satellite galaxies is insufficient to build up the central galaxies in halos with masses $\la 10^{12.5}\msunh$, and so the quenching of star formation must occur in halos with higher masses. Abridged.

preprint2007arXiv

The cross-correlation between galaxies of different luminosities and Colors

We study the cross-correlation between galaxies of different luminosities and colors, using a sample selected from the SDSS Dr 4. Galaxies are divided into 6 samples according to luminosity, and each of these samples is divided into red and blue subsamples. Projected auto-correlation and cross-correlation is estimated for these subsample. At projected separations r_p > 1\mpch, all correlation functions are roughly parallel, although the correlation amplitude depends systematically on luminosity and color. On r_p < 1\mpch, the auto- and cross-correlation functions of red galaxies are significantly enhanced relative to the corresponding power laws obtained on larger scales. Such enhancement is absent for blue galaxies and in the cross-correlation between red and blue galaxies. We esimate the relative bias factor on scales r > 1\mpch for each subsample using its auto-correlation function and cross-correlation functions. The relative bias factors obtained from different methods are similar. For blue galaxies the luminosity-dependence of the relative bias is strong over the luminosity range probed (-23.0<M_r < -18.0),but for red galaxies the dependence is weaker and becomes insignificant for luminosities below L^*. To examine whether a significant stochastic/nonlinear component exists in the bias relation, we study the ratio R_ij= W_{ii}W_{jj}/W_{ij}^2, where W_{ij} is the projected correlation between subsample i and j. We find that the values of R_ij are all consistent with 1 for all-all, red-red and blue-blue samples, however significantly larger than 1 for red-blue samples. For faint red - faint blue samples the values of R_{ij} are as high as ~ 2 on small scales r_p < 1 \mpch and decrease with increasing r_p.