Source author record

Shan Huang

Shan Huang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

24works
21topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

24 published item(s)

preprint2026arXiv

ARC: Active and Reflection-driven Context Management for Long-Horizon Information Seeking Agents

Large language models are increasingly deployed as research agents for deep search and long-horizon information seeking, yet their performance often degrades as interaction histories grow. This degradation, known as context rot, reflects a failure to maintain coherent and task-relevant internal states over extended reasoning horizons. Existing approaches primarily manage context through raw accumulation or passive summarization, treating it as a static artifact and allowing early errors or misplaced emphasis to persist. Motivated by this perspective, we propose ARC, which is the first framework to systematically formulate context management as an active, reflection-driven process that treats context as a dynamic internal reasoning state during execution. ARC operationalizes this view through reflection-driven monitoring and revision, allowing agents to actively reorganize their working context when misalignment or degradation is detected. Experiments on challenging long-horizon information-seeking benchmarks show that ARC consistently outperforms passive context compression methods, achieving up to an 11% absolute improvement in accuracy on BrowseComp-ZH with Qwen2.5-32B-Instruct.

preprint2026arXiv

CM-EVS: Sparse Panoramic RGB-D-Pose Data for Complete Scene Coverage

Modern 3D visual learning relies on observations sampled from metric 3D assets, yet existing scans, meshes, point clouds, simulations, and reconstructions do not directly provide a sparse, comparable, and geometry-consistent panoramic training interface. Dense trajectories duplicate nearby views, source-specific rendering policies yield heterogeneous annotations, and sparse heuristics may miss important regions or introduce depth-inconsistent observations. We study how to convert 3D assets into sparse panoramic RGB-D-pose data that preserves complete scene coverage with low redundancy and auditable provenance. We propose COVER (Coverage-Oriented Viewpoint curation with ERP Range-depth warping), a training-free ERP viewpoint curator that projects geometry observed from selected views into candidate ERP probes, scores incremental coverage, and penalizes depth conflicts. Under bounded proxy error, its greedy coverage proxy preserves the standard coverage-style approximation behavior up to an additive error term. Using COVER, we build CM-EVS (Coverage-curated Metric ERP View Set), a panoramic RGB-D-pose dataset with 36,373 curated ERP frames from 1,275 indoor scenes across Blender indoor, HM3D, and ScanNet++, complemented by outdoor panoramas from TartanGround and OB3D re-encoded into the same schema. Each frame provides full-sphere RGB, metric range depth, calibrated pose; COVER-produced indoor frames include per-step provenance logs. With a median of only 25 frames per indoor scene, CM-EVS covers all 13 unified room types while maintaining compact scene-level coverage. Experiments show that COVER improves the coverage-conflict trade-off, making CM-EVS a sparse, compact, and auditable RGB-D-pose resource for geometry-consistent panoramic 3D learning.

preprint2026arXiv

Layout optimization for the LUXE-NPOD experiment

Beam dump experiments represent an effective way to probe new physics in a parameter space, where new particles have feeble couplings to the Standard Model sector and masses below the GeV scale. The LUXE experiment, designed primarily to study strong-field quantum electrodynamics, can be used also as a photon beam dump experiment with a unique reach for new spin-0 particles in the $10-350~\mathrm{MeV}$ mass and $10^{-6}-10^{-3}~\mathrm{GeV}^{-1}$ couplings to photons ranges. This is achieved via the ``New Physics search with Optical Dump'' (NPOD) concept. While prior estimations were obtained with a simplified model of the experimental setup, in this work we present a systematic study of the new physics reach in the full, realistic experimental apparatus, including an existing detector to be used in the LUXE NPOD context. We furthermore investigate updated scenarios of LUXE's experimental plan and confirm that our results are in agreement with the original estimations of a background-free operation.

preprint2022arXiv

Axionlike-particle generation by laser-plasma interaction

The hypothetical axion and axion-like particles, feebly coupled with photon, have not yet been found in any experiment. With the improvement of laser technique, much stronger but shorter quasi-static electric and magnetic fields can be created in laboratory using laser-plasma interaction, compared to the fields of large magnets, to help the search of axion. In this article, we discuss the feasibility of ALPs exploration using planarly or cylindrically symmetric laser-plasma fields as background and an x-ray free-electron laser as probe. Both the probe and the background fields are polarized such that the existence of ALPs in the corresponding parameter space will cause polarization rotation of the probe, which can be detected with high accuracy. Besides, a structured field in the plasma creates a tunable transverse profile for the interaction and improves the signal-to-noise ratio via phase-matching mechanism. The ALP mass discussed in this article ranges from $10^{-3}$ eV to 1 keV. Some simple schemes and estimations on ALP production and polarization rotation of probe photon are given, which reveals the possibility of future laser-plasma ALP source in laboratory.

preprint2022arXiv

Building Embedded Systems Like It's 1996

Embedded devices are ubiquitous. However, preliminary evidence shows that attack mitigations protecting our desktops/servers/phones are missing in embedded devices, posing a significant threat to embedded security. To this end, this paper presents an in-depth study on the adoption of common attack mitigations on embedded devices. Precisely, it measures the presence of standard mitigations against memory corruptions in over 10k Linux-based firmware of deployed embedded devices. The study reveals that embedded devices largely omit both user-space and kernel-level attack mitigations. The adoption rates on embedded devices are multiple times lower than their desktop counterparts. An equally important observation is that the situation is not improving over time. Without changing the current practices, the attack mitigations will remain missing, which may become a bigger threat in the upcoming IoT era. Throughout follow-up analyses, we further inferred a set of factors possibly contributing to the absence of attack mitigations. The exemplary ones include massive reuse of non-protected software, lateness in upgrading outdated kernels, and restrictions imposed by automated building tools. We envision these will turn into insights towards improving the adoption of attack mitigations on embedded devices in the future.

preprint2022arXiv

Deep learning study of an electromagnetic calorimeter

The accurate and precise extraction of information from a modern particle physics detector, such as an electromagnetic calorimeter, may be complicated and challenging. In order to overcome the difficulties we propose processing the detector output using the deep-learning methodology. Our algorithmic approach makes use of a known network architecture, which is being modified to fit the problems at hand. The results are of high quality (biases of order 2%) and, moreover, indicate that most of the information may be derived from only a fraction of the detector. We conclude that such an analysis helps us understanding the essential mechanism of the detector and should be performed as a part of its designing procedure.

preprint2022arXiv

Emotions in Online Content Diffusion

Social media-transmitted online information, which is associated with emotional expressions, shapes our thoughts and actions. In this study, we incorporate social network theories and analyses and use a computational approach to investigate how emotional expressions, particularly \textit{negative discrete emotional expressions} (i.e., anxiety, sadness, anger, and disgust), lead to differential diffusion of online content in social media networks. We rigorously quantify diffusion cascades' structural properties (i.e., size, depth, maximum breadth, and structural virality) and analyze the individual characteristics (i.e., age, gender, and network degree) and social ties (i.e., strong and weak) involved in the cascading process. In our sample, more than six million unique individuals transmitted 387,486 randomly selected articles in a massive-scale online social network, WeChat. We detect the expression of discrete emotions embedded in these articles, using a newly generated domain-specific and up-to-date emotion lexicon. We apply a partial-linear instrumental variable approach with a double machine learning framework to causally identify the impact of the negative discrete emotions on online content diffusion. We find that articles with more expressions of anxiety spread to a larger number of individuals and diffuse more deeply, broadly, and virally. Expressions of anger and sadness, however, reduce cascades' size and maximum breadth. We further show that the articles with different degrees of negative emotional expressions tend to spread differently based on individual characteristics and social ties. Our results shed light on content marketing and regulation, utilizing negative emotional expressions.

preprint2022arXiv

LUXE: A new experiment to study non-perturbative QED in electron-laser and photon-laser collisions

LUXE (Laser Und XFEL Experiment) is a new experiment in planning at DESY Hamburg using the electron beam of the European XFEL. LUXE is intended to study collisions between a high-intensity optical laser and 16.5 GeV electrons from the XFEL electron beam, as well as collisions between the optical laser and high-energy secondary photons. The physics objective of LUXE are processes of quantum electrodynamics (QED) at the strong-field frontier, where the electromagnetic field of the laser is above the Schwinger limit. In this regime, QED is non-perturbative. This manifests itself in the creation of physical electron-positron pairs from the QED vacuum, similar to Hawking radiation from black holes. LUXE intends to measure the positron production rate in an unprecedented laser intensity regime. An overview of the LUXE experimental setup and its challenges will be given, followed by a discussion of the expected physics reach in the context of testing QED in the non-perturbative regime.

preprint2022arXiv

WebUAV-3M: A Benchmark for Unveiling the Power of Million-Scale Deep UAV Tracking

Unmanned aerial vehicle (UAV) tracking is of great significance for a wide range of applications, such as delivery and agriculture. Previous benchmarks in this area mainly focused on small-scale tracking problems while ignoring the amounts of data, types of data modalities, diversities of target categories and scenarios, and evaluation protocols involved, greatly hiding the massive power of deep UAV tracking. In this work, we propose WebUAV-3M, the largest public UAV tracking benchmark to date, to facilitate both the development and evaluation of deep UAV trackers. WebUAV-3M contains over 3.3 million frames across 4,500 videos and offers 223 highly diverse target categories. Each video is densely annotated with bounding boxes by an efficient and scalable semiautomatic target annotation (SATA) pipeline. Importantly, to take advantage of the complementary superiority of language and audio, we enrich WebUAV-3M by innovatively providing both natural language specifications and audio descriptions. We believe that such additions will greatly boost future research in terms of exploring language features and audio cues for multimodal UAV tracking. In addition, a fine-grained UAV tracking-under-scenario constraint (UTUSC) evaluation protocol and seven challenging scenario subtest sets are constructed to enable the community to develop, adapt and evaluate various types of advanced trackers. We provide extensive evaluations and detailed analyses of 43 representative trackers and envision future research directions in the field of deep UAV tracking and beyond. The dataset, toolkits and baseline results are available at \url{https://github.com/983632847/WebUAV-3M}.

preprint2020arXiv

A Multi-oriented Chinese Keyword Spotter Guided by Text Line Detection

Chinese keyword spotting is a challenging task as there is no visual blank for Chinese words. Different from English words which are split naturally by visual blanks, Chinese words are generally split only by semantic information. In this paper, we propose a new Chinese keyword spotter for natural images, which is inspired by Mask R-CNN. We propose to predict the keyword masks guided by text line detection. Firstly, proposals of text lines are generated by Faster R-CNN;Then, text line masks and keyword masks are predicted by segmentation in the proposals. In this way, the text lines and keywords are predicted in parallel. We create two Chinese keyword datasets based on RCTW-17 and ICPR MTWI2018 to verify the effectiveness of our method.

preprint2020arXiv

Application of Seq2Seq Models on Code Correction

We apply various seq2seq models on programming language correction tasks on Juliet Test Suite for C/C++ and Java of Software Assurance Reference Datasets(SARD), and achieve 75\%(for C/C++) and 56\%(for Java) repair rates on these tasks. We introduce Pyramid Encoder in these seq2seq models, which largely increases the computational efficiency and memory efficiency, while remain similar repair rate to their non-pyramid counterparts. We successfully carry out error type classification task on ITC benchmark examples (with only 685 code instances) using transfer learning with models pre-trained on Juliet Test Suite, pointing out a novel way of processing small programing language datasets.

preprint2019arXiv

Host Galaxies of Type Ic and Broad-lined Type Ic Supernovae from the Palomar Transient Factory: Implication for Jet Production

Unlike the ordinary supernovae (SNe) some of which are hydrogen and helium deficient (called Type Ic SNe), broad-lined Type Ic SNe (SNe Ic-bl) are very energetic events, and all SNe coincident with bona fide long duration gamma-ray bursts (LGRBs) are of Type Ic-bl. Understanding the progenitors and the mechanism driving SN Ic-bl explosions vs those of their SNe Ic cousins is key to understanding the SN-GRB relationship and jet production in massive stars. Here we present the largest set of host-galaxy spectra of 28 SNe Ic and 14 SN Ic-bl, all discovered before 2013 by the same untargeted survey, namely the Palomar Transient Factory (PTF). We carefully measure their gas-phase metallicities, stellar masses (M*s) and star-formation rates (SFRs) by taking into account recent progress in the metallicity field and propagating uncertainties correctly. We further re-analyze the hosts of 10 literature SN-GRBs using the same methods and compare them to our PTF SN hosts with the goal of constraining their progenitors from their local environments by conducting a thorough statistical comparison, including upper limits. We find that the metallicities, SFRs and M*s of our PTF SN Ic-bl hosts are statistically comparable to those of SN-GRBs, but significantly lower than those of the PTF SNe Ic. The mass-metallicity relations as defined by the SNe Ic-bl and SN-GRBs are not significantly different from the same relations as defined by the SDSS galaxies, in contrast to claims by earlier works. Our findings point towards low metallicity as a crucial ingredient for SN Ic-bl and SN-GRB production since we are able to break the degeneracy between high SFR and low metallicity. We suggest that the PTF SNe Ic-bl may have produced jets that were choked inside the star or were able break out of the star as unseen low-luminosity or off-axis GRBs.

preprint2016arXiv

Frequency Estimation of Multiple Sinusoids with Sub-Nyquist Sampling Sequences

In some applications of frequency estimation, the frequencies of multiple sinusoids are required to be estimated from sub-Nyquist sampling sequences. In this paper, we propose a novel method based on subspace techniques to estimate the frequencies by using under-sampled samples. We analyze the impact of under-sampling and demonstrate that three sub-Nyquist sequences are general enough to estimate the frequencies under some condition. The frequencies estimated from one sequence are unfolded in frequency domain, and then the other two sequences are used to pick the correct frequencies from all possible frequencies. Simulations illustrate the validity of the theory. Numerical results show that this method is feasible and accurate at quite low sampling rates.

preprint2016arXiv

HIghMass - High HI Mass, HI-Rich Galaxies at $z\sim0$: Combined HI and H$_2$ Observations

We present resolved HI and CO observations of three galaxies from the HIghMass sample, a sample of HI-massive ($M_{HI} > 10^{10} M_\odot$), gas-rich ($M_{HI}$ in top $5\%$ for their $M_*$) galaxies identified in the ALFALFA survey. Despite their high gas fractions, these are not low surface brightness galaxies, and have typical specific star formation rates (SFR$/M_*$) for their stellar masses. The three galaxies have normal star formation rates for their HI masses, but unusually short star formation efficiency scale lengths, indicating that the star formation bottleneck in these galaxies is in the conversion of HI to H$_2$, not in converting H$_2$ to stars. In addition, their dark matter spin parameters ($λ$) are above average, but not exceptionally high, suggesting that their star formation has been suppressed over cosmic time but are now becoming active, in agreement with prior H$α$ observations.

preprint2016arXiv

Line Spectral Estimation Based on Compressed Sensing with Deterministic Sub-Nyquist Sampling

As an alternative to the traditional sampling theory, compressed sensing allows acquiring much smaller amount of data, still estimating the spectra of frequency-sparse signals accurately. However, compressed sensing usually requires random sampling in data acquisition, which is difficult to implement in hardware. In this paper, we propose a deterministic and simple sampling scheme, that is, sampling at three sub-Nyquist rates which have coprime undersampled ratios. This sampling method turns out to be valid through numerical experiments. A complex-valued multitask algorithm based on variational Bayesian inference is proposed to estimate the spectra of frequency-sparse signals after sampling. Simulations show that this method is feasible and robust at quite low sampling rates.

preprint2015arXiv

A Class of Deterministic Sensing Matrices and Their Application in Harmonic Detection

In this paper, a class of deterministic sensing matrices are constructed by selecting rows from Fourier matrices. These matrices have better performance in sparse recovery than random partial Fourier matrices. The coherence and restricted isometry property of these matrices are given to evaluate their capacity as compressive sensing matrices. In general, compressed sensing requires random sampling in data acquisition, which is difficult to implement in hardware. By using these sensing matrices in harmonic detection, a deterministic sampling method is provided. The frequencies and amplitudes of the harmonic components are estimated from under-sampled data. The simulations show that this under-sampled method is feasible and valid in noisy environments.

preprint2015arXiv

Joint Frequency Estimation with Two Sub-Nyquist Sampling Sequences

In many applications of frequency estimation, the frequencies of the signals are so high that the data sampled at Nyquist rate are hard to acquire due to hardware limitation. In this paper, we propose a novel method based on subspace techniques to estimate the frequencies by using two sub-Nyquist sample sequences, provided that the two under-sampled ratios are relatively prime integers. We analyze the impact of under-sampling and expand the estimated frequencies which suffer from aliasing. Through jointing the results estimated from these two sequences, the frequencies approximate to the frequency components really contained in the signals are screened. The method requires a small quantity of hardware and calculation. Numerical results show that this method is valid and accurate at quite low sampling rates.

preprint2014arXiv

HIghMass - High HI Mass, HI-rich Galaxies at z~0: High-Resolution VLA Imaging of UGC 9037 and UGC 12506

We present resolved HI observations of two galaxies, UGC 9037 and UGC 12506, members of a rare subset of galaxies detected by the ALFALFA extragalactic HI survey characterized by high HI mass and high gas fraction for their stellar masses. Both of these galaxies have M$_*>10^{10}$ M$_\odot$ and M$_\text{HI}>$ M$_*$, as well as typical star formation rates for their stellar masses. How can such galaxies have avoided consuming their massive gas reservoirs? From gas kinematics, stability, star formation, and dark matter distributions of the two galaxies, we infer two radically different histories. UGC 9037 has high central HI surface density ($>10$ M$_\odot$ pc$^{-2}$). Its gas at most radii appears to be marginally unstable with non-circular flows across the disk. These properties are consistent with UGC 9037 having recently acquired its gas and that it will soon undergo major star formation. UGC 12506 has low surface densities of HI, and its gas is stable over most of the disk. We predict its gas to be HI-dominated at all except the smallest radii. We claim a very high dark matter halo spin parameter for UGC 12506 ($λ=0.15$), suggesting that its gas is older, and has never undergone a period of star formation significant enough to consume the bulk of its gas.

preprint2014arXiv

HIghMass -- High HI Mass, HI-rich Galaxies at z~0: Sample Definition, Optical and Halpha Imaging, and Star Formation Properties

We present first results of the study of a set of exceptional HI sources identified in the 40% ALFALFA extragalactic HI survey catalog alpha.40 as being both HI massive (M_HI > 10^10 Msun) and having high gas fractions for their stellar masses: the HIghMass galaxy sample. We analyze UV- and optical-broadband and Halpha images to understand the nature of their relatively underluminous disks in optical and to test whether their high gas fractions can be tracked to higher dark matter halo spin parameters or late gas accretion. Estimates of their star formation rates (SFRs) based on SED-fitting agree within uncertainties with the Halpha luminosity inferred SFRs. The HII region luminosity functions have standard slopes at the luminous end. The global SFRs demonstrate that the HIghMass galaxies exhibit active ongoing star formation (SF) with moderate SF efficiency, but relative to normal spirals, a lower integrated SFR in the past. Because the SF activity in these systems is spread throughout their extended disks, they have overall lower SFR surface densities and lower surface brightness in the optical bands. Relative to normal disk galaxies, the majority of HIghMass galaxies have higher Halpha equivalent widths and are bluer in their outer disks, implying an inside-out disk growth scenario. Downbending double exponential disks are more frequent than upbending disks among the gas-rich galaxies, suggesting that SF thresholds exist in the downbending disks, probably as a result of concentrated gas distribution.

preprint2012arXiv

A direct measurement of the baryonic mass function of galaxies & implications for the galactic baryon fraction

We use both an HI-selected and an optically-selected galaxy sample to directly measure the abundance of galaxies as a function of their "baryonic" mass (stars + atomic gas). Stellar masses are calculated based on optical data from the Sloan Digital Sky Survey (SDSS) and atomic gas masses are calculated using atomic hydrogen (HI) emission line data from the Arecibo Legacy Fast ALFA (ALFALFA) survey. By using the technique of abundance matching, we combine the measured baryonic function (BMF) of galaxies with the dark matter halo mass function in a LCDM universe, in order to determine the galactic baryon fraction as a function of host halo mass. We find that the baryon fraction of low-mass halos is much smaller than the cosmic value, even when atomic gas is taken into account. We find that the galactic baryon deficit increases monotonically with decreasing halo mass, in contrast with previous studies which suggested an approximately constant baryon fraction at the low-mass end. We argue that the observed baryon fractions of low mass halos cannot be explained by reionization heating alone, and that additional feedback mechanisms (e.g. supernova blowout) must be invoked. However, the outflow rates needed to reproduce our result are not easily accommodated in the standard picture of galaxy formation in a LCDM universe.

preprint2012arXiv

Gas-Bearing Early-Type Dwarf Galaxies in Virgo: Evidence for Recent Accretion

We investigate the dwarf (M_B> -16) galaxies in the Virgo cluster in the radio, optical, and ultraviolet regimes. Of the 365 galaxies in this sample, 80 have been detected in HI by the Arecibo Legacy Fast ALFA survey. These detections include 12 early-type dwarfs which have HI and stellar masses similar to the cluster dwarf irregulars and BCDs. In this sample of 12, half have star-formation properties similar to late type dwarfs, while the other half are quiescent like typical early-type dwarfs. We also discuss three possible mechanisms for their evolution: that they are infalling field galaxies that have been or are currently being evolved by the cluster, that they are stripped objects whose gas is recycled, and that the observed HI has been recently reaccreted. Evolution by the cluster adequately explains the star-forming half of the sample, but the quiescent class of early-type dwarfs is most consistent with having recently reaccreted their gas.

preprint2012arXiv

The Arecibo Legacy Fast ALFA Survey: The Galaxy Population Detected by ALFALFA

Making use of HI 21 cm line measurements from the ALFALFA survey (alpha.40) and photometry from the Sloan Digital Sky Survey (SDSS) and GALEX, we investigate the global scaling relations and fundamental planes linking stars and gas for a sample of 9417 common galaxies: the alpha.40-SDSS-GALEX sample. In addition to their HI properties derived from the ALFALFA dataset, stellar masses (M_*) and star formation rates (SFRs) are derived from fitting the UV-optical spectral energy distributions. 96% of the alpha.40-SDSS-GALEX galaxies belong to the blue cloud, with the average gas fraction f_HI = M_HI/M_* ~ 1.5. A transition in SF properties is found whereby below M_* ~ 10^9.5 M_sun, the slope of the star forming sequence changes, the dispersion in the specific star formation rate (SSFR) distribution increases and the star formation efficiency (SFE) mildly increases with M_*. The evolutionary track in the SSFR-M_* diagram, as well as that in the color magnitude diagram are linked to the HI content; below this transition mass, the star formation is regulated strongly by the HI. Comparison of HI- and optically-selected samples over the same restricted volume shows that the HI-selected population is less evolved and has overall higher SFR and SSFR at a given stellar mass, but lower SFE and extinction, suggesting either that a bottleneck exists in the HI to H_2 conversion, or that the process of SF in the very HI-dominated galaxies obeys an unusual, low efficiency star formation law. A trend is found that, for a given stellar mass, high gas fraction galaxies reside preferentially in dark matter halos with high spin parameters. Because it represents a full census of HI-bearing galaxies at z~0, the scaling relations and fundamental planes derived for the ALFALFA population can be used to assess the HI detection rate by future blind HI surveys and intensity mapping experiments at higher redshift.

preprint2011arXiv

The Arecibo Legacy Fast ALFA Survey: The alpha.40 HI Source Catalog, its Characteristics and their Impact on the Derivation of the HI Mass Function

We present a current catalog of 21 cm HI line sources extracted from the Arecibo Legacy Fast Arecibo L-band Feed Array (ALFALFA) survey over ~2800 square degrees of sky: the alpha.40 catalog. Covering 40% of the final survey area, the alpha.40 catalog contains 15855 sources in the regions 07h30m < R.A. < 16h30m, +04 deg < Dec. < +16 deg and +24 deg < Dec. < +28 deg and 22h < R.A. < 03h, +14 deg < Dec. < +16 deg and +24 deg < Dec. < +32 deg. Of those, 15041 are certainly extragalactic, yielding a source density of 5.3 galaxies per square degree, a factor of 29 improvement over the catalog extracted from the HI Parkes All Sky Survey. In addition to the source centroid positions, HI line flux densities, recessional velocities and line widths, the catalog includes the coordinates of the most probable optical counterpart of each HI line detection, and a separate compilation provides a crossmatch to identifications given in the photometric and spectroscopic catalogs associated with the Sloan Digital Sky Survey Data Release 7. Fewer than 2% of the extragalactic HI line sources cannot be identified with a feasible optical counterpart; some of those may be rare OH megamasers at 0.16 < z < 0.25. A detailed analysis is presented of the completeness, width dependent sensitivity function and bias inherent in the current alpha.40 catalog. The impact of survey selection, distance errors, current volume coverage and local large scale structure on the derivation of the HI mass function is assessed. While alpha.40 does not yet provide a completely representative sampling of cosmological volume, derivations of the HI mass function using future data releases from ALFALFA will further improve both statistical and systematic uncertainties.

preprint2011arXiv

The Survey of HI in Extremely Low-mass Dwarfs (SHIELD)

We present first results from the "Survey of HI in Extremely Low-mass Dwarfs" (SHIELD), a multi-configuration EVLA study of the neutral gas contents and dynamics of galaxies with HI masses in the 10^6-10^7 Solar mass range detected by the Arecibo Legacy Fast ALFA (ALFALFA) survey. We describe the survey motivation and concept demonstration using VLA imaging of 6 low-mass galaxies detected in early ALFALFA data products. We then describe the primary scientific goals of SHIELD and present preliminary EVLA and WIYN 3.5m imaging of the 12 SHIELD galaxies. With only a few exceptions, the neutral gas distributions of these extremely low-mass galaxies are centrally concentrated. In only 1 system have we detected HI column densities higher than 10^21 cm^-2. Despite this, the stellar populations of all of these systems are dominated by blue stars. Further, we find ongoing star formation as traced by H alpha emission in 10 of the 11 galaxies with H alpha imaging obtained to date. Taken together these results suggest that extremely low-mass galaxies are forming stars in conditions different from those found in more massive systems. While detailed dynamical analysis requires the completion of data acquisition, the most well-resolved system is amenable to meaningful position-velocity analysis. For AGC 749237, we find well-ordered rotation of 30 km/s at a distance of ~40 arcseconds from the dynamical center. At the adopted distance of 3.2 Mpc, this implies the presence of a >1x10^8 Solar mass dark matter halo and a baryon fraction < ~0.1.