Source author record

Kevin Tang

Kevin Tang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

7works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

7 published item(s)

preprint2026arXiv

Evaluating the cognitive reality of Spanish irregular morphomic patterns: Humans vs. Transformers

Do transformer models generalize morphological patterns like humans do? We investigate this by directly comparing transformers to human behavioral data on Spanish irregular morphomic patterns from \citet{Nevins2015TheRA}. We adopt the same analytical framework as the original human study. Under controlled input conditions, we evaluate whether transformer models can replicate human-like sensitivity to the morphome, a complex linguistic phenomenon. Our experiments focus on three frequency conditions: natural, low-frequency, and high-frequency distributions of verbs exhibiting irregular morphomic patterns. Transformer models achieve higher stem-accuracy than human participants. However, response preferences diverge: humans consistently favor the "natural" inflection across all items, whereas models preferred the irregular forms, and their choices are modulated by the proportion of irregular verbs present during training. Moreover, models trained on the natural and low-frequency distributions, but not the high-frequency distribution, exhibit sensitivity to phonological similarity between test items and Spanish L-shaped verbs, mirroring a limited aspect of human phonological generalization.

preprint2022arXiv

Investigating the Nature of the Luminous Ambiguous Nuclear Transient ASASSN-17jz

We present observations of the extremely luminous but ambiguous nuclear transient (ANT) ASASSN-17jz, spanning roughly 1200 days of the object's evolution. ASASSN-17jz was discovered by the All-Sky Automated Survey for Supernovae (ASAS-SN) in the galaxy SDSS J171955.84+414049.4 on UT 2017 July 27 at a redshift of $z=0.1641$. The transient peaked at an absolute $B$-band magnitude of $M_{B,{\rm peak}}=-22.81$, corresponding to a bolometric luminosity of $L_{\rm bol,peak}=8.3\times10^{44}$~erg~s$^{-1}$, and exhibited late-time ultraviolet emission that was still ongoing in our latest observations. Integrating the full light curve gives a total emitted energy of $E_{\rm tot}=(1.36\pm0.08)\times10^{52}$~erg, with $(0.80\pm0.02)\times10^{52}$~erg of this emitted within 200 days of peak light. This late-time ultraviolet emission is accompanied by increasing X-ray emission that becomes softer as it brightens. ASASSN-17jz exhibited a large number of spectral emission lines most commonly seen in active galactic nuclei (AGNs) with little evidence of evolution. It also showed transient Balmer features which became fainter and broader over time, and are still being detected $>1000$ days after peak brightness. We consider various physical scenarios for the origin of the transient, including supernovae (SNe), tidal disruption events (TDEs), AGN outbursts, and ANTs. We find that the most likely explanation is that ASASSN-17jz was an SN~IIn occurring in or near the disk of an existing AGN, and that the late-time emission is caused by the AGN transitioning to a more active state.

preprint2022arXiv

The Lick Observatory Supernova Search follow-up program: photometry data release of 70 stripped-envelope supernovae

We present BVRI and unfiltered Clear light curves of 70 stripped-envelope supernovae (SESNe), observed between 2003 and 2020, from the Lick Observatory Supernova Search (LOSS) follow-up program. Our SESN sample consists of 19 spectroscopically normal SNe~Ib, two peculiar SNe Ib, six SN Ibn, 14 normal SNe Ic, one peculiar SN Ic, ten SNe Ic-BL, 15 SNe IIb, one ambiguous SN IIb/Ib/c, and two superluminous SNe. Our follow-up photometry has (on a per-SN basis) a mean coverage of 81 photometric points (median of 58 points) and a mean cadence of 3.6d (median of 1.2d). From our full sample, a subset of 38 SNe have pre-maximum coverage in at least one passband, allowing for the peak brightness of each SN in this subset to be quantitatively determined. We describe our data collection and processing techniques, with emphasis toward our automated photometry pipeline, from which we derive publicly available data products to enable and encourage further study by the community. Using these data products, we derive host-galaxy extinction values through the empirical colour evolution relationship and, for the first time, produce accurate rise-time measurements for a large sample of SESNe in both optical and infrared passbands. By modeling multiband light curves, we find that SNe Ic tend to have lower ejecta masses and lower ejecta velocities than SNe~Ib and IIb, but higher $^{56}$Ni masses.

preprint2020arXiv

Prosody leaks into the memories of words

The average predictability (aka informativity) of a word in context has been shown to condition word duration (Seyfarth, 2014). All else being equal, words that tend to occur in more predictable environments are shorter than words that tend to occur in less predictable environments. One account of the informativity effect on duration is that the acoustic details of probabilistic reduction are stored as part of a word's mental representation. Other research has argued that predictability effects are tied to prosodic structure in integral ways. With the aim of assessing a potential prosodic basis for informativity effects in speech production, this study extends past work in two directions; it investigated informativity effects in another large language, Mandarin Chinese, and broadened the study beyond word duration to additional acoustic dimensions, pitch and intensity, known to index prosodic prominence. The acoustic information of content words was extracted from a large telephone conversation speech corpus with over 400,000 tokens and 6,000 word types spoken by 1,655 individuals and analyzed for the effect of informativity using frequency statistics estimated from a 431 million word subtitle corpus. Results indicated that words with low informativity have shorter durations, replicating the effect found in English. In addition, informativity had significant effects on maximum pitch and intensity, two phonetic dimensions related to prosodic prominence. Extending this interpretation, these results suggest that predictability is closely linked to prosodic prominence, and that the lexical representation of a word includes phonetic details associated with its average prosodic prominence in discourse. In other words, the lexicon absorbs prosodic influences on speech production.

preprint2019arXiv

SN 2017cfd: A Normal Type Ia Supernova Discovered Very Young

The Type~Ia supernova (SN~Ia) 2017cfd in IC~0511 (redshift z = 0.01209+- 0.00016$) was discovered by the Lick Observatory Supernova Search 1.6+-0.7 d after the fitted first-light time (FFLT; 15.2 d before B-band maximum brightness). Photometric and spectroscopic follow-up observations show that SN~2017cfd is a typical, normal SN~Ia with a peak luminosity MB ~ -19.2+-0.2 mag, Delta m15(B) = 1.16 mag, and reached a B-band maximum ~16.8 d after the FFLT. We estimate there to be moderately strong host-galaxy extinction (A_V = 0.39 +- 0.03 mag) based on MLCS2k2 fitting. The spectrum reveals a Si~II lambda 6355 velocity of ~11,200 kms at peak brightness. The analysis shows that SN~2017cfd is a very typical, normal SN Ia in nearly every aspect. SN~2017cfd was discovered very young, with multiband data taken starting 2 d after the FFLT, making it a valuable complement to the currently small sample (fewer than a dozen) of SNe~Ia with color data at such early times. We find that its intrinsic early-time (B - V)0 color evolution belongs to the "blue" population rather than to the distinct "red" population. Using the photometry, we constrain the companion star radius to be < 2.5 R_sun, thus ruling out a red-giant companion.

preprint2015arXiv

Improving Image Classification with Location Context

With the widespread availability of cellphones and cameras that have GPS capabilities, it is common for images being uploaded to the Internet today to have GPS coordinates associated with them. In addition to research that tries to predict GPS coordinates from visual features, this also opens up the door to problems that are conditioned on the availability of GPS coordinates. In this work, we tackle the problem of performing image classification with location context, in which we are given the GPS coordinates for images in both the train and test phases. We explore different ways of encoding and extracting features from the GPS coordinates, and show how to naturally incorporate these features into a Convolutional Neural Network (CNN), the current state-of-the-art for most image classification and recognition problems. We also show how it is possible to simultaneously learn the optimal pooling radii for a subset of our features within the CNN framework. To evaluate our model and to help promote research in this area, we identify a set of location-sensitive concepts and annotate a subset of the Yahoo Flickr Creative Commons 100M dataset that has GPS coordinates with these concepts, which we make publicly available. By leveraging location context, we are able to achieve almost a 7% gain in mean average precision.

preprint2015arXiv

Learning Temporal Embeddings for Complex Video Analysis

In this paper, we propose to learn temporal embeddings of video frames for complex video analysis. Large quantities of unlabeled video data can be easily obtained from the Internet. These videos possess the implicit weak label that they are sequences of temporally and semantically coherent images. We leverage this information to learn temporal embeddings for video frames by associating frames with the temporal context that they appear in. To do this, we propose a scheme for incorporating temporal context based on past and future frames in videos, and compare this to other contextual representations. In addition, we show how data augmentation using multi-resolution samples and hard negatives helps to significantly improve the quality of the learned embeddings. We evaluate various design decisions for learning temporal embeddings, and show that our embeddings can improve performance for multiple video tasks such as retrieval, classification, and temporal order recovery in unconstrained Internet video.