Source author record

Ryan Price

Ryan Price appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

astro-ph.IM Computation and Language eess.AS Sound

Catalog footprint

What is connected

4works

4topics

4close collaborators

Actions

Connect this record

Open graph Browse works

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

preprint2022arXiv

Cross-stitched Multi-modal Encoders

In this paper, we propose a novel architecture for multi-modal speech and text input. We combine pretrained speech and text encoders using multi-headed cross-modal attention and jointly fine-tune on the target problem. The resultant architecture can be used for continuous token-level classification or utterance-level prediction acting on simultaneous text and speech. The resultant encoder efficiently captures both acoustic-prosodic and lexical information. We compare the benefits of multi-headed attention-based fusion for multi-modal utterance-level classification against a simple concatenation of pre-pooled, modality-specific representations. Our model architecture is compact, resource efficient, and can be trained on a single consumer GPU card.

preprint2022arXiv

Seq-2-Seq based Refinement of ASR Output for Spoken Name Capture

Person name capture from human speech is a difficult task in human-machine conversations. In this paper, we propose a novel approach to capture the person names from the caller utterances in response to the prompt "say and spell your first/last name". Inspired from work on spell correction, disfluency removal and text normalization, we propose a lightweight Seq-2-Seq system which generates a name spell from a varying user input. Our proposed method outperforms the strong baseline which is based on LM-driven rule-based approach.

preprint2011arXiv

Haar wavelets as a tool for the statistical characterization of variability

In the field of gamma-ray astronomy, irregular and noisy datasets make difficult the characterization of light-curve features in terms of statistical significance while properly accounting for trial factors associated with the search for variability at different times and over different timescales. In order to address these difficulties, we propose a method based on the Haar wavelet decomposition of the data. It allows statistical characterization of possible variability, embedded in a white noise background, in terms of a confidence level. The method is applied to artificially generated data for characterization as well as to the the very high energy M87 light curve recorded with VERITAS in 2008 which serves here as a realistic application example.

preprint2010arXiv

Stellar intensity interferometry: Experimental steps toward long-baseline observations

Experiments are in progress to prepare for intensity interferometry with arrays of air Cherenkov telescopes. At the Bonneville Seabase site, near Salt Lake City, a testbed observatory has been set up with two 3-m air Cherenkov telescopes on a 23-m baseline. Cameras are being constructed, with control electronics for either off- or online analysis of the data. At the Lund Observatory (Sweden), in Technion (Israel) and at the University of Utah (USA), laboratory intensity interferometers simulating stellar observations have been set up and experiments are in progress, using various analog and digital correlators, reaching 1.4 ns time resolution, to analyze signals from pairs of laboratory telescopes.