Source author record

Ashish Arora

Ashish Arora appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2020arXiv

CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings

Following the success of the 1st, 2nd, 3rd, 4th and 5th CHiME challenges we organize the 6th CHiME Speech Separation and Recognition Challenge (CHiME-6). The new challenge revisits the previous CHiME-5 challenge and further considers the problem of distant multi-microphone conversational speech diarization and recognition in everyday home environments. Speech material is the same as the previous CHiME-5 recordings except for accurate array synchronization. The material was elicited using a dinner party scenario with efforts taken to capture data that is representative of natural conversational speech. This paper provides a baseline description of the CHiME-6 challenge for both segmented multispeaker speech recognition (Track 1) and unsegmented multispeaker speech recognition (Track 2). Of note, Track 2 is the first challenge activity in the community to tackle an unsegmented multispeaker speech recognition scenario with a complete set of reproducible open source baselines providing speech enhancement, speaker diarization, and speech recognition modules.

preprint2020arXiv

Dark trions govern the temperature-dependent optical absorption and emission of doped atomically thin semiconductors

We perform absorption and photoluminescence spectroscopy of trions in hBN-encapsulated WSe$_2$, WS$_2$, MoSe$_2$, and MoS$_2$ monolayers, depending on temperature. The different trends for W- and Mo-based materials are excellently reproduced considering a Fermi-Dirac distribution of bright and dark trions. We find a dark trion, $\rm{X_D^-}$ 19 meV $\textit{below}$ the lowest bright trion, $\rm{X}_1^-$ in WSe$_2$ and WS$_2$. In MoSe$_2$, $\rm{X_D^-}$ lies 6 meV $\textit{above}$ $\rm{X}_1^-$, while $\rm{X_D^-}$ and $\rm{X}_1^-$ almost coincide in MoS$_2$. Our results agree with GW-BSE $\textit{ab-initio}$ calculations and quantitatively explain the optical response of doped monolayers with temperature.

preprint2020arXiv

Efficient MDI Adaptation for n-gram Language Models

This paper presents an efficient algorithm for n-gram language model adaptation under the minimum discrimination information (MDI) principle, where an out-of-domain language model is adapted to satisfy the constraints of marginal probabilities of the in-domain data. The challenge for MDI language model adaptation is its computational complexity. By taking advantage of the backoff structure of n-gram model and the idea of hierarchical training method, originally proposed for maximum entropy (ME) language models, we show that MDI adaptation can be computed in linear-time complexity to the inputs in each iteration. The complexity remains the same as ME models, although MDI is more general than ME. This makes MDI adaptation practical for large corpus and vocabulary. Experimental results confirm the scalability of our algorithm on very large datasets, while MDI adaptation gets slightly worse perplexity but better word error rate results compared to simple linear interpolation.

preprint2020arXiv

The JHU Multi-Microphone Multi-Speaker ASR System for the CHiME-6 Challenge

This paper summarizes the JHU team's efforts in tracks 1 and 2 of the CHiME-6 challenge for distant multi-microphone conversational speech diarization and recognition in everyday home environments. We explore multi-array processing techniques at each stage of the pipeline, such as multi-array guided source separation (GSS) for enhancement and acoustic model training data, posterior fusion for speech activity detection, PLDA score fusion for diarization, and lattice combination for automatic speech recognition (ASR). We also report results with different acoustic model architectures, and integrate other techniques such as online multi-channel weighted prediction error (WPE) dereverberation and variational Bayes-hidden Markov model (VB-HMM) based overlap assignment to deal with reverberation and overlapping speakers, respectively. As a result of these efforts, our ASR systems achieve a word error rate of 40.5% and 67.5% on tracks 1 and 2, respectively, on the evaluation set. This is an improvement of 10.8% and 10.4% absolute, over the challenge baselines for the respective tracks.

preprint2015arXiv

Exciton band structure in layered MoSe2: from a monolayer to the bulk limit

We present the micro-photoluminescence ($μ$PL) and micro-reflectance contrast spectroscopy studies on thin films of MoSe2 with layer thicknesses ranging from a monolayer (1L) up to 5L. The thickness dependent evolution of the ground and excited state excitonic transitions taking place at various points of the Brillouin zone is determined. Temperature activated energy shifts and linewidth broadenings of the excitonic resonances in 1L, 2L and 3L flakes are accounted for by using standard formalisms previously developed for semiconductors. A peculiar shape of the optical response of the ground state (A) exciton in monolayer MoSe2 is tentatively attributed to the appearance of Fano-type resonance. Rather trivial and clearly decaying PL spectra of monolayer MoSe2 with temperature confirm that the ground state exciton in this material is optically bright in contrast to a dark exciton ground state in monolayer WSe2.

preprint2015arXiv

Excitonic resonances in thin films of WSe2: From monolayer to bulk material

We present optical spectroscopy (photoluminescence and reflectance) studies of thin layers of the transition metal dichalcogenide WSe2, with thickness ranging from mono- to tetra-layer and in the bulk limit. The investigated spectra show the evolution of excitonic resonances as a function of layer thickness, due to changes in the band structure and, importantly, due to modifications of the strength of Coulomb interaction as well. The observed temperature-activated energy shift and broadening of the fundamental direct exciton are well accounted for by standard formalisms used for conventional semiconductors. A large increase of the photoluminescence yield with temperature is observed in WSe2 monolayer, indicating the existence of competing radiative channels. The observation of absorption-type resonances due to both neutral and charged excitons in WSe2 monolayer is reported and the effect of the transfer of oscillator strength from charged to neutral exciton upon increase of temperature is demonstrated.

preprint2015arXiv

Indirect-to-direct band-gap crossover in few-layer MoTe$_2$

We study the evolution of the band-gap structure in few-layer MoTe$_2$ crystals, by means of low-temperature micro-reflectance (MR) and temperature-dependent photoluminescence (PL) measurements. The analysis of the measurements indicate that, in complete analogy with other semiconducting transition metal dichalchogenides (TMDs), the dominant PL emission peaks originate from direct transitions associated to recombination of excitons and trions. When we follow the evolution of the PL intensity as a function of layer thickness, however, we observe that MoTe$_2$ behaves differently from other semiconducting TMDs investigated earlier. Specifically, the exciton PL yield (integrated PL intensity) is identical for mono and bilayer and it starts decreasing for trilayers. A quantitative analysis of this behavior and of all our experimental observations is fully consistent with mono and bilayer MoTe$_2$ being direct band-gap semiconductors, with tetralayer MoTe$_2$ being an indirect gap semiconductor, and with trilayers having nearly identical direct and indirect gaps.This conclusion is different from the one reached for other recently investigated semiconducting transition metal dichalcogenides, for which only monolayers are found to be direct band-gap semiconductors, with thicker layers having indirect band gaps that are significantly smaller, by hundreds of meV, than the direct gap. We discuss the relevance of our findings for experiments of fundamental interest and possible future device applications.

preprint2014arXiv

Pseudo 5D HN(C)N Experiment to Facilitate the Assignment of Backbone Resonances in Proteins Exhibiting High Backbone Shift Degeneracy

Assignment of protein backbone resonances is most routinely carried out using triple resonance three dimensional NMR experiments involving amide 1H and 15N resonances. However for intrinsically unstructured proteins, alpha-helical proteins or proteins containing several disordered fragments, the assignment becomes problematic because of high degree of backbone shift degeneracy. In this backdrop, a novel reduced dimensionality (RD) experiment -(5,3)D-hNCO-CANH- is presented to facilitate (and/or to validate) the sequential backbone resonance assignment in such proteins. The proposed 3D NMR experiment makes use of the modulated amide 15N chemical shifts (resulting from the joint sampling along both its indirect dimensions) to resolve the ambiguity involved in connecting the neighboring amide resonances (i.e. HiNi and Hi-1Ni-1) for overlapping amide NH peaks. The experiment -encoding 5D spectral information- leads to a conventional 3D spectrum with significantly reduced spectral crowding and complexity. The improvisation is based on the fact that the linear combinations of intra-residue and inter-residue backbone chemical shifts along both the co-evolved indirect dimensions span a wider spectral range and produce better peak dispersion than the individual shifts themselves. Taken together, the experiment -in combination with routine triple resonance 3D NMR experiments involving backbone amide (1H and 15N) and carbon (13C-alpha and 13C') chemical shifts- will serve as a powerful complementary tool to achieve the nearly complete assignment of protein backbone resonances in a time efficient manner. The performance of the experiment and application of the method have been demonstrated here using a 15.4 kDa size folded protein and a 12 kDa size unfolded protein.