Source author record

Ahmed Ali

Ahmed Ali appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

33works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

33 published item(s)

preprint2023arXiv

Textual Data Augmentation for Arabic-English Code-Switching Speech Recognition

The pervasiveness of intra-utterance code-switching (CS) in spoken content requires that speech recognition (ASR) systems handle mixed language. Designing a CS-ASR system has many challenges, mainly due to data scarcity, grammatical structure complexity, and domain mismatch. The most common method for addressing CS is to train an ASR system with the available transcribed CS speech, along with monolingual data. In this work, we propose a zero-shot learning methodology for CS-ASR by augmenting the monolingual data with artificially generating CS text. We based our approach on random lexical replacements and Equivalence Constraint (EC) while exploiting aligned translation pairs to generate random and grammatically valid CS content. Our empirical results show a 65.5% relative reduction in language model perplexity, and 7.7% in ASR WER on two ecologically valid CS test sets. The human evaluation of the generated text using EC suggests that more than 80% is of adequate quality.

preprint2022arXiv

Balanced End-to-End Monolingual pre-training for Low-Resourced Indic Languages Code-Switching Speech Recognition

The success in designing Code-Switching (CS) ASR often depends on the availability of the transcribed CS resources. Such dependency harms the development of ASR in low-resourced languages such as Bengali and Hindi. In this paper, we exploit the transfer learning approach to design End-to-End (E2E) CS ASR systems for the two low-resourced language pairs using different monolingual speech data and a small set of noisy CS data. We trained the CS-ASR, following two steps: (i) building a robust bilingual ASR system using a convolution-augmented transformer (Conformer) based acoustic model and n-gram language model, and (ii) fine-tuned the entire E2E ASR with limited noisy CS data. We tested our method on MUCS 2021 challenge and achieved 3rd place in the CS track. We then tested the proposed method using noisy CS data released for Hindi-English and Bengali-English pairs in Multilingual and Code-Switching ASR Challenges for Low Resource Indian Languages (MUCS 2021) and achieved 3rd place in the CS track. Unlike, the leading two systems that benefited from crawling YouTube and learning transliteration pairs, our proposed transfer learning approach focused on using only the limited CS data with no data-cleaning or data re-segmentation. Our approach achieved 14.1% relative gain in word error rate (WER) in Hindi-English and 27.1% in Bengali-English. We provide detailed guidelines on the steps to finetune the self-attention based model for limited data for ASR. Moreover, we release the code and recipe used in this paper.

preprint2022arXiv

ClassSPLOM -- A Scatterplot Matrix to Visualize Separation of Multiclass Multidimensional Data

In multiclass classification of multidimensional data, the user wants to build a model of the classes to predict the label of unseen data. The model is trained on the data and tested on unseen data with known labels to evaluate its quality. The results are visualized as a confusion matrix which shows how many data labels have been predicted correctly or confused with other classes. The multidimensional nature of the data prevents the direct visualization of the classes so we design ClassSPLOM to give more perceptual insights about the classification results. It uses the Scatterplot Matrix (SPLOM) metaphor to visualize a Linear Discriminant Analysis projection of the data for each pair of classes and a set of Receiving Operating Curves to evaluate their trustworthiness. We illustrate ClassSPLOM on a use case in Arabic dialects identification.

preprint2022arXiv

Creating Speech-to-Speech Corpus from Dubbed Series

Dubbed series are gaining a lot of popularity in recent years with strong support from major media service providers. Such popularity is fueled by studies that showed that dubbed versions of TV shows are more popular than their subtitled equivalents. We propose an unsupervised approach to construct speech-to-speech corpus, aligned on short segment levels, to produce a parallel speech corpus in the source- and target- languages. Our methodology exploits video frames, speech recognition, machine translation, and noisy frames removal algorithms to match segments in both languages. To verify the performance of the proposed method, we apply it on long and short dubbed clips. Out of 36 hours TR-AR dubbed series, our pipeline was able to generate 17 hours of paired segments, which is about 47% of the corpus. We applied our method on another language pair, EN-AR, to ensure it is robust enough and not tuned for a specific language or a specific corpus. Regardless of the language pairs, the accuracy of the paired segments was around 70% when evaluated using human subjective evaluation. The corpus will be freely available for the research community.

preprint2020arXiv

Analyzing Phonetic and Graphemic Representations in End-to-End Automatic Speech Recognition

End-to-end neural network systems for automatic speech recognition (ASR) are trained from acoustic features to text transcriptions. In contrast to modular ASR systems, which contain separately-trained components for acoustic modeling, pronunciation lexicon, and language modeling, the end-to-end paradigm is both conceptually simpler and has the potential benefit of training the entire system on the end task. However, such neural network models are more opaque: it is not clear how to interpret the role of different parts of the network and what information it learns during training. In this paper, we analyze the learned internal representations in an end-to-end ASR model. We evaluate the representation quality in terms of several classification tasks, comparing phonemes and graphemes, as well as different articulatory features. We study two languages (English and Arabic) and three datasets, finding remarkable consistency in how different properties are represented in different layers of the deep neural network.

preprint2020arXiv

Interpretation of $Y_b (10750)$ as a tetraquark and its production mechanism

Recently, the Belle Collaboration has updated the analysis of the cross sections for the processes $e^+ e^- \to Υ(nS)\, π^+ π^-$ ($n = 1,\, 2,\, 3$) in the $e^+ e^-$ center-of-mass energy range from 10.52 to 11.02 GeV. A new structure, called here $Y_b (10750)$, with the mass $M (Y_b) = (10752.7 \pm 5.9^{+0.7}_{-1.1})$ MeV and the Breit-Wigner width $Γ(Y_b) = (35.5^{+17.6 +3.9}_{-11.3 -3.3})$ MeV was observed \cite{Abdesselam:2019gth}. We interpret $Y_b (10750)$ as a compact $J^{PC} = 1^{--}$ state with a dominant tetraquark component. The mass eigenstate $Y_b (10750)$ is treated as a linear combination of the diquark-antidiquark and $b \bar b$ components due to the mixing via gluonic exchanges shown recently to arise in the limit of large number of quark colors. The mixing angle between $Y_b$ and $Υ(5S)$ can be estimated from the electronic width, recently determined to be $Γ_{ee} (Y_b) = (13.7 \pm 1.8)$ eV. The mixing provides a plausible mechanism for $Y_b (10750)$ production in high energy collisions from its $b \bar b$ component and we work out the Drell-Yan and prompt production cross sections for $p p \to Y_b (10750) \to Υ(nS)\, π^+ π^-$ at the LHC. The resonant part of the dipion invariant mass spectrum in $Y_b (10750) \to Υ(1S)\, π^+ π^-$ and the corresponding angular distribution of $π^+$-meson in the dipion rest frame are presented as an example.

preprint2020arXiv

What Was Written vs. Who Read It: News Media Profiling Using Text Analysis and Social Media Context

Predicting the political bias and the factuality of reporting of entire news outlets are critical elements of media profiling, which is an understudied but an increasingly important research direction. The present level of proliferation of fake, biased, and propagandistic content online, has made it impossible to fact-check every single suspicious claim, either manually or automatically. Alternatively, we can profile entire news outlets and look for those that are likely to publish fake or biased content. This approach makes it possible to detect likely "fake news" the moment they are published, by simply checking the reliability of their source. From a practical perspective, political bias and factuality of reporting have a linguistic aspect but also a social context. Here, we study the impact of both, namely (i) what was written (i.e., what was published by the target medium, and how it describes itself on Twitter) vs. (ii) who read it (i.e., analyzing the readers of the target medium on Facebook, Twitter, and YouTube). We further study (iii) what was written about the target medium on Wikipedia. The evaluation results show that what was written matters most, and that putting all information sources together yields huge improvements over the current state-of-the-art.

preprint2020arXiv

Word Error Rate Estimation Without ASR Output: e-WER2

Measuring the performance of automatic speech recognition (ASR) systems requires manually transcribed data in order to compute the word error rate (WER), which is often time-consuming and expensive. In this paper, we continue our effort in estimating WER using acoustic, lexical and phonotactic features. Our novel approach to estimate the WER uses a multistream end-to-end architecture. We report results for systems using internal speech decoder features (glass-box), systems without speech decoder features (black-box), and for systems without having access to the ASR system (no-box). The no-box system learns joint acoustic-lexical representation from phoneme recognition results along with MFCC acoustic features to estimate WER. Considering WER per sentence, our no-box system achieves 0.56 Pearson correlation with the reference evaluation and 0.24 root mean square error (RMSE) across 1,400 sentences. The estimated overall WER by e-WER2 is 30.9% for a three hours test set, while the WER computed using the reference transcriptions was 28.5%.

preprint2019arXiv

Mass spectrum of the hidden-charm pentaquarks in the compact diquark model

The LHCb collaboration have recently updated their analysis of the resonant $J/ψp$ mass spectrum in the decay $Λ_b^0 \to J/ψp K^-$, making use of their combined Run 1 and Run 2 data. In the updated analysis, three narrow states, $P_c (4312)^+$, $P_c (4440)^+$,and $P_c (4457)^+$, are observed. The spin-parity assignments of these states are not yet known. We interpret these narrow resonances as compact hidden-charm diquark-diquark-antiquark pentaquarks. Using an effective Hamiltonian, based on constituent quarks and diquarks, we calculate the pentaquark mass spectrum for the complete $SU (3)_F$ lowest $S$- and $P$-wave multiplets, taking into account dominant spin-spin, spin-orbit, orbital and tensor interactions. The resulting spectrum is very rich and we work out the quark flavor compositions, masses, and $J^P$ quantum numbers of the pentaquarks. However, heavy quark symmetry restricts the observable states in $Λ_b$-baryon, as well as in the decays of the other weakly-decaying $b$-baryons, $Ξ_b$ and $Ω_b$. In addition, some of the pentaquark states are estimated to lie below the $J/ψp$ threshold in $Λ_b$-decays (and corresponding thresholds in $Ξ_b$- and $Ω_b$-decays). They decay via $c \bar c$ annihilation into light hadrons or a dilepton pair, and are expected to be narrower than the $P_c$-states observed. We anticipate their discovery, as well as of the other pentaquark states present in the spectrum at the LHC, and in the long-term future at a Tera-$Z$ factory.

preprint2016arXiv

Automatic Dialect Detection in Arabic Broadcast Speech

We investigate different approaches for dialect identification in Arabic broadcast speech, using phonetic, lexical features obtained from a speech recognition system, and acoustic features using the i-vector framework. We studied both generative and discriminate classifiers, and we combined these features using a multi-class Support Vector Machine (SVM). We validated our results on an Arabic/English language identification task, with an accuracy of 100%. We used these features in a binary classifier to discriminate between Modern Standard Arabic (MSA) and Dialectal Arabic, with an accuracy of 100%. We further report results using the proposed method to discriminate between the five most widely used dialects of Arabic: namely Egyptian, Gulf, Levantine, North African, and MSA, with an accuracy of 52%. We discuss dialect identification errors in the context of dialect code-switching between Dialectal Arabic and MSA, and compare the error pattern between manually labeled data, and the output from our classifier. We also release the train and test data as standard corpus for dialect identification.

preprint2016arXiv

Heavy quark symmetry and weak decays of the $b$-baryons in pentaquarks with a $c\bar{c}$ component

The discovery of the baryonic states $P_c^+(4380)$ and $P_c^+(4450)$ by the LHCb collaboration has evoked a lot of theoretical interest. These states have the minimal quark content $c \bar{c} u u d$. Interpreted as hidden charm diquark-diquark-antiquark baryons, the assigned spin and angular momentum quantum numbers are $P_c^+(4380)= \{\bar{c} [cu]_{s=1} [ud]_{s=1}; L_{\mathcal{P}}=0, J^{\rm P}=\frac{3}{2}^- \}$ and $P_c^+(4450)= \{\bar{c} [cu]_{s=1} [ud]_{s=0}; L_{\mathcal{P}}=1, J^{\rm P}=\frac{5}{2}^+ \}$, where $s=0,1$ are the spins of the diquarks and $L_{\mathcal{P}}=0,1$ are the orbital angular momentum quantum numbers of the pentaquarks. We point out that heavy quark symmetry allows only the higher mass pentaquark state $P_c^+(4450)$ having $[ud]_{s=0}$ to be produced in $Λ_b^0$ decays, whereas the lower mass state $P_c^+(4380)$ having $[ud]_{s=1}$ is disfavored. Pentaquark spectrum is rich enough to accommodate a $J^P=\frac{3}{2}^-$ state, which has the correct light diquark spin $\{\bar{c} [cu]_{s=1} [ud]_{s=0}; L_{\mathcal{P}}=0, J^{\rm P}=\frac{3}{2}^- \}$ to be produced in $Λ_b^0$ decays. Assuming that the orbital mass difference between the charmed pentaquarks is similar to the corresponding mass difference in the charmed baryons, we estimate the mass of the lower pentaquark $J^P=3/2^-$ state to be about 4110 MeV and suggest to reanalyze the LHCb data to search for this third state. We present the spectroscopy of the $S$- and $P$-wave pentaquark states having a $c\bar{c}$ pair and three light quarks using an effective Hamiltonian approach. Some of these pentaquarks can be produced in weak decays of the $b$-baryons. Combining heavy quark symmetry and the $SU(3)_F$ symmetry results in strikingly simple relations among the decay amplitudes which are presented here.

preprint2016arXiv

Multi-view Dimensionality Reduction for Dialect Identification of Arabic Broadcast Speech

In this work, we present a new Vector Space Model (VSM) of speech utterances for the task of spoken dialect identification. Generally, DID systems are built using two sets of features that are extracted from speech utterances; acoustic and phonetic. The acoustic and phonetic features are used to form vector representations of speech utterances in an attempt to encode information about the spoken dialects. The Phonotactic and Acoustic VSMs, thus formed, are used for the task of DID. The aim of this paper is to construct a single VSM that encodes information about spoken dialects from both the Phonotactic and Acoustic VSMs. Given the two views of the data, we make use of a well known multi-view dimensionality reduction technique known as Canonical Correlation Analysis (CCA), to form a single vector representation for each speech utterance that encodes dialect specific discriminative information from both the phonetic and acoustic representations. We refer to this approach as feature space combination approach and show that our CCA based feature vector representation performs better on the Arabic DID task than the phonetic and acoustic feature representations used alone. We also present the feature space combination approach as a viable alternative to the model based combination approach, where two DID systems are built using the two VSMs (Phonotactic and Acoustic) and the final prediction score is the output score combination from the two systems.

preprint2016arXiv

Multiquark Hadrons - A New Facet of QCD

I review some selected aspects of the phenomenology of multiquark states discovered in high energy experiments. They have four valence quarks (called tetraquarks) and two of them are found to have five valence quarks (called pentaquarks), extending the conventional hadron spectrum which consists of quark-antiquark $(q\bar{q})$ mesons and $qqq$ baryons. Multiquark states represent a new facet of QCD and their dynamics is both challenging and currently poorly understood. I discuss various approaches put forward to accommodate them, with emphasis on the diquark model.

preprint2016arXiv

Rare B-Meson Decays at the Crossroads

Experimental era of rare $B$-decays started with the measurement of $B \to K^* γ$ by CLEO in 1993, followed two years later by the measurement of the inclusive decay $B \to X_s γ$, which serves as the standard candle in this field. The frontier has moved in the meanwhile to the experiments at the LHC, in particular, LHCb, with the decay $B^0 \to μ^+ μ^-$ at about 1 part in $10^{10}$ being the smallest branching fraction measured so far. Experimental precision achieved in this area has put the standard model to unprecedented stringent tests and more are in the offing in the near future. I review some key measurements in radiative, semileptonic and leptonic rare $B$-decays, contrast them with their estimates in the SM, and focus on several mismatches reported recently. They are too numerous to be ignored, yet , standing alone, none of them is significant enough to warrant the breakdown of the SM. Rare $B$-decays find themselves at the crossroads, possibly pointing to new horizons, but quite likely requiring an improved theoretical description in the context of the SM. An independent precision experiment such as Belle II may help greatly in clearing some of the current experimental issues.

preprint2015arXiv

Improved Estimates of The $B_{(s)}\to V V$ Decays in Perturbative QCD Approach

We reexamine the branching ratios, $CP$-asymmetries, and other observables in a large number of $B_q\to VV(q=u,d,s)$ decays in the perturbative QCD (PQCD) approach, where $V$ denotes a light vector meson $(ρ, K^*, ω, ϕ)$. The essential difference between this work and the earlier similar works is of parametric origin and in the estimates of the power corrections related to the ratio $r_i^2=m_{V_i}^2/m_B^2(i=2,3)$ ($m_V$ and $m_B$ denote the masses of the vector and $B$ meson, respectively). In particular, we use up-to-date distribution amplitudes for the final state mesons and keep the terms proportional to the ratio $r_i^2$ in our calculations. Our updated calculations are in agreement with the experimental data, except for a limited number of decays which we discuss. We emphasize that the penguin annihilation and the hard-scattering emission contributions are essential to understand the polarization anomaly, such as in the $B\to ϕK^*$ and $B_s \to ϕϕ$ decay modes. We also compare our results with those obtained in the QCD factorization (QCDF) approach and comment on the similarities and differences, which can be used to discriminate between these approaches in future experiments.

preprint2014arXiv

Precise Calculation of the Dilepton Invariant-Mass Spectrum and the Decay Rate in $B^\pm \to π^\pm μ^+ μ^-$ in the SM

We present a precise calculation of the dilepton invariant-mass spectrum and the decay rate for $B^\pm \to π^\pm \ell^+ \ell^-$ ($\ell^\pm = e^\pm, μ^\pm $) in the Standard Model (SM) based on the effective Hamiltonian approach for the $b \to d \ell^+ \ell^-$ transitions. With the Wilson coefficients already known in the next-to-next-to-leading logarithmic (NNLL) accuracy, the remaining theoretical uncertainty in the short-distance contribution resides in the form factors $f_+ (q^2)$, $f_0 (q^2)$ and $f_T (q^2)$. Of these, $f_+ (q^2)$ is well measured in the charged-current semileptonic decays $B \to π\ell ν_\ell$ and we use the $B$-factory data to parametrize it. The corresponding form factors for the $B \to K$ transitions have been calculated in the Lattice-QCD approach for large-$q^2$ and extrapolated to the entire $q^2$-region using the so-called $z$-expansion. Using an $SU(3)_F$-breaking Ansatz, we calculate the $B \to π$ tensor form factor, which is consistent with the recently reported lattice $B \to π$ analysis obtained at large~$q^2$. The prediction for the total branching fraction ${\cal B} (B^\pm \to π^\pm μ^+ μ^-) = (1.88 ^{+0.32}_{-0.21}) \times 10^{-8}$ is in good agreement with the experimental value obtained by the LHCb Collaboration. In the low $q^2$-region, heavy-quark symmetry (HQS) relates the three form factors with each other. Accounting for the leading-order symmetry-breaking effects, and using data from the charged-current process $B \to π\ell ν_\ell$ to determine $f_+ (q^2)$, we calculate the dilepton invariant-mass distribution in the low $q^2$-region in the $B^\pm \to π^\pm \ell^+ \ell^-$ decay. This provides a model-independent and precise calculation of the partial branching ratio for this decay.

preprint2013arXiv

Hadroproduction of $Υ(nS)$ above $B\bar B$ Thresholds and Implications for $Y_b(10890)$

Based on the non-relativistic QCD factorization scheme, we study the hadroproduction of the bottomonium states $Υ(5S)$ and $Υ(6S)$. We argue to search for them in the final states $Υ(1S,2S,3S)π^+π^-$, which are found to have anomalously large production rates at $Υ(5S)$. The enhanced rates for the dipionic transitions in the $Υ(5S)$-energy region could, besides $Υ(5S)$, be ascribed to $Y_b(10890)$, a state reported by the Belle collaboration, which may be interpreted as a tetraquark. The LHC/Tevatron measurements are capable of making a case in favor of or against the existence of $Y_b(10890)$, as demonstrated here. Dalitz analysis of the $Υ(1S,2S,3S)π^+π^-$ states from the $Υ(5S)/Y_b(10890)$ decays also impacts directly on the interpretation of the charged bottomonium-like states, $Z_b(10600)$ and $Z_b(10650)$, discovered by Belle in these puzzling decays.

preprint2012arXiv

Tetraquark Interpretation of the Charged Bottomonium-like states Z_b^+-(10610) and Z_b^+-(10650) and Implications

We present a tetraquark interpretation of the charged bottomonium-like states Z+-_b(10610) and Z+-_b(10650), observed by the Belle collaboration in the pi+- Upsilon(nS) (n=1,2,3) and pi+- h_b(mP) (m=1,2) invariant mass spectra from the data taken near the peak of the Upsilon(5S). In this framework, the underlying processes involve the production and decays of a vector tetraquark Y_b(10890), e+e- --> Y_b(10890) --> [Z+-_b(10610)pi-+, Z+-_b(10650)pi-+] followed by the decays [Z+-_b (10610), Z+-_b (10650)] --> pi+- Upsilon(nS), pi+- h_b(mP). Combining the contributions from the meson loops and an effective Hamiltonian, we are able to reproduce the observed masses of the Z+-_b(10610) and Z+-_b(10650). The analysis presented here is in agreement with the Belle data and provides crucial tests of the tetraquark hypothesis. We also calculate the corresponding meson loop effects in the charm sector and find them dynamically suppressed. The charged charmonium-like states Z+-_c(3752) and Z+-_c(3882) can be searched for in the decays of the J^{PC}=1^{--} tetraquark state Y(4260) via Y(4260) --> Z+-_c(3752)pi-+ and Y(4260) --> Z+-_c(3882)pi-+, with the subsequent decays (Z+-_c(3752),Z+-_c(3882)) --> (J/psi, h_c)pi+-.

preprint2012arXiv

Transverse Energy-Energy Correlations in Next-to-Leading Order in $α_s$ at the LHC

We compute the transverse energy-energy correlation (EEC) and its asymmetry (AEEC) in next-to-leading order (NLO) in $α_s$ in proton-proton collisions at the LHC with the center-of-mass energy $E_{\rm c.m.}=7$ TeV. We show that the transverse EEC and the AEEC distributions are insensitive to the QCD factorization- and the renormalization-scales, structure functions of the proton, and for a judicious choice of the jet-size, also the underlying minimum bias events. Hence they can be used to precisely test QCD in hadron colliders and determine the strong coupling $α_s$. We illustrate these features by defining the hadron jets using the anti-$k_T$ jet algorithm and an event selection procedure employed in the analysis of jets at the LHC and show the $α_s(M_Z)$-dependence of the transverse EEC and the AEEC in the anticipated range $0.11 \leq α_s(M_Z) \leq 0.13$.

preprint2011arXiv

Improved sensitivity to charged Higgs searches in Top quark decays $t \to bH^+ \to b (τ^+ν_τ)$ at the LHC using $τ$ polarisation and multivariate tecnniques

We present an analysis with improved sensitivity to the light charged Higgs ($m_{H^+} < m_t-m_b$) searches in the top quark decays $t \to b H^+ \to b (τ^+ν_τ) + ~{\rm c.c.}$ in the $t\bar{t}$ and single $t/\bar{t}$ production processes at the LHC. In the Minimal Supersymmetric Standard Model (MSSM), one anticipates the branching ratio ${\cal B} (H^+ \to τ^+ν_τ)\simeq 1$ over almost the entire allowed $\tan β$ range. Noting that the $τ^+$ arising from the decay $H^+ \to τ^+ν_τ$ are predominantly right-polarized, as opposed to the $τ^+$ from the dominant background $W^+ \to τ^+ν_τ$, which are left-polarized, a number of $H^+/W^+ \to τ^+ν_τ$ discriminators have been proposed and studied in the literature. We consider hadronic decays of the $τ^\pm$, concentrating on the dominant one-prong decay channel $τ^\pm \to ρ^\pm ν_τ$. The energy and $p_T$ of the charged prongs normalised to the corresponding quantities of the $ρ^\pm$ are convenient variables which serve as $τ^\pm$ polariser. We use the distributions in these variables and several other kinematic quantities to train a boosted decision tree (BDT). Using the BDT classifier, and a variant of it called BDTD, which makes use of decorrelated variables, we have calculated the BDT(D)-response functions to estimate the signal efficiency vs. the rejection of the background. We argue that this chain of analysis has a high sensitivity to light charged Higgs searches up to a mass of 150 GeV in the decays $t \to b H^+$ (and charge conjugate) at the LHC. For the case of single top production, we also study the transverse mass of the system determined using Lagrange multipliers.

preprint2011arXiv

Jets and QCD: A Historical Review of the Discovery of the Quark and Gluon Jets and its Impact on QCD

The observation of quark and gluon jets has played a crucial role in establishing Quantum Chromodynamics [QCD] as the theory of the strong interactions within the Standard Model of particle physics. The jets, narrowly collimated bundles of hadrons, reflect configurations of quarks and gluons at short distances. Thus, by analysing energy and angular distributions of the jets experimentally, the properties of the basic constituents of matter and the strong forces acting between them can be explored. In this review, which is primarily a description of the discovery of the quark and gluon jets and the impact of their observation on Quantum Chromodynamics, we elaborate, in particular, the role of the gluons as the carriers of the strong force. Focusing on these basic points, jets in $e^+ e^-$ collisions will be in the foreground of the discussion and we will concentrate on the theory that was contemporary with the relevant experiments at the electron-positron colliders. In addition we will delineate the role of jets as tools for exploring other particle aspects in $ep$ and $pp/p\bar{p}$ collisions - quark and gluon densities in protons, measurements of the QCD coupling, fundamental 2-2 quark/gluon scattering processes, but also the impact of jet decays of top quarks, and $W^\pm, Z$ bosons on the electroweak sector. The presentation to a large extent is formulated in a non-technical language with the intent to recall the significant steps historically and convey the significance of this field also to communities beyond high energy physics.

preprint2011arXiv

Production of the Exotic $1^{--}$ Hadrons $ϕ(2170)$, X(4260) and $Y_b(10890)$ at the LHC and Tevatron via the Drell-Yan Mechanism

We calculate the Drell-Yan production cross sections and differential distributions in the transverse momentum and rapidity of the $J^{PC}=1^{--}$ exotic hadrons $ϕ(2170)$, X(4260) and $Y_b(10890)$ at the hadron colliders LHC and the Tevatron. These hadrons are tetraquark (four-quark) candidates, with a hidden $s\bar{s}$, $c\bar{c}$ and $b\bar{b}$ quark pair, respectively. In deriving the distributions and cross sections, we include the order $α_s$ QCD corrections, resum the large logarithms in the small transverse momentum region in the impact-parameter formalism, and use the state of the art parton distribution functions. Taking into account the data on the production and decays of these vector hadrons from the $e^+e^-$ experiments, we present the production rates for the processes $pp(\bar{p}) \to ϕ(2170)(\to ϕ(1020) π^+π^- \to K^+K^- π^+π^-)+...$, $pp(\bar{p}) \to X(4260)(\to J/ψπ^+π^- \to μ^+μ^-π^+π^-)+...$, and $pp(\bar{p}) \to Y_b(10890)(\to (Υ(1S), Υ(2S), Υ(3S)) π^+π^- \to μ^+μ^-π^+π^-)+...$. Their measurements at the hadron colliders will provide new experimental avenues to explore the underlying dynamics of these hadrons.

preprint2011arXiv

Tetraquark interpretation of $e^+ e^- \to Υπ^+π^-$ Belle data and $e^+ e^- \to b \bar{b}$ BaBar data

We summarize the main features of the spectroscopy, production and decays of the $J^{PC}=1^{--}$ tetraquarks in the $b\bar{b}$ sector, concentrating on the lowest state called $Y_b(10890)$. The tetraquark framework is used to analyze the BaBar data on the $e^+ e^- \to b\bar{b}$ cross section ($R_b$ energy scan) between $\sqrt{s}= 10.54$ and 11.20 GeV and the Belle data on the processes $e^+e^- \to Υ(1S) π^+π^-,Υ(2S) π^+π^-$ near the peak of the $Υ(5S)$ resonance. The BaBar $R_b$ energy scan is consistent with an additional state at a mass of 10.90 GeV and a width of about 28 MeV, in broad agreement with the state $Y_b(10890)$ GeV seen by Belle in the exclusive final states. We argue that the decay widths and the dipion invariant mass distributions measured by Belle are naturally explained by the tetraquark interpretation of $Y_b(10890)$.

preprint2011arXiv

Tetraquark-based analysis and predictions of the cross sections and distributions for the processes e^+ e^- --> Upsilon(1S) (pi^+ pi^-, K^+ K^-, eta pi^0) near Upsilon(5S)

We calculate the cross sections and final state distributions for the processes e^+ e^- --> Upsilon(1S) (pi^+ pi^-, K^+ K^-, eta pi^0) near the Upsilon(5S) resonance based on the tetraquark hypothesis. This framework is used to analyse the data on the Upsilon(1S) pi^+ pi^- and Upsilon(1S) K^+ K^- final states [K.F. Chen et al. (Belle Collaboration), Phys. Rev. Lett. 100, 112001 (2008); I. Adachi et al. (Belle Collaboration), arXiv:0808.2445], yielding good fits. Dimeson invariant mass spectra in these processes are shown to be dominated by the corresponding light scalar and tensor states. The resulting correlations among the cross sections are worked out. We also predict sigma(e^+ e^- --> Upsilon(1S) K^+ K^-)/sigma(e^+ e^- --> Upsilon(1S) K^0 Kbar^0) = 1/4. These features provide crucial tests of the tetraquark framework and can be searched for in the currently available and forthcoming data from the B factories.

preprint2011arXiv

Theory Overview on Spectroscopy

A theoretical overview of the exotic spectroscopy in the charm and beauty quark sector is presented. These states are unexpected harvest from the $e^+e^-$ and hadron colliders and a permanent abode for the majority of them has yet to be found. We argue that some of these states, in particular the $Y_b(10890)$ and the recently discovered states $Z_b(10610)$ and $Z_b(10650)$, discovered by the Belle collaboration are excellent candidates for tetraquark states $[bq][\bar{b}\bar{q}]$, with $q=u,d$ light quarks. Theoretical analyes of the Belle data carried out in the tetraquark context is reviewed.

preprint2010arXiv

A case for hidden $b\bar{b}$ tetraquarks based on $e^+e^- \to b\bar{b}$ cross section between $\sqrt{s}=10.54$ and 11.20 GeV

We study the spectroscopy and dominant decays of the bottomonium-like tetraquarks (bound diquarks-antidiquarks), focusing on the lowest lying P-wave $[bq][\bar{b}\bar{q}]$ states $Y_{[bq]}$ (with $q=u,d$), having $J^{PC}=1^{--}$. To search for them, we analyse the BABAR data \cite{:2008hx} obtained during an energy scan of the $e ^+ e^- \to b \bar{b}$ cross section in the range of $\sqrt{s}=10.54$ to 11.20 GeV. We find that these data are consistent with the presence of an additional $b \bar{b}$ state $Y_{[bq]}$ with a mass of 10.90 GeV and a width of about 30 MeV apart from the $Υ(5S)$ and $Υ(6S)$ resonances. A closeup of the energy region around the $Y_{[bq]}$-mass may resolve this state in terms of the two mass eigenstates, $Y_{[b,l]}$ and $Y_{[b,h]}$, with a mass difference, estimated as about 6 MeV. We tentatively identify the state $Y_{[bq]}(10900)$ from the $R_b$-scan with the state $Y_b(10890)$ observed by BELLE \cite{Abe:2007tk} in the process $e^+e^- \to Y_b(10890) \to Υ(1S, 2S) π^+ π^-$ due to their proximity in masses and decay widths.

preprint2010arXiv

Prospects of measuring the CKM matrix element $|V_{ts}|$ at the LHC

We study the prospects of measuring the CKM matrix element $\vert V_{ts}\vert$ at the LHC with the top quarks produced in the processes $p p \to t\bar{t}X$ and $p p \to t/\bar{t} X$, and the subsequent decays $t \to W^+s$ and $\bar{t} \to W^- \bar{s}$. We insist on tagging the $W^\pm$ leptonically, $W^\pm \to \ell^\pm ν_\ell$ ($\ell =e, μ, τ$), and analyse the anticipated jet profiles in the signal process $t \to W s$ and the dominant background from the decay $t \to W b$. To that end, we analyse the $V0$ ($K^0$ and $Λ$) distributions in the $s$- and $b$-quark jets concentrating on the energy and transverse momentum distributions of these particles. The $V0$s emanating from the $t \to W b$ branch have displaced decay vertexes from the interaction point due to the weak decays $b \to c \to s$ and the $b$-quark jets are rich in charged leptons. Hence, the absence of secondary vertexes and of the energetic charged leptons in the jet provide additional ($b$-jet vs. $s$-jet) discrimination in top quark decays. These distributions are used to train a boosted decision tree (BDT). Using the BDT classifier, and a variant of it called BDTD, which makes use of decorrelated variables, we calculate the BDT(D)-response functions corresponding to the signal ($t \to W s$) and background ($t \to W b$). Detailed simulations undertaken by us with the Monte Carlo generator PYTHIA are used to estimate the background rejection versus signal efficiency for three representative LHC energies $\sqrt{s}=7$ TeV, 10 TeV and 14 TeV. We argue that a benchmark with 10\% signal ($t \to W s $) efficiency and a background ($t \to W b$) rejection by a factor $10^3$ (required due to the anticipated value of the ratio $\vert V_{ts}\vert^2/\vert V_{tb} \vert^2 \simeq 1.6 \times 10^{-3}$) can be achieved at the LHC@14 TeV with an integrated luminosity of 10 fb$^{-1}$.

preprint2010arXiv

Tetraquark interpretation of the BELLE data on the anomalous $Υ(1S) π^+π^-$ and $Υ(2S) π^+π^-$ production near the $Υ(5S)$ resonance

We analyze the Belle data [K. F. Chen {\it et al.} (Belle Collaboration), Phys.\ Rev.\ Lett.\ {\bf 100}, 112001 (2008); I. Adachi {\it et al.} (Belle Collaboration), arXiv:0808.2445] on the processes $e^+ e^- \to Υ(1S)\; π^+π^-, Υ(2S)\; π^+π^-$ near the peak of the $Υ(5S)$ resonance, which are found to be anomalously large in rates compared to similar dipion transitions between the lower $Υ$ resonances. Assuming these final states arise from the production and decays of the $J^{PC}=1^{--}$ state $Y_b(10890)$, which we interpret as a bound (diquark-antidiquark) tetraquark state $[bq][\bar{b}\bar{q}]$, a dynamical model for the decays $Y_b \to Υ(1S)\; π^+π^-, Υ(2S)\; π^+π^-$ is presented. Depending on the phase space, these decays receive significant contributions from the scalar $0^{++}$ states, $f_0(600)$ and $f_0(980)$, and from the $2^{++}$ $q\bar{q}$-meson $f_2(1270)$. Our model provides excellent fits for the decay distributions, supporting $Y_b$ as a tetraquark state.

preprint2009arXiv

Radiatively corrected lepton energy distributions in top quark decays $t \to bW^+ \to b(\ell^+ ν_\ell)$ and $t \to bH^+ \to b (τ^+ ν_τ)$ and single charged prong energy distributions from subsequent $τ^+$ decays

We calculate the QED and QCD radiative corrections to the charged lepton energy distributions in the dominant semileptonic decays of the top quark $t \to b W^+ \to b(\ell^+ ν_\ell)$ $(\ell=e, μ, τ)$ in the standard model(SM), and for the decay $t \to b H^+ \to b(τ^+ ν_τ)$ in an extension of the SM having a charged Higgs boson $H^\pm$ with $m_{H^\pm} < m_t -m_b$. The QCD corrections are calculated in the leading and next-to-leading logarithmic approximations, but the QED corrections are considered in the leading logarithmic approximation only. These corrections are numerically important for precisely testing the universality of the charged current weak interactions in $t$-quark decays. As the $τ^+$ leptons arising from the decays $W^+ \to τ^+ ν_τ$ and $H^+\to τ^+ ν_τ$ are predominantly left- and right-polarised, respectively, influencing the energy distributions of the decay products in the subsequent decays of the $τ^+$, we work out the effect of the radiative corrections on such distributions in the dominant (one-charged prong) decay channels $ τ^+ \to π^+ \barν_τ, ρ^+ \barν_τ, a_1^+ \barν_τ$ and $\ell^+ ν_\ell \barν_τ$. The inclusive $π^+$ energy spectra in the decay chains $t \to b(W^+,H^+) \to b (τ^+ ν_τ) \to b (π^+ \barν_τν_τ+X)$ are calculated, which can help in searching for the induced $H^\pm$ effects at the Tevatron and the LHC.

preprint2005arXiv

Helicity analysis of the decays B --> K^* ell^+ ell^- and B --> rho ell nu_ell in the large Energy Effective Theory

We calculate the independent helicity amplitudes in the decays B --> K^* ell^+ ell^- and B --> rho ell nu_ell in the so-called Large-Energy-Effective-Theory (LEET). Taking into account the dominant O(alpha_s) and SU(3) symmetry-breaking effects, we calculate various Dalitz distributions in these decays making use of the presently available data and decay form factors calculated in the QCD sum rule approach. Differential decay rates in the dilepton invariant mass and the Forward-Backward asymmetry in B --> K^* ell^+ ell^- are worked out. We also present the decay amplitudes in the transversity basis which has been used in the analysis of data on the resonant decay B --> K^* J/psi (--> ell^+ ell^-). Measurements of the ratios R_i(s) equivalent to d Gamma_{H_i}(s)(B --> K^* ell^+ ell^-)/ d Gamma_{H_i}(s)(B --> rho ell nu_ell), involving the helicity amplitudes H_i(s), i = 0,+1, -1, as precision tests of the standard model in semileptonic rare B-decays are emphasized. We argue that R_0(s) and R_{-}(s) can be used to determine the CKM ratio |V_{ub}|/|V_{ts}| and search for new physics, where the later is illustrated by supersymmetry.

preprint2002arXiv

Implications of B -> rho gamma measurements in the Standard Model and Supersymmetric Theories

We study the implications of the recently improved upper limits on the branching ratios for the decays B -> rho gamma, expressed as R(rho gamma/K^* gamma) = BR(B -> rho gamma)/BR(B -> K^* gamma) <0.047. We work out the constraints that the current bound on R(rho gamma/K^* gamma) implies on the parameters of the quark mixing matrix in the standard model (SM). Using the present profile of the unitarity triangle, we predict this ratio to be R(rho gamma/K^* gamma) = 0.023 +/- 0.012. We also work out the correlations involving R(rho gamma/K^* gamma), the isospin-violating ratio Delta (rho gamma), and the direct CP-violating asymmetry A_CP (rho gamma) in B -> rho gamma decays in the SM, in the minimal supersymmetric extension of the SM (MSSM), and in an extension of the MSSM involving an additional flavor-changing structure in b -> d transitions.

preprint1999arXiv

Precision Flavour Physics and Supersymmetry

We review the salient features of a comparative study of the profile of the CKM unitarity triangle, and the resulting CP-violating phases $α$, $β$ and $γ$ in B decays, in the standard model and in several variants of the minimal supersymmetric standard model (MSSM), reported recently by us. These theories are characterized by a single phase in the quark flavour mixing matrix and give rise to well-defined contributions in the flavour-changing-neutral-current transitions in K and B decays. We analyse the supersymmetric contributions to the mass differences in the Bd-Bd(bar) and Bs-Bs(bar) systems, $ΔM_d$ and $ΔM_s$, respectively, and to the CP-violating quantity $|epsilon|$ in K decays. Our analysis shows that the predicted ranges of $β$ in the standard model and in MSSM models are very similar. However, precise measurements at B-factories and hadron machines may be able to distinguish these theories in terms of the other two CP-violating phases $α$ and $γ$. (Contribution to the Festschrift for L.B. Okun, to appear in a special issue of Physics Reports, eds. V.L. Telegdi and K. Winter)