Source author record

Dipankar Das

Dipankar Das appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

32works
15topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

32 published item(s)

preprint2022arXiv

Can Unsupervised Knowledge Transfer from Social Discussions Help Argument Mining?

Identifying argument components from unstructured texts and predicting the relationships expressed among them are two primary steps of argument mining. The intrinsic complexity of these tasks demands powerful learning models. While pretrained Transformer-based Language Models (LM) have been shown to provide state-of-the-art results over different NLP tasks, the scarcity of manually annotated data and the highly domain-dependent nature of argumentation restrict the capabilities of such models. In this work, we propose a novel transfer learning strategy to overcome these challenges. We utilize argumentation-rich social discussions from the ChangeMyView subreddit as a source of unsupervised, argumentative discourse-aware knowledge by finetuning pretrained LMs on a selectively masked language modeling task. Furthermore, we introduce a novel prompt-based strategy for inter-component relation prediction that compliments our proposed finetuning method while leveraging on the discourse context. Exhaustive experiments show the generalization capability of our method on these two tasks over within-domain as well as out-of-domain datasets, outperforming several existing and employed strong baselines.

preprint2022arXiv

Diluting quark flavor hierarchies using dihedral symmetry

We present a $D_4$ flavored extension of the SM which provides an intuitive reasoning for the masses and mixing patterns in the quark sector. In our model, the Cabibbo mixing angle stems purely from the scalar sector dynamics. In fact, the orders of magnitude of the CKM matrix elements are readily obtained from the hierarchical nature of the vacuum expectation values. Moreover, we also show that the smallness of the off-Cabibbo elements in the CKM matrix is strongly connected to the heaviness of the third generation of quarks.

preprint2022arXiv

JU_NLP at HinglishEval: Quality Evaluation of the Low-Resource Code-Mixed Hinglish Text

In this paper we describe a system submitted to the INLG 2022 Generation Challenge (GenChal) on Quality Evaluation of the Low-Resource Synthetically Generated Code-Mixed Hinglish Text. We implement a Bi-LSTM-based neural network model to predict the Average rating score and Disagreement score of the synthetic Hinglish dataset. In our models, we used word embeddings for English and Hindi data, and one hot encodings for Hinglish data. We achieved a F1 score of 0.11, and mean squared error of 6.0 in the average rating score prediction task. In the task of Disagreement score prediction, we achieve a F1 score of 0.18, and mean squared error of 5.0.

preprint2022arXiv

Measuring frequency and period separations in red-giant stars using machine learning

Asteroseismology is used to infer the interior physics of stars. The \textit{Kepler} and TESS space missions have provided a vast data set of red-giant light curves, which may be used for asteroseismic analysis. These data sets are expected to significantly grow with future missions such as \textit{PLATO}, and efficient methods are therefore required to analyze these data rapidly. Here, we describe a machine learning algorithm that identifies red giants from the raw oscillation spectra and captures \textit{p} and \textit{mixed} mode parameters from the red-giant power spectra. We report algorithmic inferences for large frequency separation ($Δν$), frequency at maximum amplitude ($ν_{max}$), and period separation ($ΔΠ$) for an ensemble of stars. In addition, we have discovered $\sim$25 new probable red giants among 151,000 \textit{Kepler} long-cadence stellar-oscillation spectra analyzed by the method, among which four are binary candidates which appear to possess red-giant counterparts. To validate the results of this method, we selected $\sim$ 3,000 \textit{Kepler} stars, at various evolutionary stages ranging from subgiants to red clumps, and compare inferences of $Δν$, $ΔΠ$, and $ν_{max}$ with estimates obtained using other techniques. The power of the machine-learning algorithm lies in its speed: it is able to accurately extract seismic parameters from 1,000 spectra in $\sim$5 seconds on a modern computer (single core of the Intel Xeon Platinum 8280 CPU).

preprint2021arXiv

A three Higgs doublet model with symmetry-suppressed flavour changing neutral currents

We construct a three-Higgs doublet model with a flavour non-universal ${\rm U}(1)\times \mathbb{Z}_2$ symmetry. That symmetry induces suppressed flavour-changing interactions mediated by neutral scalars. New scalars with masses below the TeV scale can still successfully negotiate the constraints arising from flavour data. Such a model can thus encourage direct searches for extra Higgs bosons in the future collider experiments, and includes a non-trivial flavour structure.

preprint2020arXiv

Crossed two Higgs-doublet models: reduction of Yukawa parameters in the low-scale limit of left-right symmetry and other avatars

We present new variants of the Two Higgs-Doublet Model where all Yukawa couplings with physical Higgs bosons are controlled by the quark mixing matrices of both chiralities, as well as, in one case, the ratio between the two scalar doublets' vacuum expectation values. We obtain these by imposing approximate symmetries on the Lagrangian which, in one of the cases, clearly reveals the model to be the electroweak remnant of the Minimal Left-Right Symmetric Model. We also argue for the benefits of the bidoublet notation in the Two Higgs-Doublet Model context for uncovering new models.

preprint2020arXiv

Development of POS tagger for English-Bengali Code-Mixed data

Code-mixed texts are widespread nowadays due to the advent of social media. Since these texts combine two languages to formulate a sentence, it gives rise to various research problems related to Natural Language Processing. In this paper, we try to excavate one such problem, namely, Parts of Speech tagging of code-mixed texts. We have built a system that can POS tag English-Bengali code-mixed data where the Bengali words were written in Roman script. Our approach initially involves the collection and cleaning of English-Bengali code-mixed tweets. These tweets were used as a development dataset for building our system. The proposed system is a modular approach that starts by tagging individual tokens with their respective languages and then passes them to different POS taggers, designed for different languages (English and Bengali, in our case). Tags given by the two systems are later joined together and the final result is then mapped to a universal POS tag set. Our system was checked using 100 manually POS tagged code-mixed sentences and it returned an accuracy of 75.29%

preprint2020arXiv

Double Higgs boson production as an exclusive probe for a sequential fourth generation with wrong-sign Yukawa couplings

It has been shown that the data from the Large Hadron Collider (LHC) does not rule out a chiral sequential fourth generation of fermions that obtain their masses through an identical mechanism as the other three generations do. However, this is possible only if the scalar sector of the Standard Model is suitably enhanced, like embedding it in a type-II two-Higgs doublet model. In this article, we try to show that double Higgs production (DHP) can unveil the existence of such a hidden fourth generation in a very efficient way. While the DHP cross-section in the SM is quite small, it is significantly enhanced with a fourth generation. We perform a detailed analysis of the dependence of the DHP cross-section on the model parameters, and show that either a positive signal of DHP is seen in the early next run of the LHC, or the model is ruled out.

preprint2020arXiv

Investigating Deep Learning Approaches for Hate Speech Detection in Social Media

The phenomenal growth on the internet has helped in empowering individual's expressions, but the misuse of freedom of expression has also led to the increase of various cyber crimes and anti-social activities. Hate speech is one such issue that needs to be addressed very seriously as otherwise, this could pose threats to the integrity of the social fabrics. In this paper, we proposed deep learning approaches utilizing various embeddings for detecting various types of hate speeches in social media. Detecting hate speech from a large volume of text, especially tweets which contains limited contextual information also poses several practical challenges. Moreover, the varieties in user-generated data and the presence of various forms of hate speech makes it very challenging to identify the degree and intention of the message. Our experiments on three publicly available datasets of different domains shows a significant improvement in accuracy and F1-score.

preprint2020arXiv

JUNLP@SemEval-2020 Task 9:Sentiment Analysis of Hindi-English code mixed data using Grid Search Cross Validation

Code-mixing is a phenomenon which arises mainly in multilingual societies. Multilingual people, who are well versed in their native languages and also English speakers, tend to code-mix using English-based phonetic typing and the insertion of anglicisms in their main language. This linguistic phenomenon poses a great challenge to conventional NLP domains such as Sentiment Analysis, Machine Translation, and Text Summarization, to name a few. In this work, we focus on working out a plausible solution to the domain of Code-Mixed Sentiment Analysis. This work was done as participation in the SemEval-2020 Sentimix Task, where we focused on the sentiment analysis of English-Hindi code-mixed sentences. our username for the submission was "sainik.mahata" and team name was "JUNLP". We used feature extraction algorithms in conjunction with traditional machine learning algorithms such as SVR and Grid Search in an attempt to solve the task. Our approach garnered an f1-score of 66.2\% when tested using metrics prepared by the organizers of the task.

preprint2020arXiv

K-TanH: Efficient TanH For Deep Learning

We propose K-TanH, a novel, highly accurate, hardware efficient approximation of popular activation function TanH for Deep Learning. K-TanH consists of parameterized low-precision integer operations, such as, shift and add/subtract (no floating point operation needed) where parameters are stored in very small look-up tables that can fit in CPU registers. K-TanH can work on various numerical formats, such as, Float32 and BFloat16. High quality approximations to other activation functions, e.g., Sigmoid, Swish and GELU, can be derived from K-TanH. Our AVX512 implementation of K-TanH demonstrates $>5\times$ speed up over Intel SVML, and it is consistently superior in efficiency over other approximations that use floating point arithmetic. Finally, we achieve state-of-the-art Bleu score and convergence results for training language translation model GNMT on WMT16 data sets with approximate TanH obtained via K-TanH on BFloat16 inputs.

preprint2020arXiv

Preparation of Sentiment tagged Parallel Corpus and Testing its effect on Machine Translation

In the current work, we explore the enrichment in the machine translation output when the training parallel corpus is augmented with the introduction of sentiment analysis. The paper discusses the preparation of the same sentiment tagged English-Bengali parallel corpus. The preparation of raw parallel corpus, sentiment analysis of the sentences and the training of a Character Based Neural Machine Translation model using the same has been discussed extensively in this paper. The output of the translation model has been compared with a base-line translation model using automated metrics such as BLEU and TER as well as manually.

preprint2016arXiv

Authorship Verification - An Approach based on Random Forest

Authorship attribution, being an important problem in many areas in-cluding information retrieval, computational linguistics, law and journalism etc., has been identified as a subject of increasingly research interest in the re-cent years. In case of Author Identification task in PAN at CLEF 2015, the main focus was given on cross-genre and cross-topic author verification tasks. We have used several word-based and style-based features to identify the dif-ferences between the known and unknown problems of one given set and label the unknown ones accordingly using a Random Forest based classifier.

preprint2016arXiv

Distributed Deep Learning Using Synchronous Stochastic Gradient Descent

We design and implement a distributed multinode synchronous SGD algorithm, without altering hyper parameters, or compressing data, or altering algorithmic behavior. We perform a detailed analysis of scaling, and identify optimal design points for different networks. We demonstrate scaling of CNNs on 100s of nodes, and present what we believe to be record training throughputs. A 512 minibatch VGG-A CNN training run is scaled 90X on 128 nodes. Also 256 minibatch VGG-A and OverFeat-FAST networks are scaled 53X and 42X respectively on a 64 node cluster. We also demonstrate the generality of our approach via best-in-class 6.5X scaling for a 7-layer DNN on 16 nodes. Thereafter we attempt to democratize deep-learning by training on an Ethernet based AWS cluster and show ~14X scaling on 16 nodes.

preprint2016arXiv

Labeling of Query Words using Conditional Random Field

This paper describes our approach on Query Word Labeling as an attempt in the shared task on Mixed Script Information Retrieval at Forum for Information Retrieval Evaluation (FIRE) 2015. The query is written in Roman script and the words were in English or transliterated from Indian regional languages. A total of eight Indian languages were present in addition to English. We also identified the Named Entities and special symbols as part of our task. A CRF based machine learning framework was used for labeling the individual words with their corresponding language labels. We used a dictionary based approach for language identification. We also took into account the context of the word while identifying the language. Our system demonstrated an overall accuracy of 75.5% for token level language identification. The strict F-measure scores for the identification of token level language labels for Bengali, English and Hindi are 0.7486, 0.892 and 0.7972 respectively. The overall weighted F-measure of our system was 0.7498.

preprint2016arXiv

Updated scalar sector constraints in Higgs triplet model

We show that in the Higgs triplet model, after the Higgs discovery, the mixing angle in the CP-even sector can be strongly constrained from unitarity. We also discuss how large quantum effects in $h\toγγ$ may arise in a SM-like scenario and a certain part of the parameter space can be ruled out from the diphoton signal strength. Using $T$-parameter and diphoton signal strength measurements, we update the bounds on the nonstandard scalar masses.

preprint2015arXiv

$S_3$ symmetry and the quark mixing matrix

We impose an $S_3$ symmetry on the quark fields under which two of three quarks transform like a doublet and the remaining one as singlet, and use a scalar sector with the same structure of $SU(2)$ doublets. After gauge symmetry breaking, a $\mathbb{Z}_2$ subgroup of the $S_3$ remains unbroken. We show that this unbroken subgroup can explain the approximate block structure of the CKM matrix. By allowing soft breaking of the $S_3$ symmetry in the scalar sector, we show that one can generate the small elements, of quadratic or higher order in the Wolfenstein parametrization of the CKM matrix. We also predict the existence of exotic new scalars, with unconventional decay properties, which can be used to test our model experimentally.

preprint2015arXiv

GraphMat: High performance graph analytics made productive

Given the growing importance of large-scale graph analytics, there is a need to improve the performance of graph analysis frameworks without compromising on productivity. GraphMat is our solution to bridge this gap between a user-friendly graph analytics framework and native, hand-optimized code. GraphMat functions by taking vertex programs and mapping them to high performance sparse matrix operations in the backend. We get the productivity benefits of a vertex programming framework without sacrificing performance. GraphMat is in C++, and we have been able to write a diverse set of graph algorithms in this framework with the same effort compared to other vertex programming frameworks. GraphMat performs 1.2-7X faster than high performance frameworks such as GraphLab, CombBLAS and Galois. It achieves better multicore scalability (13-15X on 24 cores) than other frameworks and is 1.2X off native, hand-optimized code on a variety of different graph algorithms. Since GraphMat performance depends mainly on a few scalable and well-understood sparse matrix operations, GraphMatcan naturally benefit from the trend of increasing parallelism on future hardware.

preprint2015arXiv

Implications Of The Higgs Discovery On Physics Beyond The Standard Model

In this thesis, we investigate the implications of the LHC Higgs data on different BSM scenarios. Since the data seem to agree with the SM expectations, any nonstandard couplings will be strongly constrained. First we investigate, in a model independent way, the constraints on the nonstandard Higgs couplings with the fermions and the vector bosons in view of high energy unitarity and the measured value of the Higgs to diphoton signal strength. Then we concentrate on a particular BSM scenario, namely, the two Higgs-doublet models (2HDMs). Consistency of the Higgs data with the corresponding SM predictions strongly motivates us to work in the {\em alignment limit}. In this limit, including the informations of Higgs mass and its SM-like nature, we find many new constraints on the nonstandard masses and $\tanβ$. We also study the constraints on the charged scalar mass arising from the $h\to γγ$ signal strength measurements and observe that the charged scalar does not necessarily decouple from the diphoton decay width. We then move on to some particular variants of 2HDMs, known as BGL models, and study the flavor constraints on these models. Here we find that lighter than conventionally allowed nonstandard scalars can successfully negotiate the stringent bounds coming from flavor physics data and can leave unconventional decay signatures that can be used as distinctive features of these models. We also analyze the stability and unitarity constraints in a three Higgs-doublet model (3HDM) with $S_3$ symmetry and find that there must be many more nonstandard particles below 1~TeV. We also observe that the nondecoupling feature of the charged scalar in the context of $h\to γγ$ is not unique to the 2HDMs only, instead it is a general property of the multi doublet extensions of the SM with an exact discrete symmetry.

preprint2015arXiv

New limits on $\mathbf{\tanβ}$ for 2HDMs with $\mathbf{Z_2}$ symmetry

In two-Higgs-doublet models with exact $Z_2$ symmetry, putting $m_h \simeq 125$ GeV at the alignment limit, the following limits on the heavy scalar masses are obtained from the conditions of unitarity and stability of the scalar potential: $m_H,~m_A,~m_{H^+} < 1$ TeV and $1/8 <\tanβ<8$. The constraints from $b \to s γ$ and neutral meson mass differences, when superimposed on the unitarity constraints, put a tighter lower limit on $\tanβ$ depending on $m_{H^+}$. It has also been shown that larger values of $\tanβ$ can be allowed by introducing soft breaking term in the potential at the expense of a correlation between $m_H$ and the soft breaking parameter.

preprint2015arXiv

Scalar sector of Two-Higgs-Doublet models: A mini-review

A vast literature on the theory and phenomenology of Two-Higgs-Doublet models (2HDM) exists since long. However, the present situation demands a revisit of some 2HDM properties. Now that a 125 GeV scalar resonance has been discovered at the LHC, with its couplings to other particles showing increasing affinity to the Standard Model Higgs-like behavior, the 2HDM parameter space is more squeezed than ever. We briefly review the different parametrizations of the 2HDM potential and discuss the constraints on the parameter space arising from the unitarity and stability of the potential together with constraints from the oblique electroweak $T$-parameter. We also differentiate the consequences of imposing a global continuous U(1) symmetry on the potential from a discrete $Z_2$ symmetry.

preprint2015arXiv

Search for a 'stable alignment limit' in two Higgs-doublet models

We study the conditions required to make the 2HDM scalar potential stable up to the Planck scale. The lightest CP-even scalar is assumed to have been found at the LHC and the {\em alignment limit} is imposed in view of the LHC Higgs data. We find that ensuring stability up to scales $\gtrsim 10^{10}$~GeV necessitates the introduction of a soft breaking parameter in the theory. Even then, some interesting correlations between the nonstandard masses and the soft breaking parameter need to be satisfied. Consequently, a 2HDM becomes completely determined by only two nonstandard parameters, namely, $\tb$ and a mass parameter, $m_0$, with $\tb \gtrsim 3$. These observations make a 2HDM, in the {\em stable alignment limit}, more predictive than ever.

preprint2015arXiv

Two more hidden scalars around 125 GeV and $h \to μτ$

We show that the $2.4σ$ signal of the leptonic flavor violating (LFV) Higgs boson decay $h\toμτ$, as observed by the CMS collaboration recently, can be explained by a certain class of two-Higgs doublet models that allow controllable flavor-changing neutral current with minimal number of free parameters. We postulate that (i) the alignment limit is maintained, which means the lightest neutral scalar ($h$) has identical couplings to that of the Standard Model Higgs boson and (ii) the signal comes from two other neutral scalars, the CP-even $H$ and the CP-odd $A$, almost degenerate with $h$ at 125 GeV. We also show that (i) it is entirely possible that these scalars are hidden, apart from this LFV signal; (ii) the signal strengths of $b\bar{b}$, $τ^+τ^-$ and $γγ$ around 125 GeV put severe constraints on the parameter space of such models; (iii) the constraint is further enhanced by the non-observation of processes like $μ\to eγ$, and we predict that the branching ratio of $μ\to eγ$ cannot be even an order below the present experimental limit, highlighting the role it plays in forcing $H$ and $A$ to be near-degenerate; (iv) an enhancement in the $τ^+τ^-$ production cross-section at around 125 GeV is expected in the gluon fusion channel, and should be observed during the next run of the LHC; (v) the branching ratio in the $eτ$ channel is enhanced and is expected to be at least about $2\%$. The constrained parameter space and minimum number of free parameters, along with such strong predictions, make this model easily testable and falsifiable.

preprint2014arXiv

Analysis of an extended scalar sector with $S_3$ symmetry

We investigate the scalar potential of a general $S_3$-symmetric three-Higgs-doublet model. The outcome of our analysis does not depend on the fermionic sector of the model. We identify a decoupling limit for the scalar spectrum of this scenario. In view of the recent LHC Higgs data, we show our numerical results only in the decoupling limit. Unitarity and stability of the scalar potential demand that many new scalars must be lurking below 1 TeV. We provide numerical predictions for $h\to γγ$ and $h\to Z γ$ signal strengths which can be used to falsify the theory.

preprint2014arXiv

Feasibility of light scalars in a class of two-Higgs-doublet models and their decay signatures

We demonstrate that light charged and extra neutral scalars in the (100-200) GeV mass range pass the potentially dangerous flavor constraints in a particular class of two-Higgs-doublet model which has appropriately suppressed flavor-changing neutral currents at tree level. We study their decay branching ratios into various fermionic final states and comment on the possibility of their detection in the collider experiments. We also remark on how their trademark decay signatures can be used to discriminate them from the light nonstandard scalars predicted in other two-Higgs-doublet models.

preprint2014arXiv

Identifying Bengali Multiword Expressions using Semantic Clustering

One of the key issues in both natural language understanding and generation is the appropriate processing of Multiword Expressions (MWEs). MWEs pose a huge problem to the precise language processing due to their idiosyncratic nature and diversity in lexical, syntactical and semantic properties. The semantics of a MWE cannot be expressed after combining the semantics of its constituents. Therefore, the formalism of semantic clustering is often viewed as an instrument for extracting MWEs especially for resource constraint languages like Bengali. The present semantic clustering approach contributes to locate clusters of the synonymous noun tokens present in the document. These clusters in turn help measure the similarity between the constituent words of a potentially candidate phrase using a vector space model and judge the suitability of this phrase to be a MWE. In this experiment, we apply the semantic clustering approach for noun-noun bigram MWEs, though it can be extended to any types of MWEs. In parallel, the well known statistical models, namely Point-wise Mutual Information (PMI), Log Likelihood Ratio (LLR), Significance function are also employed to extract MWEs from the Bengali corpus. The comparative evaluation shows that the semantic clustering approach outperforms all other competing statistical models. As a by-product of this experiment, we have started developing a standard lexicon in Bengali that serves as a productive Bengali linguistic thesaurus.

preprint2014arXiv

Nondecoupling of charged scalars in Higgs decay to two photons and symmetries of the scalar potential

A large class of two- and three-Higgs-doublet models with discrete symmetries has been employed in the literature to address various aspects of flavor physics. We analyse how the precision measurement of the Higgs to diphoton signal strength would severely constrain these scenarios due to the nondecoupling behavior of the charged scalars, to the extent that the number of additional scalar doublets can be constrained no matter how heavy the nonstandard scalars are. We demonstrate that if the scalar potential is endowed with appropriate global continuous symmetries together with soft breaking parameters, decoupling can be achieved thanks to the unitarity constraints on the mass-square differences of the heavy scalars.

preprint2013arXiv

High precision measurement of ultraweak transitions of the a${^1}{\triangle}_g{\leftarrow}$X $^3Σ^{-}_g$ $(0,0)$ band of molecular oxygen

We have used a highly sensitive cavity-enhanced frequency modulation spectroscopy technique to measure different parameters of the ultraweak transitions of a${^1}{\triangle}_g{\leftarrow}$X $^3Σ^{-}_g$ $(0,0)$ band of molecular oxygen in the range 7640 cm$^{-1}$ to 7917 cm$^{-1}$. The self-broadened half-width and air-broadened half-width of the transitions have been measured for three different pressures for both $^{16}$O$_2$ and $^{18}$O$_2$. To measure the line intensity and self-broadened half-width we have used ultra pure oxygen sample and air-broadened half-width was measured with dry air sample. The $(0,0)$ band of $^{16}$O$_2$ and $^{18}$O$_2$ show weak quadrupole transitions with line intensities ranging from $1\times10^{-30}$ to $1.9\times10^{-28}$ cm/molecule. The measurements are in excellent agreement with the recent measurement of Rothman et. al

preprint2013arXiv

Scalar sector properties of two-Higgs-doublet models with a global U(1) symmetry

We analyze the scalar sector properties of a general class of two-Higgs-doublet models which has a global U(1) symmetry in the quartic terms. We find constraints on the parameters of the potential from the considerations of unitarity of scattering amplitudes, the global stability of the potential and the $ρ$-parameter. We concentrate on the spectrum of the non-standard scalar masses in the decoupling limit which is preferred by the Higgs data at the LHC. We exhibit charged-Higgs induced contributions to the diphoton decay width of the 125\,GeV Higgs boson and its correlation with the corresponding $Zγ$ width.

preprint2010arXiv

High-precision measurement of hyperfine structure in the $D$ lines of alkali atoms

We have measured hyperfine structure in the first-excited $P$ state ($D$ lines) of all the naturally-occurring alkali atoms. We use high-resolution laser spectroscopy to resolve hyperfine transitions, and measure intervals by locking the frequency shift produced by an acousto-optic modulator to the difference between two transitions. In most cases, the hyperfine coupling constants derived from our measurements improve previous values significantly.

preprint2010arXiv

Very long optical path-length from a compact multi-pass cell

The multiple-pass optical cell is an important tool for laser absorption spectroscopy and its many applications. For most practical applications, such as trace-gas detection, a compact and robust design is essential. Here we report an investigation into a multi-pass cell design based on a pair of cylindrical mirrors, with a particular focus on achieving very long optical paths. We demonstrate a path-length of 50.31 m in a cell with 40 mm diameter mirrors spaced 88.9 mm apart - a 3-fold increase over the previously reported longest path-length obtained with this type of cell configuration. We characterize the mechanical stability of the cell and describe the practical conditions necessary to achieve very long path-lengths.