Source author record

Alexey Sorokin

Alexey Sorokin appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2020arXiv

LowResourceEval-2019: a shared task on morphological analysis for low-resource languages

The paper describes the results of the first shared task on morphological analysis for the languages of Russia, namely, Evenki, Karelian, Selkup, and Veps. For the languages in question, only small-sized corpora are available. The tasks include morphological analysis, word form generation and morpheme segmentation. Four teams participated in the shared task. Most of them use machine-learning approaches, outperforming the existing rule-based ones. The article describes the datasets prepared for the shared tasks and contains analysis of the participants' solutions. Language corpora having different formats were transformed into CONLL-U format. The universal format makes the datasets comparable to other language corpura and facilitates using them in other NLP tasks.

preprint2014arXiv

Monoid automata for displacement context-free languages

In 2007 Kambites presented an algebraic interpretation of Chomsky-Schutzenberger theorem for context-free languages. We give an interpretation of the corresponding theorem for the class of displacement context-free languages which are equivalent to well-nested multiple context-free languages. We also obtain a characterization of k-displacement context-free languages in terms of monoid automata and show how such automata can be simulated on two stacks. We introduce the simultaneous two-stack automata and compare different variants of its definition. All the definitions considered are shown to be equivalent basing on the geometric interpretation of memory operations of these automata.

preprint2014arXiv

Pumping lemma and Ogden lemma for displacement context-free grammars

The pumping lemma and Ogden lemma offer a powerful method to prove that a particular language is not context-free. In 2008 Kanazawa proved an analogue of pumping lemma for well-nested multiple-context free languages. However, the statement of lemma is too weak for practical usage. We prove a stronger variant of pumping lemma and an analogue of Ogden lemma for this language family. We also use these statements to prove that some natural context-sensitive languages cannot be generated by tree-adjoining grammars.

preprint2012arXiv

Non-invertibility in Some Heteroscedastic Models

In order to calculate the unobserved volatility in conditional heteroscedastic time series models, the natural recursive approximation is very often used. Following \cite{StraumannMikosch2006}, we will call the model \emph{invertible} if this approximation (based on true parameter vector) converges to the real volatility. Our main results are necessary and sufficient conditions for invertibility. We will show that the stationary GARCH($p$, $q$) model is always invertible, but certain types of models, such as EGARCH of \cite{Nelson1991} and VGARCH of \cite{EngleNg1993} may indeed be non-invertible. Moreover, we will demonstrate it's possible for the pair (true volatility, approximation) to have a non-degenerate stationary distribution. In such cases, the volatility estimate given by the recursive approximation with the true parameter vector is inconsistent.