Source author record

Svetlana Puzynina

Svetlana Puzynina appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

21works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

21 published item(s)

preprint2023arXiv

Finite and infinite closed-rich words

A word is called closed if it has a prefix which is also its suffix and there is no internal occurrences of this prefix in the word. In this paper we study words that are rich in closed factors, i.e., which contain the maximal possible number of distinct closed factors. As the main result, we show that for finite words the asymptotics of the maximal number of distinct closed factors in a word of length $n$ is $\frac{n^2}{6}$. For infinite words, we show there exist words such that each their factor of length $n$ contains a quadratic number of distinct closed factors, with uniformly bounded constant; we call such words infinite closed-rich. We provide several necessary and some sufficient conditions for a word to be infinite closed rich. For example, we show that all linearly recurrent words are closed-rich. We provide a characterization of rich words among Sturmian words. Certain examples we provide involve non-constructive methods.

preprint2022arXiv

Abelian Combinatorics on Words: a Survey

We survey known results and open problems in abelian combinatorics on words. Abelian combinatorics on words is the extension to the commutative setting of the classical theory of combinatorics on words. The extension is based on \emph{abelian equivalence}, which is the equivalence relation defined in the set of words by having the same Parikh vector, that is, the same number of occurrences of each letter of the alphabet. In the past few years, there was a lot of research on abelian analogues of classical definitions and properties in combinatorics on words. This survey aims to gather these results.

preprint2020arXiv

On Abelian Closures of Infinite Non-binary Words

Two finite words $u$ and $v$ are called abelian equivalent if each letter occurs equally many times in both $u$ and $v$. The abelian closure $\mathcal{A}(\mathbf{x})$ of an infinite word $\mathbf{x}$ is the set of infinite words $\mathbf{y}$ such that, for each factor $u$ of $\mathbf{y}$, there exists a factor $v$ of $\mathbf{x}$ which is abelian equivalent to $u$. The notion of an abelian closure gives a characterization of Sturmian words: among uniformly recurrent binary words, periodic and aperiodic Sturmian words are exactly those words for which $\mathcal{A}(\mathbf{x})$ equals the shift orbit closure $Ω(\mathbf{x})$. Furthermore, for an aperiodic binary word that is not Sturmian, its abelian closure contains infinitely many minimial subshifts. In this paper we consider the abelian closures of well-known families of non-binary words, such as balanced words and minimal complexity words. We also consider abelian closures of general subshifts and make some initial observations of their abelian closures and pose some related open questions.

preprint2020arXiv

Recurrence along directions in multidimensional words

In this paper we introduce and study new notions of uniform recurrence in multidimensional words. A $d$-dimensional word is called \emph{uniformly recurrent} if for all $(s_1,\ldots,s_d)\in\mathbb{N}^d$ there exists $n\in\mathbb{N}$ such that each block of size $(n,\ldots,n)$ contains the prefix of size $(s_1,\ldots,s_d)$. We are interested in a modification of this property. Namely, we ask that for each rational direction $(q_1,\ldots,q_d)$, each rectangular prefix occurs along this direction in positions $\ell(q_1,\ldots,q_d)$ with bounded gaps. Such words are called \emph{uniformly recurrent along all directions}. We provide several constructions of multidimensional words satisfying this condition, and more generally, a series of four increasingly stronger conditions. In particular, we study the uniform recurrence along directions of multidimentional rotation words and of fixed points of square morphisms.

preprint2016arXiv

Cost and dimension of words of zero topological entropy

Let $A^*$ denote the free monoid generated by a finite nonempty set $A.$ In this paper we introduce a new measure of complexity of languages $L\subseteq A^*$ defined in terms of the semigroup structure on $A^*.$ For each $L\subseteq A^*,$ we define its {\it cost} $c(L)$ as the infimum of all real numbers $α$ for which there exist a language $S\subseteq A^*$ with $p_S(n)=O(n^α)$ and a positive integer $k$ with $L\subseteq S^k.$ We also define the {\it cost dimension} $d_c(L)$ as the infimum of the set of all positive integers $k$ such that $L\subseteq S^k$ for some language $S$ with $p_S(n)=O(n^{c(L)}).$ We are primarily interested in languages $L$ given by the set of factors of an infinite word $x=x_0x_1x_2\cdots \in A^ω$ of zero topological entropy, in which case $c(L)<+\infty.$ We establish the following characterisation of words of linear factor complexity: Let $x\in A^ω$ and $L=$Fac$(x)$ be the set of factors of $x.$ Then $p_x(n)=Θ(n)$ if and only $c(L)=0$ and $d_c(L)=2.$ In other words, $p_x(n)=O(n)$ if and only if Fac$(x)\subseteq S^2$ for some language $S\subseteq A^+$ of bounded complexity (meaning $\limsup p_S(n)<+\infty).$ In general the cost of a language $L$ reflects deeply the underlying combinatorial structure induced by the semigroup structure on $A^*.$ For example, in contrast to the above characterisation of languages generated by words of sub-linear complexity, there exist non factorial languages $L$ of complexity $p_L(n)=O(\log n)$ (and hence of cost equal to $0)$ and of cost dimension $+\infty.$ In this paper we investigate the cost and cost dimension of languages defined by infinite words of zero topological entropy.

preprint2016arXiv

Minimal complexity of equidistributed infinite permutations

An infinite permutation is a linear ordering of the set of natural numbers. An infinite permutation can be defined by a sequence of real numbers where only the order of elements is taken into account. In the paper we investigate a new class of {\it equidistributed} infinite permutations, that is, infinite permutations which can be defined by equidistributed sequences. Similarly to infinite words, a complexity $p(n)$ of an infinite permutation is defined as a function counting the number of its subpermutations of length $n$. For infinite words, a classical result of Morse and Hedlund, 1938, states that if the complexity of an infinite word satisfies $p(n) \leq n$ for some $n$, then the word is ultimately periodic. Hence minimal complexity of aperiodic words is equal to $n+1$, and words with such complexity are called Sturmian. For infinite permutations this does not hold: There exist aperiodic permutations with complexity functions growing arbitrarily slowly, and hence there are no permutations of minimal complexity. We show that, unlike for permutations in general, the minimal complexity of an equidistributed permutation $α$ is $p_α(n)=n$. The class of equidistributed permutations of minimal complexity coincides with the class of so-called Sturmian permutations, directly related to Sturmian words.

preprint2016arXiv

On cardinalities of $k$-abelian equivalence classes

Two words $u$ and $v$ are $k$-abelian equivalent if, for each word $x$ of length at most $k$, $x$ occurs equally many times as a factor in both $u$ and $v$. The notion of $k$-abelian equivalence is an intermediate notion between the abelian equivalence and the equality of words. In this paper, we study the equivalence classes induced by the $k$-abelian equivalence, mainly focusing on the cardinalities of the classes. In particular, we are interested in the number of singleton $k$-abelian classes, i.e., classes containing only one element. We find a connection between the singleton classes and cycle decompositions of the de Bruijn graph. We show that the number of classes of words of length $n$ containing one single element is of order $\mathcal O(n^{N_m(k-1)-1})$, where $N_m(l) = \tfrac{1}{l}\sum_{d\mid l} φ(d)m^{l/d}$ is the number of necklaces of length $l$ over an $m$-ary alphabet. We conjecture that the upper bound is sharp. We also remark that, for $k$ even and $m = 2$, the lower bound $Ω(n^{N_m(k-1)-1})$ follows from an old conjecture on the existence of Gray codes for necklaces of odd length. We verify this conjecture for necklaces of length up to 15.

preprint2015arXiv

Abelian bordered factors and periodicity

A finite word u is said to be bordered if u has a proper prefix which is also a suffix of u, and unbordered otherwise. Ehrenfeucht and Silberger proved that an infinite word is purely periodic if and only if it contains only finitely many unbordered factors. We are interested in abelian and weak abelian analogues of this result; namely, we investigate the following question(s): Let w be an infinite word such that all sufficiently long factors are (weakly) abelian bordered; is w (weakly) abelian periodic? In the process we answer a question of Avgustinovich et al. concerning the abelian critical factorization theorem.

preprint2015arXiv

Canonical Representatives of Morphic Permutations

An infinite permutation can be defined as a linear ordering of the set of natural numbers. In particular, an infinite permutation can be constructed with an aperiodic infinite word over $\{0,\ldots,q-1\}$ as the lexicographic order of the shifts of the word. In this paper, we discuss the question if an infinite permutation defined this way admits a canonical representative, that is, can be defined by a sequence of numbers from [0, 1], such that the frequency of its elements in any interval is equal to the length of that interval. We show that a canonical representative exists if and only if the word is uniquely ergodic, and that is why we use the term ergodic permutations. We also discuss ways to construct the canonical representative of a permutation defined by a morphic word and generalize the construction of Makarov, 2009, for the Thue-Morse permutation to a wider class of infinite words.

preprint2015arXiv

On a group theoretic generalization of the Morse-Hedlund theorem

In their 1938 seminal paper on symbolic dynamics, Morse and Hedlund proved that every aperiodic infinite word $x\in A^N,$ over a non empty finite alphabet $A,$ contains at least $n+1$ distinct factors of each length $n.$ They further showed that an infinite word $x$ has exactly $n+1$ distinct factors of each length $n$ if and only if $x$ is binary, aperiodic and balanced, i.e., $x$ is a Sturmian word. In this paper we obtain a broad generalization of the Morse-Hedlund theorem via group actions. Given a subgroup $G$ of the symmetric group $S_n, $ let $1\leq ε(G)\leq n$ denote the number of distinct $G$-orbits of $\{1,2,\ldots ,n\}.$ Since $G$ is a subgroup of $S_n,$ it acts on $A^n=\{a_1a_2\cdots a_n\,|\,a_i\in A\}$ by permutation. Thus, given an infinite word $x\in A^N$ and an infinite sequence $ω=(G_n)_{n\geq 1}$ of subgroups $G_n \subseteq S_n,$ we consider the complexity function $p_{ω,x}:N \rightarrow N$ which counts for each length $n$ the number of equivalence classes of factors of $x$ of length $n$ under the action of $G_n.$ We show that if $x$ is aperiodic, then $p_{ω, x}(n)\geqε(G_n)+1$ for each $n\geq 1,$ and moreover, if equality holds for each $n,$ then $x$ is Sturmian. Conversely, let $x$ be a Sturmian word. Then for every infinite sequence $ω=(G_n)_{n\geq 1}$ of Abelian subgroups $G_n \subseteq S_n,$ there exists $ω'=(G_n')_{n\geq 1}$ such that for each $n\geq 1:$ $G_n'\subseteq S_n$ is isomorphic to $G_n$ and $p_{ω',x}(n)=ε(G'_n)+1.$ Applying the above results to the sequence $(Id_n)_{n\geq 1},$ where $Id_n$ is the trivial subgroup of $S_n$ consisting only of the identity, we recover both directions of the Morse-Hedland theorem.

preprint2014arXiv

Aperiodic pseudorandom number generators based on infinite words

In this paper we study how certain families of aperiodic infinite words can be used to produce aperiodic pseudorandom number generators (PRNGs) with good statistical behavior. We introduce the \emph{well distributed occurrences} (WELLDOC) combinatorial property for infinite words, which guarantees absence of the lattice structure defect in related pseudorandom number generators. An infinite word $u$ on a $d$-ary alphabet has the WELLDOC property if, for each factor $w$ of $u$, positive integer $m$, and vector $\mathbf v\in\mathbb Z_{m}^{d}$, there is an occurrence of $w$ such that the Parikh vector of the prefix of $u$ preceding such occurrence is congruent to $\mathbf v$ modulo $m$. (The Parikh vector of a finite word $v$ over an alphabet $\mathcal A$ has its $i$-th component equal to the number of occurrences of the $i$-th letter of $\mathcal A$ in $v$.) We prove that Sturmian words, and more generally Arnoux-Rauzy words and some morphic images of them, have the WELLDOC property. Using the TestU01 and PractRand statistical tests, we moreover show that not only the lattice structure is absent, but also other important properties of PRNGs are improved when linear congruential generators are combined using infinite words having the WELLDOC property.

preprint2014arXiv

Infinite Self-Shuffling Words

In this paper we introduce and study a new property of infinite words: An infinite word $x\in A^\mathbb{N}$, with values in a finite set $A$, is said to be $k$-self-shuffling $(k\geq 2)$ if $x$ admits factorizations: $x=\prod_{i=0}^\infty U_i^{(1)}\cdots U_i^{(k)}=\prod_{i=0}^\infty U_i^{(1)}=\cdots =\prod_{i=0}^\infty U_i^{(k)}$. In other words, there exists a shuffle of $k$-copies of $x$ which produces $x$. We are particularly interested in the case $k=2$, in which case we say $x$ is self-shuffling. This property of infinite words is shown to be an intrinsic property of the word and not of its language (set of factors). For instance, every aperiodic word contains a non self-shuffling word in its shift orbit closure. While the property of being self-shuffling is a relatively strong condition, many important words arising in the area of symbolic dynamics are verified to be self-shuffling. They include for instance the Thue-Morse word and all Sturmian words of intercept $0<ρ<1$ (while those of intercept $ρ=0$ are not self-shuffling). Our characterization of self-shuffling Sturmian words can be interpreted arithmetically in terms of a dynamical embedding and defines an arithmetic process we call the {\it stepping stone model}. One important feature of self-shuffling words stems from its morphic invariance, which provides a useful tool for showing that one word is not the morphic image of another. The notion of self-shuffling has other unexpected applications particularly in the area of substitutive dynamical systems. For example, as a consequence of our characterization of self-shuffling Sturmian words, we recover a number theoretic result, originally due to Yasutomi, on a classification of pure morphic Sturmian words in the orbit of the characteristic.

preprint2014arXiv

Infinite square-free self-shuffling words

In this paper we answer two recent questions from Charlier et al. and Harju about self-shuffling words. An infinite word $w$ is called self-shuffling, if $w=\prod_{i=0}^\infty U_iV_i=\prod_{i=0}^\infty U_i=\prod_{i=0}^\infty V_i$ for some finite words $U_i$, $V_i$. Harju recently asked whether square-free self-shuffling words exist. We answer this question affirmatively. Besides that, we build an infinite word such that no word in its shift orbit closure is self-shuffling, answering positively a question from Charlier et al.

preprint2013arXiv

Central sets defined by words of low factor complexity

A subset $A$ of $\mathbb{N}$ is called an IP-set if $A$ contains all finite sums of distinct terms of some infinite sequence $(x_n)_{n\in \mathbb{N}} $ of natural numbers. Central sets, first introduced by Furstenberg using notions from topological dynamics, constitute a special class of IP-sets possessing additional nice combinatorial properties: Each central set contains arbitrarily long arithmetic progressions, and solutions to all partition regular systems of homogeneous linear equations. In this paper we show how certain families of aperiodic words of low factor complexity may be used to generate a wide assortment of central sets having additional nice properties inherited from the rich combinatorial structure of the underlying word. We consider Sturmian words and their extensions to higher alphabets (so-called Arnoux-Rauzy words), as well as words generated by substitution rules including the famous Thue-Morse word. We also describe a connection between central sets and the strong coincidence condition for fixed points of primitive substitutions which represents a new approach to the strong coincidence conjecture for irreducible Pisot substitutions. Our methods simultaneously exploit the general theory of combinatorics on words, the arithmetic properties of abstract numeration systems defined by substitution rules, notions from topological dynamics including proximality and equicontinuity, the spectral theory of symbolic dynamical systems, and the beautiful and elegant theory, developed by N. Hindman, D. Strauss and others, linking IP-sets to the algebraic/topological properties of the Stone-Čech compactification of $\mathbb{N}.$ Using the key notion of $p$-$\lim_n,$ regarded as a mapping from words to words, we apply ideas from combinatorics on words in the framework of ultrafilters.

preprint2013arXiv

Central sets generated by uniformly recurrent words

A subset $A$ of $\nats$ is called an IP-set if $A$ contains all finite sums of distinct terms of some infinite sequence $(x_n)_{n\in \nats} $ of natural numbers. Central sets, first introduced by Furstenberg using notions from topological dynamics, constitute a special class of IP-sets possessing rich combinatorial properties: Each central set contains arbitrarily long arithmetic progressions, and solutions to all partition regular systems of homogeneous linear equations. In this paper we investigate central sets in the framework of combinatorics on words. Using various families of uniformly recurrent words, including Sturmian words, the Thue-Morse word and fixed points of weak mixing substitutions, we generate an assortment of central sets which reflect the rich combinatorial structure of the underlying words. The results in this paper rely on interactions between different areas of mathematics, some of which had not previously been directly linked. They include the general theory of combinatorics on words, abstract numeration systems, and the beautiful theory, developed by Hindman, Strauss and others, linking IP-sets and central sets to the algebraic/topological properties of the Stone-Čech compactification of $\nats .$

preprint2013arXiv

On additive properties of sets defined by the Thue-Morse word

In this paper we study some additive properties of subsets of the set $\nats$ of positive integers: A subset $A$ of $\nats$ is called {\it $k$-summable} (where $k\in\ben$) if $A$ contains $\textstyle \big{\sum_{n\in F}x_n | \emp\neq F\subseteq {1,2,...,k\} \big}$ for some $k$-term sequence of natural numbers $x_1<x_2 < ... < x_k$. We say $A \subseteq \nats$ is finite FS-big if $A$ is $k$-summable for each positive integer $k$. We say is $A \subseteq \nats$ is infinite FS-big if for each positive integer $k,$ $A$ contains ${\sum_{n\in F}x_n | \emp\neq F\subseteq \nats and #F\leq k}$ for some infinite sequence of natural numbers $x_1<x_2 < ... $. We say $A\subseteq \nats $ is an IP-set if $A$ contains ${\sum_{n\in F}x_n | \emp\neq F\subseteq \nats and #F<\infty}$ for some infinite sequence of natural numbers $x_1<x_2 < ... $. By the Finite Sums Theorem [5], the collection of all IP-sets is partition regular, i.e., if $A$ is an IP-set then for any finite partition of $A$, one cell of the partition is an IP-set. Here we prove that the collection of all finite FS-big sets is also partition regular. Let $\TM =011010011001011010... $ denote the Thue-Morse word fixed by the morphism $0\mapsto 01$ and $1\mapsto 10$. For each factor $u$ of $\TM$ we consider the set $\TM\big|_u\subseteq \nats$ of all occurrences of $u$ in $\TM$. In this note we characterize the sets $\TM\big|_u$ in terms of the additive properties defined above. Using the Thue-Morse word we show that the collection of all infinite FS-big sets is not partition regular.

preprint2013arXiv

Weak abelian periodicity of infinite words

We say that an infinite word w is weak abelian periodic if it can be factorized into finite words with the same frequencies of letters. In the paper we study properties of weak abelian periodicity, its relations with balance and frequency. We establish necessary and sufficient conditions for weak abelian periodicity of fixed points of uniform binary morphisms. Also, we discuss weak abelian periodicity in minimal subshifts.

preprint2012arXiv

A regularity lemma and twins in words

For a word $S$, let $f(S)$ be the largest integer $m$ such that there are two disjoints identical (scattered) subwords of length $m$. Let $f(n, Σ) = \min \{f(S): S \text{is of length} n, \text{over alphabet} Σ\}$. Here, it is shown that \[2f(n, \{0,1\}) = n-o(n)\] using the regularity lemma for words. I.e., any binary word of length $n$ can be split into two identical subwords (referred to as twins) and, perhaps, a remaining subword of length $o(n)$. A similar result is proven for $k$ identical subwords of a word over an alphabet with at most $k$ letters.

preprint2012arXiv

Abelian returns in Sturmian words

Return words constitute a powerful tool for studying symbolic dynamical systems. They may be regarded as a discrete analogue of the first return map in dynamical systems. In this paper we investigate two abelian variants of the notion of return word, each of them gives rise to a new characterization of Sturmian words. We prove that a recurrent infinite word is Sturmian if and only if each of its factors has two or three abelian (or semi-abelian) returns. We study the structure of abelian returns in Sturmian words and give a characterization of those factors having exactly two abelian returns. Finally we discuss connections between abelian returns and periodicity in words.

preprint2012arXiv

On minimal factorizations of words as products of palindromes

Given a finite word u, we define its palindromic length |u|_{pal} to be the least number n such that u=v_1v_2... v_n with each v_i a palindrome. We address the following open question: Does there exist an infinite non ultimately periodic word w and a positive integer P such that |u|_{pal}<P for each factor u of w? We give a partial answer to this question by proving that if an infinite word w satisfies the so-called (k,l)-condition for some k and l, then for each positive integer P there exists a factor u of w whose palindromic length |u|_{pal}>P. In particular, the result holds for all the k-power-free words and for the Sierpinski word.