arXiv++ Combinatorics

Browse math.CO papers from arXiv

Papers by Michelangelo Bucci

9 paper(s) by this author · All BibTeX
Aperiodic pseudorandom number generators based on infinite words
Published in Theoret. Comput. Sci. 647 (2016), 85 -- 100 • View PublicationBIB
In this paper we study how certain families of aperiodic infinite words can be used to produce aperiodic pseudorandom number generators (PRNGs) with good statistical behavior. We introduce the \emph{well distributed occurrences} (WELLDOC) combinatorial property for infinite words, which guarantees absence of the lattice structure defect in related pseudorandom number generators. An infinite word $u$ on a $d$-ary alphabet has the WELLDOC property if, for each factor $w$ of $u$, positive integer $m$, and vector $\mathbf v\in\mathbb Z_{m}^{d}$, there is an occurrence of $w$ such that the Parikh vector of the prefix of $u$ preceding such occurrence is congruent to $\mathbf v$ modulo $m$. (The Parikh vector of a finite word $v$ over an alphabet $\mathcal A$ has its $i$-th component equal to the number of occurrences of the $i$-th letter of $\mathcal A$ in $v$.) We prove that Sturmian words, and more generally Arnoux-Rauzy words and some morphic images of them, have the WELLDOC property. Using the TestU01 and PractRand statistical tests, we moreover show that not only the lattice structure is absent, but also other important properties of PRNGs are improved when linear congruential generators are combined using infinite words having the WELLDOC property.
Central sets generated by uniformly recurrent words
Published • View PublicationBIB
A subset $A$ of $\nats$ is called an IP-set if $A$ contains all finite sums of distinct terms of some infinite sequence $(x_n)_{n\in \nats} $ of natural numbers. Central sets, first introduced by Furstenberg using notions from topological dynamics, constitute a special class of IP-sets possessing rich combinatorial properties: Each central set contains arbitrarily long arithmetic progressions, and solutions to all partition regular systems of homogeneous linear equations. In this paper we investigate central sets in the framework of combinatorics on words. Using various families of uniformly recurrent words, including Sturmian words, the Thue-Morse word and fixed points of weak mixing substitutions, we generate an assortment of central sets which reflect the rich combinatorial structure of the underlying words. The results in this paper rely on interactions between different areas of mathematics, some of which had not previously been directly linked. They include the general theory of combinatorics on words, abstract numeration systems, and the beautiful theory, developed by Hindman, Strauss and others, linking IP-sets and central sets to the algebraic/topological properties of the Stone-Čech compactification of $\nats .$
On additive properties of sets defined by the Thue-Morse word
Published • View PublicationBIB
In this paper we study some additive properties of subsets of the set $\nats$ of positive integers: A subset $A$ of $\nats$ is called {\it $k$-summable} (where $k\in\ben$) if $A$ contains $\textstyle \big{\sum_{n\in F}x_n | \emp\neq F\subseteq {1,2,...,k\} \big}$ for some $k$-term sequence of natural numbers $x_1<x_2 < ... < x_k$. We say $A \subseteq \nats$ is finite FS-big if $A$ is $k$-summable for each positive integer $k$. We say is $A \subseteq \nats$ is infinite FS-big if for each positive integer $k,$ $A$ contains ${\sum_{n\in F}x_n | \emp\neq F\subseteq \nats and #F\leq k}$ for some infinite sequence of natural numbers $x_1<x_2 < ... $. We say $A\subseteq \nats $ is an IP-set if $A$ contains ${\sum_{n\in F}x_n | \emp\neq F\subseteq \nats and #F<\infty}$ for some infinite sequence of natural numbers $x_1<x_2 < ... $. By the Finite Sums Theorem [5], the collection of all IP-sets is partition regular, i.e., if $A$ is an IP-set then for any finite partition of $A$, one cell of the partition is an IP-set. Here we prove that the collection of all finite FS-big sets is also partition regular. Let $\TM =011010011001011010... $ denote the Thue-Morse word fixed by the morphism $0\mapsto 01$ and $1\mapsto 10$. For each factor $u$ of $\TM$ we consider the set $\TM\big|_u\subseteq \nats$ of all occurrences of $u$ in $\TM$. In this note we characterize the sets $\TM\big|_u$ in terms of the additive properties defined above. Using the Thue-Morse word we show that the collection of all infinite FS-big sets is not partition regular.
Reversible Christoffel factorizations
Published in Theoretical Computer Science 495 (2013) 17-24 • View PublicationBIB
We define a family of natural decompositions of Sturmian words in Christoffel words, called *reversible Christoffel* (RC) factorizations. They arise from the observation that two Sturmian words with the same language have (almost always) arbitrarily long Abelian equivalent prefixes. Using the three gap theorem, we prove that in each RC factorization, only 2 or 3 distinct Christoffel words may occur. We begin the study of such factorizations, considered as infinite words over 2 or 3 letters, and show that in the general case they are either Sturmian words, or obtained by a three-interval exchange transformation.
Some characterizations of Sturmian words in terms of the lexicographic order
Published in Fundamenta Informaticae 116 (2012) 25-33 • View PublicationBIB
In this paper we present three new characterizations of Sturmian words based on the lexicographic ordering of their factors.
Enumeration and Structure of Trapezoidal Words
Published in Theoretical Computer Science 468, 12-22 (2013) • View PublicationBIB
Trapezoidal words are words having at most $n+1$ distinct factors of length $n$ for every $n\ge 0$. They therefore encompass finite Sturmian words. We give combinatorial characterizations of trapezoidal words and exhibit a formula for their enumeration. We then separate trapezoidal words into two disjoint classes: open and closed. A trapezoidal word is closed if it has a factor that occurs only as a prefix and as a suffix; otherwise it is open. We investigate open and closed trapezoidal words, in relation with their special factors. We prove that Sturmian palindromes are closed trapezoidal words and that a closed trapezoidal word is a Sturmian palindrome if and only if its longest repeated prefix is a palindrome. We also define a new class of words, \emph{semicentral words}, and show that they are characterized by the property that they can be written as $uxyu$, for a central word $u$ and two different letters $x,y$. Finally, we investigate the prefixes of the Fibonacci word with respect to the property of being open or closed trapezoidal words, and show that the sequence of open and closed prefixes of the Fibonacci word follows the Fibonacci sequence.
Central sets defined by words of low factor complexity
A subset $A$ of $\mathbb{N}$ is called an IP-set if $A$ contains all finite sums of distinct terms of some infinite sequence $(x_n)_{n\in \mathbb{N}} $ of natural numbers. Central sets, first introduced by Furstenberg using notions from topological dynamics, constitute a special class of IP-sets possessing additional nice combinatorial properties: Each central set contains arbitrarily long arithmetic progressions, and solutions to all partition regular systems of homogeneous linear equations. In this paper we show how certain families of aperiodic words of low factor complexity may be used to generate a wide assortment of central sets having additional nice properties inherited from the rich combinatorial structure of the underlying word. We consider Sturmian words and their extensions to higher alphabets (so-called Arnoux-Rauzy words), as well as words generated by substitution rules including the famous Thue-Morse word. We also describe a connection between central sets and the strong coincidence condition for fixed points of primitive substitutions which represents a new approach to the strong coincidence conjecture for irreducible Pisot substitutions. Our methods simultaneously exploit the general theory of combinatorics on words, the arithmetic properties of abstract numeration systems defined by substitution rules, notions from topological dynamics including proximality and equicontinuity, the spectral theory of symbolic dynamical systems, and the beautiful and elegant theory, developed by N. Hindman, D. Strauss and others, linking IP-sets to the algebraic/topological properties of the Stone-Čech compactification of $\mathbb{N}.$ Using the key notion of $p$-$\lim_n,$ regarded as a mapping from words to words, we apply ideas from combinatorics on words in the framework of ultrafilters.
A new characteristic property of rich words
Published in Theoretical Computer Science 410 (2009) 2860-2863 • View PublicationBIB
Originally introduced and studied by the third and fourth authors together with J. Justin and S. Widmer in arXiv:0801.1656, rich words constitute a new class of finite and infinite words characterized by containing the maximal number of distinct palindromes. Several characterizations of rich words have already been established. A particularly nice characteristic property is that all 'complete returns' to palindromes are palindromes. In this note, we prove that rich words are also characterized by the property that each factor is uniquely determined by its longest palindromic prefix and its longest palindromic suffix.
A connection between palindromic and factor complexity using return words
Published in Advances In Applied Mathematics 42 (2009) 60--74 • View PublicationBIB
In this paper we prove that for any infinite word W whose set of factors is closed under reversal, the following conditions are equivalent: (I) all complete returns to palindromes are palindromes; (II) P(n) + P(n+1) = C(n+1) - C(n) + 2 for all n, where P (resp. C) denotes the palindromic complexity (resp. factor complexity) function of W, which counts the number of distinct palindromic factors (resp. factors) of each length in W.