Papers by Alessandro De Luca
18 paper(s) by this author
· All BibTeX
Dorst-Smeulders Coding for Arbitrary Binary Words
Published
• View Publication
• BIB
A binary word is Sturmian if the occurrences of each letter are balanced, in the sense that in any two factors of the same length, the difference between the number of occurrences of the same letter is at most 1. In digital geometry, Sturmian words correspond to discrete approximations of straight line segments in the Euclidean plane. The Dorst-Smeulders coding, introduced in 1984, is a 4-tuple of integers that uniquely represents a Sturmian word $w$, enabling its reconstruction using $|w|$ modular operations, making it highly efficient in practice. In this paper, we present a linear-time algorithm that, given a binary input word $w$, computes the Dorst-Smeulders coding of its longest Sturmian prefix. This forms the basis for computing the Dorst-Smeulders coding of an arbitrary binary word $w$, which is a minimal decomposition (in terms of the number of factors) of $w$ into Sturmian words, each represented by its Dorst-Smeulders coding. This coding could be leveraged in compression schemes where the input is transformed into a binary word composed of long Sturmian segments. Although the algorithm is conceptually simple and can be implemented in just a few lines of code, it is grounded in a deep analysis of the structural properties of Sturmian words.
Digital Convexity and Combinatorics on Words
An upward (resp. downward) digitally convex word is a binary word that best approximates from below (resp. from above) an upward (resp. downward) convex curve in the plane. We study these words from the combinatorial point of view, formalizing their geometric properties and highlighting connections with Christoffel words and finite Sturmian words. In particular, we study from the combinatorial perspective the operations of inflation and deflation on digitally convex words.
Some Results on Digital Segments and Balanced Words
Published in Theoretical Computer Science 1021 (2024) 114935
• View Publication
• BIB
We exhibit combinatorial results on Christoffel words and binary balanced words that are motivated by their geometric interpretation as approximations of digital segments. We give a closed formula for counting the exact number of balanced words with $a$ zeroes and $b$ ones. We also study minimal non-balanced words.
On the Lie complexity of Sturmian words
Published in Theoretical Computer Science 938 (2022) 81-85
• View Publication
• BIB
Bell and Shallit recently introduced the Lie complexity of an infinite word $s$ as the function counting for each length the number of conjugacy classes of words whose elements are all factors of $s$. They proved, using algebraic techniques, that the Lie complexity is bounded above by the first difference of the factor complexity plus one; hence, it is uniformly bounded for words with linear factor complexity, and, in particular, it is at most 2 for Sturmian words, which are precisely the words with factor complexity $n+1$ for every $n$. In this note, we provide an elementary combinatorial proof of the result of Bell and Shallit and give an exact formula for the Lie complexity of any Sturmian word.
On the extremal values of the cyclic continuants of Motzkin and Straus
In a 1983 paper, G. Ramharter asks what are the extremal arrangements for the cyclic analogues of the regular and semi-regular continuants first introduced by T.S. Motzkin and E.G. Straus in 1956. In this paper we answer this question by showing that for each set $A$ consisting of positive integers $1<a_1<a_2<\cdots <a_k$ and a $k$-term partition $P: n_1+n_2 + \cdots + n_k=n$, there exists a unique (up to reversal) cyclic word $x$ which maximizes (resp. minimizes) the regular cyclic continuant $K^{\circlearrowright}(\cdot)$ amongst all cyclic words over $A$ with Parikh vector $(n_1,n_2,\ldots,n_k)$. We also show that the same is true for the minimizing arrangement for the semi-regular cyclic continuant $\dot K^{\circlearrowright}(\cdot)$. As in the non-cyclic case, the main difficulty is to find the maximizing arrangement for the semi-regular continuant, which is not unique in general and may depend on the integers $a_1,\ldots,a_k$ and not just on their relative order. We show that if a cyclic word $x$ maximizes $\dot K^{\circlearrowright}(\cdot)$ amongst all permutations of $x$, then it verifies a strong combinatorial condition which we call the singular property. We develop an algorithm for constructing all singular cyclic words having a prescribed Parikh vector.
Extremal values of semi-regular continuants and codings of interval exchange transformations
Published in Mathematika 69 (2023) 432-457
• View Publication
• BIB
Given a set $A$ of positive integers $a_1<\cdots<a_k$ and a partition $P: n_1+\cdots+n_k=n$, find the extremal denominators of the regular and semi-regular continued fraction $[0;x_1,\ldots,x_n]$ with partial quotients $x_i\in A$ and where each $a_i$ occurs exactly $n_i$ times in $x_1,\ldots,x_n$. In 1983, G. Ramharter gave an explicit description of the extremal arrangements of the regular continued fraction and the minimizing arrangement for the semi-regular continued fraction and showed that in each case the arrangement is unique up to reversal and independent of the actual values of the integers $a_i$. However, an explicit determination of a maximizing arrangement for the semi-regular continuant turned out to be more difficult. Ramharter conjectured that as in the other three cases, the maximizing arrangement is unique up to reversal and depends only on the partition $P$ and not on the values of the $a_i$. He further verified the conjecture in the case of a binary $A$. In this paper we confirm Ramharter's conjecture for sets $A$ with $|A|=3$ and give an algorithmic construction for the unique maximizing arrangement. We also show that Ramharter's conjecture fails for sets with $|A|\geq 4$, as the maximizing arrangement is in general neither unique nor independent of the values of the digits in $A$. The central idea is that the extremal arrangements satisfy a strong combinatorial condition, which may also be stated in the context of infinite sequences on an ordered set. We show that for bi-infinite binary words, this condition coincides with the Markoff property, discovered by A.A. Markoff in 1879 in his study of minima of binary quadratic forms. We further show that this same combinatorial condition is the fundamental property which describes the orbit structure of the natural codings of points under a symmetric $k$-interval exchange transformation.
Characteristic Parameters and Special Trapezoidal Words
Following earlier work by Aldo de Luca and others, we study trapezoidal words and their prefixes, with respect to their characteristic parameters $K$ and $R$ (length of shortest unrepeated suffix, and shortest length without right special factors, respectively), as well as their symmetric versions $H$ and $L$. We consider the distinction between closed (i.e., periodic-like) and open prefixes, and between Sturmian and non-Sturmian ones. Our main results characterize right special and strictly bispecial trapezoidal words, as done by de Luca and Mignosi for Sturmian words.
The sequence of open and closed prefixes of a Sturmian word
Published in Advances in Applied Mathematics Volume 90, September 2017, Pages 27-45
• View Publication
• BIB
A finite word is closed if it contains a factor that occurs both as a prefix and as a suffix but does not have internal occurrences, otherwise it is open. We are interested in the {\it oc-sequence} of a word, which is the binary sequence whose $n$-th element is $0$ if the prefix of length $n$ of the word is open, or $1$ if it is closed. We exhibit results showing that this sequence is deeply related to the combinatorial and periodic structure of a word. In the case of Sturmian words, we show that these are uniquely determined (up to renaming letters) by their oc-sequence. Moreover, we prove that the class of finite Sturmian words is a maximal element with this property in the class of binary factorial languages. We then discuss several aspects of Sturmian words that can be expressed through this sequence. Finally, we provide a linear-time algorithm that computes the oc-sequence of a finite word, and a linear-time algorithm that reconstructs a finite Sturmian word from its oc-sequence.
On Christoffel and standard words and their derivatives
We introduce and study natural derivatives for Christoffel and finite standard words, as well as for characteristic Sturmian words. These derivatives, which are realized as inverse images under suitable morphisms, preserve the aforementioned classes of words. In the case of Christoffel words, the morphisms involved map $a$ to $a^{k+1}b$ (resp.,~$ab^{k}$) and $b$ to $a^{k}b$ (resp.,~$ab^{k+1}$) for a suitable $k>0$. As long as derivatives are longer than one letter, higher-order derivatives are naturally obtained. We define the depth of a Christoffel or standard word as the smallest order for which the derivative is a single letter. We give several combinatorial and arithmetic descriptions of the depth, and (tight) lower and upper bounds for it.
Sturmian words and the Stern sequence
Published
• View Publication
• BIB
Central, standard, and Christoffel words are three strongly interrelated classes of binary finite words which represent a finite counterpart of characteristic Sturmian words. A natural arithmetization of the theory is obtained by representing central and Christoffel words by irreducible fractions labeling respectively two binary trees, the Raney (or Calkin-Wilf) tree and the Stern-Brocot tree. The sequence of denominators of the fractions in Raney's tree is the famous Stern diatomic numerical sequence. An interpretation of the terms $s(n)$ of Stern's sequence as lengths of Christoffel words when $n$ is odd, and as minimal periods of central words when $n$ is even, allows one to interpret several results on Christoffel and central words in terms of Stern's sequence and, conversely, to obtain a new insight in the combinatorics of Christoffel and central words by using properties of Stern's sequence. One of our main results is a non-commutative version of the "alternating bit sets theorem" by Calkin and Wilf. We also study the length distribution of Christoffel words corresponding to nodes of equal height in the tree, obtaining some interesting bounds and inequalities.
Aperiodic pseudorandom number generators based on infinite words
Published in Theoret. Comput. Sci. 647 (2016), 85 -- 100
• View Publication
• BIB
In this paper we study how certain families of aperiodic infinite words can be used to produce aperiodic pseudorandom number generators (PRNGs) with good statistical behavior. We introduce the \emph{well distributed occurrences} (WELLDOC) combinatorial property for infinite words, which guarantees absence of the lattice structure defect in related pseudorandom number generators. An infinite word $u$ on a $d$-ary alphabet has the WELLDOC property if, for each factor $w$ of $u$, positive integer $m$, and vector $\mathbf v\in\mathbb Z_{m}^{d}$, there is an occurrence of $w$ such that the Parikh vector of the prefix of $u$ preceding such occurrence is congruent to $\mathbf v$ modulo $m$. (The Parikh vector of a finite word $v$ over an alphabet $\mathcal A$ has its $i$-th component equal to the number of occurrences of the $i$-th letter of $\mathcal A$ in $v$.) We prove that Sturmian words, and more generally Arnoux-Rauzy words and some morphic images of them, have the WELLDOC property. Using the TestU01 and PractRand statistical tests, we moreover show that not only the lattice structure is absent, but also other important properties of PRNGs are improved when linear congruential generators are combined using infinite words having the WELLDOC property.
Open and Closed Prefixes of Sturmian Words
Published in Lecture Notes in Computer Science 8079: 132-142 (2013)
• View Publication
• BIB
A word is closed if it contains a proper factor that occurs both as a prefix and as a suffix but does not have internal occurrences, otherwise it is open. We deal with the sequence of open and closed prefixes of Sturmian words and prove that this sequence characterizes every finite or infinite Sturmian word up to isomorphisms of the alphabet. We then characterize the combinatorial structure of the sequence of open and closed prefixes of standard Sturmian words. We prove that every standard Sturmian word, after swapping its first letter, can be written as an infinite product of squares of reversed standard words.
Reversible Christoffel factorizations
Published in Theoretical Computer Science 495 (2013) 17-24
• View Publication
• BIB
We define a family of natural decompositions of Sturmian words in Christoffel words, called *reversible Christoffel* (RC) factorizations. They arise from the observation that two Sturmian words with the same language have (almost always) arbitrarily long Abelian equivalent prefixes. Using the three gap theorem, we prove that in each RC factorization, only 2 or 3 distinct Christoffel words may occur. We begin the study of such factorizations, considered as infinite words over 2 or 3 letters, and show that in the general case they are either Sturmian words, or obtained by a three-interval exchange transformation.
Some characterizations of Sturmian words in terms of the lexicographic order
Published in Fundamenta Informaticae 116 (2012) 25-33
• View Publication
• BIB
In this paper we present three new characterizations of Sturmian words based on the lexicographic ordering of their factors.
Enumeration and Structure of Trapezoidal Words
Published in Theoretical Computer Science 468, 12-22 (2013)
• View Publication
• BIB
Trapezoidal words are words having at most $n+1$ distinct factors of length $n$ for every $n\ge 0$. They therefore encompass finite Sturmian words. We give combinatorial characterizations of trapezoidal words and exhibit a formula for their enumeration. We then separate trapezoidal words into two disjoint classes: open and closed. A trapezoidal word is closed if it has a factor that occurs only as a prefix and as a suffix; otherwise it is open. We investigate open and closed trapezoidal words, in relation with their special factors. We prove that Sturmian palindromes are closed trapezoidal words and that a closed trapezoidal word is a Sturmian palindrome if and only if its longest repeated prefix is a palindrome. We also define a new class of words, \emph{semicentral words}, and show that they are characterized by the property that they can be written as $uxyu$, for a central word $u$ and two different letters $x,y$. Finally, we investigate the prefixes of the Fibonacci word with respect to the property of being open or closed trapezoidal words, and show that the sequence of open and closed prefixes of the Fibonacci word follows the Fibonacci sequence.
A generalized palindromization map in free monoids
Published in Theoretical Computer Science 454 (2012) 109-128
• View Publication
• BIB
The palindromization map $ψ$ in a free monoid $A^*$ was introduced in 1997 by the first author in the case of a binary alphabet $A$, and later extended by other authors to arbitrary alphabets. Acting on infinite words, $ψ$ generates the class of standard episturmian words, including standard Arnoux-Rauzy words. In this paper we generalize the palindromization map, starting with a given code $X$ over $A$. The new map $ψ_X$ maps $X^*$ to the set $PAL$ of palindromes of $A^*$. In this way some properties of $ψ$ are lost and some are saved in a weak form. When $X$ has a finite deciphering delay one can extend $ψ_X$ to $X^ω$, generating a class of infinite words much wider than standard episturmian words. For a finite and maximal code $X$ over $A$, we give a suitable generalization of standard Arnoux-Rauzy words, called $X$-AR words. We prove that any $X$-AR word is a morphic image of a standard Arnoux-Rauzy word and we determine some suitable linear lower and upper bounds to its factor complexity.
For any code $X$ we say that $ψ_X$ is conservative when $ψ_X(X^{*})\subseteq X^{*}$. We study conservative maps $ψ_X$ and conditions on $X$ assuring that $ψ_X$ is conservative. We also investigate the special case of morphic-conservative maps $ψ_{X}$, i.e., maps such that $φ\circ ψ= ψ_X\circ φ$ for an injective morphism $φ$. Finally, we generalize $ψ_X$ by replacing palindromic closure with $θ$-palindromic closure, where $θ$ is any involutory antimorphism of $A^*$. This yields an extension of the class of $θ$-standard words introduced by the authors in 2006.
A new characteristic property of rich words
Published in Theoretical Computer Science 410 (2009) 2860-2863
• View Publication
• BIB
Originally introduced and studied by the third and fourth authors together with J. Justin and S. Widmer in arXiv:0801.1656, rich words constitute a new class of finite and infinite words characterized by containing the maximal number of distinct palindromes. Several characterizations of rich words have already been established. A particularly nice characteristic property is that all 'complete returns' to palindromes are palindromes. In this note, we prove that rich words are also characterized by the property that each factor is uniquely determined by its longest palindromic prefix and its longest palindromic suffix.
A connection between palindromic and factor complexity using return words
Published in Advances In Applied Mathematics 42 (2009) 60--74
• View Publication
• BIB
In this paper we prove that for any infinite word W whose set of factors is closed under reversal, the following conditions are equivalent:
(I) all complete returns to palindromes are palindromes;
(II) P(n) + P(n+1) = C(n+1) - C(n) + 2 for all n, where P (resp. C) denotes the palindromic complexity (resp. factor complexity) function of W, which counts the number of distinct palindromic factors (resp. factors) of each length in W.