Papers by Michel Rigo
29 paper(s) by this author
· All BibTeX
Symbols frequencies in the Thue--Morse word in base $3/2$ and related conjectures
We study a binary Thue--Morse-type sequence arising from the base-$3/2$ expansion of integers, an archetypal automatic sequence in a rational base numeration system. Because the sequence is generated by a periodic iteration of morphisms rather than a single primitive substitution, classical Perron--Frobenius methods do not directly apply to determine symbol frequencies. We prove that both symbols ${\tt 0},{\tt 1}$ occur with frequency $1/2$ and we show uniform recurrence and symmetry properties of its set of factors. The proof reveals a structural bridge between combinatorics on words and harmonic analysis: the first difference sequence is shown to be Toeplitz, providing dynamical rigidity, while filtered frequencies naturally encode a dyadic structure that lifts to the compact group of $2$-adic integers. In this $2$-adic setting, desubstitution becomes a linear operator on Fourier coefficients, and a spectral contraction argument enforces uniqueness of limiting densities. Our results answer several conjectures of Dekking (on a sibling sequence) and illustrate how harmonic analysis on compact groups can be fruitfully combined with substitution dynamics.
Variants of Wythoff game with terminal positions or blocking maneuvers
We show how the software Walnut can be used to obtain concise proofs of results concerning variants of the famous Wythoff game, in which blocking maneuvers or terminal positions are added, as discussed respectively by Larsson (2011) and Komak et al. (2025). Our approach provides automatic proofs that both confirm and extend their results, and the same techniques apply to newly introduced variants as well.
Then, using classic techniques, we obtain new recursive and morphic characterizations of Wythoff-type games where the set of terminal positions $(x,y)$ satisfy $x+y\le\ell$.
The use of Walnut in combinatorial game theory is relatively recent, and only a few examples have been explored so far. The Wythoff game, being directly connected to the Fibonacci numeration system, proves especially well-suited to this kind of approach. It permits us to solve instances for a fixed value of a parameter.
Computing Expansions in Infinitely Many Cantor Real Bases via a Single Transducer
Representing real numbers using convenient numeration systems (integer bases, $β$-numeration, Cantor bases, etc.) has been a longstanding mathematical challenge. This paper focuses on Cantor real bases and, specifically, on automatic Cantor real bases and the properties of expansions of real numbers in this setting. We develop a new approach where a single transducer associated with a fixed real number $r$, computes the $\mathbf{B}$-expansion of $r$ but for an infinite family of Cantor real bases $\mathbf{B}$ given as input. This point of view contrasts with traditional computational models for which the numeration system is fixed. Under some assumptions on the finitely many Pisot numbers occurring in the Cantor real base, we show that only a finite part of the transducer is visited. We obtain fundamental results on the structure of this transducer and on decidability problems about these expansions, proving that for certain classes of Cantor real bases, key combinatorial properties such as greediness of the expansion or periodicity can be decided algorithmically.
Automatic Abelian Complexities of Parikh-Collinear Fixed Points
Parikh-collinear morphisms have the property that all the Parikh vectors of the images of letters are collinear, i.e., the associated adjacency matrix has rank 1. In the conference DLT-WORDS 2023 we showed that fixed points of Parikh-collinear morphisms are automatic. We also showed that the abelian complexity function of a binary fixed point of such a morphism is automatic under some assumptions. In this note, we fully generalize the latter result. Namely, we show that the abelian complexity function of a fixed point of an arbitrary, possibly erasing, Parikh-collinear morphism is automatic. Furthermore, a deterministic finite automaton with output generating this abelian complexity function is provided by an effective procedure. To that end, we discuss the constant of recognizability of a morphism and the related cutting set.
q-Parikh Matrices and q-deformed binomial coefficients of words
We have introduced a q-deformation, i.e., a polynomial in q with natural coefficients, of the binomial coefficient of two finite words u and v counting the number of occurrences of v as a subword of u. In this paper, we examine the q-deformation of Parikh matrices as introduced by Eğecioğlu in 2004.
Many classical results concerning Parikh matrices generalize to this new framework: Our first important observation is that the elements of such a matrix are in fact q-deformations of binomial coefficients of words. We also study their inverses and as an application, we obtain new identities about q-binomials.
For a finite word z and for the sequence $(p_n)_{n\ge 0}$ of prefixes of an infinite word, we show that the polynomial sequence $\binom{p_n}{z}_q$ converges to a formal series. We present links with additive number theory and k-regular sequences. In the case of a periodic word $u^ω$, we generalize a result of Salomaa: the sequence $\binom{u^n}{z}_q$ satisfies a linear recurrence relation with polynomial coefficients. Related to the theory of integer partition, we describe the growth and the zero set of the coefficients of the series associated with $u^ω$.
Finally, we show that the minors of a q-Parikh matrix are polynomials with natural coefficients and consider a generalization of Cauchy's inequality. We also compare q-Parikh matrices associated with an arbitrary word with those associated with a canonical word $12\cdots k$ made of pairwise distinct symbols.
Introducing q-deformed binomial coefficients of words
Gaussian binomial coefficients are q-analogues of the binomial coefficients of integers. On the other hand, binomial coefficients have been extended to finite words, i.e., elements of the finitely generated free monoids. In this paper we bring together these two notions by introducing q-analogues of binomial coefficients of words. We study their basic properties, e.g., by extending classical formulas such as the q-Vandermonde and Manvel's et al. identities to our setting. As a consequence, we get information about the structure of the considered words: these q-deformations of binomial coefficients of words contain much richer information than the original coefficients. From an algebraic perspective, we introduce a q-shuffle and a family q-infiltration products for non-commutative formal power series. Finally, we apply our results to generalize a theorem of Eilenberg characterizing so-called p-group languages. We show that a language is of this type if and only if it is a Boolean combination of specific languages defined through q-binomial coefficients seen as polynomials over $\mathbb{F}_p$.
On extended boundary sequences of morphic and Sturmian words
Published
• View Publication
• BIB
Generalizing the notion of the boundary sequence introduced by Chen and Wen, the $n$th term of the $\ell$-boundary sequence of an infinite word is the finite set of pairs $(u,v)$ of prefixes and suffixes of length $\ell$ appearing in factors $uyv$ of length $n+\ell$ ($n\ge \ell\ge 1$). Otherwise stated, for increasing values of $n$, one looks for all pairs of factors of length $\ell$ separated by $n-\ell$ symbols.
For the large class of addable abstract numeration systems $S$, we show that if an infinite word is $S$-automatic, then the same holds for its $\ell$-boundary sequence. In particular, they are both morphic (or generated by an HD0L system). To precise the limits of this result, we discuss examples of non-addable numeration systems and $S$-automatic words for which the boundary sequence is nevertheless $S$-automatic and conversely, $S$-automatic words with a boundary sequence that is not $S$-automatic. In the second part of the paper, we study the $\ell$-boundary sequence of a Sturmian word. We show that it is obtained through a sliding block code from the characteristic Sturmian word of the same slope. We also show that it is the image under a morphism of some other characteristic Sturmian word.
On digital sequences associated with Pascal's triangle
Published
• View Publication
• BIB
We consider the sequence of integers whose $n$th term has base-$p$ expansion given by the $n$th row of Pascal's triangle modulo $p$ (where $p$ is a prime number). We first present and generalize well-known relations concerning this sequence. Then, with the great help of Sloane's On-Line Encyclopedia of Integer Sequences, we show that it appears naturally as a subsequence of a $2$-regular sequence. Its study provides interesting relations and surprisingly involves odious and evil numbers, Nim-sum and even Gray codes. Furthermore, we examine similar sequences emerging from prime numbers involving alternating sum-of-digits modulo~$p$. This note ends with a discussion about Pascal's pyramid involving trinomial coefficients.
Characterizations of families of morphisms and words via binomial complexities
Published
• View Publication
• BIB
Two words are $k$-binomially equivalent if each subword of length at most $k$ occurs the same number of times in both words. The $k$-binomial complexity of an infinite word is a counting function that maps $n$ to the number of $k$-binomial equivalence classes represented by its factors of length $n$. Cassaigne et al. [Int. J. Found. Comput. S., 22(4) (2011)] characterized a family of morphisms, which we call Parikh-collinear, as those morphisms that map all words to words with bounded $1$-binomial complexity. Firstly, we extend this characterization: they map words with bounded $k$-binomial complexity to words with bounded $(k+1)$-binomial complexity. As a consequence, fixed points of Parikh-collinear morphisms are shown to have bounded $k$-binomial complexity for all $k$. Secondly, we give a new characterization of Sturmian words with respect to their $k$-binomial complexity. Then we characterize recurrent words having, for some $k$, the same $j$-binomial complexity as the Thue-Morse word for all $j\le k$. Finally, inspired by questions raised by Lejeune, we study the relationships between the $k$- and $(k+1)$-binomial complexities of infinite words; as well as the link with the usual factor complexity.
Revisiting regular sequences in light of rational base numeration systems
Published
• View Publication
• BIB
Regular sequences generalize the extensively studied automatic sequences. Let $S$ be an abstract numeration system. When the numeration language $L$ is prefix-closed and regular, a sequence is said to be $S$-regular if the module generated by its $S$-kernel is finitely generated.
In this paper, we give a new characterization of such sequences in terms of the underlying numeration tree $T(L)$ whose nodes are words of $L$. We may decorate these nodes by the sequence of interest following a breadth-first enumeration. For a prefix-closed regular language $L$, we prove that a sequence is $S$-regular if and only if the tree $T(L)$ decorated by the sequence is linear, i.e., the decoration of a node depends linearly on the decorations of a fixed number of ancestors.
Next, we introduce and study regular sequences in a rational base numeration system, whose numeration language is known to be highly non-regular. We motivate and comment our definition that a sequence is $\frac{p}{q}$-regular if the underlying numeration tree decorated by the sequence is linear. We give the first few properties of such sequences, we provide a few examples of them, and we propose a method for guessing $\frac{p}{q}$-regularity. Then we discuss the relationship between $\frac{p}{q}$-automatic sequences and $\frac{p}{q}$-regular sequences. We finally present a graph directed linear representation of a $\frac{p}{q}$-regular sequence. Our study permits us to highlight the places where the regularity of the numeration language plays a predominant role.
Automatic sequences: from rational bases to trees
Published in Discrete Mathematics & Theoretical Computer Science, vol. 24, no. 1, Automata, Logic and Semantics (July 19, 2022) dmtcs:8455
• View Publication
• BIB
The $n$th term of an automatic sequence is the output of a deterministic finite automaton fed with the representation of $n$ in a suitable numeration system. In this paper, instead of considering automatic sequences built on a numeration system with a regular numeration language, we consider those built on languages associated with trees having periodic labeled signatures and, in particular, rational base numeration systems. We obtain two main characterizations of these sequences. The first one is concerned with $r$-block substitutions where $r$ morphisms are applied periodically. In particular, we provide examples of such sequences that are not morphic. The second characterization involves the factors, or subtrees of finite height, of the tree associated with the numeration system and decorated by the terms of the sequence.
On the binomial equivalence classes of finite words
Published
• View Publication
• BIB
Two finite words $u$ and $v$ are $k$-binomially equivalent if, for each word $x$ of length at most $k$, $x$ appears the same number of times as a subsequence (i.e., as a scattered subword) of both $u$ and $v$. This notion generalizes abelian equivalence. In this paper, we study the equivalence classes induced by the $k$-binomial equivalence with a special focus on the cardinalities of the classes. We provide an algorithm generating the $2$-binomial equivalence class of a word. For $k \geq 2$ and alphabet of $3$ or more symbols, the language made of lexicographically least elements of every $k$-binomial equivalence class and the language of singletons, i.e., the words whose $k$-binomial equivalence class is restricted to a single element, are shown to be non context-free. As a consequence of our discussions, we also prove that the submonoid generated by the generators of the free nil-$2$ group on $m$ generators is isomorphic to the quotient of the free monoid $\{ 1, \ldots , m\}^{*}$ by the $2$-binomial equivalence.
Reconstructing Words from Right-Bounded-Block Words
Published
• View Publication
• BIB
A reconstruction problem of words from scattered factors asks for the minimal information, like multisets of scattered factors of a given length or the number of occurrences of scattered factors from a given set, necessary to uniquely determine a word. We show that a word $w \in \{a, b\}^{*}$ can be reconstructed from the number of occurrences of at most $\min(|w|_a, |w|_b)+ 1$ scattered factors of the form $a^{i} b$. Moreover, we generalize the result to alphabets of the form $\{1,\ldots,q\}$ by showing that at most $ \sum^{q-1}_{i=1} |w|_i (q-i+1)$ scattered factors suffices to reconstruct $w$. Both results improve on the upper bounds known so far. Complexity time bounds on reconstruction algorithms are also considered here.
The carry propagation of the successor function
Given any numeration system, we call carry propagation at a number $N$ the number of digits that are changed when going from the representation of $N$ to the one of $N+1$, and amortized carry propagation the limit of the mean of the carry propagations at the first $N$ integers, when $N$ tends to infinity, if this limit exists.
In the case of the usual base $p$ numeration system, it can be shown that the limit indeed exists and is equal to $p/(p-1)$. We recover a similar value for those numeration systems we consider and for which the limit exists.
We address the problem of the existence of the amortized carry propagation in non-standard numeration systems of various kinds: abstract numeration systems, rational base numeration systems, greedy numeration systems and beta-numeration. We tackle the problem by three different types of techniques: combinatorial, algebraic, and ergodic. For each kind of numeration systems that we consider, the relevant method allows for establishing sufficient conditions for the existence of the carry propagation and examples show that these conditions are close to being necessary conditions.
Taking-and-merging games as rewrite games
Published in Discrete Mathematics & Theoretical Computer Science, vol. 22 no. 4 (September 23, 2020) dmtcs:5200
• View Publication
• BIB
This work is a contribution to the study of rewrite games. Positions are finite words, and the possible moves are defined by a finite number of local rewriting rules. We introduce and investigate taking-and-merging games, that is, where each rule is of the form a^k->epsilon.
We give sufficient conditions for a game to be such that the losing positions (resp. the positions with a given Grundy value) form a regular language or a context-free language. We formulate several related open questions in parallel with the famous conjecture of Guy about the periodicity of the Grundy function of octal games.
Finally we show that more general rewrite games quickly lead to undecidable problems. Namely, it is undecidable whether there exists a winning position in a given regular language, even if we restrict to games where each move strictly reduces the length of the current position. We formulate several related open questions in parallel with the famous conjecture of Guy about the periodicity of the Grundy function of octal games.
Computing the $k$-binomial complexity of the Thue--Morse word
Two words are $k$-binomially equivalent whenever they share the same subwords, i.e., subsequences, of length at most $k$ with the same multiplicities. This is a refinement of both abelian equivalence and the Simon congruence. The $k$-binomial complexity of an infinite word $\mathbf{x}$ maps the integer $n$ to the number of classes in the quotient, by this $k$-binomial equivalence relation, of the set of factors of length $n$ occurring in $\mathbf{x}$. This complexity measure has not been investigated very much. In this paper, we characterize the $k$-binomial complexity of the Thue--Morse word. The result is striking, compared to more familiar complexity functions. Although the Thue--Morse word is aperiodic, its $k$-binomial complexity eventually takes only two values. In this paper, we first obtain general results about the number of occurrences of subwords appearing in iterates of the form $Ψ^\ell(w)$ for an arbitrary morphism $Ψ$. We also thoroughly describe the factors of the Thue--Morse word by introducing a relevant new equivalence relation.
Automatic sequences based on Parry or Bertrand numeration systems
Published
• View Publication
• BIB
We study the factor complexity and closure properties of automatic sequences based on Parry or Bertrand numeration systems. These automatic sequences can be viewed as generalizations of the more typical $k$-automatic sequences and Pisot-automatic sequences. We show that, like $k$-automatic sequences, Parry-automatic sequences have sublinear factor complexity while there exist Bertrand-automatic sequences with superlinear factor complexity. We prove that the set of Parry-automatic sequences with respect to a fixed Parry numeration system is not closed under taking images by uniform substitutions or periodic deletion of letters. These closure properties hold for $k$-automatic sequences and Pisot-automatic sequences, so our result shows that these properties are lost when generalizing to Parry numeration systems and beyond. Moreover, we show that a multidimensional sequence is $U$-automatic with respect to a positional numeration system $U$ with regular language of numeration if and only if its $U$-kernel is finite.
Counting Subwords Occurrences in Base-b Expansions
Published in Integers 18A (2018), no. A13, 32 pp
• Search Publication
We count the number of distinct (scattered) subwords occurring in the base-b expansion of the non-negative integers. More precisely, we consider the sequence $(S_b(n))_{n\ge 0}$ counting the number of positive entries on each row of a generalization of the Pascal triangle to binomial coefficients of base-$b$ expansions. By using a convenient tree structure, we provide recurrence relations for $(S_b(n))_{n\ge 0}$ leading to the $b$-regularity of the latter sequence. Then we deduce the asymptotics of the summatory function of the sequence $(S_b(n))_{n\ge 0}$.
Generalized Pascal triangle for binomial coefficients of words
Published in Adv. Appl. Math. 80 (2016) 24-27
• View Publication
• BIB
We introduce a generalization of Pascal triangle based on binomial coefficients of finite words. These coefficients count the number of times a word appears as a subsequence of another finite word. Similarly to the Sierpiński gasket that can be built as the limit set, for the Hausdorff distance, of a convergent sequence of normalized compact blocks extracted from Pascal triangle modulo $2$, we describe and study the first properties of the subset of $[0, 1] \times [0, 1]$ associated with this extended Pascal triangle modulo a prime $p$.
Behavior of digital sequences through exotic numeration systems
Published in Electron. J. Combin. 24 (2017), no. 1, Paper 1.44, 36 pp
• View Publication
• BIB
Many digital functions studied in the literature, e.g., the summatory function of the base-$k$ sum-of-digits function, have a behavior showing some periodic fluctuation. Such functions are usually studied using techniques from analytic number theory or linear algebra. In this paper we develop a method based on exotic numeration systems and we apply it on two examples motivated by the study of generalized Pascal triangles and binomial coefficients of words.