Papers by Manon Stipulanti
31 paper(s) by this author
· All BibTeX
Symbols frequencies in the Thue--Morse word in base $3/2$ and related conjectures
We study a binary Thue--Morse-type sequence arising from the base-$3/2$ expansion of integers, an archetypal automatic sequence in a rational base numeration system. Because the sequence is generated by a periodic iteration of morphisms rather than a single primitive substitution, classical Perron--Frobenius methods do not directly apply to determine symbol frequencies. We prove that both symbols ${\tt 0},{\tt 1}$ occur with frequency $1/2$ and we show uniform recurrence and symmetry properties of its set of factors. The proof reveals a structural bridge between combinatorics on words and harmonic analysis: the first difference sequence is shown to be Toeplitz, providing dynamical rigidity, while filtered frequencies naturally encode a dyadic structure that lifts to the compact group of $2$-adic integers. In this $2$-adic setting, desubstitution becomes a linear operator on Fourier coefficients, and a spectral contraction argument enforces uniqueness of limiting densities. Our results answer several conjectures of Dekking (on a sibling sequence) and illustrate how harmonic analysis on compact groups can be fruitfully combined with substitution dynamics.
Factor-balancedness, linear recurrence, and factor complexity
In the study of infinite words, various notions of balancedness provide quantitative measures for how regularly letters or factors occur, and they find applications in several areas of mathematics and theoretical computer science. In this paper, we study factor-balancedness and uniform factor-balancedness, making two main contributions. First, we establish general sufficient conditions for an infinite word to be (uniformly) factor-balanced, applicable in particular to any given linearly recurrent word. These conditions are formulated in terms of $\mathcal{S}$-adic representations and generalize results of Adamczewski on primitive substitutive words, which show that balancedness of length-2 factors already implies uniform factor-balancedness. As an application of our criteria, we characterize the Sturmian words and ternary Arnoux--Rauzy words that are uniformly factor-balanced as precisely those with bounded weak partial quotients. Our second main contribution is a study of the relationship between factor-balancedness and factor complexity. In particular, we analyze the non-primitive substitutive case and construct an example of a factor-balanced word with exponential factor complexity, thereby making progress on a question raised in 2025 by Arnoux, Berthé, Minervino, Steiner, and Thuswaldner on the relation between balancedness and discrete spectrum.
Effective Computation of Generalized Abelian Complexity for Pisot Type Substitutive Sequences
Generalized abelian equivalence compares words by their factors up to a certain bounded length. The associated complexity function counts the equivalence classes for factors of a given size of an infinite sequence. How practical is this notion? When can these equivalence relations and complexity functions be computed efficiently? We study the fixed points of substitution of Pisot type. Each of their $k$-abelian complexities is bounded and the Parikh vectors of their length-$n$ prefixes form synchronized sequences in the associated Dumont--Thomas numeration system. Therefore, the $k$-abelian complexity of Pisot substitution fixed points is automatic in the same numeration system. Two effective generic construction approaches are investigated using the \texttt{Walnut} theorem prover and are applied to several examples. We obtain new properties of the Tribonacci sequence, such as a uniform bound for its factor balancedness together with a two-dimensional linear representation of its generalized abelian complexity functions.
Positionality of Dumont--Thomas numeration systems for integers
Introduced in 2001 by Lecomte and Rigo, abstract numeration systems provide a way of expressing natural numbers with words from a language $L$ accepted by a finite automaton. As it turns out, these numeration systems are not necessarily positional, i.e., we cannot always find a sequence $U=(U_i)_{i\ge 0}$ of integers such that the value of every word in the language $L$ is determined by the position of its letters and the first few values of $U$. Finding the conditions under which an abstract numeration system is positional seems difficult in general. In this paper, we thus consider this question for a particular sub-family of abstract numeration systems called Dumont--Thomas numeration systems. They are derived from substitutions and were introduced in 1989 by Dumont and Thomas. We exhibit conditions on the underlying substitution so that the corresponding Dumont--Thomas numeration is positional. We first work in the most general setting, then particularize our results to some practical cases. Finally, we link our numeration systems to existing literature, notably properties studied by Rényi in 1957, Parry in 1960, Bertrand-Mathis in 1989, and Fabre in 1995.
Additive word complexity and Walnut
In combinatorics on words, a classical topic of study is the number of specific patterns appearing in infinite sequences. For instance, many works have been dedicated to studying the so-called factor complexity of infinite sequences, which gives the number of different factors (contiguous subblocks of their symbols), as well as abelian complexity, which counts factors up to a permutation of letters. In this paper, we consider the relatively unexplored concept of additive complexity, which counts the number of factors up to additive equivalence. We say that two words are additively equivalent if they have the same length and the total weight of their letters is equal. Our contribution is to expand the general knowledge of additive complexity from a theoretical point of view and consider various famous examples. We show a particular case of an analog of the long-standing conjecture on the regularity of the abelian complexity of an automatic sequence. In particular, we use the formalism of logic, and the software Walnut, to decide related properties of automatic sequences. We compare the behaviors of additive and abelian complexities, and we also consider the notion of abelian and additive powers. Along the way, we present some open questions and conjectures for future work.
On the pseudorandomness of Parry--Bertrand automatic sequences
The correlation measure is a testimony of the pseudorandomness of a sequence $\infw{s}$ and provides information about the independence of some parts of $\infw{s}$ and their shifts. Combined with the well-distribution measure, a sequence possesses good pseudorandomness properties if both measures are relatively small. In combinatorics on words, the famous $b$-automatic sequences are quite far from being pseudorandom, as they have small factor complexity on the one hand and large well-distribution and correlation measures on the other. This paper investigates the pseudorandomness of a specific family of morphic sequences, including classical $b$-automatic sequences. In particular, we show that such sequences have large even-order correlation measures; hence, they are not pseudorandom. We also show that even- and odd-order correlation measures behave differently when considering some simple morphic sequences.
The reflection complexity of sequences over finite alphabets
In combinatorics on words, the well-studied factor complexity function $ρ_{\infw{x}}$ of a sequence $\infw{x}$ over a finite
alphabet counts, for every nonnegative integer $n$, the number of distinct length-$n$ factors of $\infw{x}$. In this paper, we
introduce the \emph{reflection complexity} function $r_{\infw{x}}$ to enumerate the factors occurring in a sequence $\infw{x}$, up
to reversing the order of symbols in a word. We prove a number of results about the growth properties of $r_{\infw{x}}$
and its relationship with other complexity functions. We also prove a Morse--Hedlund-type result characterizing eventually periodic
sequences in terms of their reflection complexity, and we deduce a characterization of Sturmian sequences. We investigate
the reflection complexity of quasi-Sturmian, episturmian, $(s+1)$-dimensional billiard, complementation-symmetric Rote, and rich
sequences. Furthermore, we prove that if $\infw{x}$ is $k$-automatic, then $r_{\infw{x}}$ is computably $k$-regular, and we use the
software \texttt{Walnut} to evaluate the reflection complexity of some automatic sequences, such as the Thue--Morse sequence. We
note that there are still many unanswered questions about this reflection measure.
Automatic Abelian Complexities of Parikh-Collinear Fixed Points
Parikh-collinear morphisms have the property that all the Parikh vectors of the images of letters are collinear, i.e., the associated adjacency matrix has rank 1. In the conference DLT-WORDS 2023 we showed that fixed points of Parikh-collinear morphisms are automatic. We also showed that the abelian complexity function of a binary fixed point of such a morphism is automatic under some assumptions. In this note, we fully generalize the latter result. Namely, we show that the abelian complexity function of a fixed point of an arbitrary, possibly erasing, Parikh-collinear morphism is automatic. Furthermore, a deterministic finite automaton with output generating this abelian complexity function is provided by an effective procedure. To that end, we discuss the constant of recognizability of a morphism and the related cutting set.
Exploring the Crochemore and Ziv-Lempel factorizations of some automatic sequences with the software Walnut
We explore the Ziv-Lempel and Crochemore factorizations of some classical automatic sequences making an extensive use of the theorem prover Walnut.
Combinatorics on words and generating Dirichlet series of automatic sequences
Generating series are crucial in enumerative combinatorics, analytic combinatorics, and combinatorics on words. Though it might seem at first view that generating Dirichlet series are less used in these fields than ordinary and exponential generating series, there are many notable papers where they play a fundamental role, as can be seen in particular in the work of Flajolet and several of his co-authors. In this paper, we study Dirichlet series of integers with missing digits or blocks of digits in some integer base $b$; i.e., where the summation ranges over the integers whose expansions form some language strictly included in the set of all words over the alphabet $\{0, 1, \dots, b-1\}$ that do not begin with a $0$. We show how to unify and extend results proved by Nathanson in 2021 and by Köhler and Spilker in 2009. En route, we encounter several sequences from Sloane's On-Line Encyclopedia of Integer Sequences, as well as some famous $b$-automatic sequences or $b$-regular sequences. We also consider a specific sequence that is not $b$-regular.
Summing the sum of digits
Published in Communications in Mathematics, Volume 33 (2025), Issue 2 (Special issue: Numeration, Liège 2023, dedicated to the 75th birthday of professor Christiane Frougny) (May 9, 2024) cm:12610
• View Publication
• BIB
We revisit and generalize inequalities for the summatory function of the sum of digits in a given integer base. We prove that several known results can be deduced from a theorem in a 2023 paper by Mohanty, Greenbury, Sarkany, Narayanan, Dingle, Ahnert, and Louis, whose primary scope is the maximum mutational robustness in genotype-phenotype maps.
String attractors of some simple-Parry automatic sequences
Firstly studied by Kempa and Prezza in 2018 as the cement of text compression algorithms, string attractors have become a compelling object of theoretical research within the community of combinatorics on words. In this context, they have been studied for several families of finite and infinite words. In this paper, we obtain string attractors of prefixes of particular infinite words generalizing k-bonacci words (including the famous Fibonacci word) and related to simple Parry numbers. In fact, our description involves the numeration systems classically derived from the considered morphisms. This extends our previous work published in the international conference WORDS 2023.
On extended boundary sequences of morphic and Sturmian words
Published
• View Publication
• BIB
Generalizing the notion of the boundary sequence introduced by Chen and Wen, the $n$th term of the $\ell$-boundary sequence of an infinite word is the finite set of pairs $(u,v)$ of prefixes and suffixes of length $\ell$ appearing in factors $uyv$ of length $n+\ell$ ($n\ge \ell\ge 1$). Otherwise stated, for increasing values of $n$, one looks for all pairs of factors of length $\ell$ separated by $n-\ell$ symbols.
For the large class of addable abstract numeration systems $S$, we show that if an infinite word is $S$-automatic, then the same holds for its $\ell$-boundary sequence. In particular, they are both morphic (or generated by an HD0L system). To precise the limits of this result, we discuss examples of non-addable numeration systems and $S$-automatic words for which the boundary sequence is nevertheless $S$-automatic and conversely, $S$-automatic words with a boundary sequence that is not $S$-automatic. In the second part of the paper, we study the $\ell$-boundary sequence of a Sturmian word. We show that it is obtained through a sliding block code from the characteristic Sturmian word of the same slope. We also show that it is the image under a morphism of some other characteristic Sturmian word.
A full characterization of Bertrand numeration systems
Published
• View Publication
• BIB
Among all positional numeration systems, the widely studied Bertrand numeration systems are defined by a simple criterion in terms of their numeration languages. In 1989, Bertrand-Mathis characterized them via representations in a real base $β$. However, the given condition turns to be not necessary. Hence, the goal of this paper is to provide a correction of Bertrand-Mathis' result. The main difference arises when $β$ is a Parry number, in which case are derived two associated Bertrand numeration systems. Along the way, we define a non-canonical $β$-shift and study its properties analogously to those of the usual canonical one.
On digital sequences associated with Pascal's triangle
Published
• View Publication
• BIB
We consider the sequence of integers whose $n$th term has base-$p$ expansion given by the $n$th row of Pascal's triangle modulo $p$ (where $p$ is a prime number). We first present and generalize well-known relations concerning this sequence. Then, with the great help of Sloane's On-Line Encyclopedia of Integer Sequences, we show that it appears naturally as a subsequence of a $2$-regular sequence. Its study provides interesting relations and surprisingly involves odious and evil numbers, Nim-sum and even Gray codes. Furthermore, we examine similar sequences emerging from prime numbers involving alternating sum-of-digits modulo~$p$. This note ends with a discussion about Pascal's pyramid involving trinomial coefficients.
Characterizations of families of morphisms and words via binomial complexities
Published
• View Publication
• BIB
Two words are $k$-binomially equivalent if each subword of length at most $k$ occurs the same number of times in both words. The $k$-binomial complexity of an infinite word is a counting function that maps $n$ to the number of $k$-binomial equivalence classes represented by its factors of length $n$. Cassaigne et al. [Int. J. Found. Comput. S., 22(4) (2011)] characterized a family of morphisms, which we call Parikh-collinear, as those morphisms that map all words to words with bounded $1$-binomial complexity. Firstly, we extend this characterization: they map words with bounded $k$-binomial complexity to words with bounded $(k+1)$-binomial complexity. As a consequence, fixed points of Parikh-collinear morphisms are shown to have bounded $k$-binomial complexity for all $k$. Secondly, we give a new characterization of Sturmian words with respect to their $k$-binomial complexity. Then we characterize recurrent words having, for some $k$, the same $j$-binomial complexity as the Thue-Morse word for all $j\le k$. Finally, inspired by questions raised by Lejeune, we study the relationships between the $k$- and $(k+1)$-binomial complexities of infinite words; as well as the link with the usual factor complexity.
Closed Ziv-Lempel factorization of the $m$-bonacci words
Published
• View Publication
• BIB
A word $w$ is said to be closed if it has a proper factor $x$ which occurs exactly twice in $w$, as a prefix and as a suffix of $w$. Based on the concept of Ziv-Lempel factorization, we define the closed $z$-factorization of finite and infinite words. Then we find the closed $z$-factorization of the infinite $m$-bonacci words for all $m \geq 2$. We also classify closed prefixes of the infinite $m$-bonacci words.
Revisiting regular sequences in light of rational base numeration systems
Published
• View Publication
• BIB
Regular sequences generalize the extensively studied automatic sequences. Let $S$ be an abstract numeration system. When the numeration language $L$ is prefix-closed and regular, a sequence is said to be $S$-regular if the module generated by its $S$-kernel is finitely generated.
In this paper, we give a new characterization of such sequences in terms of the underlying numeration tree $T(L)$ whose nodes are words of $L$. We may decorate these nodes by the sequence of interest following a breadth-first enumeration. For a prefix-closed regular language $L$, we prove that a sequence is $S$-regular if and only if the tree $T(L)$ decorated by the sequence is linear, i.e., the decoration of a node depends linearly on the decorations of a fixed number of ancestors.
Next, we introduce and study regular sequences in a rational base numeration system, whose numeration language is known to be highly non-regular. We motivate and comment our definition that a sequence is $\frac{p}{q}$-regular if the underlying numeration tree decorated by the sequence is linear. We give the first few properties of such sequences, we provide a few examples of them, and we propose a method for guessing $\frac{p}{q}$-regularity. Then we discuss the relationship between $\frac{p}{q}$-automatic sequences and $\frac{p}{q}$-regular sequences. We finally present a graph directed linear representation of a $\frac{p}{q}$-regular sequence. Our study permits us to highlight the places where the regularity of the numeration language plays a predominant role.
Automatic sequences: from rational bases to trees
Published in Discrete Mathematics & Theoretical Computer Science, vol. 24, no. 1, Automata, Logic and Semantics (July 19, 2022) dmtcs:8455
• View Publication
• BIB
The $n$th term of an automatic sequence is the output of a deterministic finite automaton fed with the representation of $n$ in a suitable numeration system. In this paper, instead of considering automatic sequences built on a numeration system with a regular numeration language, we consider those built on languages associated with trees having periodic labeled signatures and, in particular, rational base numeration systems. We obtain two main characterizations of these sequences. The first one is concerned with $r$-block substitutions where $r$ morphisms are applied periodically. In particular, we provide examples of such sequences that are not morphic. The second characterization involves the factors, or subtrees of finite height, of the tree associated with the numeration system and decorated by the terms of the sequence.
Regular sequences and synchronized sequences in abstract numeration systems
Published
• View Publication
• BIB
The notion of $b$-regular sequences was generalized to abstract numeration systems by Maes and Rigo in 2002. Their definition is based on a notion of $\mathcal{S}$-kernel that extends that of $b$-kernel. However, this definition does not allow us to generalize all of the many characterizations of $b$-regular sequences. In this paper, we present an alternative definition of $\mathcal{S}$-kernel, and hence an alternative definition of $\mathcal{S}$-regular sequences, which enables us to use recognizable formal series in order to generalize most (if not all) known characterizations of $b$-regular sequences to abstract numeration systems. We then give two characterizations of $\mathcal{S}$-automatic sequences as particular $\mathcal{S}$-regular sequences. Next, we present a general method to obtain various families of $\mathcal{S}$-regular sequences by enumerating $\mathcal{S}$-recognizable properties of $\mathcal{S}$-automatic sequences. As an example of the many possible applications of this method, we show that, provided that addition is $\mathcal{S}$-recognizable, the factor complexity of an $\mathcal{S}$-automatic sequence defines an $\mathcal{S}$-regular sequence. In the last part of the paper, we study $\mathcal{S}$-synchronized sequences. Along the way, we prove that the formal series obtained as the composition of a synchronized relation and a recognizable series is recognizable. As a consequence, the composition of an $\mathcal{S}$-synchronized sequence and a $\mathcal{S}$-regular sequence is shown to be $\mathcal{S}$-regular. All our results are presented in an arbitrary dimension $d$ and for an arbitrary semiring $\mathbb{K}$.