rooted tree
363 papers tagged with this keyword
Lengths of paths in rooted trees
We provide formulas for generating functions of many types of paths in various rooted tree structures. We compute the $k$th moment of the generating functions for various types of vertical paths. In two specific familes of trees we find exact closed formulas for expectations and their asymptotic values. Some of these closed formulas are surprisingly simple.
The topological trees with extreme Matula numbers
Denote by $p_m$ the $m$-th prime number ($p_1=2,~p_2=3,~p_3=5,~ p_4=7,~\ldots$). Let $T$ be a rooted tree with branches $T_1,T_2,\ldots,T_r$. The Matula number $M(T)$ of $T$ is $p_{M(T_1)}\cdot p_{M(T_2)}\cdot \ldots \cdot p_{M(T_r)}$, starting with $M(K_1)=1$. This number was put forward half a century ago by the American mathematician David Matula. In this paper, we prove that the star (consisting of a root and leaves attached to it) and the binary caterpillar (a binary tree whose internal vertices form a path starting at the root) have the smallest and greatest Matula number, respectively, over all topological trees (rooted trees without vertices of outdegree $1$) with a prescribed number of leaves -- the extreme values are also derived.
Signature Catalan Combinatorics
Published
• View Publication
• BIB
The Catalan numbers constitute one of the most important sequences in combinatorics. Catalan objects have been generalized in various directions, including the classical Fuss-Catalan objects and the rational Catalan generalization of Armstrong-Rhoades-Williams. We propose a wider generalization of these families indexed by a composition $s$ which is motivated by the combinatorics of planar rooted trees; when $s=(2,...,2)$ and $s=(k+1,...,k+1)$ we recover the classical Catalan and Fuss-Catalan combinatorics, respectively. Furthermore, to each pair $(a,b)$ of relatively prime numbers we can associate a signature that recovers the combinatorics of rational Catalan objects. We present explicit bijections between the resulting $s$-Catalan objects, and a fundamental recurrence that generalizes the fundamental recurrence of the classical Catalan numbers. Our framework allows us to define signature generalizations of parking functions which coincide with the generalized parking functions studied by Pitman-Stanley and Yan, as well as generalizations of permutations which coincide with the notion of Stirling multipermutations introduced by Gessel-Stanley. Some of our constructions differ from the ones of Armstrong-Rhoades-Williams, however as a byproduct of our extension, we obtain the additional notions of rational permutations and rational trees.
The Cavender-Farris-Neyman Model with a Molecular Clock
Published
• View Publication
• BIB
We give a combinatorial description of the toric ideal of invariants of the Cavender-Farris-Neyman model with a molecular clock (CFN-MC) on a rooted binary phylogenetic tree and prove results about the polytope associated to this toric ideal. Key results about the polyhedral structure include that the number of vertices of this polytope is a Fibonacci number, the facets of the polytope can be described using the combinatorial "cluster" structure of the underlying rooted tree, and the volume is equal to an Euler zig-zag number. The toric ideal of invariants of the CFN-MC model has a quadratic Groebner basis with squarefree initial terms. Finally, we show that the Ehrhart polynomial of these polytopes, and therefore the Hilbert series of the ideals, depends only on the number of leaves of the underlying binary tree, and not on the topology of the tree itself. These results are analogous to classic results for the Cavender-Farris-Neyman model without a molecular clock. However, new techniques are required because the molecular clock assumption destroys the toric fiber product structure that governs group-based models without the molecular clock.
Prime Parking Functions on Rooted Trees
Published
• View Publication
• BIB
For a labeled, rooted tree with edges oriented towards the root, we consider the vertices as parking spots and the edge orientation as a one-way street. Each driver, starting with her preferred parking spot, searches for and parks in the first unoccupied spot along the directed path to the root. If all $n$ drivers park, the sequence of spot preferences is called a parking function. We consider the sequences, called \emph{prime} parking functions, for which each driver parks and each edge in the tree is traversed by some driver after failing to park at her preferred spot. We prove that the total number of prime parking functions on trees with $n$ vertices is $(2n-2)!$. Additionally, we generalize \emph{increasing} parking functions, those in which the drivers park with a weakly-increasing order of preference, to trees and prove that the total number of increasing prime parking functions on trees with $n$ vertices is $(n-1)!S_{n-1}$, where $\{S_i\}_{i \geq 0}$ are the large Schröder numbers.
k-Ary spanning trees contained in tournaments
A rooted tree is called a $k$-ary tree, if all non-leaf vertices have exactly $k$ children, except possibly one non-leaf vertex has at most $k-1$ children. Denote by $h(k)$ the minimum integer such that every tournament of order at least $h(k)$ contains a $k$-ary spanning tree. It is well-known that every tournament contains a Hamiltonian path, which implies that $h(1)=1$. Lu et al. [J. Graph Theory {\bf 30}(1999) 167--176] proved the existence of $h(k)$, and showed that $h(2)=4$ and $h(3)=8$. The exact values of $h(k)$ remain unknown for $k\geq 4$. A result of Erdős on the domination number of tournaments implies $h(k)=Ω(k\log k)$. In this paper, we prove that $h(4)=10$ and $h(5)\geq13$.
A polynomial associated with rooted trees and specific posets
We investigate a trivariate polynomial associated with rooted trees. It generalises a bivariate polynomial for rooted trees that was recently introduced by Liu. We show that this polynomial satisfies a deletion-contraction recursion and can be expressed as a sum over maximal antichains. Several combinatorial quantities can be obtained as special values, in particular the number of antichains, maximal antichains and cutsets. We prove that two of the three possible bivariate specialisations characterise trees uniquely up to isomorphism. One of these has already been established by Liu, the other is new. For the third specialisation, we construct non-isomorphic trees with the same associated polynomial.
We finally find that our polynomial can be generalised in a natural way to a family of posets that we call $\mathcal{V}$-posets. These posets are obtained recursively by either disjoint unions or adding a greatest/least element to existing $\mathcal{V}$-posets.
The complexity of comparing multiply-labelled trees by extending phylogenetic-tree metrics
Published
• View Publication
• BIB
A multilabeled tree (or MUL-tree) is a rooted tree in which every leaf is labelled by an element from some set, but in which more than one leaf may be labelled by the same element of that set. In phylogenetics, such trees are used in biogeographical studies, to study the evolution of gene families, and also within approaches to construct phylogenetic networks. A multilabelled tree in which no leaf-labels are repeated is called a phylogenetic tree, and one in which every label is the same is also known as a tree-shape. In this paper, we consider the complexity of computing metrics on MUL-trees that are obtained by extending metrics on phylogenetic trees. In particular, by restricting our attention to tree shapes, we show that computing the metric extension on MUL-trees is NP complete for two well-known metrics on phylogenetic trees, namely, the path-difference and Robinson Foulds distances. We also show that the extension of the Robinson Foulds distance is fixed parameter tractable with respect to the distance parameter. The path distance complexity result allows us to also answer an open problem concerning the complexity of solving the quadratic assignment problem for two matrices that are a Robinson similarity and a Robinson dissimilarity, which we show to be NP-complete. We conclude by considering the maximum agreement subtree (MAST) distance on phylogenetic trees to MUL-trees. Although its extension to MUL-trees can be computed in polynomial time, we show that computing its natural generalization to more than two MUL-trees is NP-complete, although fixed-parameter tractable in the maximum degree when the number of given trees is bounded.
Limiting probabilities for vertices of a given rank in rooted trees
We consider two varieties of labeled rooted trees, and the probability that a vertex chosen from all vertices of all trees of a given size uniformly at random has a given rank. We prove that this probability converges to a limit as the tree size goes to infinity.
Inducibility of d-ary trees
Imitating a recently introduced invariant of trees, we initiate the study of the inducibility of $d$-ary trees (rooted trees whose vertex outdegrees are bounded from above by $d\geq 2$) with a given number of leaves. We determine the exact inducibility for stars and binary caterpillars. For $T$ in the family of strictly $d$-ary trees (every vertex has $0$ or $d$ children), we prove that the difference between the maximum density of a $d$-ary tree $D$ in $T$ and the inducibility of $D$ is of order $\mathcal{O}(|T|^{-1/2})$ compared to the general case where it is shown that the difference is $\mathcal{O}(|T|^{-1})$ which, in particular, responds positively to an existing conjecture on the inducibility in binary trees. We also discover that the inducibility of a binary tree in $d$-ary trees is independent of $d$. Furthermore, we establish a general lower bound on the inducibility and also provide a bound for some special trees. Moreover, we find that the maximum inducibility is attained for binary caterpillars for every $d$.
Generalized Fitch Graphs: Edge-labeled Graphs that are explained by Edge-labeled Trees
Published
• View Publication
• BIB
Fitch graphs $G=(X,E)$ are di-graphs that are explained by $\{\otimes,1\}$-edge-labeled rooted trees with leaf set $X$: there is an arc $xy\in E$ if and only if the unique path in $T$ that connects the least common ancestor $\textrm{lca}(x,y)$ of $x$ and $y$ with $y$ contains at least one edge with label $1$. In practice, Fitch graphs represent xenology relations, i.e., pairs of genes $x$ and $y$ for which a horizontal gene transfer happened along the path from $\textrm{lca}(x,y)$ to $y$. In this contribution, we generalize the concept of xenology and Fitch graphs and consider complete di-graphs $K_{|X|}$ with vertex set $X$ and a map $ε$ that assigns to each arc $xy$ a unique label $ε(x,y)\in M\cup \{\otimes\}$, where $M$ denotes an arbitrary set of symbols. A di-graph $(K_{|X|},ε)$ is a generalized Fitch graph if there is an $M\cup \{\otimes\}$-edge-labeled tree $(T,λ)$ that can explain $(K_{|X|},ε)$. We provide a simple characterization of generalized Fitch graphs $(K_{|X|},ε)$ and give an $O(|X|^2)$-time algorithm for their recognition as well as for the reconstruction of the unique least resolved phylogenetic tree that explains $(K_{|X|},ε)$.
Extremal values of the Sackin tree balance index
Published
• View Publication
• BIB
Tree balance plays an important role in different research areas like theoretical computer science and mathematical phylogenetics. For example, it has long been known that under the Yule model, a pure birth process, imbalanced trees are more likely than balanced ones. Also, concerning ordered search trees, more balanced ones allow for more efficient data structuring than imbalanced ones. Therefore, different methods to measure the balance of trees were introduced. The Sackin index is one of the most frequently used measures for this purpose. In many contexts, statements about the minimal and maximal values of this index have been discussed, but formal proofs have only been provided for some of them, and only in the context of ordered binary (search) trees, not for general rooted trees. Moreover, while the number of trees with maximal Sackin index as well as the number of trees with minimal Sackin index when the number of leaves is a power of 2 are relatively easy to understand, the number of trees with minimal Sackin index for all other numbers of leaves has been completely unknown. In this manuscript, we extend the findings on trees with minimal and maximal Sackin indices from the literature on ordered trees and subsequently use our results to provide formulas to explicitly calculate the numbers of such trees. We also extend previous studies by analyzing the case when the underlying trees need not be binary. Finally, we use our results to contribute both to the phylogenetic as well as the computer scientific literature by using the new findings on Sackin minimal and maximal trees in order to derive formulas to calculate the number of both minimal and maximal phylogenetic trees as well as minimal and maximal ordered trees both in the binary and non-binary settings. All our results have been implemented in the Mathematica package SackinMinimizer, which has been made publicly available.
Rooted tree maps and the Kawashima relations for multiple zeta values
Published
• View Publication
• BIB
Recently, inspired by the Connes-Kreimer Hopf algebra of rooted trees, the second named author introduced rooted tree maps as a family of linear maps on the noncommutative polynomial algebra in two letters. These give a class of relations among multiple zeta values, which are known to be a subclass of the so-called linear part of the Kawashima relations. In this paper we show the opposite implication, that is the linear part of the Kawashima relations is implied by the relations coming from rooted tree maps.
Graphs with prescribed local neighborhoods of their universal coverings
Published
• View Publication
• BIB
Given a collection of n rooted trees with depth h, we give a necessary and sufficient condition for this collection to be the collection of h-depth universal covering neighborhoods at each vertex.
Asymptotic results on Hoppe trees and its variations
Published
• View Publication
• BIB
A uniform recursive tree on $n$ vertices is a random tree where each possible $(n-1)!$ labeled recursive rooted tree is selected with equal probability. In this paper we introduce and study weighted trees, a non-uniform recursive tree model departing from the recently introduced Hoppe trees. This class generalizes both uniform recursive trees and Hoppe trees. The generalization provides diversity among the nodes, making the model more flexible for applications. We also analyze the number of leaves, the height, the depth, the number of branches, and the size of the largest branch in these weighted trees.
Rooted tree maps and the derivation relation for multiple zeta values
Published
• View Publication
• BIB
Rooted tree maps assign to an element of the Connes-Kreimer Hopf algebra of rooted trees a linear map on the noncommutative polynomial algebra in two letters. Evaluated at any admissible word these maps induce linear relations between multiple zeta values. In this note we show that the derivation relations for multiple zeta values are contained in this class of linear relations.
Rooted Tree Maps
Published
• View Publication
• BIB
Based on Hopf algebra of rooted trees introduced by Connes and Kreimer, we construct a class of linear maps on noncommutative polynomial algebra in two indeterminates, namely rooted tree maps. We also prove that their maps induce a class of relations among multiple zeta values.
The local limit of the uniform spanning tree on dense graphs
Published in Journal of Statistical Physics 173 (2018), no. 3-4, 502-545
• View Publication
• BIB
Let $G$ be a connected graph in which almost all vertices have linear degrees and let $T$ be a uniform spanning tree of $G$. For any fixed rooted tree $F$ of height $r$ we compute the asymptotic density of vertices $v$ for which the $r$-ball around $v$ in $T$ is isomorphic to $F$. We deduce from this that if $\{G_n\}$ is a sequence of such graphs converging to a graphon $W$, then the uniform spanning tree of $G_n$ locally converges to a multi-type branching process defined in terms of $W$.
As an application, we prove that in a graph with linear minimum degree, with high probability, the density of leaves in a uniform spanning tree is at least $1/e-o(1)$, the density of vertices of degree $2$ is at most $1/e+o(1)$ and the density of vertices of degree $k\geq 3$ is at most ${(k-2)^{k-2} \over (k-1)! e^{k-2}} + o(1)$. These bounds are sharp.
Constructing Directed Cayley Graphs of Small Diameter: A Potent Solovay-Kitaev Procedure
Published
• View Publication
• BIB
Let $Γ$ be a group and $(Γ_n)_{n=1} ^{\infty}$ be a descending sequence of finite-index normal subgroups. We establish explicit upper bounds on the diameters of the directed Cayley graphs of the $Γ/Γ_n$ , under some natural hypotheses on the behaviour of power and commutator words in $Γ$. The bounds we obtain do not depend on a choice of generating set. Moreover under reasonable conditions our method provides a fast algorithm for constructing directed Cayley graphs of diameter satisfying our bounds. The proof is closely analogous to the the Solovay-Kitaev procedure, which only uses commutator words, but also only constructs small-diameter undirected Cayley graphs. As an application we give directed diameter bounds on finite quotients of two very different groups: $SL_2 (\mathbb{F}_q [[t]])$ (for $q$ even) and a group of automorphisms of the ternary rooted tree introduced by Fabrykowski and Gupta.
An infinite class of unsaturated rooted trees corresponding to designable RNA secondary structures
Published
• View Publication
• BIB
An RNA secondary structure is designable if there is an RNA sequence which can attain its maximum number of base pairs only by adopting that structure. The combinatorial RNA design problem, introduced by Haleš et al. in 2016, is to determine whether or not a given RNA secondary structure is designable. Haleš et al. identified certain classes of designable and non-designable secondary structures by reference to their corresponding rooted trees. We introduce an infinite class of rooted trees containing unpaired nucleotides at the greatest depth, and prove constructively that their corresponding secondary structures are designable. This complements previous results for the combinatorial RNA design problem.