labeled tree
89 papers tagged with this keyword
The probability of random trees being isomorphic
We study the fundamental question of how likely it is that two randomly chosen trees are isomorphic to each other for different models of random trees. We show that the probability decays exponentially for rooted labeled trees as well as for Galton--Watson trees with bounded degrees but that this is not true for plane trees, thus providing a counterexample in the general case of Galton--Watson trees without degree restrictions. We also derive limiting distributions for some related parameters: the number of vertices of given degrees in pairs of labeled trees conditioned on being isomorphic, thus showing that they have a different shape than usual labeled trees, as well as the number of labelings and plane representations of Pólya trees. The results rely on both probabilistic and analytic tools.
Automated Counting and Statistical Analysis of Labeled Trees with Degree Restrictions
Arthur Cayley famously proved that there are n to the power n-2 labeled trees on n vertices. Here we go much further and show how to enumerate, fully automatically, labeled trees such that every vertex has a number of neighbors that belongs to a specified finite set, and also count trees where the number of neighbors is not allowed to be in a given finite set. We also give detailed statistical analysis, and show that in the sample space of labeled trees with n vertices, the random variable "number of vertices with d neighbors" is asymptotically normal, and for any different degrees, are jointly asymptotically normal, but of course, not independently so (except for the pair (1,3), i.e. the number of leaves and the number of degree-3 vertices, where there are asymptotically independent). We also give new proofs to Amram Meir and John Noon's expressions for the limiting expectation and variance for these, and derive an explicit expression for the covariance.
From Modular Decomposition Trees to Level-1 Networks: Pseudo-Cographs, Polar-Cats and Prime Polar-Cats
Published
• View Publication
• BIB
The modular decomposition of a graph $G$ is a natural construction to capture key features of $G$ in terms of a labeled tree $(T,t)$ whose vertices are labeled as "series" ($1$), "parallel" ($0$) or "prime". However, full information of $G$ is provided by its modular decomposition tree $(T,t)$ only, if $G$ does not contain prime modules. In this case, $(T,t)$ explains $G$, i.e., $\{x,y\}\in E(G)$ if and only if the lowest common ancestor $\mathrm{lca}_T(x,y)$ of $x$ and $y$ has label "$1$". This information, however, gets lost whenever $(T,t)$ contains vertices with label "prime". In this contribution, we aim at replacing "prime" vertices in $(T,t)$ by simple 0/1-labeled cycles, which leads to the concept of rooted labeled level-1 networks $(N,t)$.
We characterize graphs that can be explained by such level-1 networks $(N,t)$, which generalizes the concept of graphs that can be explained by labeled trees, that is, cographs. We provide three novel graph classes: \emph{polar-cats} are a proper subclass of \emph{pseudo-cographs} which forms a proper subclass of \emph{prime polar-cats}. In particular, every cograph is a pseudo-cograph and prime polar-cats are precisely those graphs that can be explained by a labeled level-1 network. The class of prime polar-cats is defined in terms of the modular decomposition of graphs and the property that all prime modules "induce" polar-cats. We provide a plethora of structural results and characterizations for graphs of these new classes. In addition, we show under which conditions there is a unique least-resolved labeled level-1 network that explains a given graph and provide linear-time algorithms to recognize all these types of graphs and to construct level-1 networks to explain them.
Persistence for a class of order-one autoregressive processes and Mallows-Riordan polynomials
Published
• View Publication
• BIB
We establish exact formulae for the persistence probabilities of an AR(1) sequence with symmetric uniform innovations in terms of certain families of polynomials, most notably a family introduced by Mallows and Riordan as enumerators of finite labeled trees when ordered by inversions. The connection of these polynomials with the volumes of certain polytopes is also discussed. Two further results provide factorizations of general AR(1) models, one for negative drifts with continuous innovations, and one for positive drifts with continuous and symmetric innovations. The second factorization extends a classical universal formula of Sparre Andersen for symmetric random walks. Our results also lead to explicit asymptotic estimates for the persistence probabilities.
A branch statistic for trees: Interpreting coefficients of the characteristic polynomial of braid deformations
Published in Enumerative Combinatorics and Applications 3:1 (2023) Article S2R5
• View Publication
• BIB
A hyperplane arrangement in $\mathbb{R}^n$ is a finite collection of affine hyperplanes. The regions are the connected components of the complement of these hyperplanes. By a theorem of Zaslavsky, the number of regions of a hyperplane arrangement is the sum of coefficients of its characteristic polynomial. Arrangements that contain hyperplanes parallel to subspaces whose defining equations are $x_i - x_j = 0$ form an important class called the deformations of the braid arrangement. In a recent work, Bernardi showed that regions of certain deformations are in one-to-one correspondence with certain labeled trees. In this article, we define a statistic on these trees such that the distribution is given by the coefficients of the characteristic polynomial. In particular, our statistic applies to well-studied families like extended Catalan, Shi, Linial and semiorder.
Rooted quasi-Stirling permutations of general multisets
Published
• View Publication
• BIB
Given a general multiset $\mathcal{M}=\{1^{m_1},2^{m_2},\ldots,n^{m_n}\}$, where $i$ appears $m_i$ times, a multipermutation $π$ of $\mathcal{M}$ is called {\em quasi-Stirling}, if it contains no subword of the form $abab$ with $a\neq b$. We designate exactly one entry of $π$, say $k\in \mathcal{M}$, which is not the leftmost entry among all entries with the same value, by underlining it in $π$, and we refer to the pair $(π,k)$ as a quasi-Stirling multipermutation of $\mathcal{M}$ rooted at $k$. By introducing certain vertex and edge labeled trees, we give a new bijective proof of an identity due to Yan, Yang, Huang and Zhu, which links the enumerator of rooted quasi-Stirling multipermutations by the numbers of ascents, descents, and plateaus, with the exponential generating function of the {\em bivariate Eulerian polynomials}. This identity can be viewed as a natural extension of Elizalde's result on $k$-quasi-Stirling permutations, and our bijective approach to proving it enables us to: (1) prove bijectively a Carlitz type identity involving quasi-Stirling polynomials on multisets that was first obtained by Yan and Zhu; (2) confirm a recent partial $γ$-positivity conjecture due to Lin, Ma and Zhang, and find a combinatorial interpretation of the $γ$-coefficients in terms of two new statistics defined on quasi-Stirling multipermutations called sibling descents and double sibling descents.
Combining Orthology and Xenology Data in a Common Phylogenetic Tree
Published
• View Publication
• BIB
A rooted tree $T$ with vertex labels $t(v)$ and set-valued edge labels $λ(e)$ defines maps $δ$ and $\varepsilon$ on the pairs of leaves of $T$ by setting $δ(x,y)=q$ if the last common ancestor $\text{lca}(x,y)$ of $x$ and $y$ is labeled $q$, and $m\in \varepsilon(x,y)$ if $m\inλ(e)$ for at least one edge $e$ along the path from $\text{lca}(x,y)$ to $y$. We show that a pair of maps $(δ,\varepsilon)$ derives from a tree $(T,t,λ)$ if and only if there exists a common refinement of the (unique) least-resolved vertex labeled tree $(T_δ,t_δ)$ that explains $δ$ and the (unique) least resolved edge labeled tree $(T_{\varepsilon},λ_{\varepsilon})$ that explains $\varepsilon$ (provided both trees exist). This result remains true if certain combinations of labels at incident vertices and edges are forbidden.
Partial $γ$-Positivity for Quasi-Stirling Permutations of Multisets
Published
• View Publication
• BIB
We prove that the enumerative polynomials of quasi-Stirling permutations of multisets with respect to the statistics of plateaux, descents and ascents are partial $γ$-positive, thereby confirming a recent conjecture posed by Lin, Ma and Zhang. This is accomplished by proving the partial $γ$-positivity of the enumerative polynomials of certain ordered labeled trees, which are in bijection with quasi-Stirling permutations of multisets. As an application, we provide an alternative proof of the partial $γ$-positivity of the enumerative polynomials on Stirling permutations of multisets.
Quasi-Stirling Polynomials on Multisets
Published
• View Publication
• BIB
A permutation $π$ of a multiset is said to be a {\em quasi-Stirling} permutation if there does not exist four indices $i<j<k<\ell$ such that $π_i=π_k$ and $π_j=π_{\ell}$. For a multiset $\mathcal{M}$, denote by $\overline{\mathcal{Q}}_{\mathcal{M}}$ the set of quasi-Stirling permutations of $\mathcal{M}$. The {\em qusi-Stirling polynomial} on the multiset $\mathcal{M}$ is defined by $ \overline{Q}_{\mathcal{M}}(t)=\sum_{π\in \overline{\mathcal{Q}}_{\mathcal{M}}}t^{des(π)}$, where $des(π)$ denotes the number of descents of $π$. By employing generating function arguments, Elizalde derived an elegant identity involving quasi-Stirling polynomials on the multiset $\{1^2, 2^2, \ldots, n^2\}$, in analogy to the identity on Stirling polynomials. In this paper, we derive an identity involving quasi-Stirling polynomials $\overline{Q}_{\mathcal{M}}(t)$ for any multiset $\mathcal{M}$, which is a generalization of the identity on Eulerian polynomial and Elizalde's identity on quasi-Stirling polynomials on the multiset $\{1^2, 2^2, \ldots, n^2\}$. We provide a combinatorial proof the identity in terms of certain ordered labeled trees. Specializing $\mathcal{M}=\{1^2, 2^2, \ldots, n^2\}$ implies a combinatorial proof of Elizalde's identity in answer to the problem posed by Elizalde. As an application, our identity enables us to show that the quasi-Stirling polynomial $\overline{Q}_{\mathcal{M}}(t)$ has only real roots and the coefficients of $\overline{Q}_{\mathcal{M}}(t)$ are unimodal and log-concave for any multiset $\mathcal{M}$, in analogy to Brenti's result for Stirling polynomials on multisets.
Total positivity of some polynomial matrices that enumerate labeled trees and forests, I. Forests of rooted labeled trees
Published in Monatsh. Math. 200, 389--452 (2023)
• View Publication
• BIB
We consider the lower-triangular matrix of generating polynomials that enumerate $k$-component forests of rooted trees on the vertex set $[n]$ according to the number of improper edges (generalizations of the Ramanujan polynomials). We show that this matrix is coefficientwise totally positive and that the sequence of its row-generating polynomials is coefficientwise Hankel-totally positive. More generally, we define the generic rooted-forest polynomials by introducing also a weight $m! \, φ_m$ for each vertex with $m$ proper children. We show that if the weight sequence $φ$ is Toeplitz-totally positive, then the two foregoing total-positivity results continue to hold. Our proofs use production matrices and exponential Riordan arrays.
Isomorphic unordered labeled trees up to substitution ciphering
Published
• View Publication
• BIB
Given two messages - as linear sequences of letters, it is immediate to determine whether one can be transformed into the other by simple substitution cipher of the letters. On the other hand, if the letters are carried as labels on nodes of topologically isomorphic unordered trees, determining if a substitution exists is referred to as marked tree isomorphism problem in the literature and has been show to be as hard as graph isomorphism. While the left-to-right direction provides the cipher of letters in the case of linear messages, if the messages are carried by unordered trees, the cipher is given by a tree isomorphism. The number of isomorphisms between two trees is roughly exponential in the size of the trees, which makes the problem of finding a cipher difficult by exhaustive search. This paper presents a method that aims to break the combinatorics of the isomorphisms search space. We show that in a linear time (in the size of the trees), we reduce the cardinality of this space by an exponential factor on average.
Two Metrics on Rooted Unordered Trees with Labels
Published in Algorithms Mol Biol 17, 13 (2022)
• View Publication
• BIB
The early development of a zygote can be mathematically described by a developmental tree. To compare developmental trees of different species, we need to define distances on trees. If children cells after a division are not distinguishable, developmental trees are represented by the space $\mathcal{T}$ of rooted trees with possibly repeated labels, where all vertices are unordered. If children cells after a division are partially distinguishable, developmental trees are represented by the space $\mathcal{P}$ of rooted trees with possibly repeated labels, where vertices can be ordered or unordered. On $\mathcal{T}$, the space of rooted unordered trees with possibly repeated labels, we define two metrics: the best-match metric and the left-regular metric, which show some advantages over existing methods. On $\mathcal{P}$, the space of rooted labeled trees with ordered or unordered vertices, there is no metric, and we define a semimetric, which is a variant of the best-match metric. To compute the best-match distance between two trees, the expected time complexity and worst-case time complexity are both $\mathcal{O}(n^2)$, where $n$ is the tree size. To compute the left-regular distance between two trees, the expected time complexity is $\mathcal{O}(n)$, and the worst-case time complexity is $\mathcal{O}(n\log n)$. For rooted labeled trees with (fully/partially) unordered vertices, we define metrics (semimetric) that have fast algorithms to compute and have advantages over existing methods. Such trees also appear outside of developmental biology, and such metrics can be applied to other types of trees which have more extensive applications, especially in molecular biology.
From Modular Decomposition Trees to Rooted Median Graphs
Published
• View Publication
• BIB
The modular decomposition of a symmetric map $δ\colon X\times X \to Υ$ (or, equivalently, a set of symmetric binary relations, a 2-structure, or an edge-colored undirected graph) is a natural construction to capture key features of $δ$ in labeled trees. A map $δ$ is explained by a vertex-labeled rooted tree $(T,t)$ if the label $δ(x,y)$ coincides with the label of the last common ancestor of $x$ and $y$ in $T$, i.e., if $δ(x,y)=t(\mathrm{lca}(x,y))$. Only maps whose modular decomposition does not contain prime nodes, i.e., the symbolic ultrametrics, can be exaplained in this manner. Here we consider rooted median graphs as a generalization to (modular decomposition) trees to explain symmetric maps. We first show that every symmetric map can be explained by "extended" hypercubes and half-grids. We then derive a a linear-time algorithm that stepwisely resolves prime vertices in the modular decomposition tree to obtain a rooted and labeled median graph that explains a given symmetric map $δ$. We argue that the resulting "tree-like" median graphs may be of use in phylogenetics as a model of evolutionary relationships.
Labeled trees generating complete, compact, and discrete ultrametric spaces
Published
• View Publication
• BIB
We investigate the interrelations between labeled trees and ultrametric spaces generated by these trees. The labeled trees, which generate complete ultrametrics, totally bounded ultrametrics, and discrete ones, are characterized up to isomorphism. As corollary, we obtain a characterization of labeled trees generating compact ultrametrics, and discrete totally bounded ultrametrics. It is also shown that every ultrametric space generated by labeled tree contains a dense discrete subspace.
Symmetric Dyck tilings, ballot tableaux and tree-like tableaux of shifted shapes
Symmetric Dyck tilings and ballot tilings are certain tilings in the region surrounded by two ballot paths. We study the relations of combinatorial objects which are bijective to symmetric Dyck tilings such as labeled trees, Hermite histories, and perfect matchings. We also introduce two operations on labeled trees for symmetric Dyck tilings: symmetric Dyck tiling strip (symDTS) and symmetric Dyck tiling ribbon (symDTR). We give two definitions of Hermite histories for symmetric Dyck tilings, and show that they are equivalent by use of the correspondence between symDTS operation and an Hermite history. Since ballot tilings form a subset in the set of symmetric Dyck tilings, we construct an inclusive map from labeled trees for ballot tilings to labeled trees for symmetric Dyck tilings. By this inclusive map, the results for symmetric Dyck tilings can be applied to those of ballot tilings. We introduce and study the notions of ballot tableaux and tree-like tableaux of shifted shapes, which are generalizations of Dyck tableaux and tree-like tableaux, respectively. The correspondence between ballot tableaux and tree-like tableaux of shifted shapes is given by using the symDTR operation and the structure of labeled trees for symmetric Dyck tilings.
Poset topology of $s$-weak order via SB-labelings
Published
• View Publication
• BIB
Ceballos and Pons generalized weak order on permutations to a partial order on certain labeled trees, thereby introducing a new class of lattices called $s$-weak order. They also generalized the Tamari lattice by defining a particular sublattice of $s$-weak order called the $s$-Tamari lattice. We prove that the homotopy type of each open interval in $s$-weak order and in the $s$-Tamari lattice is either a ball or sphere. We do this by giving $s$-weak order and the $s$-Tamari lattice a type of edge labeling known as an SB-labeling. We characterize which intervals are homotopy equivalent to spheres and which are homotopy equivalent to balls; we also determine the dimension of the spheres for the intervals yielding spheres.
The distributions under two species-tree models of the number of root ancestral configurations for matching gene trees and species trees
Published
• View Publication
• BIB
For a pair consisting of a gene tree and a species tree, the ancestral configurations at an internal node of the species tree are the distinct sets of gene lineages that can be present at that node. Ancestral configurations appear in computations of gene tree probabilities under evolutionary models conditional on fixed species trees, and the enumeration of root ancestral configurations -- ancestral configurations at the root of the species tree -- assists in describing the complexity of these computations. In the case that the gene tree matches the species tree in topology, we study the distribution of the number of root ancestral configurations of a random labeled tree topology under each of two models.
Exact-$2$-Relation Graphs
Published
• View Publication
• BIB
Pairwise compatibility graphs (PCGs) with non-negative integer edge weights recently have been used to describe rare evolutionary events and scenarios with horizontal gene transfer. Here we consider the case that vertices are separated by exactly two discrete events: Given a tree $T$ with leaf set $L$ and edge-weights $λ: E(T)\to\mathbb{N}_0$, the non-negative integer pairwise compatibility graph $\textrm{nniPCG}(T,λ,2,2)$ has vertex set $L$ and $xy$ is an edge whenever the sum of the non-negative integer weights along the unique path from $x$ to $y$ in $T$ equals $2$. A graph $G$ has a representation as $\textrm{nniPCG}(T,λ,2,2)$ if and only if its point-determining quotient $G/\!\rthin$ is a block graph, where two vertices are in relation $\rthin$ if they have the same neighborhood in $G$. If $G$ is of this type, a labeled tree $(T,λ)$ explaining $G$ can be constructed efficiently. In addition, we consider an oriented version of this class of graphs.
Bijections on $r$-Shi and $r$-Catalan Arrangements
Published
• View Publication
• BIB
Associated with the $r$-Shi arrangement and $r$-Catalan arrangement in $\Bbb{R}^n$, we introduce a cubic matrix for each region to establish two bijections in a uniform way. Firstly, the positions of minimal positive entries in column slices of the cubic matrix will give a bijection from regions of the $r$-Shi arrangement to $O$-rooted labeled $r$-trees. Secondly, the numbers of positive entries in column slices of the cubic matrix will give a bijection from regions of the $r$-Catalan arrangement to pairings of permutation and $r$-Dyck path. Moreover, the numbers of positive entries in row slices of the cubic matrix will recover the Pak-Stanley labeling, a celebrated bijection from regions of the $r$-Shi arrangement to $r$-parking functions.
On recursively defined combinatorial classes and labelled trees
Published
• View Publication
• BIB
We define and prove isomorphisms between three combinatorial classes involving labeled trees. We also give an alternative proof by means of generating functions.