Papers by Stephan Wagner
70 paper(s) by this author
· All BibTeX
Refined enumeration of $k$-plane trees and $k$-noncrossing trees
Published
• View Publication
• BIB
A $k$-plane tree is a plane tree whose vertices are assigned labels between $1$ and $k$ in such a way that the sum of the labels along any edge is no greater than $k+1$. These trees are known to be related to $(k+1)$-ary trees, and they are counted by a generalised version of the Catalan numbers. We prove a surprisingly simple refined counting formula, where we count trees with a prescribed number of labels of each kind. Several corollaries are derived from this formula, and an analogous theorem is proven for $k$-noncrossing trees, a similarly defined family of labelled noncrossing trees that are related to $(2k+1)$-ary trees.
Enumeration of Generalized Dyck Paths Based on the Height of Down-Steps Modulo $k$
Published in Electron. J. Combin.30(2023), no.1, Paper No. 1.26, 18 pp
• View Publication
• BIB
For fixed non-negative integers $k$, $t$, and $n$, with $t < k$, a $k_t$-Dyck path of length $(k+1)n$ is a lattice path that starts at $(0, 0)$, ends at $((k+1)n, 0)$, stays weakly above the line $y = -t$, and consists of steps from the step-set $\{(1, 1), (1, -k)\}$. We enumerate the family of $k_t$-Dyck paths by considering the number of down-steps at a height of $i$ modulo $k$. Given a tuple $(a_1, a_2, \ldots, a_k)$ we find an exact enumeration formula for the number of $k_t$-Dyck paths of length $(k+1)n$ with $a_i$ down-steps at a height of $i$ modulo $k$, $1 \leq i \leq k$. The proofs given are done via bijective means or with generating functions.
Convex characters, algorithms and matchings
Published
• View Publication
• BIB
Phylogenetic trees are used to model evolution: leaves are labelled to represent contemporary species ("taxa") and interior vertices represent extinct ancestors. Informally, convex characters are measurements on the contemporary species in which the subset of species (both contemporary and extinct) that share a given state, form a connected subtree. In \cite{KelkS17} it was shown how to efficiently count, list and sample certain restricted subfamilies of convex characters, and algorithmic applications were given. We continue this work in a number of directions. First, we show how combining the enumeration of convex characters with existing parameterised algorithms can be used to speed up exponential-time algorithms for the \emph{maximum agreement forest problem} in phylogenetics. Second, we re-visit the quantity $g_2(T)$, defined as the number of convex characters on $T$ in which each state appears on at least 2 taxa. We use this to give an algorithm with running time $O( φ^{n} \cdot \text{poly}(n) )$, where $φ\approx 1.6181$ is the golden ratio and $n$ is the number of taxa in the input trees, for computation of \emph{maximum parsimony distance on two state characters}. By further restricting the characters counted by $g_2(T)$ we open an interesting bridge to the literature on enumeration of matchings. By crossing this bridge we improve the running time of the aforementioned parsimony distance algorithm to $O( 1.5895^{n} \cdot \text{poly}(n) )$, and obtain a number of new results in themselves relevant to enumeration of matchings on at-most binary trees.
Broadcasting induced colourings of random recursive trees and preferential attachment trees
Published in Random Structures & Algorithms, 2023
• View Publication
• BIB
In this work we consider random two-colourings of random linear preferential attachment trees, which includes random recursive trees, random plane-oriented recursive trees, random binary search trees, and a class of random $d$-ary trees. The random colouring is defined by assigning the root of the tree the colour red or blue with equal probability, and all other vertices are assigned the colour of their parent with probability $p$ and the other colour otherwise. These colourings have been previously studied in other contexts, including Ising models and broadcasting, and can be considered as generalizations of bond percolation. With the help of Pólya urns, we prove limiting distributions, after proper rescalings, for the number of vertices of each colour, the number of monochromatic subtrees of each colour, as well as the number of leaves and fringe subtrees with two-colourings. Using methods from analytic combinatorics, we also provide precise descriptions of the limiting distribution after proper rescaling of the size of the root cluster; the largest monochromatic subtree containing the root. The description of the limiting distributions extends previous work on bond percolation in random preferential attachment trees.
Distinct Fringe Subtrees in Random Trees
Published
• View Publication
• BIB
A fringe subtree of a rooted tree is a subtree induced by one of the vertices and all its descendants. We consider the problem of estimating the number of distinct fringe subtrees in two types of random trees: simply generated trees and families of increasing trees (recursive trees, $d$-ary increasing trees and generalized plane-oriented recursive trees). We prove that the order of magnitude of the number of distinct fringe subtrees (under rather mild assumptions on what `distinct' means) in random trees with $n$ vertices is $n/\sqrt{\log n}$ for simply generated trees and $n/\log n$ for increasing trees.
On the maximum mean subtree order of trees
Published in European Journal of Combinatorics 2021
• View Publication
• BIB
A subtree of a tree is any induced subgraph that is again a tree (i.e., connected). The mean subtree order of a tree is the average number of vertices of its subtrees. This invariant was first analyzed in the 1980s by Jamison. An intriguing open question raised by Jamison asks whether the maximum of the mean subtree order, given the order of the tree, is always attained by some caterpillar. While we do not completely resolve this conjecture, we find some evidence in its favor by proving different features of trees that attain the maximum. For example, we show that the diameter of a tree of order $n$ with maximum mean subtree order must be very close to $n$. Moreover, we show that the maximum mean subtree order is equal to $n - 2\log_2 n + O(1)$. For the local mean subtree order, which is the average order of all subtrees containing a fixed vertex, we can be even more precise: we show that its maximum is always attained by a broom and that it is equal to $n - \log_2 n + O(1)$.
The birth of the strong components
Published
• View Publication
• BIB
Random directed graphs $D(n,p)$ undergo a phase transition around the point $p = 1/n$, and the width of the transition window has been known since the works of Luczak and Seierstad. They have established that as $n \to \infty$ when $p = (1 + μn^{-1/3})/n$, the asymptotic probability that the strongly connected components of a random directed graph are only cycles and single vertices decreases from 1 to 0 as $μ$ goes from $-\infty$ to $\infty$.
By using techniques from analytic combinatorics, we establish the exact limiting value of this probability as a function of $μ$ and provide more properties of the structure of a random digraph around, below and above its transition point. We obtain the limiting probability that a random digraph is acyclic and the probability that it has one strongly connected complex component with a given difference between the number of edges and vertices (called excess). Our result can be extended to the case of several complex components with given excesses as well in the whole range of sparse digraphs.
Our study is based on a general symbolic method which can deal with a great variety of possible digraph families, and a version of the saddle point method which can be systematically applied to the complex contour integrals appearing from the symbolic method. While the technically easiest model is the model of random multidigraphs, in which multiple edges are allowed, and where edge multiplicities are sampled independently according to a Poisson distribution with a fixed parameter $p$, we also show how to systematically approach the family of simple digraphs, where multiple edges are forbidden, and where 2-cycles are either allowed or not.
Our theoretical predictions are supported by numerical simulations, and we provide tables of numerical values for the integrals of Airy functions that appear in this study.
Trees with minimum number of infima closed sets
Published
• View Publication
• BIB
Let $T$ be a rooted tree, and $V(T)$ its set of vertices. A subset $X$ of $V(T)$ is called an infima closed set of $T$ if for any two vertices $u,v\in X$, the first common ancestor of $u$ and $v$ is also in $X$. This paper determines the trees with minimum number of infima closed sets among all rooted trees of given order, thereby answering a question of Klazar. It is shown that these trees are essentially complete binary trees, with the exception of vertices at the last levels. Moreover, an asymptotic estimate for the minimum number of infima closed sets in a tree with $n$ vertices is also provided.
Extremal trees with fixed degree sequence
Published
• View Publication
• BIB
The greedy tree $\mathcal{G}(D)$ and the $\mathcal{M}$-tree $\mathcal{M}(D)$ are known to be extremal among trees with degree sequence $D$ with respect to various graph invariants. This paper provides a general theorem that covers a large family of invariants for which $\mathcal{G}(D)$ or $\mathcal{M}(D)$ is extremal. Many known results, for example on the Wiener index, the number of subtrees, the number of independent subsets and the number of matchings follow as corollaries, as do some new results on invariants such as the number of rooted spanning forests, the incidence energy and the solvability. We also extend our results on trees with fixed degree sequence $D$ to the set of trees whose degree sequence is majorised by a given sequence $D$, which also has a number of applications.
On the Collection of Fringe Subtrees in Random Binary Trees
Published
• View Publication
• BIB
A fringe subtree of a rooted tree is a subtree consisting of one of the nodes and all its descendants. In this paper, we are specifically interested in the number of non-isomorphic trees that appear in the collection of all fringe subtrees of a binary tree. This number is analysed under two different random models: uniformly random binary trees and random binary search trees.
In the case of uniformly random binary trees, we show that the number of non-isomorphic fringe subtrees lies between $c_1n/\sqrt{\ln n}(1+o(1))$ and $c_2n/\sqrt{\ln n}(1+o(1))$ for two constants $c_1 \approx 1.0591261434$ and $c_2 \approx 1.0761505454$, both in expectation and with high probability, where $n$ denotes the size (number of leaves) of the uniformly random binary tree. A similar result is proven for random binary search trees, but the order of magnitude is $n/\ln n$ in this case.
Our proof technique can also be used to strengthen known results on the number of distinct fringe subtrees (distinct in the sense of ordered trees). This quantity is of the same order of magnitude in both cases, but with slightly different constants in the upper and lower bounds.
On the probability that a random subtree is spanning
Published
• View Publication
• BIB
We consider the quantity $P(G)$ associated with a graph $G$ that is defined as the probability that a randomly chosen subtree of $G$ is spanning. Motivated by conjectures due to Chin, Gordon, MacPhee and Vincent on the behaviour of this graph invariant depending on the edge density, we establish first that $P(G)$ is bounded below by a positive constant provided that the minimum degree is bounded below by a linear function in the number of vertices. Thereafter, the focus is shifted to the classical Erdős-Rényi random graph model $G(n,p)$. It is shown that $P(G)$ converges in probability to $e^{-1/(ep_{\infty})}$ if $p \to p_{\infty} > 0$ and to $0$ if $p \to 0$.
The average size of matchings in graphs
In this paper, we consider the average size of independent edge sets, also called matchings, in a graph. We characterize the extremal graphs for the average size of matchings in general graphs and trees. In addition, we obtain inequalities between the average size of matchings and the number of matchings as well as the matching energy, which is defined as the sum of the absolute values of the zeros of the matching polynomial.
On two subclasses of Motzkin paths and their relation to ternary trees
Two subclasses of Motzkin paths, S-Motzkin and T-Motzkin paths, are introduced. We provide bijections between S-Motzkin paths and ternary trees, S-Motzkin paths and non-crossing trees, and T-Motzkin paths and ordered pairs of ternary trees. Symbolic equations for both paths, and thus generating functions for the paths, are provided. Using these, various parameters involving the two paths are analyzed.
On the inducibility of small trees
Published in Discrete Mathematics & Theoretical Computer Science, vol. 21 no. 4, Combinatorics (October 17, 2019) dmtcs:5381
• View Publication
• BIB
The quantity that captures the asymptotic value of the maximum number of appearances of a given topological tree (a rooted tree with no vertices of outdegree $1$) $S$ with $k$ leaves in an arbitrary tree with sufficiently large number of leaves is called the inducibility of $S$. Its precise value is known only for some specific families of trees, most of them exhibiting a symmetrical configuration. In an attempt to answer a recent question posed by Czabarka, Székely, and the second author of this article, we provide bounds for the inducibility $J(A_5)$ of the $5$-leaf binary tree $A_5$ whose branches are a single leaf and the complete binary tree of height $2$. It was indicated before that $J(A_5)$ appears to be `close' to $1/4$. We can make this precise by showing that $0.24707\ldots \leq J(A_5) \leq 0.24745\ldots$. Furthermore, we also consider the problem of determining the inducibility of the tree $Q_4$, which is the only tree among $4$-leaf topological trees for which the inducibility is unknown.
Further results on the inducibility of $d$-ary trees
A subset of leaves of a rooted tree induces a new tree in a natural way. The density of a tree $D$ inside a larger tree $T$ is the proportion of such leaf-induced subtrees in $T$ that are isomorphic to $D$ among all those with the same number of leaves as $D$. The inducibility of $D$ measures how large this density can be as the size of $T$ tends to infinity. In this paper, we explicitly determine the inducibility in some previously unknown cases and find general upper and lower bounds, in particular in the case where $D$ is balanced, i.e., when its branches have at least almost the same size. Moreover, we prove a result on the speed of convergence of the maximum density of $D$ in strictly $d$-ary trees $T$ (trees where every internal vertex has precisely $d$ children) of a given size $n$ to the inducibility as $n \to \infty$, which supports an open conjecture.
A central limit theorem for almost local additive tree functionals
Published
• View Publication
• BIB
An additive functional of a rooted tree is a functional that can be calculated recursively as the sum of the values of the functional over the branches, plus a certain toll function. Janson recently proved a central limit theorem for additive functionals of conditioned Galton-Watson trees under the assumption that the toll function is local, i.e. only depends on a fixed neighbourhood of the root. We extend his result to functionals that are "almost local" in a certain sense, thus covering a wider range of functionals. The notion of almost local functional intuitively means that the toll function can be approximated well by considering only a neighbourhood of the root. Our main result is illustrated by several explicit examples including natural graph theoretic parameters such as the number of independent sets, the number of matchings, and the number of dominating sets. We also cover a functional stemming from a tree reduction process that was studied by Hackl, Heuberger, Kropf, and Prodinger.
On the Number of Increasing Trees with Label Repetitions
Published
• View Publication
• BIB
We study the asymptotic number of certain monotonically labeled increasing trees arising from a generalized evolution process. The main difference between the presented model and the classical model of binary increasing trees is that the same label can appear in distinct branches of the tree. In the course of the analysis we develop a method to extract asymptotic information on the coefficients of purely formal power series. The method is based on an approximate Borel transform (or, more generally, Mittag-Leffler transform) which enables us to quickly guess the exponential growth rate. With this guess the sequence is then rescaled and a singularity analysis of the generating function of the scaled counting sequence yields accurate asymptotics. The actual analysis is based on differential equations and a Tauberian argument. The counting problem for trees of size n exhibits interesting asymptotics involving powers of n with irrational exponents.
The ancestral matrix of a rooted tree
Published
• View Publication
• BIB
Given a rooted tree $T$ with leaves $v_1,v_2,\ldots,v_n$, we define the ancestral matrix $C(T)$ of $T$ to be the $n \times n$ matrix for which the entry in the $i$-th row, $j$-th column is the level (distance from the root) of the first common ancestor of $v_i$ and $v_j$. We study properties of this matrix, in particular regarding its spectrum: we obtain several upper and lower bounds for the eigenvalues in terms of other tree parameters. We also find a combinatorial interpretation for the coefficients of the characteristic polynomial of $C(T)$, and show that for $d$-ary trees, a specific value of the characteristic polynomial is independent of the precise shape of the tree.
Reducing Simply Generated Trees by Iterative Leaf Cutting
Published in Proceedings of the Sixteenth Workshop on Analytic Algorithmics and Combinatorics (ANALCO) (Philadelphia PA), SIAM, 2019, pp. 36-44
• View Publication
• BIB
We consider a procedure to reduce simply generated trees by iteratively removing all leaves. In the context of this reduction, we study the number of vertices that are deleted after applying this procedure a fixed number of times by using an additive tree parameter model combined with a recursive characterization.
Our results include asymptotic formulas for mean and variance of this quantity as well as a central limit theorem.
The average size of independent sets of graphs
Published
• View Publication
• BIB
In this paper, we study the average size of independent (vertex) sets of a graph. This invariant can be regarded as the logarithmic derivative of the independence polynomial evaluated at $1$. We are specifically concerned with extremal questions. The maximum and minimum for general graphs are attained by the empty and complete graph respectively, while for trees we prove that the path minimises the average size of independent sets and the star maximises it. While removing a vertex does not always decrease the average size of independent sets, we prove that there always exists a vertex for which this is the case.