arXiv++ Combinatorics

Browse math.CO papers from arXiv

species

265 papers tagged with this keyword
2026-09-04
The Molecular Species $\mathbf{C}_α$: Geometric Realization and a Closed Formula for Kronecker Coefficients
In this paper we first introduce the \emph{infinite multi-row periodic pattern of shape $α$} as a geometric realization of $\mathbf{C}_α$. We then give an explicit formula for the coefficients $b^λ_{α,β}$ appearing in the species decomposition \[ \mathbf{C}_α\times \mathbf{C}_β = \sum_{λ\vdash n} b^λ_{α,β}\,\mathbf{C}_λ. \]
2026-09-04
Layered mixed matrices and reaction networks
The purpose of this work is twofold. In the first part, we consider layered mixed matrices introduced by Murota, relate them to existing notions in combinatorial commutative algebra, and investigate the irreducibility of their determinants. Furthermore, for a layered mixed matrix in combinatorial canonical form, we determine the sparsity structure of its inverse. That is, we characterize which entries of the inverse are nonzero. In the second part, we establish for the first time a formal connection between these algebraic results and the theory of buffering structures for reaction networks developed by Mochizuki and Okada. We identify the lattice of buffering structures with the lattice of order ideals of the block poset of the combinatorial canonical form of the associated layered mixed matrix. This allows us to characterize the reducibility of the symbolic Jacobian determinant as a polynomial in the reaction-rate derivatives, as well as the nonzero sensitivity responses of species concentrations to reaction-rate perturbations.
2026-08-22
The species of interval orders
We show that, in the ring of virtual species, \[ \mathcal{I}=\sum_{m\geq 0}(-1)^m\prod_{i=1}^{m}\bigl((E^{-1})^i-1\bigr), \] where $\mathcal{I}$ is the species of interval orders and $E^{-1}$ is the multiplicative inverse of the species $E$ of sets. The right-hand side is the virtual species of signed ballot matrices introduced by Claesson and Hannah. They showed that its signed cardinality counts labeled interval orders. We strengthen this to a species identity, which we prove twice: first algebraically and then bijectively, using a natural sign-reversing involution. The cycle index series of $\mathcal{I}$ specializes to the generating series for labeled and unlabeled interval orders. We describe the automorphism group of an interval order as a Young subgroup and prove the identity $\mathcal{I}=\mathcal{R}\circ E_+$, where $\mathcal{R}$ is the species of rigid interval orders. We also show that Glaisher's T-number $T_n$ counts the $24$-colored interval orders on $[n]$ in which no isolated element has color $24$.
How many cherry-picking sequences are needed to reduce all subtrees of a phylogenetic tree?
Phylogenetic networks are graphs that represent the evolutionary history of species. Recently, the class of orchard phylogenetic networks, which can be reduced by so-called cherry-picking sequences, has gained attention for its computational and biological aspects. In this paper, we study a fundamental question on orchards and their cherry-picking sequences by considering the CoveringNumber problem: given an orchard network $N$, how many cherry-picking sequences are needed to reduce all subnetworks of $N$? We initiate this study by considering the problem for trees. We then show that the covering number can be computed for binary trees recursively using a similar but more fine-grained notion of survival covering number. We also give a recursive formula for the survival covering number of non-binary trees. However, computing the covering number for non-binary trees appears to be considerably more challenging. For this case, we show that the covering number of star trees (whose root is adjacent to all leaves) is equivalent to the so-called SubsetConnectivity problem, which we introduce in this paper. Finally, we show that if there is no restriction on the sequence length, a single sequence of minimum length $\binom{n}{2}$ suffices to reduce all subtrees of a tree on $n$ leaves.
Type $C$ multiline queues and the open-boundary TASEP
The totally asymmetric simple exclusion process (TASEP) with open boundaries is a finite Markov chain describing particles hopping between adjacent sites on a one-dimensional lattice with left and right boundary transitions governed by parameters $α$ and $β$. The multispecies TASEP is a higher-rank generalization in which particles have different species. Multiline queues were introduced by Ferrari and Martin (2007) to compute the stationary distribution of the multispecies TASEP on a circle. It has remained an open problem to find a combinatorial formula for the stationary distribution of the multispecies open-boundary TASEP. Using Kirillov--Reshetikhin crystals of type $C$, we construct type $C$ multiline queues and a corresponding Ferrari--Martin pairing algorithm that projects them to TASEP configurations. This yields a combinatorial formula for the stationary distribution of the multispecies open-boundary TASEP for the $α=β=1$ specialization.
Enumerating monophyletic characters in mathematical phylogenetics
Grouping species according to their phylogenetic relationships often results in different groups than grouping them according to their shared traits. Monophyletic groups play an important role in this regard, as they are groups of species sharing the same trait and being uniquely defined by a joint phylogenetic subtree. This immediately leads to the question of how to identify possible monophyletic groups in characters, which assign each present-day species a certain trait and which are typically used for phylogenetic tree reconstruction. In our manuscript, we provide a general formula to quantify how many different characters are monophyletic on any given tree and provide simple formulae for binary characters and for certain tree shapes. We also investigate relations between monophyly and the well-known phylogenetic tree reconstruction criterion maximum parsimony by providing a linear-time algorithm which determines the parsimony score together with the monophyly type of a character on a tree.
Proximity Measures for Classes of Phylogenetic Networks
Phylogenetic networks are used to represent the evolutionary history of species. Due to biological interpretations and computational advantages, researchers have focused on restricted classes of phylogenetic networks, such as tree-child, orchard, and tree-based. These classes capture different notions of tree-likeness: tree-child networks require every internal vertex to have a taxon reachable by a tree path, orchard networks are trees with horizontal arcs (for modelling histories rife with horizontal gene transfers), and tree-based networks are trees with additional (not-necessarily horizontal) arcs. A natural question to ask is ``how far is a given network from belonging to a particular class?'' This motivates the study of proximity measures, which measure the minimum number of graph modifications required to transform a network into one belonging to a particular class. In this paper, we consider three proximity measures based on leaf addition, valid arc deletion, and arc deletion. We study pairwise comparability of the proximity measures, prove complexity results, and derive extremal bounds for the classes of tree, tree-child, orchard, and tree-based networks.
2026-07-12
Extended generalized permutahedra, and cointeracting bialgebras
A Hopf monoid structure on extended generalized permutahedra (EGP) was recently introduced by M.Aguiar and F.Ardila. We investigate the existence of a cointeracting bialgebra structure on EGP's. We show that a suitable notion of cointeraction exists, not in the classical comodule sense, but via the framework of measuring algebras. The comodule-type map assigns to each polyhedron the sum of pairs of face and tangent cone at the face. EGP's and affine cone EGP's form the cointeracting bimonoids in species with EGP as a third measuring structure. EGP's are in bijection to extended submodular functions. For an EGP, we also describe explicitly the submodular functions of its faces and tangent cones. The braid fan and its relation to preorders play a key role in this description.
2026-05-24
Asymptotic probability of irreducibles III: Anti-SEQ
In this paper, we study the structure of the complete asymptotic expansion of the probability that a large combinatorial object is connected or consists of a given number of connected components. For rapidly growing labeled families of structures, the coefficients involved in these expansions are possibly negative integers. Using species theory, we interpret these coefficients as the difference between the counting sequences of two derivative species of structures. In particular, we show that this difference can be viewed as the counting sequence of the virtual species obtained with the help of an "anti-$\mathrm{SEQ}$" operator applied to the initial family of structures. Applications include $P$-angulated discrete surfaces, quadratic square-tiled surfaces, and non-orientable graph encoded manifolds, which were not reachable with our previous methods. Moving on to the weighted species, we establish the whole structure of the asymptotic expansion of the probability that a graph is connected in the Erdős-Rényi model $G(n,p)$. Here, the asymptotic coefficients are polynomials in $\frac{p}{1-p}$ and can be described both in terms of simple graphs and irreducible tournaments with ties. We also provide general asymptotic results for sequence and cycle decomposition, as well as the complete asymptotic expansion of the probability that a random labeled tournament with ties is irreducible.
2026-05-20
An axiomatic framework from splitting and merging in MAT-labeled graphs, vines, and single-peaked domains
In recent work (Forum Math.~Sigma, 2024), we established a correspondence between MAT-labeled graphs arising from hyperplane arrangements and regular vines from probability theory. In this paper, we extend this connection to Arrow's single-peaked domains in social choice theory. We show that MAT-labeled complete graphs, regular vines, and maximal Arrow's single-peaked domains arise from the same recursive combinatorial structure. Our main result gives an axiomatic characterization of these objects using the language of combinatorial species. At the heart of this characterization are two fundamental operations, called splitting and merging, together with natural compatibility conditions that uniquely determine the structures. As consequences, we obtain explicit correspondences between maximal Arrow's single-peaked domains, MAT-labeled complete graphs, and regular vines, resolving an open problem in the economics literature concerning the combinatorial characterization of single-peaked domains. We further show, by a direct proof, that regular vines are equivalent to $(n,3)$-extremal lattices from formal concept analysis. Consequently, these extremal lattices also fit naturally into the same splitting and merging framework, providing another example from a different area that satisfies our axiomatic characterization.
2026-05-07
A $μ$-distance for semidirected orchard phylogenetic networks
In evolutionary biology, phylogenetic networks are now widely used to represent the historical relationships between species and population, when this history includes reticulation events such as hybridization, gene flow and admixture between populations. Semidirected phylogenetic networks are appropriate models when the direction of some edges and the root position are not identifiable from data. Comparing semidirected networks is important in many applications. For rooted and directed networks, a $μ$-representation was originally introduced to distinguish tree-child networks, and has since been extended in two different directions: to the larger class of orchard directed networks by adding an extra component that counts paths to reticulations; and to semidirected networks, through an edge-based variant. However, the latter does not provide a distance between semidirected and orchard networks. We introduce here a new edge-based $μ$-representation capable of distinguishing distinct orchard binary semidirected networks. For this class, we provide a reconstruction algorithm and therefore obtain a true distance that is computable in polynomial time.
2026-04-12
Hopf substitutions in Species
In the theory of species, the species $\mathbf{L}$ of linear orders and the substitution operation $\boldsymbol{\circ}$ combine for a compelling result: given any positive comonoid $\mathbf{p}$, $\mathbf{L}\boldsymbol{\circ}\mathbf{p}$ carries the structure of Hopf monoid, freely generated by $\mathbf{p}$. Leaving aside the universal property this implies, we ask, "for which $\mathbf{b}$ does $\mathbf{b}\boldsymbol{\circ}\mathbf{p}$ carry the structure of Hopf monoid?" After answering this question, we look at basic properties of our construction. We also extend a result of the present authors, on interpolation in species, to this new context.
2026-04-11
Species, Symmetric Functions, and Kronecker Product
We study two new families of symmetric functions arising from a species-theoretic construction motivated by cycle structure. For each partition of $n$, we define two combinatorial species that decompose into molecules indexed by the same partition, giving rise to two corresponding basis of the homogeneous symmetric functions of degree $n$. We prove that each of these families forms a basis by exhibiting explicit cycle-index formulas and triangular transition matrices to the power-sum basis. Using these constructions, we generalize a classical result describing the Kronecker (Hadamard) product in the homogeneous basis to the two new settings. In particular, we show that the categories generated by these species are closed under the Kronecker product, and that the product of two basis elements expands with nonnegative integer coefficients. Our results provide a new combinatorial framework for studying the Kronecker product and suggest avenues toward interpreting its structure constants
2026-02-19
Multispecies inhomogeneous $t$-PushTASEP with general capacity
We study an $n$-species $t$-PushTASEP, an integrable long-range stochastic process, on a one-dimensional periodic lattice with inhomogeneities $x_1,\ldots,x_L$ and arbitrary capacity $l$ at each lattice site. The Markov matrix is identified with an alternating sum of commuting transfer matrices over all fundamental representations of $U_t(\widehat{sl}_{n+1})$. Stationary probabilities are expressed in a matrix product form involving a fusion of quantized corner transfer matrices for the strange five-vertex model introduced by Okado, Scrimshaw, and the second author. The resulting partition function, which serves as the normalization factor of the stationary probabilities, is obtained from the $l=1$ case by a finite plethystic substitution of length $l$.
2026-01-20
Bialgebraic structures on boolean functions
We study several bialgebraic structures on boolean functions, that is to say maps defined on the set of subsets of a finite set $X$, taking the value $0$ on $\emptyset$. Examples of boolean functions are given by the indicator function of the hyperedges of a given hypergraph, or the rank function of a matroid. We give the species of boolean functions a two-parameters family of products and a coproduct, and this defines a two-parameters family of twisted bialgebras. We then try to define a second coproduct on boolean functions, based on contractions, in order to obtain a double bialgebra. We show that this is not possible on the whole species of boolean functions, but that there exists a maximal subspecies where this is possible. This subspecies being rather mysterious, we introduce rigid boolean functions and show that this subspecies has indeed a second coproduct, as wished, and that it contains rank functions of matroids and indicator functions associated to hypergraphs. As a consequence, we obtain a unique polynomial invariant on rigid boolean functions, which is a generalization of the chromatic polynomial of graphs.
Characterizations of undirected 2-quasi best match graphs
Bipartite best match graphs (BMG) and their generalizations arise in mathematical phylogenetics as combinatorial models describing evolutionary relationships among related genes in a pair of species. In this work, we characterize the class of \emph{undirected 2-quasi-BMGs} (un2qBMGs), which form a proper subclass of the $P_6$-free chordal bipartite graphs. We show that un2qBMGs are exactly the class of bipartite graphs free of $P_6$, $C_6$, and the eight-vertex Sunlet$_4$ graph. Equivalently, a bipartite graph $G$ is un2qBMG if and only if every connected induced subgraph contains a ``heart-vertex'' which is adjacent to all the vertices of the opposite color. We further provide a $O(|V(G)|^3)$ algorithm for the recognition of un2qBMGs that, in the affirmative case, constructs a labeled rooted tree that ``explains'' $G$. Finally, since un2qBMGs coincide with the $(P_6,C_6)$-free bi-cographs, they can also be recognized in linear time.
2025-10-10 v2
Parameterized Algorithms for Diversity of Networks with Ecological Dependencies
For a phylogenetic tree, the phylogenetic diversity of a set A of taxa is the total weight of edges on paths to A. Finding small sets of maximal diversity is crucial for conservation planning, as it indicates where limited resources can be invested most efficiently. In recent years, efficient algorithms have been developed to find sets of taxa that maximize phylogenetic diversity either in a phylogenetic network or in a phylogenetic tree subject to ecological constraints, such as a food web. However, these aspects have mostly been studied independently. Since both factors are biologically important, it seems natural to consider them together. In this paper, we introduce decision problems where, given a phylogenetic network, a food web, and integers k, and D, the task is to find a set of k taxa with phylogenetic diversity of at least D under the maximize all paths measure, while also satisfying viability conditions within the food web. Here, we consider different definitions of viability, which all demand that a "sufficient" number of prey species survive to support surviving predators. We investigate the parameterized complexity of these problems and present several fixed-parameter tractable (FPT) algorithms. Specifically, we provide a complete complexity dichotomy characterizing which combinations of parameters - out of the size constraint k, the acceptable diversity loss D, the scanwidth of the food web, the maximum in-degree in the network, and the network height h - lead to W[1]-hardness and which admit FPT algorithms. Our primary methodological contribution is a novel algorithmic framework for solving phylogenetic diversity problems in networks where dependencies (such as those from a food web) impose an order, using a color coding approach.
2025-07-30 v2
An explicit power series result for the two type ASEP
Research in combinatorics has often focused on the ASEP (asymmetric simple exclusion process). The ASEP is inspired by processes in statistical mechanics, and involves particles of various species moving around a lattice. The particles do not change species. In the present paper, based on earlier results of Mortimer and Prellberg and others, we obtain a new power series results for the two type ASEP.
Characterizing semi-directed phylogenetic networks and their multi-rootable variants
In evolutionary biology, phylogenetic networks are graphs that provide a flexible framework for representing complex evolutionary histories that involve reticulate evolutionary events. Recently phylogenetic studies have started to focus on a special class of such networks called semi-directed networks. These graphs are defined as mixed graphs that can be obtained by de-orienting some of the arcs in some rooted phylogenetic network, that is, a directed acyclic graph whose leaves correspond to a collection of species and that has a single source or root vertex. However, this definition of semi-directed networks is implicit in nature since it is not clear when a mixed-graph enjoys this property or not. In this paper, we introduce novel, explicit mathematical characterizations of semi-directed networks, and also multi-semi-directed networks, that is, mixed graphs that can be obtained from directed phylogenetic networks that may have more than one root. In addition, through extending foundational tools from the theory of rooted networks into the semi-directed setting - such as cherry picking sequences, omnians, and path partitions - we characterize when a (multi-)semi-directed network can be obtained by de-orienting some rooted network that is contained in one of the well-known classes of tree-child, orchard, tree-based or forest-based networks. These results address structural aspects of (multi-)semi-directed networks and pave the way to improved theoretical and computational analyses of such networks, for example, within the development of algebraic evolutionary models that are based on such networks.
2025-07-12 v2
Counting fixed-point-free Cayley permutations
Two-sort species yield differential equations for functional digraphs of Cayley permutations. From these we obtain an explicit formula for fixed-point-free Cayley permutations and conjecture that their proportion tends to $1/e$, as for permutations and endofunctions. Our approach also yields counting formulas when the functional digraph is a tree, forest, or connected.