arXiv++ Combinatorics

Browse math.CO papers from arXiv

statistical model

46 papers tagged with this keyword
Torus Actions on Matrix Schubert and Kazhdan-Lusztig Varieties, and their Links to Statistical Models
We investigate the toric geometry of two families of generalised determinantal varieties arising from permutations: Matrix Schubert varieties ($\overline{X_w}$) and Kazhdan-Lusztig varieties ($\mathcal{N}_{v,w}$). Matrix Schubert varieties can be written as $\overline{X_w} = Y_w \times \mathbb C^d$, where $d$ is maximal. We are especially interested in the structure and complexity of these varieties $Y_w$ and $\mathcal{N}_{v,w}$ under the so-called usual torus actions. In the case when $Y_w$ is toric, we provide a full characterisation of the simple reflections $s_i$ that render ${Y_{w \cdot s_i}}$ toric, as well as the corresponding changes to the weight cone. For Kazhdan-Lusztig varieties, we consider how moving one of the two permutations $v,w$ along a chain in the Bruhat poset affects their complexity. Additionally, we study the complexity of these varieties, for permutations $v$ and $w$ of a specific structure. Finally, we consider the links between these determinantal varieties and two classes of statistical models; namely conditional independence and quasi-independence models.
Causal Models for Growing Networks
Real-world networks grow over time; statistical models based on node exchangeability are not appropriate. Instead of constraining the structure of the \textit{distribution} of edges, we propose that the relevant symmetries refer to the \textit{causal structure} between them. We first enumerate the 96 causal directed acyclic graph (DAG) models over pairs of nodes (dyad variables) in a growing network with finite ancestral sets that are invariant to node deletion. We then partition them into 21 classes with ancestral sets that are closed under node marginalization. Several of these classes are remarkably amenable to distributed and asynchronous evaluation. As an example, we highlight a simple model that exhibits flexible power-law degree distributions and emergent phase transitions in sparsity, which we characterize analytically. With few parameters and much conditional independence, our proposed framework provides natural baseline models for causal inference in relational data.
2025-02-18 v2
Learning the symmetric group: large from small
Machine learning explorations can make significant inroads into solving difficult problems in pure mathematics. One advantage of this approach is that mathematical datasets do not suffer from noise, but a challenge is the amount of data required to train these models and that this data can be computationally expensive to generate. Key challenges further comprise difficulty in a posteriori interpretation of statistical models and the implementation of deep and abstract mathematical problems. We propose a method for scalable tasks, by which models trained on simpler versions of a task can then generalize to the full task. Specifically, we demonstrate that a transformer neural-network trained on predicting permutations from words formed by general transpositions in the symmetric group $S_{10}$ can generalize to the symmetric group $S_{25}$ with near 100\% accuracy. We also show that $S_{10}$ generalizes to $S_{16}$ with similar performance if we only use adjacent transpositions. We employ identity augmentation as a key tool to manage variable word lengths, and partitioned windows for training on adjacent transpositions. Finally we compare variations of the method used and discuss potential challenges with extending the method to other tasks.
2025-01-15 v3
Fermions and Zeta Function on the Graph
We propose a novel fermionic model on the graphs. The Dirac operator of the model consists of deformed incidence matrices on the graph and the partition function is given by the inverse of the graph zeta function. We find that the coefficients of the inverse of the graph zeta function, which is a polynomial of finite degree in the coupling constant, count the number of fermionic cycles on the graph. We also construct the model on grid graphs by using the concept of the covering graph and the Artin-Ihara $L$-function. In connection with this, we show that the fermion doubling is absent, and the overlap fermions can be constructed on a general graph. Furthermore, we relate our model to statistical models by introducing the winding number around cycles, where the distribution of the poles of the graph zeta function (the zeros of the partition function) plays a crucial role. Finally, we formulate gauge theory including fermions on the graph from the viewpoint of the covering graph derived from the gauge group in a unified way.
2024-05-30
Phylogenetic degrees for Jukes-Cantor model
Jukes-Cantor model is one of the most meaningful statistical models from a biological perspective. We are interested in computing the algebraic degrees for phylogenetic varieties, which we call phylogenetic degrees, associated to the Jukes-Cantor model and any tree. As these varieties are toric, their geometry is hidden in the associated polytopes. For this reason, we provide two different combinatorial approaches to compute the volume for these polytopes.
Geometry of rational quasi-independence models as toric fiber products
Published in Alg. Stat. 17 (2026) 1-31 • View PublicationBIB
We investigate the geometry of a family of log-linear statistical models called quasi-independence models. The toric fiber product is useful for understanding the geometry of parameter inference in these models because the maximum likelihood degree is multiplicative under the TFP. We define the coordinate toric fiber product, or cTFP, and give necessary and sufficient conditions under which a quasi-independence model is a cTFP of lower-order models. We show that the vanishing ideal of every 2-way quasi-independence model with ML-degree 1 can be realized as an iterated toric fiber product of linear ideals. We also classify which Lawrence lifts of 2-way quasi-independence models are cTFPs and give a necessary condition under which a $k$-way model has ML-degree 1 using its facial submodels.
Irreducible Markov Chains on spaces of graphs with fixed degree-color sequences
We study a colored generalization of the famous simple-switch Markov chain for sampling the set of graphs with a fixed degree sequence. Here we consider the space of graphs with colored vertices, in which we fix the degree sequence and another statistic arising from the vertex coloring, and prove that the set can be connected with simple color-preserving switches or moves. These moves form a basis for defining an irreducible Markov chain necessary for testing statistical model fit to block-partitioned network data. Our methods further generalize well-known algebraic results from the 1990s: namely, that the corresponding moves can be used to construct a regular triangulation for a generalization of the second hypersimplex. On the other hand, in contrast to the monochromatic case, we show that for simple graphs, the 1-norm of the moves necessary to connect the space increases with the number of colors.
2024-02-08 v3
Homaloidal Polynomials and Gaussian Models of Maximum Likelihood Degree One
Published in Alg. Stat. 15 (2024) 167-198 • View PublicationBIB
We study the Gaussian statistical models whose log-likelihood function has a unique complex critical point, i.e., has maximum likelihood degree one. We exploit the connection developed by Améndola et. al. between the models having maximum likelihood degree one and homaloidal polynomials. We study the spanning tree generating function of a graph and show this polynomial is homaloidal when the graph is chordal. When the graph is a cycle on $n$ vertices, $n \geq 4$, we prove the polynomial is not homaloidal, and show that the maximum likelihood degree of the resulting model is the $n$th Eulerian number. These results support our conjecture that the spanning tree generating function is a homaloidal polynomial if and only if the graph is chordal. We also provide an algebraic formulation for the defining equations of these models. Using existing results, we provide a computational study on constructing new families of homaloidal polynomials. In the end, we analyze the symmetric determinantal representation of such polynomials and provide an upper bound on the size of the matrices involved.
2023-11-22 v2
Likelihood Geometry of Reflexive Polytopes
Published in Alg. Stat. 15 (2024) 113-143 • View PublicationBIB
We study the problem of maximum likelihood (ML) estimation for statistical models defined by reflexive polytopes. Our focus is on the maximum likelihood degree of these models as an algebraic measure of complexity of the corresponding optimization problem. We compute the ML degrees of all 4319 classes of three-dimensional reflexive polytopes, and observe some surprising behavior in terms of the presence of gaps between ML degrees and degrees of the associated toric varieties. We interpret these drops in the context of discriminants and prove formulas for the ML degree for families of reflexive polytopes, including the hypercube and its dual, the cross polytope, in arbitrary dimension. In particular, we determine a family of embeddings for the $d$-cube that implies ML degree one. Finally, we discuss generalized constructions of families of reflexive polytopes in terms of their ML degrees.
2023-10-17
The Codegree, Weak Maximum Likelihood Threshold, and the Gorenstein Property of Hierarchical Models
Published in Alg. Stat. 16 (2025) 201-215 • View PublicationBIB
The codegree of a lattice polytope is the smallest integer dilate that contains a lattice point in the relative interior. The weak maximum likelihood threshold of a statistical model is the smallest number of data points for which there is a non-zero probability that the maximum likelihood estimate exists. The codegree of a marginal polytope is a lower bound on the maximum likelihood threshold of the associated log-linear model, and they are equal when the marginal polytope is normal. We prove a lower bound on the codegree in the case of hierarchical log-linear models and provide a conjectural formula for the codegree in general. As an application, we study when the marginal polytopes of hierarchical models are Gorenstein, including a classification of Gorenstein decomposable models, and a conjectural classification of Gorenstein binary hierarchical models.
2023-04-24
The Theory of Gene Family Histories
Most genes are part of larger families of evolutionary related genes. The history of gene families typically involves duplications and losses of genes as well as horizontal transfers into other organisms. The reconstruction of detailed gene family histories, i.e., the precise dating of evolutionary events relative to phylogenetic tree of the underlying species has remained a challenging topic despite their importance as a basis for detailed investigations into adaptation and functional evolution of individual members of the gene family. The identification of orthologs, moreover, is a particularly important subproblem of the more general setting considered here. In the last few years, an extensive body of mathematical results has appeared that tightly links orthology, a formal notion of best matches among genes, and horizontal gene transfer. The purpose of this chapter is the broadly outline some of the key mathematical insights and to discuss their implication for practical applications. In particular, we focus on tree-free methods, i.e., methods to infer orthology or horizontal gene transfer as well as gene trees, species trees and reconciliations between them without using \emph{a priori} knowledge of the underlying trees or statistical models for the inference of phylogenetic trees. Instead, the initial step aims to extract binary relations among genes.
Toric Fiber Products in Geometric Modeling
An important challenge in Geometric Modeling is to classify polytopes with rational linear precision. Equivalently, in Algebraic Statistics one is interested in classifying scaled toric varieties, also known as discrete exponential families, for which the maximum likelihood estimator can be written in closed form as a rational function of the data (rational MLE). The toric fiber product (TFP) of statistical models is an operation to iteratively construct new models with rational MLE from lower dimensional ones. In this paper we introduce TFPs to the Geometric Modeling setting to construct polytopes with rational linear precision and give explicit formulae for their blending functions. A special case of the TFP is taking the Cartesian product of two polytopes and their blending functions. The Horn matrix of a statistical model with rational MLE is a key player in both Geometric Modeling and Algebraic Statistics; it proved to be fruitful providing a characterisation of those polytopes having the more restrictive property of strict linear precision. We give an explicit description of the Horn matrix of a TFP.
One-connection rule for structural equation models
Linear structural equation models are multivariate statistical models encoded by mixed graphs. In particular, the set of covariance matrices for distributions belonging to a linear structural equation model for a fixed mixed graph $G=(V, D,B)$ is parameterized by a rational function with parameters for each vertex and edge in $G$. This rational parametrization naturally allows for the study of these models from an algebraic and combinatorial point of view. Indeed, this point of view has led to a collection of results in the literature, mainly focusing on questions related to identifiability and determining relationships between covariances (i.e., finding polynomials in the Gaussian vanishing ideal). So far, a large proportion of these results has focused on the case when $D$, the directed part of the mixed graph $G$, is acyclic. This is due to the fact that in the acyclic case, the parametrization becomes polynomial and there is a description of the entries of the covariance matrices in terms of a finite sum. We move beyond the acyclic case and give a closed form expression for the entries of the covariance matrices in terms of the one-connections in a graph obtained from $D$ through some small operations. This closed form expression then allows us to show that if $G$ is simple, then the parametrization map is generically finite-to-one. Finally, having a closed form expression for the covariance matrices allows for the development of an algorithm for systematically exploring possible polynomials in the Gaussian vanishing ideal.
2022-05-19 v3
Classifying one-dimensional discrete models with maximum likelihood degree one
Published • View PublicationBIB
We propose a classification of all one-dimensional discrete statistical models with maximum likelihood degree one based on their rational parametrization. We show how all such models can be constructed from members of a smaller class of 'fundamental models' using a finite number of simple operations. We introduce 'chipsplitting games', a class of combinatorial games on a grid which we use to represent fundamental models. This combinatorial perspective enables us to show that there are only finitely many fundamental models in the probability simplex $Δ_n$ for $n\leq 4$.
2022-02-04 v2
Rearrangement Events on Circular Genomes
Published • View PublicationBIB
Early literature on genome rearrangement modelling views the problem of computing evolutionary distances as an inherently combinatorial one. In particular, attention was given to estimating distances using the minimum number of events required to transform one genome into another. In hindsight, this approach is analogous to early methods for inferring phylogenetic trees from DNA sequences such as maximum parsimony -- both are motivated by the principle that the true distance minimises evolutionary change, and both are effective if this principle is a true reflection of reality. Recent literature considers genome rearrangement under statistical models, continuing this parallel with DNA-based methods; the goal here is to use model-based methods (for example maximum likelihood techniques) to compute distance estimates that incorporate the large number of rearrangement paths that can transform one genome into another. Crucially, this approach requires one to decide upon a set of feasible rearrangement events and, in this paper, we focus on characterising well-motivated models for signed, uni-chromosomal circular genomes, where the number of regions remains fixed. Since rearrangements are often mathematically described using permutations, we isolate the sets of permutations representing rearrangements that are biologically reasonable in this context, for example inversions and translocations. We provide precise mathematical expressions for these rearrangements, and then describe them in terms of the set of cuts made in the genome when they are applied. We directly compare cuts to breakpoints, and use this concept to count the distinct rearrangement actions which apply a given number of cuts. Finally, we provide some examples of rearrangement models, and include a discussion of some questions that arise when defining plausible models.
2021-12-29 v4
Logarithmic Voronoi polytopes for discrete linear models
Published in Alg. Stat. 15 (2024) 1-13 • View PublicationBIB
We study logarithmic Voronoi cells for linear statistical models and partial linear models. The logarithmic Voronoi cells at points on such model are polytopes. To any $d$-dimensional linear model inside the probability simplex $Δ_{n-1}$, we can associate an $n\times d$ matrix $B$. For interior points, we describe the vertices of these polytopes in terms of co-circuits of $B$. We also show that these polytopes are combinatorially isomorphic to the dual of a vector configuration with Gale diagram $B$. This means that logarithmic Voronoi cells at all interior points on a linear model have the same combinatorial type. We also describe logarithmic Voronoi cells at points on the boundary of the simplex. Finally, we study logarithmic Voronoi cells of partial linear models, where the points on the boundary of the model are especially of interest.
2021-11-26 v4
Perturbing Isoradial Triangulations
Published • View PublicationBIB
We consider an infinite, planar, Delaunay graph which is obtained by locally deforming the embedding of a general, isoradial graph, w.r.t. a real deformation parameter $ε$. This entails a careful analysis of edge-flips induced by the deformation and the Delaunay constraints. Using Kenyon's exact and asymptotic results for the Green's function on an isoradial graph, we calculate the leading asymptotics of the first and second order terms in the perturbative expansion of the log-determinant of the Beltrami-Laplace operator $Δ(ε)$, the David-Eynard Kähler operator $\mathcal{D}(ε)$, and the conformal Laplacian $\underlineΔ(ε)$ on the deformed graph. We show that the scaling limits of the second order {\it bi-local} term for both the Beltrami-Laplace and David-Eynard operators exist and coincide, with a value independent of the choice of initial isoradial graph. Our results allow to define a discrete analogue of the stress energy tensor for each of the three operators. Furthermore we can identify a central charge ($c$) in the case of both the Beltrami-Laplace and David-Eynard operators. While the scaling limit is consistent with the stress-energy tensor and value of the central charge for the Gaussian free field (GFF), the discrete central charge value of $c=-2$ for the David-Eynard operator is, however, at odds with the value of $c=-26$ expected by Polyakov's theory of 2D quantum gravity; moreover there are problems with convergence of the scaling limit of the discrete stress energy tensor for the David-Eynard operator. The bi-local term for the conformal Laplacian involves anomalous terms corresponding to the creation of discrete {\it curvature dipoles} in the deformed Delaunay graph; we examine the difficulties in defining a convergent scaling limit in this case. Connections with some discrete statistical models at criticality are explored.
Families of polytopes with rational linear precision in higher dimensions
Published in Foundations of Computational Mathematics, 2022 • View PublicationBIB
In this article we introduce a new family of lattice polytopes with rational linear precision. For this purpose, we define a new class of discrete statistical models that we call multinomial staged tree models. We prove that these models have rational maximum likelihood estimators (MLE) and give a criterion for these models to be log-linear. Our main result is then obtained by applying Garcia-Puente and Sottile's theorem that establishes a correspondence between polytopes with rational linear precision and log-linear models with rational MLE. Throughout this article we also study the interplay between the primitive collections of the normal fan of a polytope with rational linear precision and the shape of the Horn matrix of its corresponding statistical model. Finally, we investigate lattice polytopes arising from toric multinomial staged tree models, in terms of the combinatorics of their tree representations.
An Optimal Algorithm for Strict Circular Seriation
Published • View PublicationBIB
We study the problem of circular seriation, where we are given a matrix of pairwise dissimilarities between $n$ objects, and the goal is to find a {\em circular order} of the objects in a manner that is consistent with their dissimilarity. This problem is a generalization of the classical {\em linear seriation} problem where the goal is to find a {\em linear order}, and for which optimal ${\cal O}(n^2)$ algorithms are known. Our contributions can be summarized as follows. First, we introduce {\em circular Robinson matrices} as the natural class of dissimilarity matrices for the circular seriation problem. Second, for the case of {\em strict circular Robinson dissimilarity matrices} we provide an optimal ${\cal O}(n^2)$ algorithm for the circular seriation problem. Finally, we propose a statistical model to analyze the well-posedness of the circular seriation problem for large $n$. In particular, we establish ${\cal O}(\log(n)/n)$ rates on the distance between any circular ordering found by solving the circular seriation problem to the underlying order of the model, in the Kendall-tau metric.
2021-04-19 v5
Tree Topologies along a Tropical Line Segment
Published • View PublicationBIB
Tropical geometry with the max-plus algebra has been applied to statistical learning models over tree spaces because geometry with the tropical metric over tree spaces has some nice properties such as convexity in terms of the tropical metric. One of the challenges in applications of tropical geometry to tree spaces is the difficulty interpreting outcomes of statistical models with the tropical metric. This paper focuses on combinatorics of tree topologies along a tropical line segment, an intrinsic geodesic with the tropical metric, between two phylogenetic trees over the tree space and we show some properties of a tropical line segment between two trees. Specifically we show that a probability of a tropical line segment of two randomly chosen trees going through the origin (the star tree) is zero if the number of leave is greater than four, and we also show that if two given trees differ only one nearest neighbor interchange (NNI) move, then the tree topology of a tree in the tropical line segment between them is the same tree topology of one of these given two trees with possible zero branch lengths.