arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

6952 papers tagged with this keyword
2023-10-12
Algebraic Connectivity Characterization of Ensemble Random Hypergraphs
Random hypergraph is a broad concept used to describe probability distributions over hypergraphs, which are mathematical structures with applications in various fields, e.g., complex systems in physics, computer science, social sciences, and network science. Ensemble methods, on the other hand, are crucial both in physics and machine learning. In physics, ensemble theory helps bridge the gap between the microscopic and macroscopic worlds, providing a statistical framework for understanding systems with a vast number of particles. In machine learning, ensemble methods are valuable because they improve predictive accuracy, reduce overfitting, lower prediction variance, mitigate bias, and capture complex relationships in data. However, there is limited research on applying ensemble methods to a set of random hypergraphs. This work aims to study the connectivity behavior of an ensemble of random hypergraphs. Specifically, it focuses on quantifying the random behavior of the algebraic connectivity of these ensembles through tail bounds. We utilize Laplacian tensors to represent these ensemble random hypergraphs and establish mathematical theorems, such as Courant-Fischer and Lieb-Seiringer theorems for tensors, to derive tail bounds for the algebraic connectivity. We derive three different tail bounds, i.e., Chernoff, Bennett, and Bernstein bounds, for the algebraic connectivity of ensemble hypergraphs with respect to different random hypergraphs assumptions.
2023-10-11
Sampling triangulations of manifolds using Monte Carlo methods
We propose a Monte Carlo method to efficiently find, count, and sample abstract triangulations of a given manifold M. The method is based on a biased random walk through all possible triangulations of M (in the Pachner graph), constructed by combining (bi-stellar) moves with suitable chosen accept/reject probabilities (Metropolis-Hastings). Asymptotically, the method guarantees that samples of triangulations are drawn at random from a chosen probability. This enables us not only to sample (rare) triangulations of particular interest but also to estimate the (extremely small) probability of obtaining them when isomorphism types of triangulations are sampled uniformly at random. We implement our general method for surface triangulations and 1-vertex triangulations of 3-manifolds. To showcase its usefulness, we present a number of experiments: (a) we recover asymptotic growth rates for the number of isomorphism types of simplicial triangulations of the 2-dimensional sphere; (b) we experimentally observe that the growth rate for the number of isomorphism types of 1-vertex triangulations of the 3-dimensional sphere appears to be singly exponential in the number of their tetrahedra; and (c) we present experimental evidence that a randomly chosen isomorphism type of 1-vertex n-tetrahedra 3-sphere triangulation, for n tending to infinity, almost surely shows a fixed edge-degree distribution which decays exponentially for large degrees, but shows non-monotonic behaviour for small degrees.
2023-10-10 v2
Finding cliques and dense subgraphs using edge queries
We consider the problem of finding a large clique in an Erdős--Rényi random graph where we are allowed unbounded computational time but can only query a limited number of edges. Recall that the largest clique in $G \sim G(n,1/2)$ has size roughly $2\log_{2} n$. Let $α_{\star}(δ,\ell)$ be the supremum over $α$ such that there exists an algorithm that makes $n^δ$ queries in total to the adjacency matrix of $G$, in a constant $\ell$ number of rounds, and outputs a clique of size $α\log_{2} n$ with high probability. We give improved upper bounds on $α_{\star}(δ,\ell)$ for every $δ\in [1,2)$ and $\ell \geq 3$. We also study analogous questions for finding subgraphs with density at least $η$ for a given $η$, and prove corresponding impossibility results.
Combinatorics of pruned Hurwitz numbers
Hurwitz numbers enumerate branched morphisms between Riemannn surfaces with fixed numerical data. They represent important objects in enumerative geometry that are accessible by combinatorial techniques. In the past decade, many variants of Hurwitz numbers have appeared in the literature. In this paper, we focus on an exciting such variant that arises naturally from the theory of topological recursion: Pruned Hurwitz numbers. These are defined as an enumeration of a relevant subset of branched morphisms between Riemann surfaces, that yield smaller numbers than their classical counterparts while retaining maximal information. Thus, pruned Hurwitz numbers may be viewed as the core of the Hurwitz problem. In this paper, we develop the combinatorial theory of pruned Hurwitz numbers. In particular, motivated by the successful application of combinatorial techniques to classical Hurwitz numbers, we derive two new combinatorial expressions of pruned Hurwitz numbers. Firstly, we show that they may be expressed in terms of Hurwitz mobiles which are tree-like structure that arise from the theory of random planar maps. Secondly, we prove a tropical correspondence theorem which allows the enumeration of pruned Hurwitz numbers in terms of tropical covers.
2023-10-06 v2
On rainbow thresholds
Resolving a recent problem of Bell, Frieze, and Marbach, we establish both the threshold result of Frankston--Kahn--Narayanan--Park, and its strengthening by Spiro, in the rainbow setting. This has applications to the thresholds for rainbow structures in random graphs where each edge is given a uniformly random color from a set of given colors.
2023-10-06
Asymptotic distribution of degree--based topological indices
Published in MATCH Commun. Math. Comput. Chem. 2023 • Search Publication
Topological indices play a significant role in mathematical chemistry. Given a graph $\mathcal{G}$ with vertex set $\mathcal{V}=\{1,2,\dots,n\}$ and edge set $\mathcal{E}$, let $d_i$ be the degree of node $i$. The degree-based topological index is defined as $\mathcal{I}_n=$ $\sum_{\{i,j\}\in \mathcal{E}}f(d_i,d_j)$, where $f(x,y)$ is a symmetric function. In this paper, we investigate the asymptotic distribution of the degree-based topological indices of a heterogeneous Erdős-Rényi random graph. We show that after suitably centered and scaled, the topological indices converges in distribution to the standard normal distribution. Interestingly, we find that the general Randić index with $f(x,y)=(xy)^τ$ for a constant $τ$ exhibits a phase change at $τ=-\frac{1}{2}$.
2023-10-05
Periodic $q$-Whittaker and Hall-Littlewood processes
We study the periodic $q$-Whittaker and Hall-Littlewood processes, two probability measures on sequences of partitions. We prove that a certain observable of the periodic $q$-Whittaker process exhibits a $(q,u)$ symmetry after a random shift, generalizing a previous result of Imamura, Mucciconi, and Sasamoto who showed a matching between the periodic Schur and $q$-Whittaker measures, and also give a vertex model formulation of their result. As part of our proof of the $(q,u)$ symmetry, we obtain contour integral formulas for both the periodic $q$-Whittaker and Hall-Littlewood processes. We also show a matching between certain observables in the periodic Hall-Littlewood process and in a quasi-periodic stochastic six vertex model after a suitable random shift, and discuss a limit to the stationary periodic stochastic six vertex model.
Scaling limit of the cluster size distribution for the random current measure on the complete graph
We study the percolation configuration arising from the random current representation of the near-critical Ising model on the complete graph. We compute the scaling limit of the cluster size distribution for an arbitrary set of sources in the single and the double current measures. As a byproduct, we compute the tangling probabilities recently introduced by Gunaratnam, Panagiotis, Panis, and Severo in [GPPS22]. This provides a new perspective on the switching lemma for the $\varphi^4$ model introduced in the same paper: in the Gaussian limit we recover Wick's law, while in the Ising limit we recover the corresponding tool for the Ising model.
2023-10-02 v2
Subgraph densities and scaling limits of random graphs with a prescribed modular decomposition
We consider large uniform labeled random graphs in different classes with prescribed decorations in their modular decomposition. Our main result is the estimation of the number of copies of every graph as an induced subgraph. As a consequence, we obtain the convergence of a uniform random graph in such classes to a Brownian limit object in the space of graphons. Our proofs rely on combinatorial arguments, computing generating series using the symbolic method and deriving asymptotics using singularity analysis.
2023-09-30 v2
The Lovász Theta Function for Recovering Planted Clique Covers and Graph Colorings
The problems of computing graph colorings and clique covers are central challenges in combinatorial optimization. Both of these are known to be NP-hard, and thus computationally intractable in the worst-case instance. A prominent approach for computing approximate solutions to these problems is the celebrated Lovász theta function $\vartheta(G)$, which is specified as the solution of a semidefinite program (SDP), and hence tractable to compute. In this work, we move beyond the worst-case analysis and set out to understand whether the Lovász theta function recovers clique covers for random instances that have a latent clique cover structure, possibly obscured by noise. We answer this question in the affirmative and show that for graphs generated from the planted clique model we introduce in this work, the SDP formulation of $\vartheta(G)$ has a unique solution that reveals the underlying clique-cover structure with high-probability. The main technical step is an intermediate result where we prove a deterministic condition of recovery based on an appropriate notion of sparsity.
2023-09-29 v2
Joint extremes of inversions and descents of random permutations
We provide asymptotic theory for the joint distribution of $X_{\mathrm{inv}}$ and $X_{\mathrm{des}}$, the numbers of inversions and descents of random permutations. Recently, Dörr & Kahle (2022) proved that $X_{\mathrm{inv}}$, respectively, $X_{\mathrm{des}}$ is in the maximum domain of attraction of the Gumbel distribution. To tackle the dependency between these two permutation statistics, we use Hájek projections and a suitable quantitative Gaussian approximation. We show that $(X_{\mathrm{inv}}, X_{\mathrm{des}})$ is in the maximum domain of attraction of the two-dimensional Gumbel distribution with independent margins. This result can be stated in the broader combinatorial framework of finite Coxeter groups, on which our method also yields the central limit theorem for $(X_{\mathrm{inv}}, X_{\mathrm{des}})$ and various other permutation statistics as a novel contribution. In particular, signed permutation groups with random biased signs and products of classical Weyl groups are investigated.
2023-09-29 v2
Pólya urns on hypergraphs
We study Pólya urns on hypergraphs and prove that, when the incidence matrix of the hypergraph is injective, there exists a point $v=v(H)$ such that the random process converges to $v$ almost surely. We also provide a partial result when the incidence matrix is not injective.
Resilience for Loose Hamilton Cycles
We study the emergence of loose Hamilton cycles in subgraphs of random hypergraphs. Our main result states that the minimum $d$-degree threshold for loose Hamiltonicity relative to the random $k$-uniform hypergraph $H_k(n,p)$ coincides with its dense analogue whenever $p \geq n^{- (k-1)/2+o(1)}$. The value of $p$ is approximately tight for $d>(k+1)/2$. This is particularly interesting because the dense threshold itself is not known beyond the cases when $d \geq k-2$.
2023-09-24
Probabilistic Bounds for Data Storage with Feature Selection and Undersampling
In this paper we consider data storage from a probabilistic point of view and obtain bounds for efficient storage in the presence of feature selection and undersampling, both of which are important from the data science perspective. First, we consider encoding of correlated sources for nonstationary data and obtain a Slepian-Wolf type result for the probability of error. We then reinterpret our result by allowing one source to be the set of features to be discarded and other source to be remaining data to be encoded. Next, we consider neighbourhood domination in random graphs where we impose the condition that a fraction of neighbourhood must be present for each vertex and obtain optimal bounds on the minimum size of such a set. We show how such sets are useful for data undersampling in the presence of imbalanced datasets and briefly illustrate our result using~\(k-\)nearest neighbours type classification rules as an example.
2023-09-23 v3
Runs in Random Sequences over Ordered Sets
Published • View PublicationBIB
We determine the distributions of lengths of runs in random sequences of elements from a totally ordered set (total order) or partially ordered set (partial order). In particular, we produce novel formulae for the expected value, variance, and probability generating function (PGF) of such lengths in the case of an arbitrary total order. Our focus is on the case of distributions with both atoms and diffuse (absolutely or singularly continuous) mass which has not been addressed in this generality before. We also provide a method of calculating the PGF of run lengths for countably series-parallel partial orders. Additionally, we prove a strong law of large numbers for the distribution of run lengths in a particular realization of an infinite sequence.
2023-09-22 v2
Robust Hamiltonicity in families of Dirac graphs
A graph is called Dirac if its minimum degree is at least half of the number of vertices in it. Joos and Kim showed that every collection $\mathbb{G}=\{G_1,\ldots,G_n\}$ of Dirac graphs on the same vertex set $V$ of size $n$ contains a Hamilton cycle transversal, i.e., a Hamilton cycle $H$ on $V$ with a bijection $φ:E(H)\rightarrow [n]$ such that $e\in G_{φ(e)}$ for every $e\in E(H)$. In this paper, we determine up to a multiplicative constant, the threshold for the existence of a Hamilton cycle transversal in a collection of random subgraphs of Dirac graphs in various settings. Our proofs rely on constructing a spread measure on the set of Hamilton cycle transversals of a family of Dirac graphs. As a corollary, we obtain that every collection of $n$ Dirac graphs on $n$ vertices contains at least $(cn)^{2n}$ different Hamilton cycle transversals $(H,φ)$ for some absolute constant $c>0$. This is optimal up to the constant $c$. Finally, we show that if $n$ is sufficiently large, then every such collection spans $n/2$ pairwise edge-disjoint Hamilton cycle transversals, and this is best possible. These statements generalize classical counting results of Hamilton cycles in a single Dirac graph.
2023-09-22 v2
Sidorenko Hypergraphs and Random Turán Numbers
Let $\mathrm{ex}(G_{n,p}^r,F)$ denote the maximum number of edges in an $F$-free subgraph of the random $r$-uniform hypergraph $G_{n,p}^r$, and let $s(F):=\sup\{s: \exists H,\ t_F(H)=t_{K_r^r}(H)^{s+e(F)}>0\}$. Following recent work of Conlon, Lee, and Sidorenko, we prove non-trivial lower bounds on $\mathrm{ex}(G_{n,p}^r,F)$ whenever $s(F)>0$, i.e. $F$ is not Sidorenko. This connection between Sidorenko's conjecture and random Turán problems gives new lower bounds on $\mathrm{ex}(G_{n,p}^r,F)$ whenever $s(F)>0$, and further allows us to establish upper bounds for $s(F)$ whenever upper bounds for $\mathrm{ex}(G_{n,p}^r,F)$ are known. As a consequence, we prove that $s(\mathrm{E}^r(K_{k+1}^k))=\frac{1}{r-k}$ where $\mathrm{E}^r(K_{k+1}^k)$ is the $r$-expansion of $K_{k+1}^k$.
2023-09-22 v3
Zero-One Laws for Random Feasibility Problems
We introduce a general random model of a combinatorial optimization problem with geometric structure that encapsulates both linear programming and integer linear programming. Let $Q$ be a bounded set called the feasible set, $E$ be an arbitrary set called the constraint set, and $A$ be a random linear transform. We define and study the $\ell^q$-margin, $M_q := d_q(AQ, E)$. The margin quantifies the feasibility of finding $y \in AQ$ satisfying the constraint $y \in E$. Our contribution is to establish strong concentration of the margin for any $q \in (2,\infty]$, assuming only that $E$ has permutation symmetry. The case of $q = \infty$ is of particular interest in applications -- specifically to combinatorial ``balancing'' problems -- and is markedly out of the reach of the classical isoperimetric and concentration-of-measure tools that suffice for $q \le 2$. Generality is a key feature of this result: we assume permutation symmetry of the constraint set and nothing else. This allows us to encode many optimization problems in terms of the margin, including random versions of: the closest vector problem, integer linear feasibility, perceptron-type problems, $\ell^q$-combinatorial discrepancy for $2 \le q \le \infty$, and matrix balancing. Concentration of the margin implies a host of new sharp threshold results in these models, and also greatly simplifies and extends some key known results.
Four universal growth regimes in degree-dependent first passage percolation on spatial random graphs I
One-dependent first passage percolation is a spreading process on a graph where the transmission time through each edge depends on the direct surroundings of the edge. In particular, the classical iid transmission time $L_{xy}$ is multiplied by $(W_xW_y)^μ$, a polynomial of the expected degrees $W_x, W_y$ of the endpoints of the edge $xy$, which we call the penalty function. Beyond the Markov case, we also allow any distribution for $L_{xy}$ with regularly varying distribution near $0$. We then run this process on three spatial scale-free random graph models: finite and infinite Geometric Inhomogeneous Random Graphs, and Scale-Free Percolation. In these spatial models, the connection probability between two vertices depends on their spatial distance and on their expected degrees. We show that as the penalty-function, i.e., $μ$ increases, the transmission time between two far away vertices sweeps through four universal phases: explosive (with tight transmission times), polylogarithmic, polynomial but strictly sublinear, and linear in the Euclidean distance. The strictly polynomial growth phase here is a new phenomenon that so far was extremely rare in spatial graph models. The four growth phases are highly robust in the model parameters and are not restricted to phase boundaries. Further, the transition points between the phases depend non-trivially on the main model parameters: the tail of the degree distribution, a long-range parameter governing the presence of long edges, and the behaviour of the distribution $L$ near $0$. In this paper we develop new methods to prove the upper bounds in all sub-explosive phases. Our companion paper complements these results by providing matching lower bounds in the polynomial and linear regimes.
Polynomial growth in degree-dependent first passage percolation on spatial random graphs
In this paper we study a version of (non-Markovian) first passage percolation on graphs, where the transmission time between two connected vertices is non-iid, but increases by a penalty factor polynomial in their expected degrees. Based on the exponent of the penalty-polynomial, this makes it increasingly harder to transmit to and from high-degree vertices. This choice is motivated by awareness or time-limitations. For the iid part of the transmission times we allow any nonnegative distribution with regularly varying behaviour at $0$. For the underlying graph models we choose spatial random graphs that have power-law degree distributions, so that the effect of the penalisation becomes visible: (finite and infinite) Geometric Inhomogeneous Random Graphs, and Scale-Free Percolation. In these spatial models, the connection probability between two vertices depends on their spatial distance and on their expected degrees. We prove that upon increasing the penalty exponent, the transmission time between two far away vertices $x,y$ sweeps through four universal phases even for a single underlying graph: explosive (tight transmission times), polylogarithmic, polynomial but sublinear ($|x-y|^{η_0+o(1)}$ for an explicit $η_0<1$), and linear ($Θ(|x-y|)$) in their Euclidean distance. Further, none of these phases are restricted to phase boundaries, and those are non-trivial in the main model parameters: the tail of the degree-distribution, a long-range parameter, and the exponent of regular variation of the iid part of the transmission times. In this paper we present proofs of lower bounds for the latter two phases and the upper bound for the linear phase. These complement the matching upper bounds for the polynomial regime in our companion paper.