arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

6952 papers tagged with this keyword
Homogeneous substructures in random ordered uniform matchings
An ordered $r$-uniform matching of size $n$ is a collection of $n$ pairwise disjoint $r$-subsets of a linearly ordered set of $rn$ vertices. For $n=2$, such a matching is called an $r$-pattern, as it represents one of $\tfrac12\binom{2r}r$ ways two disjoint edges may intertwine. Given a set $\mathcal{P}$ of $r$-patterns, a $\mathcal{P}$-clique is a matching with all pairs of edges belonging to $\mathcal{P}$. In this paper we determine the order of magnitude of the size of a largest $\mathcal{P}$-clique in a random ordered $r$-uniform matching for several sets $\mathcal{P}$, including all sets of size $|\mathcal{P}|\le2$ and the set $\mathcal{R}^{(r)}$ of all $2^{r-1}$ $r$-partite $r$-patterns.
2026-01-20
Wasserstein distances between ERGMs and Erdős-Rényi models
Ferromagnetic exponential random graph models (ERGMs) are random graph models under which the presence of certain small structures (such as triangles) is encouraged; they can be constructed by tilting an Erdős--Rényi model by the exponential of a particular nonlinear Hamiltonian. These models are mixtures of metastable wells which each behave macroscopically like an Erdős--Rényi model, exhibiting the same laws of large numbers for subgraph counts [CD13]. However, on the microscopic scale these metastable wells are very different from Erdős--Rényi models, with the total variation distance between the two measures tending to 1 [MX23]. In this article we clarify this situation by providing a sharp (up to constants) bound on the Hamming-Wasserstein distance between the two models, which is the average number of edges at which they differ, under the coupling which minimizes this average. In particular, we show that this distance is $Θ(n^{3/2})$, quantifying exactly how these models differ. An upper bound of this form has appeared in the past [RR19], but this was restricted to the subcritical (high-temperature) regime of parameters. We extend this bound, using a new proof technique, to the supercritical (low-temperature) regime, and prove a matching lower bound which has only previously appeared in the subcritical regime of special cases of ERGMs satisfying a "triangle-free" condition [DF25]. To prove the lower bound in the presence of triangles, we introduce an approximation of the discrete derivative of the Hamiltonian, which controls the dynamical properties of the ERGM, in terms of local counts of triangles and wedges (two-stars) near an edge. This approximation is the main technical and conceptual contribution of the article, and we expect it will be useful in a variety of other contexts as well. Along the way, we also prove a bound on the marginal edge probability under the ERGM via a new bootstrapping argument. Such a bound has already appeared [FLSW25], but again only in the subcritical regime and using a different proof strategy.
2026-01-19
Explicit Entropic Constructions for Coverage, Facility Location, and Graph Cuts
Shannon entropy is a polymatroidal set function and lies at the foundation of information theory, yet the class of entropic polymatroids is strictly smaller than the class of all submodular functions. In parallel, submodular and combinatorial information measures (SIMs) have recently been proposed as a principled framework for extending entropy, mutual information, and conditional mutual information to general submodular functions, and have been used extensively in data subset selection, active learning, domain adaptation, and representation learning. This raises a natural and fundamental question: are the monotone submodular functions most commonly used in practice entropic? In this paper, we answer this question in the affirmative for a broad class of widely used polymatroid functions. We provide explicit entropic constructions for set cover and coverage functions, facility location, saturated coverage, concave-over-modular functions via truncations, and monotone graph-cut-type objectives. Our results show that these functions can be realized exactly as Shannon entropies of appropriately constructed random variables. As a consequence, for these functions, submodular mutual information coincides with classical mutual information, conditional gain specializes to conditional entropy, and submodular conditional mutual information reduces to standard conditional mutual information in the entropic sense. These results establish a direct bridge between combinatorial information measures and classical information theory for many of the most common submodular objectives used in applications.
A Lower Bound on the Expected Number of Distinct Patterns in a Random Permutation
Let $π_n$ be a uniformly chosen random permutation on $[n]$. The authors of [2] showed that the expected number of distinct consecutive patterns of all lengths $k\in\{1,2,\ldots,n\}$ in $π_n$ was $\frac{n^2}{2}(1-o(1))$ as $n\to\infty$, exhibiting the fact that random permutations pack consecutive patterns near-perfectly. A conjecture was made in [11] that the same is true for non-consecutive patterns, i.e., that there are $2^n(1-o(1))$ distinct non-consecutive patterns expected in a random permutation. This conjecture is false, but, in this paper, we prove that a random permutation contains an expected number of at least $2^{n-1}(1+o(1))$ distinct permutations; this number is half of the range of the number of distinct permutations.
2026-01-17
Analysis of a Random Local Search Algorithm for Dominating Set
Dominating Set is a well-known combinatorial optimization problem which finds application in computational biology or mobile communication. Because of its $\mathrm{NP}$-hardness, one often turns to heuristics for good solutions. Many such heuristics have been empirically tested and perform rather well. However, it is not well understood why their results are so good or even what guarantees they can offer regarding their runtime or the quality of their results. For this, a strong theoretical foundation has to be established. We contribute to this by rigorously analyzing a Random Local Search (RLS) algorithm that aims to find a minimum dominating set on a graph. We consider its performance on cycle graphs with $n$ vertices. We prove an upper bound for the expected runtime until an optimum is found of $\mathcal{O}\left(n^4\log^2(n)\right)$. In doing so, we introduce several models to represent dominating sets on cycles that help us understand how RLS explores the search space to find an optimum. For our proof we use techniques which are already quite popular for the analysis of randomized algorithms. We further apply a special method to analyze a reversible Markov Chain, which arises as a result of our modeling. This method has not yet found wide application in this kind of runtime analysis.
2026-01-15
Coarsening Causal DAG Models
Directed acyclic graphical (DAG) models are a powerful tool for representing causal relationships among jointly distributed random variables, especially concerning data from across different experimental settings. However, it is not always practical or desirable to estimate a causal model at the granularity of given features in a particular dataset. There is a growing body of research on causal abstraction to address such problems. We contribute to this line of research by (i) providing novel graphical identifiability results for practically-relevant interventional settings, (ii) proposing an efficient, provably consistent algorithm for directly learning abstract causal graphs from interventional data with unknown intervention targets, and (iii) uncovering theoretical insights about the lattice structure of the underlying search space, with connections to the field of causal discovery more generally. As proof of concept, we apply our algorithm on synthetic and real datasets with known ground truths, including measurements from a controlled physical system with interacting light intensity and polarization.
Source localisation in simple random walks
We consider the problem of locating the source (starting vertex) of a simple random walk, given a snapshot of the set of edges (or vertices) visited in the first $n$ steps. Considering lattices $\mathbb{Z}^d$, in dimensions $d \geq 5$, we show that the source can be identified (a) with probability bounded away from $0$ using one guess, and (b) with probability arbitrarily close to $1$ using a constant number of guesses. On the other hand, for dimensions $d \leq 2$, we show that one cannot locate the source with positive constant probability. Our arguments apply more generally to strongly transient and recurrent simple random walks on vertex-transitive graphs.
2026-01-15
Universality results for random matrices over finite local rings
Let $R$ be a finite local ring. We prove a quantitative universality statement for the cokernel of random matrices with i.i.d. entries valued in $R$. Rather than use the moment method, we use the Lindeberg replacement technique. This approach also yields a universality result for several invariants that are finer than the cokernel, such as the span and the determinant.
2026-01-14
$q$-deformation of the Marchenko-Pastur law
We study a $q$-deformed random unitary ensemble associated with the little-$q$ Laguerre weight, which provides a discrete analogue of the classical Laguerre unitary ensemble. In the double scaling regime $q=e^{-λ/N}$, where $N$ is the system size and $λ\ge 0$, we derive the limiting spectral distribution as $N\to \infty$, which yields a $q$-deformation of the Marchenko-Pastur law. The limiting density undergoes a phase transition at an explicitly determined critical value $λ_c$: for $λ<λ_c$, the support consists of a single band region, whereas for $λ>λ_c$ an additional saturated region emerges adjacent to the band region. Our derivation of the limiting distribution is based on three complementary approaches: the method of moments, the analysis of a constrained equilibrium problem, and the asymptotic zero distribution of orthogonal polynomials. As a consequence, we establish the convergence of the empirical measure as well as a large deviation principle. In addition, we derive closed-form expressions for the spectral moments using the combinatorial structure of orthogonal polynomials, and obtain large-$N$ expansions for these moments.
2026-01-14
Quantative universality for cokernels of matrices with symmetries
We prove universality for cokernels of random integral matrices with symmetries via an approach different from the classical surjection moment method introduced by Wood (arXiv:1402.5149). In the symmetric case, we reprove Hodges' universality theorem (arXiv:2311.07078), i.e. the version incorporating the canonical pairing from Wood's setting, and in the alternating case we reprove the local universality theorem of Nguyen-Wood (arXiv:2210.08526). A key advantage of our method is that it is quantitative: we obtain explicit error bounds, which are exponentially small in most regimes, thereby addressing Wood's question on effective convergence rates. Our argument is inspired by Maples' exposure-process and coupling viewpoint (arXiv:1301.1239) and uses a generalized form of Fourier-analytic estimates in the exponentially sharp style of Ferber-Jain-Sah-Sawhney (arXiv:2106.04049).
2026-01-13
Asymptotic distribution of the Betti numbers of $\overline{\mathcal{M}}_{0,n}$
Asymptotic normality is frequently observed in large combinatorial structures, rigorously established for many quantities such as cycles or inversions in random permutations, the number of prime factors of random integers, and various parameters of random graphs. In this paper, we investigate whether this normal limit behavior extends to the topological invariants of geometric spaces. We show that the Betti numbers of the moduli space of rational curves with $n$ marked points $\overline{\mathcal{M}}_{0,n}$ and the Fulton-MacPherson configuration space $\mathbb{P}^1[n]$ are asymptotically normally distributed. Based on numerical evidence and established log-concavity, we conjecture that the Betti numbers of the quotients of these spaces by the symmetric group $\mathbb{S}_n$ are also asymptotically normally distributed. In contrast, we provide examples of geometric spaces that do not follow this Gaussian law.
2026-01-13
Beta distribution and associated Stirling numbers of the second kind
Published in Probability and Mathematical Statistics, 2024, Vol. 44, Fasc. 1, 119--132 • View PublicationBIB
This article gives a formula for associated Stirling numbers of the second kind based on the moment of a sum of independent random variables having a beta distribution. From this formula we deduce, using probabilistic approaches, lower and upper bounds for these numbers.
Fluctuations of the Ising free energy on Erdős-Rényi graphs
We investigate the ferromagnetic Ising model on the Erdős-Rényi random graph $\mathbb{G}(n,m)$ with bounded average degree $d=2m/n$. Specifically, we determine the limiting distribution of $\log Z_{\mathbb{G}(n,m)}(β,B)$, where $Z_{\mathbb{G}(n,m)}(β,B)$ is the partition function at inverse temperature $β>0$ and external field $B\geq0$. If either $B>0$, or $B=0$, $d>1$ and $β>\operatorname{ath}(1/d)$ the limiting distribution is a Gaussian whose variance is of order $Θ(n)$ and is described by a family of stochastic fixed point problems that encode the root magnetisation of two correlated Galton-Watson trees. By contrast, if $B=0$ and either $d\leq1$ or $β<\operatorname{ath}(1/d)$ the limiting distribution is an infinite sum of independent random variables and has bounded variance.
2026-01-12
Approximate FKG inequalities for phase-bound spin systems
The FKG inequality is an invaluable tool in monotone spin systems satisfying the FKG lattice condition, which provides positive correlations for all coordinate-wise increasing functions of spins. However, the FKG lattice condition is somewhat brittle and is not preserved when confining a spin system to a particular phase. For instance, consider the Curie-Weiss model, which is a model of a ferromagnet with two phases at low temperature corresponding to positive and negative overall magnetization. It is not a priori clear if each phase internally has positive correlations for increasing functions, or if the positive correlations in the model arise primarily from the global choice of positive or negative magnetization. In this article, we show that the individual phases do indeed satisfy an approximate form of the FKG inequality in a class of generalized higher-order Curie-Weiss models (including the standard Curie-Weiss model as a special case), as well as in ferromagnetic exponential random graph models (ERGMs). To cover both of these settings, we present a general result which allows for the derivation of such approximate FKG inequalities in a straightforward manner from inputs related to metastable mixing; we expect that this general result will be widely applicable. In addition, we derive some consequences of the approximate FKG inequality, including a version of a useful covariance inequality originally due to Newman as well as Bulinski and Shabanovich. We use this to extend the proof of the central limit theorem for ERGMs within a phase at low temperatures, due to the second author, to the non-forest phase-coexistence regime, answering a question posed by Bianchi, Collet, and Magnanini for the edge-triangle model.
2026-01-12
The random stable roommates problem typically has no solution
Assume that $n = 2k$ potential roommates each have an ordered preference of the $n-1$ others. A stable matching is a perfect matching of the $n$ roommates in which no two unmatched people prefer each other to their matched partners. In their seminal 1962 stable marriage paper, Gale and Shapley noted that not every instance of the stable roommates problem admits a stable matching. In the case when the preferences are chosen uniformly at random, Gusfield and Irving predicted in 1989 that there is no stable matching with high probability for large $n$. We prove this conjecture and show that for $n$ sufficiently large, the probability there is a stable matching is at most $n^{-1/17}$.
2026-01-12
Anticoncentration of random spanning trees in almost regular graphs
The celebrated formula of Otter \emph{[Ann. of Math. (2) 49 (1948), 583--599]} asserts that the complete graph contains exponentially many non-isomorphic spanning trees. In this paper, we show that every connected almost regular graph with sufficiently large degree already contains exponentially many non-isomorphic spanning trees. Indeed, we prove a stronger statement: for every fixed $n$-vertex tree $T$, $$ \Pr\bigl[\mathcal{T} \simeq_{\mathrm{iso}} T\bigr] = e^{-Ω(n)}, $$ where $\mathcal{T}$ is a uniformly random spanning tree of a connected $n$-vertex almost regular graph with sufficiently large degree. To prove this, we introduce a graph-theoretic variant of the classical balls--into--bins model, which may be of independent interest.
Note on edge expansion and modularity in preferential attachment graphs
Edge expansion is a parameter indicating how well-connected a graph is. It is useful for designing robust networks, analysing random walks or information flow through a network and is an important notion in theoretical computer science. Modularity is a measure of how well a graph can be partitioned into communities and is widely used in clustering applications. We study these two parameters in two commonly considered models of random preferential attachment graphs, with $h \geq 2$ edges added per step. We establish new bounds for the likely edge expansion for both random models. Using bounds for edge expansion of small subsets of vertices, we derive new upper bounds also for the modularity values for small $h$.
2026-01-09
A Halász-type theorem for permutation anticoncentration
Given a set $A=\{a_1,\ldots,a_n\}$ of real numbers and real coefficients $b_1,\ldots,b_n$, consider the distribution of the sum obtained by pairing the $a_i$'s with the $b_i$'s according to a uniformly random permutation. A recent theorem of Pawlowski shows that as soon as the coefficients are not all equal, this distribution is always spread out at scale $n^{-1}$: no single value can occur with probability larger than $\frac{1}{2\lceil n/2\rceil + 1}$, and this bound is sharp in general. We show that stronger anticoncentration holds when the coefficients have additional diversity. We quantify the structure of the coefficient multiset by a simple statistic depending on its multiplicity profile, and prove that the maximum point mass of the permuted sum decays polynomially faster as this statistic grows. In particular, when the coefficients are all distinct we obtain a bound of $n^{-5/2+o(1)}$, which can be regarded as an analogue of a classical theorem of Erdős and Moser.
2026-01-09
Eigenvalues of $p$-adic random matrices
We develop the basic theory of eigenvalues of $p$-adic random matrices, analogous to the classical theory for random matrices over $\mathbb{R}$ and $\mathbb{C}$. Such eigenvalue statistics were proposed as a model for the zeroes of $p$-adic $L$-functions by Ellenberg-Jain-Venkatesh, who computed the limiting distribution of the number of eigenvalues in a unit disc. We compute the full joint distribution of the $n$ eigenvalues of an $n \times n$ matrix with Haar distribution, obtaining Coulomb gas type formulas as in the archimedean case, with Vandermonde terms leading to eigenvalue repulsion. From these Coulomb gas density functions we derive asymptotics of eigenvalue statistics as $n \to \infty$. These include exact computations, such as a closed form $$ρ(x,y) = 1 - θ_3(-\sqrt{p};||x-y||^2/p)$$ for the limiting pair correlation of eigenvalues in $\mathbb{Z}_p$, and similar results in quadratic extensions. Such formulas yield concrete numerical predictions on zeroes of $p$-adic $L$-functions. For eigenvalues in arbitrary extensions of $\mathbb{Q}_p$ we also give precise estimates on their pair-repulsion and expected number of eigenvalues in each extension. Finally, we compute the asymptotic probability that all eigenvalues lie in $\mathbb{Z}_p$. Our proofs combine results from several distinct areas: $p$-adic orbital integrals, roots of random $p$-adic polynomials, the Sawin-Wood moment method for random modules, and Markov chains associated with measures on integer partitions.
Asymptotic enumeration of constrained bipartite, directed and oriented graphs by degree sequence
In the sufficiently sparse case, we find the probability that a uniformly random bipartite graph with given degree sequence contains no edge from a specified set of edges. This enables us to enumerate loop-free digraphs and oriented graphs with given in-degree and out-degree sequences, and obtain subgraph probabilities. Our theorems are not restricted to the near-regular case. As an application, we determine the expected permanent of sparse or very dense random matrices with given row and column sums; in the regular case, our formula holds over all densities. We also draw conclusions about the degrees of a random orientation of a random undirected graph with given degrees, including its number of Eulerian orientations.