arXiv++ Combinatorics

Browse math.CO papers from arXiv

math.PR ↗ arXiv

284 papers in this category
2026-04-10
Random 0/1-polytopes expand rapidly
A 0/1-polytope is the convex hull of a subset $V\subseteq \{0,1\}^n$. A celebrated conjecture of Mihail and Vazirani asserts that the graph of every 0/1-polytope has edge-expansion at least 1. In this paper, we show that typical 0/1-polytopes have significantly stronger expansion. Specifically, if $V$ is formed by sampling each vertex of $\{0,1\}^n$ independently with constant probability $p$, then with high probability the edge-expansion is $Θ(n)$ for $p \in (1/2, 1)$, and $n^{Θ(\log \log n)}$ for $p \in (0, 1/2)$. This improves the previously best known bound $Ω(1)$ due to Ferber, Krivelevich, Sales and Samotij.
2026-04-08
Random permutations from $q$-Demazure products
We study the $q$-deformation of the Demazure product model from arXiv:2407.21653. Consider the longest element $w_0$ in $S_n$ written as a reduced word in simple transpositions. Independently delete each transposition with probability $1-p$ and apply the $q$-Demazure product to the remaining ones. We show that the law of the resulting permutation converges as $n \to \infty$ to a deterministic permuton, which coincides with the $q=0$ case studied in arXiv:2407.21653 for adjusted probability $p'=p(1-q)/(1-qp)$. This resolves Conjecture 1.13 from arXiv:2407.21653 and identifies the limiting permuton explicitly.
Short proofs in combinatorics, probability and number theory II
We give a quintet of proofs resulting from questions posed by Erdős. These questions concern ordinary lines in planar point sets, sequences with uniformly small exponential sums, $K_4$-free $4$-critical graphs with few chords in any cycle, a counterexample to a "fewnomial" version of the Erdős--Turán discrepancy bound, and a finiteness theorem for integers $n$ such that $n-a k^2$ is prime for all $k\leq \sqrt{n/a}$ coprime to $n$ (for fixed $a\in\mathbb Z_+$). Each proof is due to an internal model at OpenAI.
2026-04-08
The Random Subsequence Model and Uniform Codes for the Deletion Channel
We introduce the Random Subsequence Model, a spin glass model on pairs of random strings $(X,Y) \in \{0,1\}^N \times \{0,1\}^M$ whose partition function counts subsequence embeddings of $Y$ into $X$. We study two variants: the null model, where $X$ and $Y$ are independent and uniform, and the planted model, where $X$ is uniform and $Y$ is a uniformly-random length-$M$ subsequence of $X$. We connect the Random Subsequence Model to longstanding problems in various fields, including the best rate achievable by uniformly-random codes in the deletion channel, the longest common subsequence problem between two random strings, and models of directed polymers in statistical physics. In the regime where $N,M\to\infty$ at a fixed ratio $α= M/N \in (0,1)$, we exhibit strict asymptotic separations between the null annealed free energy and the quenched free energies of the null and planted models at all values of the density parameter $α$. This suggests that these models are in a spin glass phase at zero temperature throughout the entire dense regime. As a consequence, we show that uniformly-random codes achieve a positive rate in the deletion channel for all deletion probabilities $p\in [0,1),$ settling multiple conjectures of the second author, Isik and Weissman (2024) and proving the first such positive rate result for the regime $p \geq 1/2$. We also give an exact analytic formula for the annealed free energy of the planted model for all values of the density parameter. This implies a corresponding analytic upper bound on the best rate achievable by uniformly-random codes in the deletion channel, complementing the lower bound from our first result. Our upper and lower bounds for the capacity of the deletion channel under uniform codes are far closer to each other than the best known upper and lower bounds for the capacity of the deletion channel.
2026-04-07
Simplicity of random hypergraphs
Random hypergraphs extend the classical notion of random graphs by allowing hyperedges to join more than two vertices, making them well-suited for modeling higher-order interactions in complex systems. Despite their broad applicability, many structural properties of random hypergraphs remain less understood than in the graph setting. One such property is simplicity: the absence of self-loops, multi-hyperedges, and, in the hypergraph context, degenerate hyperedges where hyperedges contain a copy of the same vertex at least twice. While the behaviour of the number of such self-loops and multi-hyperedges is well understood for random graphs through the configuration model, analogous results for hypergraphs are comparatively sparse. In this work, we study both undirected and directed hypergraphs generated by the configuration model with prescribed vertex and hyperedge degrees. We derive exact, explicit expressions for the expected number of self-loops, multi-hyperedges and degenerate hyperedges, extending classical results from the graph setting. In addition, an asymptotical analysis shows that, under mild moment conditions on the degree distribution, the expected fraction of self-loops, multi-hyperedges and degenerate hyperedges vanishes as the number of vertices grows. Our results provide a systematic understanding of simplicity in directed and undirected hypergraph models.
Large fringe trees for random trees with given vertex degrees
This paper extends the study of fringe trees in random plane trees with a given degree statistic. While previous work established the asymptotic normality of the count of fringe trees isomorphic to a fixed tree, we investigate the case where the target tree grows with the size of the random tree. We consider three primary subtree counts: the number of fringe trees isomorphic to a specific growing tree, the number of fringe trees sharing a given growing degree statistic, and the number of fringe trees of a specific growing size. To establish our results, we employ and compare four distinct probabilistic frameworks: the method of moments with the Gao-Wormald theorem, Stein's method with coupling (to provide explicit error bounds in total variation distance), the Cai-Devroye method, and Stein's method with exchangeable pairs. Our findings provide conditions for Poisson and normal convergence for these subtree counts. Additionally, we provide a local limit theorem for sums of values obtained via sampling without replacement that may be of independent interest. Finally, our results and methods are also applied to conditioned critical Galton-Watson trees.
Non-existence probabilities and lower tails in the critical regime via Belief Propagation
We compute the logarithmic asymptotics of the non-existence probability (and more generally the lower-tail probability) for a wide variety of combinatorial problems for a range of parameters in the `critical regime' between the regime amenable to hypergraph container methods and that amenable to Janson's inequality. Examples include lower tails and non-existence probabilities for subgraphs of random graphs and for $k$-term arithmetic progressions in random sets of integers. Our methods apply in the general framework of estimating the probability that a $p$-random subset of vertices in a $k$-uniform hypergraph induces significantly fewer hyperedges than expected. We show that under some simple structural conditions on the hypergraph and an upper bound on $p$ determined by a phase transition in the hard-core model on the infinite $k$-uniform, $Δ$-regular, linear hypertree, this probability can be accurately approximated by the Bethe free energy evaluated at the unique fixed point of a Belief Propagation operator on the hypergraph.
2026-04-05
Equality in Fill's spectral gap problem
We study the adjacent-transposition chain on the symmetric group $\mathfrak{S}_n$ with a regular parameter vector $\vec{p} = (p_{i,j})_{i\neq j}$. Fill's spectral gap conjecture, recently resolved in the affirmative by Greaves-Zhu, states that among all regular parameter vectors, the spectral gap of the transition matrix is minimized by the uniform vector $p_{i,j}= 1/2$ for all $i\neq j$. We prove the stronger statement that among all regular parameter vectors, the spectral gap is minimized if and only if $\vec{p}$ has a neutral label, i.e., there exists $c \in [n]$ such that $p_{c,i} = 1/2$ for all $i\neq c$. Moreover, in this case, we show that the multiplicity of the second largest eigenvalue is equal to the number of neutral labels, unless the number of neutral labels is $n-2$ or $n$, in which case the multiplicity is $n-1$. This confirms a conjecture of Fill.
2026-04-03
The record statistic and forward stability of Schubert products
We initiate a probabilistic study of forward stability for products of Schubert polynomials through the record statistic (left-to-right maxima) of permutations. Building on the explicit record formula for forward stability obtained by Hardt and Wallach, we study random pairs of permutations drawn from three natural families: uniform permutations, Grassmannian permutations, and Boolean permutations. For each family, we determine record probabilities and use them to analyze the asymptotic behavior of forward stability. For uniform and Grassmannian permutations, we obtain asymptotics for the mean together with limiting distribution results. For Boolean permutations, we prove linear-order growth of the mean, and our analysis also produces an explicit time-inhomogeneous Markov chain that yields an exact linear-time uniform sampler. Beyond these cases, we prove that the record-set statistic is equidistributed on the avoidance classes of $132$ and $231$, and consequently the corresponding forward stability distributions coincide. We conclude with conjectures for numerous further permutation classes and a conjectural recursive criterion for when two avoidance classes have the same record-set distribution.
2026-04-02
Semicircle laws with combined variance for non-uniform Erdős-Rényi hypergraphs
We consider Erdős-Rényi-type random hypergraphs that are non-uniform, in the sense that hyperedges of different sizes may coexist, and inhomogeneous, in that connection probabilities may depend on the hyperedge size. All parameters are allowed to scale with the hypergraph size. We study the random adjacency matrix whose $(u,v)$-entry counts the number of hyperedges containing both vertices $u$ and $v$, and characterize its expected limiting spectral distribution in terms of the connection probabilities and the hyperedge sizes. We provide a Pastur-type condition, in the sense of Chatterjee (2005), under which the matrix can be Gaussianized, as well as a more restrictive but simpler sufficient condition in terms of the generalized average degree of the model. As a second main result, based on such a Gaussianization, we characterize the limiting spectral distributions under non-sparse conditions as semicircle laws with an explicit parametric variance. The latter can be expressed as a convex combination of the variances arising in the uniform cases, with coefficients determined by the trade-off between the different sources of inhomogeneity.
2026-04-01
An Unconditional Barrier for Proving Multilinear Algebraic Branching Program Lower Bounds
Since the breakthrough superpolynomial multilinear formula lower bounds of Raz (Theory of Computing 2006), proving such lower bounds against multilinear algebraic branching programs (mABPs) has been a longstanding open problem in algebraic complexity theory. All known multilinear lower bounds rely on the min-partition rank method, and the best bounds against mABPs have remained quadratic (Alon, Kumar, and Volk, Combinatorica 2020). We show that the min-partition rank method cannot prove superpolynomial mABP lower bounds: there exists a full-rank multilinear polynomial computable by a polynomial-size mABP. This is an unconditional barrier: new techniques are needed to separate $\mathsf{mVBP}$ from higher classes in the multilinear hierarchy. Our proof resolves an open problem of Fabris, Limaye, Srinivasan, and Yehudayoff (ECCC 2026), who showed that the power of this method is governed by the minimum size $N(n)$ of a combinatorial object called a $1$-balanced-chain set system, and proved $N(n) \le n^{O(\log n/\log\log n)}$. We prove $N(n) = n^{O(1)}$ by giving the chain-builder a binary choice at each step, biasing what was a symmetric random walk into one where the imbalance increases with probability at most $1/4$; a supermartingale argument combined with a multi-scale recursion yields the polynomial bound.
2026-03-31
Disordered Schur Measures
In this paper, we introduce and study random Schur measures whose parameters are sampled from the Circular Unitary Ensemble. We show that Schur measures with CUE disorder exhibit behavior reminiscent of spin glasses.
Randomstrasse101: Open Problems of 2025
Randomstrasse101 is a blog dedicated to Open Problems in Mathematics, with a focus on Probability Theory, Computation, Combinatorics, Statistics, and related topics. This manuscript serves as a stable record of the Open Problems posted in 2025, with the goal of easing academic referencing. The blog can currently be accessed at randomstrasse101.math.ethz.ch
2026-03-31
Counting partial Hadamard matrices in the cubic regime
We give a precise asymptotic formula for the number of $n\times 4t$ partial Hadamard matrices in the regimes $t/n^3\to\infty$ and $t/n^3\toΘ$ for sufficiently large fixed $Θ$. This strengthens earlier results of de~Launey and Levin, who obtained the asymptotic for $t/n^{12}\to\infty$, and of Canfield, who extended this to $t/n^4\to\infty$.
2026-03-30
Composition of random functions and word reconstruction
Given two functions $\mathbf{a}\!:\! [n] \rightarrow [n]$ and $\mathbf{b}\!:\! [n] \rightarrow [n]$ chosen uniformly at random, any word $w=w_1w_2\dots w_k\in \{a,b\}^k$ induces a random function $\mathbf{w}\!:\! [n] \rightarrow [n]$ by composition, i.e. $\mathbf{w}=φ_{w_k}\circ \dots \circ φ_{w_1}$ with $φ_a=\mathbf{a}$ and $φ_b=\mathbf{b}$. We study the following question: assuming $w$ is fixed but unknown, and $n$ goes to infinity, does one sample of $\mathbf{w}$ carry enough information to (partially) recover the word $w$ with good enough probability? We show that the length of $w$, and its exponent (largest $d$ such that $w={u}^d$ for some word ${u}$) can be recovered with high probability. We also prove that the random functions stemming from two different words are separated in total variation distance, provided that certain ``auto-correlation'' word-depending constant $c(w)$ is different for each of them. We give an explicit expression for $c(w)$ and conjecture that non-isomorphic words have different constants. We prove that this is the case assuming a major conjecture in transcendental number theory, Schanuel's conjecture.
2026-03-30
Limit Laws for the Distance to Fréchet Means of Random Graphs
This paper investigates the Fréchet mean of the Erdős-Rényi random graph $G_{n,p}$ with respect to the Frobenius distance on graph Laplacians, a metric that captures global structural information beyond local edge flips. We first characterize the Fréchet mean set as consisting of quasi-regular graphs (i.e., graphs where all vertex degrees differ by at most one). We then analyze the asymptotic behavior of the Frobenius distance $F_n=d_{\mathrm{F}}(G_{n,p},R)$ as $n\to\infty$, where $R$ is any Fréchet mean. Closed-form expressions for the mean and variance of $F_n^2$ are derived, which are invariant to the choice of $R$. Leveraging these results, we establish several weak convergence laws for the Frobenius distance over all regimes of $p \in (0,1)$ as $n \to \infty$. Finally, under the scaling condition $n^2 p(1-p) \to \infty$ we prove the asymptotic normality of this distance, which exhibits a phase transition governed by the growth rate of $np(1-p)$. Our results reveal how metric selection fundamentally shapes Fréchet mean geometry in random graphs.
2026-03-28
On the critical fugacity of the hard-core model on regular bipartite graphs
We establish long-range order for the hard-core model on a finite, regular bipartite graph above a threshold fugacity given in terms of expansion parameters of the graph. The result applies to the $d$-dimensional hypercube graph and, more generally, to $d$-dimensional discrete tori of fixed side length, proving long-range order at fugacities $λ\geΩ(\frac{\log d}{d})$. Furthermore, we use reflection positivity to transfer the result to the lattice $\mathbb{Z}^{d}$, verifying the long-standing belief that its critical fugacity is of the form $d^{-1+o(1)}$ as $d\to\infty$.
2026-03-28
The $k$-cycle shuffling with repeated cards
We investigate the $k$-cycle shuffle on repeated cards, namely on a deck consisting of $l$ identical copies of each of $m$ card types, with total size $n=ml$. We establish asymptotic results for the total variation mixing of this shuffle, including cutoff and explicit limiting profiles. For fixed $l$, we show that the walk exhibits cutoff at time $\frac{n}{k}\log n$ with window of order $\frac{n}{k}$, and we identify the limiting profile in terms of the total variation distance between Poisson distributions arising from quotient fixed-point statistics. When $l\to\infty$ with sufficiently slow growth, more precisely when $l=o(\log n)$, we prove that the cutoff location shifts to $\frac{n}{k}\left(\log n-\frac 12\log l\right)$, again with window of order $\frac{n}{k}$, and that the limiting profile is asymptotically Gaussian, arising from a Poisson comparison after normal approximation. The proof is based on an approximation of the shuffling measure by an explicitly tractable auxiliary measure, generalizing the $k=2$ case from Jain and Sawhney (arXiv:2410.23944). The representation-theoretic framework underlying the analysis of this auxiliary measure follows from the work of Hough (arXiv:1605.00911) and Nestoridi and Olesker-Taylor (arXiv:2005.13437)
2026-03-27
Complete Causal Identification from Ancestral Graphs under Selection Bias
Many causal discovery algorithms, including the celebrated FCI algorithm, output a Partial Ancestral Graph (PAG). PAGs serve as an abstract graphical representation of the underlying causal structure, modeled by directed acyclic graphs with latent and selection variables. This paper develops a characterization of the set of extended-type conditional independence relations that are invariant across all causal models represented by a PAG. This theory allows us to formulate a general measure-theoretic version of Pearl's causal calculus and a sound and complete identification algorithm for PAGs under selection bias. Our results also apply when PAGs are learned by certain algorithms that integrate observational data with experimental data and incorporate background knowledge.
2026-03-27
A proof of Fill's spectral gap conjecture
We prove a quantitative lower bound on the spectral gap of the adjacent-transposition chain on the symmetric group with a general probability vector. As a consequence, among all regular probability vectors, the spectral gap of the transition matrix is minimised by the uniform probability vector, i.e., $p_{i,j}\equiv {\frac 1 2}$ for all $i \ne j$. A second consequence is a uniform polynomial bound on the inverse spectral gap in the regular case. This resolves a longstanding conjecture known as Fill's Gap Problem.