arXiv++ Combinatorics

Browse math.CO papers from arXiv

math.PR ↗ arXiv

284 papers in this category
2026-04-28
Asymptotic height of Plancherel random trees
We study a natural analogue of Ulam's problem for random rooted trees distributed according to a Plancherel-type measure. This probability measure is closely related to the classical Plancherel measure on integer partitions. For a Plancherel random tree $T_n$ with $n$ vertices, we investigate the asymptotic behavior of its height $H_n$, defined as the maximal distance from the root to a leaf. We prove that this height grows logarithmically. More precisely, there is a one-parameter family of random trees $(T_n(θ))_{n \in \mathbb{N}}$ indexed by $θ>0$ such that $\frac{H_n}{\log n}$ converges in probability to $c_\star(θ)$, where $c_\star(θ)$ is an explicit constant depending on the parameter $θ$. The case of Plancherel trees corresponds to the parameter $θ=2$. The proof is based on the fact that the Plancherel random trees can be viewed as Ewens fragmentation trees, for which the height exhibits a sharp threshold phenomenon. An upper bound is obtained via $s$-mass functionals and contraction estimates, while the lower bound is derived by embedding the model into a branching random walk with logarithmic displacements governed by a Poisson--Dirichlet distribution. The constant $c_\star(θ)$ is characterized through a variational principle associated with this branching random walk.
2026-04-27
Liouville Quantum Duality and Random Planar Maps II
This is Part II of our project on block-weighted planar maps and Liouville quantum duality. Focusing on the scaling properties at the dual critical point, we derive the conditional distribution of the root block size given the total size, as well as, conversely, the distribution of the total size for a fixed root block size. We show that these laws are in perfect agreement with the results of Liouville quantum gravity (LQG), obtained by modifying the standard Liouville random measure with additional atomic contributions representing localized quantum areas. The ratio of dual and direct partition functions with punctures is shown to be universal, its explicit LQG expression exactly matching its combinatorial analogue. We also investigate the block distance profile for doubly rooted maps, which is here rigorously related to the distance profile of maps consisting of a single block. Finally, we analyze the multifractal properties of the usual and dual Liouville measures, predicting the associated spectra, from both quantum and Euclidean perpectives. We illustrate our results through specific realizations of block-weighted planar maps, i.e., quadrangulations decomposed into simple blocks, tree-like structures formed by attaching quartic maps, and bicubic maps decomposed into 3-connected blocks. For each model, we give the single non-universal constant which uniquely determines the strength of the corresponding atomic Liouville measure.
2026-04-27
Local Limit of Random Regular Bipartite Planar Maps
We prove the existence of the local limit of uniform random d-regular bipartite planar maps, for every $d\geq 3$, as the number of vertices tends to infinity. The proof relies on a bijection between maps and so-called blossoming trees established in a previous work. After proving local convergence of the associated decorated trees, we extend the bijection to infinite trees and transfer the convergence to planar maps. The limiting object is almost surely one-ended and recurrent for the simple random walk.
Large deviation principles for pattern-avoiding permutations, and limit shapes for constrained Mallows permutations
We study Mallows random permutations conditioned to avoid a given pattern $α$ of length~$3$. When the bias parameter is of the form $e^{β/n}$, we prove that these permutations converge to a non-trivial explicit deterministic permuton that depends on the pattern $α$ and on the parameter $β$. Along the way, we provide parametrizations for $α$-avoiding permutons, and establish a large deviation principle for uniform $α$-avoiding permutations. As a byproduct of the proof, we also obtain asymptotic estimates of two versions of $q$-Catalan numbers in the regime $q=e^{β/n}$.
2026-04-26
Recursive Record Filtering and Longest Decreasing Subsequences
We consider a recursive record-filtering procedure, which we informally call Disappear-Sort. Let $D_n$ denote the random variable giving the required number of passes in Disappear-Sort to eliminate a sequence of length $n$ sampled as i.i.d. copies of a continuous random variable $X$, where each pass retains the left-to-right records and discards all remaining entries. We show that this procedure admits two natural probabilistic interpretations. For the resampling variant we prove that $d_n=\mathbb{E}[D_n]$ satisfies an exact recurrence involving the unsigned Stirling numbers of the first kind. For the non-resampling variant, we associate to a permutation $p_n\in S_n$ a natural poset and prove that the recursive Disappear-Sort layers form an antichain decomposition of this poset. We deduce that the total number of passes equals $L(p_n)$, where $L(p_n)$ is the length of the longest decreasing subsequence of $p_n$. We then show that for a uniform random permutation of size $n$, the expectation $\mathbb{E}[D_n]$ of this second variant coincides with the expected first-column length of a Plancherel-random Young diagram. Using the Robinson--Schensted correspondence, we obtain an exact formula for this expectation in terms of partitions and standard Young tableaux, and classical Plancherel asymptotics then yield $\mathbb{E}[D_n]\sim 2\sqrt{n}$, with fluctuations on the $n^{1/6}$ scale governed by the Tracy--Widom law derived by Baik, Deift and Johansson. We conclude with an $O(n\log n)$ implementation.
2026-04-26
The Cutoff Profile for Random Transpositions on Repeated Cards in the Full Range of Parameters
The random transposition shuffle on repeated cards induces a Markov chain on the quotient space of arrangements with multiplicities, and is equivalent to the many-urn mean-field Bernoulli-Laplace model introduced by Scarabotti. Writing $n=ml$, where there are $m$ card types and each type appears $l$ times, we determine the limiting profile for the total variation distance to stationarity at times $t=\frac{n}{2}\left(\log n-\frac{1}{2}\log l+c\right)$, under the assumption $l=ω(1)$. Scarabotti previously established that this process exhibits cutoff at time $\frac{n}{2}(\log n-\frac{1}{2}\log l)$; our result refines this by identifying the precise asymptotic shape of convergence inside the cutoff window. We show that the limiting profile is asymptotically Gaussian, with different explicit forms in the regimes $m$ fixed and $m=ω(1)$. Together with our previous work on the fixed-$l$ regime, where the limiting profile is of Poisson type, this yields the cutoff profile for the random transposition shuffle on $n=ml$ repeated cards for the full range of parameters $m$ and $l$. Our argument has two main steps. First, we combine Scarabotti's Fourier-analytic framework for the many-urn Bernoulli-Laplace model with the approximation method of Jain-Sawhney (arXiv:2410.23944). More precisely, we compare the original shuffling measure with an explicitly tractable auxiliary measure directly on the repeated card quotient space, rather than passing through an intermediate comparison on the full symmetric group; this step relies in particular on our new estimates for Kostka numbers. Second, we reduce the limiting-profile problem to quotient fixed-point statistics and analyze them via Hoeffding-type combinatorial central limit theorems.
High-Precision Framework for Expected Hitting Times Analysis in the Dice-Sum Process
We study the expected number of rolls required for the cumulative sum of a fair six-sided die to first enter a prescribed target set $H\subset\mathbb{Z}_{\ge0}$. A one-variable dynamic-programming formulation is introduced that removes dependence on the roll count. Within this framework, the infinite process is truncated at a large cutoff $N$ and corrected by an analytically derived overshoot term that accounts for the rare event of exceeding $N$ before entering $H$. Explicit bounds on this residual yield a strict two-sided estimate of the truncation error. The method is numerically efficient, requiring constant memory and linear time in the cutoff. For the perfect-square target set $H=\{n^2:n\in\mathbb{N}\}$, all quantities are evaluated explicitly, yielding \[ \mathbb{E}[T]=7.07976423755110510389555305690818489468\ldots, \] provably correct to 1,017 decimal places. This constitutes the most precise result known to date and establishes a general framework for high-accuracy computation of discrete hitting times.
2026-04-25
Scaling limit of Sinkhorn-rescaled Random Matrices via Stability of Static Schrödinger Bridges
We analyze the asymptotic behavior and scaling limits of large random matrices rescaled via the Sinkhorn algorithm to match prescribed row and column margins. For a random matrix with independent sub-exponential entries, we show that its Sinkhorn rescaling concentrates around the rescaling of its mean matrix, both at the level of the Schrödinger potentials and as random measures on the unit square, with explicit non-asymptotic rates. As the dimensions grow, the rescaled random matrix converges to the continuous static Schrödinger bridge (SSB) determined by the limiting margins and reference density. Around this scaling limit we develop a fluctuation theory: bulk rigidity for the empirical spectral distribution of the associated sample covariance matrix, and a central limit theorem for the empirical Schrödinger potentials of the rescaled empirical mean. Our analysis is driven by a new quantitative stability theory for the SSB, developed in three forms: Lipschitz continuity in the Hellinger distance under perturbations of the reference measure (kernel stability); Hölder-$1/2$ continuity in the Hellinger distance under $L^1$ perturbations of the margins (margin stability); and $L^\infty$ stability of the discrete Schrödinger potentials under margin perturbation (potential stability). Translated to the discrete random-matrix setting, these bounds yield the concentration and scaling-limit results, while a local law for random Gram matrices with a non-uniform variance profile drives the bulk rigidity. Our SSB stability theory may be of independent interest.
The Mihail-Vazirani conjecture and strong edge-expansion in random $0/1$ polytopes
We study the edge-expansion of the graph of a random $0/1$ polytope $P^d_p$, defined as the convex hull of a random subset of the points in $\{0,1\}^d$ where every point is retained independently and with probability $p$. This problem was introduced more than twenty years ago in a work of Gillmann and Kaibel, and has been extensively studied ever since. We prove that, for every fixed $\varepsilon>0$ and every $p\in(0,1-\varepsilon]$, with high probability the graph of $P^d_p$ has edge-expansion $Θ(d)$. This improves the previously best known bound due to Ferber, Krivelevich, Sales and Samotij, and verifies, in a strong form, the celebrated Mihail-Vazirani conjecture for random $0/1$ polytopes. Although the expansion factor $Θ(d)$ is typically best possible for $p\ge 1/2+\varepsilon$, we also show that the behaviour changes drastically at $p=1/2$. Namely, for every fixed $\varepsilon>0$ and every integer $k\ge 2$, if $p\le 1/2-\varepsilon$, then with high probability the graph of $P^d_p$ has edge-expansion $Ω(d^k)$. Thus, random $0/1$ polytopes exhibit an interesting phase transition at $p=1/2$.
2026-04-21
Achieving the Kesten-Stigum bound in the non-uniform hypergraph stochastic block model
We study the community detection problem in the non-uniform hypergraph stochastic block model (HSBM), where hyperedges of varying sizes coexist. This setting captures higher-order and multi-view interactions and raises a fundamental question: can multiple uniform hypergraph layers below the detection threshold be combined to enable weak recovery? We answer this question by establishing a Kesten--Stigum-type bound for weak recovery in a general class of non-uniform HSBMs with $r$ blocks, generated according to multiple symmetric probability tensors. In the case $r=2$, we show that weak recovery is possible whenever the sum of the signal-to-noise ratios across all uniform hypergraph layers exceeds one, thereby confirming the positive part of a conjecture in (Chodrow et al., 2023). Moreover, we provide a polynomial-time spectral algorithm that achieves this threshold via an optimally weighted non-backtracking operator. For the unweighted non-backtracking matrix, our spectral method attains a different algorithmic threshold, also conjectured in (Chodrow et al., 2023). Our approach develops a spectral theory for weighted non-backtracking operators on non-uniform hypergraphs, including a precise characterization of outlier eigenvalues and eigenvector overlaps. We introduce a novel Ihara--Bass formula tailored to weighted non-uniform hypergraphs, which yields an efficient low-dimensional representation and leads to a provable spectral reconstruction algorithm. Taken together, these results provide a principled and computationally efficient approach to clustering in non-uniform hypergraphs, and highlight the role of optimal weighting in aggregating heterogeneous higher-order interactions.
2026-04-17
Universal dualities for Wilson loops in lattice Yang-Mills
We identify a universal finite-$N$ structure underlying Wilson loop expectations in lattice Yang-Mills, in any dimension $d\geq 2$, for gauge group $\mathrm{U}(N)$, and for arbitrary smooth central plaquette actions. The starting point is a state-sum expansion in plaquette labels by irreducible representations, in which each term factorizes into an action-dependent spectral weight and an action-independent topological coefficient. We then analyze these coefficients in three exact ways: as a gauge/string expansion over decorated spanning surfaces, as a local spin-foam/channel model on the dual incidence graph, and as a universal finite-$N$ master loop equation that closes on the coefficient side. As a consequence, several recent Wilson-action results are recovered as specializations of our broader action-agnostic framework.
2026-04-16
A counter-example to persistence in generalised preferential attachment trees
Consider a generalised preferential attachment tree with attachment function $f$, that is a random tree, where at each time-step a node connects to an existing node $v$ with probability proportional to $f(\mathrm{deg}(v))$, where $\mathrm{deg}(v)$ denotes the degree of the node in the existing tree. We provide a counter-example to a conjecture of the author asserting that under the assumption $\sum_{j=1}^{\infty} \frac{1}{f(j)^2} < \infty$ there is a persistent hub in the model, that is, a single node that has the maximal degree for all but finitely many time-steps. The counter-example is a minor modification of a related counter-example due to Galganov and Ilienko.
2026-04-15
Convolution, cumulants and infinitesimal generators in the formal power series ring
We extend the notions of finite free convolution and finite free cumulants to the setting of formal power series by introducing their natural analogues, namely $t$-deformed convolution and $t$-deformed cumulants. In this framework, we establish $t$-deformed analogues of the law of large numbers and the central limit theorem, revealing structural parallels with classical, free, and finite free probability theories. We show that the case $t=-1$ recovers classical convolution at the level of moment generating functions, thereby connecting the theory directly to classical probability. We further investigate the infinitesimal generators associated with $\boxplus^t$-continuous semigroups, deriving explicit representation formulas that clarify how these generators describe the infinitesimal evolution of the semigroup. In the case $t = d$, our results yield explicit formulas for finite free infinitesimal generators. In the case $t = -1$, we relate these generators to those of one-dimensional Lévy processes by identifying the corresponding terms in their representations. This establishes a direct connection between $\boxplus^t$-convolution semigroups and classical Lévy-Khintchine-type generators.
Sweet Trims are made of Threes: A càdlàg erasure of the Brownian tree
We present a simple trimming algorithm that generates nested uniform binary plane trees by removing leaves one-by-one using a best-of-three-match procedure. While its one-step transition specializes to the Luczak-Winkler & Caraceni-Stauffer coupling, its scaling limit provides a suprising càdlàg erasure of Brownian trees, reminiscent of SLE theory.
2026-04-14
On additive averaging kernels for finite Markov chains
We study additive mixtures of Markov kernels of the form $A_α= αP + (1-α)G$, where $α\in [0,1]$, $P$ is a baseline sampler and $G$ is a Gibbs kernel induced by a partition of the state space. We first motivate the study of $A_α$, which can be interpreted as the projection of a lifted Markov chain. We then consider the minimisation of distance to stationarity under two objectives: the squared Frobenius norm and the Kullback-Leibler (KL) divergence. For the Frobenius objective, we derive explicit trace formulas and identify a Cheeger-type functional that characterises optimal two-block partitions. This yields a structured combinatorial optimisation problem admitting a difference-of-submodular decomposition, enabling efficient approximation via majorisation-minimisation. We also obtain geometric decay rates governed by the absolute spectral gap of $P$. For the KL divergence, we establish convexity-based bounds showing that the divergence of $A_α$ is controlled by those of both $P$ and $G$, thereby reducing partition selection to the Gibbs component. Numerical experiments on the Curie-Weiss model demonstrate that suitable choice of both the partition and the parameter $α$ can significantly accelerate convergence in total variation distance. We observe a consistent trade-off between local exploration and global averaging, with intermediate values of $α$ achieving the best performance across regimes.
2026-04-13
FlowBoost Reveals Phase Transitions and Spectral Structure in Finite Free Information Inequalities
Using FlowBoost, a closed-loop deep generative optimization framework for extremal structure discovery, we investigate $\ell^p$-generalizations of the finite free Stam inequality for real-rooted polynomials under finite free additive convolution $\boxplus_n$. At $p=2$, FlowBoost finds the Hermite pair as the unique equality case and reveals the spectral structure of the linearized convolution map at this extremal point. As a result, we conjecture that the singular values of the doubly stochastic coupling matrix $E_n$ on the mean-zero subspace are ${2^{-k/2}:k=1,\ldots,n-1}$, independent of $n$. Conditional on this conjecture, we obtain a sharp local stability constant and the finite free CLT convergence rate, both uniform in $n$. We introduce a one-parameter family of $p$-Stam inequalities using $\ell^p$-Fisher information and prove that the Hermite pair itself violates the inequality for every $p>2$, with the sign of the deficit governed by the $\ell^p$-contraction ratio of $E_n$. Systematic computation via FlowBoost supports the conjecture that $p^*\!=2$ is the sharp critical exponent. For $p<2$, the extremal configurations undergo a bifurcation, meaning that they become non-matching pairs with bimodal root structure, converging back to the Hermite diagonal only as $p\to 2^-$. Our findings demonstrate that FlowBoost, can be an effective tool of mathematical discovery in infinite-dimensional extremal problems.
2026-04-13
The smallest singular value of signed random combinatorial matrices
Let $M_n$ be an $n\times n$ signed random combinatorial matrix whose rows are independent and uniformly distributed over the set of $\{-1,0,1\}$-vectors with exactly $n/2$ zero coordinates. Despite the dependence induced by the row constraints, we prove that there exist constants $C,c > 0$ such that for any $\varepsilon\ge0$, \begin{align*} \textbf{P}\left(s_{n}(M_n)\le {\varepsilon}{n^{-1/2}}\right)\le C\varepsilon+e^{-cn}. \end{align*} In particular, the probability that $M_n$ is singular is exponentially small. Our approach builds on the Combinatorial Least Common Denominator (CLCD) introduced by Tran and develops the method in the present constrained setting.
2026-04-12
A Strict Gap Between Relaxed and Partition-Constrained Spectral Compression in a Six-State Lumpable Markov Chain
This paper studies a finite reversible lumpable Markov chain for which relaxed spectral compression yields a larger determinant than partition-constrained compression. For a symmetric six-state lumpable chain and the positive operator $T=P^2$, I compare the relaxed benchmark \begin{equation*} \mathfrak D^{\mathrm{rel}}_3(T):=\sup_{U^*U=I_3}\det(U^*TU) \end{equation*} and the partition-constrained benchmark \begin{equation*} \sup_{\mathcal A\,\mathrm{3\text{-}partition}}\det Q_{\mathcal A}(T), \qquad Q_{\mathcal A}(T)=H_{\mathcal A}^*TH_{\mathcal A}. \end{equation*} Here the partition-constrained benchmark is the compression induced by normalized indicator vectors of genuine partitions of the state space. I derive closed formulas for the two analytically central partition families, prove strict upper bounds for both in a local-mode-dominated regime, and combine these bounds with an exhaustive enumeration of all $90$ partitions into three nonempty cells in an explicit six-state model. For this model, one obtains a strict global gap: \begin{equation*} \sup_{\mathcal A}\det Q_{\mathcal A}(T)<\mathfrak D^{\mathrm{rel}}_3(T). \end{equation*} Thus, in this model, indicator-based partition frames are strictly weaker than relaxed orthonormal frames even after global partition-constrained optimization.
2026-04-10
Asymptotic enumeration of admixed arrays and a different independence heuristic
We introduce a class of paired binary matrices called admixed arrays, which arise in analyses of large-scale genetic data and can be viewed as weighted edge colorings of complete bipartite graphs. This combinatorial structure gives rise to two natural families of marginal constraints: a row-sum constraint and a paired column-sum constraint, the latter inducing an inequality among entries of the matrix pair. We study the enumeration of admixed arrays under these constraints in dense regimes. First, we obtain exact formulas for the sizes of the families defined by each constraint in isolation and derive a finite-size criterion characterizing when one constraint is more restrictive than the other. In the large-dimension limit, this comparison simplifies to an entropy inequality, yielding an information-theoretic interpretation and a quantifiable error bound in the semi-regular case. We then analyze the asymptotic enumeration of the doubly constrained family in a semi-regular setting. Using saddle-point approximation and probabilistic techniques, we derive a detailed asymptotic expansion for the logarithm of the count, isolating an explicit fourth-moment contribution and establishing quantitative control of the higher-order remainder. A consequence of this analysis is a phenomenon absent from classical binary and integer matrix models: in the regime $N=Θ(P)$ with uniform margins and density bounded away from zero, the two constraint families obey the independence heuristic with a correction factor $1/\sqrt[4]{e}$ rather than the familiar $e^{\pm1/2}$. Numerical experiments corroborate the analytical approximations, and we implement and extend an algorithm of Miller and Harrison (2013) as open-source software to enumerate constrained admixed arrays.
Limit laws for longest edges in empty region graphs
Empty region graphs are graphs whose vertices are points in $\mathbb{R}^d$ and where two vertices are connected by an edge whenever some associated region does not contain any other vertices. We investigate the asymptotic behaviour of long edges in empty region graphs generated by a stationary Poisson process in $\mathbb{R}^d$. {Letting} the intensity of the underlying Poisson process tend to infinity, we consider the associated point process of edge midpoints, suitably transformed edge lengths, and directions of the edges. We prove that it converges in distribution to a Poisson process on $\mathbb{R}^d \times \mathbb{R}\times\mathbb{L}^d$, where $\mathbb{L}^d$ is the space of lines in $\mathbb{R}^d$ through the origin, and that the suitably transformed length of the longest edge with midpoint in an observation window converges in distribution to a Gumbel distributed random variable. Our approach yields explicit error bounds in Kantorovich--Rubinstein distance for the point process convergence {when restricting to an observation window} and in Kolmogorov distance for the maximal edge length. The results apply uniformly to a broad class of empty region graphs, including the Gabriel graph, the relative neighbourhood graph, the beta-skeleton graph, the Mastercard graph, and the Pacman graph.