math.PR ↗ arXiv
284 papers in this category
Tackling the 6/49 Lottery and Debunking Common Myths with Probabilistic Methods and Combinatorial Designs
At the end, the house always wins! This simple truth holds for all public games of chance. Nevertheless, since lotteries have existed, people have tried everything to give luck a helping hand. This article compares objective scientific approaches to tackle the 6/49 lottery: probabilistic methods and combinatorial designs. The mathematical models developed herein can be modified and applied to other lotteries. The newly constructed (49, 6, 5) covering design is introduced, which meets the Schönheim bound. For lottery designs and for covering designs, a benchmark based on probabilistic methods is presented. It is demonstrated that common attempts to outwit the odds correspond to limitations of numbers to subsets, which disproportionately reduce the chances of winning.
Cycle structure of random standardized permutations
In this article, we study a model of random permutations, which we call random standardized permutations, based on a sequence of i.i.d. random variables. This model generalizes others, such as the riffle-shuffle and the major-index-biased permutations. We first establish an exact result on the joint distribution of the number of cycles of given lengths, involving the notion of primitive words. From this result, we obtain various convergence results, most of which are proved using the method of moments. First we prove that the number of small cycles may have either a Poisson limit distribution, or a limit distribution given by a countable sum of independent geometric distributions. Then we establish a limit distribution for large cycles, which is the Poisson-Dirichlet process. Finally we prove a central limit theorem for the total number of cycles.
Derivation of optimal stochastic Runge-Kutta methods with exotic and decorated Butcher series for the weak integration of stochastic dynamics
The design of numerical integrators for solving stochastic dynamics with high weak order relies on tedious calculations and is subject to a high number of order conditions. The original approaches from the literature consider strong approximations and adapt them for the weak approximation by replacing the iterated stochastic integrals by appropriate random variables. The methods obtained this way are sub-optimal in their number of function evaluations and the analysis of order conditions is unnecessarily complicated. We provide in this paper a novel approach, relying on well-chosen sets of random Runge-Kutta coefficients, that greatly reduce the number of order conditions. The approach is successfully applied to the creation of a collection of new stochastic Runge-Kutta methods of second weak order with an optimal number of function evaluations and a smaller number of random variables. The efficiency of the new methods is confirmed with numerical experiments and a modern algebraic approach using Hopf algebras is provided for the derivation and the study of the order conditions.
Stein's method and the modular behavior of Eulerian numbers
The Eulerian number A(n,k) counts permutations of n symbols with exactly k descents. Motivated by questions in cryptography, several authors have studied the proportion of permutations whose number of descents lies in a fixed congruence class mod b, and its convergence to 1/b. We give an explicit error bound for this convergence using Stein's method for translated Poisson approximation.
Trace identities for quiver representations
We give an expression for the determinant of the twisted Laplacian associated with any linear representation of a finite quiver in terms of traces of the holonomy of its cycles. To establish this expression, we prove a general identity for the determinant of a block matrix in terms of traces of products of its blocks. We give two proofs, one purely enumerative and one using generating series.
In the special case of a finite graph equipped with a vector bundle and a connection, the twisted Laplacian determinant admits a combinatorial interpretation as a weighted count of tuples of oriented cycle-rooted spanning forests, where the weights involve traces of holonomies along cycles formed by combining the edges of the forests.
Toeplitz matrices from permutation displacements and the triangular kernel
Toeplitz matrices arise naturally in harmonic analysis, operator theory, and numerical analysis. In this note we investigate Toeplitz matrices whose coefficients depend on the matrix size through a scaled kernel $a_k=f(k/n)$. We show that the empirical mean of their eigenvalues converges to a weighted integral of $f$, where the weight $1-|x|$ reflects the density of diagonals in Toeplitz matrices. We then introduce a combinatorial construction associating a Toeplitz matrix to a permutation via its displacement counts. For a uniformly random permutation, the expected matrix converges to the Toeplitz matrix generated by the triangular kernel $1-|x|$. Interestingly, the triangular kernel also appears as the covariance function of the integrated Brownian motion, providing a probabilistic interpretation of the same operator. Finally, we analyze the integral operator with kernel $(1-|x-y|)$ on $[0,1]$ and determine its eigenfunctions and eigenvalues explicitly. This operator describes the limiting spectral structure associated with the averaged Toeplitz matrices arising from permutation displacements. These results highlight a natural bridge between Toeplitz matrix theory, permutation statistics, and classical integral operators.
Computation and sampling for Schubert specializations
We present computational results on principal specializations $\mathfrak{S}_w(1^n)$ of Schubert polynomials, which count reduced pipe dreams and reduced bumpless pipe dreams (RBPD). We find the first counterexample, at $n=17$, to the Merzon-Smirnov conjecture (arXiv:1410.6857) that the maximum of $\mathfrak{S}_w(1^n)$ over $S_n$ is attained at a layered permutation. The simulations suggest that $\lim_{n \to \infty} \log(\max_{w\in S_n}\mathfrak{S}_w(1^n))/n^2$ equals the maximal layered permutations' constant from Morales-Pak-Panova (arXiv:1805.04341). We also explore the random permutation drawn from the distribution proportional to $\mathfrak{S}_w(1^n)$, revealing permuton-like asymptotics similar to those for Grothendieck polynomials by Morales-Panova-Petrov-Yeliussizov (arXiv:2407.21653).
We implement and compare three recurrences for $\mathfrak{S}_w(1^n)$: the descent formula (Macdonald), transition formula (Lascoux--Schutzenberger), and cotransition formula (Knutson). For sampling uniformly random RBPDs (whose count is $\sum_{w\in S_n} \mathfrak{S}_w(1^n)$), we show that reducedness breaks the sublattice property of the ASM lattice, preventing monotone CFTP and causing false coalescence. We develop an efficient MCMC sampler with macroscopic "droop" updates for connectivity and fast mixing. Our code computes $\mathfrak{S}_w(1^n)$ up to $n\sim 20$ and samples random RBPDs up to $n\sim 60$ on a personal computer ($n\sim 100$ on a cluster).
Maximising homomorphism counts between digraphs
We prove a Sidorenko-type inequality for directed trees: for every oriented tree $T$ on $k$ vertices and every finite directed graph $G$, the homomorphism count hom$(T,G)$ is bounded above by the maximum of the two pure star counts hom$(S_{0,k-1},G)$ and hom$(S_{k-1,0},G)$. In other words, among all directed trees on $k$ vertices, the pure in- and out-stars maximise the homomorphism count into host digraphs. The proof is purely combinatorial, based on an iterative leaf-reallocation scheme combined with Hölder's inequality. We further investigate the corresponding homomorphism order on directed trees, discuss refinements via tail-truncation and pointwise bounds for rooted host graphs, and record several consequences, e.g. for random directed graph models and local weak limits, where the inequality reduces tree statistics to controlled pure in- and out-degree moments.
Uniqueness and locality of the ground state of the disordered Monomer-Dimer models on independently weighted Unimodular Bienaymé-Galton-Watson trees
Consider a finite graph $G=(V(G),E(G))$ and two continuous weight distributions $ω$ and $ξ$, for which we only assume that $ξ$ is lower bounded. Next, independently draw weights $(w(e))_{e \in E(G)}$ with distribution $ω$ on edges and $(x(v))_{v \in V(G)}$ with distribution $ξ$ on vertices. The ground state of the monomer-dimer model on the weighted graph $G$ is a collection of edges (dimers) and vertices (monomers) such that every vertex is included in at most one monomer or dimer, and such that the sum of weights on its dimers and monomers is maximised.
Take $(G_n,o_n)_{n \in \mathbb{N}}$ to be a sequence of random rooted weighted graphs that converges locally to an independently weighted unimodular Bienaymé-Galton-Watson tree $(\mathbb{T},o)$ with vertex-weight distribution $ξ$ and edge-weight distribution $ω$ . By proving that the ground state of the monomer-dimer model on the tree $(\mathbb{T},o)$ is almost surely unique and locally approximable, we prove that the ground state of the monomer-dimer model on $(G_n,o_n)$ must converge locally to the ground state of the monomer-dimer model on $(\mathbb{T},o)$. This also implies a strong decorrelation property on monomer-dimer models on unimodular Bienaymé-Galton-Watson trees.
Supercritical Site Percolation on Regular Graphs
We consider site (vertex) percolation on $d$-regular graphs, for both constant-degree and growing-degree cases. We give sufficient, and relatively tight, conditions for the emergence of the ``Erdős-Rényi component phenomenon" in the supercritical regime $p=\frac{1+ε}{d-1}$: namely, the appearance of a unique giant component of order $n/d$ in the percolated subgraph, with all other components being of size $O(\log n)$. Our main results apply both to the $d$-dimensional hypercube and to pseudo-random graphs, and resolve two open questions in these cases. We further discuss differences (and similarities) between bond (edge) percolation setting and site percolation setting.
Anticoncentration of random spanning trees in graphs with large minimum degree
A classical result by Otter shows that the complete graph has an exponential number of non-isomorphic spanning trees. This was recently extended by Lee to every almost regular graph of sufficiently large degree.
In this paper, we consider graphs of large minimum degree. We show that every connected graph $G$ with $n$ vertices and minimum degree $d$ has at least $n^{Ω(d)}$ non-isomorphic spanning trees. This is tight up to the constant factor in the exponent. In fact, we prove the following anticoncentration result: if $\mathcal{T}$ is a uniformly random spanning tree of $G$, then for every tree $T$, the probability that $\mathcal{T}$ is isomorphic to $T$ is at most $n^{-Ω(d)}$. This proves a conjecture of Lee in a strong form.
Decay of correlations and zeros for the hard-core model
In a recent paper the last author proved that absence of complex zeros of the partition function of the hard-core model near a parameter $λ>0$ implies a form of correlation decay called strong spacial mixing. In this paper we investigate the reverse implication. We introduce a strengthening of strong spatial mixing that we call very strong spatial mixing (VSSM).
Our main result is that if VSSM holds at a parameter $λ>0$ for a family of graphs, this implies that the partition function has no zeros near that parameter for each graph in the family. We also demonstrate that a closely related variant of very strong spatial mixing does not imply zero-freeness. As a consequence of our main result, we moreover obtain that VSSM implies spectral independence. Our proof relies on transforming the problem to the analysis of an induced non-autonomous dynamical system given by Möbius transformations.
The largest $K_r$-free set of vertices in a random graph
For $r \ge 2$ and a graph $G$, let $α_{r}(G)$ be the maximum number of vertices in a $K_r$-free subgraph of $G$. We investigate the value $α_{r}(G)$ when $G$ is the random graph $G \sim G_{n, 1/2}$ and discover the following phenomenon: with high probability, $α_r(G)$ lies in an interval of constant length that varies in a non-monotonic fashion from $1$ to $\lfloor r/2\rfloor+1$ depending on the value of $n$. The special case $r=2$ corresponds to the independence number of random graphs which is well-known to have two-point concentration; our results therefore extend and generalize this basic fact in random graph theory, showing more complicated behavior when $r>2$. We also prove similar results where $K_r$ is replaced by any color critical graph like $C_5$.
Permanents of random matrices over finite fields
Fix a finite field $\mathbb F_q$ and let $A\in \mathbb F_q^{n\times n}$ be a uniformly random $n\times n$ matrix over $\mathbb F_q$. The asymptotic distribution of the determinant $\det(A)$ is well-understood, but the asymptotic distribution of the permanent $\operatorname{per}(A)$ is still something of a mystery. In this paper we make a first step in this direction, proving that $\operatorname{per}(A)$ is significantly more uniform than $\det(A)$.
Non-uniform Kahn-Kalai, spread, variants, and applications
Building on B.Park and Vondrak's recent generalization of the J.Park-Pham Theorem (formerly known as Kahn-Kalai conjecture) to non-uniform probability measures, this paper introduces the notion of "spread" for the non-uniform setting. This provides a framework to establish 1-statements for subgraph containment in inhomogeneous random graphs with or without a set of forced edges. Using this approach, we derived conditions for the emergence of perfect matchings in the Stochastic Block Model and the Chung-Lu model, and verified that these conditions are in general not tight, but they capture thresholds across a broad range of regimes. Finally, we bridge this non-uniform framework with $\mathcal{G}(n,\textbf{d})$, utilizing a coupling argument to demonstrate thresholds for perfect matchings in $\mathcal{G}(n,\textbf{d})$ for a broad range of degree sequences $\textbf{d}$.
Iterated Graph Systems (I): random walks and diffusion limits
This paper investigates random walks and diffusion limits on a broad class of fractal graphs generated by Edge Iterated Graph Systems (EIGS). We study random walks on combinatorial limit graphs and establish the connections among several dimensions, including the Einstein relation. Building on this, we prove that the rescaled simple random walks converge in the Gromov-Hausdorff-Prokhorov-Skorokhod topology to the limiting diffusion, which further coincides with Brownian motion when the resistance dimension is positive. Moreover, we use the degree dimension to unify the on-diagonal heat-kernel estimates in the locally finite and locally infinite (scale-free) regimes. Finally, we solve the open problem on the quenched resistance exponent for the DHL percolation cluster left in [27].
Sharp threshold for universality of cokernels of classical random matrix models over the $p$-adic integers
We prove that $\frac{\log n}{n}$ is the sharp threshold for universality of the distribution of cokernels of random matrices over $\mathbb{Z}_p$. More precisely, let $α_n = \frac{c\log n}{n}$ for a constant $c>0$ and let $A(n)$ be an $α_n$-balanced random matrix over $\mathbb{Z}_p$. For non-symmetric, symmetric, and alternating matrix models, we prove that if $c>1$, then the limiting distribution of the cokernel of $A(n)$ coincides with the universal distribution of the corresponding symmetry type, whereas universality fails at the critical scale $c=1$. This improves earlier universality results, which required $α_n \gg \frac{\log n}{n}$, to the optimal threshold. As an application, we generalize the universality result for Sylow $p$-subgroups of sandpile groups of Erdős-Rényi random graphs to a broader class of Erdős-Rényi graph sequences. Our approach is based on a unified framework that simultaneously treats all symmetry types of random matrices as well as the random graph model, rather than handling each case separately.
On the structure of the sandpile identity element on Sierpinski gasket graphs
We consider the identity of the abelian sandpile group of finite approximation graphs of the Sierpinski gasket, and we show that the second-order term in the scaling limit converges to the path distance to the nearest corner on the Sierpinski gasket. The proof relies on a decomposition of the identity of the sandpile group into the sum of a constant function and the Laplacian of the graph distance on the approximating graphs.
Optimising two-block averaging kernels to speed up Markov chains
We study the problem of selecting optimal two-block partitions to accelerate the mixing of finite Markov chains under group-averaging transformations. The main objectives considered are the Kullback-Leibler (KL) divergence and the Frobenius distance to stationarity. We establish explicit connections between these objectives and the induced projection chain. In the case of the KL divergence, this reduction yields explicit decay rates in terms of the log-Sobolev constant. For the Frobenius distance, we identify a Cheeger-type functional that characterises optimal cuts. This formulation recasts two-block selection as a structured combinatorial optimisation problem admitting difference-of-submodular decompositions. We further propose several algorithmic approximations, including majorisation-minimisation and coordinate descent schemes, as computationally feasible alternatives to exhaustive combinatorial search. Our numerical experiments reveal that optimal cuts under the two objectives can substantially reduce total variation distance to stationarity and demonstrate the practical effectiveness of the proposed approximation algorithms.
Central limit theorems for high dimensional lattice polytopes: symmetric edge polytopes
We investigate symmetric edge polytopes generated by Erdős--Rényi random graphs in a high-dimensional regime. These objects provide a natural and largely unexplored model of random lattice polytopes, in which geometric properties are governed by graph-theoretic structure. Focusing on the number of polytope edges and on the number of edges in unimodular triangulations, we derive precise asymptotics for expectations and variances and establish central limit theorems with explicit rates of convergence. Our analysis combines a detailed combinatorial-geometric study of the graph configurations determining the facial structure with the discrete Malliavin--Stein method for normal approximation. In particular, we identify a distinguished parameter value at which the leading variance term cancels, producing an atypical fluctuation regime. To the best of our knowledge, the results obtained here constitute the first distributional limit theorems for random lattice polytopes