math.PR ↗ arXiv
284 papers in this category
Limit Profiles for Separation Distance
This paper studies limit profiles for the separation distance. A limit profile records the limiting shape of the distance to stationarity inside the cutoff window, at times of the form $t_n+cw_n$. We start with two famous card shuffles, a general setup for inverse riffle shuffles and random transpositions, and we determine their separation distance limit profiles. We then develop a spectral comparison technique and study continuity properties in the style of [Nes24; Nes25], adapted to separation distance. The comparison method is illustrated through random transpositions, as well as random walks on product groups and the hypercube.
Limit Laws for Consensus Protocols on the Complete Graph
We study a distributed consensus problem on a complete communication network of $n$ vertices, each holding one of two opinions. The vertices communicate in rounds, possibly in the presence of adversarial noise, and exchange information until they all agree on a single opinion. We consider a general class of protocols, where the vertices randomly sample neighbors and update their own opinion according to an update function $f$ depending on the sampled opinions. A prominent example is the $k$-maj protocol, where every vertex adopts the majority opinion of $k$ randomly sampled neighbors.
We consider the runtime $R_n$ that is the number of rounds until all vertices agree on the same opinion, which we call the dominating opinion $D_n$. In our main result we describe the limiting distributions of these two key quantities for a large class of update functions $f$, for arbitrary initial configurations and under the presence of an adversary who may alter the opinions of up to $o(\sqrt{n})$ vertices in each round. We show that there are $f$-specific constants $γ, m > 0$ such that $R_n$ centers around $μ_n = \frac{1}{2}\log_γn + \log_m\ln n$, and we describe the asymptotic distribution of $R_n - μ_n$. In particular, we show that it does not converge, and that it becomes asymptotically periodic both in the $\log n$ as well as the $\log\log n$ scale. Applied to $k$-maj, our results show, among other things, that $γ_{k\text{-maj}} = \binom{k-1}{\lfloor k/2 \rfloor}2^{1-k}k \sim ({2k}/π)^{1/2}$.
The Sharma-Mittal Entropy is Subadditive and Supermodular on the Majorization Lattice
We prove that Sharma-Mittal entropy is a subadditive and supermodular function on the lattice of all $n$-dimensional probability distributions, ordered according to the partial order relation defined by majorization among vectors. Our result unifies and extends analogous results presented in the literature for the Shannon entropy, the Tsallis entropy, and the Rényi entropy.
Faster random walks via infrequent steering
Random walks on graphs can be slow. To speed them up, imagine that at each step instead of choosing the neighbor at random, there is a small probability $\varepsilon>0$ that we can choose it. We show that in this case, at least for graphs of bounded degree, there is a way to steer the walk so that it visits every vertex in $n^{1+o(1)}$ steps with high probability. The key to this result is a way to decompose arbitrary graphs into small-diameter pieces.
Star-collision in random hypergraphs
We study star-based symmetries in uniform hypergraphs and their consequences for matrices whose entries depend only on vertex stars. Such matrices admit a deterministic decomposition into a global component and a local component supported on equivalence classes of vertices with identical stars, known as units. While nontrivial units may exist at finite size in hypergraphs of uniformity greater than two, their persistence in random settings has remained unclear.
We analyze star collisions in random $k$-uniform hypergraphs and show that, in some particular regimes, nontrivial units disappear with high probability as the number of vertices grows. As a consequence, star-dependent matrices exhibit asymptotically trivial local structure, and their spectral behavior, invariant subspaces, and associated linear dynamics are governed by a reduced quotient object obtained by contracting vertex stars.
These results identify star collisions as a finite-size phenomenon in random hypergraphs and clarify the asymptotic irrelevance of star-based symmetries for operator behavior in large random systems in particular regimes.
${\mathrm{ASL}_n}(\mathbb Z)$ invariant random subsets of $\mathbb Z^n$
We classify measures on $\{0,1\}^{\mathbb{Z}^d}$, $d \geq 3$, the space of subsets of $\mathbb{Z}^d$, which are invariant under all affine special linear transformations. In other words, we classify simple point processes on $\mathbb{Z}^d$ whose law is invariant under affine special linear transformations.
We show that every such process is built from a random equivariant polynomial together with independent random sampling, a higher-order generalisation of the cut-and-project method: a random polynomial map is drawn from a distribution invariant under a natural action of $\mathrm{SL}_d(\mathbb{Z})$, each site is then retained independently with a probability determined by a measurable function of the polynomial's value, and the classical cut-and-project construction is recovered in the degree-one case. As a corollary, when the underlying $\mathbb{Z}^d$-action is weakly mixing the measure must be a convex combination of Bernoulli shifts, in the spirit of de Finetti's theorem on exchangeable processes. Our theorem also makes precise how the Howe--Moore theorem fails for the pair $(\mathrm{ASL}_d(\mathbb{Z}), \mathrm{SL}_d(\mathbb{Z}))$.
Motivated by this classification, we formulate a conjecture for $\mathrm{ASL}_d(\mathbb{R})$-invariant point processes on $\mathbb{R}^d$, predicting that any such set decomposes into a Poisson part and a quasicrystal part. The proofs rely on the interaction between the Host--Kra theory of characteristic factors, Zimmer's theory of dynamical cocycles of simple Lie groups, and the dynamics of $\mathrm{SL}_d(\mathbb{Z})$-actions on homogeneous spaces.
Classification aggregation: a quantitative impossibility theorem
A group of individuals wishes to classify $m$ objects into $n$ categories in such a way that no class is left empty, a condition known as surjectivity. The opinions of the individuals are aggregated separately for each object using an aggregation function that can depend on the object.
Maniquet and Mongin showed that if the aggregation functions are unanimous and the outcome must always be surjective, then the aggregation mechanism is dictatorial. Cailloux et al. showed that the same holds even if unanimity is relaxed to citizen sovereignty (each object can be classified into any category).
We show that similar results hold even if we only require the outcome to be surjective with probability $1-ε$ (with respect to an arbitrary symmetric i.i.d. distribution), provided that the aggregation functions are far from being constant.
On the way, we characterize all aggregation mechanisms whose outcome is always surjective without any assumptions on the aggregation functions.
Our approach uses a general result of Alekseev and Filmus which has wider applicability. We illustrate this by proving a similar impossibility result for aggregating equivalence relations.
Multicritical Scaling Limit of Shifted Schur Measure
We investigate the multicritical scaling limit of the shifted Schur measures. Under an appropriate scaling limit and specific conditions on the continuous parameters, we explicitly determine the limit shape of strict partitions distributed according to the shifted Schur measure. We then show that, under a multicritical condition, the edge scaling limit of the correlation function converges to a determinant of the higher-order Airy kernel. This rigorously demonstrates a transition from a Pfaffian point process to a determinantal distribution in the scaling limit.
Burnside process on parking functions and Dyck paths
Let $G$ be a finite group acting on a finite set $X$. This group action splits $X$ into disjoint orbits. The Burnside process is a Markov chain on $X$ which has a uniform stationary distribution when the chain is projected to orbits. We initiate the study of the Burnside process on Catalan structures. We consider two special cases: the first where the state space is the set of parking functions of length $n$ and $G = S_n$ is the symmetric group on $[n]$, such that $G$ acts by permuting coordinates, and the second where the state space is the set of labeled Dyck paths of length $2n$ and $G = S_n$ acts by permuting labels. The resulting Burnside processes give novel algorithms for sampling, respectively, an increasing parking function and a Dyck path approximately uniformly at random. Our main result shows that both processes are rapidly mixing, with mixing times upper bounded by $O(n \log n)$. As an application, we show how our Burnside process can be used to sample triangulations of an $(n+2)$-gon approximately uniformly at random.
Clumsy and Careless: Stationary-Entry Flux in Non-monotone Coupon Collectors
We study three nonmonotone coupon-collector models through a stationary-entry viewpoint. In such models the all-present state is not absorbing, so completion is governed not by the disappearance of a monotone terminal cloud but by rare new entries into a target state, except in the reset-button model, where exact regeneration gives a separate reduction.
We prove a finite stationary-entry theorem: a mixing estimate, a one-block clump-control estimate, and the stationary entry flux imply an exponential hitting law. For the reset-button collector, regeneration gives an exact probability-generating function in terms of the ordinary coupon-collector transform and recovers the known beta-function expectation, while also yielding rare-success exponential limits and negligible-reset Gumbel limits.
For the clumsy collector with fixed loss probability $p$ and $q=1-p$, the stationary-entry flux is $p q^n$, and $p q^n T_n$ converges to $\operatorname{Exp}(1)$. Thus the fixed-loss standardized limit is exponential rather than Gumbel. For the post-loss careless collector, we compute the sharp stationary-entry flux $$ μ_n\sim (q;q)_\infty^{-1}\frac{n!}{n^n}q^{n(n+1)/2} $$ and prove $μ_nT_n\Rightarrow\operatorname{Exp}(1)$, with matching moment asymptotics. This shows that the careless scale is governed by a stationary high tail, or ordered lucky climb, rather than by the independent one-point marginal heuristic. We also analyze a combined clumsy-careless model, confirming stability of the high-tail entry mechanism.
The critical activation density in graph bootstrap percolation
In graph bootstrap percolation, edges of an Erdős-Rényi random graph ${\mathcal G}_{n,p}$ are initially active. Activation spreads to other edges of the complete graph $K_n$ by an iterative process governed by a fixed graph $H$, whereby an edge becomes active whenever it is the only inactive edge in a copy of $H$. If all edges of $K_n$ are eventually activated, we say the process $H$-percolates. The case $H=K_3$ corresponds to the classical sharp threshold for connectivity in ${\mathcal G}_{n,p}$. When $H=K_4$, there are close connections with $2$-neighbor bootstrap percolation from statistical physics. Varying $H$ produces a wide range of behaviors.
In this work, for every graph $H$, we locate the critical $H$-percolation threshold $p_c(n,H)$, answering a question of Balogh, Bollobás, and Morris. Our general methods recover and improve several previous results. The location of $p_c(n,H)$ is related to a critical limiting density $ρ(H)$ of graphs that most efficiently activate a given edge. Introducing the parameter $ρ(H)$ raises several questions. For instance, it remains open whether $ρ(H)$ is computable in general, and its expression appears to indicate when the $H$-percolation threshold is sharp.
Generalization and Probabilistic Proofs of Some Combinatorial Identities
Using a probabilistic approach, we derive some interesting combinatorial identities involving gamma and beta functions. These results generalize certain well-known combinatorial identities involving binomial coefficients and special functions. In particular, by studying moments of the difference of two gamma and beta random variables, both in the dependent and independent cases, we obtain new combinatorial identities. This approach provides a systematic method to derive further combinatorial identities from probabilistic transformations.
When Does the Dice Sum Become Prime?
Given a (possibly infinite) subset $A$ of the natural numbers, we ask how many times a fair six-sided die must be rolled until the rolled numbers add up to an element of $A$. Using a one-dimensional dynamic programming recursion together with truncation and rigorous error bounds, we compute the expected number of rolls efficiently and with very high accuracy. When $A$ is the set of prime numbers, the irregular distribution of primes makes it difficult to obtain explicit error estimates. Nevertheless, the density of primes implies that the associated survival probability decays exponentially fast, which enables highly accurate truncation estimates. As a result, our calculations yield significantly sharper estimates for this expectation and its higher moments than the original results of Conroy, Alon, and Malinovsky. In particular, we determine the expectation to more than $1000$ decimal places.
A noisy min-max game on trees
We study a noisy version of a min-max type zero-sum game on the $d$-ary tree. Each edge of the tree is assigned an i.i.d.\ cookie, distributed uniformly on $\{+1,-1\}$. The game is played as follows: starting at the root, two players alternate turns in choosing a child to move to, with the game ending after each player took $n$ turns. Both players have full knowledge of the cookies on the whole tree. The cookies along the traversed edges are picked up and placed in a shared cookie jar. The first player's payoff is the sum of the cookies in the cookie jar, while the second player pays that sum. The value $V_n$ of the $n$-round game is the largest signed sum which can be guaranteed by the first player.
We analyze the value $V_n$ and show that as $n \to \infty$, the value is tight for $d=2$, converges in distribution for $d \ge 3$, and converges almost surely for $d \ge 15$. Along the way, we prove various tightness and double exponential tail decay results.
The analysis is a mix of percolation-type arguments for large $d$, and iterations on distributions combined with interval arithmetic for small $d$. For $d=2$ we prove the existence of a continuum of fixed points for this iteration, highlighting surprising qualitative differences with the case $d \ge 3$. The question of convergence for $d=2$ remains open.
On Talagrand's Convexity Conjecture
We prove that any centered $1$-subgaussian random vector in $\mathbb{R}^{n}$ can be written as the sum of a universal number of standard Gaussian vectors. Following the work of the second-named author, this solves M. Talagrand's convexity problem, which in turn implies a combinatorial analogue of the problem.
The stochastic block model has the overlap graph property for modularity
The overlap gap property (OGP) is a statement about the geometry of near-optimal solutions. Exhibiting OGP implies failure of a class of local algorithms; and has been observed to coincide with conjectured algorithmic limits in problems with statistical computational gap.
We consider the Stochastic Block Model (SBM), where the graph has a planted partition with $k$ equal-size blocks which form the `communities', and where, for parameters $p>q$, vertices within the same community connect with probability $p$, while vertices in different communities connect with probability $q$, independently across pairs of vertices. Modularity--based clustering algorithms have become ubiquitous in applications. This article studies theoretical limits of local algorithms based on the modularity score on the SBM.
We establish that modularity exhibits OGP on the SBM. This rules out a class of local algorithms based on modularity for recovery in the SBM, and shows slow mixing time for a related Markov Chain. Theoretically this is one of the few instances where OGP has been established for a `planted' model, as most such analyses to date consider the `null' model.
As part of our analysis, we extend a result by Bickel and Chen 2009, who established that with high probability, the modularity optimal partition of SBM is $o(n)$ local moves away from the planted partition, where $n$ is the graph size. We show that, with high probability, any partition with modularity score sufficiently near the optimal value is close to the planted partition.
Counting subgraphs in bounded-size Achlioptas processes
Achlioptas processes such as the Bohman--Frieze process are much harder to analyse than the classical Erdős--Rényi process, due to the dependence between edges added at different stages. This dependence means that most analysis so far is dynamic, often based on the differential equation method. In the Erdős--Rényi case there is an alternative static approach, pioneered by Erdős, Rényi and Bollobás, based on evaluating the expectation (and higher moments) of various subgraph counts, and using this to study the component structure. Here we show that this latter approach can be applied (with some complications) to the Bohman--Frieze process. For example, we are able to show that the expected number $μ_{k,t,n}$ of $k$-vertex tree components after $tn$ steps satisfies (essentially) $μ_{k,t,n}=c_{k,t}n(1+O(k/\sqrt{n}))$. Our method gives a very complicated formula for $c_{k,t}$, which seems to be unusable. However, since $c_{k,t}$ does not depend on $n$, we may use recent results obtained by the differential equation method and branching process analysis to find the asymptotics of $c_{k,t}$ as $k\to\infty$. The latter results also give a formula for $μ_{k,t,n}$ of the form $c_{k,t}n$ plus an error term, with a much more usable description of $c_{k,t}$ but a much worse error term. We combine the best of both worlds to prove a number of new results about the process near criticality. In particular, we obtain extremely sharp bounds on the size of the largest non-giant component near criticality, including the limiting distribution of its fluctuations.
The Ballot Event for Two-Player Coupon Collection: A Renewal--Catalan Asymptotic
We study the two-player coupon-collector competition in which two independent collectors draw one coupon each per round from a set of $d$ equally likely coupon types. Myers and Wilf gave finite formulae for several two-player events and explicitly left open the ballot-type problem of finding the probability that the ultimate winner was never behind. We prove that this probability satisfies $$ b_d \sim \frac{2}{d}, \qquad d\to\infty .$$ The proof uses a renewal decomposition at the tie boundary. The first one-sided tie-break has an explicit entrance distribution; its level, scaled by $d^{1/2}$, converges to a Rayleigh law; and, after the break, the leader's survival probability is governed by a Catalan, or gambler's-ruin, harmonic. The main estimate shows that the accumulated defect of this comparison harmonic in the exact simultaneous-round chain is negligible.
Geometry of Rényi Entropy on the Majorization Lattice
Majorization is a stochastic ordering relation that compares the relative diversity of probability distributions with numerous applications in econometrics, spectral theory, and ecology. It is well-known that the majorization partial order forms a complete lattice on the set of ordered probability distributions. In this work, we study the properties of Rényi entropy on the majorization lattice. We establish a fundamental relation between the comonotone coupling and the independent coupling associated with a collection of marginal distributions. Consequently, we show that, for every order $ α\in [0,\infty] $, the Rényi entropy is subadditive on the majorization lattice. We further characterize the supermodular regime, showing that Rényi entropy is supermodular on the majorization lattice for $ α\in \{0\} \cup [1,\infty] $.
Normal approximation of the numbers of isolated edges and isolated 2-stars in uniform simple graphs with given vertex degrees
We consider the configuration model and the uniform simple graph with given degree sequence $\boldsymbol{d}=\left(d_i\right)_{i=1}^n$. We derive quantitative bounds for the errors in (i) joint normal-Poisson approximation to the numbers of isolated edges, isolated 2-stars, self-loops and double edges in the configuration model, and (ii) normal approximation to the numbers of isolated edges and isolated 2-stars conditioned on that the configuration model is simple. The latter provides the first finite sample normal approximation results for the uniform simple graph with given vertex degrees. To achieve this, we develop a new Stein's method for joint normal-Poisson approximation and a new coupling approach to sums of indicators, which may be of independent interest.