probabilistic combinatorics
37 papers tagged with this keyword
Graph bootstrap percolation -- a discovery of slowness
Graph bootstrap percolation is a discrete-time process capturing the spread of a virus on the edges of $K_n$. Given an initial set $G\subseteq K_n$ of infected edges, the transmission of the virus is governed by a fixed graph $H$: in each round of the process any edge $e$ of $K_n$ that is the last uninfected edge in a copy of $H$ in $K_n$ gets infected as well. Once infected, edges remain infected forever. The process was introduced by Bollobás in 1968 in the context of weak saturation and has since inspired a vast array of beautiful mathematics. The main focus of this survey is the extremal question of how long the infection process can last before stabilising. We give an exposition of our recent systematic study of this maximum running time and the influence of the infection rule $H$. The topic turns out to possess a wide variety of interesting behaviour, with connections to additive, extremal and probabilistic combinatorics. Along the way we encounter a number of surprises and attractive open problems.
Probabilistic combinatorics at exponentially small scales
In many applications of the probabilistic method, one looks to study phenomena that occur ``with high probability''. More recently however, in an attempt to understand some of the most fundamental problems in combinatorics, researchers have been diving deeper into these probability spaces and understanding phenomena that occur at much smaller probability scales. Here I will survey a few of these ideas from the perspective of my own work in the area.
Asymptotic enumeration via graph containers and entropy
The container methods are powerful tools to bound the number of independent sets of graphs and hypergraphs, and they have been extremely influential in the area of extremal and probabilistic combinatorics. We will focus on more specialized graph container methods due to Sapozhenko (1987) that deal with sets in expander graphs. Entropy, first introduced by Shannon (1948) in the area of information theory, is a measure of the expected amount of information contained in a random variable. Entropy has seen lots of fascinating applications in a wide range of enumeration problems. In this survey article, we will discuss recent developments that exploit a combination of the two methods on enumerating graph homomorphisms.
On Graham's rearrangement conjecture over $\mathbb{F}_2^n$
A sequence $s_1,s_2,\ldots, s_k$ of elements of a group $G$ is called a valid ordering if the partial products $s_1, s_1 s_2, \ldots, s_1\cdots s_k$ are all distinct. A long-standing problem in combinatorial group theory asks whether, for a given group $G$, every subset $S \subseteq G\setminus \{\mathrm{id}\}$ admits a valid ordering; the instance of the additive group $\mathbb{F}_p$ is the content of a well-known 1971 conjecture of Graham. Most partial progress to date has concerned the edge cases where either $S$ or $G \setminus S$ is quite small. Our main result is an essentially complete resolution of the problem for $G=\mathbb{F}_2^n$: we show that there is an absolute constant $C>0$ such that every subset $S\subseteq \mathbb{F}_2^n \setminus \{0\}$ of size at least $C$ admits a valid ordering. Our proof combines techniques from additive and probabilistic combinatorics, including the Freiman--Ruzsa theorem and the absorption method.
Along the way, we also solve the general problem for moderately large subsets: there is a constant $c>0$ such that for every group $G$ (not necessarily abelian), every subset $S \subseteq G\setminus \{\mathrm{id}\}$ of size at least $|G|^{1-c}$ admits a valid ordering. Previous work in this direction concerned only sets of size at least $(1-o(1))|G|$. A main ingredient in our proof is a structural result, similar in spirit to the Arithmetic Regularity Lemma, showing that every Cayley graph can be efficiently decomposed into mildly quasirandom components.
The Fundamental Limits of Recovering Planted Subgraphs
Given an arbitrary subgraph $H=H_n$ and $p=p_n \in (0,1)$, the planted subgraph model is defined as follows. A statistician observes the union a random copy $H^*$ of $H$, together with random noise in the form of an instance of an Erdos-Renyi graph $G(n,p)$. Their goal is to recover the planted $H^*$ from the observed graph. Our focus in this work is to understand the minimum mean squared error (MMSE) for sufficiently large $n$.
A recent paper [MNSSZ23] characterizes the graphs for which the limiting MMSE curve undergoes a sharp phase transition from $0$ to $1$ as $p$ increases, a behavior known as the all-or-nothing phenomenon, up to a mild density assumption on $H$. In this paper, we provide a formula for the limiting MMSE curve for any graph $H=H_n$, up to the same mild density assumption. This curve is expressed in terms of a variational formula over pairs of subgraphs of $H$, and is inspired by the celebrated subgraph expectation thresholds from the probabilistic combinatorics literature [KK07]. Furthermore, we give a polynomial-time description of the optimizers of this variational problem. This allows one to efficiently approximately compute the MMSE curve for any dense graph $H$ when $n$ is large enough. The proof relies on a novel graph decomposition of $H$ as well as a new minimax theorem which may be of independent interest.
Our results generalize to the setting of minimax rates of recovering arbitrary monotone boolean properties planted in random noise, where the statistician observes the union of a planted minimal element $A \subseteq [N]$ of a monotone property and a random $Ber(p)^{\otimes N}$ vector. In this setting, we provide a variational formula inspired by the so-called "fractional" expectation threshold [Tal10], again describing the MMSE curve (in this case up to a multiplicative constant) for large enough $n$.
Connectivity for square percolation and coarse cubical rigidity in random right-angled Coxeter groups
We consider random right-angled Coxeter groups, $W_Γ$, whose presentation graph $Γ$ is taken to be an Erdős--Rényi random graph, i.e., $Γ\sim \mathcal{G}_{n,p}$. We use techniques from probabilistic combinatorics to establish several new results about the geometry of these random groups.
We resolve a conjecture of Susse and determine the connectivity threshold for square percolation on the random graph $Γ\sim \mathcal{G}_{n,p}$. We use this result to determine a large range of $p$ for which the random right-angled Coxeter group $W_Γ$ has a unique cubical coarse median structure. Until recent work of Fioravanti, Levcovitz and Sageev, there were no non-hyperbolic examples of groups with cubical coarse rigidity; our present results show the property is in fact typically satisfied by a random RACG for a wide range of the parameter $p$, including $p=1/2$.
Borel Local Lemma: arbitrary random variables and limited exponential growth
The Lovász Local Lemma (the LLL for short) is a powerful tool in probabilistic combinatorics that is used to verify the existence of combinatorial objects with desirable properties. Recent years saw the development of various "constructive" versions of the LLL. A major success of this research direction is the Borel version of the LLL due to Csóka, Grabowski, Máthé, Pikhurko, and Tyros, which holds under a subexponential growth assumption. A drawback of their approach is that it only applies when the underlying random variables take values in a finite set. We present an alternative proof of a Borel version of the LLL that holds even if the underlying random variables are continuous and applies to dependency graphs of limited exponential growth.
Sampling and counting triangle-free graphs near the critical density
We study the following combinatorial counting and sampling problems: can we efficiently sample from the Erdős-Rényi random graph $G(n,p)$ conditioned on triangle-freeness? Can we efficiently approximate the probability that $G(n,p)$ is triangle-free? These are prototypical instances of forbidden substructure problems ubiquitous in combinatorics. The algorithmic questions are instances of approximate counting and sampling for a hypergraph hard-core model.
Estimating the probability that $G(n,p)$ has no triangles is a fundamental question in probabilistic combinatorics and one that has led to the development of many important tools in the field. Through the work of several authors, the asymptotics of the logarithm of this probability are known if $p =o( n^{-1/2})$ or if $p =ω( n^{-1/2})$. The regime $p = Θ(n^{-1/2})$ is more mysterious, as this range witnesses a dramatic change in the the typical structural properties of $G(n,p)$ conditioned on triangle-freeness. As we show, this change in structure has a profound impact on the performance of sampling algorithms.
We give two different efficient sampling algorithms for triangle-free graphs (and complementary algorithms to approximate the triangle-freeness large deviation probability), one that is efficient when $p < c/\sqrt{n}$ and one that is efficient when $p > C/\sqrt{n}$ for constants $c, C>0$. The latter algorithm involves a new approach for dealing with large defects in the setting of sampling from low-temperature spin models.
Towards an optimal hypergraph container lemma
The hypergraph container lemma is a powerful tool in probabilistic combinatorics that has found many applications since it was first proved a decade ago. Roughly speaking, it asserts that the family of independent sets of every uniform hypergraph can be covered by a small number of almost-independent sets, called containers. In this article, we formulate and prove two new versions of the lemma that display the following three attractive features. First, they both admit short and simple proofs that have surprising connections to other well-studied topics in probabilistic combinatorics. Second, they use alternative notions of almost-independence in order to describe the containers. Third, they yield improved dependence of the number of containers on the uniformity of the hypergraph, hitting a natural barrier for second-moment-type approaches.
Game Connectivity and Adaptive Dynamics
We analyse the typical structure of games in terms of the connectivity properties of their best-response graphs. Our central result shows that, among games that are `generic' (without indifferences) and that have a pure Nash equilibrium, all but a small fraction are \emph{connected}, meaning that every action profile that is not a pure Nash equilibrium can reach every pure Nash equilibrium via best-response paths. This has important implications for dynamics in games. In particular, we show that there are simple, uncoupled, adaptive dynamics for which period-by-period play converges almost surely to a pure Nash equilibrium in all but a small fraction of generic games that have one (which contrasts with the known fact that there is no such dynamic that leads almost surely to a pure Nash equilibrium in \emph{every} generic game that has one). We build on recent results in probabilistic combinatorics for our characterisation of game connectivity.
Borel versions of the Local Lemma and LOCAL algorithms for graphs of finite asymptotic separation index
Asymptotic separation index is a parameter that measures how easily a Borel graph can be approximated by its subgraphs with finite components. In contrast to the more classical notion of hyperfiniteness, asymptotic separation index is well-suited for combinatorial applications in the Borel setting. The main result of this paper is a Borel version of the Lovász Local Lemma -- a powerful general-purpose tool in probabilistic combinatorics -- under a finite asymptotic separation index assumption. As a consequence, we show that locally checkable labeling problems that are solvable by efficient randomized distributed algorithms admit Borel solutions on bounded degree Borel graphs with finite asymptotic separation index. From this we derive a number of corollaries, for example a Borel version of Brooks's theorem for graphs with finite asymptotic separation index.
Anticoncentration in Ramsey graphs and a proof of the Erdős-McKay conjecture
Published
• View Publication
• BIB
An $n$-vertex graph is called $C$-Ramsey if it has no clique or independent set of size $C\log_2 n$ (i.e., if it has near-optimal Ramsey behavior). In this paper, we study edge-statistics in Ramsey graphs, in particular obtaining very precise control of the distribution of the number of edges in a random vertex subset of a $C$-Ramsey graph. This brings together two ongoing lines of research: the study of "random-like" properties of Ramsey graphs and the study of small-ball probabilities for low-degree polynomials of independent random variables.
The proof proceeds via an "additive structure" dichotomy on the degree sequence, and involves a wide range of different tools from Fourier analysis, random matrix theory, the theory of Boolean functions, probabilistic combinatorics, and low-rank approximation. One of the consequences of our result is the resolution of an old conjecture of Erdős and McKay, for which Erdős offered one of his notorious monetary prizes.
A Smoother Notion of Spread Hypergraphs
Published
• View Publication
• BIB
Alweiss, Lovett, Wu, and Zhang introduced $q$-spread hypergraphs in their breakthrough work regarding the sunflower conjecture, and since then $q$-spread hypergraphs have been used to give short proofs of several outstanding problems in probabilistic combinatorics. A variant of $q$-spread hypergraphs was implicitly used by Kahn, Narayanan, and Park to determine the threshold for when a square of a Hamiltonian cycle appears in the random graph $G_{n,p}$. In this paper we give a common generalization of the original notion of $q$-spread hypergraphs and the variant used by Kahn et al.
Thresholds versus fractional expectation-thresholds
Published
• View Publication
• BIB
Proving a conjecture of Talagrand, a fractional version of the 'expectation-threshold' conjecture of Kalai and the second author, we show for any increasing family $F$ on a finite set $X$ that $p_c (F) =O( q_f (F) \log \ell(F))$, where $p_c(F)$ and $q_f(F)$ are the threshold and 'fractional expectation-threshold' of $F$, and $\ell(F)$ is the largest size of a minimal member of $F$. This easily implies several heretofore difficult results and conjectures in probabilistic combinatorics, including thresholds for perfect hypergraph matchings (Johansson--Kahn--Vu), bounded-degree spanning trees (Montgomery), and bounded-degree spanning graphs (new). We also resolve (and vastly extend) the 'axial' version of the random multi-dimensional assignment problem (earlier considered by Martin--Mézard--Rivoire and Frieze--Sorkin). Our approach builds on a recent breakthrough of Alweiss, Lovett, Wu and Zhang on the Erdős--Rado 'Sunflower Conjecture'.
A Short Proof of Bernoulli Disjointness via the Local Lemma
Published
• View Publication
• BIB
Recently, Glasner, Tsankov, Weiss, and Zucker showed that if $Γ$ is an infinite discrete group, then every minimal $Γ$-flow is disjoint from the Bernoulli shift $2^Γ$. Their proof is somewhat involved; in particular, it invokes separate arguments for different classes of groups. In this note, we give a short and self-contained proof of their result using purely combinatorial methods applicable to all groups at once. Our proof relies on the Lovász Local Lemma, an important tool in probabilistic combinatorics that has recently found several applications in the study of dynamical systems.
Upper tails via high moments and entropic stability
Suppose that $X$ is a bounded-degree polynomial with nonnegative coefficients on the $p$-biased discrete hypercube. Our main result gives sharp estimates on the logarithmic upper tail probability of $X$ whenever an associated extremal problem satisfies a certain entropic stability property. We apply this result to solve two long-standing open problems in probabilistic combinatorics: the upper tail problem for the number of arithmetic progressions of a fixed length in the $p$-random subset of the integers and the upper tail problem for the number of cliques of a fixed size in the random graph $G_{n,p}$. We also make significant progress on the upper tail problem for the number of copies of a fixed regular graph $H$ in $G_{n,p}$. To accommodate readers who are interested in learning the basic method, we include a short, self-contained solution to the upper tail problem for the number of triangles in $G_{n,p}$ for all $p=p(n)$ satisfying $n^{-1}\log n\ll p \ll 1$.
Triangle resilience of the square of a Hamilton cycle in random graphs
Published
• View Publication
• BIB
Since first introduced by Sudakov and Vu in 2008, the study of resilience problems in random graphs received a lot of attention in probabilistic combinatorics. Of particular interest are resilience problems of spanning structures. It is known that for spanning structures which contain many triangles, local resilience cannot prevent an adversary from destroying all copies of the structure by removing a negligible amount of edges incident to every vertex. In this paper we generalise the notion of local resilience to $H$-resilience and demonstrate its usefulness on the containment problem of the square of a Hamilton cycle. In particular, we show that there exists a constant $C > 0$ such that if $p \geq C\log^3 n/\sqrt{n}$ then w.h.p. in every subgraph $G$ of a random graph $G_{n, p}$ there exists the square of a Hamilton cycle, provided that every vertex of $G$ remains on at least a $(4/9 + o(1))$-fraction of its triangles from $G_{n, p}$. The constant $4/9$ is optimal and the value of $p$ slightly improves on the best-known appearance threshold of such a structure and is optimal up to the logarithmic factor.
Threshold Progressions in a Variety of Covering and Packing Contexts
Using standard methods (due to Janson, Stein-Chen, and Talagrand) from probabilistic combinatorics, we explore the following general theme: As one progresses from each member of a family of objects ${\cal A}$ being "covered" by at most one object in a random collection ${\cal C}$, to being covered at most $λ$ times, to being covered at least once, to being covered at least $λ$ times, a hierarchy of thresholds emerge. We will then see how such results vary according to the context, and level of dependence introduced. Examples will be from extremal set theory, combinatorics, and additive number theory.
Building Large Free Subshifts Using the Local Lemma
Published
• View Publication
• BIB
Gao, Jackson, and Seward proved that every countably infinite group $Γ$ admits a nonempty free subshift $X \subseteq 2^Γ$. Here we strengthen this result by showing that free subshifts can be "large" in various senses. Specifically, we prove that for any $k \geqslant 2$ and $h < \log_2 k$, there exists a free subshift $X \subseteq k^Γ$ of Hausdorff dimension and, if $Γ$ is sofic, entropy at least $h$, answering two questions attributed by Gao, Jackson, and Seward to Juan Souto. Furthermore, we establish a general lower bound on the largest "size" of a free subshift $X'$ contained inside a given subshift $X$. A central role in our arguments is played by the Lovász Local Lemma, an important tool in probabilistic combinatorics, whose relevance to the problem of finding free subshifts was first recognized by Aubrun, Barbieri, and Thomassé.
Positive association of the oriented percolation cluster in randomly oriented graphs
Published in Combinatorics, Probability and Computing, 28(6):811-815 (2019)
• View Publication
• BIB
Consider any fixed graph whose edges have been randomly and independently oriented, and write $\{S \leadsto i\}$ to indicate that there is an oriented path going from a vertex $s \in S$ to vertex $i$. Narayanan (2016) proved that for any set $S$ and any two vertices $i$ and $j$, $\{S \leadsto i\}$ and $\{S \leadsto j\}$ are positively correlated. His proof relies on the Ahlswede-Daykin inequality, a rather advanced tool of probabilistic combinatorics.
In this short note, I give an elementary proof of the following, stronger result: writing $V$ for the vertex set of the graph, for any source set $S$, the events $\{S \leadsto i\}$, $i \in V$, are positively associated -- meaning that the expectation of the product of increasing functionals of the family $\{S \leadsto i\}$ for $i \in V$ is greater than the product of their expectations.