random
6952 papers tagged with this keyword
A canonical Ramsey theorem with list constraints in random (hyper-)graphs
The celebrated canonical Ramsey theorem of Erdős and Rado implies that for a given $k$-uniform hypergraph (or $k$-graph) $H$, if $n$ is sufficiently large then any colouring of the edges of the complete $k$-graph $K^{(k)}_n$ gives rise to copies of $H$ that exhibit certain colour patterns. We are interested in sparse random versions of this result and the threshold at which the random $k$-graph ${\mathbf{G}}^{(k)}(n,p)$ inherits the canonical Ramsey properties of $K^{(k)}_n$. Our main result here pins down this threshold when we focus on colourings that are constrained by some prefixed lists. This result is applied in an accompanying work of the authors on the threshold for the canonical Ramsey property (with no list constraints) in the case that $H$ is a (2-uniform) even cycle.
Strong spatial mixing for colorings on trees and its algorithmic applications
Strong spatial mixing (SSM) is an important quantitative notion of correlation decay for Gibbs distributions arising in statistical physics, probability theory, and theoretical computer science. A longstanding conjecture is that the uniform distribution on proper $q$-colorings on a $Δ$-regular tree exhibits SSM whenever $q \ge Δ+1$. Moreover, it is widely believed that as long as SSM holds on bounded-degree trees with $q$ colors, one would obtain an efficient sampler for $q$-colorings on all bounded-degree graphs via simple Markov chain algorithms. It is surprising that such a basic question is still open, even on trees, but then again it also highlights how much we still have to learn about random colorings. In this paper, we show the following:
(1) For any $Δ\ge 3$, SSM holds for random $q$-colorings on trees of maximum degree $Δ$ whenever $q \ge Δ+ 3$. Thus we almost fully resolve the aforementioned conjecture. Our result substantially improves upon the previously best bound which requires $q \ge 1.59Δ+γ^*$ for an absolute constant $γ^* > 0$.
(2) For any $Δ\ge 3$ and girth $g = Ω_Δ(1)$, we establish optimal mixing of the Glauber dynamics for $q$-colorings on graphs of maximum degree $Δ$ and girth $g$ whenever $q \ge Δ+3$. Our approach is based on a new general reduction from spectral independence on large-girth graphs to SSM on trees that is of independent interest.
Using the same techniques, we also prove near-optimal bounds on weak spatial mixing (WSM), a closely-related notion to SSM, for the antiferromagnetic Potts model on trees.
Hypergraph Animals
Here we introduce simple structures for the analysis of complex hypergraphs, hypergraph animals. These structures are designed to describe the local node neighbourhoods of nodes in hypergraphs. We establish their relationships to lattice animals and network motifs, and we develop their combinatorial properties for sparse and uncorrelated hypergraphs. We make use of the tight link of hypergraph animals to partition numbers, which opens up a vast mathematical framework for the analysis of hypergraph animals. We then study their abundances in random hypergraphs. Two transferable insights result from this analysis: (i) it establishes the importance of high-cardinality edges in ensembles of random hypergraphs that are inspired by the classical Erdös-Renyí random graphs; and (ii) there is a close connection between degree and hyperedge cardinality in random hypergraphs that shapes animal abundances and spectra profoundly. Both findings imply that hypergraph animals can have the potential to affect information flow and processing in complex systems. Our analysis of also suggests that we need to spend more effort on investigating and developing suitable conditional ensembles of random hypergraphs that can capture real-world structures and their complex dependency structures.
Maximum Agreement Subtrees and Hölder homeomorphisms between Brownian trees
We prove that the size of the largest common subtree between two uniform, independent, leaf-labelled random binary trees of size $n$ is typically less than $n^{1/2-\varepsilon}$ for some $\varepsilon>0$. Our proof relies on the coupling between discrete random trees and the Brownian tree and on a recursive decomposition of the Brownian tree due to Aldous. Along the way, we also show that almost surely, there is no $(1-\varepsilon)$-Hölder homeomorphism between two independent copies of the Brownian tree.
On a generalisation of the coupon collector problem
We consider a generalisation of the classical coupon collector problem. We define a super-coupon to be any $s$-subset of a universe of $n$ coupons. In each round, a random $r$-subset from the universe is drawn and all its $s$-subsets are marked as collected. We show that the time to collect all super-coupons is $\binom{r}{s}^{-1}\binom{n}{s} \log \binom{n}{s}(1 + o(1))$ on average and has a Gumbel limit after a suitable normalisation. In a similar vein, we show that for any $α\in (0, 1)$, the expected time to collect $(1 - α)$ proportion of all super-coupons is $\binom{r}{s}^{-1}\binom{n}{s} \log \big(\frac{1}α\big)(1 + o(1))$. The $r = s$ case of this model is equivalent to the classical coupon collector model.
We also consider a temporally dependent model where the $r$-subsets are drawn according to the following Markovian dynamics: the $r$-subset at round $k + 1$ is formed by replacing a random coupon from the $r$-subset drawn at round $k$ with another random coupon from outside this $r$-subset. We link the time it takes to collect all super-coupons in the $r = s$ case of this model to the cover time of random walk on a certain finite regular graph and conjecture that in general, it takes $\frac{r}{s} \binom{r}{s}^{-1}\binom{n}{s}\log\binom{n}{s}(1 + o(1))$ time on average to collect all super-coupons.
Universality in prelimiting tail behavior for regular subgraph counts in the Poisson regime
Let $N$ be the number of copies of a small subgraph $H$ in an Erdős-Rényi graph $G \sim \mathcal{G}(n, p_n)$ where $p_n \to 0$ is chosen so that $\mathbb{E} N = c$, a constant. Results of Bollobás show that for regular graphs $H$, the count $N$ weakly converges to a Poisson random variable. For large but finite $n$, and for the specific case of the triangle, investigations of the upper tail $\mathbb{P}(N \geq k_n)$ by Ganguly, Hiesmayr and Nam (2022) revealed that there is a phase transition in the tail behavior and the associated mechanism. Smaller values of $k_n$ correspond to disjoint occurrences of $H$, leading to Poisson tails, with a different behavior emerging when $k_n$ is large, guided by the appearance of an almost clique. We show that a similar phase transition also occurs when $H$ is any regular graph, at the point where $k_n^{1 -2/q}\log k_n = \log n$ ($q$ is the number of vertices in $H$). This establishes universality of this transition, previously known only for the case of the triangle.
Asymptotics of dimer coverings on free boundary rail-yard graphs
Rail-yard graphs are a general class of graphs introduced in \cite{bbccr} on which the random dimer coverings form Schur processes. We study asymptotic limits of random dimer coverings on rail yard graphs with free boundary conditions on both the left boundary and the right boundary (double-sided free boundary) when the mesh sizes of the graphs go to 0. Each dimer covering corresponds to a sequence of interlacing partitions starting with an arbitrary partition and ending in an arbitrary partition. Under the assumption that the probability of each dimer covering is proportional to the product of weights of present edges, we obtain the moment formula for the height function which includes an infinite product. By passing down to the scaling limit, we compute the limit shape (law of large numbers) of the rescaled height functions and prove the convergence of unrescaled height fluctuations to a diffeomorphic image of the restriction of the 0-boundary Gaussian free field (central limit theorem) on the upper half plane to a subset. Applications include the limit shape and height fluctuations for free boundary steep tilings as proposed in \cite{BCC17}. The technique to obtain these results is to analyze a class of Macdonald processes with dual specializations, subject to further complexities arising from the infinite product in the moment formula.
We also obtain a new algorithm to sample double-sided free boundary dimer coverings on rail-yard graphs, which fulfills an open problem in \cite{bbbccv14}.
Matrix Perturbation: Davis-Kahan in the Infinity Norm
Perturbation theory is developed to analyze the impact of noise on data and has been an essential part of numerical analysis. Recently, it has played an important role in designing and analyzing matrix algorithms. One of the most useful tools in this subject, the Davis-Kahan sine theorem, provides an $\ell_2$ error bound on the perturbation of the leading singular vectors (and spaces).
We focus on the case when the signal matrix has low rank and the perturbation is random, which occurs often in practice. In an earlier paper, O'Rourke, Wang, and the second author showed that in this case, one can obtain an improved theorem. In particular, the noise-to-gap ratio condition in the original setting can be weakened considerably.
In the current paper, we develop an infinity norm version of the O'Rourke-Vu-Wang result. The key ideas in the proof are a new bootstrapping argument and the so-called iterative leave-one-out method, which may be of independent interest.
Applying the new bounds, we develop new, simple, and quick algorithms for several well-known problems, such as finding hidden partitions and matrix completion. The core of these new algorithms is the fact that one is now able to quickly approximate certain key objects in the infinity norm, which has critical advantages over approximations in the $\ell_2$ norm, Frobenius norm, or spectral norm.
Dispersion entropy: A Measure of Irregularity for Graph Signals
We introduce a novel method, called Dispersion Entropy for Graph Signals, $DE_G$, as a powerful tool for analysing the irregularity of signals defined on graphs. We demonstrate the effectiveness of $DE_G$ in detecting changes in the dynamics of signals defined on synthetic and real-world graphs, by defining mixed processing on random geometric graphs or those exhibiting with small-world properties. Remarkably, $DE_G$ generalises the classical dispersion entropy for univariate time series, enabling its application in diverse domains such as image processing, time series analysis, and network analysis, as well as in establishing theoretical relationships (i.e., graph centrality measures, spectrum). Our results indicate that $DE_G$ effectively captures the irregularity of graph signals across various network configurations, successfully differentiating between distinct levels of randomness and connectivity. Consequently, $DE_G$ provides a comprehensive framework for entropy analysis of various data types, enabling new applications of dispersion entropy not previously feasible, and revealing relationships between graph signals and its graph topology.
Isoperimetric Inequalities and Supercritical Percolation on High-dimensional Graphs
It is known that many different types of finite random subgraph models undergo quantitatively similar phase transitions around their percolation thresholds, and the proofs of these results rely on isoperimetric properties of the underlying host graph. Recently, the authors showed that such a phase transition occurs in a large class of regular high-dimensional product graphs, generalising a classic result for the hypercube.
In this paper we give new isoperimetric inequalities for such regular high-dimensional product graphs, which generalise the well-known isoperimetric inequality of Harper for the hypercube, and are asymptotically sharp for a wide range of set sizes. We then use these isoperimetric properties to investigate the structure of the giant component $L_1$ in supercritical percolation on these product graphs, that is, when $p=\frac{1+ε}{d}$, where $d$ is the degree of the product graph and $ε>0$ is a small enough constant.
We show that typically $L_1$ has edge-expansion $Ω\left(\frac{1}{d\ln d}\right)$. Furthermore, we show that $L_1$ likely contains a linear-sized subgraph with vertex-expansion $Ω\left(\frac{1}{d\ln d}\right)$. These results are best possible up to the logarithmic factor in $d$.
Using these likely expansion properties, we determine, up to small polylogarithmic factors in $d$, the likely diameter of $L_1$ as well as the typical mixing time of a lazy random walk on $L_1$. Furthermore, we show the likely existence of a path of length $Ω\left(\frac{n}{d\ln d}\right)$. These results not only generalise, but also improve substantially upon the known bounds in the case of the hypercube, where in particular the likely diameter and typical mixing time of $L_1$ were previously only known to be polynomial in $d$.
Random clique complex process inside the critical window
We consider the random clique complex process - the process of clique complexes induced by the complete graph with i.i.d. Uniform edge weights. We investigate the evolution of the Betti numbers of the clique complex process in the critical window and in particular, show a process-level convergence of the Betti numbers to a Poisson process. Our proof technique gives easily an hitting time result i.e, with high probability, the $k$th cohomology becomes trivial when there are no more isolated $k$-faces. Our results imply that the thresholds for vanishing of cohomology of the clique complex process coincides with that of the threshold for vanishing of `instantaneous' homology determined by \citet{SVT}. We also give a lower bound for the probability of clique complex process to have Kazhdan's property $(T)$. These results show a different behaviour for the clique complex process compared to the Čech complex process investigated in the geometric setting by \citet{B19}.
Pseudorandom Linear Codes are List Decodable to Capacity
We introduce a novel family of expander-based error correcting codes. These codes can be sampled with randomness linear in the block-length, and achieve list-decoding capacity (among other local properties). Our expander-based codes can be made starting from any family of sufficiently low-bias codes, and as a consequence, we give the first construction of a family of algebraic codes that can be sampled with linear randomness and achieve list-decoding capacity. We achieve this by introducing the notion of a pseudorandom puncturing of a code, where we select $n$ indices of a base code $C\subset \mathbb{F}_q^m$ via an expander random walk on a graph on $[m]$. Concretely, whereas a random linear code (i.e. a truly random puncturing of the Hadamard code) requires $O(n^2)$ random bits to sample, we sample a pseudorandom linear code with $O(n)$ random bits. We show that pseudorandom puncturings satisfy several desirable properties exhibited by truly random puncturings. In particular, we extend a result of (Guruswami Mosheiff FOCS 2022) and show that a pseudorandom puncturing of a small-bias code satisfies the same local properties as a random linear code with high probability. As a further application of our techniques, we also show that pseudorandom puncturings of Reed Solomon codes are list-recoverable beyond the Johnson bound, extending a result of (Lund Potukuchi RANDOM 2020). We do this by instead analyzing properties of codes with large distance, and show that pseudorandom puncturings still work well in this regime.
Spectral pseudorandomness and the road to improved clique number bounds for Paley graphs
We study subgraphs of Paley graphs of prime order $p$ induced on the sets of vertices extending a given independent set of size $a$ to a larger independent set. Using a sufficient condition proved in the author's recent companion work, we show that a family of character sum estimates would imply that, as $p \to \infty$, the empirical spectral distributions of the adjacency matrices of any sequence of such subgraphs have the same weak limit (after rescaling) as those of subgraphs induced on a random set including each vertex independently with probability $2^{-a}$, namely, a Kesten-McKay law with parameter $2^a$. We prove the necessary estimates for $a = 1$, obtaining in the process an alternate proof of a character sum equidistribution result of Xi (2022), and provide numerical evidence for this weak convergence for $a \geq 2$. We also conjecture that the minimum eigenvalue of any such sequence converges (after rescaling) to the left edge of the corresponding Kesten-McKay law, and provide numerical evidence for this convergence. Finally, we show that, once $a \geq 3$, this (conjectural) convergence of the minimum eigenvalue would imply bounds on the clique number of the Paley graph improving on the current state of the art due to Hanson and Petridis (2021), and that this convergence for all $a \geq 1$ would imply that the clique number is $o(\sqrt{p})$.
Power-law bounds for increasing subsequences in Brownian separable permutons and homogeneous sets in Brownian cographons
The Brownian separable permutons are a one-parameter family -- indexed by $p\in(0,1)$ -- of universal limits of random constrained permutations. We show that for each $p\in (0,1)$, there are explicit constants $1/2 < α_*(p) \leq β^*(p) < 1$ such that the length of the longest increasing subsequence in a random permutation of size $n$ sampled from the Brownian separable permuton is between $n^{α_*(p) - o(1)}$ and $n^{β^*(p) + o(1)}$ with probability tending to 1 as $n\to\infty$. In the symmetric case $p=1/2$, we have $α_*(p) \approx 0.812$ and $β^*(p)\approx 0.975$. We present numerical simulations which suggest that the lower bound $α_*(p)$ is close to optimal in the whole range $p\in(0,1)$.
Our results work equally well for the closely related Brownian cographons. In this setting, we show that for each $p\in (0,1)$, the size of the largest clique (resp. independent set) in a random graph on $n$ vertices sampled from the Brownian cographon is between $n^{α_*(p) - o(1)}$ and $n^{β^*(p) + o(1)}$ (resp. $n^{α_*(1-p) - o(1)}$ and $n^{β^*(1-p) + o(1)}$) with probability tending to 1 as $n\to\infty$.
Our proofs are based on the analysis of a fragmentation process embedded in a Brownian excursion introduced by Bertoin (2002). We expect that our techniques can be extended to prove similar bounds for uniform separable permutations and uniform cographs.
Signal processing on large networks with group symmetries
Current methods of graph signal processing rely heavily on the specific structure of the underlying network: the shift operator and the graph Fourier transform are both derived directly from a specific graph. In many cases, the network is subject to error or natural changes over time. This motivated a new perspective on GSP, where the signal processing framework is developed for an entire class of graphs with similar structures. This approach can be formalized via the theory of graph limits, where graphs are considered as random samples from a distribution represented by a graphon.
When the network under consideration has underlying symmetries, they may be modeled as samples from Cayley graphons. In Cayley graphons, vertices are sampled from a group, and the link probability between two vertices is determined by a function of the two corresponding group elements. Infinite groups such as the 1-dimensional torus can be used to model networks with an underlying spatial reality. Cayley graphons on finite groups give rise to a Stochastic Block Model, where the link probabilities between blocks form a (edge-weighted) Cayley graph. This manuscript summarizes some work on graph signal processing on large networks, in particular samples of Cayley graphons.
Uniformly Random Colourings of Sparse Graphs
Published
• View Publication
• BIB
We analyse uniformly random proper $k$-colourings of sparse graphs with maximum degree $Δ$ in the regime $Δ< k\ln k $. This regime corresponds to the lower side of the shattering threshold for random graph colouring, a paradigmatic example of the shattering threshold for random Constraint Satisfaction Problems. We prove a variety of results about the solution space geometry of colourings of fixed graphs, generalising work of Achlioptas, Coja-Oghlan, and Molloy on random graphs, and justifying the performance of stochastic local search algorithms in this regime. Our central proof relies only on elementary techniques, namely the first-moment method and a quantitative induction, yet it strengthens list-colouring results due to Vu, and more recently Davies, Kang, P., and Sereni, and generalises state-of-the-art bounds from Ramsey theory in the context of sparse graphs. It further yields an approximately tight lower bound on the number of colourings, also known as the partition function of the Potts model, with implications for efficient approximate counting.
Limits of polyhedral multinomial distributions
We consider limits of certain measures supported on lattice points in lattice polyhedra defined as the intersection of half-spaces $\{m\in\mathbb{R}^n|\langle v_i,x\rangle+a_i \geq 0\}$, where $\sum_i v_i = 0$. The measures are densities associated to lattice random variables obtained by restriction of multinomial random variables. We find the limiting Gaussian distributions explicitly.
Asymptotic analysis and efficient random sampling of directed ordered acyclic graphs
Directed acyclic graphs (DAGs) are directed graphs in which there is no path from a vertex to itself. DAGs are an omnipresent data structure in computer science and the problem of counting the DAGs of given number of vertices and to sample them uniformly at random has been solved respectively in the 70's and the 00's. In this paper, we propose to explore a new variation of this model where DAGs are endowed with an independent ordering of the out-edges of each vertex, thus allowing to model a wide range of existing data structures. We provide efficient algorithms for sampling objects of this new class, both with or without control on the number of edges, and obtain an asymptotic equivalent of their number. We also show the applicability of our method by providing an effective algorithm for the random generation of classical labelled DAGs with a prescribed number of vertices and edges, based on a similar approach. This is the first known algorithm for sampling labelled DAGs with full control on the number of edges, and it meets a need in terms of applications, that had already been acknowledged in the literature.
Rectangular matrix additions in low and high temperatures
We study the addition of two random independent $M\times N$ rectangular random matrices with invariant distributions in two limit regimes, where the parameter beta (inverse temperature) goes to infinity and zero. In low temperature regime the random singular values of the sum concentrate at deterministic points, while in high temperature regime we obtain a Law of Large Numbers for the empirical measures. As a consequence, we deliver a duality between low and high temperatures. Our proof uses the type BC Bessel function as characteristic function of rectangular matrices, and through the analysis of this function we introduce a new family of cumulants, that linearize the addition in high temperature limit, and degenerate to the classical or free cumulants in special cases.
Sharp threshold for embedding balanced spanning trees in random geometric graphs
A rooted tree is balanced if the degree of a vertex depends only on its distance to the root. In this paper we determine the sharp threshold for the appearance of a large family of balanced spanning trees in the random geometric graph $\mathcal{G}(n,r,d)$. In particular, we find the sharp threshold for balanced binary trees. More generally, we show that all sequences of balanced trees with uniformly bounded degrees and height tending to infinity appear above a sharp threshold, and none of these appears below the same value. Our results hold more generally for geometric graphs satisfying a mild condition on the distribution of their vertex set, and we provide a polynomial time algorithm to find such trees.