random
6952 papers tagged with this keyword
Smoothed analysis for graph isomorphism
There is no known polynomial-time algorithm for graph isomorphism testing, but elementary combinatorial "refinement" algorithms seem to be very efficient in practice. Some philosophical justification is provided by a classical theorem of Babai, Erdős and Selkow: an extremely simple polynomial-time combinatorial algorithm (variously known as "naïve refinement", "naïve vertex classification", "colour refinement" or the "1-dimensional Weisfeiler-Leman algorithm") yields a so-called canonical labelling scheme for "almost all graphs". More precisely, for a typical outcome of a random graph $G(n,1/2)$, this simple combinatorial algorithm assigns labels to vertices in a way that easily permits isomorphism-testing against any other graph.
We improve the Babai-Erdős-Selkow theorem in two directions. First, we consider randomly perturbed graphs, in accordance with the smoothed analysis philosophy of Spielman and Teng: for any graph $G$, naïve refinement becomes effective after a tiny random perturbation to $G$ (specifically, the addition and removal of $O(n\log n)$ random edges). Actually, with a twist on naïve refinement, we show that $O(n)$ random additions and removals suffice. These results significantly improve on previous work of Gaudio-Rácz-Sridhar, and are in certain senses best-possible.
Second, we complete a long line of research on canonical labelling of random graphs: for any $p$ (possibly depending on $n$), we prove that a random graph $G(n,p)$ can typically be canonically labelled in polynomial time. This is most interesting in the extremely sparse regime where $p$ has order of magnitude $c/n$; denser regimes were previously handled by Bollobás, Czajka-Pandurangan, and Linial-Mosheiff. Our proof also provides a description of the automorphism group of a typical outcome of $G(n,p_n)$ (slightly correcting a prediction of Linial-Mosheiff).
Spread blow-up lemma with an application to perturbed random graphs
Combining ideas of Pham, Sah, Sawhney, and Simkin on spread perfect matchings in super-regular bipartite graphs with an algorithmic blow-up lemma, we prove a spread version of the blow-up lemma. Intuitively, this means that there exists a probability measure over copies of a desired spanning graph $H$ in a given system of super-regular pairs which does not heavily pin down any subset of vertices. This allows one to complement the use of the blow-up lemma with the recently resolved Kahn-Kalai conjecture. As an application, we prove an approximate version of a conjecture of Böttcher, Parczyk, Sgueglia, and Skokan on the threshold for appearance of powers of Hamilton cycles in perturbed random graphs.
On the $H$-space of a random graph
The edge space $\mathcal{E}(G)$ of a graph $G$ is the vector space $\mathbb{F}_2^{E(G)}$ with members naturally identified with subgraphs of $G$, and the $H$-space is the subspace $\mathcal{C}_H(G)$ of $ \mathcal{E}(G)$ spanned by copies of the graph $H$. We are interested in when the random graph $G = G_{n,p}$ is likely to satisfy \[\mathcal{C}_H(G) = \mathcal{W}_H(G),\] where $\mathcal{W}_H(G)$ takes one of four natural values, depending on the value of $\mathcal{C}_H(K_n)$. We show that for strictly $2$-balanced $H$, w.h.p. the above equality holds whenever every edge of $G$ is in a copy of $H$.
On the local convergence of integer-valued Lipschitz functions on regular trees
We study random integer-valued Lipschitz functions on regular trees. It was shown by Peled, Samotij and Yehudayoff that such functions are localized, however, finer questions about the structure of Gibbs measures remain unanswered. Our main result is that the weak limit of a uniformly chosen 1-Lipschitz function with 0 boundary condition on a $d$-ary tree of height $n$ exists as $n \to \infty$ if $2 \le d \le 7$, but not if $d \ge 8$, thereby partially answering a question posed by Peled, Samotij and Yehudayoff. For large $d$, the value at the root alternates between being almost entirely concentrated on 0 for even $n$ and being roughly uniform on $\{-1,0,1\}$ for odd $n$, leading to different limits as $n$ approaches infinity along evens or odds. For $d \ge 8$, the essence of this phenomenon is preserved, which obstructs the convergence. For $d \le 7$, this phenomenon ceases to exist, and the law of the value at the root loses its connection with the parity of $n$. Along the way, we also obtain an alternative proof of localization. The key idea is a fixed point convergence result for a related operator on $\ell^\infty$, and a procedure to show that the iterations get into a `basin of attraction' of the fixed point. We also prove some accompanying analogous `even-odd phenomenon' type results about $M$-lipschitz functions on general non-amenable graphs with high enough expansion (this includes for example the large $d$ case for regular trees). We also prove a convergence result for 1-Lipschitz functions with $\{0,1\}$ boundary condition. This last result relies on an absolute value FKG for uniform 1-Lipschitz functions when shifted by $1/2$.
A combinatorial approach to nonlinear spectral gaps
A seminal open question of Pisier and Mendel--Naor asks whether every degree-regular graph which satisfies the classical discrete Poincaré inequality for scalar functions, also satisfies an analogous inequality for functions taking values in \textit{any} normed space with non-trivial cotype. Motivated by applications, it is also greatly important to quantify the dependence of the corresponding optimal Poincaré constant on the cotype $q$. Works of Odell--Schlumprecht (1994), Ozawa (2004), and Naor (2014) make substantial progress on the former question by providing a positive answer for normed spaces which also have an unconditional basis, in addition to finite cotype. However, little is known in the way of quantitative estimates: the mentioned results imply a bound on the Poincaré constant depending super-exponentially on $q$.
We introduce a novel combinatorial framework for proving quantitative nonlinear spectral gap estimates. The centerpiece is a property of regular graphs that we call \emph{long range expansion}, which holds with high probability for random regular graphs. Our main result is that any regular graph with the long-range expansion property satisfies a discrete Poincaré inequality for any normed space with an unconditional basis and cotype $q$, with a Poincaré constant that depends \emph{polynomially} on $q$, which is optimal. As an application, any normed space with an unconditional basis which admits a low distortion embedding of an $n$-vertex random regular graph, must have cotype at least polylogarithmic in $n$. This extends a celebrated lower-bound of Matoušek for low distortion embeddings of random graphs into $\ell_q$ spaces.
Hypergeometric Functions of Random Matrices and Quasimodular Forms
Hypergeometric functions of complex matrices were introduced by James in multivariate statistics. These special functions play many roles in random matrix theory. The main goal of this paper is to suggest a new use for them as holomorphic observables of the Circular Unitary Ensemble. We analyze the high-dimensional behavior of the expected derivatives of these random analytic functions, and show that they admit asymptotic expansions which can be described in terms of quasimodular forms, giving an apparently new connection between the CUE and number theory.
Polyhedral volume ratios, Izmestiev's Colin de Verdiere matrices and Spectral Gaps
We present a relation between volumes of certain lower dimensional simplices associated to a full-dimensional primal and polar dual polytope in R^k. We then discuss an application of this relation to a geometric construction of a Colin de Verdiere matrix by Ivan Izmestiev. In the second part of the paper, we introduce a variation of vertex transitive polytopes, translate their associated Colin de Verdiere matrices into random walk matrices, and investigate extremality properties of the spectral gaps of these random walk matrices in two concrete examples - permutahedra of Coxeter groups and polytopes associated to the pure rotational tetrahedral group - where maximal spectral gaps correspond to equilateral polytopes.
Simulating Simple Random Walks With a Deck of Cards
When we want to simulate the realization of a symmetric simple random walk on $\mathbb Z^d$, we use $(2d)$-side fair dice to decide to which neighbor it jumps at each step if $d\geq 2$ or we simply use a fair coin when $d=1$. Assume that instead of using a dice or a coin we want to do a simulation using a well shuffled deck with $K$ cards of each of the $2d$ suits. In the first step the probability of jumping to each neighbor is $(2d)^{-1}$, but from the second step it becomes biased. Of course if we continue performing this simulation, the total variation distance between its law and the law of the random walk will increase until all cards are used. In this paper we investigate the minimum number of cards $N=2d K$ that a deck must contain so that the total variation distance between the law of a $n$-step simulation and the law of a $n$-step realization of the random walk is smaller than a chosen threshold $\varepsilon \in (0,1)$. More generally, we prove that when $N=cn$ this distance converges, as $n \to \infty$, to a Gaussian profile which depends on $c\geq 2d$. Furthermore, our analysis shows that this Gaussian profile vanishes as $c \to \infty$, proving the convergence of a multivariate hypergeometric distribution to a multinomial distribution in total variation.
Free cumulants and freeness for unitarily invariant random tensors
We address the question of the asymptotic description of random tensors that are local-unitary invariant, that is, invariant by conjugation by tensor products of independent unitary matrices. We consider both the mixed case of a tensor with $D$ inputs and $D$ outputs, and the case where there is a factorization between the inputs and outputs, called pure, which includes the random tensor models extensively studied in the physics literature.
The finite size and asymptotic moments are defined using correlations of certain invariant polynomials encoded by $D$-tuples of permutations, up to relabeling equivalence. Finite size free cumulants associated to the expectations of these invariants are defined through invertible finite size moment-cumulants formulas.
Two important cases are considered asymptotically: pure random tensors that scale like a complex Gaussian, and mixed random tensors that scale like a Wishart tensor. In both cases, we derive a notion of tensorial free cumulants associated to first order invariants, through moment-cumulant formulas involving summations over non-crossing permutations. The pure and mixed cases involve the same combinatorics, but differ by the invariants that define the distribution at first order. In both cases, the tensorial free-cumulants of a sum of two independent tensors are shown to be additive. A preliminary discussion of higher orders is provided.
Tensor freeness is then defined as the vanishing of mixed first order tensorial free cumulants. The equivalent formulation at the level of asymptotic moments is derived in the pure and mixed cases, and we provide an algebraic construction of tensorial probability spaces, which generalize non-commutative probability spaces: random tensors converge in distribution to elements of these spaces, and tensor freeness of random variables corresponds to tensor freeness of the subspaces they generate.
New matrix perturbation bounds with relative norm: Perturbation of eigenspaces
Matrix perturbation bounds (such as Weyl and Davis-Kahan) are used abundantly in many areas of mathematics and data science. Many bounds (such as the above two) involve the spectral norm of the noise matrix and are sharp in worst case analysis.
In order to refine these classical bounds, we introduce a new parameter, which we refer to as the relative norm. This parameter measures the strength of the action of the noise matrix on the relevant eigenvectors of the ground matrix. It has turned out that in a number of situations, we can use the relative norm as a replacement for the spectral norm. This has led to a number of notable improvements under certain sets of assumptions, which are frequently met in practice. For instance, our new results apply very well in the case when the noise matrix is random.
For the purpose of our study, we introduce a new method of analysis, which combines the classical contour integral argument with new (combinatorial) ideas. This method is robust and of independent interest.
In the current paper, we focus on the perturbation of eigenspaces (Davis-Kahan type results). Perturbation bounds for eigenspaces are essential in statistics and theoretical computer science, and thus deserve a special treatment. Furthermore, this will lay the ground for the more technical treatment of general matrix functionals, which appears in a future paper.
A combinatorial approach to phase transitions in random graph isomorphism problems
We consider two independent Erdős-Rényi random graphs, with possibly different parameters, and study two isomorphism problems, a graph embedding problem and a common subgraph problem. Under certain conditions on the graph parameters we show a sharp asymptotic phase transition as the graph sizes tend to infinity. This extends known results for the case of uniform Erdős-Rényi random graphs. Our approach is primarily combinatorial, naturally leading to several related problems for further exploration.
Combinatorics of a dissimilarity measure for pairs of draws from discrete probability vectors on finite sets of objects
Motivated by a problem in population genetics, we examine the combinatorics of dissimilarity for pairs of random unordered draws of multiple objects, with replacement, from a collection of distinct objects. Consider two draws of size $K$ taken with replacement from a set of $I$ objects, where the two draws represent samples from potentially distinct probability distributions over the set of $I$ objects. We define the set of \emph{identity states} for pairs of draws via a series of actions by permutation groups, describing the enumeration of all such states for a given $K \geq 2$ and $I \geq 2$. Given two probability vectors for the $I$ objects, we compute the probability of each identity state. From the set of all such probabilities, we obtain the expectation for a dissimilarity measure, finding that it has a simple form that generalizes a result previously obtained for the case of $K=2$. We determine when the expected dissimilarity between two draws from the same probability distribution exceeds that of two draws taken from different probability distributions. We interpret the results in the setting of the genetics of polyploid organisms, those whose genetic material contains many copies of the genome ($K > 2$).
Creating Subgraphs in Semi-Random Hypergraph Games
The semi-random hypergraph process is a natural generalisation of the semi-random graph process, which can be thought of as a one player game. For fixed $r < s$, starting with an empty hypergraph on $n$ vertices, in each round a set of $r$ vertices $U$ is presented to the player independently and uniformly at random. The player then selects a set of $s-r$ vertices $V$ and adds the hyperedge $U \cup V$ to the $s$-uniform hypergraph. For a fixed (monotone) increasing graph property, the player's objective is to force the graph to satisfy this property with high probability in as few rounds as possible.
We focus on the case where the player's objective is to construct a subgraph isomorphic to an arbitrary, fixed hypergraph $H$. In the case $r=1$ the threshold for the number of rounds required was already known in terms of the degeneracy of $H$. In the case $2 \le r < s$, we give upper and lower bounds on this threshold for general $H$, and find further improved upper bounds for cliques in particular. We identify cases where the upper and lower bounds match. We also demonstrate that the lower bounds are not always tight by finding exact thresholds for various paths and cycles.
Bivariate exponential integrals and edge-bicolored graphs
Published in Le Matematiche, 80 (1), 167-187 (2025)
• View Publication
• BIB
We show that specific exponential bivariate integrals serve as generating functions of labeled edge-bicolored graphs. Based on this, we prove an asymptotic formula for the number of regular edge-bicolored graphs with arbitrary weights assigned to different vertex structures. The asymptotic behavior is governed by the critical points of a polynomial. As an application, we discuss the Ising model on a random 4-regular graph and show how its phase transitions arise from our formula.
Tree height and the asymptotic mean of the Colijn-Plazzotta rank of unlabeled binary rooted trees
The Colijn--Plazzotta ranking is a bijective encoding of the unlabeled binary rooted trees with positive integers. We show that the rank $f(t)$ of a tree $t$ is closely related to its height $h$, the length of the longest path from a leaf to the root. We consider the rank $f(τ_n)$ of a random $n$-leaf tree $τ_n$ under each of three models: (i) uniformly random unlabeled unordered binary rooted trees, or unlabeled topologies; (ii) uniformly random leaf-labeled binary trees, or labeled topologies under the uniform model; and (iii) random binary search trees, or labeled topologies under the Yule--Harding model. Relying on the close relationship between tree rank and tree height, we obtain results concerning the asymptotic properties of $\log \log f(τ_n)$. In particular, we find $\mathbb{E} \{\log_2 \log f(τ_n)\} \sim 2 \sqrt{πn}$ for uniformly random unlabeled ordered binary rooted trees and uniformly random leaf-labeled binary trees, and for a constant $α\approx 4.31107$, $\mathbb{E}\{\log_2 \log f(τ_n)\} \sim α\log n $ for leaf-labeled binary trees under the Yule--Harding model. We show that the mean of $f(τ_n)$ itself under the three models is largely determined by the rank $c_{n-1}$ of the highest-ranked tree -- the caterpillar -- obtaining an asymptotic relationship with $π_n c_{n-1}$, where $π_n$ is a model-specific function of $n$. The results resolve open problems, providing a new class of results on an encoding useful in mathematical phylogenetics.
The difference between the chromatic and the cochromatic number of a random graph
The cochromatic number $ζ(G)$ of a graph $G$ is the minimum number of colours needed for a vertex colouring where every colour class is either an independent set or a clique. Let $χ(G)$ denote the usual chromatic number. Around 1991 Erdős and Gimbel asked: For the random graph $G \sim G_{n, 1/2}$, does $χ(G)-ζ(G) \rightarrow \infty$ whp? Erdős offered \$100 for a positive and \$1,000 for a negative answer.
We give a positive answer to this question for roughly 95% of all values $n$.
The hitting time of nice factors
Consider the random $u$-uniform hypergraph (or $u$-graph) process on $n$ vertices, where $n$ is divisible by $r>u\ge 2$. It was recently shown that with high probability, as soon as every vertex is covered by a copy of the complete $u$-graph $K_r$, it also contains a $K_r$-factor (RSA, Vol. 65 II, Sept. 2024). The hitting time result is obtained using a process coupling, which is based on the proof of the corresponding sharp threshold result (RSA, Vol. 61 IV, Dec. 2022). The latter, however, was not only derived for complete $u$-graphs, but for a broader class of so-called nice $u$-graphs.
The purpose of this article is to extend the process coupling for complete $u$-graphs to the full scope of the sharp threshold result: nice $u$-graphs. As a byproduct, we obtain the extension of the hitting time result to nice $u$-graphs. Since the relevant combinatorial bounds in the proof for the $K_r$-case cannot be generalized, we introduce new arguments that do not only apply to nice u-graphs, but will be relevant for the broader class of strictly 1-balanced u-graphs. Further, we show how the remainder of the process coupling for the $K_r$-case can be utilized in a black-box manner for any u-graph. These advances pave the way for future generalizations.
Canonical labelling of sparse random graphs
We show that if $p=O(1/n)$, then the Erdős-Rényi random graph $G(n,p)$ with high probability admits a canonical labeling computable in time $O(n\log n)$. Combined with the previous results on the canonization of random graphs, this implies that $G(n,p)$ with high probability admits a polynomial-time canonical labeling whatever the edge probability function $p$. Our algorithm combines the standard color refinement routine with simple post-processing based on the classical linear-time tree canonization. Noteworthy, our analysis of how well color refinement performs in this setting allows us to complete the description of the automorphism group of the 2-core of $G(n,p)$.
Concentration of information on discrete groups
Motivated by the Asymptotic Equipartition Property and its recently discovered role in the cutoff phenomenon, we initiate the systematic study of varentropy on discrete groups. Our main result is an approximate tensorization inequality which asserts that the varentropy of any conjugacy-invariant random walk is, up to a universal multiplicative constant, at most that of the free Abelian random walk with the same jump rates. In particular, it is always bounded by the number d of generators, uniformly in time and in the size of the group. This universal estimate is sharp and can be seen as a discrete analogue of a celebrated result of Bobkov and Madiman concerning random d-dimensional vectors with a log-concave density (AOP 2011). A key ingredient in our proof is the fact that conjugacy-invariant random walks have non-negative Bakry-Émery curvature, a result which seems new and of independent interest.
Cutoff for the Biased Random Transposition Shuffle
In this paper, we study the biased random transposition shuffle, a natural generalization of the classical random transposition shuffle studied by Diaconis and Shahshahani. We diagonalize the transition matrix of the shuffle and use these eigenvalues to prove that the shuffle exhibits total variation cutoff at time $t_N = \frac{1}{2b} N \log N$ with window $N$. We also prove that the limiting distribution of the number of fixed cards near the cutoff time is Poisson.