arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

6952 papers tagged with this keyword
2021-04-16 v2
On the power of random greedy algorithms
Published in European Journal of Combinatorics 105 (2022), 103551 • View PublicationBIB
In this paper we solve two problems of Esperet, Kang and Thomasse as well as Li concerning (i) induced bipartite subgraphs in triangle-free graphs and (ii) van der Waerden numbers. Each time random greedy algorithms allow us to go beyond the Lovasz Local Lemma or alteration method used in previous work, illustrating the power of the algorithmic approach to the probabilistic method.
Rankings in directed configuration models with heavy tailed in-degrees
Published • View PublicationBIB
We consider the extremal values of the stationary distribution of sparse directed random graphs with given degree sequences and their relation to the extremal values of the in-degree sequence. The graphs are generated by the directed configuration model. Under the assumption of bounded $(2+η)$-moments on the in-degrees and of bounded out-degrees, we obtain tight comparisons between the maximum value of the stationary distribution and the maximum in-degree. Under the further assumption that the order statistics of the in-degrees have a power-law behavior, we show that the extremal values of the stationary distribution also have a power-law behavior with the same index. In the same setting, we prove that these results extend to the PageRank scores of the random digraph, thus confirming a version of the so-called power-law hypothesis. Along the way, we establish several facts about the model, including the mixing time cutoff and the characterization of the typical values of the stationary distribution, which were previously obtained under the assumption of bounded in-degrees.
Getting the Lay of the Land in Discrete Space: A Survey of Metric Dimension and its Applications
Published • View PublicationBIB
The metric dimension of a graph is the smallest number of nodes required to identify all other nodes based on shortest path distances uniquely. Applications of metric dimension include discovering the source of a spread in a network, canonically labeling graphs, and embedding symbolic data in low-dimensional Euclidean spaces. This survey gives a self-contained introduction to metric dimension and an overview of the quintessential results and applications. We discuss methods for approximating the metric dimension of general graphs, and specific bounds and asymptotic behavior for deterministic and random families of graphs. We conclude with related concepts and directions for future work.
Fluctuations of Subgraph Counts in Graphon Based Random Graphs
Published • View PublicationBIB
Given a graphon $W$ and a finite simple graph $H$, with vertex set $V(H)$, denote by $X_n(H, W)$ the number of copies of $H$ in a $W$-random graph on $n$ vertices. The asymptotic distribution of $X_n(H, W)$ was recently obtained by Hladký, Pelekis, and Šileikis (2021) in the case where $H$ is a clique. In this paper, we extend this result to any fixed graph $H$. Towards this we introduce a notion of $H$-regularity of graphons and show that if the graphon $W$ is not $H$-regular, then $X_n(H, W)$ has Gaussian fluctuations with scaling $n^{|V(H)|-\frac{1}{2}}$. On the other hand, if $W$ is $H$-regular, then the fluctuations are of order $n^{|V(H)|-1}$ and the limiting distribution of $X_n(H, W)$ can have both Gaussian and non-Gaussian components, where the non-Gaussian component is a (possibly) infinite weighted sum of centered chi-squared random variables with the weights determined by the spectral properties of a graphon derived from $W$. Our proofs use the asymptotic theory of generalized $U$-statistics developed by Janson and Nowicki (1991). We also investigate the structure of $H$-regular graphons for which either the Gaussian or the non-Gaussian component of the limiting distribution (but not both) is degenerate. Interestingly, there are also $H$-regular graphons $W$ for which both the Gaussian or the non-Gaussian components are degenerate, that is, $X_n(H, W)$ has a degenerate limit even under the scaling $n^{|V(H)|-1}$. We give an example of this degeneracy with $H=K_{1, 3}$ (the 3-star) and also establish non-degeneracy in a few examples. This naturally leads to interesting open questions on higher-order degeneracies.
Linear-sized independent sets in random cographs and increasing subsequences in separable permutations
Published • View PublicationBIB
This paper is interested in independent sets (or equivalently, cliques) in uniform random cographs. We also study their permutation analogs, namely, increasing subsequences in uniform random separable permutations. First, we prove that, with high probability as $n$ gets large, the largest independent set in a uniform random cograph with $n$ vertices has size $o(n)$. This answers a question of Kang, McDiarmid, Reed and Scott. Using the connection between graphs and permutations via inversion graphs, we also give a similar result for the longest increasing subsequence in separable permutations. These results are proved using the self-similarity of the Brownian limits of random cographs and random separable permutations, and actually apply more generally to all families of graphs and permutations with the same limit. Second, and unexpectedly given the above results, we show that for $β>0$ sufficiently small, the expected number of independent sets of size $βn$ in a uniform random cograph with $n$ vertices grows exponentially fast with $n$. We also prove a permutation analog of this result. This time the proofs rely on singularity analysis of the associated bivariate generating functions.
2021-04-15
A Tale of Two Limits: An Extremal Pagerank Problem
For a directed graph, the Pagerank algorithm emulates a random walker on the graph that occasionally "jumps" to a random vertex based on a jumping parameter $α$. Upon completion, the algorithm generates a stochastic vector whose entries correspond to the limiting probability that the walker will be at that vertex. This vector is a right eigenvector of a corresponding Markov trasition matrix. Undoubtedly, this vector can drastically change based upon the jumping parameter $α$. In this article, we investigate the maximum possible discrepancy for different Pagerank vectors on the same unweighted directed (perhaps with loops) graph as measured by the 2-norm. We show that the limsup of this discrepancy can be as large as $\sqrt{\frac{67}{50}}$ using a very specific construction. (For contrast, the norm of the difference for any two stochastic vectors is at most $\sqrt{2}$.) Interestingly, on this construction this discrepancy occurs when $α= 1$ and when $α$ is very close to 1.
2021-04-14
New results for MaxCut in $H$-free graphs
Published • View PublicationBIB
The MaxCut problem asks for the size ${\rm mc}(G)$ of a largest cut in a graph $G$. It is well known that ${\rm mc}(G)\ge m/2$ for any $m$-edge graph $G$, and the difference ${\rm mc}(G)-m/2$ is called the surplus of $G$. The study of the surplus of $H$-free graphs was initiated by Erdős and Lovász in the 70s, who in particular asked what happens for triangle-free graphs. This was famously resolved by Alon, who showed that in the triangle-free case the surplus is $Ω(m^{4/5})$, and found constructions matching this bound. We prove several new results in this area. Firstly, we show that for every fixed odd $r\ge 3$, any $C_r$-free graph with $m$ edges has surplus $Ω_r\big(m^{\frac{r+1}{r+2}}\big)$. This is tight, as is shown by a construction of pseudorandom $C_r$-free graphs due to Alon and Kahale. It improves previous results of several researchers, and complements a result of Alon, Krivelevich and Sudakov which is the same bound when $r$ is even. Secondly, generalizing the result of Alon, we allow the graph to have triangles, and show that if the number of triangles is a bit less than in a random graph with the same density, then the graph has large surplus. For regular graphs our bounds on the surplus are sharp. Thirdly, we prove that an $n$-vertex graph with few copies of $K_r$ and average degree $d$ has surplus $Ω_r(d^{r-1}/n^{r-3})$, which is tight when $d$ is close to $n$ provided that a conjectured dense pseudorandom $K_r$-free graph exists. This result is used to improve the best known lower bound (as a function of $m$) on the surplus of $K_r$-free graphs. Our proofs combine techniques from semidefinite programming, probabilistic reasoning, as well as combinatorial and spectral arguments.
2021-04-13 v3
Measured expanders
Published • View PublicationBIB
By measured graphs we mean graphs endowed with a measure on the set of vertices. In this context, we explore the relations between the appropriate Cheeger constant and Poincaré inequalities. We prove that the so-called Cheeger inequality holds in two cases: when the measure comes from a random walk, or when the measure has a bounded measure ratio. Moreover, we also prove that our measured (asymptotic) expanders are generalised expanders introduced by Tessera. Finally, we present some examples to demonstrate relations and differences between classical expander graphs and the measured ones. The current paper is motivated primarily by our previous work on the rigidity problem for Roe algebras.
2021-04-13
Permanent of random matrices from representation theory: moments, numerics, concentration, and comments on hardness of boson-sampling
Computing the distribution of permanents of random matrices has been an outstanding open problem for several decades. In quantum computing, "anti-concentration" of this distribution is an unproven input for the proof of hardness of the task of boson-sampling. We study permanents of random i.i.d. complex Gaussian matrices, and more broadly, submatrices of random unitary matrices. Using a hybrid representation-theoretic and combinatorial approach, we prove strong lower bounds for all moments of the permanent distribution. We provide substantial evidence that our bounds are close to being tight and constitute accurate estimates for the moments. Let $U(d)^{k\times k}$ be the distribution of $k\times k$ submatrices of $d\times d$ random unitary matrices, and $G^{k\times k}$ be the distribution of $k\times k$ complex Gaussian matrices. (1) Using the Schur-Weyl duality (or the Howe duality), we prove an expansion formula for the $2t$-th moment of $|Perm(M)|$ when $M$ is drawn from $U(d)^{k\times k}$ or $G^{k\times k}$. (2) We prove a surprising size-moment duality: the $2t$-th moment of the permanent of random $k\times k$ matrices is equal to the $2k$-th moment of the permanent of $t\times t$ matrices. (3) We design an algorithm to exactly compute high moments of the permanent of small matrices. (4) We prove lower bounds for arbitrary moments of permanents of matrices drawn from $G^{ k\times k}$ or $U(k)$, and conjecture that our lower bounds are close to saturation up to a small multiplicative error. (5) Assuming our conjectures, we use the large deviation theory to compute the tail of the distribution of log-permanent of Gaussian matrices for the first time. (6) We argue that it is unlikely that the permanent distribution can be uniquely determined from the integer moments and one may need to supplement the moment calculations with extra assumptions to prove the anti-concentration conjecture.
2021-04-10
Lower tails via relative entropy
Published • View PublicationBIB
We show that the naive mean-field approximation correctly predicts the leading term of the logarithmic lower tail probabilities for the number of copies of a given subgraph in $G(n,p)$ and of arithmetic progressions of a given length in random subsets of the integers in the entire range of densities where the mean-field approximation is viable. Our main technical result provides sufficient conditions on the maximum degrees of a uniform hypergraph $\mathcal{H}$ that guarantee that the logarithmic lower tail probabilities for the number of edges induced by a binomial random subset of the vertices of $\mathcal{H}$ can be well-approximated by considering only product distributions. This may be interpreted as a weak, probabilistic version of the hypergraph container lemma that is applicable to all sparser-than-average (and not only independent) sets.
2021-04-07 v6
Goodness of fit for log-linear ERGMs
Many popular models from the networks literature can be viewed through a common lens of contingency tables on network dyads, resulting in \emph{log-linear ERGMs}: exponential family models for random graphs whose sufficient statistics are linear on the dyads. We propose a new model in this family, the \emph{$p_1$-SBM}, which combines node and group effects common in network formation mechanisms. In particular, it is a generalization of several well-known ERGMs including the stochastic blockmodel for undirected graphs with known block assignment, the degree-corrected version of it, and the directed $p_1$ model without group structure. We frame the problem of testing model fit for the log-linear ERGM class through an exact conditional test whose $p$-value can be approximated efficiently in networks of both small and moderately large sizes. The sampling methods we build rely on a dynamic adaptation of Markov bases. We use quick estimation algorithms adapted from the contingency table literature and effective sampling methods rooted in graph theory and algebraic statistics. The performance and scalability of the method is demonstrated on two data sets from biology: the connectome of \emph{C. elegans} and the interactome of \emph{Arabidopsis thaliana}. These two networks -- a network and a protein-protein interaction network -- have been popular examples in the network science literature. Our work provides a model-based approach to studying them.
2021-04-06
The sum of powers of subtree sizes for conditioned Galton-Watson trees
Published • View PublicationBIB
We study the additive functional $X_n(α)$ on conditioned Galton-Watson trees given, for arbitrary complex $α$, by summing the $α$th power of all subtree sizes. Allowing complex $α$ is advantageous, even for the study of real $α$, since it allows us to use powerful results from the theory of analytic functions in the proofs. For $\Reα< 0$, we prove that $X_n(α)$, suitably normalized, has a complex normal limiting distribution; moreover, as processes in $α$, the weak convergence holds in the space of analytic functions in the left half-plane. We establish, and prove similar process-convergence extensions of, limiting distribution results for $α$ in various regions of the complex plane. We focus mainly on the case where $\Reα> 0$, for which $X_n(α)$, suitably normalized, has a limiting distribution that is not normal but does not depend on the offspring distribution $ξ$ of the conditioned Galton-Watson tree, assuming only that $E[ξ] = 1$ and $0 < \mathrm{Var} [ξ] < \infty$. Under a weak extra moment assumption on $ξ$, we prove that the convergence extends to moments, ordinary and absolute and mixed, of all orders. At least when $\Reα> \frac12$, the limit random variable $Y(α)$ can be expressed as a function of a normalized Brownian excursion.
2021-04-05
Which Sampling Densities are Suitable for Spectral Clustering on Unbounded Domains?
We consider a random geometric graph with vertices sampled from a probability measure supported on $\mathbb R^d$, and study its connectivity. We show the graph is typically disconnected, unless the sampling density has superexponential decay. In the later setting, we identify an asymptotic threshold value for the radius parameter of the graph such that, for radius values beyond the threshold, some concentration properties hold for the sampled points of the graph, while the graph is disconnected for radius values below the same threshold. Properties of point processes are well-known to be closely related to the analysis of geometric learning problems, such as spectral clustering. This work can be seen as a first step towards understanding the consistency of spectral clustering when the probability measure has unbounded support. In particular, we narrow down the setting under which spectral clustering algorithms on $\mathbb R^d$ may be expected to achieve consistency, to a sufficiently fast decay of the sampling density (superexponential) and a sufficiently slowly decaying radius parameter value as a function of $n$, the number of sampled points.
What does a typical metric space look like?
Published • View PublicationBIB
The collection $\mathcal{M}_n$ of all metric spaces on $n$ points whose diameter is at most $2$ can naturally be viewed as a compact convex subset of $\mathbb{R}^{\binom{n}{2}}$, known as the metric polytope. In this paper, we study the metric polytope for large $n$ and show that it is close to the cube $[1,2]^{\binom{n}{2}} \subseteq \mathcal{M}_n$ in the following two senses. First, the volume of the polytope is not much larger than that of the cube, with the following quantitative estimates: \[ \left(\tfrac{1}{6}+o(1)\right)n^{3/2} \le \log \mathrm{Vol}(\mathcal{M}_n)\le O(n^{3/2}). \] Second, when sampling a metric space from $\mathcal{M}_n$ uniformly at random, the minimum distance is at least $1 - n^{-c}$ with high probability, for some $c > 0$. Our proof is based on entropy techniques. We discuss alternative approaches to estimating the volume of $\mathcal{M}_n$ using exchangeability, Szemerédi's regularity lemma, the hypergraph container method, and the Kővári--Sós--Turán theorem.
The modularity of random graphs on the hyperbolic plane
Published • View PublicationBIB
Modularity is a quantity which has been introduced in the context of complex networks in order to quantify how close a network is to an ideal modular network in which the nodes form small interconnected communities that are joined together with relatively few edges. In this paper, we consider this quantity on a recent probabilistic model of complex networks introduced by Krioukov et al. (Phys. Rev. E 2010). This model views a complex network as an expression of hidden hierarchies, encapsulated by an underlying hyperbolic space. For certain parameters, this model was proved to have typical features that are observed in complex networks such as power law degree distribution, bounded average degree, clustering coefficient that is asymptotically bounded away from zero, and ultra-small typical distances. In the present work, we investigate its modularity and we show that, in this regime, it converges to 1 in probability.
2021-04-02 v2
Convergence rates of limit theorems in random chord diagrams
We study the asymptotic distributions of the number of crossings and the number of simple chords in a random chord diagram. Using size-bias coupling and Stein's method, we obtain bounds on the Kolmogorov distance between the distribution of the number of crossings and a standard normal random variable, and on the total variation distance between the distribution of the number of simple chords and a Poisson random variable. As an application, we provide explicit error bounds on the number of chord diagrams containing no simple chords.
2021-04-02
On random compact sets, equidecomposition, and domains of expansion in R^3
We study random compact subsets of R^3 which can be described as "random Menger sponges". We use those random sets to construct a pair of compact sets A and B in R^3 which are of the same positive measure, such that A can be covered by finitely many translates of B, B can be covered by finitely many translates of A, and yet A and B are not equidecomposable. Furthermore, we construct the first example of a compact subset of R^3 of positive measure which is not a domain of expansion. This answers a question of Adrian Ioana.
2021-03-31
Parking functions: From combinatorics to probability
Published • View PublicationBIB
Suppose that $m$ drivers each choose a preferred parking space in a linear car park with $n$ spots. In order, each driver goes to their chosen spot and parks there if possible, and otherwise takes the next available spot if it exists. If all drivers park successfully, the sequence of choices is called a parking function. Classical parking functions correspond to the case $m=n$; we study here combinatorial and probabilistic aspects of this generalized case. We construct a family of bijections between parking functions $\text{PF}(m, n)$ with $m$ cars and $n$ spots and spanning forests $\mathscr{F}(n+1, n+1-m)$ with $n+1$ vertices and $n+1-m$ distinct trees having specified roots. This leads to a bijective correspondence between $\text{PF}(m, n)$ and monomial terms in the associated Tutte polynomial of a disjoint union of $n-m+1$ complete graphs. We present an identity between the "inversion enumerator" of spanning forests with fixed roots and the "displacement enumerator" of parking functions. The displacement is then related to the number of graphs on $n+1$ labeled vertices with a fixed number of edges, where the graph has $n+1-m$ disjoint rooted components with specified roots. We investigate various probabilistic properties of a uniform parking function, giving a formula for the law of a single coordinate. As a side result we obtain a recurrence relation for the displacement enumerator. Adapting known results on random linear probes, we further deduce the covariance between two coordinates when $m=n$.
2021-03-30
Paths, cycles and sprinkling in random hypergraphs
Published • View PublicationBIB
We prove a lower bound on the length of the longest $j$-tight cycle in a $k$-uniform binomial random hypergraph for any $2 \le j \le k-1$. We first prove the existence of a $j$-tight path of the required length. The standard "sprinkling" argument is not enough to show that this path can be closed to a $j$-tight cycle -- we therefore show that the path has many extensions, which is sufficient to allow the sprinkling to close the cycle.
2021-03-30
Boundary of the boundary for random walks on groups
We study fine structure related to finitely supported random walks on infinite finitely generated discrete groups, largely motivated by dimension group techniques. The unfaithful extreme harmonic functions (defined only on proper space-time cones), aka unfaithful pure traces, can be represented on systems of finite support, avoiding dead ends. This motivates properties of the random walk (WC) and of the group (SWC) which become of interest in their own right. While all abelian groups satisfy WC, the do not satisfy SWC; however some abelian by finite groups do satisfy the latter, and we characterize when this occurs. In general, we determine the maximal order ideals, aka, maximal proper space-time subcones of that generated by the group element $1$ at time zero), and show that the corresponding quotients are stationary simple dimension groups, and that all such can occur for the free group on two generators. We conclude with a case study of the discrete Heisenberg group, determining among other things, the pure traces (these are the unfaithful ones, not arising from characters).