random
6952 papers tagged with this keyword
Community Detection in Hypergraphs via Mutual Information Maximization
The hypergraph community detection problem seeks to identify groups of related nodes in hypergraph data. We propose an information-theoretic hypergraph community detection algorithm which compresses the observed data in terms of community labels and community-edge intersections. This algorithm can also be viewed as maximum-likelihood inference in a degree-corrected microcanonical stochastic blockmodel. We perform the inference/compression step via simulated annealing. Unlike several recent algorithms based on canonical models, our microcanonical algorithm does not require inference of statistical parameters such as node degrees or pairwise group connection rates. Through synthetic experiments, we find that our algorithm succeeds down to recently-conjectured thresholds for sparse random hypergraphs. We also find competitive performance in cluster recovery tasks on several hypergraph data sets.
Counting two-forests and random cut size via potential theory
We prove a lower bound on the number of spanning two-forests in a graph, in terms of the number of vertices, edges, and spanning trees. This implies an upper bound on the average cut size of a random two-forest. The main tool is an identity relating the number of spanning trees and two-forests to pairwise effective resistances in a graph. Along the way, we make connections to potential theoretic invariants on metric graphs.
Solving a Random Asymmetric TSP Exactly in Quasi-Polynomial Time w.h.p
Let the costs $C(i,j)$ for an instance of the Asymmetric Traveling Salesperson Problem (ATSP) be independent copies of a non-negative random variable $C$ from a class of distributions that include the uniform $[0,1]$ distribution and the exponential mean 1 distribution with mean 1. We describe an algorithm that solves ATSP exactly in time $e^{\log^{2+o(1)}n}$, w.h.p.
Sparse pancyclic subgraphs of random graphs
It is known that the complete graph $K_n$ contains a pancyclic subgraph with $n+(1+o(1))\cdot \log _2 n$ edges, and that there is no pancyclic graph on $n$ vertices with fewer than $n+\log _2 (n-1) -1$ edges. We show that, with high probability, $G(n,p)$ contains a pancyclic subgraph with $n+(1+o(1))\log_2 n$ edges for $p \ge p^*$, where $p^*=(1+o(1))\ln n/n$, right above the threshold for pancyclicity.
Reconstruction of graph colourings
A $k$-deck of a (coloured) graph is a multiset of its induced $k$-vertex subgraphs. Given a graph $G$, when is it possible to reconstruct with high probability a uniformly random colouring of its vertices in $r$ colours from its $k$-deck? In this paper, we study this question for grids and random graphs. Reconstruction of random colourings of $d$-dimensional $n$-grids from the deck of their $k$-subgrids is one of the most studied colour reconstruction questions. The 1-dimensional case is motivated by the problem of reconstructing DNA sequences from their `shotgunned' stretches. It was comprehensively studied and the above reconstruction question was completely answered in the '90s. In this paper, we get a very precise answer for higher $d$. For every $d\geq 2$ and every $r\geq 2$, we present an almost linear algorithm that reconstructs with high probability a random $r$-colouring of vertices of a $d$-dimensional $n$-grid from the deck of all its $k$-subgrids for every $k\geq(d\log_r n)^{1/d}+1/d+\varepsilon$ and prove that the random $r$-colouring is not reconstructible with high probability if $k\leq (d\log_r n)^{1/d}-\varepsilon$. This answers the question of Narayanan and Yap (that was asked for $d\geq 3$) on "two-point concentration" of the minimum $k$ so that $k$-subgrids determine the entire colouring. Next, we prove that with high probability a uniformly random $r$-colouring of vertices of a uniformly random graph $G(n,1/2)$ is reconstructible from its full $k$-deck if $k\geq 2\log_2 n+8$ and is not reconstructible with high probability if $k\leq\sqrt{2\log_2 n}$. We further show that the colour reconstruction algorithm for random graphs can be modified and used for graph reconstruction: we prove that with high probability $G(n,1/2)$ is reconstructible from its full $k$-deck if $k\geq 2\log_2 n+11$ while it is not reconstructible with high probability if $k\leq 2\sqrt{\log_2 n}$.
Meeting, coalescence and consensus time on random directed graphs
We consider Markovian dynamics on a typical realization of the so-called Directed Configuration Model (DCM), that is, a random directed graph with prescribed in- and out-degrees. In this random geometry, we study the meeting time of two random walks on a typical realization of the graph starting at stationarity, the coalescence time for a system of coalescent random walks, and the consensus time of the voter model. Indeed, it is known that the latter three quantities are related to each other when the underlying sequence of graphs satisfies certain mean field conditions. Such conditions can be summarized by requiring a fast mixing time of the random walk and some anti-concentration of its stationary distribution: properties that a typical random directed graph is known to have under natural assumptions on the degree sequence. In this paper we show that, for a typical large graph from the DCM ensemble, the distribution of the meeting time is well-approximated by an exponential random variable and we provide the first-order approximation of its expectation, showing that the latter is linear in the size of the graph, and the preconstant depends on some easy statistics of the degree sequence. As a byproduct, we are able to analyze the effect of the degree sequence in changing the meeting, coalescence and consensus time. Our approach follows the classical idea of converting meeting into hitting times of a proper collapsed chain, which we control by the so-called First Visit Time Lemma.
Upper bounds on the $2$-colorability threshold of random $d$-regular $k$-uniform hypergraphs for $k\geq 3$
For a large class of random constraint satisfaction problems (CSP), deep but non-rigorous theory from statistical physics predict the location of the sharp satisfiability transition. The works of Ding, Sly, Sun (2014, 2016) and Coja-Oghlan, Panagiotou (2014) established the satisfiability threshold for random regular $k$-NAE-SAT, random $k$-SAT, and random regular $k$-SAT for large enough $k\geq k_0$ where $k_0$ is a large non-explicit constant. Establishing the same for small values of $k\geq 3$ remains an important open problem in the study of random CSPs.
In this work, we study two closely related models of random CSPs, namely the $2$-coloring on random $d$-regular $k$-uniform hypergraphs and the random $d$-regular $k$-NAE-SAT model. For every $k\geq 3$, we prove that there is an explicit $d_{\ast}(k)$ which gives a satisfiability upper bound for both of the models. Our upper bound $d_{\ast}(k)$ for $k\geq 3$ matches the prediction from statistical physics for the hypergraph $2$-coloring by Dall'Asta, Ramezanpour, Zecchina (2008), thus conjectured to be sharp. Moreover, $d_{\ast}(k)$ coincides with the satisfiability threshold of random regular $k$-NAE-SAT for large enough $k\geq k_0$ by Ding, Sly, Sun (2014).
Random Walk Labelings of Perfect Trees and Other Graphs
A Random walk labeling of a graph $G$ is any labeling of $G$ that could have been obtained by performing a random walk on $G$. Continuing two recent works, we calculate the number of random walk labelings of perfect trees, combs, and double combs, the torus $C_2\times C_n$, and the graph obtained by connecting three path graphs to form two cycles.
Uniform attachment with freezing
In the classical model of random recursive trees, trees are recursively built by attaching new vertices to old ones. What happens if vertices are allowed to freeze, in the sense that new vertices cannot be attached to already frozen ones? We are interested in the impact of freezing on the height of such trees.
Structure and Noise in Dense and Sparse Random Graphs: Percolated Stochastic Block Model via the EM Algorithm and Belief Propagation with Non-Backtracking Spectra
In this survey paper it is illustrated how spectral clustering methods for unweighted graphs are adapted to the dense and sparse regimes. Whereas Laplacian and modularity based spectral clustering is apt to dense graphs, recent results show that for sparse ones, the non-backtracking spectrum is the best candidate to find assortative clusters of nodes. Here belief propagation in the sparse stochastic block model is derived with arbitrarily given model parameters that results in a non-linear system of equations; with linear approximation, the spectrum of the non-backtracking matrix is able to specify the number $k$ of clusters. Then the model parameters themselves can be estimated by the EM algorithm.
Bond percolation in the assortative model is considered in the following two senses: the within- and between-cluster edge probabilities decrease with the number of nodes and edges coming into existence in this way are retained with probability $β$. As a consequence, the optimal $k$ is the number of the structural real eigenvalues (greater than $\sqrt{c}$, where $c$ is the average degree) of the non-backtracking matrix of the graph. Assuming, these eigenvalues $μ_1 >\dots > μ_k$ are distinct, the multiple phase transitions obtained for $β$ are $β_i =\frac{c}{μ_i^2}$; further, at $β_i$ the number of detectable clusters is $i$, for $i=1,\dots ,k$. Inflation-deflation techniques are also discussed to classify the nodes themselves, which can be the base of the sparse spectral clustering. Simulation results, as well as real life examples are presented.
On the Kohayakawa-Kreuter conjecture
Let us say that a graph $G$ is Ramsey for a tuple $(H_1,\dots,H_r)$ of graphs if every $r$-coloring of the edges of $G$ contains a monochromatic copy of $H_i$ in color $i$, for some $i \in [r]$. A famous conjecture of Kohayakawa and Kreuter, extending seminal work of Rödl and Ruciński, predicts the threshold at which the binomial random graph $G_{n,p}$ becomes Ramsey for $(H_1,\dots,H_r)$ asymptotically almost surely. In this paper, we resolve the Kohayakawa-Kreuter conjecture for almost all tuples of graphs. Moreover, we reduce its validity to the truth of a certain deterministic statement, which is a clear necessary condition for the conjecture to hold. All of our results actually hold in greater generality, when one replaces the graphs $H_1,\dots,H_r$ by finite families $\mathcal{H}_1,\dots,\mathcal{H}_r$. Additionally, we pose a natural (deterministic) graph-partitioning conjecture, which we believe to be of independent interest, and whose resolution would imply the Kohayakawa-Kreuter conjecture.
On Approximability of Satisfiable k-CSPs: IV
We prove a stability result for general $3$-wise correlations over distributions satisfying mild connectivity properties. More concretely, we show that if $Σ,Γ$ and $Φ$ are alphabets of constant size, and $μ$ is a pairwise connected distribution over $Σ\timesΓ\timesΦ$ with no $(\mathbb{Z},+)$ embeddings in which the probability of each atom is $Ω(1)$, then the following holds. Any triplets of $1$-bounded functions $f\colon Σ^n\to\mathbb{C}$, $g\colon Γ^n\to\mathbb{C}$, $h\colon Φ^n\to\mathbb{C}$ satisfying
\[
\left|\mathbb{E}_{(x,y,z)\sim μ^{\otimes n}}\big[f(x)g(y)h(z)\big]\right|\geq \varepsilon
\]
must arise from an Abelian group associated with the distribution $μ$. More specifically, we show that there is an Abelian group $(H,+)$ of constant size such that for any such $f,g$ and $h$, the function $f$ (and similarly $g$ and $h$) is correlated with a function of the form $\tilde{f}(x) = χ(σ(x_1),\ldots,σ(x_n)) L (x)$, where $σ\colon Σ\to H$ is some map, $χ\in \hat{H}^{\otimes n}$ is a character, and $L\colon Σ^n\to\mathbb{C}$ is a low-degree function with bounded $2$-norm.
En route we prove a few additional results that may be of independent interest, such as an improved direct product theorem, as well as a result we refer to as a ``restriction inverse theorem'' about the structure of functions that, under random restrictions, with noticeable probability have significant correlation with a product function. In companion papers, we show applications of our results to the fields of Probabilistically Checkable Proofs, as well as various areas in discrete mathematics such as extremal combinatorics and additive combinatorics.
Locked Polyomino Tilings
A locked $t$-omino tiling is a grid tiling by $t$-ominoes such that, if you remove any pair of tiles, the only way to fill in the remaining $2t$ grid cells with $t$-ominoes is to use the same two tiles in the exact same configuration as before. We exclude degenerate cases where there is only one tiling overall due to small dimensions. It is a classic (and straightforward) result that finite grids do not admit locked 2-omino tilings. In this paper, we construct explicit locked $t$-omino tilings for $t \geq 3$ on grids of various dimensions. Most notably, we show that locked 3- and 4-omino tilings exist on finite square grids of arbitrarily large size, and locked $t$-omino tilings of the infinite grid exist for arbitrarily large $t$. The result for 4-omino tilings in particular is remarkable because they are so rare and difficult to construct: Only a single tiling is known to exist on any grid up to size $40 \times 40$. In a weighted version of the problem where vertices of the grid may have weights from the set $\{1, 2\}$ that count toward the total tile size, we demonstrate the existence of locked tilings on arbitrarily large square weighted grids with only 6 tiles.
Locked $t$-omino tilings arise as obstructions to widely used political redistricting algorithms in a model of redistricting where the underlying census geography is a grid graph. Most prominent is the ReCom Markov chain, which takes a random walk on the space of redistricting plans by iteratively merging and splitting pairs of districts (tiles) at a time. Locked $t$-omino tilings are isolated states in the state space of ReCom. The constructions in this paper are counterexamples to the meta-conjecture that ReCom is irreducible on graphs of practical interest.
Catching a robber on a random $k$-uniform hypergraph
Published
• View Publication
• BIB
The game of \emph{Cops and Robber} is usually played on a graph, where a group of cops attempt to catch a robber moving along the edges of the graph. The \emph{cop number} of a graph is the minimum number of cops required to win the game. An important conjecture in this area, due to Meyniel, states that the cop number of an $n$-vertex connected graph is $O(\sqrt{n})$. In 2016, Prałat and Wormald [Meyniel's conjecture holds for random graphs, Random Structures Algorithms. 48 (2016), no. 2, 396-421. MR3449604] showed that this conjecture holds with high probability for random graphs above the connectedness threshold. Moreoever, Łuczak and Prałat [Chasing robbers on random graphs: Zigzag theorem, Random Structures Algorithms. 37 (2010), no. 4, 516-524. MR2760362] showed that on a $\log$-scale the cop number demonstrates a surprising \emph{zigzag} behaviour in dense regimes of the binomial random graph $G(n,p)$. In this paper, we consider the game of Cops and Robber on a hypergraph, where the players move along hyperedges instead of edges. We show that with high probability the cop number of the $k$-uniform binomial random hypergraph $G^k(n,p)$ is $O\left(\sqrt{\frac{n}{k}}\, \log n \right)$ for a broad range of parameters $p$ and $k$ and that on a $\log$-scale our upper bound on the cop number arises as the minimum of \emph{two} complementary zigzag curves, as opposed to the case of $G(n,p)$. Furthermore, we conjecture that the cop number of a connected $k$-uniform hypergraph on $n$ vertices is $O\left(\sqrt{\frac{n}{k}}\,\right)$.
Minors of matroids represented by sparse random matrices over finite fields
Consider a random $n\times m$ matrix $A$ over the finite field of order $q$ where every column has precisely $k$ nonzero elements, and let $M[A]$ be the matroid represented by $A$. In the case that q=2, Cooper, Frieze and Pegden (RS\&A 2019) proved that given a fixed binary matroid $N$, if $k\ge k_N$ and $m/n\ge d_N$ where $k_N$ and $d_N$ are sufficiently large constants depending on N, then a.a.s. $M[A]$ contains $N$ as a minor. We improve their result by determining the sharp threshold (of $m/n$) for the appearance of a fixed matroid $N$ as a minor of $M[A]$, for every $k\ge 3$, and every finite field.
Correspondence coloring of random graphs
We show that Erdős-Rényi random graphs $G(n,p)$ with constant density $p<1$ have correspondence chromatic number $O(n/\sqrt{\log n})$; this matches a prediction from linear Hadwiger's conjecture for correspondence coloring. The proof follows from a simple sufficient condition for correspondence colorability in terms of the numbers of independent sets.
On the hardness of finding balanced independent sets in random bipartite graphs
We consider the algorithmic problem of finding large \textit{balanced} independent sets in sparse random bipartite graphs, and more generally the problem of finding independent sets with specified proportions of vertices on each side of the bipartition. In a bipartite graph it is trivial to find an independent set of density at least half (take one of the partition classes). In contrast, in a random bipartite graph of average degree $d$, the largest balanced independent sets (containing equal number of vertices from each class) are typically of density $(2+o_d(1)) \frac{\log d}{d}$. Can we find such large balanced independent sets in these graphs efficiently? By utilizing the overlap gap property and the low-degree algorithmic framework, we prove that local and low-degree algorithms (even those that know the bipartition) cannot find balanced independent sets of density greater than $(1+ε) \frac{\log d}{d}$ for any $ε>0$ fixed and $d$ large but constant. This factor $2$ statistical--computational gap between what exists and what local algorithms can achieve is analogous to the gap for finding large independent sets in (non-bipartite) random graphs. Our results therefor suggest that this gap is pervasive in many models, and that hard computational problems can lurk inside otherwise tractable ones. A particularly striking aspect of the gap in bipartite graphs is that the algorithm achieving the lower bound is extremely simple and can be implemented as a $1$-local algorithm and a degree-$1$ polynomial (a linear function).
On the distribution of the entries of a fixed-rank random matrix over a finite field
Let $r > 0$ be an integer, let $\mathbb{F}_q$ be a finite field of $q$ elements, and let $\mathcal{A}$ be a nonempty proper subset of $\mathbb{F}_q$. Moreover, let $\mathbf{M}$ be a random $m \times n$ rank-$r$ matrix over $\mathbb{F}_q$ taken with uniform distribution. We prove, in a precise sense, that, as $m, n \to +\infty$ and $r,q,\mathcal{A}$ are fixed, the number of entries of $\mathbf{M}$ that belong to $\mathcal{A}$ approaches a normal distribution.
All These Approximate Ramsey Properties
We consider finitary approximations of the (embedding) Ramsey property. Using a class of homogeneous reducts of random ordered hypergraphs, we prove that these properties form a strict hierarchy. We also show that every class of finite structures in which every structure of size at most 2 is a "Ramsey object" essentially consists of ordered structures, generalising a known result for countable Ramsey classes.
Tight Approximations for Graphical House Allocation
The Graphical House Allocation problem asks: how can $n$ houses (each with a fixed non-negative value) be assigned to the vertices of an undirected graph $G$, so as to minimize the "aggregate local envy", i.e., the sum of absolute differences along the edges of $G$? This problem generalizes the classical Minimum Linear Arrangement problem, as well as the well-known House Allocation Problem from Economics, the latter of which has notable practical applications in organ exchanges. Recent work has studied the computational aspects of Graphical House Allocation and observed that the problem is NP-hard and inapproximable even on particularly simple classes of graphs, such as vertex disjoint unions of paths. However, the dependence of any approximations on the structural properties of the underlying graph had not been studied.
In this work, we give a complete characterization of the approximability of the Graphical House Allocation problem. We present algorithms to approximate the optimal envy on general graphs, trees, planar graphs, bounded-degree graphs, bounded-degree planar graphs, and bounded-degree trees. For each of these graph classes, we then prove matching lower bounds, showing that in each case, no significant improvement can be attained unless P = NP. We also present general approximation ratios as a function of structural parameters of the underlying graph, such as treewidth; these match the aforementioned tight upper bounds in general, and are significantly better approximations for many natural subclasses of graphs. Finally, we present constant factor approximation schemes for the special classes of complete binary trees and random graphs.