arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

6952 papers tagged with this keyword
2021-12-26
Mallows permutation models with $L^1$ and $L^2$ distances I: hit and run algorithms and mixing times
Mallows permutation model, introduced by Mallows in statistical ranking theory, is a class of non-uniform probability measures on the symmetric group $S_n$. The model depends on a distance metric $d(σ,τ)$ on $S_n$, which can be chosen from a host of metrics on permutations. In this paper, we focus on Mallows permutation models with $L^1$ and $L^2$ distances, respectively known in the statistics literature as Spearman's footrule and Spearman's rank correlation. Unlike most of the random permutation models that have been analyzed in the literature, Mallows permutation models with $L^1$ and $L^2$ distances do not have an explicit expression for their normalizing constants. This poses challenges to the task of sampling from these Mallows models. In this paper, we consider hit and run algorithms for sampling from both models. Hit and run algorithms are a unifying class of Markov chain Monte Carlo (MCMC) algorithms including the celebrated Swendsen-Wang and data augmentation algorithms. For both models, we show order $\log{n}$ mixing time upper bounds for the hit and run algorithms. This demonstrates much faster mixing of the hit and run algorithms compared to local MCMC algorithms such as the Metropolis algorithm. The proof of the results on mixing times is based on the path coupling technique, for which a novel coupling for permutations with one-sided restrictions is involved. Extensions of the hit and run algorithms to weighted versions of the above models, a two-parameter permutation model that involves the $L^1$ distance and Cayley distance, and lattice permutation models in dimensions greater than or equal to $2$ are also discussed. The order $\log{n}$ mixing time upper bound pertains to the two-parameter permutation model.
2021-12-25 v2
Modularity and partially observed graphs
Suppose that there is an unknown underlying graph $G$ on a large vertex set, and we can test only a proportion of the possible edges to check whether they are present in $G$. If $G$ has high modularity, is the observed graph $G'$ likely to have high modularity? We see that this is indeed the case under a mild condition, in a natural model where we test edges at random. We find that $q^*(G') \geq q^*(G)-\varepsilon$ with probability at least $1-\varepsilon$, as long as the expected number edges in $G'$ is large enough. Similarly, $q^*(G') \leq q^*(G)+\varepsilon$ with probability at least $1-\varepsilon$, under the stronger condition that the expected average degree in $G'$ is large enough. Further, under this stronger condition, finding a good partition for $G'$ helps us to find a good partition for $G$. We also consider the vertex sampling model for partially observing the underlying graph: we find that for dense underlying graphs we may estimate the modularity by sampling constantly many vertices and observing the corresponding induced subgraph, but this does not hold for underlying graphs with a subquadratic number of edges. Finally we deduce some related results, for example showing that under-sampling tends to lead to overestimation of modularity.
2021-12-25 v2
Rainbow connectivity of randomly perturbed graphs
Published • View PublicationBIB
In this note we examine the following random graph model: for an arbitrary graph $H$, with quadratic many edges, construct a graph $G$ by randomly adding $m$ edges to $H$ and randomly coloring the edges of $G$ with $r$ colors. We show that for $m$ a large enough constant and $r \geq 5$, every pair of vertices in $G$ are joined by a rainbow path, i.e., $G$ is {\it rainbow connected}, with high probability. This confirms a conjecture of Anastos and Frieze [{\it J. Graph Theory} {\bf 92} (2019)] who proved the statement for $r \geq 7$ and resolved the case when $r \leq 4$ and $m$ is a function of $n$.
Asymptotic Bounds on the Combinatorial Diameter of Random Polytopes
Published • View PublicationBIB
The combinatorial diameter $\operatorname{diam}(P)$ of a polytope $P$ is the maximum shortest path distance between any pair of vertices. In this paper, we provide upper and lower bounds on the combinatorial diameter of a random "spherical" polytope, which is tight to within one factor of dimension when the number of inequalities is large compared to the dimension. More precisely, for an $n$-dimensional polytope $P$ defined by the intersection of $m$ i.i.d.\ half-spaces whose normals are chosen uniformly from the sphere, we show that $\operatorname{diam}(P)$ is $Ω(n m^{\frac{1}{n-1}})$ and $O(n^2 m^{\frac{1}{n-1}} + n^5 4^n)$ with high probability when $m \geq 2^{Ω(n)}$. For the upper bound, we first prove that the number of vertices in any fixed two dimensional projection sharply concentrates around its expectation when $m$ is large, where we rely on the $Θ(n^2 m^{\frac{1}{n-1}})$ bound on the expectation due to Borgwardt [Math. Oper. Res., 1999]. To obtain the diameter upper bound, we stitch these ``shadows paths'' together over a suitable net using worst-case diameter bounds to connect vertices to the nearest shadow. For the lower bound, we first reduce to lower bounding the diameter of the dual polytope $P^\circ$, corresponding to a random convex hull, by showing the relation $\operatorname{diam}(P) \geq (n-1)(\operatorname{diam}(P^\circ)-2)$. We then prove that the shortest path between any ``nearly'' antipodal pair vertices of $P^\circ$ has length $Ω(m^{\frac{1}{n-1}})$.
2021-12-23 v4
A combinatorial proof of the Gaussian product inequality beyond the MTP${}_2$ case
Published in Dependence Modeling (2022), 10 (1), 236-244 • View PublicationBIB
A combinatorial proof of the Gaussian product inequality (GPI) is given under the assumption that each component of a centered Gaussian random vector $\boldsymbol{X} = (X_1, \ldots, X_d)$ of arbitrary length can be written as a linear combination, with coefficients of identical sign, of the components of a standard Gaussian random vector. This condition on $\boldsymbol{X}$ is shown to be strictly weaker than the assumption that the density of the random vector $(|X_1|, \ldots, |X_d|)$ is multivariate totally positive of order $2$, abbreviated MTP${}_2$, for which the GPI is already known to hold. Under this condition, the paper highlights a new link between the GPI and the monotonicity of a certain ratio of gamma functions.
2021-12-23 v2
The threshold for stacked triangulations
Published • View PublicationBIB
A \emph{stacked triangulation} of a $d$-simplex $\mathbf{o}=\{1,\ldots,d+1\}$ ($d\geq 2$) is a triangulation obtained by repeatedly subdividing a $d$-simplex into $d+1$ new ones via a new vertex (the case $d=2$ is known as an Appolonian network). We study the occurrence of such a triangulation in the Linial--Meshulam model, i.e., for which $p$ does the random simplicial complex $Y\sim \mathcal{Y}_d(n,p)$ contain the faces of a stacked triangulation of the $d$-simplex $\mathbf{o}$, with its internal vertices labeled in $[n]$. In the language of bootstrap percolation in hypergraphs, it pertains to the threshold for $K_{d+2}^{d+1}$, the $(d+1)$-uniform clique on $d+2$ vertices. Our main result identifies this threshold for every $d\geq 2$, showing it is asymptotically $(α_d n)^{-1/d}$, where $α_d$ is the growth rate of the Fuss--Catalan numbers of order $d$. The proof hinges on a second moment argument in the supercritical regime, and on Kalai's algebraic shifting in the subcritical regime.
2021-12-22 v3
Constructions and bounds for subspace codes
Subspace codes are the $q$-analog of binary block codes in the Hamming metric. Here the codewords are vector spaces over a finite field. They have e.g. applications in random linear network coding, distributed storage, and cryptography. In this chapter we survey known constructions and upper bounds for subspace codes.
2021-12-22 v2
An algorithm for generating random mixed-arity trees
Inspired by [4] we present a new algorithm for uniformly random generation of ordered trees in which all occuring outdegrees can be specified by a given sequence of numbers. The method can be used for random generation of binary or n-ary trees, or ones with various arities. We show that the algorithm is correct and has $O(n)$ time complexity for $n$ being the desired number of nodes in the resulting tree. In the discussion part we show how some selected formulas can be derived with the use of ideas developed in the proof of correctness of the algorithm.
Comparing balanced $\mathbb{Z}_v$-sequences obtained from ElGamal function to random balanced sequences
Published • View PublicationBIB
In this paper, we investigate the randomness properties of sequences in $\mathbb{Z}_v$ derived from permutations in $\mathbb{Z}_{p}^*$ using the remainder function modulo $v$, where $p$ is a prime integer. Motivated by earlier studies with a cryptographic focus we compare sequences constructed from the ElGamal function $x \to g^x$ for $x\in\mathbb{Z}_{>0}$ and $g$ a primitive element of $\mathbb{Z}_{p}^*$, to sequences constructed from random permutations of $\mathbb{Z}_{p}^*$. We prove that sequences obtained from ElGamal have maximal period and behave similarly to random permutations with respect to the balance and run properties of Golomb's postulates for pseudo-random sequences. Additionally we show that they behave similarly to random permutations for the tuple balance property. This requires some significant work determining properties of random balanced periodic sequences. In general, for these properties and excepting for very unlikely events, the ElGamal sequences behave the same as random balanced sequences.
2021-12-21 v2
Exponential decay of intersection volume with applications on list-decodability and Gilbert-Varshamov type bound
Published • View PublicationBIB
We give some natural sufficient conditions for balls in a metric space to have small intersection. Roughly speaking, this happens when the metric space is (i) expanding and (ii) well-spread, and (iii) a certain random variable on the boundary of a ball has a small tail. As applications, we show that the volume of intersection of balls in Hamming, Johnson spaces and symmetric groups decay exponentially as their centers drift apart. To verify condition (iii), we prove some large deviation inequalities `on a slice' for functions with Lipschitz conditions. We then use these estimates on intersection volumes to $\bullet$ obtain a sharp lower bound on list-decodability of random $q$-ary codes, confirming a conjecture of Li and Wootters; and $\bullet$ improve the classical bound of Levenshtein from 1971 on constant weight codes by a factor linear in dimension, resolving a problem raised by Jiang and Vardy. Our probabilistic point of view also offers a unified framework to obtain improvements on other Gilbert--Varshamov type bounds, giving conceptually simple and calculation-free proofs for $q$-ary codes, permutation codes, and spherical codes. Another consequence is a counting result on the number of codes, showing ampleness of large codes.
The mesoscopic geometry of sparse random maps
Published in Journal de l{\textquoteright}École polytechnique {\textemdash} Mathématiques, Tome 9 (2022), pp. 1305-1345 • View PublicationBIB
We investigate the structure of large uniform random maps with $n$ edges, $\mathrm{f}_n$ faces, and with genus $\mathrm{g}_n$ in the so-called sparse case, where the ratio between the number vertices and edges tends to $1$. We focus on two regimes: the planar case $(\mathrm{f}_n, 2\mathrm{g}_n) = (\mathrm{s}_n, 0)$ and the unicellular case with moderate genus $(\mathrm{f}_n, 2 \mathrm{g}_n) = (1, \mathrm{s}_n-1)$, both when $1 \ll \mathrm{s}_n \ll n$. Albeit different at first sight, these two models can be treated in a unified way using a probabilistic version of the classical core-kernel decomposition. In particular, we show that the number of edges of the core of such maps, obtained by iteratively removing degree $1$ vertices, is concentrated around $\sqrt{n \mathrm{s}_{n}}$. Further, their kernel, obtained by contracting the vertices of the core with degree $2$, is such that the sum of the degree of its vertices exceeds that of a trivalent map by a term of order $\sqrt{\mathrm{s}_{n}^{3}/n}$; in particular they are trivalent with high probability when $\mathrm{s}_{n} \ll n^{1/3}$. This enables us to identify a mesoscopic scale $\sqrt{n/\mathrm{s}_n}$ at which the scaling limits of these random maps can be seen as the local limit of their kernels, which is the dual of the UIPT in the planar case and the infinite three-regular tree in the unicellular case, where each edge is replaced by an independent (biased) Brownian tree with two marked points.
2021-12-17
The strong component structure of the barely subcritical directed configuration model
We study the behaviour of the largest components of the directed configuration model in the barely subcritical regime. We show that with high probability all strongly connected components in this regime are either cycles or isolated vertices and give an asymptotic distribution of the size of the $k$th largest cycle. This gives a configuration model analogue of a result of Łuczak and Seierstad for the binomial random digraph.
Hat guessing numbers of strongly degenerate graphs
Published • View PublicationBIB
Assume $n$ players are placed on the $n$ vertices of a graph $G$. The following game was introduced by Winkler: An adversary puts a hat on each player, where each hat has a colour out of $q$ available colours. The players can see the hat of each of their neighbours in $G$, but not their own hat. Using a prediscussed guessing strategy, the players then simultaneously guess the colour of their hat. The players win if at least one of them guesses correctly, else the adversary wins. The largest integer $q$ such that there is a winning strategy for the players is denoted by $\text{HG}(G)$, and this is called the hat guessing number of $G$. Although this game has received a lot of attention in the recent years, not much is known about how the hat guessing number relates to other graph parameters. For instance, a natural open question is whether the hat guessing number can be bounded from above in terms of degeneracy. In this paper, we prove that the hat guessing number of a graph can be bounded from above in terms of a related notion, which we call strong degeneracy. We further give an exact characterisation of graphs with bounded strong degeneracy. As a consequence, we significantly improve the best known upper bound on the hat guessing number of outerplanar graphs from $2^{125000}$ to $40$, and further derive upper bounds on the hat guessing number for any class of $K_{2,s}$-free graphs with bounded expansion, such as the class of $C_4$-free planar graphs, more generally $K_{2,s}$-free graphs with bounded Hadwiger number or without a $K_t$-subdivision, and for Erdős-Rényi random graphs with constant average degree.
2021-12-17 v4
A central limit theorem for cycles of Mallows permutations
Fix $q\neq 1$, and sample $w\in S_n$ from the Mallows measure. We study the distribution of $C_i(w)$, the number of $i$-cycles, as $n$ grows large. When $q<1$, they are jointly Gaussian, and this more or less follows from known ideas, but the regime $q>1$ behaves quite differently. In particular, we show that the even cycles $C_{2i}(w)$ have a mean and variance of order $n$, and jointly converge to Gaussian random variables, while the odd cycles $C_{2i+1}(w)$ have a bounded mean and variance, and converge to $C_{2i+1}(w^{even})$ or $C_{2i+1}(w^{odd})$ for some explicit random permutations $w^{even}$ and $w^{odd}$, depending on whether $n$ is even or odd. An extension to a larger class of functions is also given. The proof utilizes a two-sided stationary regenerative process associated to Mallows permutations constructed by Gnedin and Olshanski, extending the ideas of Basu and Bhatnagar.
2021-12-16
Random Walk Models for Nontrivial Identities of Bernoulli and Euler Polynomials
We consider the $1$-dimensional reflected Brownian motion and $3$-dimensional Bessel process and the general models. By decomposing the hitting times of consecutive sites into loops, we obtain identities, called loop identities, for the generating functions of the hitting times. After proving this decomposition both combinatorially and inductively, we consider the case that sites are equally distributed. Then, from loop identities, we derive expressions of Bernoulli and Euler polynomials, in terms of Euler polynomials of higher-orders.
2021-12-16 v2
Multi-orbit cyclic subspace codes and linear sets
Published • View PublicationBIB
Cyclic subspace codes gained a lot of attention especially because they may be used in random network coding for correction of errors and erasures. Roth, Raviv and Tamo in 2018 established a connection between cyclic subspace codes (with certain parameters) and Sidon spaces. These latter objects were introduced by Bachoc, Serra and Zémor in 2017 in relation with the linear analogue of Vosper's Theorem. This connection allowed Roth, Raviv and Tamo to construct large classes of cyclic subspace codes with one or more orbits. In this paper we will investigate cyclic subspace codes associated to a set of Sidon spaces, that is cyclic subspace codes with more than one orbit. Moreover, we will also use the geometry of linear sets to provide some bounds on the parameters of a cyclic subspace code. Conversely, cyclic subspace codes are used to construct families of linear sets which extend a class of linear sets recently introduced by Napolitano, Santonastaso, Polverino and the author. This yields large classes of linear sets with a special pattern of intersection with the hyperplanes, defining rank metric and Hamming metric codes with only three distinct weights.
2021-12-14
On the Gauss-Epple homomorphism of the braid group $B_n$, and generalizations to Artin groups of crystallographic type
In this paper, we introduce a broad family of group homomorphisms that we name the Gauss-Epple homomorphisms. In the setting of braid groups, the Gauss-Epple invariant was originally defined by Epple based on a note of Gauss as an action of the braid group $B_n$ on the set $\{1, \dots, n\}\times\mathbb{Z}$; we prove that it is well-defined. We consider the associated group homomorphism from $B_n$ to the symmetric group $\text{Sym}(\{1, \dots, n\}\times\mathbb{Z})$. We prove that this homomorphism factors through $\mathbb{Z}^n\rtimes S_n$ (in fact, its image is an order 2 subgroup of the previous group). We also describe the kernel of the homomorphism and calculate the asymptotic probability that it contains a random braid of a given length. Furthermore, we discuss the super-Gauss-Epple homomorphism, a homomorphism which extends the generalization of the Gauss-Epple homomorphism and describe a related 1-cocycle of the symmetric group $S_n$ on the set of antisymmetric $n\times n$ matrices over the integers. We then generalize the super-Gauss-Epple homomorphism and the associated 1-cocycle to Artin groups of finite type. For future work, we suggest studying possible generalizations to complex reflection groups and computing the vector spaces of Gauss-Epple analogues.
2021-12-14 v2
Fixed points, descents, and inversions in parabolic double cosets of the symmetric group
We consider statistics on permutations chosen uniformly at random from fixed parabolic double cosets of the symmetric group. We show that the distribution of fixed points is asymptotically Poisson and establish central limit theorems for the distribution of descents and inversions. Our proofs use Stein's method with size-bias coupling and dependency graphs, which also gives convergence rates for our distributional approximations. As applications of our size-bias coupling and dependency graph constructions, we obtain concentration of measure results on the number of fixed points, descents, and inversions.
Sparse random graphs with many triangles
Published • View PublicationBIB
In this paper we consider the Erdős-Rényi random graph in the sparse regime in the limit as the number of vertices $n$ tends to infinity. We are interested in what this graph looks like when it contains many triangles, in two settings. First, we derive asymptotically sharp bounds on the probability that the graph contains a large number of triangles. We show that conditionally on this event, with high probability the graph contains an almost complete subgraph, i.e., the triangles form a near-clique, and has the same local limit as the original Erdős-Rényi random graph. Second, we derive asymptotically sharp bounds on the probability that the graph contains a large number of vertices that are part of a triangle. If order $n$ vertices are in triangles, then the local limit (provided it exists) is different from that of the Erdős-Rényi random graph. Our results shed light on the challenges that arise in the description of real-world networks, which often are sparse, yet highly clustered, and on exponential random graphs, which often are used to model such networks.
2021-12-10
Expected value of letters of permutations with a given number of $k$-cycles
In this paper, we study permutations $π\in S_n$ with exactly $m$ transpositions. In particular, we are interested in the expected value of $π(1)$ when such permutations are chosen uniformly at random. When $n$ is even, this expected value is approximated closely by $(n+1)/2$, with an error term that is related to the number isometries of the $(n/2-m)$-dimensional hypercube that move every face. Furthermore, when $k \mid n$, this construction generalizes to allow us to compute the expected value of $π(1)$ for permutations with exactly $m$ $k$-cycles. In this case, the expected value has an error term which is related instead to the number derangements of the generalized symmetric group $S(k,n/k-m)$. When $k$ does not divide $n$, the expected value of $π(1)$ is precisely $(n+1)/2$. Indirectly, this suggests the existence of a reversible algorithm to insert a letter into a permutation which preserves the number of $k$-cycles, which we construct.