arXiv++ Combinatorics

Browse math.CO papers from arXiv

uniform distribution

173 papers tagged with this keyword
2019-04-15 v3
Making multigraphs simple by a sequence of double edge swaps
We show that any loopy multigraph with a graphical degree sequence can be transformed into a simple graph by a finite sequence of double edge swaps with each swap involving at least one loop or multiple edge. Our result answers a question of Janson motivated by random graph theory, and it adds to the rich literature on reachability of double edge swaps with applications in Markov chain Monte Carlo sampling from the uniform distribution of graphs with prescribed degrees.
2019-03-14
Keyed hash function from large girth expander graphs
In this paper we present an algorithm to compute keyed hash function (message authentication code MAC). Our approach uses a family of expander graphs of large girth denoted $D(n,q)$, where $n$ is a natural number bigger than one and $q$ is a prime power. Expander graphs are known to have excellent expansion properties and thus they also have very good mixing properties. All requirements for a good MAC are satisfied in our method and a discussion about collisions and preimage resistance is also part of this work. The outputs closely approximate the uniform distribution and the results we get are indistinguishable from random sequences of bits. Exact formulas for timing are given in term of number of operations per bit of input. Based on the tests, our method for implementing DMAC shows good efficiency in comparison to other techniques. 4 operations per bit of input can be achieved. The algorithm is very flexible and it works with messages of any length. Many existing algorithms output a fixed length tag, while our constructions allow generation of an arbitrary length output, which is a big advantage.
2019-02-26 v3
Polynomial bound for the partition rank vs the analytic rank of tensors
Published in Discrete Analysis, 2020:7, 18pp • View PublicationBIB
A tensor defined over a finite field $\mathbb{F}$ has low analytic rank if the distribution of its values differs significantly from the uniform distribution. An order $d$ tensor has partition rank 1 if it can be written as a product of two tensors of order less than $d$, and it has partition rank at most $k$ if it can be written as a sum of $k$ tensors of partition rank 1. In this paper, we prove that if the analytic rank of an order $d$ tensor is at most $r$, then its partition rank is at most $f(r,d,|\mathbb{F}|)$, where, for fixed $d$ and $\mathbb{F}$, $f$ is a polynomial in $r$. This is an improvement of a recent result of the author, where he obtained a tower-type bound. Prior to our work, the best known bound was an Ackermann-type function in $r$ and $d$, though it did not depend on $\mathbb{F}$. It follows from our results that a biased polynomial has low rank; there too we obtain a polynomial dependence improving the previously known Ackermann-type bound. A similar polynomial bound for the partition rank was obtained independently and simultaneously by Milićević.
2019-01-28 v2
Random graphs with given vertex degrees and switchings
Random graphs with a given degree sequence are often constructed using the configuration model, which yields a random multigraph. We may adjust this multigraph by a sequence of switchings, eventually yielding a simple graph. We show that, assuming essentially a bounded second moment of the degree distribution, this construction with the simplest types of switchings yields a simple random graph with an almost uniform distribution, in the sense that the total variation distance is $o(1)$. This construction can be used to transfer results on distributional convergence from the configuration model multigraph to the uniform random simple graph with the given vertex degrees. As examples, we give a few applications to asymptotic normality. We show also a weaker result yielding contiguity when the maximum degree is too large for the main theorem to hold.
2018-12-31 v4
Cohen-Lenstra distributions via random matrices over complete discrete valuation rings with finite residue fields
Let $(R, \mathfrak{m})$ be a complete discrete valuation ring with the finite residue field $R/\mathfrak{m} = \mathbb{F}_{q}$. Given a monic polynomial $P(t) \in R[t]$ whose reduction modulo $\mathfrak{m}$ gives an irreducible polynomial $\bar{P}(t) \in \mathbb{F}_{q}[t]$, we initiate the investigation of the distribution of $\mathrm{coker}(P(A))$, where $A \in \mathrm{Mat}_{n}(R)$ is randomly chosen with respect to the Haar probability measure on the additive group $\mathrm{Mat}_{n}(R)$ of $n \times n$ $R$-matrices. One of our main results generalizes two results of Friedman and Washington. Our other results are related to the distribution of the $\bar{P}$-part of a random matrix $\bar{A} \in \mathrm{Mat}_{n}(\mathbb{F}_{q})$ with respect to the uniform distribution, and one of them generalizes a result of Fulman. We heuristically relate our results to a celebrated conjecture of Cohen and Lenstra, which predicts that given an odd prime $p$, any finite abelian $p$-group (i.e., $\mathbb{Z}_{p}$-module) $H$ occurs as the $p$-part of the class group of a random imaginary quadratic field extension of $\mathbb{Q}$ with a probability inversely proportional to $|\mathrm{Aut}_{\mathbb{Z}}(H)|$. We review three different heuristics for the conjecture of Cohen and Lenstra, and they are all related to special cases of our main conjecture, which we prove as our main theorems. For proofs, we use some concrete combinatorial connections between $\mathrm{Mat}_{n}(R)$ and $\mathrm{Mat}_{n}(\mathbb{F}_{q})$ to translate our problems about a Haar-random matrix in $\mathrm{Mat}_{n}(R)$ into problems about a random matrix in $\mathrm{Mat}_{n}(\mathbb{F}_{q})$ with respect to the uniform distribution.
2018-11-06
The entropy of lies: playing twenty questions with a liar
`Twenty questions' is a guessing game played by two players: Bob thinks of an integer between $1$ and $n$, and Alice's goal is to recover it using a minimal number of Yes/No questions. Shannon's entropy has a natural interpretation in this context. It characterizes the average number of questions used by an optimal strategy in the distributional variant of the game: let $μ$ be a distribution over $[n]$, then the average number of questions used by an optimal strategy that recovers $x\sim μ$ is between $H(μ)$ and $H(μ)+1$. We consider an extension of this game where at most $k$ questions can be answered falsely. We extend the classical result by showing that an optimal strategy uses roughly $H(μ) + k H_2(μ)$ questions, where $H_2(μ) = \sum_x μ(x)\log\log\frac{1}{μ(x)}$. This also generalizes a result by Rivest et al. for the uniform distribution. Moreover, we design near optimal strategies that only use comparison queries of the form `$x \leq c$?' for $c\in[n]$. The usage of comparison queries lends itself naturally to the context of sorting, where we derive sorting algorithms in the presence of adversarial noise.
2018-10-12
Uniform random posets
We propose a simple algorithm generating labelled posets of given size according to the almost uniform distribution. By "almost uniform" we understand that the distribution of generated posets converges in total variation to the uniform distribution. Our method is based on a Markov chain generating directed acyclic graphs.
2018-10-04 v2
The Four Point Permutation Test for Latent Block Structure in Incidence Matrices
Transactional data may be represented as a bipartite graph $G:=(L \cup R, E)$, where $L$ denotes agents, $R$ denotes objects visible to many agents, and an edge in $E$ denotes an interaction between an agent and an object. Unsupervised learning seeks to detect block structures in the adjacency matrix $Z$ between $L$ and $R$, thus grouping together sets of agents with similar object interactions. New results on quasirandom permutations suggest a non-parametric \textbf{four point test} to measure the amount of block structure in $G$, with respect to vertex orderings on $L$ and $R$. Take disjoint 4-edge random samples, order these four edges by left endpoint, and count the relative frequencies of the $4!$ possible orderings of the right endpoint. When these orderings are equiprobable, the edge set $E$ corresponds to a quasirandom permutation $π$ of $|E|$ symbols. Total variation distance of the relative frequency vector away from the uniform distribution on 24 permutations measures the amount of block structure. Such a test statistic, based on $\lfloor |E|/4 \rfloor$ samples, is computable in $O(|E|/p)$ time on $p$ processors. Possibly block structure may be enhanced by precomputing \textbf{natural orders} on $L$ and $R$, related to the second eigenvector of graph Laplacians. In practice this takes $O(d |E|)$ time, where $d$ is the graph diameter. Five open problems are described.
2018-09-28
Low analytic rank implies low partition rank for tensors
A tensor defined over a finite field $\mathbb{F}$ has low analytic rank if the distribution of its values differs significantly from the uniform distribution. An order $d$ tensor has partition rank 1 if it can be written as a product of two tensors of order less than $d$, and it has partition rank at most $k$ if it can be written as a sum of $k$ tensors of partition rank 1. In this paper, we prove that if the analytic rank of an order $d$ tensor is at most $r$, then its partition rank is at most $f(r,d,|\mathbb{F}|)$. Previously, this was known with $f$ being an Ackermann-type function in $r$ and $d$ but not depending on $\mathbb{F}$. The novelty of our result is that $f$ has only tower-type dependence on its parameters. It follows from our results that a biased polynomial has low rank; there too we obtain a tower-type dependence improving the previously known Ackermann-type bound.
The Best-or-Worst and the Postdoc problems with random number of candidates
In this paper we consider two variants of the Secretary problem: The Best-or-Worst and the Postdoc problems. We extend previous work by considering that the number of objects is not known and follows either a discrete Uniform distribution $\mathcal{U}[1,n]$ or a Poisson distribution $\mathcal{P}(λ)$. We show that in any case the optimal strategy is a threshold strategy, we provide the optimal cutoff values and the asymptotic probabilities of success. We also put our results in relation with closely related work.
2018-06-21
A note on log-concave random graphs
Published in Electron. J. Combin. 26 (2019), no. 3, Paper No. 3.36, 9 pp • View PublicationBIB
We establish a threshold for the connectivity of certain random graphs whose (dependent) edges are determined by the uniform distributions on generalized Orlicz balls, crucially using their negative correlation properties. We also show the existence of a unique giant component for such random graphs.
2018-04-30
A large deviation principle for the Erdős-Rényi uniform random graph
Published • View PublicationBIB
Starting with the large deviation principle (LDP) for the Erdős-Rényi binomial random graph $\mathcal{G}(n,p)$ (edge indicators are i.i.d.), due to Chatterjee and Varadhan (2011), we derive the LDP for the uniform random graph $\mathcal{G}(n,m)$ (the uniform distribution over graphs with $n$ vertices and $m$ edges), at suitable $m=m_n$. Applying the latter LDP we find that tail decays for subgraph counts in $\mathcal{G}(n,m_n)$ are controlled by variational problems, which up to a constant shift, coincide with those studied by Kenyon et al. and Radin et al. in the context of constrained random graphs, e.g., the edge/triangle model.
2018-03-20 v3
A probabilistic variant of Sperner's theorem and of maximal $r$-cover free families
Published in Discrete Mathematics, October 2020, volume 343, issue 10, article 112027 • View PublicationBIB
A family of sets is called $r$-\emph{cover free} if no set in the family is contained in the union of $r$ (or less) other sets in the family. A $1$-cover free family is simply an antichain with respect to set inclusion. Thus, Sperner's classical result determines the maximal cardinality of a $1$-cover free family of subsets of an $n$-element set. Estimating the maximal cardinality of an $r$-cover free family of subsets of an $n$-element set for $r>1$ was also studied. In this note we are interested in the following probabilistic variant of this problem. Let $S_0,S_1,\ldots, S_r$ be independent and identically distributed random subsets of an $n$-element set. Which distribution minimizes the probability that $S_0\subseteq {\bigcup_{i=1}^r S_i}$? A natural candidate is the uniform distribution on an $r$-cover-free family of maximal cardinality. We show that for $r=1$ such distribution is indeed best possible. In a complete contrast, we also show that this is far from being true for every $r>1$ and $n$ large enough.
2018-02-19 v3
Further results on random cubic planar graphs
Published • View PublicationBIB
We provide precise asymptotic estimates for the number of several classes of labelled cubic planar graphs, and we analyze properties of such random graphs under the uniform distribution. This model was first analyzed by Bodirsky et al. (Random Structures Algorithms 2007). We revisit their work and obtain new results on the enumeration of cubic planar graphs and on random cubic planar graphs. In particular, we determine the exact probability of a random cubic planar graph being connected, and we show that the distribution of the number of triangles in random cubic planar graphs is asymptotically normal with linear expectation and variance. To the best of our knowledge, this is the first time one is able to determine the asymptotic distribution for the number of copies of a fixed graph containing a cycle in classes of random planar graphs arising from planar maps.
2017-09-28 v3
Hypergraph expanders from Cayley graphs
Published • View PublicationBIB
We present a simple mechanism, which can be randomised, for constructing sparse $3$-uniform hypergraphs with strong expansion properties. These hypergraphs are constructed using Cayley graphs over $\mathbb{Z}_2^t$ and have vertex degree which is polylogarithmic in the number of vertices. Their expansion properties, which are derived from the underlying Cayley graphs, include analogues of vertex and edge expansion in graphs, rapid mixing of the random walk on the edges of the skeleton graph, uniform distribution of edges on large vertex subsets and the geometric overlap property.
2017-09-25 v2
Asymptotic Properties of Random Restricted Partitions
Published • View PublicationBIB
We study two types of probability measures on the set of integer partitions of $n$ with at most $m$ parts. The first one chooses the random partition with a chance related to its largest part only. We then obtain the limiting distributions of all of the parts together and that of the largest part as $n$ tends to infinity while $m$ is fixed or tends to infinity. In particular, if $m$ goes to infinity not too fast, the largest part satisfies the central limit theorem. The second measure is very general. It includes the Dirichlet distribution and the uniform distribution as special cases. We derive the asymptotic distributions of the parts jointly by taking limits of $n$ and $m$ in the same manner as that in the first probability measure.
2017-09-04 v3
The combinatorics of the colliding bullets problem
Published • View PublicationBIB
The finite colliding bullets problem is the following simple problem: consider a gun, whose barrel remains in a fixed direction; let $(V_i)_{1\le i\le n}$ be an i.i.d.\ family of random variables with uniform distribution on $[0,1]$; shoot $n$ bullets one after another at times $1,2,\dots, n$, where the $i$th bullet has speed $V_i$. When two bullets collide, they both annihilate. We give the distribution of the number of surviving bullets, and in some generalisation of this model. While the distribution is relatively simple (and we found a number of bold claims online), our proof is surprisingly intricate and mixes combinatorial and geometric arguments; we argue that any rigorous argument must very likely be rather elaborate.
2017-06-19 v2
Bernoulli Correlations and Cut Polytopes
Published • View PublicationBIB
Given $n$ symmetric Bernoulli variables, what can be said about their correlation matrix viewed as a vector? We show that the set of those vectors $R(\mathcal{B}_n)$ is a polytope and identify its vertices. Those extreme points correspond to correlation vectors associated to the discrete uniform distributions on diagonals of the cube $[0,1]^n$. We also show that the polytope is affinely isomorphic to a well-known cut polytope ${\rm CUT}(n)$ which is defined as a convex hull of the cut vectors in a complete graph with vertex set $\{1,\ldots,n\}$. The isomorphism is obtained explicitly as $R(\mathcal{B}_n)= {\mathbf{1}}-2~{\rm CUT}(n)$. As a corollary of this work, it is straightforward using linear programming to determine if a particular correlation matrix is realizable or not. Furthermore, a sampling method for multivariate symmetric Bernoullis with given correlation is obtained. In some cases the method can also be used for general, not exclusively Bernoulli, marginals.
On the Maximum Size of Block Codes Subject to a Distance Criterion
Published • View PublicationBIB
We establish a general formula for the maximum size of finite length block codes with minimum pairwise distance no less than $d$. The achievability argument involves an iterative construction of a set of radius-$d$ balls, each centered at a codeword. We demonstrate that the number of such balls that cover the entire code alphabet cannot exceed this maximum size. Our approach can be applied to codes $i)$ with elements over arbitrary code alphabets, and $ii)$ under a broad class of distance measures, thereby ensuring the generality of our formula. Our formula indicates that the maximum code size can be fully characterized by the cumulative distribution function of the distance measure evaluated at two independent and identically distributed random codewords. When the two random codewords assume a uniform distribution over the entire code alphabet, our formula recovers and obtains a natural generalization of the Gilbert-Varshamov (GV) lower bound. We also establish a general formula for the zero-error capacity of any sequence of channels. Finally, we extend our study to the asymptotic setting, where we establish first- and second-order bounds on the asymptotic code rate subject to a normalized minimum distance constraint.
2017-05-26
Probabilistic and Geometrical Applications to Graph Theory
This paper consists of two halves. In the first half of the paper, we consider real-valued functions $f$ whose domain is the vertex set of a graph $G$ and that are Lipschitz with respect to the graph distance. By placing a uniform distribution on the vertex set, we treat $f$ as a random variable. We investigate the link between the isoperimetric function of $G$ and the functions $f$ that have maximum variance or meet the bound established by the subgaussian inequality. We present several results describing the extremal functions, and use those results to resolve: (A) a conjecture by Bobkov, Houdré, and Tetali characterizing the extremal functions of the subgaussian inequality of the odd cycle, and (B) a conjecture by Alon, Boppana, and Spencer on the relationship between maximum variance functions and the isoperimetric function of product graphs. While establishing a discrete analogue of the curved Brunn-Minkowski inequality for the discrete hypercube, Ollivier and Villani suggested several avenues for research. We resolve them in second half of the paper as follows. (1) They propose that a bound on $t$-midpoints can be obtained by repeated application of the bound on midpoints, if the original sets are convex. We construct a specific example where this reasoning fails, and then prove our construction is general by characterizing the convex sets in the discrete hypercube. (2) A second proposed technique to bound $t$-midpoints involves new results in concentration of measure. We follow through on this proposal, with heavy use on results from the first half of the paper. (3) We show that the curvature of the discrete hypercube is not positive or zero.