uniform distribution
173 papers tagged with this keyword
A generalization of Kátai's orthogonality criterion with applications
Published in Discrete and Continuous Dynamical Systems, Volume 39 (2019), Number 5, pp. 2581-2612
• View Publication
• BIB
We study properties of arithmetic sets coming from multiplicative number theory and obtain applications in the theory of uniform distribution and ergodic theory. Our main theorem is a generalization of Kátai's orthogonality criterion. Here is a special case of this theorem:
Let $a\colon\mathbb{N}\to\mathbb{C}$ be a bounded sequence satisfying $$ \sum_{n\leq x} a(pn)\overline{a(qn)} = {\rm o}(x),~\text{for all distinct primes $p$ and $q$.} $$ Then for any multiplicative function $f$ and any $z\in\mathbb{C}$ the indicator function of the level set $E=\{n\in\mathbb{N}:f(n)=z\}$ satisfies $$ \sum_{n\leq x} \mathbb{1}_E(n)a(n)={\rm o}(x). $$
With the help of this theorem one can show that if $E=\{n_1<n_2<\ldots\}$ is a level set of a multiplicative function having positive upper density, then for a large class of sufficiently smooth functions $h\colon(0,\infty)\to\mathbb{R}$ the sequence $(h(n_j))_{j\in\mathbb{N}}$ is uniformly distributed $\bmod~1$. This class of functions $h(t)$ includes: all polynomials $p(t)=a_kt^k+\ldots+a_1t+a_0$ such that at least one of the coefficients $a_1,a_2,\ldots,a_k$ is irrational, $t^c$ for any $c>0$ with $c\notin \mathbb{N}$, $\log^r(t)$ for any $r>2$, $\log(Γ(t))$, $t\log(t)$, and $\frac{t}{\log t}$. The uniform distribution results, in turn, allow us to obtain new examples of ergodic sequences, i.e. sequences along which the ergodic theorem holds.
The flip Markov chain for connected regular graphs
Published
• View Publication
• BIB
Mahlmann and Schindelhauer (2005) defined a Markov chain which they called $k$-Flipper, and showed that it is irreducible on the set of all connected regular graphs of a given degree (at least 3). We study the 1-Flipper chain, which we call the flip chain, and prove that the flip chain converges rapidly to the uniform distribution over connected $2r$-regular graphs with $n$ vertices, where $n\geq 8$ and $r = r(n)\geq 2$. Formally, we prove that the distribution of the flip chain will be within $\varepsilon$ of uniform in total variation distance after $\text{poly}(n,r,\log(\varepsilon^{-1}))$ steps. This polynomial upper bound on the mixing time is given explicitly, and improves markedly on a previous bound given by Feder et al.(2006). We achieve this improvement by using a direct two-stage canonical path construction, which we define in a general setting.
This work has applications to decentralised networks based on random regular connected graphs of even degree, as a self-stabilising protocol in which nodes spontaneously perform random flips in order to repair the network.
Tournament limits: Degree distributions, score functions and self-converseness
Motivated by known results for finite tournaments, we define and study the score functions of tournament kernels and the degree distributions of tournament limits. Our main theorem completely characterises those distributions that appear as the degree distribution of some tournament limit and those functions that appear as the score function of some tournament kernel. We also show that only the uniform distribution can be realised as the outdegree distribution of a unique tournament limit. Finally we define self-converse tournament limits and kernels and characterise their degree distributions and score functions.
The probability of avoiding consecutive patterns in the Mallows distribution
Published
• View Publication
• BIB
We use various combinatorial and probabilistic techniques to study growth rates for the probability that a random permutation from the Mallows distribution avoids consecutive patterns. The Mallows distribution behaves like a $q$-analogue of the uniform distribution by weighting each permutation $π$ by $q^{inv(π)}$, where $inv(π)$ is the number of inversions in $π$ and $q$ is a positive, real-valued parameter. We prove that the growth rate exists for all patterns and all $q>0$, and we generalize Goulden and Jackson's cluster method to keep track of the number of inversions in permutations avoiding a given consecutive pattern. Using singularity analysis, we approximate the growth rates for length-3 patterns, monotone patterns, and non-overlapping patterns starting with 1, and we compare growth rates between different patterns. We also use Stein's method to show that, under certain assumptions on $q$, the length of $σ$, and $inv(σ)$, the number of occurrences of a given pattern $σ$ is well approximated by the normal distribution.
Longest monotone subsequences and rare regions of pattern-avoiding permutations
Published in Electronic Journal of Combinatorics, Volume 24 (2017), Issue 4, Paper #P4.13
• View Publication
• BIB
We consider the distributions of the lengths of the longest monotone and alternating subsequences in classes of permutations of size $n$ that avoid a specific pattern or set of patterns, with respect to the uniform distribution on each such class. We obtain exact results for any class that avoids two patterns of length 3, as well as results for some classes that avoid one pattern of length 4 or more. In our results, the longest monotone subsequences have expected length proportional to $n$ for pattern-avoiding classes, in contrast with the $\sqrt n$ behaviour that holds for unrestricted permutations.
In addition, for a pattern $τ$ of length $k$, we scale the plot of a random $τ$-avoiding permutation down to the unit square and study the "rare region," which is the part of the square that is exponentially unlikely to contain any points. We prove that when $τ_1>τ_k$, the complement of the rare region is a closed set that contains the main diagonal of the unit square. For the case $τ_1=k,$ we also show that the lower boundary of the part of the rare region above the main diagonal is a curve that is Lipschitz continuous and strictly increasing on $[0,1]$.
Uniform measures on braid monoids and dual braid monoids
Published in Journal of Algebra, Elsevier, Volume 473, March 2017, pages 627-666
• View Publication
• BIB
We aim at studying the asymptotic properties of typical positive braids, respectively positive dual braids. Denoting by $μ_k$ the uniform distribution on positive (dual) braids of length $k$, we prove that the sequence $(μ_k)_k$ converges to a unique probability measure $μ_{\infty}$ on infinite positive (dual) braids. The key point is that the limiting measure $μ_{\infty}$ has a Markovian structure which can be described explicitly using the combinatorial properties of braids encapsulated in the Möbius polynomial. As a by-product, we settle a conjecture by Gebhardt and Tawn (J. Algebra, 2014) on the shape of the Garside normal form of large uniform braids.
The distribution of minimum-weight cliques and other subgraphs in graphs with random edge weights
Published
• View Publication
• BIB
We determine, asymptotically in $n$, the distribution and mean of the weight of a minimum-weight $k$-clique (or any strictly balanced graph $H$) in a complete graph $K_n$ whose edge weights are independent random values drawn from the uniform distribution or other continuous distributions. For the clique, we also provide explicit (non-asymptotic) bounds on the distribution's CDF in a form obtained directly from the Stein-Chen method, and in a looser but simpler form. The direct form extends to other subgraphs and other edge-weight distributions. We illustrate the clique results for various values of $k$ and $n$. The results may be applied to evaluate whether an observed minimum-weight copy of a graph $H$ in a network provides statistical evidence that the network's edge weights are not independently distributed but have some structure.
Walking on the Edge and Cosystolic Expansion
Random walks on regular bounded degree expander graphs have numerous applications. A key property of these walks is that they converge rapidly to the uniform distribution on the vertices. The recent study of expansion of high dimensional simplicial complexes, which are the high dimensional analogues of graphs, calls for the natural generalization of random walks to higher dimensions. In particular, a high order random walk on a $2$-dimensional simplicial complex moves at random between neighboring edges of the complex, where two edges are considered neighbors if they share a common triangle. We show that if a regular $2$-dimensional simplicial complex is a cosystolic expander and the underlying graph of the complex has a spectral gap larger than $1/2$, then the random walk on the edges of the complex converges rapidly to the uniform distribution on the edges.
Poset edge densities, nearly reduced words, and barely set-valued tableaux
Published in J. Combin. Theory Ser. A 158 (2018), 66-125
• View Publication
• BIB
In certain finite posets, the expected down-degree of their elements is the same whether computed with respect to either the uniform distribution or the distribution weighting an element by the number of maximal chains passing through it. We show that this coincidence of expectations holds for Cartesian products of chains, connected minuscule posets, weak Bruhat orders on finite Coxeter groups, certain lower intervals in Young's lattice, and certain lower intervals in the weak Bruhat order below dominant permutations. Our tools involve formulas for counting nearly reduced factorizations in 0-Hecke algebras; that is, factorizations that are one letter longer than the Coxeter group length.
Size biased couplings and the spectral gap for random regular graphs
Published in Ann. Probab., 46(1):72-125, 2018
• View Publication
• BIB
Let $λ$ be the second largest eigenvalue in absolute value of a uniform random $d$-regular graph on $n$ vertices. It was famously conjectured by Alon and proved by Friedman that if $d$ is fixed independent of $n$, then $λ=2\sqrt{d-1} +o(1)$ with high probability. In the present work we show that $λ=O(\sqrt{d})$ continues to hold with high probability as long as $d=O(n^{2/3})$, making progress towards a conjecture of Vu that the bound holds for all $1\le d\le n/2$. Prior to this work the best result was obtained by Broder, Frieze, Suen and Upfal (1999) using the configuration model, which hits a barrier at $d=o(n^{1/2})$. We are able to go beyond this barrier by proving concentration of measure results directly for the uniform distribution on $d$-regular graphs. These come as consequences of advances we make in the theory of concentration by size biased couplings. Specifically, we obtain Bennett-type tail estimates for random variables admitting certain unbounded size biased couplings.
Symmetric Graphs with respect to Graph Entropy
Published
• View Publication
• BIB
Let $F_G(P)$ be a functional defined on the set of all the probability distributions on the vertex set of a graph $G$. We say that $G$ is \emph{symmetric with respect to $F_G(P)$} if the uniform distribution on $V(G)$ maximizes $F_G(P)$. Using the combinatorial definition of the entropy of a graph in terms of its vertex packing polytope and the relationship between the graph entropy and fractional chromatic number, we characterize all graphs which are symmetric with respect to graph entropy. We show that a graph is symmetric with respect to graph entropy if and only if its vertex set can be uniformly covered by its maximum size independent sets. Furthermore, given any strictly positive probability distribution $P$ on the vertex set of a graph $G$, we show that $P$ is a maximizer of the entropy of graph $G$ if and only if its vertex set can be uniformly covered by its maximum weighted independent sets. We also show that the problem of deciding if a graph is symmetric with respect to graph entropy, where the weight of the vertices is given by probability distribution $P$, is co-NP-hard.
Random Interval Graphs
Published in J. Discrete Math. Sci. Cryptography 20 (8): 1697-1720, 2017
• View Publication
• BIB
In this thesis, which is supervised by Dr. David Penman, we examine random interval graphs. Recall that such a graph is defined by letting $X_{1},\ldots X_{n},Y_{1},\ldots Y_{n}$ be $2n$ independent random variables, with uniform distribution on $[0,1]$. We then say that the $i$th of the $n$ vertices is the interval $[X_{i},Y_{i}]$ if $X_{i}<Y_{i}$ and the interval $[Y_{i},X_{i}]$ if $Y_{i}<X_{i}$. We then say that two vertices are adjacent if and only if the corresponding intervals intersect.
We recall from our MA902 essay that fact that in such a graph, each edge arises with probability $2/3$, and use this fact to obtain estimates of the number of edges. Next, we turn to how these edges are spread out, seeing that (for example) the range of degrees for the vertices is much larger than classically, by use of an interesting geometrical lemma. We further investigate the maximum degree, showing it is always very close to the maximum possible value $(n-1)$, and the striking result that it is equal to $(n-1)$ with probability exactly $2/3$. We also recall a result on the minimum degree, and contrast all these results with the much narrower range of values obtained in the alternative \lq comparable\rq\, model $G(n,2/3)$ (defined later).
We then study clique numbers, chromatic numbers and independence numbers in the Random Interval Graphs, presenting (for example) a result on independence numbers which is proved by considering the largest chain in the associated interval order.
Last, we make some brief remarks about other ways to define random interval graphs, and extensions of random interval graphs, including random dot product graphs and other ways to define random interval graphs. We also discuss some areas these ideas should be usable in. We close with a summary and some comments.
Cutoff on all Ramanujan graphs
Published
• View Publication
• BIB
We show that on every Ramanujan graph $G$, the simple random walk exhibits cutoff: when $G$ has $n$ vertices and degree $d$, the total-variation distance of the walk from the uniform distribution at time $t=\frac{d}{d-2}\log_{d-1} n + s\sqrt{\log n}$ is asymptotically $\mathbb{P}(Z > c\, s)$ where $Z$ is a standard normal variable and $c=c(d)$ is an explicit constant. Furthermore, for all $1 \leq p \leq \infty$, $d$-regular Ramanujan graphs minimize the asymptotic $L^p$-mixing time for SRW among all $d$-regular graphs. Our proof also shows that, for every vertex $x$ in $G$ as above, its distance from $n-o(n)$ of the vertices is asymptotically $\log_{d-1} n$.
Generic properties of subgroups of free groups and finite presentations
Published in Contemporary Mathematics 677 (2016) 1-44
• View Publication
• BIB
Asymptotic properties of finitely generated subgroups of free groups, and of finite group presentations, can be considered in several fashions, depending on the way these objects are represented and on the distribution assumed on these representations: here we assume that they are represented by tuples of reduced words (generators of a subgroup) or of cyclically reduced words (relators). Classical models consider fixed size tuples of words (e.g. the few-generator model) or exponential size tuples (e.g. Gromov's density model), and they usually consider that equal length words are equally likely. We generalize both the few-generator and the density models with probabilistic schemes that also allow variability in the size of tuples and non-uniform distributions on words of a given length.Our first results rely on a relatively mild prefix-heaviness hypothesis on the distributions, which states essentially that the probability of a word decreases exponentially fast as its length grows. Under this hypothesis, we generalize several classical results: exponentially generically a randomly chosen tuple is a basis of the subgroup it generates, this subgroup is malnormal and the tuple satisfies a small cancellation property, even for exponential size tuples. In the special case of the uniform distribution on words of a given length, we give a phase transition theorem for the central tree property, a combinatorial property closely linked to the fact that a tuple freely generates a subgroup. We then further refine our results when the distribution is specified by a Markovian scheme, and in particular we give a phase transition theorem which generalizes the classical results on the densities up to which a tuple of cyclically reduced words chosen uniformly at random exponentially generically satisfies a small cancellation property, and beyond which it presents a trivial group.
Bias vs structure of polynomials in large fields, and applications in information theory
Published
• View Publication
• BIB
Let $f$ be a polynomial of degree $d$ in $n$ variables over a finite field $\mathbb{F}$. The polynomial is said to be unbiased if the distribution of $f(x)$ for a uniform input $x \in \mathbb{F}^n$ is close to the uniform distribution over $\mathbb{F}$, and is called biased otherwise. The polynomial is said to have low rank if it can be expressed as a composition of a few lower degree polynomials. Green and Tao [Contrib. Discrete Math 2009] and Kaufman and Lovett [FOCS 2008] showed that bias implies low rank for fixed degree polynomials over fixed prime fields. This lies at the heart of many tools in higher order Fourier analysis. In this work, we extend this result to all prime fields (of size possibly growing with $n$). We also provide a generalization to nonprime fields in the large characteristic case. However, we state all our applications in the prime field setting for the sake of simplicity of presentation.
Using the above generalization to large fields as a starting point, we are also able to settle the list decoding radius of fixed degree Reed-Muller codes over growing fields. The case of fixed size fields was solved by Bhowmick and Lovett [STOC 2015], which resolved a conjecture of Gopalan-Klivans-Zuckerman [STOC 2008]. Here, we show that the list decoding radius is equal the minimum distance of the code for all fixed degrees, even when the field size is possibly growing with $n$.
Additionally, we effectively resolve the weight distribution problem for Reed-Muller codes of fixed degree over all fields, first raised in 1977 in the classic textbook by MacWilliams and Sloane [Research Problem 15.1 in Theory of Error Correcting Codes].
Onset of the Asymptotic Regime for Finite Orders
Published
• View Publication
• BIB
We describe a Markov-Chain-Monte-Carlo algorithm which can be used to generate naturally labeled n-element posets at random with a probability distribution of one's choice. Implementing this algorithm for the uniform distribution, we explore the approach to the asymptotic regime in which almost every poset takes on the three-layer structure described by Kleitman and Rothschild (KR). By tracking the n-dependence of several order-invariants, among them the height of the poset, we observe an oscillatory behavior which is very unlike a monotonic approach to the KR regime. Only around n=40 or so does this "finite size dance" appear to give way to a gradual crossover to asymptopia which lasts until n=85, the largest n we have simulated.
The stripping process can be slow: part I
Published
• View Publication
• BIB
Given an integer k, we consider the parallel k-stripping process applied to a hypergraph H: removing all vertices with degree less than k in each iteration until reaching the k-core of H. Take H as H_r(n,m): a random r-uniform hypergraph on n vertices and m hyperedges with the uniform distribution. Fixing k,r\ge 2 with (k,r)\neq (2,2), it has previously been proved that there is a constant c_{r,k} such that for all m=cn with constant c\neq c_{r,k}, with high probability, the parallel k-stripping process takes O(\log n) iterations. In this paper we investigate the critical case when c=c_{r,k}+o(1). We show that the number of iterations that the process takes can go up to some power of n, as long as c approaches c_{r,k} sufficiently fast. A second result we show involves the depth of a non-k-core vertex v: the minimum number of steps required to delete v from H_r(n,m) where in each step one vertex with degree less than k is removed. We will prove lower and upper bounds on the maximum depth over all non-k-core vertices.
Matchings in Benjamini-Schramm convergent graph sequences
Published in Trans. Amer. Math. Soc. 368 (2016), no. 6, 4197--4218
• Search Publication
We introduce the matching measure of a finite graph as the uniform distribution on the roots of the matching polynomial of the graph. We analyze the asymptotic behavior of the matching measure for graph sequences with bounded degree.
A graph parameter is said to be estimable if it converges along every Benjamini-Schramm convergent sparse graph sequence. We prove that the normalized logarithm of the number of matchings is estimable. We also show that the analogous statement for perfect matchings already fails for d-regular bipartite graphs for any fixed d at least 3. The latter result relies on analyzing the probability that a randomly chosen perfect matching contains a particular edge.
However, for any sequence of d-regular bipartite graphs converging to the d-regular tree, we prove that the normalized logarithm of the number of perfect matchings converges. This applies to random d-regular bipartite graphs. We show that the limit equals to the exponent in Schrijver's lower bound on the number of perfect matchings.
Our analytic approach also yields a short proof for the Nguyen-Onak (also Elek--Lippner) theorem saying that the matching ratio is estimable. In fact, we prove the slightly stronger result that the independence ratio is estimable for claw-free graphs.
Property Testing on Product Distributions: Optimal Testers for Bounded Derivative Properties
Published
• View Publication
• BIB
The primary problem in property testing is to decide whether a given function satisfies a certain property, or is far from any function satisfying it. This crucially requires a notion of distance between functions. The most prevalent notion is the Hamming distance over the {\em uniform} distribution on the domain. This restriction to uniformity is more a matter of convenience than of necessity, and it is important to investigate distances induced by more general distributions.
In this paper, we make significant strides in this direction. We give simple and optimal testers for {\em bounded derivative properties} over {\em arbitrary product distributions}. Bounded derivative properties include fundamental properties such as monotonicity and Lipschitz continuity. Our results subsume almost all known results (upper and lower bounds) on monotonicity and Lipschitz testing.
We prove an intimate connection between bounded derivative property testing and binary search trees (BSTs). We exhibit a tester whose query complexity is the sum of expected depths of optimal BSTs for each marginal. Furthermore, we show this sum-of-depths is also a lower bound. A fundamental technical contribution of this work is an {\em optimal dimension reduction theorem} for all bounded derivative properties, which relates the distance of a function from the property to the distance of restrictions of the function to random lines. Such a theorem has been elusive even for monotonicity for the past 15 years, and our theorem is an exponential improvement to the previous best known result.
Probability that n random points in a disk are in convex position
Published
• View Publication
• BIB
In this paper we give a formula for the probability that $n$ random points chosen under the uniform distribution in a disk are in convex position. While close, the formula is recursive and is totally explicit only for the first values of $n$.