gaussian distribution
38 papers tagged with this keyword
Kruskal-Katona for convex sets, with applications
The well-known Kruskal-Katona theorem in combinatorics says that (under mild conditions) every monotone Boolean function $f: \{0,1\}^n \to \{0,1\}$ has a nontrivial "density increment." This means that the fraction of inputs of Hamming weight $k+1$ for which $f=1$ is significantly larger than the fraction of inputs of Hamming weight $k$ for which $f=1.$
We prove an analogous statement for convex sets. Informally, our main result says that (under mild conditions) every convex set $K \subset \mathbb{R}^n$ has a nontrivial density increment. This means that the fraction of the radius-$r$ sphere that lies within $K$ is significantly larger than the fraction of the radius-$r'$ sphere that lies within $K$, for $r'$ suitably larger than $r$. For centrally symmetric convex sets we show that our density increment result is essentially optimal.
As a consequence of our Kruskal-Katona type theorem, we obtain the first efficient weak learning algorithm for convex sets under the Gaussian distribution. We show that any convex set can be weak learned to advantage $Ω(1/n)$ in $\mathsf{poly}(n)$ time under any Gaussian distribution and that any centrally symmetric convex set can be weak learned to advantage $Ω(1/\sqrt{n})$ in $\mathsf{poly}(n)$ time. We also give an information-theoretic lower bound showing that the latter advantage is essentially optimal for $\mathsf{poly}(n)$ time weak learning algorithms. As another consequence of our Kruskal-Katona theorem, we give the first nontrivial Gaussian noise stability bounds for convex sets at high noise rates. Our results extend the known correspondence between monotone Boolean functions over $ \{0,1\}^n$ and convex bodies in Gaussian space.
Gaps of Summands of the Zeckendorf Lattice
Published
• View Publication
• BIB
A beautiful theorem of Zeckendorf states that every positive integer has a unique decomposition as a sum of non-adjacent Fibonacci numbers. Such decompositions exist more generally, and much is known about them. First, for any positive linear recurrence {Gn} the number of summands in the legal decompositions for integers in [Gn, Gn+1) converges to a Gaussian distribution. Second, Bower, Insoft, Li, Miller, and Tosteson proved that the probability of a gap between summands in a decomposition which is larger than the recurrence length converges to geometric decay. While most of the literature involves one-dimensional sequences, some recent work by Chen, Guo, Jiang, Miller, Siktar, and Yu have extended these decompositions to d-dimensional lattices, where a legal decomposition is a chain of points such that one moves in all d dimensions to get from one point to the next. They proved that some but not all properties from 1-dimensional sequences still hold. We continue this work and look at the distribution of gaps between terms of legal decompositions, and prove similar to the 1-dimensional cases that when d = 2 the gap vectors converge to a bivariate geometric random variable.
Asymptotic joint spectra of Cartesian powers of strongly regular graphs and bivariate Charlier-Hermite polynomials
Published in Colloq. Math. 162 (2020) 1-22
• View Publication
• BIB
Generalizing previous work of Hora (1998) on the asymptotic spectral analysis for the Hamming graph $H(n,q)$ which is the $n^{\mathrm{th}}$ Cartesian power $K_q^{\square n}$ of the complete graph $K_q$ on $q$ vertices, we describe the possible limits of the joint spectral distribution of the pair $(G^{\square n},\overline{G}\vphantom{G}^{\square n})$ of the $n^{\mathrm{th}}$ Cartesian powers of a strongly regular graph $G$ and its complement $\overline{G}$, where we let $n\rightarrow\infty$, and $G$ may vary with $n$. This result is an analogue of the bivariate central limit theorem, and we obtain in this way the bivariate Poisson distributions and the standard bivariate Gaussian distribution, together with the product measures of univariate Poisson and Gaussian distributions. We also report a family of bivariate hypergeometric orthogonal polynomials with respect to the last distributions, which we call the bivariate Charlier-Hermite polynomials, and prove basic formulas for them. This family of orthogonal polynomials seems previously unnoticed, possibly because of its peculiarity.
Half-space Macdonald processes
Published in Forum of Mathematics, Pi 8 (2020) e11
• View Publication
• BIB
Macdonald processes are measures on sequences of integer partitions built using the Cauchy summation identity for Macdonald symmetric functions. These measures are a useful tool to uncover the integrability of many probabilistic systems, including the Kardar-Parisi-Zhang (KPZ) equation and a number of other models in its universality class. In this paper we develop the structural theory behind half-space variants of these models and the corresponding half-space Macdonald processes. These processes are built using a Littlewood summation identity instead of the Cauchy identity, and their analysis is considerably harder than their full-space counterparts.
We compute moments and Laplace transforms of observables for general half-space Macdonald measures. Introducing new dynamics preserving this class of measures, we relate them to various stochastic processes, in particular the log-gamma polymer in a half-quadrant (they are also related to the stochastic six-vertex model in a half-quadrant and the half-space ASEP). For the polymer model, we provide explicit integral formulas for the Laplace transform of the partition function. Non-rigorous saddle point asymptotics yield convergence of the directed polymer free energy to either the Tracy-Widom GOE, GSE or the Gaussian distribution depending on the average size of weights on the boundary.
Statistical Analysis of Binary Functional Graphs of the Discrete Logarithm
The increased use of cryptography to protect our personal information makes us want to understand the security of cryptosystems. The security of many cryptosystems relies on solving the discrete logarithm, which is thought to be relatively difficult. Therefore, we focus on the statistical analysis of certain properties of the graph of the discrete logarithm. We discovered the expected value and variance of a certain property of the graph and compared the expected value to experimental data. Our finding did not coincide with our intuition of the data following a Gaussian distribution given a large sample size. Thus, we found the theoretical asymptotic distributions of certain properties of the graph.
On the almost eigenvectors of random regular graphs
Published
• View Publication
• BIB
Let $d\geq 3$ be fixed and $G$ be a large random $d$-regular graph on $n$ vertices. We show that if $n$ is large enough then the entry distribution of every almost eigenvector $v$ of $G$ (with entry sum 0 and normalized to have length $\sqrt{n}$) is close to some Gaussian distribution $N(0,σ)$ in the weak topology where $0\leqσ\leq 1$. Our theorem holds even in the stronger sense when many entries are looked at simultaneously in small random neighborhoods of the graph. Furthermore, we also get the Gaussianity of the joint distribution of several almost eigenvectors if the corresponding eigenvalues are close. Our proof uses graph limits and information theory. Our results have consequences for factor of i.i.d.\ processes on the infinite regular tree.
A family of sequences of binomial type
Published in Probability and Mathematical Statistics, (2013) 33.2, 401-408
• Search Publication
For delta operator $aD-bD^{p+1}$ we find the corresponding polynomial sequence of binomial type and relations with Fuss numbers. In the case $D-\frac{1}{2}D^2$ we show that the corresponding Bessel-Carlitz polynomials are moments of the convolution semigroup of inverse Gaussian distributions. We also find probability distributions $ν_{t}$, $t>0$, for which $\left\{y_{n}(t)\right\}$, the Bessel polynomials at $t$, is the moment sequence.
On the growth of permutation classes
We study aspects of the enumeration of permutation classes, sets of permutations closed downwards under the subpermutation order.
First, we consider monotone grid classes of permutations. We present procedures for calculating the generating function of any class whose matrix has dimensions $m \times 1$ for some $m$, and of acyclic and unicyclic classes of gridded permutations. We show that almost all large permutations in a grid class have the same shape, and determine this limit shape.
We prove that the growth rate of a grid class is given by the square of the spectral radius of an associated graph and deduce some facts relating to the set of grid class growth rates. In the process, we establish a new result concerning tours on graphs. We also prove a similar result relating the growth rate of a geometric grid class to the matching polynomial of a graph, and determine the effect of edge subdivision on the matching polynomial. We characterise the growth rates of geometric grid classes in terms of the spectral radii of trees.
We then investigate the set of growth rates of permutation classes and establish a new upper bound on the value above which every real number is the growth rate of some permutation class. In the process, we prove new results concerning expansions of real numbers in non-integer bases in which the digits are drawn from sets of allowed values.
Finally, we introduce a new enumeration technique, based on associating a graph with each permutation, and determine the generating functions for some previously unenumerated classes. We conclude by using this approach to provide an improved lower bound on the growth rate of the class of permutations avoiding the pattern $1324$. In the process, we prove that, asymptotically, patterns in Łukasiewicz paths exhibit a concentrated Gaussian distribution.
Tensor models from the viewpoint of matrix models: the case of the Gaussian distribution
Observables in random tensor theory are polynomials in the entries of a tensor of rank $d$ which are invariant under $U(N)^d$. It is notoriously difficult to evaluate the expectations of such polynomials, even in the Gaussian distribution. In this article, we introduce singular value decompositions to evaluate the expectations of polynomial observables of Gaussian random tensors. Performing the matrix integrals over the unitary group leads to a notion of effective observables which expand onto regular, matrix trace invariants. Examples are given to illustrate that both asymptotic and exact new calculations of expectations can be performed this way.
Permutations avoiding 1324 and patterns in Łukasiewicz paths
Published in J. London Math. Soc., 92(1):105-122, 2015
• View Publication
• BIB
The class Av(1324), of permutations avoiding the pattern 1324, is one of the simplest sets of combinatorial objects to define that has, thus far, failed to reveal its enumerative secrets. By considering certain large subsets of the class, which consist of permutations with a particularly regular structure, we prove that the growth rate of the class exceeds 9.81. This improves on a previous lower bound of 9.47. Central to our proof is an examination of the asymptotic distributions of certain substructures in the Hasse graphs of the permutations. In this context, we consider occurrences of patterns in Łukasiewicz paths and prove that in the limit they exhibit a concentrated Gaussian distribution.
The Distribution of Gaps between Summands in Generalized Zeckendorf Decompositions
Published
• View Publication
• BIB
Zeckendorf proved that any integer can be decomposed uniquely as a sum of non-adjacent Fibonacci numbers, $F_n$. Using continued fractions, Lekkerkerker proved the average number of summands of an $m \in [F_n, F_{n+1})$ is essentially $n/(\varphi^2 +1)$, with $\varphi$ the golden ratio. Miller-Wang generalized this by adopting a combinatorial perspective, proving that for any positive linear recurrence the number of summands in decompositions for integers in $[G_n, G_{n+1})$ converges to a Gaussian distribution. We prove the probability of a gap larger than the recurrence length converges to decaying geometrically, and that the distribution of the smaller gaps depends in a computable way on the coefficients of the recurrence. These results hold both for the average over all $m \in [G_n, G_{n+1})$, as well as holding almost surely for the gap measure associated to individual $m$. The techniques can also be used to determine the distribution of the longest gap between summands, which we prove is similar to the distribution of the longest gap between heads in tosses of a biased coin. It is a double exponential strongly concentrated about the mean, and is on the order of $\log n$ with computable constants depending on the recurrence.
The calculation of expectation values in Gaussian random tensor theory via meanders
Published
• View Publication
• BIB
A difficult problem in the theory of random tensors is to calculate the expectation values of polynomials in the tensor entries, even in the large N limit and in a Gaussian distribution. Here we address this issue, focusing on a family of polynomials labeled by permutations, which naturally generalize the single-trace invariants of random matrix models. Through Wick's theorem, we show that the Feynman graph expansion of the expectation values of those polynomials enumerates meandric systems whose lower arch configuration is obtained from the upper arch configuration by a permutation on half of the arch feet. Our main theorem reduces the calculation of expectation values to those of polynomials labeled by stabilized-interval-free permutations (SIF) which are proved to enumerate irreducible meandric systems. This together with explicit calculations of expectation values associated to SIF permutations allows to exactly evaluate large N expectation values beyond the so-called melonic polynomials.
A Note on Discrete Gaussian Combinations of Lattice Vectors
We analyze the distribution of $\sum_{i=1}^m v_i \bx_i$ where $\bx_1,...,\bx_m$ are fixed vectors from some lattice $\cL \subset \R^n$ (say $\Z^n$) and $v_1,...,v_m$ are chosen independently from a discrete Gaussian distribution over $\Z$. We show that under a natural constraint on $\bx_1,...,\bx_m$, if the $v_i$ are chosen from a wide enough Gaussian, the sum is statistically close to a discrete Gaussian over $\cL$. We also analyze the case of $\bx_1,...,\bx_m$ that are themselves chosen from a discrete Gaussian distribution (and fixed).
Our results simplify and qualitatively improve upon a recent result by Agrawal, Gentry, Halevi, and Sahai \cite{AGHS13}.
A Gaussian distribution for refined DT invariants and 3D partitions
Published
• View Publication
• BIB
We show that the refined Donaldson-Thomas invariants of C3, suitably normalized, have a Gaussian distribution as limit law. Combinatorially these numbers are given by weighted counts of 3D partitions. Our technique is to use the Hardy-Littlewood circle method to analyze the bivariate asymptotics of a q-deformation of MacMahon's function. The proof is based on that of E.M. Wright who explored the single variable case.
Large Deviations for the Empirical Distribution in the Branching Random Walk
Published
• View Publication
• BIB
We consider the branching random walk on the real line where the underlying motion is of a simple random walk and branching is at least binary and at most decaying exponentially in law. It is well known that the normalized empirical measure converges to the Gaussian distribution for typical sets A. We therefore analyze the probability that at step n the empirical distribution differs from the Gaussian distribution by a constant ε. We show that the decay is doubly exponential in either n or \sqrt{n}, depending on the set A and ε, and we find the leading coefficient in the top exponent. To the best of our knowledge, this is the first time such large deviation probabilities are treated in this model.
Circuit partitions and #P-complete products of inner products
We present a simple, natural #P-complete problem. Let G be a directed graph, and let k be a positive integer. We define q(G;k) as follows. At each vertex v, we place a k-dimensional complex vector x_v. We take the product, over all edges (u,v), of the inner product <x_u,x_v>. Finally, q(G;k) is the expectation of this product, where the x_v are chosen uniformly and independently from all vectors of norm 1 (or, alternately, from the Gaussian distribution). We show that q(G;k) is proportional to G's cycle partition polynomial, and therefore that it is #P-complete for any k>1.
Isomorphism and Symmetries in Random Phylogenetic Trees
The probability that two randomly selected phylogenetic trees of the same size are isomorphic is found to be asymptotic to a decreasing exponential modulated by a polynomial factor. The number of symmetrical nodes in a random phylogenetic tree of large size obeys a limiting Gaussian distribution, in the sense of both central and local limits. The probability that two random phylogenetic trees have the same number of symmetries asymptotically obeys an inverse square-root law. Precise estimates for these problems are obtained by methods of analytic combinatorics, involving bivariate generating functions, singularity analysis, and quasi-powers approximations.
Gaussian Bounds for Noise Correlation of Functions
Published
• View Publication
• BIB
In this paper we derive tight bounds on the expected value of products of {\em low influence} functions defined on correlated probability spaces. The proofs are based on extending Fourier theory to an arbitrary number of correlated probability spaces, on a generalization of an invariance principle recently obtained with O'Donnell and Oleszkiewicz for multilinear polynomials with low influences and bounded degree and on properties of multi-dimensional Gaussian distributions. The results derived here have a number of applications to the theory of social choice in economics, to hardness of approximation in computer science and to additive combinatorics problems.