arXiv++ Combinatorics

Browse math.CO papers from arXiv

Papers by Van Vu

45 paper(s) by this author · All BibTeX
2026-01-07
Anti-concentration with respect to random permutations
Classical anti-concentration results focus on the random sum $S := \sum _{i=1}^n ξ_i v_i$, where $ξ_i$ are independent random variables and $v_i$ are real numbers. In this paper, we prove new concentration results concerning the random sum $S := \sum_{i=1}^n w_{π_i } v_i $, where $w_i , v_i$ are real numbers and $π$ is a random permutation.
2025-01-31 v2
Fast exact recovery of noisy matrix from few entries: the infinity norm approach
The matrix recovery (completion) problem, a central problem in data science and theoretical computer science, is to recover a matrix $A$ from a relatively small sample of entries. While such a task is impossible in general, it has been shown that one can recover $A$ exactly in polynomial time, with high probability, from a random subset of entries, under three (basic and necessary) assumptions: (1) the rank of $A$ is very small compared to its dimensions (low rank), (2) $A$ has delocalized singular vectors (incoherence), and (3) the sample size is sufficiently large. There are many different algorithms for the task, including convex optimization by Candes, Tao and Recht (2009), alternating projection by Hardt and Wooters (2014) and low rank approximation with gradient descent by Keshavan, Montanari and Oh (2009, 2010). In applications, it is more realistic to assume that data is noisy. In this case, these approaches provide an approximate recovery with small root mean square error. However, it is hard to transform such an approximate recovery to an exact one. Recently, results by Abbe et al. (2017) and Bhardwaj et al. (2023) concerning approximation in the infinity norm showed that we can achieve exact recovery even in the noisy case, given that the ground matrix has bounded precision. Beyond the three basic assumptions above, they required either the condition number of $A$ is small (Abbe et al.) or the gap between consecutive singular values is large (Bhardwaj et al.). In this paper, we remove these extra spectral assumptions. As a result, we obtain a simple algorithm for exact recovery in the noisy case, under only the three basic assumptions. This is the first such algorithm. To analyse this algorithm, we introduce a contour integration argument which is totally different from all previous methods and may be of independent interest.
2024-09-30 v6
New matrix perturbation bounds with relative norm: Perturbation of eigenspaces
Matrix perturbation bounds (such as Weyl and Davis-Kahan) are used abundantly in many areas of mathematics and data science. Many bounds (such as the above two) involve the spectral norm of the noise matrix and are sharp in worst case analysis. In order to refine these classical bounds, we introduce a new parameter, which we refer to as the relative norm. This parameter measures the strength of the action of the noise matrix on the relevant eigenvectors of the ground matrix. It has turned out that in a number of situations, we can use the relative norm as a replacement for the spectral norm. This has led to a number of notable improvements under certain sets of assumptions, which are frequently met in practice. For instance, our new results apply very well in the case when the noise matrix is random. For the purpose of our study, we introduce a new method of analysis, which combines the classical contour integral argument with new (combinatorial) ideas. This method is robust and of independent interest. In the current paper, we focus on the perturbation of eigenspaces (Davis-Kahan type results). Perturbation bounds for eigenspaces are essential in statistics and theoretical computer science, and thus deserve a special treatment. Furthermore, this will lay the ground for the more technical treatment of general matrix functionals, which appears in a future paper.
2023-04-01 v3
Matrix Perturbation: Davis-Kahan in the Infinity Norm
Perturbation theory is developed to analyze the impact of noise on data and has been an essential part of numerical analysis. Recently, it has played an important role in designing and analyzing matrix algorithms. One of the most useful tools in this subject, the Davis-Kahan sine theorem, provides an $\ell_2$ error bound on the perturbation of the leading singular vectors (and spaces). We focus on the case when the signal matrix has low rank and the perturbation is random, which occurs often in practice. In an earlier paper, O'Rourke, Wang, and the second author showed that in this case, one can obtain an improved theorem. In particular, the noise-to-gap ratio condition in the original setting can be weakened considerably. In the current paper, we develop an infinity norm version of the O'Rourke-Vu-Wang result. The key ideas in the proof are a new bootstrapping argument and the so-called iterative leave-one-out method, which may be of independent interest. Applying the new bounds, we develop new, simple, and quick algorithms for several well-known problems, such as finding hidden partitions and matrix completion. The core of these new algorithms is the fact that one is now able to quickly approximate certain key objects in the infinity norm, which has critical advantages over approximations in the $\ell_2$ norm, Frobenius norm, or spectral norm.
2023-02-11 v2
The "Power of Few" Phenomenon: The Sparse Case
Published in Random Structures & Algorithms. 66 (2024) • View PublicationBIB
The "majority dynamics" process on a social network begins with an initial phase, where the individuals are split into two competing parties, Red and Blue. Every day, everyone updates their affiliation to match the majority among those of their friends. While studying this process on Erdos-Renyi G(n, p) random graph (with constant density), the authors discovered the "Power of Few" phenomenon, showing that a very small advantage to one side already guarantees that everybody will unanimously join that side after just a few days with overwhelming probability. For example, when p = 1/2, then 10 extra members guarantee this unanimity with a 90% chance, regardless of the value of n. In this paper, we study this phenomenon for sparse random graphs. It is clear that below the connectivity threshold, the phenomenon ceases to hold, as the isolated vertices never change their colors. We show that it holds for every density above the threshold. To make the process more realistic, we also assume that individuals can randomly activate their accounts to post their opinions and observe their neighbors (just as we do on social media). We prove that the phenomenon is robust under this assumption.
2020-05-06
Recent progress in combinatorial random matrix theory
Published • View PublicationBIB
We discuss recent progress many problems in random matrix theory of a combinatorial nature, including several breakthroughs that solve long standing famous conjectures.
2019-11-22 v2
Reaching a Consensus on Random Networks: The Power of Few
A community of $n$ individuals splits into two camps, Red and Blue. The individuals are connected by a social network, which influences their colors. Everyday, each person changes his/her color according to the majority among his/her neighbors. Red (Blue) wins if everyone in the community becomes Red (Blue) at some point. We study this process when the underlying network is the random Erdos-Renyi graph $G(n, p)$. With a balanced initial state ($n/2$ person in each camp), it is clear that each color wins with the same probability. Our study reveals that for any constants $p$ and $\varepsilon$, there is a constant $C$ such that if one camp has $n/2 +C$ individuals, then it wins with probability at least $1 - \varepsilon$. The surprising key fact here is that $C$ does not depend on $n$, the population of the community. When $p=1/2$ and $\varepsilon =.1$, one can set $C$ as small as 6. If the aim of the process is to choose a candidate, then this means it takes only $6$ "defectors" to win an election unanimously with overwhelming odd.
2018-09-14 v2
Spectrum of complex networks
Published • View PublicationBIB
The study of complex networks has been one of the most active fields in science in recent decades. Spectral properties of networks (or graphs that represent them) are of fundamental importance. Researchers have been investigating these properties for many years, and, based on numerical data, have raised a number of questions about the distribution of the eigenvalues and eigenvectors. In this paper, we give the solution to some of these questions. In particular, we determine the limiting distribution of (the bulk of) the spectrum as the size of the network grows to infinity and show that the leading eigenvectors are strongly localized. We focus on the preferential attachment graph, which is the most popular mathematical model for growing complex networks. Our analysis is, on the other hand, general and can be applied to other models.
2018-02-10 v2
Sparse Random Matrices have Simple Spectrum
Published • View PublicationBIB
Let $M_n$ be a class of symmetric sparse random matrices, with independent entries $M_{ij} = δ_{ij} ξ_{ij}$ for $i \leq j$. $δ_{ij}$ are i.i.d. Bernoulli random variables taking the value $1$ with probability $p \geq n^{-1+δ}$ for any constant $δ> 0$ and $ξ_{ij}$ are i.i.d. centered, subgaussian random variables. We show that with high probability this class of random matrices has simple spectrum (i.e. the eigenvalues appear with multiplicity one). We can slightly modify our proof to show that the adjacency matrix of a sparse Erdős-Rényi graph has simple spectrum for $n^{-1+δ} \leq p \leq 1- n^{-1+δ}$. These results are optimal in the exponent. The result for graphs has connections to the notorious graph isomorphism problem.
2017-11-08 v2
Random matrices: Probability of Normality
Published • View PublicationBIB
In this paper, we investigate the following question: How often is a random matrix normal? We consider a random $n\times n$ matrix, $M_n$, whose entries are i.i.d. Rademacher random variables (taking values $\{ \pm1 \}$ with probability $1/2$) and prove $$2^{-\left(0.5+o(1)\right)n^2} \le P\left(M_n \text{ is normal}\right) \le 2^{-(0.302+o(1))n^{2}}. $$ We conjecture that the lower bound is sharp.
2017-07-17
Random eigenfunctions on flat tori: universality for the number of intersections
Published • View PublicationBIB
We show that several statistics of the number of intersections between random eigenfunctions of general eigenvalues with a given smooth curve in flat tori are universal under various families of randomness.
2016-07-29 v3
Law of Iterated Logarithm for random graphs
Published • View PublicationBIB
A milestone in Probability Theory is the law of the iterated logarithm (LIL), proved by Khinchin and independently by Kolmogorov in the 1920s, which asserts that for iid random variables $\{t_i\}_{i=1}^{\infty}$ with mean $0$ and variance $1$ $$ \Pr \left[ \limsup_{n\rightarrow \infty} \frac{ \sum_{i=1}^n t_i }{σ_n \sqrt {2 \log \log n }} =1 \right] =1 . $$ In this paper we prove that LIL holds for various functionals of random graphs and hypergraphs models. We first prove LIL for the number of copies of a fixed subgraph $H$. Two harder results concern the number of global objects: perfect matchings and Hamiltonian cycles. The main new ingredient in these results is a large deviation bound, which may be of independent interest. For random $k$-uniform hypergraphs, we obtain the Central Limit Theorem (CLT) and LIL for the number of Hamilton cycles.
2016-06-30 v2
Packing perfect matchings in random hypergraphs
Published • View PublicationBIB
We introduce a new procedure for generating the binomial random graph/hypergraph models, referred to as \emph{online sprinkling}. As an illustrative application of this method, we show that for any fixed integer $k\geq 3$, the binomial $k$-uniform random hypergraph $H^{k}_{n,p}$ contains $N:=(1-o(1))\binom{n-1}{k-1}p$ edge-disjoint perfect matchings, provided $p\geq \frac{\log^{C}n}{n^{k-1}}$, where $C:=C(k)$ is an integer depending only on $k$. Our result for $N$ is asymptotically best optimal and for $p$ is optimal up to the $polylog(n)$ factor.
2016-03-09 v5
Sum-avoiding sets in groups
Published in Discrete Analysis 2016:15, 31 pp • View PublicationBIB
Let $A$ be a finite subset of an arbitrary additive group $G$, and let $φ(A)$ denote the cardinality of the largest subset $B$ in $A$ that is sum-avoiding in $A$ (that is to say, $b_1+b_2 \not \in A$ for all distinct $b_1,b_2 \in B$). The question of controlling the size of $A$ in terms of $φ(A)$ in the case when $G$ was torsion-free was posed by Erdős and Moser. When $G$ has torsion, $A$ can be arbitrarily large for fixed $φ(A)$ due to the presence of subgroups. Nevertheless, we provide a qualitative answer to an analogue of the Erdős-Moser problem in this setting, by establishing a structure theorem, which roughly speaking asserts that $A$ is either efficiently covered by $φ(A)$ finite subgroups of $G$, or by fewer than $φ(A)$ finite subgroups of $G$ together with a residual set of bounded cardinality. In order to avoid a large number of nested inductive arguments, our proof uses the language of nonstandard analysis. We also answer negatively a question of Erdős regarding large subsets $A$ of finite additive groups $G$ with $φ(A)$ bounded, but give a positive result when $|G|$ is not divisible by small primes.
2016-03-09 v2
Sumfree sets in groups: a survey
Published • View PublicationBIB
We discuss several questions concerning sum-free sets in groups, raised by Erdős in his survey "Extremal problems in number theory" (Proceedings of the Symp. Pure Math. VIII AMS) published in 1965. Among other things, we give a characterization for large sets $A$ in an abelian group $G$ which do not contain a subset $B$ of fixed size $k$ such that the sum of any two different elements of $B$ do not belong to $A$ (in other words, $B$ is sum-free with respect to $A$). Erdős, in the above mentioned survey, conjectured that if $|A|$ is sufficiently large compared to $k$, then $A$ contains two elements that add up to zero. This is known to be true for $k \leq 3$. We give counterexamples for all $k \ge 4$. On the other hand, using the new characterization result, we are able to prove a positive result in the case when $|G|$ is not divisible by small primes.
2016-01-14 v3
Eigenvectors of random matrices: A survey
Eigenvectors of large matrices (and graphs) play an essential role in combinatorics and theoretical computer science. The goal of this survey is to provide an up-to-date account on properties of eigenvectors when the matrix (or graph) is random.
2015-04-01 v3
Random matrices: tail bounds for gaps between eigenvalues
Published • View PublicationBIB
Gaps (or spacings) between consecutive eigenvalues are a central topic in random matrix theory. The goal of this paper is to study the tail distribution of these gaps in various random matrix models. We give the first repulsion bound for random matrices with discrete entries and the first super-polynomial bound on the probability that a random graph has simple spectrum, along with several applications.
2014-12-03
Random matrices have simple spectrum
Published • View PublicationBIB
Let $M_n = (ξ_{ij})_{1 \leq i,j \leq n}$ be a real symmetric random matrix in which the upper-triangular entries $ξ_{ij}, i<j$ and diagonal entries $ξ_{ii}$ are independent. We show that with probability tending to 1, $M_n$ has no repeated eigenvalues. As a corollary, we deduce that the Erd{\H o}s-Renyi random graph has simple spectrum asymptotically almost surely, answering a question of Babai.
2014-09-29
Random walks with different directions: Drunkards beware !
Published • View PublicationBIB
As an extension of Polya's classical result on random walks on the square grids ($\Z^d$), we consider a random walk where the steps, while still have unit length, point to different directions. We show that in dimensions at least 4, the returning probability after $n$ steps is at most $n^{-d/2 - d/(d-2) +o(1)}$, which is sharp. The real surprise is in dimensions 2 and 3. In dimension 2, where the traditional grid walk is recurrent, our upper bound is $n^{-ω(1)}$, which is much worse than higher dimensions. In dimension 3, we prove an upper bound of order $n^{-4 +o(1)}$. We discover a new conjecture concerning incidences between spheres and points in $\R^3$, which, if holds, would improve the bound to $n^{-9/2 +o(1)}$, which is consistent % with the $d \ge 4$ case. to the $d \ge 4$ case. This conjecture resembles Szemerédi-Trotter type results and is of independent interest.
2014-09-15
Real roots of random polynomials: expectation and repulsion
Published • View PublicationBIB
Let $P_{n}(x)= \sum_{i=0}^n ξ_i x^i$ be a Kac random polynomial where the coefficients $ξ_i$ are iid copies of a given random variable $ξ$. Our main result is an optimal quantitative bound concerning real roots repulsion. This leads to an optimal bound on the probability that there is a double root. As an application, we consider the problem of estimating the number of real roots of $P_n$, which has a long history and in particular was the main subject of a celebrated series of papers by Littlewood and Offord from the 1940s. We show, for a large and natural family of atom variables $ξ$, that the expected number of real roots of $P_n(x)$ is exactly $\frac{2}π \log n +C +o(1)$, where $C$ is an absolute constant depending on the atom variable $ξ$. Prior to this paper, such a result was known only for the case when $ξ$ is Gaussian.