Papers by Vishesh Jain
37 paper(s) by this author
· All BibTeX
Singularity of discrete random matrices
Published
• View Publication
• BIB
Let $ξ$ be a non-constant real-valued random variable with finite support, and let $M_{n}(ξ)$ denote an $n\times n$ random matrix with entries that are independent copies of $ξ$. For $ξ$ which is not uniform on its support, we show that \begin{align*} \mathbb{P}[M_{n}(ξ)\text{ is singular}] &= \mathbb{P}[\text{zero row or column}] + (1+o_n(1))\mathbb{P}[\text{two equal (up to sign) rows or columns}], \end{align*} thereby confirming a folklore conjecture. As special cases, we obtain:
(1) For $ξ= \text{Bernoulli}(p)$ with fixed $p \in (0,1/2)$, \[\mathbb{P}[M_{n}(ξ)\text{ is singular}] = 2n(1-p)^{n} + (1+o_n(1))n(n-1)(p^2 + (1-p)^2)^{n},\] which determines the singularity probability to two asymptotic terms. Previously, no result of such precision was available in the study of the singularity of random matrices.
(2) For $ξ= \text{Bernoulli}(p)$ with fixed $p \in (1/2,1)$, \[\mathbb{P}[M_{n}(ξ)\text{ is singular}] = (1+o_n(1))n(n-1)(p^2 + (1-p)^2)^{n}.\] Previously, only the much weaker upper bound of $(\sqrt{p} + o_n(1))^{n}$ was known due to the work of Bourgain-Vu-Wood.
For $ξ$ which is uniform on its support:
(1) We show that \begin{align*} \mathbb{P}[M_{n}(ξ)\text{ is singular}] &= (1+o_n(1))^{n}\mathbb{P}[\text{two rows or columns are equal}]. \end{align*} (2) Perhaps more importantly, we provide a sharp analysis of the contribution of the `compressible' part of the unit sphere to the lower tail of the smallest singular value of $M_{n}(ξ)$.
The smallest singular value of dense random regular digraphs
Published
• View Publication
• BIB
Let $A$ be the adjacency matrix of a uniformly random $d$-regular digraph on $n$ vertices, and suppose that $\min(d,n-d)\geqλn$. We show that for any $κ\geq 0$, \[\mathbb{P}[s_n(A)\leqκ]\leq C_λκ\sqrt{n}+2e^{-c_λn}.\] Up to the constants $C_λ, c_λ> 0$, our bound matches optimal bounds for $n\times n$ random matrices, each of whose entries is an i.i.d $\text{Ber}(d/n)$ random variable. The special case $κ= 0$ of our result confirms a conjecture of Cook regarding the probability of singularity of dense random regular digraphs.
Perfectly Sampling $k\geq (8/3 +o(1))Δ$-Colorings in Graphs
Published
• View Publication
• BIB
We present a randomized algorithm which takes as input an undirected graph $G$ on $n$ vertices with maximum degree $Δ$, and a number of colors $k \geq (8/3 + o_Δ(1))Δ$, and returns -- in expected time $\tilde{O}(nΔ^{2}\log{k})$ -- a proper $k$-coloring of $G$ distributed perfectly uniformly on the set of all proper $k$-colorings of $G$. Notably, our sampler breaks the barrier at $k = 3Δ$ encountered in recent work of Bhandari and Chakraborty [STOC 2020]. We also sketch how to modify our methods to relax the restriction on $k$ to $k \geq (8/3 - ε_0)Δ$ for an absolute constant $ε_0 > 0$.
As in the work of Bhandari and Chakraborty, and the pioneering work of Huber [STOC 1998], our sampler is based on Coupling from the Past [Propp&Wilson, Random Struct. Algorithms, 1995] and the bounding chain method [Huber, STOC 1998; Häggström&Nelander, Scand. J. Statist., 1999]. Our innovations include a novel bounding chain routine inspired by Jerrum's analysis of the Glauber dynamics [Random Struct. Algorithms, 1995], as well as a preconditioning routine for bounding chains which uses the algorithmic Lovász Local Lemma [Moser&Tardos, J.ACM, 2010].
The probability of selecting $k$ edge-disjoint Hamilton cycles in the complete graph
Let $H_1,\dots,H_k$ be Hamilton cycles in $K_n$, chosen independently and uniformly at random. We show, for $k = o(n^{1/100})$, that the probability of $H_1,\dots,H_k$ being edge-disjoint is $(1+o(1))e^{-2\binom{k}{2}}$. This extends a corresponding estimate obtained by Robbins in the case $k=2$.
Quantitative invertibility of random matrices: a combinatorial perspective
We study the lower tail behavior of the least singular value of an $n\times n$ random matrix $M_n := M+N_n$, where $M$ is a fixed complex matrix with operator norm at most $\exp(n^{c})$ and $N_n$ is a random matrix, each of whose entries is an independent copy of a complex random variable with mean $0$ and variance $1$. Motivated by applications, our focus is on obtaining bounds which hold with extremely high probability, rather than on the least singular value of a typical such matrix.
This setting has previously been considered in a series of influential works by Tao and Vu, most notably in connection with the strong circular law, and the smoothed analysis of the condition number, and our results improve upon theirs in two ways:
(i) We are able to handle $\|M\| = O(\exp(n^{c}))$, whereas the results of Tao and Vu are applicable only for $M = O(\text{poly(n)})$.
(ii) Even for $M = O(\text{poly(n)})$, we are able to extract more refined information -- for instance, our results show that for such $M$, the probability that $M_n$ is singular is $O(\exp(-n^{c}))$, whereas even in the case when $ξ$ is a Bernoulli random variable, the results of Tao and Vu only give a bound of the form $O_{C}(n^{-C})$ for any constant $C>0$.
As opposed to all previous works obtaining such bounds with error rate better than $n^{-1}$, our proof makes no use either of the inverse Littlewood--Offord theorems, or of any sophisticated net constructions. Instead, we show how to reduce the problem from the (complex) sphere to (Gaussian) integer vectors, where it is solved directly by utilizing and extending a combinatorial approach to the singularity problem for random discrete matrices, recently developed by Ferber, Luh, Samotij, and the author.
The strong circular law: a combinatorial view
Let $N_n$ be an $n\times n$ complex random matrix, each of whose entries is an independent copy of a centered complex random variable $z$ with finite non-zero variance $σ^{2}$. The strong circular law, proved by Tao and Vu, states that almost surely, as $n\to \infty$, the empirical spectral distribution of $N_n/(σ\sqrt{n})$ converges to the uniform distribution on the unit disc in $\mathbb{C}$.
A crucial ingredient in the proof of Tao and Vu, which uses deep ideas from additive combinatorics, is controlling the lower tail of the least singular value of the random matrix $xI - N_{n}/(σ\sqrt{n})$ (where $x\in \mathbb{C}$ is fixed) with failure probability that is inverse polynomial. In this paper, using a simple and novel approach (in particular, not using tools from additive combinatorics or any net arguments), we show that for any fixed matrix $M$ with operator norm at most $n^{0.51}$ and for all $η\geq 0$, $$\Pr\left(s_n(M+N_n) \leq η\right) \lesssim n^{C}η+ \exp(-n^{c}),$$ where $s_n(M+N_n)$ is the least singular value of $M+N_n$ and $C,c$ are absolute constants. Our result is optimal up to the constants $C,c$ and the inverse exponential-type error rate improves upon the inverse polynomial error rate due to Tao and Vu.
During the course of our proof, we extend the solution of the counting problem in inverse Littlewood-Offord theory, recently isolated by the author along with Ferber, Luh, and Samotij, from Rademacher variables to general complex random variables. This significantly improves on estimates for this problem obtained using the optimal inverse Littlewood-Offord theorem of Nguyen and Vu, and may be of independent interest.
Approximate Spielman-Teng theorems for the least singular value of random combinatorial matrices
An approximate Spielman-Teng theorem for the least singular value $s_n(M_n)$ of a random $n\times n$ square matrix $M_n$ is a statement of the following form: there exist constants $C,c >0$ such that for all $η\geq 0$, $\Pr(s_n(M_n) \leq η) \lesssim n^{C}η+ \exp(-n^{c})$. The goal of this paper is to develop a simple and novel framework for proving such results for discrete random matrices. As an application, we prove an approximate Spielman-Teng theorem for $\{0,1\}$-valued matrices, each of whose rows is an independent vector with exactly $n/2$ zero components. This improves on previous work of Nguyen and Vu, and is the first such result in a `truly combinatorial' setting.
On the counting problem in inverse Littlewood--Offord theory
Let $ε_1, \dotsc, ε_n$ be i.i.d. Rademacher random variables taking values $\pm 1$ with probability $1/2$ each. Given an integer vector $\boldsymbol{a} = (a_1, \dotsc, a_n)$, its concentration probability is the quantity $ρ(\boldsymbol{a}):=\sup_{x\in \mathbb{Z}}\Pr(ε_1 a_1+\dots+ε_n a_n = x)$. The Littlewood-Offord problem asks for bounds on $ρ(\boldsymbol{a})$ under various hypotheses on $\boldsymbol{a}$, whereas the inverse Littlewood-Offord problem, posed by Tao and Vu, asks for a characterization of all vectors $\boldsymbol{a}$ for which $ρ(\boldsymbol{a})$ is large. In this paper, we study the associated counting problem: How many integer vectors $\boldsymbol{a}$ belonging to a specified set have large $ρ(\boldsymbol{a})$? The motivation for our study is that in typical applications, the inverse Littlewood-Offord theorems are only used to obtain such counting estimates. Using a more direct approach, we obtain significantly better bounds for this problem than those obtained using the inverse Littlewood--Offord theorems of Tao and Vu and of Nguyen and Vu. Moreover, we develop a framework for deriving upper bounds on the probability of singularity of random discrete matrices that utilizes our counting result. To illustrate the methods, we present the first `exponential-type' (i.e., $\exp(-n^c)$ for some positive constant $c$) upper bounds on the singularity probability for the following two models: (i) adjacency matrices of dense signed random regular digraphs, for which the previous best known bound is $O(n^{-1/4})$ due to Cook; and (ii) dense row-regular $\{0,1\}$-matrices, for which the previous best known bound is $O_{C}(n^{-C})$ for any constant $C>0$ due to Nguyen.
Uniformity-independent minimum degree conditions for perfect matchings in hypergraphs
In this note, we prove that there exists a universal constant $c=\frac{43}{50}$ such that for every $k\in \mathbb{N}$ and every $d<k/2$, every $k$-uniform hypergraph on $n$ vertices and with minimum $d$-degree at least $(c+o_n(1))\binom{n-d}{k-d}$ contains a perfect matching. This is the first such bound which is independent of $k$, and therefore, improves all previously known bounds when $k$ is large. Our approach is based on combining the seminal work of Alon et al. with known bounds on a conjectured probabilistic inequality due to Feige.
Towards the linear arboricity conjecture
Published
• View Publication
• BIB
The linear arboricity of a graph $G$, denoted by $\text{la}(G)$, is the minimum number of edge-disjoint linear forests (i.e. forests in which every connected component is a path) in $G$ whose union covers all the edges of $G$. A famous conjecture due to Akiyama, Exoo, and Harary from 1980 asserts that $\text{la}(G)\leq \lceil (Δ(G)+1)/2 \rceil$, where $Δ(G)$ denotes the maximum degree of $G$. This conjectured upper bound would be best possible, as is easily seen by taking $G$ to be a regular graph. In this paper, we show that for every graph $G$, $\text{la}(G)\leq \fracΔ{2}+O(Δ^{2/3-α})$ for some $α> 0$, thereby improving the previously best known bound due to Alon and Spencer from 1992. For graphs which are sufficiently good spectral expanders, we give even better bounds. Our proofs of these results further give probabilistic polynomial time algorithms for finding such decompositions into linear forests.
Singularity of random symmetric matrices -- a combinatorial approach to improved bounds
Published in Forum of Mathematics, Sigma, vol. 7, e22, 29 pages (2019)
• View Publication
• BIB
Let $M_n$ denote a random symmetric $n \times n$ matrix whose upper diagonal entries are independent and identically distributed Bernoulli random variables (which take values $1$ and $-1$ with probability $1/2$ each). It is widely conjectured that $M_n$ is singular with probability at most $(2+o(1))^{-n}$. On the other hand, the best known upper bound on the singularity probability of $M_n$, due to Vershynin (2011), is $2^{-n^c}$, for some unspecified small constant $c > 0$. This improves on a polynomial singularity bound due to Costello, Tao, and Vu (2005), and a bound of Nguyen (2011) showing that the singularity probability decays faster than any polynomial. In this paper, improving on all previous results, we show that the probability of singularity of $M_n$ is at most $2^{-n^{1/4}\sqrt{\log{n}}/1000}$ for all sufficiently large $n$. The proof utilizes and extends a novel combinatorial approach to discrete random matrix theory, which has been recently introduced by the authors together with Luh and Samotij.
On the number of Hadamard matrices via anti-concentration
Published
• View Publication
• BIB
Many problems in combinatorial linear algebra require upper bounds on the number of solutions to an underdetermined system of linear equations $Ax = b$, where the coordinates of the vector $x$ are restricted to take values in some small subset (e.g. $\{\pm 1\}$) of the underlying field. The classical ways of bounding this quantity are to use either a rank bound observation due to Odlyzko or a vector anti-concentration inequality due to Halász. The former gives a stronger conclusion except when the number of equations is significantly smaller than the number of variables; even in such situations, the hypotheses of Halász's inequality are quite hard to verify in practice. In this paper, using a novel approach to the anti-concentration problem for vector sums, we obtain new Halász-type inequalities which beat the Odlyzko bound even in settings where the number of equations is comparable to the number of variables. In addition to being stronger, our inequalities have hypotheses which are considerably easier to verify. We present two applications of our inequalities to combinatorial (random) matrix theory: (i) we obtain the first non-trivial upper bound on the number of $n\times n$ Hadamard matrices, and (ii) we improve a recent bound of Deneanu and Vu on the probability of normality of a random $\{\pm 1\}$ matrix.
On the k-planar local crossing number
Published
• View Publication
• BIB
Given a fixed positive integer $k$, the $k$-planar local crossing number of a graph $G$, denoted by $\text{LCR}_k(G)$, is the minimum positive integer $L$ such that $G$ can be decomposed into $k$ subgraphs, each of which can be drawn in a plane such that no edge is crossed more than $L$ times. In this note, we show that under certain natural restrictions, the ratio $\text{LCR}_k(G)/\text{LCR}_1(G)$ is of order $1/k^2$, which is analogous to a recent result of Pach et al. for the $k$-planar crossing number (defined as the minimum positive integer $C$ for which there is a $k$-planar drawing of $G$ with $C$ total edge crossings). As a corollary of our proof we show that, under similar restrictions, one may obtain a $k$-planar drawing of $G$ with \emph{both} the total number of edge crossings as well as the maximum number of times any edge is crossed essentially matching the best known bounds. Our proof relies on the crossing number inequality and several probabilistic tools such as concentration of measure and the Lovász local lemma.
Number of 1-factorizations of regular high-degree graphs
A $1$-factor in an $n$-vertex graph $G$ is a collection of $\frac{n}{2}$ vertex-disjoint edges and a $1$-factorization of $G$ is a partition of its edges into edge-disjoint $1$-factors. Clearly, a $1$-factorization of $G$ cannot exist unless $n$ is even and $G$ is regular (that is, all vertices are of the same degree). The problem of finding $1$-factorizations in graphs goes back to a paper of Kirkman in 1847 and has been extensively studied since then. Deciding whether a graph has a $1$-factorization is usually a very difficult question. For example, it took more than 60 years and an impressive tour de force of Csaba, Kühn, Lo, Osthus and Treglown to prove an old conjecture of Dirac from the 1950s, which says that every $d$-regular graph on $n$ vertices contains a $1$-factorization, provided that $n$ is even and $d\geq 2\lceil \frac{n}{4}\rceil-1$. In this paper we address the natural question of estimating $F(n,d)$, the number of $1$-factorizations in $d$-regular graphs on an even number of vertices, provided that $d\geq \frac{n}{2}+\varepsilon n$. Improving upon a recent result of Ferber and Jain, which itself improved upon a result of Cameron from the 1970s, we show that $F(n,d)\geq \left((1+o(1))\frac{d}{e^2}\right)^{nd/2}$, which is asymptotically best possible.
1-factorizations of pseudorandom graphs
A $1$-factorization of a graph $G$ is a collection of edge-disjoint perfect matchings whose union is $E(G)$. A trivial necessary condition for $G$ to admit a $1$-factorization is that $|V(G)|$ is even and $G$ is regular; the converse is easily seen to be false. In this paper, we consider the problem of finding $1$-factorizations of regular, pseudorandom graphs. Specifically, we prove that an $(n,d,λ)$-graph $G$ (that is, a $d$-regular graph on $n$ vertices whose second largest eigenvalue in absolute value is at most $λ$) admits a $1$-factorization provided that $n$ is even, $C_0\leq d\leq n-1$ (where $C_0$ is a universal constant), and $λ\leq d^{1-o(1)}$. In particular, since (as is well known) a typical random $d$-regular graph $G_{n,d}$ is such a graph, we obtain the existence of a $1$-factorization in a typical $G_{n,d}$ for all $C_0\leq d\leq n-1$, thereby extending to all possible values of $d$ results obtained by Janson, and independently by Molloy, Robalewska, Robinson, and Wormald for fixed $d$. Moreover, we also obtain a lower bound for the number of distinct $1$-factorizations of such graphs $G$ which is off by a factor of $2$ in the base of the exponent from the known upper bound. This lower bound is better by a factor of $2^{nd/2}$ than the previously best known lower bounds, even in the simplest case where $G$ is the complete graph. Our proofs are probabilistic and can be easily turned into polynomial time (randomized) algorithms.
The Mean-Field Approximation: Information Inequalities, Algorithms, and Complexity
The mean field approximation to the Ising model is a canonical variational tool that is used for analysis and inference in Ising models. We provide a simple and optimal bound for the KL error of the mean field approximation for Ising models on general graphs, and extend it to higher order Markov random fields. Our bound improves on previous bounds obtained in work in the graph limit literature by Borgs, Chayes, Lovász, Sós, and Vesztergombi and another recent work by Basak and Mukherjee. Our bound is tight up to lower order terms. Building on the methods used to prove the bound, along with techniques from combinatorics and optimization, we study the algorithmic problem of estimating the (variational) free energy for Ising models and general Markov random fields. For a graph $G$ on $n$ vertices and interaction matrix $J$ with Frobenius norm $\| J \|_F$, we provide algorithms that approximate the free energy within an additive error of $εn \|J\|_F$ in time $\exp(poly(1/ε))$. We also show that approximation within $(n \|J\|_F)^{1-δ}$ is NP-hard for every $δ> 0$. Finally, we provide more efficient approximation algorithms, which find the optimal mean field approximation, for ferromagnetic Ising models and for Ising models satisfying Dobrushin's condition.
The Vertex Sample Complexity of Free Energy is Polynomial
We study the following question: given a massive Markov random field on $n$ nodes, can a small sample from it provide a rough approximation to the free energy $\mathcal{F}_n = \log{Z_n}$?
Results in graph limit literature by Borgs, Chayes, Lovász, Sós, and Vesztergombi show that for Ising models on $n$ nodes and interactions of strength $Θ(1/n)$, an $ε$ approximation to $\log Z_n / n$ can be achieved by sampling a randomly induced model on $2^{O(1/ε^2)}$ nodes. We show that the sampling complexity of this problem is {\em polynomial in} $1/ε$. We further show a polynomial dependence on $ε$ cannot be avoided.
Our results are very general as they apply to higher order Markov random fields. For Markov random fields of order $r$, we obtain an algorithm that achieves $ε$ approximation using a number of samples polynomial in $r$ and $1/ε$ and running time that is $2^{O(1/ε^2)}$ up to polynomial factors in $r$ and $ε$. For ferromagnetic Ising models, the running time is polynomial in $1/ε$.
Our results are intimately connected to recent research on the regularity lemma and property testing, where the interest is in finding which properties can tested within $ε$ error in time polynomial in $1/ε$. In particular, our proofs build on results from a recent work by Alon, de la Vega, Kannan and Karpinski, who also introduced the notion of polynomial vertex sample complexity. Another critical ingredient of the proof is an effective bound by the authors of the paper relating the variational free energy and the free energy.