Papers by Dmitriy Kunisky
15 paper(s) by this author
· All BibTeX
Inequalities for rank-two permanents and finite free convolutions
Bang (1976) proved the inequality for matrix permanents $\mathrm{per}^2(A) \geq 2^{-2n}\mathrm{per}(A \otimes J_2)$, where $J_2$ is the $2 \times 2$ all-ones matrix and $A$ is any $n \times n$ matrix with non-negative entries.
We show that, if $A$ is any $n \times n$ real-valued matrix with rank at most two (possibly having negative entries), this inequality can be sharpened, replacing the constant $2^{-2n}$ by $1 / \binom{2n}{n} = (n!)^2 / (2n)! > 2^{-2n}$.
We then show that this sharpened inequality also implies new inequalities for finite free convolutions of polynomials: if $p$ and $q$ are monic real-rooted polynomials of degree $n$, then $(p \boxplus_n q)(x)^2 \geq (p^2 \boxplus_{2n} q^2)(x)$ and $(p \boxtimes_n q)(x)^2 \geq (p^2 \boxtimes_{2n} q^2)(x)$ for all $x \in \mathbb{R}$, for $\boxplus_n$ and $\boxtimes_n$ the finite free additive and multiplicative convolution operations, respectively, on polynomials of degree $n$.
The even-uniform hypergraph Moore bound
The hypergraph Moore bound conjectured by Feige (2008) controls the size of the smallest even cover in a $k$-uniform hypergraph in terms of the average density of hyperedges. An even cover is a set of hyperedges covering each vertex an even number of times, generalizing the notion of a cycle in a graph, so the size of the smallest non-trivial even cover provides a notion of hypergraph girth. Recent work starting from the breakthrough result of Guruswami, Kothari, and Manohar (2022) proved the conjecture up to polylogarithmic factors, whose exponents were later gradually improved. We give a simple proof of Feige's original hypergraph Moore bound conjecture for all even $k\ge 4$, with no superfluous polylogarithmic factors. Our proof roughly follows the proof of the graph Moore bound, but works with colored walks in a Kikuchi graph built from a hypergraph and controls their growth using a polynomial interpolation method.
The Lovász number of random circulant graphs
This paper addresses the behavior of the Lovász number for dense random circulant graphs. The Lovász number is a well-known semidefinite programming upper bound on the independence number. Circulant graphs, an example of a Cayley graph, are highly structured vertex-transitive graphs on integers modulo $n$, where the connectivity of pairs of vertices depends only on the difference between their labels. While for random circulant graphs the asymptotics of fundamental quantities such as the clique and the chromatic number are well-understood, characterizing the exact behavior of the Lovász number remains open. In this work, we provide upper and lower bounds on the expected value of the Lovász number and show that it scales as the square root of the number of vertices, up to a log log factor. Our proof relies on a reduction of the semidefinite program formulation of the Lovász number to a linear program with random objective and constraints via diagonalization of the adjacency matrix of a circulant graph by the discrete Fourier transform (DFT). This leads to a problem about controlling the norms of vectors with sparse Fourier coefficients, which we study using results on the restricted isometry property of subsampled DFT matrices.
Low coordinate degree algorithms II: Categorical signals and generalized stochastic block models
We study when low coordinate degree functions (LCDF) -- linear combinations of functions depending on small subsets of entries of a vector -- can test for the presence of categorical structure, including community structure and generalizations thereof, in high-dimensional data. This complements the first paper of this series, which studied the power of LCDF in testing for continuous structure like real-valued signals perturbed by additive noise. We apply the tools developed there to a general form of stochastic block model (SBM), where a population is assigned random labels and every $p$-tuple of the population generates an observation according to an arbitrary probability measure associated to the $p$ labels of its members. We show that the performance of LCDF admits a unified analysis for this class of models. As applications, we prove tight lower bounds against LCDF (and therefore also against low degree polynomials) for nearly arbitrary graph and regular hypergraph SBMs, always matching suitable generalizations of the Kesten-Stigum threshold. We also prove tight lower bounds for group synchronization and abelian group sumset problems under the "truth-or-Haar" noise model, and use our technical results to give an improved analysis of Gaussian multi-frequency group synchronization. In most of these models, for some parameter settings our lower bounds give new evidence for conjectural statistical-to-computational gaps. Finally, interpreting some of our findings, we propose a precise analogy between categorical and continuous signals: a general SBM as above behaves, in terms of the tradeoff between subexponential runtime cost of testing algorithms and the signal strength needed for a testing algorithm to succeed, like a spiked $p_*$-tensor model of a certain order $p_*$ that may be computed from the parameters of the SBM.
Statistical inference of a ranked community in a directed graph
We study the problem of detecting or recovering a planted ranked subgraph from a directed graph, an analog for directed graphs of the well-studied planted dense subgraph model. We suppose that, among a set of $n$ items, there is a subset $S$ of $k$ items having a latent ranking in the form of a permutation $π$ of $S$, and that we observe a fraction $p$ of pairwise orderings between elements of $\{1, \dots, n\}$ which agree with $π$ with probability $\frac{1}{2} + q$ between elements of $S$ and otherwise are uniformly random. Unlike in the planted dense subgraph and planted clique problems where the community $S$ is distinguished by its unusual density of edges, here the community is only distinguished by the unusual consistency of its pairwise orderings. We establish computational and statistical thresholds for both detecting and recovering such a ranked community. In the log-density setting where $k$, $p$, and $q$ all scale as powers of $n$, we establish the exact thresholds in the associated exponents at which detection and recovery become statistically and computationally feasible. These regimes include a rich variety of behaviors, exhibiting both statistical-computational and detection-recovery gaps. We also give finer-grained results for two extreme cases: (1) $p = 1$, $k = n$, and $q$ small, where a full tournament is observed that is weakly correlated with a global ranking, and (2) $p = 1$, $q = \frac{1}{2}$, and $k$ small, where a small "ordered clique" (totally ordered directed subgraph) is planted in a random tournament.
Asymptotic Bounds and Online Algorithms for Average-Case Matrix Discrepancy
We study the matrix discrepancy problem in the average-case setting. Given a sequence of $m \times m$ symmetric matrices $A_1,\ldots,A_n$, its discrepancy is defined as the minimal spectral norm over all signed sums $\sum_{i=1}^n x_iA_i$ with $x_1,\ldots,x_n \in \{\pm1\}$. Our contributions are twofold. First, we study the asymptotic discrepancy of random matrices. When the matrices belong to the Gaussian orthogonal ensemble, we provide a sharp characterization of the asymptotic discrepancy and show that the limiting distribution is concentrated around $Θ(\sqrt{nm}4^{-(1 + o(1))n/m^2})$, under the assumption $m^2 \ll n/\log{n}$. We observe that the trivial bound $O(\sqrt{nm})$ cannot be improved when $n \ll m^2$ and show that this phenomenon occurs for a broad class of random matrices. In the case $n = Ω(m^2)$, we provide a matching upper bound. Second, we analyse the matrix hyperbolic cosine algorithm, an online algorithm for matrix discrepancy minimization due to Zouzias (2011), in the average-case setting. We show that the algorithm achieves with high probability a discrepancy of $O(m\log{m})$ for a broad class of random matrices, including Wigner matrices with entries satisfying a hypercontractive inequality and Gaussian Wishart matrices.
On the Structure of Bad Science Matrices
The bad science matrix problem consists in finding, among all matrices $A \in \mathbb{R}^{n \times n}$ with rows having unit $\ell^2$ norm, one that maximizes $β(A) = \frac{1}{2^n} \sum_{x \in \{-1, 1\}^n} \|Ax\|_\infty$. Our main contribution is an explicit construction of an $n \times n$ matrix $A$ showing that $β(A) \geq \sqrt{\log_2(n+1)}$, which is only 18% smaller than the asymptotic rate. We prove that every entry of any optimal matrix is a square root of a rational number, and we find provably optimal matrices for $n \leq 4$.
Inference of rankings planted in random tournaments
We consider the problem of inferring an unknown ranking of $n$ items from a random tournament on $n$ vertices whose edge directions are correlated with the ranking. We establish, in terms of the strength of these correlations, the computational and statistical thresholds for detection (deciding whether an observed tournament is purely random or drawn correlated with a hidden ranking) and recovery (estimating the hidden ranking with small error in Spearman's footrule or Kendall's tau metric on permutations). Notably, we find that this problem provides a new instance of a detection-recovery gap: solving the detection problem requires much weaker correlations than solving the recovery problem. In establishing these thresholds, we also identify simple algorithms for detection (thresholding a degree 2 polynomial) and recovery (outputting a ranking by the number of "wins" of a tournament vertex, i.e., the out-degree) that achieve optimal performance up to constants in the correlation strength. For detection, we find that the above low-degree polynomial algorithm is superior to a natural spectral algorithm. We also find that, whenever it is possible to achieve strong recovery (i.e., to estimate with vanishing error in the above metrics) of the hidden ranking, then the above "Ranking By Wins" algorithm not only does so, but also outputs a close approximation of the maximum likelihood estimator, a task that is NP-hard in the worst case.
Computational hardness of detecting graph lifts and certifying lift-monotone properties of random regular graphs
We introduce a new conjecture on the computational hardness of detecting random lifts of graphs: we claim that there is no polynomial-time algorithm that can distinguish between a large random $d$-regular graph and a large random lift of a Ramanujan $d$-regular base graph (provided that the lift is corrupted by a small amount of extra noise), and likewise for bipartite random graphs and lifts of bipartite Ramanujan graphs. We give evidence for this conjecture by proving lower bounds against the local statistics hierarchy of hypothesis testing semidefinite programs. We then explore the consequences of this conjecture for the hardness of certifying bounds on numerous functions of random regular graphs, expanding on a direction initiated by Bandeira, Banks, Kunisky, Moore, and Wein (2021). Conditional on this conjecture, we show that no polynomial-time algorithm can certify tight bounds on the maximum cut of random 3- or 4-regular graphs, the maximum independent set of random 3- or 4-regular graphs, or the chromatic number of random 7-regular graphs. We show similar gaps asymptotically for large degree for the maximum independent set and for any degree for the minimum dominating set, finding that naive spectral and combinatorial bounds are optimal among all polynomial-time certificates. Likewise, for small-set vertex and edge expansion in the limit of very small sets, we show that the spectral bounds of Kahale (1995) are optimal among all polynomial-time certificates.
Average-Case Matrix Discrepancy: Asymptotics and Online Algorithms
We study the operator norm discrepancy of i.i.d. random matrices, initiating the matrix-valued analog of a long line of work on the $\ell^{\infty}$ norm discrepancy of i.i.d. random vectors. First, using repurposed results on vector discrepancy and new first moment method calculations, we give upper and lower bounds on the discrepancy of random matrices. We treat i.i.d. matrices drawn from the Gaussian orthogonal ensemble (GOE) and low-rank Gaussian Wishart distributions. In both cases, for what turns out to be the "critical" number of $Θ(n^2)$ matrices of dimension $n \times n$, we identify the discrepancy up to constant factors. Second, we give a new analysis of the matrix hyperbolic cosine algorithm of Zouzias (2011), a matrix version of an online vector discrepancy algorithm of Spencer (1977) studied for average-case inputs by Bansal and Spencer (2020), for the case of i.i.d. random matrix inputs. We both give a general analysis and extract concrete bounds on the discrepancy achieved by this algorithm for matrices with independent entries (including GOE matrices) and Gaussian Wishart matrices.
Spectral pseudorandomness and the road to improved clique number bounds for Paley graphs
We study subgraphs of Paley graphs of prime order $p$ induced on the sets of vertices extending a given independent set of size $a$ to a larger independent set. Using a sufficient condition proved in the author's recent companion work, we show that a family of character sum estimates would imply that, as $p \to \infty$, the empirical spectral distributions of the adjacency matrices of any sequence of such subgraphs have the same weak limit (after rescaling) as those of subgraphs induced on a random set including each vertex independently with probability $2^{-a}$, namely, a Kesten-McKay law with parameter $2^a$. We prove the necessary estimates for $a = 1$, obtaining in the process an alternate proof of a character sum equidistribution result of Xi (2022), and provide numerical evidence for this weak convergence for $a \geq 2$. We also conjecture that the minimum eigenvalue of any such sequence converges (after rescaling) to the left edge of the corresponding Kesten-McKay law, and provide numerical evidence for this convergence. Finally, we show that, once $a \geq 3$, this (conjectural) convergence of the minimum eigenvalue would imply bounds on the clique number of the Paley graph improving on the current state of the art due to Hanson and Petridis (2021), and that this convergence for all $a \geq 1$ would imply that the clique number is $o(\sqrt{p})$.
Generic MANOVA limit theorems for products of projections
We study the convergence of the empirical spectral distribution of $\mathbf{A} \mathbf{B} \mathbf{A}$ for $N \times N$ orthogonal projection matrices $\mathbf{A}$ and $\mathbf{B}$, where $\frac{1}{N}\mathrm{Tr}(\mathbf{A})$ and $\frac{1}{N}\mathrm{Tr}(\mathbf{B})$ converge as $N \to \infty$, to Wachter's MANOVA law. Using free probability, we show mild sufficient conditions for convergence in moments and in probability, and use this to prove a conjecture of Haikin, Zamir, and Gavish (2017) on random subsets of unit-norm tight frames. This result generalizes previous ones of Farrell (2011) and Magsino, Mixon, and Parshall (2021). We also derive an explicit recursion for the difference between the empirical moments $\frac{1}{N}\mathrm{Tr}((\mathbf{A} \mathbf{B} \mathbf{A})^k)$ and the limiting MANOVA moments, and use this to prove a sufficient condition for convergence in probability of the largest eigenvalue of $\mathbf{A} \mathbf{B} \mathbf{A}$ to the right edge of the support of the limiting law in the special case where that law belongs to the Kesten-McKay family. As an application, we give a new proof of convergence in probability of the largest eigenvalue when $\mathbf{B}$ is unitarily invariant; equivalently, this determines the limiting operator norm of a rectangular submatrix of size $\frac{1}{2}N \times αN$ of a Haar-distributed $N \times N$ unitary matrix for any $α\in (0, 1)$. Unlike previous proofs, we use only moment calculations and non-asymptotic bounds on the unitary Weingarten function, which we believe should pave the way to analyzing the largest eigenvalue for products of random projections having other distributions.
The spectrum of the Grigoriev-Laurent pseudomoments
Published
• View Publication
• BIB
Grigoriev (2001) and Laurent (2003) independently showed that the sum-of-squares hierarchy of semidefinite programs does not exactly represent the hypercube $\{\pm 1\}^n$ until degree at least $n$ of the hierarchy. Laurent also observed that the pseudomoment matrices her proof constructs appear to have surprisingly simple and recursively structured spectra as $n$ increases. While several new proofs of the Grigoriev-Laurent lower bound have since appeared, Laurent's observations have remained unproved. We give yet another, representation-theoretic proof of the lower bound, which also yields exact formulae for the eigenvalues of the Grigoriev-Laurent pseudomoments. Using these, we prove and elaborate on Laurent's observations.
Our arguments have two features that may be of independent interest. First, we show that the Grigoriev-Laurent pseudomoments are a special case of a Gram matrix construction of pseudomoments proposed by Bandeira and Kunisky (2020). Second, we find a new realization of the irreducible representations of the symmetric group corresponding to Young diagrams with two rows, as spaces of multivariate polynomials that are multiharmonic with respect to an equilateral simplex.
The discrepancy of unsatisfiable matrices and a lower bound for the Komlós conjecture constant
Published
• View Publication
• BIB
We construct simple, explicit matrices with columns having unit $\ell^2$ norm and discrepancy approaching $1 + \sqrt{2} \approx 2.414$. This number gives a lower bound, the strongest known as far as we are aware, on the constant appearing in the Komlós conjecture. The "unsatisfiable matrices" giving this bound are built by scaling the entries of clause-variable matrices of certain unsatisfiable Boolean formulas. We show that, for a given formula, such a scaling maximizing a lower bound on the discrepancy may be computed with a convex second-order cone program. Using a dual certificate for this program, we show that our lower bound is optimal among those using unsatisfiable matrices built from formulas admitting read-once resolution proofs of unsatisfiability. We also conjecture that a generalization of this certificate shows that our bound is optimal among all bounds using unsatisfiable matrices.
Spectral Planting and the Hardness of Refuting Cuts, Colorability, and Communities in Random Graphs
We study the problem of efficiently refuting the k-colorability of a graph, or equivalently certifying a lower bound on its chromatic number. We give formal evidence of average-case computational hardness for this problem in sparse random regular graphs, showing optimality of a simple spectral certificate. This evidence takes the form of a computationally-quiet planting: we construct a distribution of d-regular graphs that has significantly smaller chromatic number than a typical regular graph drawn uniformly at random, while providing evidence that these two distributions are indistinguishable by a large class of algorithms. We generalize our results to the more general problem of certifying an upper bound on the maximum k-cut.
This quiet planting is achieved by minimizing the effect of the planted structure (e.g. colorings or cuts) on the graph spectrum. Specifically, the planted structure corresponds exactly to eigenvectors of the adjacency matrix. This avoids the pushout effect of random matrix theory, and delays the point at which the planting becomes visible in the spectrum or local statistics. To illustrate this further, we give similar results for a Gaussian analogue of this problem: a quiet version of the spiked model, where we plant an eigenspace rather than adding a generic low-rank perturbation.
Our evidence for computational hardness of distinguishing two distributions is based on three different heuristics: stability of belief propagation, the local statistics hierarchy, and the low-degree likelihood ratio. Of independent interest, our results include general-purpose bounds on the low-degree likelihood ratio for multi-spiked matrix models, and an improved low-degree analysis of the stochastic block model.