arXiv++ Combinatorics

Browse math.CO papers from arXiv

Papers by Cole Franks

5 paper(s) by this author · All BibTeX
2021-02-12 v3
Barriers for recent methods in geodesic optimization
Published • View PublicationBIB
We study a class of optimization problems including matrix scaling, matrix balancing, multidimensional array scaling, operator scaling, and tensor scaling that arise frequently in theory and in practice. Some of these problems, such as matrix and array scaling, are convex in the Euclidean sense, but others such as operator scaling and tensor scaling are geodesically convex on a different Riemannian manifold. Trust region methods, which include box-constrained Newton's method, are known to produce high precision solutions very quickly for matrix scaling and matrix balancing (Cohen et. al., FOCS 2017, Allen-Zhu et. al. FOCS 2017), and result in polynomial time algorithms for some geodesically convex problems like operator scaling (Garg et. al. STOC 2018, Bürgisser et. al. FOCS 2019). One is led to ask whether these guarantees also hold for multidimensional array scaling and tensor scaling. We show that this is not the case by exhibiting instances with exponential diameter bound: we construct polynomial-size instances of 3-dimensional array scaling and 3-tensor scaling whose approximate solutions all have doubly exponential condition number. Moreover, we study convex-geometric notions of complexity known as margin and gap, which are used to bound the running times of all existing optimization algorithms for such problems. We show that margin and gap are exponentially small for several problems including array scaling, tensor scaling and polynomial scaling. Our results suggest that it is impossible to prove polynomial running time bounds for tensor scaling based on diameter bounds alone. Therefore, our work motivates the search for analogues of more sophisticated algorithms, such as interior point methods, for geodesically convex optimization that do not rely on polynomial diameter bounds.
2018-11-02 v3
A simplified disproof of Beck's three permutations conjecture and an application to root-mean-squared discrepancy
Published • View PublicationBIB
A $k$-permutation family on $n$ vertices is a set system consisting of the intervals of $k$ permutations of the integers $1$ through $n$. The discrepancy of a set system is the minimum over all red-blue vertex colorings of the maximum difference between the number of red and blue vertices in any set in the system. In 2011, Newman and Nikolov disproved a conjecture of Beck that the discrepancy of any $3$-permutation family is at most a constant independent of $n$. Here we give a simpler proof that Newman and Nikolov's sequence of $3$-permutation families has discrepancy $Ω(\log n)$. We also exhibit a sequence of $6$-permutation families with root-mean-squared discrepancy $Ω(\sqrt{\log n})$; that is, in any red-blue vertex coloring, the square root of the expected difference between the number of red and blue vertices in an interval of the system is $Ω(\sqrt{\log n})$.
2018-07-11 v2
On the Discrepancy of Random Matrices with Many Columns
Published • View PublicationBIB
Motivated by the Komlós conjecture in combinatorial discrepancy, we study the discrepancy of random matrices with $m$ rows and $n$ independent columns drawn from a bounded lattice random variable. It is known that for $n$ tending to infinity and $m$ fixed, with high probability the $\ell_\infty$-discrepancy is at most twice the $\ell_\infty$-covering radius of the integer span of the support of the random variable. However, the easy argument for the above fact gives no concrete bounds on the failure probability in terms of $n$. We prove that the failure probability is inverse polynomial in $m, n$ and some well-motivated parameters of the random variable. We also obtain the analogous bounds for the discrepancy in arbitrary norms. We apply these results to two random models of interest. For random $t$-sparse matrices, i.e. uniformly random matrices with $t$ ones and $m-t$ zeroes in each column, we show that the $\ell_\infty$-discrepancy is at most $2$ with probability $1 - O(\sqrt{ \log n/n})$ for $n = Ω(m^3 \log^2 m)$. This improves on a bound proved by Ezra and Lovett (Ezra and Lovett, Approx+Random, 2016) showing that the same is true for $n$ at least $m^t$. For matrices with random unit vector columns, we show that the $\ell_\infty$-discrepancy is $O(\exp(\sqrt{ n/m^3}))$ with probability $1 - O(\sqrt{ \log n/n})$ for $n = Ω(m^3 \log^2 m).$ Our approach, in the spirit of Kuperberg, Lovett and Peled (G. Kuperberg, S. Lovett and R. Peled, STOC 2012), uses Fourier analysis to prove that for $m \times n$ matrices $M$ with i.i.d. columns, and $m$ sufficiently large, the distribution of $My$ for random $y \in \{-1,1\}^n$ obeys a local limit theorem.
2018-01-01 v2
Operator scaling with specified marginals
Published • View PublicationBIB
The completely positive maps, a generalization of the nonnegative matrices, are a well-studied class of maps from $n\times n$ matrices to $m\times m$ matrices. The existence of the operator analogues of doubly stochastic scalings of matrices is equivalent to a multitude of problems in computer science and mathematics, such rational identity testing in non-commuting variables, noncommutative rank of symbolic matrices, and a basic problem in invariant theory (Garg, Gurvits, Oliveira and Wigderson, FOCS, 2016). We study operator scaling with specified marginals, which is the operator analogue of scaling matrices to specified row and column sums. We characterize the operators which can be scaled to given marginals, much in the spirit of the Gurvits' algorithmic characterization of the operators that can be scaled to doubly stochastic (Gurvits, Journal of Computer and System Sciences, 2004). Our algorithm produces approximate scalings in time poly(n,m) whenever scalings exist. A central ingredient in our analysis is a reduction from the specified marginals setting to the doubly stochastic setting. Operator scaling with specified marginals arises in diverse areas of study such as the Brascamp-Lieb inequalities, communication complexity, eigenvalues of sums of Hermitian matrices, and quantum information theory. Some of the known theorems in these areas, several of which had no effective proof, are straightforward consequences of our characterization theorem. For instance, we obtain a simple algorithm to find, when they exist, a tuple of Hermitian matrices with given spectra whose sum has a given spectrum. We also prove new theorems such as a generalization of Forster's theorem (Forster, Journal of Computer and System Sciences, 2002) concerning radial isotropic position.
2013-09-01 v3
Graph Labeling with Distance Conditions and the Delta Squared Conjecture
We give bounds on the L(2,1)-labeling number of a simple graph in terms of its order and its maximum degree. We also describe an infinite class of graphs of which the elements have the highest L(2,1)-labeling numbers in terms of their maximum degrees of any known infinite class of graphs.