Papers by Mark Sellke
7 paper(s) by this author
· All BibTeX
Short proofs in combinatorics, probability and number theory II
We give a quintet of proofs resulting from questions posed by Erdős. These questions concern ordinary lines in planar point sets, sequences with uniformly small exponential sums, $K_4$-free $4$-critical graphs with few chords in any cycle, a counterexample to a "fewnomial" version of the Erdős--Turán discrepancy bound, and a finiteness theorem for integers $n$ such that $n-a k^2$ is prime for all $k\leq \sqrt{n/a}$ coprime to $n$ (for fixed $a\in\mathbb Z_+$). Each proof is due to an internal model at OpenAI.
Short proofs in combinatorics and number theory
We give a triplet of short proofs, each of which answers a question raised by Erdős. The first concerns the small prime factors of $\binom{n}{k}$, the second concerns whether an additive basis $A$ can always be split into pieces $A_1$ and $A_2$ such that each of $A_i + A_i$ has bounded gaps, and the final concerns whether $\{αp\}$ is "well-distributed" in the sense introduced by Hlawka and Petersen. In each case, the proof is due entirely to an internal model at OpenAI.
Improved Lower Bound for Frankl's Union-Closed Sets Conjecture
Published
• View Publication
• BIB
We verify an explicit inequality conjectured recently by Gilmer, thus proving that for any nonempty union-closed family $F \subseteq 2^{[n]}$, some $i\in [n]$ is contained in at least a $\frac{3-\sqrt{5}}{2} \approx 0.38$ fraction of the sets in $F$. One case, an explicit one-variable inequality, is checked by computer calculation.
Local algorithms for Maximum Cut and Minimum Bisection on locally treelike regular graphs of large degree
Published
• View Publication
• BIB
Given a graph $G$ of degree $k$ over $n$ vertices, we consider the problem of computing a near maximum cut or a near minimum bisection in polynomial time. For graphs of girth $2L$, we develop a local message passing algorithm whose complexity is $O(nkL)$, and that achieves near optimal cut values among all $L$-local algorithms. Focusing on max-cut, the algorithm constructs a cut of value $nk/4+ n\mathsf{P}_\star\sqrt{k/4}+\mathsf{err}(n,k,L)$, where $\mathsf{P}_\star\approx 0.763166$ is the value of the Parisi formula from spin glass theory, and $\mathsf{err}(n,k,L)=o_n(n)+no_k(\sqrt{k})+n \sqrt{k} o_L(1)$ (subscripts indicate the asymptotic variables). Our result generalizes to locally treelike graphs, i.e., graphs whose girth becomes $2L$ after removing a small fraction of vertices.
Earlier work established that, for random $k$-regular graphs, the typical max-cut value is $nk/4+ n\mathsf{P}_\star\sqrt{k/4}+o_n(n)+no_k(\sqrt{k})$. Therefore our algorithm is nearly optimal on such graphs. An immediate corollary of this result is that random regular graphs have nearly minimum max-cut, and nearly maximum min-bisection among all regular locally treelike graphs. This can be viewed as a combinatorial version of the near-Ramanujan property of random regular graphs.
Covering $\mathsf{Irrep}(S_n)$ With Tensor Products and Powers
Published
• View Publication
• BIB
We study when a tensor product of irreducible representations of the symmetric group $S_n$ contains all irreducibles as subrepresentations; we say such a tensor product covers $\mathsf{Irrep}(S_n)$. Our results show that this behavior is typical. We first give a general sufficient criterion for tensor products to have this property, which holds asymptotically almost surely for constant-sized collections of (Plancherel or uniformly) random irreducibles. We also consider the minimal tensor power of a single fixed irreducible representation needed to cover $\mathsf{Irrep}(S_n)$. Here a simple lower bound comes from considering dimensions, and we show it is always tight up to a universal constant factor as was recently conjectured by Liebeck, Shalev, and Tiep.
Approximating Continuous Functions by ReLU Nets of Minimal Width
This article concerns the expressive power of depth in deep feed-forward neural nets with ReLU activations. Specifically, we answer the following question: for a fixed $d_{in}\geq 1,$ what is the minimal width $w$ so that neural nets with ReLU activations, input dimension $d_{in}$, hidden layer widths at most $w,$ and arbitrary depth can approximate any continuous, real-valued function of $d_{in}$ variables arbitrarily well? It turns out that this minimal width is exactly equal to $d_{in}+1.$ That is, if all the hidden layer widths are bounded by $d_{in}$, then even in the infinite depth limit, ReLU nets can only express a very limited class of functions, and, on the other hand, any continuous function on the $d_{in}$-dimensional unit cube can be approximated to arbitrary precision by ReLU nets in which all hidden layers have width exactly $d_{in}+1.$ Our construction in fact shows that any continuous function $f:[0,1]^{d_{in}}\to\mathbb R^{d_{out}}$ can be approximated by a net of width $d_{in}+d_{out}$. We obtain quantitative depth estimates for such an approximation in terms of the modulus of continuity of $f$.
The Saxl Conjecture for Fourth Powers via the Semigroup Property
Published in J. Algebra. Comb. (2017) 45:33-80
• View Publication
• BIB
The tensor square conjecture states that for $n \geq 10$, there is an irreducible representation $V$ of the symmetric group $S_n$ such that $V \otimes V$ contains every irreducible representation of $S_n$. Our main result is that for large enough $n$, there exists an irreducible representation $V$ such that $V^{\otimes 4}$ contains every irreducible representation. We also show that tensor squares of certain irreducible representations contain $(1-o(1))$-fraction of irreducible representations with respect to two natural probability distributions. Our main tool is the semigroup property, which allows us to break partitions down into smaller ones.