Papers by Konstantin Tikhomirov
22 paper(s) by this author
· All BibTeX
Online Permutation Embedding: Optimal Stopping and Scaling Laws
We study optimal online algorithms for embedding a permutation $π$ of $[k]$ into an iid stream of uniform $[0,1]$ random variables. This problem is a broad generalization of the classical online monotone subsequence selection problem, recovered in the special case $π=\mathrm{Id}_k$. Our first contribution is an efficiently solvable dynamic program for the optimal embedding time of any $k$-permutation $π$. This dynamic program also yields an explicit optimal online embedding algorithm. We then investigate the asymptotic scaling of the optimal embedding time for uniformly random target permutations, as well as the extremal problem of identifying the permutations with largest expected online embedding time. Our second main result shows that, to first order, random permutations are strictly faster to embed than monotone permutations, which in turn are strictly faster to embed than the extremal permutations. This separation stands in sharp contrast to prevailing conjectures and heuristics in the offline theory of permutation embeddings.
Online Beck--Fiala Down to Logarithmic Sparsity
The Beck--Fiala conjecture asserts that every matrix $A\in\{0,1\}^{n\times T}$ with at most $d$ nonzero entries in each column has discrepancy $O(\sqrt d)$. A major breakthrough result of Bansal and Jiang recently established the validity of the conjecture for $d \ge \log(T)^2$. The present article extends the validity of the classical \textit{offline} Beck--Fiala conjecture to $d \ge \log(T)^{1+o(1)}$; moreover, the main thrust of the result is that it is actually obtained by an efficient \textit{online} algorithm that minimizes prefix discrepancy. The result is also essentially optimal, since online prefix discrepancy is known to scale as $ω(\sqrt{d})$ for $d =o(\log T)$. As an immediate corollary, the open question of online vector balancing in the Spencer setting is also resolved.
The algorithm is based on a compactly supported Metropolis fixed-point walk, constructed by combining ideas from several recent works on the online Komlós problem. The proof was generated in conversation with ChatGPT 5.6 Pro; the authors provided high-level guidance in several rounds of prompting, followed by manual checking and rewriting of the proof.
Metric Poincaré inequalities for graphs
This article obtains purely metric counterparts of cornerstone results in the theory of embedding graphs into normed spaces. Our first main result is a metric analogue of Matoušek's extrapolation relating the Poincaré constants $γ(G,\varrho^p)$ and $γ(G,\varrho^q)$ for any exponents $0 < p,q < \infty$, any bounded-degree expander graph $G$, and any target metric space $\mathcal{M}=(M,\varrho)$. Our second main result provides a sharp estimate of the Poincaré constant $γ(G,\varrho)$ in terms of the cardinalities of the vertex set of $G$ and the metric space $\mathcal{M}=(M,\varrho)$, in the setting of \textit{random} graphs. This yields optimal estimates on the minimum cardinality of (bi-Lipschitz) universal metric spaces for graphs, finally establishing a nonlinear analogue of Matoušek's celebrated "incompressibility" theorem (1996). Further, we obtain estimates on the nonlinear spectral gap of metric snowflakes and sharp lower bounds on the distortion of random regular graphs into arbitrary metric spaces. Our proofs develop new nonlinear techniques, including random compression methods and a novel structural dichotomy for metric embeddings.
A threshold for online balancing of sparse i.i.d. vectors
Consider the task of \textit{online} vector balancing for stochastic arrivals $(X_i)_{i \in [T]}$, where the time horizon satisfies $T = Θ(n)$, and the $X_i$ are i.i.d uniform $d$--sparse $n$--dimensional binary vectors, with $2\leq d \le (\log\log n)^2/\log\log\log n$. We show that for this range of parameters, every online algorithm incurs discrepancy at least $Ω(\log \log n)$, and there is an efficient algorithm which achieves a matching discrepancy bound of $O(\log\log n)$ w.h.p. This establishes an asymptotic gap, both existential and algorithmic, between the online and offline versions of the average--case Beck--Fiala problem. Strikingly, the optimal online discrepancy in the considered setting is order $\log \log n$, independent of $d$ and the norms of the vectors $(X_i)_i$. Our assumptions on $d$ are nearly optimal, as this independence ceases when $d=ω((\log\log n)^2)$.
Metric dimension reduction modulus for superlogarithmic distortion
The metric dimension reduction modulus $k^α_n(\ell_\infty)$ is the smallest $k$ such that every $n$--point metric space can be embedded into some $k$-dimensional normed space, with bi--Lipschitz distortion at most $α$. Determining sharp asymptotics for $k^α_n(\ell_\infty)$ is a fundamental task in metric geometry, with $α=Θ(\log n)$ bearing particular interest. A line of advances over the past decades has led to an upper bound on $k^α_n(\ell_\infty)$ for $α= Ω(\log n)$, but a matching lower bound has remained open. We close this gap, establishing: for every fixed $β> 0$, $$ k^α_n(\ell_\infty) =Θ\bigg(\frac{\log n}{\log(\fracα{\log n}+1)}\bigg)\quad \mbox{for every $α\geq β\log n$}. $$ This resolves a question from Naor's 2018 ICM plenary lecture. Our result is obtained by characterizing the minimum dimension $d$ for which, with high probability, a random regular graph admits an $α$--embedding into some $d$--dimensional normed space.
Discrete Poincaré inequalities and universal approximators for random graphs
Nonlinear Poincaré inequalities are indispensable tools in the study of dimension reduction and low-distortion embeddings of graphs into metric spaces, and have found remarkable algorithmic applications. A basic open problem, posed by Jon Kleinberg (2013), asks whether the optimal nonlinear Poincaré constant for maps between two independent $3$-regular random graphs is dimension-free, i.e., independent of vertex-set sizes. We give a complete and affirmative resolution to Kleinberg's problem, also allowing for arbitrary graph degrees. As a corollary, we obtain a stochastic construction of $O(1)\text{-universal}$ approximators for random graphs, answering a question of Mendel and Naor.
A universal threshold for geometric embeddings of trees
A graph $G=(V,E)$ is geometrically embeddable into a normed space $X$ when there is a mapping $ζ: V\to X$ such that $\|ζ(v)-ζ(w)\|_X\leqslant 1$ if and only if $\{v,w\}\in E$, for all distinct $v,w\in V$. Our result is the following universal threshold for the embeddability of trees. Let $Δ\geqslant 3$, and let $N$ be sufficiently large in terms of $Δ$. Every $N$--vertex tree of maximal degree at most $Δ$ is embeddable into any normed space of dimension at least $64\,\frac{\log N}{\log\log N}$, and complete trees are non-embeddable into any normed space of dimension less than $\frac{1}{2}\,\frac{\log N}{\log\log N}$. In striking contrast, spectral expanders and random graphs are known to be non-embeddable in sublogarithmic dimension. Our result is based on a randomized embedding whose analysis utilizes the recent breakthroughs on Bourgain's slicing problem.
Universal geometric non-embedding of random regular graphs
Let $Δ\ge 3$ be fixed, $n \ge n_Δ$ be a large integer. It is a classical result that $Δ$--regular expanders on $n$ vertices are not embeddable as geometric (distance) graphs into Euclidean space of dimension less than $c \log n$, for some universal constant $c$. We show that for typical $Δ$-regular graphs, this obstruction is universal with respect to the choice of norm. More precisely, for a uniform random $Δ$-regular graph $G$ on $n$ vertices, it holds with high probability: there is no normed space of dimension less than $c\log n$ which admits a geometric graph isomorphic to $G$. The proof is based on a seeded multiscale $\varepsilon$--net argument.
Locally seeded embeddings, and Ramsey numbers of bipartite graphs with sublinear bandwidth
A seminal result of Lee asserts that the Ramsey number of any bipartite $d$-degenerate graph $H$ satisfies $\log r(H) = \log n + O(d)$. In particular, this bound applies to every bipartite graph of maximal degree $Δ$. It remains a compelling challenge to identify conditions that guarantee that an $n$-vertex graph $H$ has Ramsey number linear in $n$, independently of $Δ$. Our contribution is a characterization of bipartite graphs with linear-size Ramsey numbers in terms of graph bandwidth, a notion of local connectivity. We prove that for any $n$-vertex bipartite graph $H$ with maximal degree at most $Δ$ and bandwidth $b(H)$ at most $\exp(-CΔ\logΔ)\,n$, we have $\log r(H) = \log n + O(1)$. This characterization is nearly optimal: for every $Δ$ there exists an $n$-vertex bipartite graph $H$ of degree at most $Δ$ and $b(H) \leq \exp(-cΔ)\,n$, such that $\log r(H) = \log n + Ω(Δ)$. We also provide bounds interpolating between these two bandwidth regimes.
A combinatorial approach to nonlinear spectral gaps
A seminal open question of Pisier and Mendel--Naor asks whether every degree-regular graph which satisfies the classical discrete Poincaré inequality for scalar functions, also satisfies an analogous inequality for functions taking values in \textit{any} normed space with non-trivial cotype. Motivated by applications, it is also greatly important to quantify the dependence of the corresponding optimal Poincaré constant on the cotype $q$. Works of Odell--Schlumprecht (1994), Ozawa (2004), and Naor (2014) make substantial progress on the former question by providing a positive answer for normed spaces which also have an unconditional basis, in addition to finite cotype. However, little is known in the way of quantitative estimates: the mentioned results imply a bound on the Poincaré constant depending super-exponentially on $q$.
We introduce a novel combinatorial framework for proving quantitative nonlinear spectral gap estimates. The centerpiece is a property of regular graphs that we call \emph{long range expansion}, which holds with high probability for random regular graphs. Our main result is that any regular graph with the long-range expansion property satisfies a discrete Poincaré inequality for any normed space with an unconditional basis and cotype $q$, with a Poincaré constant that depends \emph{polynomially} on $q$, which is optimal. As an application, any normed space with an unconditional basis which admits a low distortion embedding of an $n$-vertex random regular graph, must have cotype at least polylogarithmic in $n$. This extends a celebrated lower-bound of Matoušek for low distortion embeddings of random graphs into $\ell_q$ spaces.
Upgrading MLSI to LSI for reversible Markov chains
Published
• View Publication
• BIB
For reversible Markov chains on finite state spaces, we show that the modified log-Sobolev inequality (MLSI) can be upgraded to a log-Sobolev inequality (LSI) at the surprisingly low cost of degrading the associated constant by $\log (1/p)$, where $p$ is the minimum non-zero transition probability. We illustrate this by providing the first log-Sobolev estimate for Zero-Range processes on arbitrary graphs. As another application, we determine the modified log-Sobolev constant of the Lamplighter chain on all bounded-degree graphs, and use it to provide negative answers to two open questions by Montenegro and Tetali (2006) and Hermon and Peres (2018). Our proof builds upon the `regularization trick' recently introduced by the last two authors.
On bounded degree graphs with large size-Ramsey numbers
Published
• View Publication
• BIB
The size-Ramsey number $\hat r(G')$ of a graph $G'$ is defined as the smallest integer $m$ so that there exists a graph $G$ with $m$ edges such that every $2$-coloring of the edges of $G$ contains a monochromatic copy of $G'$. Answering a question of Beck, Rodl and Szemeredi showed that for every $n\geq 1$ there exists a graph $G'$ on $n$ vertices each of degree at most three, with the size-Ramsey number at least $cn\log^{\frac{1}{60}}n$ for a universal constant $c>0$. In this note we show that a modification of Rodl and Szemeredi's construction leads to a bound $\hat r(G')\geq cn\,\exp(c\sqrt{\log n})$.
A remark on the Ramsey number of the hypercube
Published
• View Publication
• BIB
A well known conjecture of Burr and Erdos asserts that the Ramsey number $r(Q_n)$ of the hypercube $Q_n$ on $2^n$ vertices is of the order $O(2^n)$. In this paper, we show that $r(Q_n)=O(2^{2n-c n})$ for a universal constant $c>0$, improving upon the previous best known bound $r(Q_n)=O(2^{2n})$, due to Conlon, Fox and Sudakov.
Shotgun assembly of unlabeled Erdos-Renyi graphs
Published
• View Publication
• BIB
Given a positive integer $n$, an unlabeled graph $G$ on $n$ vertices, and a vertex $v$ of $G$, let $N_G(v)$ be the subgraph of $G$ induced by vertices of $G$ of distance at most one from $v$. We show that there are universal constants $C,c>0$ with the following property. Let the sequence $(p_n)_{n=1}^\infty$ satisfy $n^{-1/2}\log^C n\leq p_n\leq c$. For each $n$, let $Γ_n$ be an unlabeled $G(n,p_n)$ Erdös-Rényi graph. Then with probability $1-o_n(1)$, any unlabeled graph $\tilde Γ_n$ on $n$ vertices with $\{N_{\tilde Γ_n}(v)\}_{v}=\{N_{Γ_n}(v)\}_{v}$ must coincide with $Γ_n$. This establishes $\tilde Θ(n^{-1/2})$ as the transition range for the density parameter $p_n$ between reconstructability and non-reconstructability of Erdös-Rényi graphs from their $1$-neighborhoods, and resolves a problem of Gaudio and Mossel.
Sharp Poincaré and log-Sobolev inequalities for the switch chain on regular bipartite graphs
Published
• View Publication
• BIB
Consider the switch chain on the set of $d$-regular bipartite graphs on $n$ vertices with $3\leq d\leq n^{c}$, for a small universal constant $c>0$. We prove that the chain satisfies a Poincaré inequality with a constant of order $O(nd)$; moreover, when $d$ is fixed, we establish a log-Sobolev inequality for the chain with a constant of order $O_d(n\log n)$. We show that both results are optimal. The Poincaré inequality implies that in the regime $3\leq d\leq n^c$ the mixing time of the switch chain is at most $O\big((nd)^2 \log(nd)\big)$, improving on the previously known bound $O\big((nd)^{13} \log(nd)\big)$ due to Kannan, Tetali and Vempala and $O\big(n^7d^{18} \log(nd)\big)$ obtained by Dyer et al. The log-Sobolev inequality that we establish for constant $d$ implies a bound $O(n\log^2 n)$ on the mixing time of the chain which, up to the $\log n$ factor, captures a conjectured optimal bound. Our proof strategy relies on building, for any fixed function on the set of $d$-regular bipartite simple graphs, an appropriate extension to a function on the set of multigraphs given by the configuration model. We then establish a comparison procedure with the well studied random transposition model in order to obtain the corresponding functional inequalities. While our method falls into a rich class of comparison techniques for Markov chains on different state spaces, the crucial feature of the method - dealing with chains with a large distortion between their stationary measures - is a novel addition to the theory.
Outliers in spectrum of sparse Wigner matrices
In this paper, we study the effect of sparsity on the appearance of outliers in the semi-circular law. Let $(W_n)_{n=1}^\infty$ be a sequence of random symmetric matrices such that each $W_n$ is $n\times n$ with i.i.d entries above and on the main diagonal equidistributed with the product $b_nξ$, where $ξ$ is a real centered uniformly bounded random variable of unit variance and $b_n$ is an independent Bernoulli random variable with a probability of success $p_n$. Assuming that $\lim\limits_{n\to\infty}n p_n=\infty$, we show that for the random sequence $(ρ_n)_{n=1}^\infty$ given by $$ρ_n:=θ_n+\frac{n p_n}{θ_n},\quad θ_n:=\sqrt{\max\big(\max\limits_{i\leq n}\|{\rm Row_i}(W_n)\|_2^2-np_n,n p_n\big)},$$ the ratio $\frac{\|W_n\|}{ρ_n}$ converges to one in probability. A non-centered counterpart of the theorem allows to obtain asymptotic expressions for eigenvalues of the Erdős--Renyi graphs, which were unknown in the regime $n p_n=Θ(\log n)$. In particular, denoting by $A_n$ the adjacency matrix of $\mathcal{G}(n,p_n)$ and by $λ_{|k|}(A_n)$ its $k$-th largest (by the absolute value) eigenvalue, under the assumptions $\lim\limits_{n\to\infty }n p_n=\infty$ and $\lim\limits_{n\to\infty}p_n=0$ we have:
-(No non-trivial outliers) If $\liminf\frac{n p_n}{\log n}\geq\frac{1}{\log (4/e)}$ then for any fixed $k\geq2$, $\frac{|λ_{|k|}(A_n)|}{2\sqrt{n p_n}}$ converges to $1$ in probability.
-(Outliers) If $\limsup\frac{n p_n}{\log n}<\frac{1}{\log (4/e)}$ then there is $\varepsilon>0$ such that for any $k\in\mathbb{N}$, we have $\lim\limits_{n\to\infty}\mathbb{P}\Big\{\frac{|λ_{|k|}(A_n)|}{2\sqrt{n p_n}}>1+\varepsilon\Big\}=1$.
On a conceptual level, our result highlights similarities in appearance of outliers in spectrum of sparse matrices and the so-called BBP phase transition phenomenon in deformed Wigner matrices.
Singularity of random Bernoulli matrices
For each $n$, let $M_n$ be an $n\times n$ random matrix with independent $\pm 1$ entries. We show that ${\mathbb P}\{\mbox{$M_n$ is singular}\}=(1/2+o_n(1))^n$, which settles an old problem. Some generalizations are considered.
The smallest singular value of a shifted $d$-regular random square matrix
Published in Probability Theory and Related Fields, 2018
• View Publication
• BIB
We derive a lower bound on the smallest singular value of a random $d$-regular matrix, that is, the adjacency matrix of a random $d$-regular directed graph. More precisely, let $C_1<d< c_1 n/\log^2 n$ and let $\mathcal{M}_{n,d}$ be the set of all $0/1$-valued square $n\times n$ matrices such that each row and each column of a matrix $M\in \mathcal{M}_{n,d}$ has exactly $d$ ones. Let $M$ be uniformly distributed on $\mathcal{M}_{n,d}$. Then the smallest singular value $s_{n} (M)$ of $M$ is greater than $c_2 n^{-6}$ with probability at least $1-C_2\log^2 d/\sqrt{d}$, where $c_1$, $c_2$, $C_1$, and $C_2$ are absolute positive constants independent of any other parameters.
On the norm of a random jointly exchangeable matrix
Published in Journal of Theoretical Probability, 2018
• View Publication
• BIB
In this note, we show that the norm of an $n\times n$ random jointly exchangeable matrix with zero diagonal can be estimated in terms of the norm of its $n/2\times n/2$ submatrix located in the top right corner. As a consequence, we prove a relation between the second largest singular values of a random matrix with constant row and column sums and its top right $n/2\times n/2$ submatrix. The result has an application to estimating the spectral gap of random undirected $d$-regular graphs in terms of the second singular value of {\it directed} random graphs with predefined degree sequences.
The spectral gap of dense random regular graphs
Published in Annals of Probability, Volume 47, Number 1 (2019), 362-419
• View Publication
• BIB
For any $α\in (0,1)$ and any $n^α\leq d\leq n/2$, we show that $λ(G)\leq C_α\sqrt{d}$ with probability at least $1-\frac{1}{n}$, where $G$ is the uniform random $d$-regular graph on $n$ vertices, $λ(G)$ denotes its second largest eigenvalue (in absolute value) and $C_α$ is a constant depending only on $α$. Combined with earlier results in this direction covering the case of sparse random graphs, this completely settles the problem of estimating the magnitude of $λ(G)$, up to a multiplicative constant, for all values of $n$ and $d$, confirming a conjecture of Vu. The result is obtained as a consequence of an estimate for the second largest singular value of adjacency matrices of random {\it directed} graphs with predefined degree sequences. As the main technical tool, we prove a concentration inequality for arbitrary linear forms on the space of matrices, where the probability measure is induced by the adjacency matrix of a random directed graph with prescribed degree sequences. The proof is a non-trivial application of the Freedman inequality for martingales, combined with boots-trapping and tensorization arguments. Our method bears considerable differences compared to the approach used by Broder, Frieze, Suen and Upfal (1999) who established the upper bound for $λ(G)$ for $d=o(\sqrt{n})$, and to the argument of Cook, Goldstein and Johnson (2015) who derived a concentration inequality for linear forms and estimated $λ(G)$ in the range $d= O(n^{2/3})$ using size-biased couplings.