random
6952 papers tagged with this keyword
Tilted biorthogonal ensembles, Grothendieck random partitions, and determinantal tests
We study probability measures on partitions based on symmetric Grothendieck polynomials. These deformations of Schur polynomials introduced in the K-theory of Grassmannians share many common properties. Our Grothendieck measures are analogs of the Schur measures on partitions introduced by Okounkov (arXiv:math/9907127 [math.RT]). Despite the similarity of determinantal formulas for the probability weights of Schur and Grothendieck measures, we demonstrate that Grothendieck measures are \emph{not} determinantal point processes. This question is related to the principal minor assignment problem in algebraic geometry, and we employ a determinantal test first obtained by Nanson in 1897 for the $4\times4$ problem. We also propose a procedure for getting Nanson-like determinantal tests for matrices of any size $n\ge4$ which appear new for $n\ge 5$.
By placing the Grothendieck measures into a new framework of tilted biorthogonal ensembles generalizing a rich class of determinantal processes introduced by Borodin (arXiv:math/9804027 [math.CA]), we identify Grothendieck random partitions as a cross-section of a Schur process, a determinantal process in two dimensions. This identification expresses the correlation functions of Grothendieck measures through sums of Fredholm determinants, which are not immediately suitable for asymptotic analysis. A more direct approach allows us to obtain a limit shape result for the Grothendieck random partitions. The limit shape curve is not particularly explicit as it arises as a cross-section of the limit shape surface for the Schur process. The gradient of this surface is expressed through the argument of a complex root of a cubic equation.
The Random Turán Problem for Theta Graphs
Given a graph $F$, we define $\operatorname{ex}(G_{n,p},F)$ to be the maximum number of edges in an $F$-free subgraph of the random graph $G_{n,p}$. Very little is known about $\operatorname{ex}(G_{n,p},F)$ when $F$ is bipartite, with essentially tight bounds known only when $F$ is either $C_4, C_6, C_{10}$, or $K_{s,t}$ with $t$ sufficiently large in terms of $s$, due to work of Füredi and of Morris and Saxton. We extend this work by establishing essentially tight bounds when $F$ is a theta graph with sufficiently many paths. Our main innovation is in proving a balanced supersaturation result for vertices, which differs from the standard approach of proving balanced supersaturation for edges.
Irreducibility of Recombination Markov Chains in the Triangular Lattice
Published
• View Publication
• BIB
In the United States, regions are frequently divided into districts for the purpose of electing representatives. How the districts are drawn can affect who's elected, and drawing districts to give an advantage to a certain group is known as gerrymandering. It can be surprisingly difficult to detect gerrymandering, but one algorithmic method is to compare a current districting plan to a large number of randomly sampled plans to see whether it is an outlier. Recombination Markov chains are often used for this random sampling: randomly choose two districts, consider their union, and split this union in a new way. This works well in practice, but the theory behind it remains underdeveloped. For example, it's not known if recombination Markov chains are irreducible, that is, if recombination moves suffice to move from any districting plan to any other.
Irreducibility of recombination Markov chains can be formulated as a graph problem: for a graph $G$, is the space of all partitions of $G$ into $k$ connected subgraphs ($k$ districts) connected by recombination moves? We consider three simply connected districts and district sizes $k_1\pm 1$ vertices, $k_2\pm 1$ vertices, and $k3\pm 1$ vertices. We prove for arbitrarily large triangular regions in the triangular lattice, recombination Markov chains are irreducible. This is the first proof of irreducibility under tight district size constraints for recombination Markov chains beyond small or trivial examples.
Probabilistic enumeration and equivalence of nonisomorphic trees
We present a new probabilistic proof of Otter's asymptotic formula for the number of unlabelled trees with a given number of vertices. We additionally prove a new approximation result, showing that the total variation distance between random Pólya trees and random unlabelled trees tends to zero when the number of vertices tends to infinity. In order to demonstrate that our approach is not restricted to trees we extend our results to tree-like classes of graphs.
Poisson-Dirichlet scaling limits of Kemp's supertrees
We determine the Gromov--Hausdorff--Prokhorov scaling limits and local limits of Kemp's $d$-dimensional binary trees and other models of supertrees. The limits exhibit a root vertex with infinite degree and are constructed by rescaling infinitely many independent stable trees or other spaces according to a function of a two-parameter Poisson--Dirichlet process and gluing them together at their roots. We discuss universality aspects of random spaces constructed in this fashion and sketch a phase diagram.
Constructions of Constant Dimension Subspace Codes
Subspace codes have important applications in random network coding. It is interesting to construct subspace codes with both sizes, and the minimum distances are as large as possible. In particular, cyclic constant dimension subspaces codes have additional properties which can be used to make encoding and decoding more efficient. In this paper, we construct large cyclic constant dimension subspace codes with minimum distances $2k-2$ and $2k$. These codes are contained in $\mathcal{G}_q(n, k)$, where $\mathcal{G}_q(n, k)$ denotes the set of all $k$-dimensional subspaces of $\mathbb{F}_{q^n}$. Consequently, some results in \cite{FW}, \cite{NXG}, and \cite{ZT} are extended.
Strongly common graphs with odd girth are cycles
A graph $H$ is called strongly common if for every coloring $φ$ of $K_n$ with two colors, the number of monochromatic copies of $H$ is at least the number of monochromatic copies of $H$ in a random coloring of $K_n$ with the same density of color classes as $φ$. In this note we prove that if a graph has odd girth but is not a cycle, then it is not strongly common. This answers a question of Chen and Ma.
The Asymptotics of the Expected Betti Numbers of Preferential Attachment Clique Complexes
The preferential attachment model is a natural and popular random graph model for a growing network that contains very well-connected ``hubs''. We study the higher-order connectivity of such a network by investigating the topological properties of its clique complex. We concentrate on the expected Betti numbers, a sequence of topological invariants of the complex related to the numbers of holes of different dimensions. We determine the asymptotic growth rates of the expected Betti numbers, and prove that the expected Betti number at dimension 1 grows linearly fast, while those at higher dimensions grow sublinearly fast. Our theoretical results are illustrated by simulations. (Changes are made in this version to generalize Proposition 14 and to streamline proofs. These changes are shown in blue.)
Further Results on Random Walk Labelings
Recently, we initiated the study of random walk labelings of graphs. These are graph labelings that are obtainable by performing a random walk on the graph, such that each vertex is labeled upon its first visit. In this work, we calculate the number of random walk labelings of several natural graph families: The wheel, fan, barbell, lollipop, tadpole, friendship, and snake graphs. Additionally, we prove several combinatorial identities that emerged during the calculations.
On the concave one-dimensional random assignment problem and Young integration theory
We investigate the one-dimensional random assignment problem in the concave case, i.e., the assignment cost is a concave power function, with exponent $0<p<1$, of the distance between $n$ source and $n$ target points, that are i.i.d. random variables with a common law on an interval. We prove that the limit of a suitable renormalization of the costs exists if the exponent $p$ is different than $1/2$. Our proof in the case $1/2<p<1$ makes use of a novel version of the Kantorovich optimal transport problem based on Young integration theory, where the difference between two measures is replaced by the weak derivative of a function with finite $q$-variation, which may be of independent interest. We also prove a similar result for the random bipartite Traveling Salesperson Problem.
Using Symbolic Computation to Explore Generalized Dyck Paths and Their Areas
We show the power of Bruno Buchberger's seminal Groebner Basis algorithm, interfaced, seamlessly, with what we call symbolic dynamical programming, to automatically generate algebraic equations satisfied by the generating functions enumerating so-called Generalized Dyck Walks, i.e. 2D walks that start and end on the x-axis, and never dip below it, for an arbitrary set of steps. More impressively, we combine it with calculus (that Maple knows very well!), to automatically compute generating functions for the sum-of-the-areas of these generalized Dyck paths, and even for the sum of any given power of the areas, enabling us to get statistical information about the area under a random generalized Dyck path.
Linear Eulerian Extensions of Inhomogenous Random Graphs
The Eulerian extension number of any graph~\(H\) (i.e. the minimum number of edges needed to be added to make~\(H\) Eulerian) is at least~\(t(H),\) half the number of odd degree vertices of~\(H.\) In this paper we consider an inhomogenous random graph~\(G\) whose edge probabilities need not all be the same and use an iterative probabilistic method to obtain sufficient conditions for the Eulerian extension number of~\(G\) to grow \emph{linearly} with~\(t(G).\) We derive our conditions in terms of the average edge probabilities and edge density and also briefly illustrate our result with an example.
Random processes for generating task-dependency graphs
We investigate random processes for generating task-dependency graphs of order $n$ with $m$ edges and a specified number of initial vertices and terminal vertices. In order to do so, we consider two random processes for generating task-dependency graphs that can be combined to accomplish this task. In the $(x, y)$ edge-removal process, we start with a maximally connected task-dependency graph and remove edges uniformly at random as long as they do not cause the number of initial vertices to exceed $x$ or the number of terminal vertices to exceed $y$. In the $(x, y)$ edge-addition process, we start with an empty task-dependency graph and add edges uniformly at random as long as they do not cause the number of initial vertices to be less than $x$ or the number of terminal vertices to be less than $y$. In the $(x, y)$ edge-addition process, we halt if there are exactly $x$ initial vertices and $y$ terminal vertices. For both processes, we determine the values of $x$ and $y$ for which the resulting task-dependency graph is guaranteed to have exactly $x$ initial vertices and $y$ terminal vertices, and we also find the extremal values for the number of edges in the resulting task-dependency graphs as a function of $x$, $y$, and the number of vertices. Furthermore, we asymptotically bound the expected number of edges in the resulting task-dependency graphs. Finally, we define a random process using only edge-addition and edge-removal, and we show that with high probability this random process generates an $(x, y)$ task-dependency graph of order $n$ with $m$ edges.
A dynamical approach to spanning and surplus edges of random graphs
Consider a finite inhomogeneous random graph running in continuous time, where each vertex has a mass, and the edge that links any pair of vertices appears with a rate equal to the product of their masses. The simultaneous breadth-first-walk introduced by Limic (2019) is extended in order to account for the surplus edge data in addition to the spanning edge data. Two different graph-based representations of the multiplicative coalescent, with different advantages and drawbacks, are discussed in detail. A canonical multi-graph from Bhamidi, Budhiraja and Wang (2014) naturally emerges. The presented framework will facilitate the understanding of scaling limits with surplus edges for near-critical random graphs in the domain of attraction of general (not necessarily standard) eternal augmented multiplicative coalescent.
A note on monotone subsequences and the RS image of invariant random permutations with macroscopic number of fixed points
The work of Vershik and Kerov [1977], Logan and Shepp [1977] established that the shape of the scaled random young diagram in Russian notation, as determined by the Plancherel measure, converges to a deterministic shape. In this article, we focus on the scenario where the number of fixed points is substantial. We provide evidence that, subject to specific requirements on the total number of cycles, the limiting shape is a scaled version of the Vershik-Kerov-Logan-Shepp limiting shape. Additionally, we identify certain limiting regimes that resemble those in Chapuy, Louf, and Walsh [2022]. Furthermore, we enhance the existing results on Tracy-Widom universality classes for $β\in \{ 1, 2, 4\}$ for monotone subsequences.
Random Algebraic Graphs and Their Convergence to Erdos-Renyi
A random algebraic graph is defined by a group $G$ with a uniform distribution over it and a connection $σ:G\longrightarrow[0,1]$ with expectation $p,$ satisfying $σ(g)=σ(g^{-1}).$ The random graph $\mathsf{RAG}(n,G,p,σ)$ with vertex set $[n]$ is formed as follows. First, $n$ independent vectors $x_1,\ldots,x_n$ are sampled uniformly from $G.$ Then, vertices $i,j$ are connected with probability $σ(x_ix_j^{-1}).$ This model captures random geometric graphs over the sphere and the hypercube, certain regimes of the stochastic block model, and random subgraphs of Cayley graphs. The main question of interest to the current paper is: when is a random algebraic graph statistically and/or computationally distinguishable from $\mathsf{G}(n,p)$? Our results fall into two categories. 1) Geometric. We focus on the case $G =\{\pm1\}^d$ and use Fourier-analytic tools. For hard threshold connections, we match [LMSY22b] for $p = ω(1/n)$ and for $1/(r\sqrt{d})$-Lipschitz connections we extend the results of [LR21b] when $d = Ω(n\log n)$ to the non-monotone setting. We study other connections such as indicators of interval unions and low-degree polynomials. 2) Algebraic. We provide evidence for an exponential statistical-computational gap. Consider any finite group $G$ and let $A\subseteq G$ be a set of elements formed by including each set of the form $\{g, g^{-1}\}$ independently with probability $1/2.$ Let $Γ_n(G,A)$ be the distribution of random graphs formed by taking a uniformly random induced subgraph of size $n$ of the Cayley graph $Γ(G,A).$ Then, $Γ_n(G,A)$ and $\mathsf{G}(n,1/2)$ are statistically indistinguishable with high probability over $A$ if and only if $\log|G|\gtrsim n.$ However, low-degree polynomial tests fail to distinguish $Γ_n(G,A)$ and $\mathsf{G}(n,1/2)$ with high probability over $A$ when $\log |G|=\log^{Ω(1)}n.$
Isomorphisms between dense random graphs
Published in Combinatorica 45 (2025), Article 35, 42 pages
• View Publication
• BIB
We consider two variants of the induced subgraph isomorphism problem for two independent binomial random graphs with constant edge-probabilities p_1,p_2. In particular, (i) we prove a sharp threshold result for the appearance of G_{n,p_1} as an induced subgraph of G_{N,p_2}, (ii) we show two-point concentration of the size of the maximum common induced subgraph of G_{N, p_1} and G_{N,p_2}, and (iii) we show that the number of induced copies of G_{n,p_1} in G_{N,p_2} has an unusual limiting distribution.
These results confirm simulation-based predictions of McCreesh, Prosser, Solnon and Trimble, and resolve several open problems of Chatterjee and Diaconis. The proofs are based on careful refinements of the first and second moment method, using extra twists to (a) take some non-standard behaviors into account, and (b) work around the large variance issues that prevent standard applications of these methods.
The distributions under two species-tree models of the total number of ancestral configurations for matching gene trees and species trees
Given a gene-tree labeled topology $G$ and a species tree $S$, the "ancestral configurations" at an internal node $k$ of $S$ represent the combinatorially different sets of gene lineages that can be present at $k$ when all possible realizations of $G$ in $S$ are considered. Ancestral configurations have been introduced as a data structure for evaluating the conditional probability of a gene-tree labeled topology given a species tree, and their enumeration assists in describing the complexity of this computation. In the case that the gene-tree labeled topology $G=t$ matches that of the species tree $S$, by techniques of analytic combinatorics, we study distributional properties of the "total" number of ancestral configurations measured across the different nodes of a random labeled topology $t$ selected under the uniform and the Yule probability models. Under both of these probabilistic scenarios, we show that the total number $T_n$ of ancestral configurations of a random labeled topology of $n$ taxa asymptotically follows a lognormal distribution. Over uniformly distributed labeled topologies, the asymptotic growth of the mean and the variance of $T_n$ are found to satisfy $\mathbb{E}_{\rm U}[T_n] \sim 2.449 \cdot 1.333^n$ and $\mathbb{V}_{\rm U}[T_n] \sim 5.050 \cdot 1.822^n$, respectively. Under the Yule model, which assigns higher probabilities to more balanced labeled topologies, we obtain the mean $\mathbb{E}_{\rm Y}[T_n] \sim 1.425^n$ and the variance $\mathbb{V}_{\rm Y}[T_n] \sim 2.045^n$.
Random Turán theorem for expansions of spanning subgraphs of tight trees
The $r$-expansion of a $k$-uniform hypergraph $H$, denoted by $H^{(+r)}$, is an $r$-uniform hypergraph obtained by enlarging each $k$-edge of $H$ with a set of $r-k$ vertices of degree one. The random Turán number $\mathrm{ex}(G^r_{n,p},H)$ is the maximum number of edges in an $H$-free subgraph of $G^r_{n,p}$, where $G^r_{n,p}$ is the Erdős-Rényi random $r$-graph with parameter $p$. In this paper, we prove an upper bound for $\mathrm{ex}(G^r_{n,p},H)$ when $H$ belongs to a large family of $r$-partite $r$-graphs: the $r$-expansion of spanning subgraphs of tight trees. This upper bound is essentially tight for at least the following two families of hypergraphs.
1. Our upper bounds are essentially tight for expansions of $K^{k-1}_{k}$, the complete $(k-1)$-graph on $k$ vertices. The proof of the lower bound makes use of a recent construction of Gowers and Janzer generalizing the famous Ruzsa-Szemerëdi construction. In particular, when $k=3$, this answers a question of the current author, Spiro and Verstraëte concerning the random Turán number of linear triangle.
2. Let $T$ be a tight tree such that the intersection of all edges of $T$ is empty. Simple construction shows that the upper bounds we have for expansions of $T$ are essentially tight.
The main technical contribution of this paper is a new way to obtain balanced supersaturation results for expansions of hypergraphs: we combine two ideas, one of Mubayi-Yepremyan and another of Balogh-Narayanan-Skokan, via codegree dichotomy. We note that neither of these two ideas alone would be enough to recover results in this paper.
Moderate deviations of triangle counts in sparse Erdős-Rényi random graphs $G(n,m)$ and $G(n,p)$
We consider the question of determining the probability of triangle count deviations in the Erdős-Rényi random graphs $G(n,m)$ and $G(n,p)$ with densities larger than $n^{-1/2}(\log{n})^{1/2}$. In particular, we determine the log probability $\log\mathbb{P}(N_{\triangle}(G)\, >\, (1+δ)p^3n^3)$ up to a constant factor across essentially the entire range of possible deviations, in both the $G(n,m)$ and $G(n,p)$ model. For the $G(n,p)$ model we also prove a stronger result, up to a $(1+o(1))$ factor, in the non-localised regime. We also obtain some results for the lower tail and for counts of cherries (paths of length $2$).