arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

6952 papers tagged with this keyword
Fast and Slow Mixing of the Kawasaki Dynamics on Bounded-Degree Graphs
Published in Random Structures & Algorithms. 67 (2025), no.4, e70038 • View PublicationBIB
We study the worst-case mixing time of the global Kawasaki dynamics for the fixed-magnetization Ising model on the class of graphs of maximum degree $Δ$. Proving a conjecture of Carlson, Davies, Kolla, and Perkins, we show that below the tree uniqueness threshold, the Kawasaki dynamics mix rapidly for all magnetizations. Disproving a conjecture of Carlson, Davies, Kolla, and Perkins, we show that the regime of fast mixing does not extend throughout the regime of tractability for this model: there is a range of parameters for which there exist efficient sampling algorithms for the fixed-magnetization Ising model on max-degree $Δ$ graphs, but the Kawasaki dynamics can take exponential time to mix. Our techniques involve showing spectral independence in the fixed-magnetization Ising model and proving a sharp threshold for the existence of multiple metastable states in the Ising model with external field on random regular graphs.
Fast Mixing in Sparse Random Ising Models
Motivated by the community detection problem in Bayesian inference, as well as the recent explosion of interest in spin glasses from statistical physics, we study the classical Glauber dynamics for sampling from Ising models with sparse random interactions. It is now well-known that when the interaction matrix has spectral diameter less than $1$, Glauber dynamics mixes in $O(n\log n)$ steps. Unfortunately, such criteria fail dramatically for interactions supported on arguably the most well-studied sparse random graph: the Erdős--Rényi random graph $G(n,d/n)$, due to the presence of almost linearly many outlier eigenvalues of unbounded magnitude. We prove that for the \emph{Viana--Bray spin glass}, where the interactions are supported on $G(n,d/n)$ and randomly assigned $\pmβ$, Glauber dynamics mixes in $n^{1+o(1)}$ time with high probability as long as $β\le O(1/\sqrt{d})$, independent of $n$. We further extend our results to random graphs drawn according to the $2$-community stochastic block model, as well as when the interactions are given by a "centered" version of the adjacency matrix. The latter setting is particularly relevant for the inference problem in community detection. Indeed, we use this to show that Glauber dynamics succeeds at recovering communities in the stochastic block model in a companion paper [LMR+24]. The primary technical ingredient in our proof is showing that with high probability, a sparse random graph can be decomposed into two parts -- a \emph{bulk} which behaves like a graph with bounded maximum degree and a well-behaved spectrum, and a \emph{near-forest} with favorable pseudorandom properties. We then use this decomposition to design a localization procedure that interpolates to simpler Ising models supported only on the near-forest, and then execute a pathwise analysis to establish a modified log-Sobolev inequality.
2024-05-09
Absolute zeta functions and periodicity of quantum walks on cycles
Published in Quantum Information and Computation, Vol. 24, No. 11&12 (2024) 901--916 • View PublicationBIB
The quantum walk is a quantum counterpart of the classical random walk. On the other hand, absolute zeta functions can be considered as zeta functions over $\mathbb{F}_1$. This study presents a connection between quantum walks and absolute zeta functions. In this paper, we focus on Hadamard walks and $3$-state Grover walks on cycle graphs. The Hadamard walks and the Grover walks are typical models of the quantum walks. We consider the periods and zeta functions of such quantum walks. Moreover, we derive the explicit forms of the absolute zeta functions of corresponding zeta functions. Also, it is shown that our zeta functions of quantum walks are absolute automorphic forms.
2024-05-09 v2
The largest subgraph without a forbidden induced subgraph
We initiate the systematic study of the following Turán-type question. Suppose $Γ$ is a graph with $n$ vertices such that the edge density between any pair of subsets of vertices of size at least $t$ is at most $1 - c$, for some $t$ and $c > 0$. What is the largest number of edges in a subgraph $G \subseteq Γ$ which does not contain a fixed graph $H$ as an induced subgraph or, more generally, which belongs to a hereditary property $\mathcal{P}$? This provides a common generalization of two recently studied cases, namely $Γ$ being a (pseudo-)random graph and a graph without a large complete bipartite subgraph. We focus on the interesting case where $H$ is a bipartite graph. We determine the answer up to a constant factor with respect to $n$ and $t$, for certain bipartite $H$ and for $Γ$ either a dense random graph or a Paley graph with a square number of vertices. In particular, our bounds match if $H$ is a tree, or if one part of $H$ has $d$ vertices complete to the other part, all other vertices in that part have degree at most $d$, and the other part has sufficiently many vertices. As applications of the latter result, we answer a question of Alon, Krivelevich, and Samotij on the largest subgraph with a hereditary property which misses a bipartite graph, and determine up to a constant factor the largest number of edges in a string subgraph of $Γ$. The proofs are based on a variant of the dependent random choice and a novel approach for finding induced copies by inductively defining probability distributions supported on induced copies of smaller subgraphs.
2024-05-08
On linear-combinatorial problems associated with subspaces spanned by $\{\pm 1\}$-vectors
A complete answer to the question about subspaces generated by $\{\pm 1\}$-vectors, which arose in the work of I.Kanter and H.Sompolinsky on associative memories, is given. More precisely, let vectors $v_1, \ldots , v_p,$ $p\leq n-1,$ be chosen at random uniformly and independently from $\{\pm 1\}^n \subset {\bf R}^n.$ Then the probability ${\mathbb P}(p, n)$ that $$span \ \langle v_1, \ldots , v_p \rangle \cap \left\{ \{\pm 1\}^n \setminus \{\pm v_1, \ldots , \pm v_p\}\right\} \ne \emptyset \ $$ is shown to be $$4{p \choose 3}\left(\frac{3}{4}\right)^n + O\left(\left(\frac{5}{8} + o_n(1)\right)^n\right) \quad \mbox{as} \quad n\to \infty,$$ where the constant implied by the $O$-notation does not depend on $p$. The main term in this estimate is the probability that some 3 vectors $v_{j_1}, v_{j_2}, v_{j_3}$ of $v_j$, $j= 1, \ldots , p,$ have a linear combination that is a $\{\pm 1\}$-vector different from $\pm v_{j_1}, \pm v_{j_2}, \pm v_{j_3}. $
2024-05-07
Multiple consecutive runs of multi-state trials: distributions of $(k_1, k_2, \dots, k_\ell)$ patterns
Published in Journal of Computational and Applied Mathematics, Volume 403, 113846 (2022) • View PublicationBIB
The pattern $(k_1, k_2, \dots, k_\ell)$ is defined to have at least $k_1$ consecutive $1$'s followed by at least $k_2$ consecutive $2$'s, $\dots$, followed by at least $k_\ell$ consecutive $\ell$'s. By iteratively applying the method that was developed previously to decouple the combinatorial complexity involved in studying complicated patterns in random sequences, the distribution of pattern $(k_1, k_2, \dots, k_\ell)$ is derived for arbitrary $\ell$. Numerical examples are provided to illustrate the results.
2024-05-07
The Large Deviation Principle for $W$-random spectral measures
The $W$-random graphs provide a flexible framework for modeling large random networks. Using the Large Deviation Principle (LDP) for $W$-random graphs from [9], we prove the LDP for the corresponding class of random symmetric Hilbert-Schmidt integral operators. Our main result describes how the eigenvalues and the eigenspaces of the integral operator are affected by the large deviations in the underlying random graphon. To prove the LDP, we demonstrate continuous dependence of the spectral measures associated with integral operators on the underlying graphons and use the Contraction Principle. To illustrate our results, we obtain leading order asymptotics of the eigenvalues of the integral operators corresponding to certain random graph sequences. These examples suggest several representative scenarios of how the eigenvalues and the eigenspaces of the integral operators are affected by large deviations. Potential implications of these observations for bifurcation analysis of Dynamical Systems and Graph Signal Processing are indicated.
2024-05-07 v2
A Constructive Winning Maker Strategy in the Maker-Breaker $C_4$-Game
Maker-Breaker subgraph games are among the most famous combinatorial games. For given $n,q \in \mathbb{N}$ and a subgraph $C$ of the complete graph $K_n$, the two players, called Maker and Breaker, alternately claim edges of $K_n$. In each round of the game Maker claims one edge and Breaker is allowed to claim up to $q$ edges. If Maker is able to claim all edges of a copy of $C$, he wins the game. Otherwise Breaker wins. In this work we introduce the first constructive strategy for Maker for the $C_4$-Maker-Breaker game and show that he can win the game if $q < 0.16 n^{2/3}$. According to the theorem of Bednarska and Luczak (2000) $n^{2/3}$ is asymptotically optimal for this game, but the constant given there for a random Maker strategy is magnitudes apart from our constant 0.16.
2024-05-07
Isomorphisms between random $d$-hypergraphs
We characterize the size of the largest common induced subgraph of two independent random uniform $d$-hypergraphs of different sizes with $d\geq 3$. More precisely, its distribution is asymptotically concentrated on two points, and we obtain as a consequence a phase transition for the inclusion of the smallest hypergraph in the largest one. This generalizes to uniform random $d$-hypergraphs the results of Chatterjee and Diaconis for uniform random graphs. Our proofs rely on the first and second moment methods.
Counting Subnetworks Under Gene Duplication in Genetic Regulatory Networks
Gene duplication is a fundamental evolutionary mechanism that contributes to biological complexity and diversity (Fortna et al., 2004). Traditionally, research has focused on the duplication of gene sequences (Zhang, 1914). However, evidence suggests that the duplication of regulatory elements may also play a significant role in the evolution of genomic functions (Teichmann and Babu, 2004; Hallin and Landry, 2019). In this work, the evolution of regulatory relationships belonging to gene-specific-substructures in a GRN are modeled. In the model, a network grows from an initial configuration by repeatedly choosing a random gene to duplicate. The likelihood that the regulatory relationships associated with the selected gene are retained through duplication is determined by a vector of probabilities. Occurrences of gene-family-specific substructures are counted under the gene duplication model. In this thesis, gene-family-specific substructures are referred to as subnetwork motifs. These subnetwork motifs are motivated by network motifs which are patterns of interconnections that recur more often in a specialized network than in a random network (Milo et al., 2002). Subnetwork motifs differ from network motifs in the way that subnetwork motifs are instances of gene-family-specific substructures while network motifs are isomorphic substructures. These subnetwork motifs are counted under Full and Partial Duplication, which differ in the way in which regulation relationships are inherited. Full duplication occurs when all regulatory links are inherited at each duplication step, and Partial Duplication occurs when regulation inheritance varies at each duplication step. Moments for the number of occurrences of subnetwork motifs are determined in each model. The results presented offer a method for discovering subnetwork motifs that are significant in a GRN under gene duplication.
The number of random 2-SAT solutions is asymptotically log-normal
We prove that throughout the satisfiable phase, the logarithm of the number of satisfying assignments of a random 2-SAT formula satisfies a central limit theorem. This implies that the log of the number of satisfying assignments exhibits fluctuations of order $\sqrt n$, with $n$ the number of variables. The formula for the variance can be evaluated effectively. By contrast, for numerous other random constraint satisfaction problems the typical fluctuations of the logarithm of the number of solutions are {\em bounded} throughout all or most of the satisfiable regime.
2024-05-05
Some sharp lower bounds for the bipartite Turán number of theta graphs
Published in Bulletin of the Hellenic Mathematical Society, Vol. 68, 2024 (1-9) • Search Publication
We expand Conlon's random algebraic construction to show that for any odd number $k \geq 3$ exists a natural number $c_k$ (the same as Conlon's) such that $\operatorname{ex}(n^a,n,θ_{k,c_k}) = Ω_{k,a}((n^{1 + a})^{\frac{k + 1}{2k}})$, with $a \in [\frac{k - 1}{k + 1}, 1)$. Where given a graph $H$, we denote by $\operatorname{ex}(n,m,H)$ the maximum number of edges an $H-$free bipartite graph can have when the cardinalities of its parts are $n$ and $m$. Also, we denote with $θ_{k,l}$ the graph where two vertices are connected through $l$ disjoint paths of length $k$.
Saturation in Random Hypergraphs
Published in Combinator. Probab. Comp. 35 (2026) 40-58 • View PublicationBIB
Let $K^r_n$ be the complete $r$-uniform hypergraph on $n$ vertices, that is, the hypergraph whose vertex set is $[n]:=\{1,2,...,n\}$ and whose edge set is $\binom{[n]}{r}$. We form $G^r(n,p)$ by retaining each edge of $K^r_n$ independently with probability $p$. An $r$-uniform hypergraph $H\subseteq G$ is $F$-saturated if $H$ does not contain any copy of $F$, but any missing edge of $H$ in $G$ creates a copy of $F$. Furthermore, we say that $H$ is weakly $F$-saturated in $G$ if $H$ does not contain any copy of $F$, but the missing edges of $H$ in $G$ can be added back one-by-one, in some order, such that every edge creates a new copy of $F$. The smallest number of edges in an $F$-saturated hypergraph in $G$ is denoted by $sat(G,F)$, and in a weakly $F$-saturated hypergraph in $G$ by $wsat(G,F)$. In 2017, Korándi and Sudakov initiated the study of saturation in random graphs, showing that for constant $p$, with high probability $sat(G(n,p),K_s)=(1+o(1))n\log_{\frac{1}{1-p}}n$, and $wsat(G(n,p),K_s)=wsat(K_n,K_s)$. Generalising their results, in this paper, we solve the suturation problem for random hypergraphs for every $2\le r < s$ and constant $p$.
A divisor generating q-series identity and its applications to probability theory and random graphs
In I981, Uchimura studied a divisor generating $q$-series that has applications in probability theory and in the analysis of data structures, called heaps. Mainly, he proved the following identity. For $|q|<1$, \begin{equation*} \sum_{n=1}^\infty n q^n (q^{n+1})_\infty =\sum_{n=1}^{\infty} \frac{(-1)^{n-1} q^{\frac{n(n+1)}{2} } }{(1-q^n) ( q)_n } = \sum_{n=1}^{\infty} \frac{ q^n }{1-q^n}. \end{equation*} Over the years, this identity has been generalized by many mathematicians in different directions. Uchimura himself in 1987, Dilcher (1995), Andrews-Crippa-Simon (1997), and recently Gupta-Kumar (2021) found a generalization of the aforementioned identity. Any generalization of the right most expression of the above identity, we name as divisor-type sum, whereas a generalization of the middle expression we say Ramanujan-type sum, and any generalization of the left most expression we refer it as Uchimura-type sum. Quite surprisingly, Simon, Crippa and Collenberg (1993) showed that the same divisor generating function has a connection with random acyclic digraphs. One of the main themes of this paper is to study these different generalizations and present a unified theory. We also discuss applications of these generalized identities in probability theory for the analysis of heaps and random acyclic digraphs.
2024-05-03
Upper tails of subgraph counts in directed random graphs
The upper tail problem in a sparse Erdős-Rényi graph asks for the probability that the number of copies of some fixed subgraph exceeds its expected value by a constant factor. We study the analogous problem for oriented subgraphs in directed random graphs. By adapting the proof of Cook, Dembo, and Pham, we reduce this upper tail problem to the asymptotic of a certain variational problem over edge weighted directed graphs. We give upper and lower bounds for the solution to the corresponding variational problem, which differ by a constant factor of at most $2$. We provide a host of subgraphs where the upper and lower bounds coincide, giving the solution to the upper tail problem. Examples of such digraphs include triangles, stars, directed $k$-cycles, and balanced digraphs.
2024-05-03
Equal Requests are Asymptotically Hardest for Data Recovery
In a distributed storage system serving hot data, the data recovery performance becomes important, captured e.g. by the service rate. We give partial evidence for it being hardest to serve a sequence of equal user requests (as in PIR coding regime) both for concrete and random user requests and server contents. We prove that a constant request sequence is locally hardest to serve: If enough copies of each vector are stored in servers, then if a request sequence with all requests equal can be served then we can still serve it if a few requests are changed. For random iid server contents, with number of data symbols constant (for simplicity) and the number of servers growing, we show that the maximum number of user requests we can serve divided by the number of servers we need approaches a limit almost surely. For uniform server contents, we show this limit is 1/2, both for sequences of copies of a fixed request and of any requests, so it is at least as hard to serve equal requests as any requests. For iid requests independent from the uniform server contents the limit is at least 1/2 and equal to 1/2 if requests are all equal to a fixed request almost surely, confirming the same. As a building block, we deduce from a 1952 result of Marshall Hall, Jr. on abelian groups, that any collection of half as many requests as coded symbols in the doubled binary simplex code can be served by this code. This implies the fractional version of the Functional Batch Code Conjecture that allows half-servers.
2024-05-02 v3
Fringe trees of Patricia tries, compressed binary search trees, and three other random full binary trees
We study the distribution of fringe trees in Patricia tries (extending earlier results by Ischebeck (2025)) and compressed binary search trees; both cases are random binary trees that have been compressed by deleting nodes of outdegree 1 so that they are random full binary trees. The main results are central limit theorems for the number of fringe trees of a given type, which imply quenched and annealed limit results for the fringe tree distribution; for Patricia tries, this is complicated by periodic oscillations in the usual manner. We also consider extended fringe trees. The results are derived from earlier results for uncompressed tries and binary search trees. In the case of compressed binary search trees, it seems difficult to give a closed formula for the asymptotic fringe tree distribution, but we provide a recursion and give examples. For comparison, we give also results, simpler and partly known, for three other models of random full binary trees: the extended binary search tree, the critical beta-spltting random tree, and the uniform random full binary tree.
On one-orbit cyclic subspace codes of $\mathcal{G}_q(n,3)$
Subspace codes have recently been used for error correction in random network coding. In this work, we focus on one-orbit cyclic subspace codes. If $S$ is an $\mathbb{F}_q$-subspace of $\mathbb{F}_{q^n}$, then the one-orbit cyclic subspace code defined by $S$ is \[ \mathrm{Orb}(S)=\{αS \colon α\in \mathbb{F}_{q^n}^*\}, \]where $αS=\lbrace αs \colon s\in S\rbrace$ for any $α\in \mathbb{F}_{q^n}^*$. Few classification results of subspace codes are known, therefore it is quite natural to initiate a classification of cyclic subspace codes, especially in the light of the recent classification of the isometries for cyclic subspace codes. We consider three-dimensional one-orbit cyclic subspace codes, which are divided into three families: the first one containing only $\mathrm{Orb}(\mathbb{F}_{q^3})$; the second one containing the optimum-distance codes; and the third one whose elements are codes with minimum distance $2$. We study inequivalent codes in the latter two families.
2024-05-02
The m-th Longest Runs of Multivariate Random Sequences
Published in Ann Inst Stat Math 69, 497-512 (2017) • View PublicationBIB
The distributions of the $m$-th longest runs of multivariate random sequences are considered. For random sequences made up of $k$ kinds of letters, the lengths of the runs are sorted in two ways to give two definitions of run length ordering. In one definition, the lengths of the runs are sorted separately for each letter type. In the second definition, the lengths of all the runs are sorted together. Exact formulas are developed for the distributions of the m-th longest runs for both definitions. The derivations are based on a two-step method that is applicable to various other runs-related distributions, such as joint distributions of several letter types and multiple run lengths of a single letter type.
2024-05-02
Joint distribution of rises, falls, and number of runs in random sequences
Published in Communications in Statistics - Theory and Methods, 48(3) (2019) • View PublicationBIB
By using the matrix formulation of the two-step approach to the distributions of runs, a recursive relation and an explicit expression are derived for the generating function of the joint distribution of rises and falls for multivariate random sequences in terms of generating functions of individual letters, from which the generating functions of the joint distribution of rises, falls, and number of runs are obtained. An explicit formula for the joint distribution of rises and falls with arbitrary specification is also obtained.