arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

6952 papers tagged with this keyword
2020-03-07 v2
Arcs in $\mathbb F_q^2$
An arc is a subset of $\mathbb F_q^2$ which does not contain any collinear triples. Let $A(q,k)$ denote the number of arcs in $\mathbb F_q^2$ with cardinality $k$. This paper is primarily concerned with estimating the size of $A(q,k)$ when $k$ is relatively large, namely $k=q^t$ for some $t>0$. Trivial estimates tell us that \[ {q \choose k} \leq A(q,k) \leq {q^2 \choose k}. \] We show that the behaviour of $A(q,k)$ changes significantly close to $t=1/2$. Below this threshold an elementary argument is used to prove that the trivial upper bound above cannot be improved significantly. On the other hand, for $t \geq 1/2+δ$, we use the theory of hypergraph containers to get an improved upper bound \[ A(q,k) \leq {q^{2-t+2δ} \choose k}. \] This technique is also used to give an upper bound for the size of the largest arc in a random subset of $\mathbb F_q^2$ which holds with high probability. For example, we prove that a $p$-random subset $Q \subset \mathbb F_q^2$ with $q^{-3/2}<p<q^{-1}$ contains an arc of size $Ω(q^{1/2})$ with high probability. The result is optimal for this range of $p$. Finally, this optimal bound for arcs in random sets is used to prove a finite field analogue of a result of Balogh and Solymosi, with a better exponent: there exists a subset $P \subset \mathbb F_q^2$ which does not contain any collinear quadruples, but with the property that for every $P' \subset P$ with $|P'| \geq |P|^{3/4+o(1)}$, $P'$ contains a collinear triple.
2020-03-07 v3
Quasi-random words and limits of word sequences
Published in European Journal of Combinatorics, Volume 98, 2021 • View PublicationBIB
Words are sequences of letters over a finite alphabet. We study two intimately related topics for this object: quasi-randomness and limit theory. With respect to the first topic we investigate the notion of uniform distribution of letters over intervals, and in the spirit of the famous Chung--Graham--Wilson theorem for graphs we provide a list of word properties which are equivalent to uniformity. In particular, we show that uniformity is equivalent to counting 3-letter subsequences. Inspired by graph limit theory we then investigate limits of convergent word sequences, those in which all subsequence densities converge. We show that convergent word sequences have a natural limit, namely Lebesgue measurable functions of the form $f:[0,1]\to[0,1]$. Via this theory we show that every hereditary word property is testable, address the problem of finite forcibility for word limits and establish as a byproduct a new model of random word sequences. Along the lines of the proof of the existence of word limits, we can also establish the existence of limits for higher dimensional structures. In particular, we obtain an alternative proof of the result by Hoppen, Kohayakawa, Moreira, Ráth and Sampaio [{\it J. Combin. Theory Ser. B 103(1):93--113, 2013}] establishing the existence of permutons.
2020-03-06 v5
Rainbow Hamilton Cycles in Random Geometric Graphs
Published • View PublicationBIB
Let $X_1,X_2,\ldots,X_n$ be chosen independently and uniformly at random from the unit $d$-dimensional cube $[0,1]^d$. Let $r$ be given and let $\cal X=\{X_1,X_2,\ldots,X_n\}$. The random geometric graph $G=G_{\cal X,r}$ has vertex set $\cal X$ and an edge $X_iX_j$ whenever $\|X_i-X_j\|\leq r$. We show that if each edge of $G$ is colored independently from one of $n+o(n)$ colors and $r$ has the smallest value such that $G$ has minimum degree at least two, then $G$ contains a rainbow Hamilton cycle a.a.s.
2020-03-06 v3
On Multitype Random Forests with a Given Degree Sequence, the Total Population of Branching Forests and Enumerations of Multitype Forests
The degree sequence $(N_{i,j}(k),1\leq i,j\leq d,k\geq 0)$ of a multitype forest with $d$ types, is the number of individuals type $i$, having $k$ children type $j$. We construct a multitype forest sampled uniformly from all multitype forest with a given degree sequence (MFGDS). For this, we use an extension of the Ballot Theorem by (Chaumont and Liu, 2016), and generalize the Vervaat transform (Vervaat, 1979) to multidimensional discrete exchangeable increment processes. We prove that MFGDS are extensions of multitype Galton-Watson (MGW) forests, since mixing the laws of the former, one obtains MGW forests with fixed sizes by type (CMGW). We also obtain the law of the total population by types in a MGW forest, generalizing Otter-Dwass formula (Otter 1949, Dwass 1969). We apply this to obtain enumerations of plane, labeled and binary multitype forests having fixed roots and individuals by types. We give an algorithm to simulate certain CMGW forests, generalizing the unitype case of (Devroye, 2012).
2020-03-06 v3
Fast calculation of the variance of edge crossings in random arrangements
The crossing number of a graph $G$, $\mathrm{cr}(G)$, is the minimum number of edge crossings arising when drawing a graph on a certain surface. Determining $\mathrm{cr}(G)$ is a problem of great importance in Graph Theory. Its maximum variant, i.e. the maximum crossing number, $\mathrm{max-cr}(G)$, is receiving growing attention. Instead of an optimization problem on the number of crossings, here we consider the variance of the number of edge crossings, when embedding the vertices of an arbitrary graph uniformly at random in some space. In his pioneering research, Moon derived this variance on random linear arrangements of complete unipartite and bipartite graphs. Given the need of efficient algorithms to support this sort of research and given also the growing interest of the number of edge crossings in spatial networks, networks where vertices are embedded in some space, here we derive an algorithm to calculate the variance in arbitrary graphs in time $o(nm^2)$ that we transform into one that runs in time $O(nm)$ by reusing computations. We also derive one for forests that runs in time $O(n)$. These algorithms work on a wide range of random layouts (not only on Moon's) and are based on novel arithmetic expressions for the calculation of the variance that we develop from previous theoretical work. This paves the way for many applications that rely on a fast but exact calculation of the variance.
2020-03-06 v2
Covering cycles in sparse graphs
Published • View PublicationBIB
Let $k \geq 2$ be an integer. Kouider and Lonc proved that the vertex set of every graph $G$ with $n \geq n_0(k)$ vertices and minimum degree at least $n/k$ can be covered by $k - 1$ cycles. Our main result states that for every $α> 0$ and $p = p(n) \in (0, 1]$, the same conclusion holds for graphs $G$ with minimum degree $(1/k + α)np$ that are sparse in the sense that \[ e_G(X,Y) \leq p|X||Y| + o(np\sqrt{|X||Y|}/\log^3 n) \qquad \forall X,Y\subseteq V(G). \] In particular, this allows us to determine the local resilience of random and pseudorandom graphs with respect to having a vertex cover by a fixed number of cycles. The proof uses a version of the absorbing method in sparse expander graphs.
2020-03-06
On the Collection of Fringe Subtrees in Random Binary Trees
Published • View PublicationBIB
A fringe subtree of a rooted tree is a subtree consisting of one of the nodes and all its descendants. In this paper, we are specifically interested in the number of non-isomorphic trees that appear in the collection of all fringe subtrees of a binary tree. This number is analysed under two different random models: uniformly random binary trees and random binary search trees. In the case of uniformly random binary trees, we show that the number of non-isomorphic fringe subtrees lies between $c_1n/\sqrt{\ln n}(1+o(1))$ and $c_2n/\sqrt{\ln n}(1+o(1))$ for two constants $c_1 \approx 1.0591261434$ and $c_2 \approx 1.0761505454$, both in expectation and with high probability, where $n$ denotes the size (number of leaves) of the uniformly random binary tree. A similar result is proven for random binary search trees, but the order of magnitude is $n/\ln n$ in this case. Our proof technique can also be used to strengthen known results on the number of distinct fringe subtrees (distinct in the sense of ordered trees). This quantity is of the same order of magnitude in both cases, but with slightly different constants in the upper and lower bounds.
Sorting networks, staircase Young tableaux and last passage percolation
Published in Séminaire Lotharingien de Combinatoire 84B - Proceedings of the 32nd International Conference on "Formal Power Series and Algebraic Combinatorics", 2020 • Search Publication
We present new combinatorial and probabilistic identities relating three random processes: the oriented swap process on $n$ particles, the corner growth process, and the last passage percolation model. We prove one of the probabilistic identities, relating a random vector of last passage percolation times to its dual, using the duality between the Robinson-Schensted-Knuth and Burge correspondences. A second probabilistic identity, relating those two vectors to a vector of "last swap times" in the oriented swap process, is conjectural. We give a computer-assisted proof of this identity for $n\le 6$ after first reformulating it as a purely combinatorial identity, and discuss its relation to the Edelman-Greene correspondence.
2020-03-06
Constraints on Brouwer's Laplacian Spectrum Conjecture
Published • View PublicationBIB
Brouwer's Conjecture states that, for any graph $G$, the sum of the $k$ largest (combinatorial) Laplacian eigenvalues of $G$ is at most $|E(G)| + \binom{k+1}{2}$, $1 \leq k \leq n$. We present several interrelated results establishing Brouwer's conjecture $\text{BC}_k(G)$ for a wide range of graphs $G$ and parameters $k$. In particular, we show that (1) $\text{BC}_k(G)$ is true for low-arboricity graphs, and in particular for planar $G$ when $k \geq 11$; (2) $\text{BC}_k(G)$ is true whenever the variance of the degree sequence is not very high, generalizing previous results for $G$ regular or random; (3) $\text{BC}_k(G)$ is true if $G$ belongs to a hereditarily spectrally-bounded class and $k$ is sufficiently large as a function of $k$, in particular $k \geq \sqrt{32n}$ for bipartite graphs; (4) $\text{BC}_k(G)$ holds unless $G$ has edge-edit distance $< k \sqrt{2n} = O(n^{3/2})$ from a split graph; (5) no $G$ violates the conjectured upper bound by more than $O(n^{5/4})$, and bipartite $G$ by no more than $O(n)$; and (6) $\text{BC}_k(G)$ holds for all $k$ outside an interval of length $O(n^{3/4})$. Furthermore, we present a surprising negative result: asymptotically almost surely, a uniform random signed complete graph violates the conjectured bound by $Ω(n)$.
2020-03-05 v2
Finding linearly generated subsequences
Published • View PublicationBIB
We develop a new algorithm to compute determinants of all possible Hankel matrices made up from a given finite length sequence over a finite field. Our algorithm fits within the dynamic programming paradigm by exploiting new recursive relations on the determinants of Hankel matrices together with new observations concerning the distribution of zero determinants among the possible matrix sizes allowed by the length of the original sequence. The algorithm can be used to isolate \emph{very} efficiently linear shift feedback registers hidden in strings with random prefix and random postfix for instance and, therefore, recovering the shortest generating vector. Our new mathematical identities can be used also in any other situations involving determinants of Hankel matrices. We also implement a parallel version of our algorithm. We compare our results empirically with the trivial algorithm which consists of computing determinants for each possible Hankel matrices made up from a given finite length sequence. Our new accelerated approach on a single processor is faster than the trivial algorithm on 160 processors for input sequences of length 16384 for instance.
2020-03-05
On zero-sum free sequences contained in random subsets of finite cyclic groups
Published in Discrete Appl. Math. 330 (2023), 118 - 127 • View PublicationBIB
Let $C_n$ be a cyclic group of order $n$. A sequence $S$ of length $\ell$ over $C_n$ is a sequence $S = a_1\boldsymbol\cdot a_2\boldsymbol\cdot \ldots\boldsymbol\cdot a_{\ell}$ of $\ell$ elements in $C_n$, where a repetition of elements is allowed and their order is disregarded. We say that $S$ is a zero-sum sequence if $Σ_{i=1}^{\ell} a_i = 0$ and that $S$ is a zero-sum free sequence if $S$ contains no zero-sum subsequence. Let $R$ be a random subset of $C_n$ obtained by choosing each element in $C_n$ independently with probability $p$. Let $N^R_{n-1-k}$ be the number of zero-sum free sequences of length $n-1-k$ in $R$. Also, let $N^R_{n-1-k,d}$ be the number of zero-sum free sequences of length $n-1-k$ having $d$ distinct elements in $R$. We obtain the expectation of $N^R_{n-1-k}$ and $N^R_{n-1-k,d}$ for $0\leq k\leq \big\lfloor \frac{n}{3} \big\rfloor$. We also show a concentration result on $N^R_{n-1-k}$ and $N^R_{n-1-k,d}$ when $k$ is fixed.
2020-03-05 v2
An improvement of the Boppana-Holzman bound for Rademacher random variables
Let $v_1,v_2,...,v_n$ be real numbers whose squares add up to $1$. Consider the $2^n$ signed sums of the form $S=\sum_{i=1}^n \pm v_i.$ Holzman and Kleitman (1992) proved that at least $\frac38=0.375$ of these sums satisfy $|S|\leq 1.$ By using bounds for appropriate moments of $S,$ Boppana and Holzman (2017) were able to improve the bound to $\frac{13}{32}=0.40625$ and even a bit better to $\frac{13}{32}+9\times10^{-6}.$ By following their approach, but using a key result of Bentkus and Dzindzalieta (2015), we will drastically improve (by more than 5\%) the latter barrier $\frac{13}{32}$ to $\frac{1}{2}-\frac{Φ(-2)}{4Φ(-\sqrt{2})}\approx 0.42768.$
2020-03-04
A homological characterization of generalized multinomial coefficients related to the entropic chain rule
Published • View PublicationBIB
There is an asymptotic relationship between the multiplicative relations among multinomial coefficients and the (additive) recurrence property of Shannon entropy known as the chain rule. We show that both types of identities are manifestations of a unique algebraic construction: a $1$-cocycle condition in \emph{information cohomology}, an algebraic invariant of phesheaves of modules on \emph{information structures} (categories of observables). Baudot and Bennequin introduced this cohomology and proved that Shannon entropy represents the only nontrivial cohomology class in degree $1$ when the coefficients are a natural presheaf of probabilistic functionals. The author obtained later a $1$-parameter family of deformations of that presheaf, in such a way that each Tsallis $α$-entropy appears as the unique $1$-cocycle associated to the parameter $α$. In this article, we introduce a new presheaf of \emph{combinatorial functionals}, which are measurable functions of finite arrays of integers; these arrays represent \emph{histograms} associated to random experiments. In this case, the only cohomology class in degree $0$ is generated by the exponential function and $1$-cocycles are Fontené-Ward generalized multinomial coefficients. As a byproduct, we get a simple combinatorial analogue of the fundamental equation of information theory that characterizes the generalized binomial coefficients. The asymptotic relationship mentioned above is extended to a correspondence between certain generalized multinomial coefficients and any $α$-entropy, that sheds new light on the meaning of the chain rule and its deformations.
2020-03-02 v2
Holes and islands in random point sets
Published • View PublicationBIB
For $d\in\mathbb{N}$, let $S$ be a set of points in $\mathbb{R}^d$ in general position. A set $I$ of $k$ points from $S$ is a $k$-island in $S$ if the convex hull $\mathrm{conv}(I)$ of $I$ satisfies $\mathrm{conv}(I) \cap S = I$. A $k$-island in $S$ in convex position is a $k$-hole in $S$. For $d,k\in\mathbb{N}$ and a convex body $K\subseteq\mathbb{R}^d$ of volume $1$, let $S$ be a set of $n$ points chosen uniformly and independently at random from $K$. We show that the expected number of $k$-holes in $S$ is in $O(n^d)$. Our estimate improves and generalizes all previous bounds. In particular, we estimate the expected number of empty simplices in $S$ by $2^{d-1}\cdot d!\cdot\binom{n}{d}$. This is tight in the plane up to a lower-order term. Our method gives an asymptotically tight upper bound $O(n^d)$ even in the much more general setting, where we estimate the expected number of $k$-islands in $S$.
2020-03-02 v2
Efficient algorithms for the Potts model on small-set expanders
An emerging trend in approximate counting is to show that certain `low-temperature' problems are easy on typical instances, despite worst-case hardness results. For the class of regular graphs one usually shows that expansion can be exploited algorithmically, and since random regular graphs are good expanders with high probability the problem is typically tractable. Inspired by approaches used in subexponential-time algorithms for Unique Games, we develop an approximation algorithm for the partition function of the ferromagnetic Potts model on graphs with a small-set expansion condition. In such graphs it may not suffice to explore the state space of the model close to ground states, and a novel feature of our method is to efficiently find a larger set of `pseudo-ground states' such that it is enough to explore the model around each pseudo-ground state.
2020-02-29 v2
Order-isomorphic twins in permutations
Published • View PublicationBIB
Let $a_1,\dotsc,a_n$ be a permutation of $[n]$. Two disjoint order-isomorphic subsequences are called \emph{twins}. We show that every permutation of $[n]$ contains twins of length $Ω(n^{3/5})$ improving the trivial bound of $Ω(n^{1/2})$. We also show that a random permutation contains twins of length $Ω(n^{2/3})$, which is sharp.
2020-02-28
On cherry and pitchfork distributions of random rooted and unrooted phylogenetic trees
Published • View PublicationBIB
Tree shape statistics are important for investigating evolutionary mechanisms mediating phylogenetic trees. As a step towards bridging shape statistics between rooted and unrooted trees, we present a comparison study on two subtree statistics known as numbers of cherries and pitchforks for the proportional to distinguishable arrangements (PDA) and the Yule-Harding-Kingman (YHK) models. Based on recursive formulas on the joint distribution of the number of cherries and that of pitchforks, it is shown that cherry distributions are log-concave for both rooted and unrooted trees under these two models. Furthermore, the mean number of cherries and that of pitchforks for unrooted trees converge respectively to those for rooted trees under the YHK model while there exists a limiting gap of 1/4 for the PDA model. Finally, the total variation distances between the cherry distributions of rooted and those of unrooted trees converge for both models. Our results indicate that caution is required for conducting statistical analysis for tree shapes involving both rooted and unrooted trees.
2020-02-27
Random stable type minimal factorizations of the $n$-cycle
We investigate random minimal factorizations of the $n$-cycle, that is, factorizations of the permutation $(1 \, 2 \cdots n)$ into a product of cycles $τ_1, \ldots, τ_k$ whose lengths $\ell(τ_1), \ldots, \ell(τ_k)$ verify the minimality condition $\sum_{i=1}^k(\ell(τ_i)-1)=n-1$. By associating to a cycle of the factorization a black polygon inscribed in the unit disk, and reading the cycles one after an other, we code a minimal factorization by a process of colored laminations of the disk, which are compact subsets made of red noncrossing chords delimiting faces that are either black or white. Our main result is the convergence of this process as $n \rightarrow \infty$, when the factorization is randomly chosen according to Boltzmann weights in the domain of attraction of an $α$-stable law, for some $α\in (1,2]$. The new limiting process interpolates between the unit circle and a colored version of Kortchemski's $α$-stable lamination. Our principal tool in the study of this process is a bijection between minimal factorizations and a model of size-conditioned labelled random trees whose vertices are colored black or white.
2020-02-26 v2
Tuning as convex optimisation: a polynomial tuner for multi-parametric combinatorial samplers
Published • View PublicationBIB
Combinatorial samplers are algorithmic schemes devised for the approximate- and exact-size generation of large random combinatorial structures, such as context-free words, various tree-like data structures, maps, tilings, RNA molecules. They can be adapted to combinatorial specifications with additional parameters, allowing for a more flexible control over the output profile of parametrised combinatorial patterns. One can control, for instance, the number of leaves, profile of node degrees in trees or the number of certain sub-patterns in generated strings. However, such a flexible control requires an additional and nontrivial tuning procedure. Using techniques of convex optimisation, we present an efficient tuning algorithm for multi-parametric combinatorial specifications. Our algorithm works in polynomial time in the system description length, the number of tuning parameters, the number of combinatorial classes in the specification, and the logarithm of the total target size. We demonstrate the effectiveness of our method on a series of practical examples, including rational, algebraic, and so-called Pólya specifications. We show how our method can be adapted to a broad range of less typical combinatorial constructions, including symmetric polynomials, labelled sets and cycles with cardinality lower bounds, simple increasing trees or substitutions. Finally, we discuss some practical aspects of our prototype tuner implementation and provide its benchmark results.
2020-02-25 v2
The asymptotic value of energy for matrices with degree-distance-based entries of random graphs
Published • View PublicationBIB
For a graph $G=(V, E)$ and $i, j\in V$, denote the distance between $i$ and $j$ in $G$ by $D(i, j)$ and the degrees of $i$, $j$ by $d_i$, $d_j$, respectively. Let $f(D(i, j), d_{i}, d_{j})$ be a function symmetric in $i$ and $j$. Define a matrix $W_f(G)$, called the weighted distance matrix, of $G$, with the $ij$-entry $W_f(G)(i, j)=f(D(i, j), d_{i}, d_{j})$ if $i\neq j$ and $W_f(G)(i, j)=0$ if $i=j$. In this paper, we prove that if the symmetric function $f$ satisfies that $f(D(i, j), (1+o(1))np, (1+o(1))np)=(1+o(1))f(D(i, j), np, np)$, then for almost all graphs $G_p$ in the $Erd\ddot{o}s$-$R\acute{e}nyi$ random graph model $\mathcal{G}_{n, p}$, the energy of $W_f(G_p)$ is $\{(\frac{8}{3π}\sqrt{p(1-p)}+o(1))\cdot|f(1, np, np)-f(2, np, np)|+o(|f(2, np, np)|)\}\cdot n^{3/2}$. As a consequence, we give the asymptotic values of energies of a variety of weighted distance matrices with function $f$ from distance-based only and mixed with degree-distance-based topological indices of chemical use. This generalizes our former result with only degree-based weights.