arXiv++ Combinatorics

Browse math.CO papers from arXiv

sequence

6845 papers tagged with this keyword
2023-03-08 v4
Six Permutation Patterns Force Quasirandomness
Published in Discrete Analysis, 2024:8, 26 pp • Search Publication
A sequence $π_1,π_2,\dots$ of permutations is said to be "quasirandom" if the induced density of every permutation $σ$ in $π_n$ converges to $1/|σ|!$ as $n\to\infty$. We prove that $π_1,π_2,\dots$ is quasirandom if and only if the density of each permutation $σ$ in the set $$\{123,321,2143,3412,2413,3142\}$$ converges to $1/|σ|!$. Previously, the smallest cardinality of a set with this property, called a "quasirandom-forcing" set, was known to be between four and eight. In fact, we show that there is a single linear expression of the densities of the six permutations in this set which forces quasirandomness and show that this is best possible in the sense that there is no shorter linear expression of permutation densities with positive coefficients with this property. In the language of theoretical statistics, this expression provides a new nonparametric independence test for bivariate continuous distributions related to Spearman's $ρ$.
2023-03-07
Complete Log Concavity of Coverage-Like Functions
We introduce an expressive subclass of non-negative almost submodular set functions, called strongly 2-coverage functions which include coverage and (sums of) matroid rank functions, and prove that the homogenization of the generating polynomial of any such function is completely log-concave, taking a step towards characterizing the coefficients of (homogeneous) completely log-concave polynomials. As a consequence we obtain that the "level sets" of any such function form an ultra-log concave sequence.
Friedman's "Long Finite Sequences'': The End of the Busy Beaver Contest
Harvey Friedman gives a comparatively short description of an ``unimaginably large'' number $n(3)$ , beyond, e.g. the values $$ A(7,184)< A({7198},158386) < n(3)$$ of Ackermann's function - but finite. We implement Friedman's combinatorial problem about subwords of words over a 3-letter alphabet on a family of Turing machines, which, starting on empty tape, run (more than) $n(3)$ steps, and then halt. Examples include a (44,8) (symbol,state count) machine as well as a (276,2) and a (2,1840) one. In total, there are at most 37022 non-trivial pairs $(n,m)$ with Busy Beaver values ${\tt BB(n,m)} < A(7198,158386).$ We give algorithms to map any $(|Q|,|E|)$ TM to another, where we can choose freely either $|Q'|\geq 2$ or $|E'|\geq 2$ (the case $|Q'|=2$ for empty initial tape is the tricky one). Given the size of $n(3)$ and the fact that these TMs are not {\it holdouts}, but assured to stop, Friedman's combinatorial problem provides a definite upper bound on what might ever be possible to achieve in the Busy Beaver contest. We also treat $n(4)> A^{(A(187196))}(1)$.
2023-03-06 v2
A Closure Lemma for tough graphs and Hamiltonian degree conditions
The closure of a graph $G$ is the graph $G^*$ obtained from $G$ by repeatedly adding edges between pairs of non-adjacent vertices whose degree sum is at least $n$, where $n$ is the number of vertices of $G$. The well-known Closure Lemma proved by Bondy and Chvátal states that a graph $G$ is Hamiltonian if and only if its closure $G^*$ is. This lemma can be used to prove several classical results in Hamiltonian graph theory. We prove a version of the Closure Lemma for tough graphs. A graph $G$ is $t$-tough if for any set $S$ of vertices of $G$, the number of components of $G-S$ is at most $t |S|$. A Hamiltonian graph must necessarily be 1-tough. Conversely, Chvátal conjectured that there exists a constant $t$ such that every $t$-tough graph is Hamiltonian. The {\it $t$-closure} of a graph $G$ is the graph $G^{t*}$ obtained from $G$ by repeatedly adding edges between pairs of non-adjacent vertices whose degree sum is at least $n-t$. We prove that, for $t\geq 2$, a $\frac{3t-1}{2}$-tough graph $G$ is Hamiltonian if and only if its $t$-closure $G^{t*}$ is. Hoàng conjectured the following: Let $G$ be a graph with degree sequence $d_1 \leq d_2 \leq \ldots \leq d_n$; then $G$ is Hamiltonian if $G$ is $t$-tough and, $\forall i <\frac{n}{2},\mbox{ if } d_i\leq i \mbox{ then } d_{n-i+t}\geq n-i$. This conjecture is analogous to the well known theorem of Chvátal on Hamiltonian ideals. Hoàng proved the conjecture for $t \leq 3$. Using the closure lemma for tough graphs, we prove the conjecture for $t = 4$.
2023-03-05 v2
Some D-finite and Some Possibly D-finite Sequences in the OEIS
Published in Journal of Integer Sequences, vol. 26, article 23.4.5, 2023 • Search Publication
In an automatic search, we found conjectural recurrences for some sequences in the OEIS that were not previously recognized as being D-finite. In some cases, we are able to prove the conjectured recurrence. In some cases, we are not able to prove the conjectured recurrence, but we can prove that a recurrence exists. In some remaining cases, we do not know where the recurrence might come from.
A combinatorial proof for the secretary problem with multiple choices
The Secretary problem is a classical sequential decision-making question that can be succinctly described as follows: a set of rank-ordered applicants are interviewed sequentially for a single position. Once an applicant is interviewed, an immediate and irrevocable decision is made if the person is to be offered the job or not and only applicants observed so far can be used in the decision process. The problem of interest is to identify the stopping rule that maximizes the probability of hiring the highest-ranked applicant. A multiple-choice version of the Secretary problem, known as the Dowry problem, assumes that one is given a fixed integer budget for the total number of selections allowed to choose the best applicant. It has been solved using tools from dynamic programming and optimal stopping theory. We provide the first combinatorial proof for a related new \emph{query-based model} for which we are allowed to solicit the response of an expert to determine if an applicant is optimal. Since the selection criteria differ from those of the Dowry problem we obtain nonidentical expected stopping times. Our result indicates that an optimal strategy is the $(a_s, a_{s-1}, \ldots, a_1)$-strategy, i.e., for the $i^{th}$ selection, where $1 \le i \le s$ and $1 \le j = s+1-i \le s$, we reject the first $a_j$ applicants, wait until the decision of the $(i-1)^{th}$ selection (if $i \ge 2$), and then accept the next applicant whose qualification is better than all previously appeared applicants. Furthermore, our optimal strategy is right-hand based, i.e., the optimal strategies for two models with $s_1$ and $s_2$ selections in total ($s_1 < s_2$) share the same sequence $a_1, a_2, \ldots, a_{s_1}$ when it is viewed from the right. When the total number of applicants tends to infinity, our result agrees with the thresholds obtained by Gilbert and Mosteller.
2023-03-04 v3
The Critical Beta-splitting Random Tree II: Overview and Open Problems
In the critical beta-splitting model of a random $n$-leaf rooted tree, clades are recursively (from the root) split into sub-clades, and a clade of $m$ leaves is split into sub-clades containing $i$ and $m-i$ leaves with probabilities $\propto 1/(i(m-i))$. Study of structure theory and explicit quantitative aspects of this model (in discrete or continuous versions) is an active research topic. For many results there are different proofs, probabilistic or analytic, so the model provides a testbed for a ``compare and contrast" discussion of techniques. This article provides an overview of results proved in the sequence of similarly-titled articles I, III, IV and related articles. We mostly do not repeat proofs given elsewhere: instead we seek to paint a ``Big Picture" via graphics and heuristics, and emphasize open problems. Our discussion is centered around three categories of results. (i) There is a CLT for leaf heights, and the analytic proofs can be extended to provide surprisingly precise analysis of other height-related aspects. (ii) There is an explicit description of the limit {\em fringe distribution} relative to a random leaf, whose graphical representation is essentially the format of the cladogram representation of biological phylogenies. (iii) There is a canonical embedding of the discrete model into a continuous-time model, that is a random tree CTCS(n) on $n$ leaves with real-valued edge lengths, and this model turns out more convenient to study. The family (CTCS(n), n \ge 2) is consistent under a ``delete random leaf and prune" operation. That leads to an explicit inductive construction of (CTCS(n), n \ge 2) as $n$ increases, and then to a limit structure CTCS($\infty$) formalized via exchangeable partitions. Many open problems remain, in particular to elucidate a relation between CTCS($\infty$) and the $β(2,1)$ coalescent.
A unified treatment of families of partition functions
We present a unified framework of combinatorial descriptions, and the analogous asymptotic growth of the coefficients of two general families of functions related to integer partitions. In particular, we resolve several conjectures and verify several claims that are posted on the On-Line Encyclopedia of Integer Sequences. We perform the asymptotic analysis by systematically applying the Mellin transform, residue analysis, and the saddle point method. The combinatorial descriptions of these families of generalized partition functions involve colorings of Young tableaux, along with their ``divisor diagrams'', denoted with sets of colors whose sizes are controlled by divisor functions.
2023-02-27
Jump-systems of $T$-paths
Published in Proceedings of the Twelfth Japanese-Hungarian Symposium on Discrete Mathematics and Its Applications (March 2023) • Search Publication
Jump systems are sets of integer vectors satisfying a simple axiom, generalizing matroids, also delta-matroids, and well-kown combinatorial examples such as degree sequences of subgraphs of a graph. It is useful to know if a set of vectors defined from combinatorial structures is a jump system: this has consequences for optimizing on the set, or on some derived sets of vectors. In this note we are mainly concerned in telling our proof of the following more than two decades old fact and its original, elementary proof for an example different from degree sequences: {\em Given an udirected graph $G=(V,E)$ and $T\subseteq V$, the vectors $m$ indexed by $T$ for which there exist a set of openly disjoint $T$-paths so that each $t\in T$ is the endpoint of exactly $m(t)$ paths forms a jump system. The same holds for edge-disjoint $T$-paths.} We are also exhibiting the context and some consequences of this fact, with some pointers to recent developments, among them ro another proof by Iwata and Yokoi, to some related jump system intersection theorems and to some open problems.
String attractors of some simple-Parry automatic sequences
Firstly studied by Kempa and Prezza in 2018 as the cement of text compression algorithms, string attractors have become a compelling object of theoretical research within the community of combinatorics on words. In this context, they have been studied for several families of finite and infinite words. In this paper, we obtain string attractors of prefixes of particular infinite words generalizing k-bonacci words (including the famous Fibonacci word) and related to simple Parry numbers. In fact, our description involves the numeration systems classically derived from the considered morphisms. This extends our previous work published in the international conference WORDS 2023.
On the Design of Codes for DNA Computing: Secondary Structure Avoidance Codes
In this work, we investigate a challenging problem, which has been considered to be an important criterion in designing codewords for DNA computing purposes, namely secondary structure avoidance in single-stranded DNA molecules. In short, secondary structure refers to the tendency of a single-stranded DNA sequence to fold back upon itself, thus becoming inactive in the computation process. While some design criteria that reduces the possibility of secondary structure formation has been proposed by Milenkovic and Kashyap (2006), the main contribution of this work is to provide an explicit construction of DNA codes that completely avoid secondary structure of arbitrary stem length. Formally, given codeword length n and arbitrary integer m>=2, we provide efficient methods to construct DNA codes of length n that avoid secondary structure of any stem length more than or equal to m. Particularly, when m = 3, our constructions yield a family of DNA codes of rate 1.3031 bits/nt, while the highest rate found in the prior art was 1.1609 bits/nt. In addition, for m>=3log n + 4, we provide an efficient encoder that incurs only one redundant symbol.
2023-02-26
Log-Concavity of Infinite Product and Infinite Sum Generating Functions
We expand on the remark by Andrews on the importance of infinite sums and products in combinatorics. Let $\{g_d(n)\}_{d\geq 0,n \geq 1}$ be the double sequences $σ_d(n)= \sum_{\ell \mid n} \ell^d$ or $ψ_d(n)= n^d$. We associate double sequences $\left\{ p^{g_{d} }\left( n\right) \right\}$ and $\left\{ q^{g_{d} }\left( n\right) \right\} $, defined as the coefficients of \begin{eqnarray*} \sum_{n=0}^{\infty} p^{g_{d} }\left( n\right) \, t^{n} & := & \prod_{n=1}^{\infty} \left( 1 - t^{n} \right)^{-\frac{ \sum_{\ell \mid n} μ(\ell) \, g_d(n/\ell) }{n} }, \\ \sum_{n=0}^{\infty} q^{g_{d} }\left( n\right) \, t^{n} & := & \frac{1}{1 - \sum_{n=1}^{\infty} g_d(n) \, t^{n} }. \end{eqnarray*} These coefficients are related to the number of partitions $\mathrm{p}\left( n\right) = p^{σ_{1 }}\left ( n\right) $, plane partitions $pp\left( n\right) = p^{σ_{2 }}\left( n\right) $ of $n$, and Fibonacci numbers $F_{2n} = q^{ψ_{1 }}\left( n\right) $. Let $n \geq 3$ and let $n \equiv 0 \pmod{3}$. Then the coefficients are log-concave at $n$ for almost all $d$ in the exponential and geometric cases. The coefficients are not log-concave for almost all $d$ in both cases, if $n \equiv 2 \pmod{3}$. Let $n\equiv 1 \pmod{3}$. Then the log-concave property flips for almost all $d$.
2023-02-25 v6
Construction numbers: How to build a graph?
A construction sequence for a graph is a listing of the elements of the graph (the set of vertices and edges) such that each edge follows both its endpoints. The construction number of the graph is the number of such sequences. We determine this number for various graph families.
2023-02-24
An application of Grothendieck theorem to the theory of multicorrelation sequences, multiple recurrence and partition regularity of quadratic equations
We use Grothendieck theorem to prove a structure theorem for multicorrelation sequences of length two, associated with two (not necessarily commuting) measure preserving actions on a probability space. We use this to deduce a multiple recurrence result concerning products of linear terms, and a partition regularity result of certain systems of quadratic equations, building on the work of Frantzikinakis and Host.
Monochromatic arithmetic progressions in automatic sequences with group structure
We determine asymptotic growth rates for lengths of monochromatic arithmetic progressions in certain automatic sequences. In particular, we look at (one-sided) fixed points of aperiodic, primitive, bijective substitutions and spin substitutions, which are generalisations of the Thue--Morse and Rudin--Shapiro substitutions, respectively. For such infinite words, we show that there exists a subsequence $\left\{d_n\right\}$ of differences along which the maximum length $A(d_n)$ of a monochromatic arithmetic progression (with fixed difference $d_n$) grows at least polynomially in $d_n$. Explicit upper and lower bounds for the growth exponent can be derived from a finite group associated to the substitution. As an application, we obtain bounds for a van der Waerden-type number for a class of colourings parametrised by the size of the alphabet and the length of the substitution.
2023-02-23
Translation of "Simplizialzerlegungen von Beschrankter Flachheit'' by Hans Freudenthal, Annals of Mathematics, Second Series, Volume 43, Number 3, July 1942, Pages 580-583
Published in German original published in the Annals of Mathematics, Second Series, Volume 43, Number 3, July 1942, Pages 580-583 • Search Publication
Translation of the paper ``Simplizialzerlegungen von Beschrankter Flachheit'' by Hans Freudenthal (https://doi.org/10.2307/1968813), in which Freudenthal answers ``a question by Brouwer about the construction of an infinite series of subdivisions of a polytope, such that the next element in the sequence is a subdivision of the previous one and such that the subsimplices that arise do not become arbitrarily flat.''
On Computing Large Temporal (Unilateral) Connected Components
A temporal (directed) graph is a graph whose edges are available only at specific times during its lifetime, $τ$. Paths are sequences of adjacent edges whose appearing times are either strictly increasing or non-strictly increasingly (i.e., non-decreasing) depending on the scenario. Then, the classical concept of connected components and also of unilateral connected components in static graphs and digraphs naturally extends to the temporal setting. In this paper, we answer to the following fundamental questions in temporal graphs. (i) What is the complexity of deciding the existence of a component of size $k$, parameterized by $τ$, by $k$, and by $k+τ$? We show that this question has a different answer depending on the considered definition of component and whether the temporal graph is directed or undirected. (ii) What is the minimum running time required to check whether a subset of vertices are pairwise reachable? A quadratic algorithm is known but, contrary to the static case, we show that a better running time is unlikely unless SETH fails. (iii) Is it possible to verify whether a subset of vertices is a component in polynomial time? We show that depending on the definition of temporal component this test is NP-complete.
2023-02-22
The intransitive dice kernel: $\frac{\mathbf{1}_{x\ge y}-\mathbf{1}_{x\le y}}{4} - \frac{3(x-y)(1+xy)}{8}$
Answering a pair of questions of Conrey, Gabbard, Grant, Liu, and Morrison, we prove that a triplet of dice drawn from the multiset model are intransitive with probability $1/4+o(1)$ and the probability a random pair of dice tie tends toward $αn^{-1}$ for an explicitly defined constant $α$. This extends and sharpens the recent results of Polymath regarding the balanced sequence model. We further show the distribution of larger tournaments converges to a universal tournamenton in both models. This limit naturally arises from the discrete spectrum of a certain skew-symmetric operator (given by the kernel in the title acting on $L^2([-1,1])$). The limit exhibits a degree of symmetry and can be used to prove that, for instance, the limiting probability that $A_i$ beats $A_{i+1}$ for $1\le i\le 4$ and that $A_5$ beats $A_1$ is $1/32+o(1)$. Furthermore, the limiting tournamenton has range contained in the discrete set $\{0,1\}$. This proves that the associated tournamenton is non-quasirandom in a dramatic fashion, vastly extending work of Cornacchia and Hązła regarding the continuous analogue of the balanced sequence model. The proof is based on a reduction to conditional central limit theorems (related to work of Polymath), the use of a "Poissonization" style method to reduce to computations with independent random variables, and the systematic use of switching-based arguments to extract cancellation in Fourier estimates when establishing local limit-type estimates.
2023-02-20 v2
Embedding theorems for random graphs with specified degrees
Published in Combinator. Probab. Comp. 34 (2025) 115-130 • View PublicationBIB
Given an $n\times n$ symmetric matrix $W\in [0,1]^{[n]\times [n]}$, let $\mathcal{G}(n,W)$ be the random graph obtained by independently including each edge $jk$ with probability $W_{jk}$. Given a degree sequence ${\bf d}=(d_1,\ldots, d_n)$, let $\mathcal{G}(n,{\bf d})$ denote a uniformly random graph with degree sequence ${\bf d}$. We couple $\mathcal{G}(n,W)$ and $\mathcal{G}(n,{\bf d})$ together so that a.a.s. $\mathcal{G}(n,W)$ is a subgraph of $\mathcal{G}(n,{\bf d})$, where $W$ is some function of ${\bf d}$. Let $Δ({\bf d})$ denote the maximum degree in ${\bf d}$. Our coupling result is optimal when $Δ({\bf d})^2\ll \|{\bf d}\|_1$, i.e.\ $W_{ij}$ is asymptotic to $\mathbb{P}(ij\in \mathcal{G}(n,{\bf d}))$ for every $i,j\in [n]$. We also have coupling results for ${\bf d}$ that are not constrained by the condition $Δ({\bf d})^2\ll \|{\bf d}\|_1$. For such ${\bf d}$ our coupling result is still close to optimal, in the sense that $W_{ij}$ is asymptotic to $\mathbb{P}(ij\in \mathcal{G}(n,{\bf d}))$ for most pairs $i,j\in [n]$.
2023-02-20
Reconstruction of Sequences Distorted by Two Insertions
Reconstruction codes are generalizations of error-correcting codes that can correct errors by a given number of noisy reads. The study of such codes was initiated by Levenshtein in 2001 and developed recently due to applications in modern storage devices such as racetrack memories and DNA storage. The central problem on this topic is to design codes with redundancy as small as possible for a given number $N$ of noisy reads. In this paper, the minimum redundancy of such codes for binary channels with exactly two insertions is determined asymptotically for all values of $N\ge 5$. Previously, such codes were studied only for channels with single edit errors or two-deletion errors.