sequence
6845 papers tagged with this keyword
From randomness in two symbols to randomness in three symbols
Published
• View Publication
• BIB
In 1909 Borel defined normality as a notion of randomness of the digits of the representation of a real number over certain base (fractional expansion). If we think the representation of a number over a base as an infinite sequence of symbols from a finite alphabet $A$, we can define normality directly for words of symbols of $A$: A word $x$ is normal to the alphabet $A$ if every finite block of symbols from $A$ appears with the same asymptotic frequency in $x$ as every other block of the same length. Many examples of normal words have been found since its definition, being Champernowne in 1933 the first to show an explicit and simple instance. Moreover, it has been characterized how we can select subsequences of a normal word $x$ preserving its normality, always leaving the alphabet $A$ fixed. In this work we consider the dual problem which consists of inserting symbols in infinite positions of a given word, in such a way that normality is preserved. Specifically, given a symbol $s$ that is not present on the original alphabet $A$ and given a word $x$ that is normal to the alphabet $A$ we solve how to insert the symbol $s$ in infinite positions of the word $x$ such that the resulting word is normal to the expanded alphabet $A\cup \{s\}$.
On Pure Degree Sequence Manipulations Forcing Long Cycles in Graphs
The well-known Hamiltonian sufficient conditions, proposed by Dirac, Faudree et al., Pósa, Bondy, Chvátal are based on pure degree manipulations without any additional conditions. In this paper, we present two new types of pure degree manipulations that produce relatively simpler (in view of verification) results. The reverse versions (long-cycle versions) of the obtained results are presented as well. The importance of pure degree manipulations is motivated by their amenability (as starting points) to generalizations and new ideas in a great variety of ways.
The packing number of the double vertex graph of the path graph
Published in Discrete Appl. Math, 247 (2018) 327-340
• View Publication
• BIB
Neil Sloane showed that the problem of determine the maximum size of a binary code of constant weight 2 that can correct a single adjacent transposition is equivalent to finding the packing number of a certain graph. In this paper we solve this open problem by finding the packing number of the double vertex graph (2-token graph) of a path graph. This double vertex graph is isomorphic to the Sloane's graph. Our solution implies a conjecture of Rob Pratt about the ordinary generating function of sequence A085680.
Inverting the Turán Problem
Published
• View Publication
• BIB
Classical questions in extremal graph theory concern the asymptotics of $\operatorname{ex}(G, \mathcal{H})$ where $\mathcal{H}$ is a fixed family of graphs and $G=G_n$ is taken from a `standard' increasing sequence of host graphs $(G_1, G_2, \dots)$, most often $K_n$ or $K_{n,n}$. Inverting the question, we can instead ask how large $e(G)$ can be with respect to $\operatorname{ex}(G,\mathcal{H})$. We show that the standard sequences indeed maximize $e(G)$ for some choices of $\mathcal{H}$, but not for others. Many interesting questions and previous results arise very naturally in this context, which also, unusually, gives rise to sensible extremal questions concerning multigraphs and non-uniform hypergraphs.
On Powers of the Catalan Number Sequence
Published in Discrete Mathematics (2018)
• View Publication
• BIB
The Catalan number sequence is one of the most famous number sequences in combinatorics and is well studied in the literature. In this paper we further investigate its fundamental properties related to the moment problem and prove for the first time that it is an infinitely divisible Stieltjes moment sequence in the sense of S.-G. Tyan. Besides, any positive real power of the sequence is still a Stieltjes determinate sequence. Some more cases including (a) the central binomial coefficient sequence (related to the Catalan sequence), (b) a double factorial number sequence and (c) the generalized Catalan (or Fuss-Catalan) sequence are also investigated. Finally, we pose two conjectures including the determinacy equivalence between powers of nonnegative random variables and powers of their moment sequences, which is supported by some existing results.
On the Computational Complexity of Non-dictatorial Aggregation
Published in Journal of Artificial Intelligence Research 72 (2021) 137-183
• View Publication
• BIB
We investigate when non-dictatorial aggregation is possible from an algorithmic perspective, where non-dictatorial aggregation means that the votes cast by the members of a society can be aggregated in such a way that there is no single member of the society that always dictates the collective outcome. We consider the setting in which the members of a society take a position on a fixed collection of issues, where for each issue several different alternatives are possible, but the combination of choices must belong to a given set X of allowable voting patterns. Such a set X is called a possibility domain if there is an aggregator that is non-dictatorial, operates separately on each issue, and returns values among those cast by the society on each issue. We design a polynomial-time algorithm that decides, given a set X of voting patterns, whether or not X is a possibility domain. Furthermore, if X is a possibility domain, then the algorithm constructs in polynomial time a non-dictatorial aggregator for X. Furthermore, we show that the question of whether a Boolean domain X is a possibility domain is in NLOGSPACE. We also design a polynomial-time algorithm that decides whether X is a uniform possibility domain, that is, whether X admits an aggregator that is non-dictatorial even when restricted to any two positions for each issue. As in the case of possibility domains, the algorithm also constructs in polynomial time a uniform non-dictatorial aggregator, if one exists. Then, we turn our attention to the case where X is given implicitly, either as the set of assignments satisfying a propositional formula, or as a set of consistent evaluations of a sequence of propositional formulas. In both cases, we provide bounds to the complexity of deciding if X is a (uniform) possibility domain.
The quasi principal rank characteristic sequence
Published in Linear Algebra and its Applications 548 (2018), 42--56
• View Publication
• BIB
A minor of a matrix is quasi-principal if it is a principal or an almost-principal minor. The quasi principal rank characteristic sequence (qpr-sequence) of an $n\times n$ symmetric matrix is introduced, which is defined as $q_1 q_2 \cdots q_n$, where $q_k$ is $\tt A$, $\tt S$, or $\tt N$, according as all, some but not all, or none of its quasi-principal minors of order $k$ are nonzero. This sequence extends the principal rank characteristic sequences in the literature, which only depend on the principal minors of the matrix. A necessary condition for the attainability of a qpr-sequence is established. Using probabilistic techniques, a complete characterization of the qpr-sequences that are attainable by symmetric matrices over fields of characteristic $0$ is given.
On $AP_3$ - covering sequences
Published in C. R. Acad. Sci. Paris, Ser. I 356 (2018), 121-124
• View Publication
• BIB
Recently, motivated by Stanley sequences, Kiss, S\' andor and Yang introduced a new type sequence: a sequence $A$ of nonnegative integers is called an $AP_k$ - covering sequence if there exists an integer $n_0$ such that if $n > n_0$, then there exist $a_1\in A, \dots , a_{k-1}\in A$, $a_1<a_2<\cdots <a_{k-1}<n$ such that $a_1, \dots , a_{k-1}, n$ form a $k$-term arithmetic progression. They prove that there exists an $AP_3$ - covering sequence $A$ such that $\limsup\limits_{n\to\infty}{A(n)}/{\sqrt n}\le 34$. In this note, we prove that there exists an $AP_3$ - covering sequence $A$ such that $\limsup\limits_{n\to\infty}{A(n)}/{\sqrt n}=\sqrt{15}$.
The number of spanning trees in circulant graphs, its arithmetic properties and asymptotic
Published
• View Publication
• BIB
In this paper, we develop a new method to produce explicit formulas for the number $τ(n)$ of spanning trees in the undirected circulant graphs $C_{n}(s_1,s_2,\ldots,s_k)$ and $C_{2n}(s_1,s_2,\ldots,s_k,n).$ Also, we prove that in both cases the number of spanning trees can be represented in the form $τ(n)=p \,n \,a(n)^2,$ where $a(n)$ is an integer sequence and $p$ is a prescribed natural number depending on the parity of $n.$ Finally, we find an asymptotic formula for $τ(n)$ through the Mahler measure of the associated Laurent polynomial $L(z)=2k-\sum\limits_{i=1}^k(z^{s_i}+z^{-s_i}).$
Combinatorial cost: a coarse setting
Published
• View Publication
• BIB
The main inspiration for this paper is a paper by Elek where he introduces combinatorial cost for graph sequences. We show that having cost equal to 1 and hyperfiniteness are coarse invariants. We also show `cost-1' for box spaces behaves multiplicatively when taking subgroups. We show that graph sequences coming from Farber sequences of a group have property A if and only if the group is amenable. The same is true for hyperfiniteness. This generalises a theorem by Elek. Furthermore we optimise this result when Farber sequences are replaced by sofic approximations. In doing so we introduce a new concept: property almost-A.
Designing RNA Secondary Structures is Hard
Published
• View Publication
• BIB
An RNA sequence is a word over an alphabet on four elements $\{A,C,G,U\}$ called bases. RNA sequences fold into secondary structures where some bases match one another while others remain unpaired. Pseudoknot-free secondary structures can be represented as well-parenthesized expressions with additional dots, where pairs of matching parentheses symbolize paired bases and dots, unpaired bases. The two fundamental problems in RNA algorithmic are to predict how sequences fold within some model of energy and to design sequences of bases which will fold into targeted secondary structures. Predicting how a given RNA sequence folds into a pseudoknot-free secondary structure is known to be solvable in cubic time since the eighties and in truly subcubic time by a recent result of Bringmann et al. (FOCS 2016). As a stark contrast, it is unknown whether or not designing a given RNA secondary structure is a tractable task; this has been raised as a challenging open question by Anne Condon (ICALP 2003). Because of its crucial importance in a number of fields such as pharmaceutical research and biochemistry, there are dozens of heuristics and software libraries dedicated to RNA secondary structure design. It is therefore rather surprising that the computational complexity of this central problem in bioinformatics has been unsettled for decades.
In this paper we show that, in the simplest model of energy which is the Watson-Crick model the design of secondary structures is NP-complete if one adds natural constraints of the form: index $i$ of the sequence has to be labeled by base $b$. This negative result suggests that the same lower bound holds for more realistic models of energy. It is noteworthy that the additional constraints are by no means artificial: they are provided by all the RNA design pieces of software and they do correspond to the actual practice.
An efficient dual sampling algorithm with Hamming distance filtration
Published
• View Publication
• BIB
Recently, a framework considering RNA sequences and their RNA secondary structures as pairs, led to some information-theoretic perspectives on how the semantics encoded in RNA sequences can be inferred. In this context, the pairing arises naturally from the energy model of RNA secondary structures. Fixing the sequence in the pairing produces the RNA energy landscape, whose partition function was discovered by McCaskill. Dually, fixing the structure induces the energy landscape of sequences. The latter has been considered for designing more efficient inverse folding algorithms.
We present here the Hamming distance filtered, dual partition function, together with a Boltzmann sampler using novel dynamic programming routines for the loop-based energy model. The time complexity of the algorithm is $O(h^2n)$, where $h,n$ are Hamming distance and sequence length, respectively, reducing the time complexity of samplers, reported in the literature by $O(n^2)$. We then present two applications, the first being in the context of the evolution of natural sequence-structure pairs of microRNAs and the second constructing neutral paths. The former studies the inverse fold rate (IFR) of sequence-structure pairs, filtered by Hamming distance, observing that such pairs evolve towards higher levels of robustness, i.e.,~increasing IFR. The latter is an algorithm that constructs neutral paths: given two sequences in a neutral network, we employ the sampler in order to construct short paths connecting them, consisting of sequences all contained in the neutral network.
The Unreasonable Rigidity of Ulam Sets
Published
• View Publication
• BIB
We give a number of results about families of Ulam sets. Generalizing behavior of Ulam sets U(1,n), we prove using an novel model theoretic approach that there is a rigidity phenomenon for Ulam sets U(a,b) as b increases. Based on this, we suggest a natural conjecture, and investigate its potential applications, including a method of proving certain families of Ulam sequences are regular, for which we also provide partial, unconditional, results. Along this same vein, we give an upper bound bound on the density of Ulam sequences U(1,n). Finally, we give classification results for higher dimensional Ulam sets.
Shifts of the prime divisor function of Alladi and Erdős
We introduce a variation on the prime divisor function $B(n)$ of Alladi and Erdős, a close relative of the sum of proper divisors function $s(n)$. After proving some basic properties regarding these functions, we study the dynamics of its iterates and discover behaviour that is reminiscent of aliquot sequences. We prove that no unbounded sequences occur, analogous to the Catalan-Dickson conjecture, and give evidence towards the analogue of the Erdős-Granville-Pomerance-Spiro conjecture on the pre-image of $s(n)$.
Level algebras and $\boldsymbol{s}$-lecture hall polytopes
Published
• View Publication
• BIB
Given a family of lattice polytopes, a common endeavor in Ehrhart theory is the classification of those polytopes in the family that are Gorenstein, or more generally level. In this article, we consider these questions for $\boldsymbol{s}$-lecture hall polytopes, which are a family of simplices arising from $\boldsymbol{s}$-lecture hall partitions. In particular, we provide concrete classifications for both of these properties purely in terms of $\boldsymbol{s}$-inversion sequences. Moreover, for a large subfamily of $\boldsymbol{s}$-lecture hall polytopes, we provide a more geometric classification of the Gorenstein property in terms of its tangent cones. We then show how one can use the classification of level $\boldsymbol{s}$-lecture hall polytopes to construct infinite families of level $\boldsymbol{s}$-lecture hall polytopes, and to describe level $\boldsymbol{s}$-lecture hall polytopes in small dimensions.
Summations of Linear Recurrent Sequences
Published
• View Publication
• BIB
We give an extension of Sister Celine's method of proving hypergeometric sum identities that allows it to handle a larger variety of input summands. We then apply this to several problems. Some give new results, and some reprove already known results in an automated way.
Improved Bounds for Testing Forbidden Order Patterns
Published
• View Publication
• BIB
A sequence $f\colon\{1,\dots,n\}\to\mathbb{R}$ contains a permutation $π$ of length $k$ if there exist $i_1<\dots<i_k$ such that, for all $x,y$, $f(i_x)<f(i_y)$ if and only if $π(x)<π(y)$; otherwise, $f$ is said to be $π$-free. In this work, we consider the problem of testing for $π$-freeness with one-sided error, continuing the investigation of [Newman et al., SODA'17].
We demonstrate a surprising behavior for non-adaptive tests with one-sided error: While a trivial sampling-based approach yields an $\varepsilon$-test for $π$-freeness making $Θ(\varepsilon^{-1/k} n^{1-1/k})$ queries, our lower bounds imply that this is almost optimal for most permutations! Specifically, for most permutations $π$ of length $k$, any non-adaptive one-sided $\varepsilon$-test requires $\varepsilon^{-1/(k-Θ(1))}n^{1-1/(k-Θ(1))}$ queries; furthermore, the permutations that are hardest to test require $Θ(\varepsilon^{-1/(k-1)}n^{1-1/(k-1)})$ queries, which is tight in $n$ and $\varepsilon$.
Additionally, we show two hierarchical behaviors here. First, for any $k$ and $l\leq k-1$, there exists some $π$ of length $k$ that requires $\tildeΘ_{\varepsilon}(n^{1-1/l})$ non-adaptive queries. Second, we show an adaptivity hierarchy for $π=(1,3,2)$ by proving upper and lower bounds for (one- and two-sided) testing of $π$-freeness with $r$ rounds of adaptivity. The results answer open questions of Newman et al. and [Canonne and Gur, CCC'17].
Uniform rank gradient, cost and local-global convergence
We analyze the rank gradient of finitely generated groups with respect to sequences of subgroups of finite index that do not necessarily form a chain, by connecting it to the cost of p.m.p. actions. We generalize several results that were only known for chains before. The connection is made by the notion of local-global convergence.
In particular, we show that for a finitely generated group $Γ$ with fixed price $c$, every Farber sequence has rank gradient $c-1$. By adapting Lackenby's trichotomy theorem to this setting, we also show that in a finitely presented amenable group, every sequence of subgroups with index tending to infinity has vanishing rank gradient.
Properties of the Fibonacci-sum graph
For each positive integer $n$, the Fibonacci-sum graph $G_n$ on vertices $1,2,\ldots,n$ is defined by two vertices forming an edge if and only if they sum to a Fibonacci number. It is known that each $G_n$ is bipartite, and all Hamiltonian paths in each $G_n$ have been classified. In this paper, it is shown that each $G_n$ has at most one non-trivial automorphism, which is given explicitly. Other properties of $G_n$ are also found, including the degree sequence, the treewidth, the nature of the bipartition, and that $G_n$ is outerplanar.
On Graph Isomorphism Problem
Let $G$ and $H$ be two simple graphs. A bijection $φ:V(G)\rightarrow V(H)$ is called an isomorphism between $G$ and $H$ if $(φv_i)(φv_j)\in E(H)$ $\Leftrightarrow$ $v_i v_j\in E(G)$, $\forall v_i,v_j \in V(G)$. In the case that $G = H$, we say $φ$ an automorphism of $G$ and denote the group consisting of all automorphisms of $G$ by $\mathrm{Aut}~G$. As well-known, the problem of determining whether or not two given graphs are isomorphic is called Graph Isomorphism Problem (GI). One of key steps in resolving GI is to work out the partition $Π^*_G$ of $V(G)$ composed of orbits of $\mathrm{Aut}~G$. By means of geometric features of $Π^*_G$ and combinatorial constructions such as the multipartite graph $[Π^*_{t_1},\cdots,Π^*_{t_s}]$, we can reduce the problem of determining $Π_G^*$ to that of working out a series of partitions of $V(G)$ each of which consists of orbits of a stabilizer that fixes a sequence of vertices of $G$, and thus the determination of the partition $Π^*_v$ is a critical transition.
On the other hand, we have for a given subspace $U \subseteq \mathbb{R}^n$ a permutation group $\mathrm{Aut}~U := \{ σ\in S_n : σ~ U = U \}$. As a matter of fact, $\mathrm{Aut}~G = \cap_{λ\in \mathrm{spec} \mathbf{A}(G) } \mathrm{Aut}~V_λ$, and moreover we can obtain a good approximation $Π[ \oplus V_λ ; v ]$ to $Π_v^*$ by analyzing a decomposition of $V_λ$ resulted from the division of $V_λ$ by subspaces $\{ \mathrm{proj}[ V_λ ]( \pmb{e}_v )^{\perp} : v \in V(G) \}$. In fact, there is a close relation among subspaces spanned by cells of $Π[ \oplus V_λ ; v ]$ of $G$, which enables us to determine $Π_v^*$ more efficiently. In virtue of that, we devise a deterministic algorithm solving GI in time $n^{ O( \log n ) }$.