Papers by Fenix W. D. Huang
8 paper(s) by this author
· All BibTeX
A topological framework for signed permutations
Published
• View Publication
• BIB
In this paper we present a topological framework for studying signed permutations and their reversal distance. As a result we can give an alternative approach and interpretation of the Hannenhalli-Pevzner formula for the reversal distance of signed permutations. Our approach utlizes the Poincaré dual, upon which reversals act in a particular way and obsoletes the notion of "padding" of the signed permutations. To this end we construct a bijection between signed permutations and an equivalence class of particular fatgraphs, called $π$-maps, and analyze the action of reversals on the latter. We show that reversals act via either slicing, gluing or half-flipping of external vertices, which implies that any reversal changes the topological genus by at most one. Finally we revisit the Hannenhalli-Pevzner formula employing orientable and non-orientable, irreducible, $π$-maps.
Shapes of topological RNA structures
Published
• View Publication
• BIB
A topological RNA structure is derived from a diagram and its shape is obtained by collapsing the stacks of the structure into single arcs and by removing any arcs of length one. Shapes contain key topological, information and for fixed topological genus there exist only finitely many such shapes. We shall express topological RNA structures as unicellular maps, i.e. graphs together with a cyclic ordering of their half-edges. In this paper we prove a bijection of shapes of topological RNA structures. We furthermore derive a linear time algorithm generating shapes of fixed topological genus. We derive explicit expressions for the coefficients of the generating polynomial of these shapes and the generating function of RNA structures of genus $g$. Furthermore we outline how shapes can be used in order to extract essential information of RNA structure databases.
Uniform generation of RNA pseudoknot structures with genus filtration
Published
• View Publication
• BIB
In this paper we present a sampling framework for RNA structures of fixed topological genus. We introduce a novel, linear time, uniform sampling algorithm for RNA structures of fixed topological genus $g$, for arbitrary $g>0$. Furthermore we develop a linear time sampling algorithm for RNA structures of fixed topological genus $g$ that are weighted by a simplified, loop-based energy functional. For this process the partition function of the energy functional has to be computed once, which has $O(n^2)$ time complexity.
On the combinatorics of sparsification
Published
• View Publication
• BIB
Background: We study the sparsification of dynamic programming folding algorithms of RNA structures. Sparsification applies to the mfe-folding of RNA structures and can lead to a significant reduction of time complexity. Results: We analyze the sparsification of a particular decomposition rule, $Λ^*$, that splits an interval for RNA secondary and pseudoknot structures of fixed topological genus. Essential for quantifying the sparsification is the size of its so called candidate set. We present a combinatorial framework which allows by means of probabilities of irreducible substructures to obtain the expected size of the set of $Λ^*$-candidates. We compute these expectations for arc-based energy models via energy-filtered generating functions (GF) for RNA secondary structures as well as RNA pseudoknot structures. For RNA secondary structures we also consider a simplified loop-energy model. This combinatorial analysis is then compared to the expected number of $Λ^*$-candidates obtained from folding mfe-structures. In case of the mfe-folding of RNA secondary structures with a simplified loop energy model our results imply that sparsification provides a reduction of time complexity by a constant factor of 91% (theory) versus a 96% reduction (experiment). For the "full" loop-energy model there is a reduction of 98% (experiment).
Topology of RNA-RNA interaction structures
Published
• View Publication
• BIB
The topological filtration of interacting RNA complexes is studied and the role is analyzed of certain diagrams called irreducible shadows, which form suitable building blocks for more general structures. We prove that for two interacting RNAs, called interaction structures, there exist for fixed genus only finitely many irreducible shadows. This implies that for fixed genus there are only finitely many classes of interaction structures. In particular the simplest case of genus zero already provides the formalism for certain types of structures that occur in nature and are not covered by other filtrations. This case of genus zero interaction structures is already of practical interest, is studied here in detail and found to be expressed by a multiple context-free grammar extending the usual one for RNA secondary structures. We show that in $O(n^6)$ time and $O(n^4)$ space complexity, this grammar for genus zero interaction structures provides not only minimum free energy solutions but also the complete partition function and base pairing probabilities.
On the uniform generation of modular diagrams
Published
• View Publication
• BIB
In this paper we present an algorithm that generates $k$-noncrossing, $σ$-modular diagrams with uniform probability. A diagram is a labeled graph of degree $\le 1$ over $n$ vertices drawn in a horizontal line with arcs $(i,j)$ in the upper half-plane. A $k$-crossing in a diagram is a set of $k$ distinct arcs $(i_1, j_1), (i_2, j_2),\ldots,(i_k, j_k)$ with the property $i_1 < i_2 < \ldots < i_k < j_1 < j_2 < \ldots< j_k$. A diagram without any $k$-crossings is called a $k$-noncrossing diagram and a stack of length $σ$ is a maximal sequence $((i,j),(i+1,j-1),\dots,(i+(σ-1),j-(σ-1)))$. A diagram is $σ$-modular if any arc is contained in a stack of length at least $σ$. Our algorithm generates after $O(n^k)$ preprocessing time,
$k$-noncrossing, $σ$-modular diagrams in $O(n)$ time and space complexity.
RNA-RNA interaction prediction: partition function and base pair pairing probabilities
Published
• View Publication
• BIB
In this paper, we study the interaction of an antisense RNA and its target mRNA, based on the model introduced by Alkan {\it et al.} (Alkan {\it et al.}, J. Comput. Biol., Vol:267--282, 2006). Our main results are the derivation of the partition function
\cite{Backhofen} (Chitsaz {\it et al.}, Bioinformatics, to appear, 2009), based on the concept of tight-structure and the computation of the base pairing probabilities. This paper contains the folding algorithm {\sf rip} which computes the partition function as well as the base pairing probabilities in $O(N^4M^2)+O(N^2M^4)$ time and $O(N^2M^2)$ space, where $N,M$ denote the lengths of the interacting sequences.
Folding 3-noncrossing RNA pseudoknot structures
Published
• View Publication
• BIB
In this paper we present a selfcontained analysis and description of the novel {\it ab initio} folding algorithm {\sf cross}, which generates the minimum free energy (mfe), 3-noncrossing, $σ$-canonical RNA structure. Here an RNA structure is 3-noncrossing if it does not contain more than three mutually crossing arcs and $σ$-canonical, if each of its stacks has size greater or equal than $σ$. Our notion of mfe-structure is based on a specific concept of pseudoknots and respective loop-based energy parameters. The algorithm decomposes into three parts: the first is the inductive construction of motifs and shadows, the second is the generation of the skeleta-trees rooted in irreducible shadows and the third is the saturation of skeleta via context dependent dynamic programming routines.