Search CORE

7 research outputs found

From Indexing Data Structures to de Bruijn Graphs

Author: A. Bankevich
A. Bowe
D. Gusfield
E.A. Rødland
J. Pell
L. Salmela
N. Bruijn de
P. Pevzner
R. Chikhi
T. Onodera
T.C. Conway
U. Manber
Y. Peng
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2014
Field of study

International audienceNew technologies have tremendously increased sequencing throughput com-pared to traditional techniques, thereby complicating DNA assembly. Hence, as-sembly programs resort to de Bruijn graphs (dBG) of k-mers of short reads to compute a set of long contigs, each being a putative segment of the sequenced molecule. Other types of DNA sequence analysis, as well as preprocessing of the reads for assembly, use classical data structures to index all substrings of the reads. It is thus interesting to exhibit algorithms that directly build a dBG of order k from a pre-existing index, and especially a contracted version of the dBG, where non branching paths are condensed into single nodes. Here, we formalise the relation-ship between suffix trees/arrays and dBGs, and exhibit linear time algorithms for constructing the full or contracted dBGs. Finally, we provide hints explaining why this bridge between indexes and dBGs enables to dynamically update the order k of the graph

HAL - Normandie Université

Crossref

From Indexing Data Structures to de Bruijn Graphs

Author: A. Bankevich
A. Bowe
D. Gusfield
E.A. Rødland
J. Pell
L. Salmela
N. Bruijn de
P. Pevzner
R. Chikhi
T. Onodera
T.C. Conway
U. Manber
Y. Peng
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2014
Field of study

HAL - Normandie Université

CiteSeerX

Crossref

INRIA a CCSD electronic archive server

HAL Descartes

Prediction of RNA Secondary Structure Including Kissing Hairpin Motifs

Author: A.E. Condon
C. Tuerk
C.Y. Chan
D. Deblasio
D.H. Mathews
D.H. Mathews
E. Rivas
E.A. Rødland
F.H.D. Batenburg van
H.L. Chen
I.L. Hofacker
J. Herold
J. Reeder
J. Ren
K.Y. Chang
M. Zuker
M.S. Andronescu
M.S. Andronescu
P.T.X. Li
R. Giegerich
R. Giegerich
R.B. Lyngsø
S. Wuchty
T. Akutsu
W.J.G. Melchers
Y. Frid
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2010
Field of study

Theis C, Janssen S, Giegerich R. Prediction of RNA Secondary Structure Including Kissing Hairpin Motifs. In: Moulton V, Singh M, eds. Algorithms in Bioinformatics. 10th international workshop (WABI 2010), proceedings. Lecture Notes in Bioinformatics. Vol 6293. Berlin: Springer; 2010: 52-64.We present three heuristic strategies for folding RNA sequences into secondary structures including kissing hairpin motifs. The new idea is to construct a kissing hairpin motif from an overlay of two simple canonical pseudoknots. The difficulty is that the overlay does not satisfy Bellman's Principle of Optimality, and the kissing hairpin cannot simply be built from optimal pseudoknots. Our strategies have time/space complexities of O(n^4)/O(n^2), O(n^4)/O(n^3), and O(n^5)/O(n^2). All strategies have been implemented in the program pKiss and were evaluated against known structures. Surprisingly, our simplest strategy performs best. As it has the same complexity as the previous algorithm for simple pseudoknots, the overlay idea opens a way to construct a variety of practically useful algorithms for pseudoknots of higher topological complexity within O(n^4) time and O(n^2) space

Crossref

Publications at Bielefeld University

K-Partite RNA Secondary Structures

Author: A. Condon
A.A. Ageev
B. Aspvall
B. Rastegari
C. Haslinger
C. Witwer
E. Rivas
E.A. Rødland
F.H.D. Batenburg van
F.H.D. Batenburg van
H. Jabbari
I.L. Hofacker
J. Cong
J. Kleinberg
J. Reeder
J. Ren
J. Ruan
M. Zuker
M.R. Garey
R. Nussinov
R.B. Lyngsø
R.B. Lyngsø
R.M. Dirks
S. Ieong
T. Akutsu
W. Unger
W. Unger
Y. Uemura
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2009
Field of study

Crossref

On the Representation of de Bruijn Graphs

Author: A. Bankevich
A. Bowe
B. Langmead
B.H. Bloom
D. Haussler
D.R. Zerbino
E.A. Rødland
G. Rizk
H. Li
H. Li
J. Pell
J. Sondow
J.R. Miller
J.T. Simpson
J.T. Simpson
J.T. Simpson
K. Salikhov
M. Burrows
M. Roberts
M. Roberts
M.G. Grabherr
P.A. Pevzner
R. Chikhi
R. Li
R. Li
R.M. Idury
S. Gnerre
S.L. Salzberg
T. Gagie
T.C. Conway
Z. Iqbal
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2014
Field of study

The de Bruijn graph plays an important role in bioinformatics, especially in the context of de novo assembly. However, the representation of the de Bruijn graph in memory is a computational bottleneck for many assemblers. Recent papers proposed a navigational data structure approach in order to improve memory usage. We prove several theoretical space lower bounds to show the limitation of these types of approaches. We further design and implement a general data structure (DBGFM) and demonstrate its use on a human whole-genome dataset, achieving space usage of 1.5 GB and a 46% improvement over previous approaches. As part of DBGFM, we develop the notion of frequency-based minimizers and show how it can be used to enumerate all maximal simple paths of the de Bruijn graph using only 43 MB of memory. Finally, we demonstrate that our approach can be integrated into an existing assembler by modifying the ABySS software to use DBGFM.Comment: Journal version (JCB). A preliminary version of this article was published in the proceedings of RECOMB 201

arXiv.org e-Print Archive

CiteSeerX

Crossref

INRIA a CCSD electronic archive server

HAL-Rennes 1

Closed Testing in Pharmaceutical Research: Historical and Recent Developments

Crossref

Two-Sphere Partition Functions and Gromov–Witten Invariants

Author: A. Bayer
A. Kapustin
A. Libgober
A. Losev
A. Strominger
A. Zamolodchikov
A.-M. Li
A.B. Givental
B. Craps
B. Kim
B. Wit de
B.H. Lian
B.R. Greene
B.R. Greene
C.H. Clemens
D.R. Morrison
D.R. Morrison
D.S. Freed
David R. Morrison
E. Witten
E. Witten
E. Witten
E.A. Rødland
E.N. Tjøtta
G. Festuccia
G.W. Moore
H. Jockers
Hans Jockers
I. Brunner
I. Ciocan-Fontanine
J. Carlson
Joshua M. Lapan
K. Hori
M. Bershadsky
M. Dine
M. Gromov
M. Gross
M. Kapustka
M. Kontsevich
M. Reid
M.-A. Bertin
M.T. Grisaru
Mauricio Romo
N. Hama
N.A. Nekrasov
P. Candelas
P. Candelas
P. Candelas
P. Candelas
P. Candelas
P. Candelas
P. Mayr
P.S. Aspinwall
P.S. Aspinwall
R. Donagi
R. Friedman
S. Cecotti
S. Cecotti
S. Hosono
S. Hosono
S. Pasquetti
T. Coates
T.H. Gulliksen
V. Bouchard
V. Periwal
V. Pestun
V.V. Batyrev
V.V. Batyrev
V.V. Batyrev
Vijay Kumar
W. Lerche
Y. Namikawa
Y. Ruan
Publication venue: 'Springer Science and Business Media LLC'
Publication date
Field of study

Crossref