Search CORE

Algorithm engineering for optimal alignment of protein structure distance matrices

Author: A. Andreeva
A. Caprara
A. Marin
A. Schrijver
C. Berbalk
D. Wu
D.A. Pelta
E. Althaus
G. Mayr
Gunnar W. Klau
H. Hasegawa
H.P. Lenhof
I. Wohlers
Inken Wohlers
L. Holm
N. Malod-Dognin
P. Di Lena
R. Andonov
R. Kolodny
R.H. Lathrop
Rumen Andonov
T. Havel
T. Kawabata
W. Xie
W.R. Taylor
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2011
Field of study

Protein structural alignment is an important problem in computational biology. In this paper, we present first successes on provably optimal pairwise alignment of protein inter-residue distance matrices, using the popular Dali scoring function. We introduce the structural alignment problem formally, which enables us to express a variety of scoring functions used in previous work as special cases in a unified framework. Further, we propose the first mathematical model for computing optimal structural alignments based on dense inter-residue distance matrices. We therefore reformulate the problem as a special graph problem and give a tight integer linear programming model. We then present algorithm engineering techniques to handle the huge integer linear programs of real-life distance matrix alignment problems. Applying these techniques, we can compute provably optimal Dali alignments for the very first time

HAL-CentraleSupelec

CiteSeerX

INRIA a CCSD electronic archive server

HAL-Rennes 1

An Exact Algorithm for Side-Chain Placement in Protein Design

Author: A. Hildebrandt
A.A. Canutescu
A.R. Leach
B. Chazelle
B. Kuhlman
C. Voigt
C. Yanover
C.L. Kingsford
D. Sontag
E. Althaus
G. Dantas
G. Dantas
Gunnar W. Klau
J. Desmet
J. Desmet
J. Xu
J.D. Bloom
K. Mehlhorn
L. Wernisch
M. Held
N.A. Pierce
N.A. Pierce
Nora C. Toussaint
P.S. Shah
R.F. Goldstein
R.L. Dunbrack
Stefan Canzar
W. Xie
Z. Xiang
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2011
Field of study

Computational protein design aims at constructing novel or improved functions on the structure of a given protein backbone and has important applications in the pharmaceutical and biotechnical industry. The underlying combinatorial side-chain placement problem consists of choosing a side-chain placement for each residue position such that the resulting overall energy is minimum. The choice of the side-chain then also determines the amino acid for this position. Many algorithms for this NP-hard problem have been proposed in the context of homology modeling, which, however, reach their limits when faced with large protein design instances. In this paper, we propose a new exact method for the side-chain placement problem that works well even for large instance sizes as they appear in protein design. Our main contribution is a dedicated branch-and-bound algorithm that combines tight upper and lower bounds resulting from a novel Lagrangian relaxation approach for side-chain placement. Our experimental results show that our method outperforms alternative state-of-the art exact approaches and makes it possible to optimally solve large protein design instances routinely

CSA: Comprehensive comparison of pairwise protein structure alignments

Author: Andonov
Barthel
Berbalk
Berman
Carugo
Emekli
G. W. Klau
GODZIK
Hamelryck
Hasegawa
Holm
Holm
I. Wohlers
Kabsch
Kawabata
Kawabata
Mayr
N. Malod-Dognin
R. Andonov
Shih
Zhang
Zhang
Publication venue
Publication date: 01/12/2011
Field of study

htmlabstractCSA is a web server for the computation, evaluation and comprehensive comparison of pairwise protein structure alignments. Its exact alignment engine computes either optimal, top-scoring alignments or heuristic alignments with quality guarantee for the inter-residue distance-based scorings of contact map overlap, PAUL, DALI and MATRAS. These and additional, uploaded alignments are compared using a number of quality measures and intuitive visualizations. CSA brings new insight into the structural relationship of the protein pairs under investigation and is a valuable tool for studying structural similarities. It is available at http://csa.project.cwi.nl

HAL-CentraleSupelec

VU Research Portal

CiteSeerX

INRIA a CCSD electronic archive server

UCL Discovery

HAL-Rennes 1

PAUL: Protein structural alignment using integer linear programming and Lagrangian relaxation

Author: A Caprara
Francisco S Domingues
G Mayr
Gunnar W Klau
IN Shindyalov
Inken Wohlers
J Jung
K Mizuguchi
L Holm
Lars Petzold
O Bachar
T Kawabata
Y Ye
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2009
Field of study

Springer - Publisher Connector

A critical evaluation of network and pathway based classifiers for outcome prediction in breast cancer

Author: A Subramanian
C Desmedt
Christine Staiger
D Hanahan
E Lee
F Reyal
G Abraham
GR Mishra
Gunnar W. Klau
HY Chuang
I Ulitsky
IW Taylor
Joaquín Dopazo
K Chin
KR Brown
L Ein-Dor
L Tian
LD Miller
LFA Wessels
LJ van’t Veer
Lodewyk F. A. Wessels
M Kanehisa
Marcus Dittrich
MH van Vliet
MJ van de Vijver
ML Gatza
MT Dittrich
P Dao
Raul Kooter
S Loi
S Ma
SA Chowdhury
Sidney Cadot
Tobias Müller
TSK Prasad
V Popovici
Y Pawitan
Y Wang
Publication venue: 'Public Library of Science (PLoS)'
Publication date: 01/10/2011
Field of study

Recently, several classifiers that combine primary tumor data, like gene expression data, and secondary data sources, such as protein-protein interaction networks, have been proposed for predicting outcome in breast cancer. In these approaches, new composite features are typically constructed by aggregating the expression levels of several genes. The secondary data sources are employed to guide this aggregation. Although many studies claim that these approaches improve classification performance over single gene classifiers, the gain in performance is difficult to assess. This stems mainly from the fact that different breast cancer data sets and validation procedures are employed to assess the performance. Here we address these issues by employing a large cohort of six breast cancer data sets as benchmark set and by performing an unbiased evaluation of the classification accuracies of the different approaches. Contrary to previous claims, we find that composite feature classifiers do not outperform simple single gene classifiers. We investigate the effect of (1) the number of selected features; (2) the specific gene set from which features are selected; (3) the size of the training set and (4) the heterogeneity of the data set on the performance of composite feature and single gene classifiers. Strikingly, we find that randomization of secondary data sources, which destroys all biological information in these sources, does not result in a deterioration in performance of composite feature classifiers. Finally, we show that when a proper correction for gene set size is performed, the stability of single gene sets is similar to the stability of composite feature sets. Based on these results there is currently no reason to prefer prognostic classifiers based on composite features over single gene classifiers for predicting outcome in breast cancer

Public Library of Science (PLOS)

VU Research Portal

Directory of Open Access Journals

Online-Publikations-Server der Universität Würzburg

FigShare

W hats

Author: Alexander Schönhuth
Gunnar W. Klau
Hartl D.
Lancia G.
Leen Stougie
Leo van Iersel
Murray Patterson
Nadia Pisanti
Tobias Marschall
Publication venue: 'Mary Ann Liebert Inc'
Publication date
Field of study

Optimizing topological cascade resilience based on the structure of terrorist networks

Author: AE Motter
AE Motter
AE Motter
Alexander Gutfraind
C Morselli
D Centola
D Kempe
DJ Watts
FO Miksche
G Gunther
G Iori
G Sharp
G Woo
GW Klau
I Dobson
J Leskovec
J Raab
J Rodriguez
J Zawodny
JK J Leskovec
JK Johnson
M Draief
M Ripeanu
M Sageman
MEJ Newman
MEJ Newman
MEJ Newman
MEJ Newman
Olaf Sporns
P Crepey
R Guimerà
R Pastor-Sarorras
RH Lindelauf
RH Lindelauf
S Battiston
SV Buldyrev
V Latora
VE Krebs
W Huang
WE Baker
YC Lai
Publication venue: 'Public Library of Science (PLoS)'
Publication date: 27/08/2010
Field of study

Complex socioeconomic networks such as information, finance and even terrorist networks need resilience to cascades - to prevent the failure of a single node from causing a far-reaching domino effect. We show that terrorist and guerrilla networks are uniquely cascade-resilient while maintaining high efficiency, but they become more vulnerable beyond a certain threshold. We also introduce an optimization method for constructing networks with high passive cascade resilience. The optimal networks are found to be based on cells, where each cell has a star topology. Counterintuitively, we find that there are conditions where networks should not be modified to stop cascades because doing so would come at a disproportionate loss of efficiency. Implementation of these findings can lead to more cascade-resilient networks in many diverse areas.Comment: 26 pages. v2: In review at Public Library of Science ON

Public Library of Science (PLOS)

Directory of Open Access Journals

Enhancing the accuracy of HMM-based conserved pathway prediction using global correspondence scores

Author: A Osman
AL Barabasi
AL Barabasi
B Srinivasan
BJ Yoon
BP Kelley
Byung-Jun Yoon
CS Liao
G Klau
J Flannick
J Flannick
M Ashburner
M Kanehisa
M Koyutürk
M Zaslavskiy
Q Yang
R Aebersold
R Pinter
R Sharan
R Sharan
R Singh
S Sahraeian
Sayed Mohammad Ebrahim Sahraeian
SME Sahraeian
W Tian
X Qian
X Qian
Xiaoning Qian
Z Li
Publication venue: 'Springer Science and Business Media LLC'
Publication date
Field of study