Search CORE

32 research outputs found

Gene network reconstruction from microarray data

Author: AV Werhli
B Efron
Florence Jaffrezic
Gwenola Tosser-Klopp
J Hausser
J Schäfer
J Whittaker
R Opgen-Rhein
W Swinkels
Publication venue: BioMed Central
Publication date: 01/01/2009
Field of study

Abstract Background Often, software available for biological pathways reconstruction rely on literature search to find links between genes. The aim of this study is to reconstruct gene networks from microarray data, using Graphical Gaussian models. Results The <it>GeneNet </it>R package was applied to the Eadgene chicken infection data set. No significant edges were found for the list of differentially expressed genes between conditions MM8 and MA8. On the other hand, a large number of significant edges were found among 85 differentially expressed genes between conditions MM8 and MM24. Conclusion Many edges were inferred from the microarray data. Most of them could, however, not be validated using other pathway reconstruction software. This was partly due to the fact that a quite large proportion of the differentially expressed genes were not annotated. Further biological validation is therefore needed for these networks, using for example in vitro invalidation of genes.</p

Crossref

Springer - Publisher Connector

Directory of Open Access Journals

PubMed Central

ProdInra

Constructing non-stationary Dynamic Bayesian Networks with a flexible lag choosing mechanism

Author: A Bernard
A Hall
A Nobile
A Para
AJ Hartemink
AV Werhli
CA Benedict
D Heckerman
D Husmeier
F Guo
H Duan
H Yu
HH McAdams
J Yu
Jr JD
Jun Huan
JW Robinson
K Honda
K Murphy
M Grzegorczy
M Zou
MF Covington
MN Arbeitman
N Friedman
N Nariai
P Mas
PA Salome
PJ Green
RM Cripps
S Chib
S Imoto
S Imoto
S Raza
SY Kim
T Mizuno
T Sandmann
W Zhao
W Zhao
Yi Jia
Publication venue: BioMed Central
Publication date: 01/01/2010
Field of study

Abstract Background Dynamic Bayesian Networks (DBNs) are widely used in regulatory network structure inference with gene expression data. Current methods assumed that the underlying stochastic processes that generate the gene expression data are stationary. The assumption is not realistic in certain applications where the intrinsic regulatory networks are subject to changes for adapting to internal or external stimuli. Results In this paper we investigate a novel non-stationary DBNs method with a potential regulator detection technique and a flexible lag choosing mechanism. We apply the approach for the gene regulatory network inference on three non-stationary time series data. For the Macrophages and Arabidopsis data sets with the reference networks, our method shows better network structure prediction accuracy. For the Drosophila data set, our approach converges faster and shows a better prediction accuracy on transition times. In addition, our reconstructed regulatory networks on the Drosophila data not only share a lot of similarities with the predictions of the work of other researchers but also provide many new structural information for further investigation. Conclusions Compared with recent proposed non-stationary DBNs methods, our approach has better structure prediction accuracy By detecting potential regulators, our method reduces the size of the search space, hence may speed up the convergence of MCMC sampling.</p

Crossref

Springer - Publisher Connector

Directory of Open Access Journals

KU ScholarWorks

PubMed Central

Using Stochastic Causal Trees to Augment Bayesian Networks for Modeling eQTL Datasets

Author: AFM Smith
Ambuj K Singh
AS Dimas
AV Werhli
BE Stranger
BJ Chen
D Heckerman
D Husmeier
D Husmeier
D Madigan
DC Kulp
DJ Lockhart
DM Ruderfer
E Chaibub Neto
EE Schadt
EO Perlstein
GA Churchill
J Pearl
J Zhu
J Zhu
J Zhu
JD Storey
JJ Faith
JJ Keurentjes
Kyle C Chipman
M Ashburner
M Morley
M Schena
MH Kutner
N Bing
N Friedman
N Friedman
O Litvin
RB Brem
RB Brem
RC Jansen
RW Doerge
S Imoto
S Mukherjee
SI Lee
W Pan
W Zhang
W Zou
Y Benjamini
Z Wang
Publication venue: BioMed Central
Publication date: 01/01/2011
Field of study

Abstract Background The combination of genotypic and genome-wide expression data arising from segregating populations offers an unprecedented opportunity to model and dissect complex phenotypes. The immense potential offered by these data derives from the fact that genotypic variation is the sole source of perturbation and can therefore be used to reconcile changes in gene expression programs with the parental genotypes. To date, several methodologies have been developed for modeling eQTL data. These methods generally leverage genotypic data to resolve causal relationships among gene pairs implicated as associates in the expression data. In particular, leading studies have augmented Bayesian networks with genotypic data, providing a powerful framework for learning and modeling causal relationships. While these initial efforts have provided promising results, one major drawback associated with these methods is that they are generally limited to resolving causal orderings for transcripts most proximal to the genomic loci. In this manuscript, we present a probabilistic method capable of learning the causal relationships between transcripts at all levels in the network. We use the information provided by our method as a prior for Bayesian network structure learning, resulting in enhanced performance for gene network reconstruction. Results Using established protocols to synthesize eQTL networks and corresponding data, we show that our method achieves improved performance over existing leading methods. For the goal of gene network reconstruction, our method achieves improvements in recall ranging from 20% to 90% across a broad range of precision levels and for datasets of varying sample sizes. Additionally, we show that the learned networks can be utilized for expression quantitative trait loci mapping, resulting in upwards of 10-fold increases in recall over traditional univariate mapping. Conclusions Using the information from our method as a prior for Bayesian network structure learning yields large improvements in accuracy for the tasks of gene network reconstruction and expression quantitative trait loci mapping. In particular, our method is effective for establishing causal relationships between transcripts located both proximally and distally from genomic loci.</p

Crossref

Springer - Publisher Connector

Directory of Open Access Journals

PubMed Central

Casual Compressive Sensing for Gene Network Inference

Author: A Butte
A Fujita
A Margolin
A Margolin
A Rao
A Shojaie
A Werhli
AC Lozano
Amin Emad
BE Perrin
BS Chen
C Olsen
C Sima
CA Penfold
CWJ Granger
D Husmeier
D Marbach
D Ruklisa
Daniele Marinazzo
DL Donoho
E Van Den Berg
EJ Candès
F Emmert-Streib
G Altay
G Della Gatta
G Stolovitzky
H de Jong
HE Samad
I Cantone
J Dingel
J Dougherty
J Watkinson
J Wright
J Yu
JF Geweke
JJ Faith
K Liang
M Bansal
M Deng
M Xu
M Zou
ML Whitfield
N Friedman
N Mukhopadhyay
Olgica Milenkovic
PE Meyer
PM Long
R Laubenbacher
R Penrose
R Tibshirani
RR Vallabhajosyula
S Becker
S Kauffman
T Chen
TS Gardner
W Dai
W Liu
W Zhao
X Cai
Y Prat
Publication venue: 'Public Library of Science (PLoS)'
Publication date: 01/01/2012
Field of study

We propose a novel framework for studying causal inference of gene interactions using a combination of compressive sensing and Granger causality techniques. The gist of the approach is to discover sparse linear dependencies between time series of gene expressions via a Granger-type elimination method. The method is tested on the Gardner dataset for the SOS network in E. coli, for which both known and unknown causal relationships are discovered

arXiv.org e-Print Archive

CiteSeerX

Public Library of Science (PLOS)

Crossref

Directory of Open Access Journals

PubMed Central

FigShare

Nonparametric identification of regulatory interactions from spatial and temporal gene expression data

Author: A Aswani
A Aswani
A Butte
A Marco
A Rao
A Savitzky
A Turing
A Werhli
Anil Aswani
C Fowlkes
C Luengo Hendriks
Charless C Fowlkes
Claire J Tomlin
D Arnosti
D Bickel
D Ruppert
David W Knowles
E Cinquemani
E Segal
F Markowetz
F Sauer
G von Dassow
H de Jong
H Janssens
H Schneeweiß
I Baianu
J Jaeger
J Lee
J Murray
J Stuart
J Yu
James Brown
K Harding
M Bansal
M Eisen
M Fujioka
M Ptashne
Mark D Biggin
MK Stephen Yeung
N Friedman
N Friedman
P Bickel
P D'haeseleer
Peter Bickel
R Bonneau
R Porreca
R Steuer
S Small
Soile VE Keränen
W Fakhouri
Xy Li
Z Xiang
Publication venue: BioMed Central
Publication date: 01/01/2010
Field of study

Abstract Background The correlation between the expression levels of transcription factors and their target genes can be used to infer interactions within animal regulatory networks, but current methods are limited in their ability to make correct predictions. Results Here we describe a novel approach which uses nonparametric statistics to generate ordinary differential equation (ODE) models from expression data. Compared to other dynamical methods, our approach requires minimal information about the mathematical structure of the ODE; it does not use qualitative descriptions of interactions within the network; and it employs new statistics to protect against over-fitting. It generates spatio-temporal maps of factor activity, highlighting the times and spatial locations at which different regulators might affect target gene expression levels. We identify an ODE model for <it>eve </it>mRNA pattern formation in the <it>Drosophila melanogaster </it>blastoderm and show that this reproduces the experimental patterns well. Compared to a non-dynamic, spatial-correlation model, our ODE gives 59% better agreement to the experimentally measured pattern. Our model suggests that protein factors frequently have the potential to behave as both an activator and inhibitor for the same <it>cis</it>-regulatory module depending on the factors' concentration, and implies different modes of activation and repression. Conclusions Our method provides an objective quantification of the regulatory potential of transcription factors in a network, is suitable for both low- and moderate-dimensional gene expression datasets, and includes improvements over existing dynamic and static models.</p

Crossref

Springer - Publisher Connector

Directory of Open Access Journals

PubMed Central

eScholarship - University of California

A negative selection heuristic to predict new transcriptional targets

Author: A Karatzoglou
A Polynikis
AA Margolin
AH Brivanlou
AV Werhli
B Liu
B Liu
B Zhang
C Elkan
C Wang
CW Hsu
F Mordelet
H Salgado
H Yu
HC Kim
HT Lin
IH Witten
JJ Faith
JJ Faith
JP Vert
JR Bock
K Basso
L Cerulo
L Cerulo
Luigi Cerulo
M Bansal
M Bansal
M Ceccarelli
M Grzegorczyk
Michele Ceccarelli
P Stegmaier
P Zoppoli
Pietro Zoppoli
RA Irizarry
S Liang
TS Gardner
U Alon
V Matys
Vincenzo Paduano
W Ci
X Li
X Li
X Wang
XL Li
Y Yamanishi
Publication venue: 'Springer Science and Business Media LLC'
Publication date
Field of study

Crossref

Evaluation and improvement of the regulatory inference for large co-expression networks with limited sample size

Author: A Fuente de la
A Reverter
AA Margolin
AL Barabasi
AV Werhli
B Zhang
B-E Perrin
C Olsen
Cristiane P. G. Calixto
D Marbach
D Marbach
F Markowetz
F Steinke
G Altay
H Hache
H Jong de
H Lahdesmaki
H Ma
H Peng
J Linde
J Schäfer
JD Allen
JJ Faith
John W. S. Brown
KP Murphy
KY Yip
L Song
LJ Kogelman
ME Studham
MV DiLeo
N Friedman
N Friedman
N Omranian
Nikoleta Tzioutziou
NS Watson-Haigh
P Bellot
P Langfelder
P Sarder
PB Madhamshettiwar
Ping Lin
R Albert
R Dehghannasiri
RJ Flassig
RJ Prill
Robbie Waugh
Runxuan Zhang
S Ballouz
S Bornholdt
S Kim
S Martin
S Rogers
S Roy
SD Walter
SM Ud-Dean
SM Ud-Dean
T Bulcke Van den
T Saito
T Schaffter
TM Cover
V Huynh-Thu
W Zhao
WC Young
Wenbin Guo
Y Tu
Y Zuo
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/06/2017
Field of study

Abstract Background Co-expression has been widely used to identify novel regulatory relationships using high throughput measurements, such as microarray and RNA-seq data. Evaluation studies on co-expression network analysis methods mostly focus on networks of small or medium size of up to a few hundred nodes. For large networks, simulated expression data usually consist of hundreds or thousands of profiles with different perturbations or knock-outs, which is uncommon in real experiments due to their cost and the amount of work required. Thus, the performances of co-expression network analysis methods on large co-expression networks consisting of a few thousand nodes, with only a small number of profiles with a single perturbation, which more accurately reflect normal experimental conditions, are generally uncharacterized and unknown. Methods We proposed a novel network inference methods based on Relevance Low order Partial Correlation (RLowPC). RLowPC method uses a two-step approach to select on the high-confidence edges first by reducing the search space by only picking the top ranked genes from an intial partial correlation analysis and, then computes the partial correlations in the confined search space by only removing the linear dependencies from the shared neighbours, largely ignoring the genes showing lower association. Results We selected six co-expression-based methods with good performance in evaluation studies from the literature: Partial correlation, PCIT, ARACNE, MRNET, MRNETB and CLR. The evaluation of these methods was carried out on simulated time-series data with various network sizes ranging from 100 to 3000 nodes. Simulation results show low precision and recall for all of the above methods for large networks with a small number of expression profiles. We improved the inference significantly by refinement of the top weighted edges in the pre-inferred partial correlation networks using RLowPC. We found improved performance by partitioning large networks into smaller co-expressed modules when assessing the method performance within these modules. Conclusions The evaluation results show that current methods suffer from low precision and recall for large co-expression networks where only a small number of profiles are available. The proposed RLowPC method effectively reduces the indirect edges predicted as regulatory relationships and increases the precision of top ranked predictions. Partitioning large networks into smaller highly co-expressed modules also helps to improve the performance of network inference methods. The RLowPC R package for network construction, refinement and evaluation is available at GitHub: https://github.com/wyguo/RLowPC

Crossref

Directory of Open Access Journals

University of Dundee Online Publications

Bagging Statistical Network Inference from Large-Scale Gene Expression Data

Modern biology and medicine aim at hunting molecular and cellular causes of biological functions and diseases. Gene regulatory networks (GRN) inferred from gene expression data are considered an important aid for this research by providing a map of molecular interactions. Hence, GRNs have the potential enabling and enhancing basic as well as applied research in the life sciences. In this paper, we introduce a new method called BC3NET for inferring causal gene regulatory networks from large-scale gene expression data. BC3NET is an ensemble method that is based on bagging the C3NET algorithm, which means it corresponds to a Bayesian approach with noninformative priors. In this study we demonstrate for a variety of simulated and biological gene expression data from S. cerevisiae that BC3NET is an important enhancement over other inference methods that is capable of capturing biochemical interactions from transcription regulation and protein-protein interaction sensibly. An implementation of BC3NET is freely available as an R package from the CRAN repository

Queen's University Belfast Research Portal

CiteSeerX

Public Library of Science (PLOS)

Crossref

Directory of Open Access Journals

PubMed Central