Search CORE

111 research outputs found

Sequence Analysis of the Genome of an Oil-Bearing Tree, Jatropha curcas L.

Author: A. Atefeh
A. Fujiyama
A. Muraki
A. Toyoda
A. Watanabe
Alvarez-Buylla
Ashburner
Aubourg
Burge
C. Minami
C. Takahashi
Chan
Costa
E. Fukai
E. Makigano
Eitas
Fairless
Gomez-Alvarez
H. Hirakawa
H. Tsuruoka
Haas
Hartmann
Huang
Huang
J. A. Cartagena
K. Fukui
K. Kawashima
Lee
Lee
Lin
M. Kato
M. Kohara
M. Yamada
N. Nakazaki
N. Ohmido
N. Shibagaki
N. Wada
Qin
Qin
Quinlan
Rijpkema
Rozen
S. Isobe
S. Kajiyama
S. Matsunaga
S. Nakayama
S. Sasamoto
S. Sato
S. Tabata
S. Tsuchimoto
S. Yuasa
Saitou
Stirpe
T. Aizu
T. Kohinata
T. Shin-i
Temnykh
Tong
Unseld
Wang
Y. Kishida
Y. Kohara
Y. Minakuchi
Ye
Zhang
Publication venue: Oxford University Press
Publication date: 13/12/2010
Field of study

The whole genome of Jatropha curcas was sequenced, using a combination of the conventional Sanger method and new-generation multiplex sequencing methods. Total length of the non-redundant sequences thus obtained was 285 858 490 bp consisting of 120 586 contigs and 29 831 singlets. They accounted for ∼95% of the gene-containing regions with the average G + C content was 34.3%. A total of 40 929 complete and partial structures of protein encoding genes have been deduced. Comparison with genes of other plant species indicated that 1529 (4%) of the putative protein-encoding genes are specific to the Euphorbiaceae family. A high degree of microsynteny was observed with the genome of castor bean and, to a lesser extent, with those of soybean and Arabidopsis thaliana. In parallel with genome sequencing, cDNAs derived from leaf and callus tissues were subjected to pyrosequencing, and a total of 21 225 unigene data have been generated. Polymorphism analysis using microsatellite markers developed from the genomic sequence data obtained was performed with 12 J. curcas lines collected from various parts of the world to estimate their genetic diversity. The genomic sequence and accompanying information presented here are expected to serve as valuable resources for the acceleration of fundamental and applied research with J. curcas, especially in the fields of environment-related research such as biofuel production. Further information on the genomic sequences and DNA markers is available at http://www.kazusa.or.jp/jatropha/

Crossref

PubMed Central

Osaka University Knowledge Archive

Integrative Identification of Arabidopsis Mitochondrial Proteome and Its Function Exploitation through Protein Interaction Network

Author: A Drawid
A Drawid
A Höglund
A Kumar
A Reinhardt
A Vianello
AA Vashisht
AH Millar
AH Millar
BJ Haas
BW Rhee SY
C Bai
C Guda
C Jonak
CG Bartoli
CG Kurland
CM Lee
D Skowyra
DA Bota
David Moore
E Delannoy
E Jambrina
E Mazzucotelli
EE Patton
EE Patton
EH Kruft V
EM Marcotte
EO Karlberg
F Rébeillé
GW Tian
H Bannai
H Fölsch
H Prokisch
HN Chua
I Small
ID Small
IM Moller
J Balk
J Bardel
J Cui
J Huang
J Kilian
JA Kreps
Jian Cui
Jinghua Liu
JK Zhu
JL Heazlewood
JL Heazlewood
JL Heazlewood
K Ishizaki
K Meierhoff
KP O'Brien
L Li
LJ Lu
M Teige
M Unseld
MG Claros
MG Claros
O Emanuelsson
O Van Aken
OA Koroleva
P Horton
P Pavlidis
R Nair
R Nair
RA Irizarry
S Hua
S Killcoyne
S Li
S Ma
S Maere
S Mahajan
S Mili
SG Andersson
T Sing
Tieliu Shi
V Gueguen
VK Mootha
Vladimir Uversky
W Werhahn
WK Huh
X Gong
Y Gavel
YD Cai
YD Cai
Yuhua Li
Z Liu
Z Yuan
Publication venue: Public Library of Science
Publication date: 31/01/2011
Field of study

Mitochondria are major players on the production of energy, and host several key reactions involved in basic metabolism and biosynthesis of essential molecules. Currently, the majority of nucleus-encoded mitochondrial proteins are unknown even for model plant Arabidopsis. We reported a computational framework for predicting Arabidopsis mitochondrial proteins based on a probabilistic model, called Naive Bayesian Network, which integrates disparate genomic data generated from eight bioinformatics tools, multiple orthologous mappings, protein domain properties and co-expression patterns using 1,027 microarray profiles. Through this approach, we predicted 2,311 candidate mitochondrial proteins with 84.67% accuracy and 2.53% FPR performances. Together with those experimental confirmed proteins, 2,585 mitochondria proteins (named CoreMitoP) were identified, we explored those proteins with unknown functions based on protein-protein interaction network (PIN) and annotated novel functions for 26.65% CoreMitoP proteins. Moreover, we found newly predicted mitochondrial proteins embedded in particular subnetworks of the PIN, mainly functioning in response to diverse environmental stresses, like salt, draught, cold, and wound etc. Candidate mitochondrial proteins involved in those physiological acitivites provide useful targets for further investigation. Assigned functions also provide comprehensive information for Arabidopsis mitochondrial proteome

Public Library of Science (PLOS)

Crossref

Directory of Open Access Journals

PubMed Central

Rapid Sequencing of the Bamboo Mitochondrial Genome Using Illumina Technology and Parallel Episodic Evolution of Organelle Genomes in Grasses

Background: Compared to their counterparts in animals, the mitochondrial (mt) genomes of angiosperms exhibit a number of unique features. However, unravelling their evolution is hindered by the few completed genomes, of which are essentially Sanger sequenced. While next-generation sequencing technologies have revolutionized chloroplast genome sequencing, they are just beginning to be applied to angiosperm mt genomes. Chloroplast genomes of grasses (Poaceae) have undergone episodic evolution and the evolutionary rate was suggested to be correlated between chloroplast and mt genomes in Poaceae. It is interesting to investigate whether correlated rate change also occurred in grass mt genomes as expected under lineage effects. A time-calibrated phylogenetic tree is needed to examine rate change. Methodology/Principal Findings: We determined a largely completed mt genome from a bamboo, Ferrocalamus rimosivaginus (Poaceae), through Illumina sequencing of total DNA. With combination of de novo and reference-guided assembly, 39.5-fold coverage Illumina reads were finally assembled into scaffolds totalling 432,839 bp. The assembled genome contains nearly the same genes as the completed mt genomes in Poaceae. For examining evolutionary rate in grass mt genomes, we reconstructed a phylogenetic tree including 22 taxa based on 31 mt genes. The topology of the wellresolved tree was almost identical to that inferred from chloroplast genome with only minor difference. The inconsistency possibly derived from long branch attraction in mtDNA tree. By calculating absolute substitution rates, we found significan

CiteSeerX

Public Library of Science (PLOS)

Crossref

Directory of Open Access Journals

PubMed Central

The Francis Crick Institute